2015
Autores
Silva, JMC; Carvalho, P; Lima, SR;
Publicação
2015 IEEE SYMPOSIUM ON COMPUTERS AND COMMUNICATION (ISCC)
Abstract
Understanding network workload through the characterization of network flows, being essential for assisting network management tasks, can benefit largely from traffic sampling as long as an accurate snapshot of network behavior is captured. This paper is devoted to evaluate the real applicability of using sampling to support flow analysis. Considering both classical and emerging sampling techniques, a comparative performance study is carried out to assess the accuracy of estimating flow parameters through sampling. After identifying the main building blocks of sampled-based measurements, a sampling framework has been implemented to provide a versatile and fair platform for carrying out the testing and comparison process. Through an encompassing coverage of representative sampling techniques, the present study aims to provide useful insights regarding the use of sampling in traffic flow analysis.
2015
Autores
Teixeira, JF; Couto, M;
Publicação
PROGRESS IN ARTIFICIAL INTELLIGENCE
Abstract
Text Mining has opened a vast array of possibilities concerning automatic information retrieval from large amounts of text documents. A variety of themes and types of documents can be easily analyzed. More complex features such as those used in Forensic Linguistics can gather deeper understanding from the documents, making possible performing difficult tasks such as author identification. In this work we explore the capabilities of simpler Text Mining approaches to author identification of unstructured documents, in particular the ability to distinguish poetic works from two of Fernando Pessoas' heteronyms: 'Alvaro de Campos and Ricardo Reis. Several processing options were tested and accuracies of 97% were reached, which encourage further developments.
2015
Autores
Ferreira, HL; Gibescu, M; Stankova, K; Kling, WL; Lopes, JP;
Publicação
2015 12TH INTERNATIONAL CONFERENCE ON THE EUROPEAN ENERGY MARKET (EEM)
Abstract
This paper deals with integrating energy storage systems (ESS) into existing electricity markets. We explain why ESS increase flexibility of power systems and energy markets and why more flexible systems and markets are desirable, particularly in a context of high integration of variable renewable energy sources (RES). The Dutch electricity markets are introduced as the case studies. As opposing to the existing literature, we focus on implementation of a dual technology ESS, which we believe is more beneficial than a single ESS. To show this, we introduce an optimal control model, in which the goal is to maximize the revenues of the dual technology energy storage system applied into two different energy markets, assuming the selling and buying electricity prices are exogenous. Subsequently, we introduce our model, using a simple strategy and present its results, showing the impact of the devices nominal rating on the potential revenues.
2015
Autores
Mickulicz, ND; Martins, R; Narasimhan, P; Gandhi, R;
Publicação
First IEEE International Conference on Big Data Computing Service and Applications, BigDataService 2015, Redwood City, CA, USA, March 30 - April 2, 2015
Abstract
Collections of time-series data appear in a wide variety of contexts. To gain insight into the underlying phenomenon (that the data represents), one must analyze the time-series data. Analysis can quickly become challenging for very large data (~terabytes or more) sets, and it may be infeasible to scan the entire data-set on each query due to time limits or resource constraints. To avoid this problem, one might pre-compute partial results by scanning the data-set (usually as the data arrives). However, for complex queries, where the value of a new data record depends on all of the data previously seen, this might be infeasible because incorporating a large amount of historical data into a query requires a large amount of storage. We present an approach to performing complex queries over very large data-sets in a manner that is (i) practical, meaning that a query does not require a scan of the entire data-set, and (ii) fixed-cost, meaning that the amount of storage required only depends on the time-range spanned by the entire data-set (and not the size of the data-set itself). We evaluate our approach with three different data-sets: (i) a 4-year commercial analytics data-set from a production content-delivery platform with over 15 million mobile users, (ii) an 18-year data-set from the Linux-kernel commit-history, and (iii) an 8-day data-set from Common Crawl HTTP logs. Our evaluation demonstrates the feasibility and practicality of our approach for a diverse set of complex queries on a diverse set of very large data-sets. © 2015 IEEE.
2015
Autores
Rivas, MA; Pirinen, M; Conrad, DF; Lek, M; Tsang, EK; Karczewski, KJ; Maller, JB; Kukurba, KR; DeLuca, DS; Fromer, M; Ferreira, PG; Smith, KS; Zhang, R; Zhao, F; Banks, E; Poplin, R; Ruderfer, DM; Purcell, SM; Tukiainen, T; Minikel, EV; Stenson, PD; Cooper, DN; Huang, KH; Sullivan, TJ; Nedzel, J; Bustamante, CD; Li, JB; Daly, MJ; Guigo, R; Donnelly, P; Ardlie, K; Sammeth, M; Dermitzakis, ET; McCarthy, MI; Montgomery, SB; Lappalainen, T; MacArthur, DG; Segre, AV; Young, TR; Gelfand, ET; Trowbridge, CA; Ward, LD; Kheradpour, P; Iriarte, B; Meng, Y; Palmer, CD; Esko, T; Winckler, W; Hirschhorn, J; Kellis, M; Getz, G; Shablin, AA; Li, G; Zhou, Y; Nobel, AB; Rusyn, I; Wright, FA; Battle, A; Mostafavi, S; Mele, M; Reverter, F; Goldmann, J; Koller, D; Gamazon, ER; Im, HK; Konkashbaev, A; Nicolae, DL; Cox, NJ; Flutre, T; Wen, X; Stephens, M; Pritchard, JK; Tu, Z; Zhang, B; Huang, T; Long, Q; Lin, L; Yang, J; Zhu, J; Liu, J; Brown, A; Mestichelli, B; Tidwell, D; Lo, E; Salvatore, M; Shad, S; Thomas, JA; Lonsdale, JT; Choi, RC; Karasik, E; Ramsey, K; Moser, MT; Foster, BA; Gillard, BM; Syron, J; Fleming, J; Magazine, H; Hasz, R; Walters, GD; Bridge, JP; Miklos, M; Sullivan, S; Barker, LK; Traino, H; Mosavel, M; Siminoff, LA; Valley, DR; Rohrer, DC; Jewel, S; Branton, P; Sobin, LH; Barcus, M; Qi, L; Hariharan, P; Wu, S; Tabor, D; Shive, C; Smith, AM; Buia, SA; Undale, AH; Robinson, KL; Roche, N; Valentino, KM; Britton, A; Burges, R; Bradbury, D; Hambright, KW; Seleski, J; Korzeniewski, GE; Erickson, K; Marcus, Y; Tejada, J; Taherian, M; Lu, C; Robles, BE; Basile, M; Mash, DC; Volpi, S; Struewing, JP; Temple, GF; Boyer, J; Colantuoni, D; Little, R; Koester, S; Carithers, LJ; Moore, HM; Guan, P; Compton, C; Sawyer, SJ; Demchok, JP; Vaught, JB; Rabiner, CA; Lockhart, NC; Friedlander, MR; 't Hoen, PAC; Monlong, J; Gonzalez-Porta, M; Kurbatova, N; Griebel, T; Barann, M; Wieland, T; Greger, L; van Iterson, M; Almlof, J; Ribeca, P; Pulyakhina, I; Esser, D; Giger, T; Tikhonov, A; Sultan, M; Bertier, G; Lizano, E; Buermans, HPJ; Padioleau, I; Schwarzmayr, T; Karlberg, O; Ongen, H; Kilpinen, H; Beltran, S; Gut, M; Kahlem, K; Amstislavskiy, V; Stegle, O; Flicek, P; Strom, TM; Lehrach, H; Schreiber, S; Sudbrak, R; Carracedo, A; Antonarakis, SE; Hasler, R; Syvanen, A; van Ommen, G; Brazma, A; Meitinger, T; Rosenstiel, P; Gut, IG; Estivill, X; The GTEx Consortium,; The Geuvadis Consortium,;
Publicação
Science
Abstract
Accurate prediction of the functional effect of genetic variation is critical for clinical genome interpretation.We systematically characterized the transcriptome effects of protein-truncating variants, a class of variants expected to have profound effects on gene function, using data from the Genotype-Tissue Expression (GTEx) and Geuvadis projects. We quantitated tissue-specific and positional effects on nonsense-mediated transcript decay and present an improved predictive model for this decay. We directly measured the effect of variants both proximal and distal to splice junctions. Furthermore, we found that robustness to heterozygous gene inactivation is not due to dosage compensation. Our results illustrate the value of transcriptome data in the functional interpretation of genetic variants.
2015
Autores
Santos, DM; Rodrigues, SSP; Oliveira, BMPM; de Almeida, MDV;
Publicação
PUBLIC HEALTH NUTRITION
Abstract
Objective To identify dietary availability and its time trends in elderly Portuguese households. Design A set of four cross-sectional studies based on the Household Budget Surveys was used. The dietary data were described using the daily per capita availability of food and beverages, energy and selected nutrients (macronutrients, different lipid fractions and simple sugars). Differences between elderly household types and time trends were studied. Setting Portuguese Household Budget Survey data from 1989/1990, 1994/1995, 2000/2001 and 2005/2006. Subjects Households with members aged 65 years were selected and categorized as solitary elderly female, solitary elderly male or couple (composed of one elderly female and one elderly male). Results While cereals, fats/oils, potatoes and sugar/sugar products decreased, an increase occurred in milk/milk products, fruits, bottled water, fruit/vegetable juices and soft drinks (P<005). The highest values for foods and beverages were mostly found in couples, while the lowest ones were from solitary males. Exceptions were observed for cereals, eggs, milk/milk products, vegetables, fruits and non-alcoholic beverages, higher in solitary females; and for sugar/sugar products and alcoholic beverages, higher in solitary males. Over time, total energy and carbohydrates decreased while proteins and saturated fatty acids increased (P<0001). Lipids increased in solitary males and couples (P<005). Simple sugars increased in solitary males but decreased in solitary females and couples (P<005). Conclusions The increases in fruits and vegetables in solitary females accord with a healthier food pattern, but overall imbalances in the macronutrient profile for all elderly households may imply a decreasing diet quality.
The access to the final selection minute is only available to applicants.
Please check the confirmation e-mail of your application to obtain the access code.