Reviving Dormant Ties In An Online Social Network Experiment,
2013
Singapore Management University
Reviving Dormant Ties In An Online Social Network Experiment, Ee Peng Lim, Denzil Correa, David Lo, Michael Finegold, Feida Zhu
Research Collection School Of Computing and Information Systems
Social network users connect and interact with one another to fulfil different kinds of social and information needs. When interaction ceases between two users, we say that their tie becomes dormant. While there are different underlying reasons of dormant ties, it is important to find means to revive such ties so as to maintain vibrancy in the relationships. In this work, we thus focus on designing an online experiment to evaluate the effectiveness of personalized social messages to revive dormant ties. The experiment carefully selects users with dormant ties so that no user gets mixed treatments and be affected by …
Mkboost: A Framework Of Multiple Kernel Boosting,
2013
Nanyang Technological University
Mkboost: A Framework Of Multiple Kernel Boosting, Hao Xia, Steven C. H. Hoi
Research Collection School Of Computing and Information Systems
Multiple kernel learning (MKL) is a promising family of machine learning algorithms using multiple kernel functions for various challenging data mining tasks. Conventional MKL methods often formulate the problem as an optimization task of learning the optimal combinations of both kernels and classifiers, which usually results in some forms of challenging optimization tasks that are often difficult to be solved. Different from the existing MKL methods, in this paper, we investigate a boosting framework of MKL for classification tasks, i.e., we adopt boosting to solve a variant of MKL problem, which avoids solving the complicated optimization tasks. Specifically, we present …
Shortlisting Top-K Assignments,
2013
Singapore Management University
Shortlisting Top-K Assignments, Yimin Lin, Kyriakos Mouratidis
Research Collection School Of Computing and Information Systems
In this paper we identify a novel query type, the top-K assignment query (αTop-K). Consider a set of objects and a set of suppliers, where each object must be assigned to one supplier. Assume that there is a cost associated with every object-supplier pair. If we allocate each object to the server with the smallest cost (for the specific object), the derived overall assignment will have the minimum total cost. In many scenarios, however, runner-up assignments may be required too, like for example when a decision maker needs to make additional considerations, not captured by individual object-supplier costs. In this …
Active Learning With Expert Advice,
2013
Nanyang Technological University
Active Learning With Expert Advice, Peilin Zhao, Steven C. H. Hoi, Jinfeng Zhuang
Research Collection School Of Computing and Information Systems
Conventional learning with expert advice methods assumes a learner is always receiving the outcome (e.g., class labels) of every incoming training instance at the end of each trial. In real applications, acquiring the outcome from oracle can be costly or time consuming. In this paper, we address a new problem of active learning with expert advice, where the outcome of an instance is disclosed only when it is requested by the online learner. Our goal is to learn an accurate prediction model by asking the oracle the number of questions as small as possible. To address this challenge, we propose …
Using Correlated Subset Structure For Compressive Sensing Recovery,
2013
Alcatel-Lucent
Using Correlated Subset Structure For Compressive Sensing Recovery, Atul Divekar, Deanna Needell
CMC Faculty Publications and Research
Compressive sensing is a methodology for the reconstruction of sparse or compressible signals using far fewer samples than required by the Nyquist criterion. However, many of the results in compressive sensing concern random sampling matrices such as Gaussian and Bernoulli matrices. In common physically feasible signal acquisition and reconstruction scenarios such as super-resolution of images, the sensing matrix has a non-random structure with highly correlated columns. Here we present a compressive sensing recovery algorithm that exploits this correlation structure. We provide algorithmic justification as well as empirical comparisons.
Brovine: Mammary Gland Gene Database,
2013
California Polytechnic State University - San Luis Obispo
Brovine: Mammary Gland Gene Database, Therin C. Irwin
Computer Science and Software Engineering
Brovine is used by the Animal Science department at Cal Poly to catalog and analyze genetic information. Brovine, or the Mammary Gland Gene Database, is a system used to store and categorize genetic information which is gathered through experimentation and through TESS, a web application that lets users search through catalogs of similar genetic information. This document describes the purpose, use, and maintenance of Brovine.
Pin: Measuring Asymmetric Information In Financial Markets With R,
2013
International Entrepreneurship Academy, Università di Bologna
Pin: Measuring Asymmetric Information In Financial Markets With R, Paolo Zagaglia
The R Journal
The package PIN computes a measure of asymmetric information in financial markets, the so-called probability of informed trading. This is obtained from a sequential trade model and is used to study the determinants of an asset price. Since the probability of informed trading depends on the number of buy- and sell-initiated trades during a trading day, this paper discusses the entire modelling cycle, from data handling to the computation of the probability of informed trading and the estimation of parameters for the underlying theoretical model.
Statistical Software From A Blind Person's Perspective,
2013
Massey University
Statistical Software From A Blind Person's Perspective, A. Jonathan R. Godfrey
The R Journal
Blind people have experienced access issues to many software applications since the advent of the Windows operating system; statistical software has proven to follow the rule and not be an exception. The ability to use R within minutes of download with next to no adaptation has opened doors for accessible production of statistical analyses for this author (himself blind) and blind students around the world. This article shows how little is required to make R the most accessible statistical software available today. There is any number of ramifications that this opportunity creates for blind students, especially in terms of their …
Multiple Factor Analysis For Contingency Tables In The Factominer Package,
2013
Transverse group for research in primary care
Multiple Factor Analysis For Contingency Tables In The Factominer Package, Belchin Kostov, Mónica Bécue-Bertaut, François Husson
The R Journal
We present multiple factor analysis for contingency tables (MFACT) and its implementation in the FactoMineR package. This method, through an option of the MFA function, allows us to deal with multiple contingency or frequency tables, in addition to the categorical and quantitative multiple tables already considered in previous versions of the package. Thanks to this revised function, either a multiple contingency table or a mixed multiple table integrating quantitative, categorical and frequency data can be tackled.
The FactoMineR package (Lê et al., 2008; Husson et al., 2011) offers the most commonly used principal component methods: principal component analysis (PCA), correspondence …
Generalized Simulated Annealing For Global Optimization: The Gensa Package,
2013
Philip Morris International R&D
Generalized Simulated Annealing For Global Optimization: The Gensa Package, Yang Xiang, Sylvain Gubian, Brain Suomela, Julia Hoeng
The R Journal
Many problems in statistics, finance, biology, pharmacology, physics, mathematics, economics, and chemistry involve determination of the global minimum of multidimensional functions. R packages for different stochastic methods such as genetic algorithms and differential evolution have been developed and successfully used in the R community. Based on Tsallis statistics, the R package GenSA was developed for generalized simulated annealing to process complicated non-linear objective functions with a large number of local minima. In this paper we provide a brief introduction to the R package and demonstrate its utility by solving a non-convex portfolio optimization problem in finance and the Thomson problem …
R Foundation News,
2013
WUWirtschaftsuniversität Wien, Austria
R Foundation News, Kurt Hornik
The R Journal
New Benefactors
Quartz, Bio, Switzerland
New supporting Institutions
Institute for Geoinformatics, Westfälische Wilhelms-Universität Münster, Germany
News From The Bioconductor Project,
2013
University of Nebraska - Lincoln
News From The Bioconductor Project, Bioconductor Team
The R Journal
Bioconductor 2.12 was released on 3 October 2012. It is compatible with R 3.0.1, and consists of 671 software packages and more than 675 up-to-date annotation packages. The release includes 65 new software packages, and enhancements to many others. Descriptions of new packages and updated NEWS files provided by current package maintainers are at http://bioconductor.org/news/bioc_2_12_release/.
Conference Review: The 6th Chinese R Conference,
2013
Renmin University of China
Conference Review: The 6th Chinese R Conference, Jing Leng, Jingjing Guan
The R Journal
The 6th Chinese R Conference (Beijing session) was held in the Sinology Pavilion of Renmin University of China (RUC), Beijing, from May 18th to 19th, 2013. The conference was orga nized by the “Capital of Statistics” (COS, http://cos.name), an online statistical community in China. It was sponsored and co-organized by the Center for Applied Statistics of RUC, the School of Statistics of RUC, and the Business Intelligence Research Center of Peking University
Possible Directions For Improving Dependency Versioning In R,
2013
University of California, Los Angeles
Possible Directions For Improving Dependency Versioning In R, Jeroen Ooms
The R Journal
One of the most powerful features of R is its infrastructure for contributed code. The built-in package manager and complementary repositories provide a great system for development and exchange of code, and have played an important role in the growth of the platform towards the de-facto standard in statistical computing that it is today. However, the number of packages on CRAN and other repositories has increased beyond what might have been foreseen, and is revealing some limitations of the current design. One such problem is the general lack of dependency versioning in the infrastructure. This paper explores this problem in …
Beadarrayfilter: An R Package To Filter Beads,
2013
Katholieke Universiteit Leuven
Beadarrayfilter: An R Package To Filter Beads, Anyiawung Chiara Forcheh, Geert Verbeke, Adetayo Kasim, Dan Lin, Ziv Shkedy, Willem Talloen, Hinrich W.H. Göhlmann, Lieven Clement
The R Journal
Microarrays enable the expression levels of thousands of genes to be measured simultaneously. However, only a small fraction of these genes are expected to be expressed under different experimental conditions. Nowadays, filtering has been introduced as a step in the microarray pre-processing pipeline. Gene filtering aims at reducing the dimensionality of data by filtering redundant features prior to the actual statistical analysis. Previous filtering methods focus on the Affymetrix platform and can not be easily ported to the Illumina platform. As such, we developed a filtering method for Illumina bead arrays. We developed an R package, beadarrayFilter, to implement …
Ggmap: Spatial Visualization With Ggplot2,
2013
Baylor University
Ggmap: Spatial Visualization With Ggplot2, David Kahle, Hadley Wickham
The R Journal
In spatial statistics the ability to visualize data and models super imposed with their basic social landmarks and geographic context is in valuable. ggmap is a new tool which enables such visualization by combining the spatial information of static maps from Google Maps,Open Street Map, Stamen Maps or Cloud Made Maps with the layered grammar of graphics implementation of ggplot2. In addition, several new utility functions are introduced which allow the user to access the Google Geocoding, Distance Matrix, and Directions APIs. The result is an easy, consistent and modular frame work for spatial graphics with several convenient tools …
Let Graphics Tell The Story: Datasets In R,
2013
Augsburg University
Let Graphics Tell The Story: Datasets In R, Antony Unwin, Heike Hofmann, Dianne Cook
The R Journal
Graphics are good for showing the information in datasets and for complementing modelling. Sometimes graphics show information models miss, sometimes graphics help to make model results more understandable, and sometimes models show whether information from graphics has statistical support or not. It is the interplay of the two approaches that is valuable. Graphics could be used a lot more in R examples and we explore this idea with some datasets available in R packages.
Stellar: A Package To Manage Stellar Evolution Tracks And Isochrones,
2013
Università di Pisa
Stellar: A Package To Manage Stellar Evolution Tracks And Isochrones, Matteo Dell’Omodarme, Giada Valle
The R Journal
We present the R package stellaR, which is designed to access and manipulate publicly available stellar evolutionary tracks and isochrones from the Pisa low-mass database. The procedures for extracting important stages in the evolution of a star from the database, for constructing isochrones from stellar tracks and for interpolating among tracks are discussed and demonstrated.
Due to the advance in the instrumentation, nowadays astronomers can deal with a huge amount of high-quality observational data. In the last decade impressive improvements of spectroscopic and photometric observational capabilities made available data which stimulated the research in the globular clusters field. The theoretical …
Osmar: Openstreetmap And R,
2013
Ludwig-Maximilians-Universität München
Osmar: Openstreetmap And R, Manuel J. Eugster, Thomas Schlesinger
The R Journal
OpenStreetMap provides freely accessible and editable geographic data. The osmar package smoothly integrates the OpenStreetMap project into the R ecosystem. The osmar package provides infrastructure to access OpenStreetMap data from different sources, to enable working with the OSM data in the familiar R idiom, and to convert the data into objects based on classes provided by existing Rpackages. This paper explains the package’s concept and shows how to use it. As an application we present a simple navigation device
Hypothesis Tests For Multivariate Linear Models Using The Car Package,
2013
McMaster University
Hypothesis Tests For Multivariate Linear Models Using The Car Package, John Fox, Michael Friendly, Sanford Weisberg
The R Journal
The multivariate linear model can be fit with the lm function in R, where the left-hand side of the model comprises a matrix of response variables, and the right-hand side is specified exactly as for a univariate linear model (i.e., with a single response variable). This paper explains how to use the Anova and linearHypothesis functions in the car package to perform convenient hypothesis tests for parameters in multivariate linear models, including models for repeated-measures data.
