Zoib: An R Package For Bayesian Inference For Beta Regression And Zero/One Inflated Beta Regression,
2015
University of Notre Dame
Zoib: An R Package For Bayesian Inference For Beta Regression And Zero/One Inflated Beta Regression, Fang Liu, Yunchuan Kong
The R Journal
The beta distribution is a versatile function that accommodates a broad range of probability distribution shapes. Beta regression based on the beta distribution can be used to model a response variable y that takes values in open unit interval (0,1). Zero/one inflated beta (ZOIB) regression models can be applied when y takes values from closed unit interval [0,1]. The ZOIB model is based a piecewise distribution that accounts for the probability mass at 0 and 1, in addition to the probability density within (0,1). This paper introduces an R package– zoib that provides Bayesian inferences for a class of ZOIB …
Fitting Conditional And Simultaneous Autoregressive Spatial Models In Hglm,
2015
Dalarna University
Fitting Conditional And Simultaneous Autoregressive Spatial Models In Hglm, Moudud Alam, Lars Rönnegård, Xia Shen
The R Journal
We present a new version ( 2.0) of the hglm package for fitting hierarchical generalized linear models (HGLMs) with spatially correlated random effects. CAR() and SAR() families for con ditional and simultaneous autoregressive random effects were implemented. Eigen decomposition of the matrix describing the spatial structure (e.g., the neighborhood matrix) was used to transform the CAR/SARrandomeffects into an independent, but heteroscedastic, Gaussian random effect. A linear predictor is fitted for the random effect variance to estimate the parameters in the CAR and SAR models. This gives a computationally efficient algorithm for moderately sized problems.
Vsurf: An R Package For Variable Selection Using Random Forests,
2015
University of Bordeaux
Vsurf: An R Package For Variable Selection Using Random Forests, Robin Genuer, Jean-Michel Poggi, Christine Tuleau-Malot
The R Journal
This paper describes the R package VSURF. Based on random forests, and for both regression and classification problems, it returns two subsets of variables. The first is a subset of important variables including some redundancy which can be relevant for interpretation, and the second one is a smaller subset corresponding to a model trying to avoid redundancy focusing more closely on the prediction objective. The two-stage strategy is based on a preliminary ranking of the explanatory variables using the random forests permutation-based score of importance and proceeds using a stepwise forward strategy for variable introduction. The two proposals can …
Generalized Hermite Distribution Modelling With The R Package Hermite,
2015
Universitat Pompeu Fabra, Unitat de Bioestadística
Generalized Hermite Distribution Modelling With The R Package Hermite, David Moriña, Manuel Higueras, Pedro Puig, María Oliveira
The R Journal
The Generalized Hermite distribution (and the Hermite distribution as a particular case) is often used for fitting count data in the presence of over-dispersion or multimodality. Despite this, to our knowledge, no standard software packages have implemented specific functions to compute basic probabilities and make simple statistical inference based on these distributions. We present here a set of computational tools that allows the user to face these difficulties by modelling with the Generalized Hermite distribution using the R package hermite. The package can also be used to generate random deviates from a Generalized Hermite distribution and to use basic …
Working With Multilabel Datasets In R: The Mldr Package,
2015
University of Granada
Working With Multilabel Datasets In R: The Mldr Package, Francisco Charte, David Charte
The R Journal
Most classification algorithms deal with datasets which have a set of input features, the variables to be used as predictors, and only one output class, the variable to be predicted. However, in late years many scenarios in which the classifier has to work with several outputs have come to life. Automatic labeling of text documents, image annotation or protein classification are among them. Multilabel datasets are the product of these new needs, and they have many specific traits. The mldr package allows the user to load datasets of this kind, obtain their characteristics, produce specialized plots, and manipulate them. The …
Numerical Evaluation Of The Gauss Hypergeometric Function With The Hypergeo Package,
2015
Auckland University of Technology
Numerical Evaluation Of The Gauss Hypergeometric Function With The Hypergeo Package, Robin K. S. Hankin
The R Journal
This paper introduces the hypergeo package of R routines for numerical calculation of hypergeometric functions. The package is focussed on efficient and accurate evaluation of the Gauss hypergeometric function over the whole of the complex plane within the constraints of fixed-precision arithmetic. The hypergeometric series is convergent only within the unit circle, so analytic continuation must be used to define the function outside the unit circle. This short document outlines the numerical and conceptual methods used in the package; and justifies the package philosophy, which is to maintain transparent and verifiable links between the software and Abramowitz and Stegun (1965). …
Treeclust: An R Package For Tree-Based Clustering Dissimilarities,
2015
Naval Postgraduate School
Treeclust: An R Package For Tree-Based Clustering Dissimilarities, Samuel E. Buttrey, Lyn R. Whitaker
The R Journal
This paper describes treeClust, an R package that produces dissimilarities useful for clustering. These dissimilarities arise from a set of classification or regression trees, one with each variable in the data acting in turn as a the response, and all others as predictors. This use of trees produces dissimilarities that are insensitive to scaling, benefit from automatic variable selection, and appear to perform well. The software allows a number of options to be set, affecting the set of objects returned in the call; the user can also specify a clustering algorithm and, optionally, return only the clustering vector. The …
Apc: An R Package For Age-Period-Cohort Analysis,
2015
Nuffield College, University of Oxford, INET
Apc: An R Package For Age-Period-Cohort Analysis, Bent Nielsen
The R Journal
The apc package includes functions for age-period-cohort analysis based on the canonical parametrisation of Kuang et al. (2008a). The package includes functions for organizing the data, descriptive plots, a deviance table, estimation of (sub-models of) the age-period-cohort model, a plot for specification testing, plots of estimated parameters, and sub-sample analysis.
News From The Bioconductor Project,
2015
University of Nebraska - Lincoln
News From The Bioconductor Project, The Bioconductor Team
The R Journal
The Bioconductor project provides tools for the analysis and comprehension of high throughput genomic data. The 1104 software packages available in Bioconductor can be viewed at http://bioconductor.org/packages/. Navigate packages using ‘biocViews’ terms and title search. Each package has an html page with a description, links to vignettes, reference manuals, and usage statistics. Start using Bioconductor version 3.2 by installing R 3.2.3 and evaluating the commands
Editorial,
2015
R Journal
Editorial, Bettina Grün
The R Journal
On behalf of the editorial board, I am pleased to publish Volume 7, Issue 2 of the R Journal. This issue contains 20 contributed research articles and several contributions to the News and Notes section.
Code Profiling In R: A Review Of Existing Methods And An Introduction To Package Guiprofiler,
2015
Universidad de Navarra
Code Profiling In R: A Review Of Existing Methods And An Introduction To Package Guiprofiler, Angel Rubio, Fernando De Villar
The R Journal
Code analysis tools are crucial to understand program behavior. Profile tools use the results of time measurements in the execution of a program to gain this understanding and thus help in the optimization of the code. In this paper, we review the different available packages to profile R code and show the advantages and disadvantages of each of them. In additon, we present GUIProfiler, a package that fulfills some unmet needs
Package GUIProfiler generates an HTML report with the timing for each code line and the relationships between different functions. This package mimics the behavior of the MATLAB profiler. …
Quantifquantile: An R Package For Performing Quantile Regression Through Optimal Quantization,
2015
Université Libre de Bruxelles, Université de Bordeaux, Inria Bordeaux Sud-Ouest
Quantifquantile: An R Package For Performing Quantile Regression Through Optimal Quantization, Isabelle Charlier, Davy Paindaveine, Jérôme Saracco
The R Journal
In quantile regression, various quantiles of a response variable Y are modelled as functions of covariates (rather than its mean). An important application is the construction of reference curves/surfaces and conditional prediction intervals for Y. Recently, a nonparametric quantile regression method based on the concept of optimal quantization was proposed. This method competes very well with k-nearest neighbor, kernel, and spline methods. In this paper, we describe an R package, called QuantifQuantile, that allows to perform quantization-based quantile regression. We describe the various functions of the package and provide examples.
Changes In R,
2015
University of Nebraska - Lincoln
The R Consortium And The R Foundation,
2015
The R Foundation
The R Consortium And The R Foundation, Martyn Plummer
The R Journal
The R Consortium was announced at the useR! 2015 conference in Aalborg, Denmark on 30 June. It is a non-profit organization set up to provide infrastructure for the R community. The purpose of this article is to explain some of the background to the setting up of the Consortium and how it interacts with the R Foundation.
An Immersive Telepresence System Using Rgb-D Sensors And Head-Mounted Display,
2015
University of Dayton
An Immersive Telepresence System Using Rgb-D Sensors And Head-Mounted Display, Xinzhong Lu, Ju Shen, Saverio Perugini, Jianjun Yang
Computer Science Faculty Publications
We present a tele-immersive system that enables people to interact with each other in a virtual world using body gestures in addition to verbal communication. Beyond the obvious applications, including general online conversations and gaming, we hypothesize that our proposed system would be particularly beneficial to education by offering rich visual contents and interactivity. One distinct feature is the integration of egocentric pose recognition that allows participants to use their gestures to demonstrate and manipulate virtual objects simultaneously. This functionality enables the instructor to effectively and efficiently explain and illustrate complex concepts or sophisticated problems in an intuitive manner. The …
Transforming C Openmp Programs For Verification In Civl,
2015
University of Nebraska-Lincoln
Transforming C Openmp Programs For Verification In Civl, Michael Rogers
School of Computing: Dissertations, Theses, and Student Research
There are numerous way to express parallelism which can make it challenging for developers to verify these programs. Many tools only target a single dialect but the Concurrency Intermediate Verification Language (CIVL) targets MPI, Pthreads, and CUDA. CIVL provides a general concurrency model that can represent pro- grams in a variety of concurrency dialects. CIVL includes a front-end that support all of the dialects mentioned above. The back-end is a verifier that uses model checking and symbolic execution to check standard properties.
In this thesis, we have designed and implemented a transformer that will take C OpenMP programs and transform …
Detecting Broken Pointcuts Using Structural Commonality And Degree Of Interest,
2015
CUNY Hunter College
Detecting Broken Pointcuts Using Structural Commonality And Degree Of Interest, Raffi Khatchadourian, Awais Rashd, Hidehiko Masuhara, Takuya Watanabe
Publications and Research
Pointcut fragility is a well-documented problem in Aspect-Oriented Programming; changes to the base code can lead to join points incorrectly falling in or out of the scope of pointcuts. Deciding which pointcuts have broken due to base-code changes is daunting, especially in large and complex systems. We present an automated approach that recommends pointcuts that are likely to require modification due to a certain base-code change and ones that do not. Our hypothesis is that join points selected by a pointcut exhibit common structural characteristics. Patterns describing such commonalities recommend pointcuts that have potentially broken to the developer. The approach …
Codehow: Effective Code Search Based On Api Understanding And Extended Boolean Model (E),
2015
Shanghai Jiaotong University
Codehow: Effective Code Search Based On Api Understanding And Extended Boolean Model (E), Fei Lv, Jian-Guang Lou, Shaowei Wang, Dongmei Zhang, Jainjun Zhao
Research Collection School Of Computing and Information Systems
Over the years of software development, a vast amount of source code has been accumulated. Many code search tools were proposed to help programmers reuse previously-written code by performing free-text queries over a large-scale codebase. Our experience shows that the accuracy of these code search tools are often unsatisfactory. One major reason is that existing tools lack of query understanding ability. In this paper, we propose CodeHow, a code search technique that can recognize potential APIs a user query refers to. Having understood the potentially relevant APIs, CodeHow expands the query with the APIs and performs code retrieval by applying …
Challenges In Analyzing Software Documentation In Portuguese,
2015
Singapore Management University
Challenges In Analyzing Software Documentation In Portuguese, Christoph Treude, Carlos A. Prolo, Fernando Figueira Filho
Research Collection School Of Computing and Information Systems
Many tools that automatically analyze, summarize, or transform software artifacts rely on natural language processing tooling for the interpretation of natural language text produced by software developers, such as documentation, code comments, commit messages, or bug reports. Processing natural language text produced by software developers is challenging because of unique characteristics not found in other texts, such as the presence of code terms and the systematic use of incomplete sentences. In addition, texts produced by Portuguese-speaking developers mix languages since many keywords and programming concepts are referred to by their English name. In this paper, we provide empirical insights into …
Leveraging Synergy Between Database And Programming Language Courses,
2015
DePauw University
Leveraging Synergy Between Database And Programming Language Courses, Brian T. Howard
Computer Science Faculty publications
Undergraduate courses in database systems and programming languages are frequently taught without much overlap. This paper argues that there is a substantial benefit to emphasizing some areas of commonality, both old and new, between the two subjects. Examples of cross-fertilization that may be used to enhance one of both of the courses include query language design and implementation, object-relational mapping, transactional memory, and various aspects of the recent "NoSQL" movement.
