Heteroscedastic Censored And Truncated Regression With Crch,
2016
Universität Innsbruck
Heteroscedastic Censored And Truncated Regression With Crch, Jakob W. Messner, Georg J. Mayr, Achim Zeileis
The R Journal
The crch package provides functions for maximum likelihood estimation of censored or truncated regression models with conditional heteroscedasticity along with suitable standard methods to summarize the fitted models and compute predictions, residuals, etc. The supported distributions include left- or right-censored or truncated Gaussian, logistic, or student-t distributions with potentially different sets of regressors for modeling the conditional location and scale. The models and their R implementation are introduced and illustrated by numerical weather prediction tasks using precipitation data for Innsbruck (Austria).
Exploring Interaction Effects In Two-Factor Studies Using The Hiddenf Package In R.,
2016
Virginia Tech Department of Statistics
Exploring Interaction Effects In Two-Factor Studies Using The Hiddenf Package In R., Christopher Franck, Jason A. Osborne
The R Journal
In crossed, two-factor studies with one observation per factor-level combination, interaction effects between factors can be hard to detect and can make the choice of a suitable statistical model difficult. This article describes hiddenf, an R package that enables users to quantify and characterize a certain form of interaction in two-factor layouts. When effects of one factor (a) fall into two groups depending on the level of another factor, and (b) are constant within these groups, the interaction pattern is deemed "hidden additivity" because within groups, the effects of the two factors are additive, while between groups the factors …
An Interactive Survey Application For Validating Social Network Analysis Techniques,
2016
Wladimirstraße 3
An Interactive Survey Application For Validating Social Network Analysis Techniques, Mitchell Joblin, Wolfgang Mauerer
The R Journal
Social network analysis is extremely well supported by the R community and is routinely used for studying the relationships between people engaged in collaborative activities. While there has been rapid development of new approaches and metrics in this field, the challenging question of validity (how well insights derived from social networks agree with reality) is often difficult to address. We propose the use of several R packages to generate interactive surveys that are specifically well suited for validating social network analyses. Using our web-based survey application, we were able to validate the results of applying community-detection algorithms to infer the …
Stylometry With R: A Package For Computational Text Analysis,
2016
Polish Academy of Sciences
Stylometry With R: A Package For Computational Text Analysis, Maciej Eder, Jan Rybicki, Mike Kestemont
The R Journal
This software paper describes ‘Stylometry with R’ (stylo), a flexible R package for the high level analysis of writing style in stylometry. Stylometry (computational stylistics) is concerned with the quantitative study of writing style, e.g. authorship verification, an application which has considerable potential in forensic contexts, as well as historical research. In this paper we introduce the possibilities of stylo for computational text analysis, via a number of dummy case studies from English and French literature. We demonstrate how the package is particularly useful in the exploratory statistical analysis of texts, e.g. with respect to authorial writing style. …
Quickpsy: An R Package To Fit Psychometric Functions For Multiple Groups,
2016
Universitat de Barcelona
Quickpsy: An R Package To Fit Psychometric Functions For Multiple Groups, Daniel Linares, Joan López-Moliner
The R Journal
quickpsy is a package to parametrically fit psychometric functions. In comparison with previous R packages, quickpsy was built to easily fit and plot data for multiple groups. Here, we describe the standard parametric model used to fit psychometric functions and the standard estimation of its parameters using maximum likelihood. We also provide examples of usage of quickpsy, including how allowing the lapse rate to vary can sometimes eliminate the bias in parameter estimation, but not in general. Finally, we describe some implementation details, such as how to avoid the problems associated to round-off errors in the maximisation of the …
The R Journal (August 2016) 8(1): Complete Issue,
2016
University of Nebraska - Lincoln
The R Journal (August 2016) 8(1): Complete Issue, The R Foundation
The R Journal
Editorial, Michael Lawrence
Contributed Research Articles
metaplus: An R Package for the Analysis of Robust Meta-Analysis and Meta-Regression, Ken J. Beath
Gender Prediction Methods Based on First Names with genderizeR, Kamil Wais
Conditional Fractional Gaussian Fields with the Package FieldSim, Alexandre Brouste, Jacques Istas, and Sophie Lambert-Lacroix
rTableICC: An R Package for Random Generation of 22K and RC Contingency Tables, Haydar Demirhan
Maps, Coordinate Reference Systems and Visualising Geographic Data with mapmisc, Patrick E. Brown
Variable Clustering in High-Dimensional Linear Regression: The R Package clere, Loïc Yengo, Julien Jacques, Christophe Biernacki, and Mickael Canouil
Stylometry with R: A Package for …
Fwdselect: An R Package For Variable Selection In Regression Models,
2016
University of Minho
Fwdselect: An R Package For Variable Selection In Regression Models, Marta Sestelo, Nora M. Villanueva, Luis Meira-Machado, Javier Roca-Pardiñas
The R Journal
In multiple regression models, when there are a large number (p) of explanatory variables which may o rmay not be relevant for predicting the response, it is useful to be able to reduce the model. To this end, it is necessary to determine the best subset of q (q p) predictors which will establish the model with the best prediction capacity. FWDselect package introduces a new forward stepwise based selection procedure to select the best model in different regression frameworks (parametric or nonparametric). The developed methodology, which can be equally applied to linear models, generalized linear models or generalized additive …
Progenyclust: An R Package For Progeny Clustering,
2016
Rice University
Progenyclust: An R Package For Progeny Clustering, Chenyue W. Hu, Amina A. Qutub
The R Journal
Identifying the optimal number of clusters is a common problem faced by data scientists in various research fields and industry applications. Though many clustering evaluation techniques have been developed to solve this problem, the recently developed algorithm Progeny Clustering is a much faster alternative and one that is relevant to biomedical applications. In this paper, we introduce an R package progenyClust that implements and extends the original Progeny Clustering algorithm for evaluating clustering stability and identifying the optimal cluster number. We illustrate its applicability using two examples: a simulated test dataset for proof-of-concept, and a cell imaging dataset for demonstrating …
Using Decipher V2.0 To Analyze Big Biological Sequence Data In R,
2016
University of Wisconsin- Madison
Using Decipher V2.0 To Analyze Big Biological Sequence Data In R, Erik S. Wright
The R Journal
In recent years, the cost of DNA sequencing has decreased at a rate that has outpaced improvements in memory capacity. It is now common to collect or have access to many gigabytes of biological sequences. This has created an urgent need for approaches that analyze sequences in subsets without requiring all of the sequences to be loaded into memory at one time. It has also opened opportunities to improve the organization and accessibility of information acquired in sequencing projects. The DECIPHER package offers solutions to these problems by assisting in the curation of large sets of biological sequences stored in …
Variable Clustering In High-Dimensional Linear Regression: The R Package Clere,
2016
FR3508 European Genomics Institute of Diabetes
Variable Clustering In High-Dimensional Linear Regression: The R Package Clere, Loïc Yengo, Julien Jacques, Christophe Biernacki, Mickael Canouil
The R Journal
Dimension reduction is one of the biggest challenges in high-dimensional regression models. We recently introduced a new methodology based on variable clustering as a means to reduce dimensionality. We present here the R package clere that implements some refinements of this methodology. An overview of the package functionalities as well as examples to run an analysis are described. Numerical experiments on real data were performed to illustrate the good predictive performance of our parsimonious method compared to standard dimension reduction approaches.
Mclust 5: Clustering, Classification And Density Estimation Using Gaussian Finite Mixture Models,
2016
Università degli Studi di Perugia
Mclust 5: Clustering, Classification And Density Estimation Using Gaussian Finite Mixture Models, Luca Scrucca, Michael Fop, T Brendan Murphy, Adrian E. Raftery
The R Journal
Finite mixture models are being used increasingly to model a wide variety of random phenomena for clustering, classification and density estimation. mclust is a powerful and popular package which allows modelling of data as a Gaussian finite mixture with different covariance structures and different numbers of mixture components, for a variety of purposes of analysis. Recently, version 5 of the package has been made available on CRAN.This updated version adds new covariance structures, dimension reduction capabilities for visualisation, model selection criteria, initialisation strategies for the EM algorithm, and bootstrap-based inference, making it a full-featured R package for data analysis via …
Schemaonread: A Package For Schema-On-Read In R,
2016
Argonne National Laboratory
Schemaonread: A Package For Schema-On-Read In R, Michael J. North
The R Journal
SchemaOnRead is a CRAN package that provides an extensible mechanism for importing a wide range of file types into R as well as support for the emerging schema-on-read paradigm in R. The schema-on-read tools within the package include a single function call that recursively reads folders with text, comma separated value, raster image, R data, HDF5, NetCDF, spreadsheet, Weka, Epi Info, Pajek network, R network, HTML, SPSS, Systat, and Stata files. It also recursively reads folders (e.g., schemaOnRead("folder")), returning a nested list of the contained elements. The provided tools can be used as-is or easily customized to implement tool chains …
Keyplayer: An R Package For Locating Key Players In Social Networks,
2016
Indiana University
Keyplayer: An R Package For Locating Key Players In Social Networks, Weihua An, Yu-Hsin Liu
The R Journal
Interest in social network analysis has exploded in the past few years, partly thanks to the advancements in statistical methods and computing for network analysis. A wide range of the methods for network analysis is already covered by existent R packages. However, no comprehensive packages are available to calculate group centrality scores and to identify key players (i.e., those players who constitute the most central group) in a network. These functionalities are important because, for example, many social and health interventions rely on key players to facilitate the intervention. Identifying key players is challenging because players who are individually the …
Safegpu: Contract- And Library-Based Gpgpu For Object-Oriented Languages,
2016
Singapore Management University
Safegpu: Contract- And Library-Based Gpgpu For Object-Oriented Languages, Alexey Kolesnichenko, Christopher M. Poskitt, Sebastian Nanz
Research Collection School Of Computing and Information Systems
Using GPUs as general-purpose processors has revolutionized parallel computing by providing, for a large and growing set of algorithms, massive data-parallelization on desktop machines. An obstacle to their widespread adoption, however, is the difficulty of programming them and the low-level control of the hardware required to achieve good performance. This paper proposes a programming approach, SafeGPU, that aims to make GPU data-parallel operations accessible through high-level libraries for object-oriented languages, while maintaining the performance benefits of lower-level code. The approach provides data-parallel operations for collections that can be chained and combined to express compound computations, with data synchronization and device …
Satisfiability Modulo Heap-Based Programs,
2016
Singapore Management University
Satisfiability Modulo Heap-Based Programs, Quang Loc Le, Jun Sun, Wei-Ngan Chin
Research Collection School Of Computing and Information Systems
In this work, we present a semi-decision procedure for a fragment of separation logic with user-defined predicates and Presburger arithmetic. To check the satisfiability of a formula, our procedure iteratively unfolds the formula and examines the derived disjuncts. In each iteration, it searches for a proof of either satisfiability or unsatisfiability. Our procedure is further enhanced with automatically inferred invariants as well as detection of cyclic proof. We also identify a syntactically restricted fragment of the logic for which our procedure is terminating and thus complete. This decidable fragment is relatively expressive as it can capture a range of sophisticated …
Examining Bridges Between Informal And Formal Learning Environments: A Sequential Mixed Method Design,
2016
University of Nebraska-Lincoln
Examining Bridges Between Informal And Formal Learning Environments: A Sequential Mixed Method Design, Dagen L. Valentine
Department of Agricultural Leadership, Education, and Communication: Dissertations, Theses, and Student Research
The purpose of this sequential mixed method study was to identify schools implementing a technology-based engineering design intervention in a way that connects or bridges formal learning environments of the school-day to informal learning environments such as afterschool programs. Further, this study investigated educators’ decisions that enabled or facilitated bridging between formal and informal learning environments. This cooperation and/or linking between informal and formal learning time is bridging. Participants included public schools (n=16) in Eastern Nebraska that incorporated the Nebraska Wearables Technology (WearTec) program at their school, club or Out-of-School-Time program during the 2015-2016 school year. Three of the schools …
Defining The Competencies, Programming Languages, And Assessments For An Introductory Computer Science Course,
2016
Old Dominion University
Defining The Competencies, Programming Languages, And Assessments For An Introductory Computer Science Course, Simon Sultana
STEMPS Theses & Dissertations
The purpose of this study was to define the competencies, programming languages, and assessments for an introductory computer science course at a small private liberal arts university. Three research questions were addressed that involved identifying the competencies, programming languages, and assessments that academic and industry experts in California’s Central Valley felt most important and appropriate for an introduction to computer science course.
The Delphi methodology was used to collect data from the two groups of experts with various backgrounds related to computing. The goal was to find consensus among the individual groups to best define aspects that would best comprise …
Fine-Grained Detection Of Programming Students’ Frustration Using Keystrokes, Mouse Clicks And Interaction Logs,
2016
Singapore Management University
Fine-Grained Detection Of Programming Students’ Frustration Using Keystrokes, Mouse Clicks And Interaction Logs, Hua Leong Fwa
Research Collection School Of Computing and Information Systems
Prolonged frustration leads to loss of confidence and eventual disinterest in the learning itself. The modelling of frustration in learning is thus important as it informs on the appropriate time to intervene to sustain the interest and motivation of students. To automatically detect learner’s frustration in a naturalistic learning environment, the novel use of keystrokes, mouse clicks and interaction patterns of students captured within the context of a tutoring system was proposed. The modelling approach was described and a comparison was made between the proposed model using Bayesian Network and the baseline Naïve Bayes model. With the formulation of an …
An Interference-Free Programming Model For Network Objects,
2016
Singapore Management University
An Interference-Free Programming Model For Network Objects, Mischael Schill, Christopher M. Poskitt, Bertrand Meyer
Research Collection School Of Computing and Information Systems
Network objects are a simple and natural abstraction for distributed object-oriented programming. Languages that support network objects, however, often leave synchronization to the user, along with its associated pitfalls, such as data races and the possibility of failure. In this paper, we present D-Scoop, a distributed programming model that allows for interference-free and transaction-like reasoning on (potentially multiple) network objects, with synchronization handled automatically, and network failures managed by a compensation mechanism. We achieve this by leveraging the runtime semantics of a multi-threaded object-oriented concurrency model, directly generalizing it with a message-based protocol for efficiently coordinating remote objects. We present …
Estimability Tools For Package Developers,
2016
The University of Iowa
Estimability Tools For Package Developers, Russell V. Lenth
The R Journal
When a linear model is rank-deficient, then predictions based on that model become questionable because not all predictions are uniquely estimable. However, some of them are, and the estimability package provides tools that package developers can use to tell which is which. With the use of these tools, a model object’s predict method could return estimable predictions as-is while f lagging non-estimable ones in some way, so that the user can know which predictions to believe. The estimability package also provides, as a demonstration, an estimability-enhanced epredict method to use in place of predict for models fitted using the stats …
