Open Access. Powered by Scholars. Published by Universities.®

Statistics and Probability Commons™

Open Access. Powered by Scholars. Published by Universities.®

2004

Discipline
Institution
Keyword
Publication
Publication Type

Articles 91 - 120 of 235

Full-Text Articles in Statistics and Probability

A Hierarchical Multivariate Two-Part Model For Profiling Providers' Effects On Healthcare Charges, John W. Robinson, Scott L. Zeger, Christopher B. Forrest Aug 2004

A Hierarchical Multivariate Two-Part Model For Profiling Providers' Effects On Healthcare Charges, John W. Robinson, Scott L. Zeger, Christopher B. Forrest

Johns Hopkins University, Dept. of Biostatistics Working Papers

Procedures for analyzing and comparing healthcare providers' effects on health services delivery and outcomes have been referred to as provider profiling. In a typical profiling procedure, patient-level responses are measured for clusters of patients treated by providers that in turn, can be regarded as statistically exchangeable. Thus, a hierarchical model naturally represents the structure of the data. When provider effects on multiple responses are profiled, a multivariate model rather than a series of univariate models, can capture associations among responses at both the provider and patient levels. When responses are in the form of charges for healthcare services and sampled …


Building A Model Of Campus Diversity: An Organization-As-The-Individual Based Approach, Ian Spencer Ray Aug 2004

Building A Model Of Campus Diversity: An Organization-As-The-Individual Based Approach, Ian Spencer Ray

Electronic Theses and Dissertations

Diversity is a vital concept when engaging with modern social science research and among higher education practitioners. Despite the pervasive discourse surrounding diversity, equity, and inclusion (DEI) initiatives, quantification of diversity remains a challenge across the social sciences. This study proposes a new conceptual framework, applies the framework to the quantification of diversity and measurement of its impacts. Further, this study demonstrates the robustness of the new conceptual framework alongside the positive impacts of diversity on institutional outcomes.

The Organization-As-The-Individual (OATI) framework arises from an integration of the social and biological sciences. The framework argues that humans, as biological organisms, …


Estimation Of Direct And Indirect Causal Effects In Longitudinal Studies, Mark J. Van Der Laan, Maya L. Petersen Aug 2004

Estimation Of Direct And Indirect Causal Effects In Longitudinal Studies, Mark J. Van Der Laan, Maya L. Petersen

U.C. Berkeley Division of Biostatistics Working Paper Series

The causal effect of a treatment on an outcome is generally mediated by several intermediate variables. Estimation of the component of the causal effect of a treatment that is mediated by a given intermediate variable (the indirect effect of the treatment), and the component that is not mediated by that intermediate variable (the direct effect of the treatment) is often relevant to mechanistic understanding and to the design of clinical and public health interventions. Under the assumption of no-unmeasured confounders, Robins & Greenland (1992) and Pearl (2000), develop two identifiability results for direct and indirect causal effects. They define an …


Linear Life Expectancy Regression With Censored Data, Ying Qing Chen, Su-Chun Cheng Aug 2004

Linear Life Expectancy Regression With Censored Data, Ying Qing Chen, Su-Chun Cheng

U.C. Berkeley Division of Biostatistics Working Paper Series

Life expectancy, i.e., mean residual life function, has been of important practical and scientific interests to characterise the distribution of residual life. Regression models are often needed to model the association between life expectancy and its covariates. In this article, we consider a linear mean residual life model and further developed some inference procedures in presence of censoring. The new model and proposed inference procedure will be demonstrated by numerical examples and application to the well-known Stanford heart transplant data. Additional semiparametric efficiency calculation and information bound are also considered.


A Comparative Study Of Interrater Reliability Coefficients Obtained From Different Statistical Procedures Using Monte Carlo Simulation Techniques, Ebrima Nying Aug 2004

A Comparative Study Of Interrater Reliability Coefficients Obtained From Different Statistical Procedures Using Monte Carlo Simulation Techniques, Ebrima Nying

Dissertations

Reliability estimation is a key research component within the global area of educational assessment. The literature reports numerous studies using different statistical techniques for estimating reliability of educational measures. However, few have focused on the estimation of interrater reliability of performance assessment (Abedi, Baker, & Herl, 1995). Specifically, this study compared three different methods for estimating interrater reliability to determine if there are differences among these estimates as a function of: sample size, measurement scale, number of raters and the theoretical population reliability (rho). The three methods of estimation were the Intraclass Correlation (ICC(2, k )) (Shrout & Fleiss, 1979), …


Enhancements To Crisp Possibilistic Reconstructability Analysis, Anas Al-Rabadi, Martin Zwick Aug 2004

Enhancements To Crisp Possibilistic Reconstructability Analysis, Anas Al-Rabadi, Martin Zwick

Complex Systems Faculty Publications and Presentations

Modified Reconstructibility Analysis (MRA), a novel decomposition within the framework of set-theoretic (crisp possibilistic) Reconstructibility Analysis, is presented. It is shown that in some cases while 3-variable NPN-classified Boolean functions are not decomposable using Conventional Reconstructibility Analysis (CRA), they are decomposable using Modified Reconstructibility Analysis (MRA). Also, it is shown that whenever a decomposition of 3-variable NPN-classified Boolean functions exists in both MRA and CRA, MRA yields simpler or equal complexity decompositions. A comparison of the corresponding complexities for Ashenhurst-Curtis decompositions, and Modified Reconstructibility Analysis (MRA) is also presented. While both AC and MRA decompose some but …


Instructing Teachers Of Children With Disabilities Within The Church Of Jesus Christ Of Latter-Day Saints, Katie E. Sampson Aug 2004

Instructing Teachers Of Children With Disabilities Within The Church Of Jesus Christ Of Latter-Day Saints, Katie E. Sampson

Theses and Dissertations

This study investigates benefits of in-service training on LDS primary teachers' ability to state an objective, obtain and keep attention, use wait time, incorporate active participation, teach to the multiple intelligences, and employ positive behavior management techniques. Two groups of 30 viewed either a video-tape or read a handout. Pre and post surveys were used to determine mean gain.
Using an ANCOVA, comparisons were made of overall mean gain for each group. Results showed participants made a gain of approximately 1/2 point per question on a 4-point scale on the video and the handout (video gain = .6032 p<.01; handout gain = .6264 p<.01). The results of this study support the hypothesis that teachers receiving one in-service will increase their perception of their ability to teach students with special needs.


Non-Parametric Estimation Of Roc Curves In The Absence Of A Gold Standard, Xiao-Hua Zhou, Pete Castelluccio, Chuan Zhou Jul 2004

Non-Parametric Estimation Of Roc Curves In The Absence Of A Gold Standard, Xiao-Hua Zhou, Pete Castelluccio, Chuan Zhou

UW Biostatistics Working Paper Series

In evaluation of diagnostic accuracy of tests, a gold standard on the disease status is required. However, in many complex diseases, it is impossible or unethical to obtain such the gold standard. If an imperfect standard is used as if it were a gold standard, the estimated accuracy of the tests would be biased. This type of bias is called imperfect gold standard bias. In this paper we develop a maximum likelihood (ML) method for estimating ROC curves and their areas of ordinal-scale tests in the absence of a gold standard. Our simulation study shows the proposed estimates for the …


The Optimal Confidence Region For A Random Parameter, Hajime Uno, Lu Tian, L.J. Wei Jul 2004

The Optimal Confidence Region For A Random Parameter, Hajime Uno, Lu Tian, L.J. Wei

Harvard University Biostatistics Working Paper Series

Under a two-level hierarchical model, suppose that the distribution of the random parameter is known or can be estimated well. Data are generated via a fixed, but unobservable realization of this parameter. In this paper, we derive the smallest confidence region of the random parameter under a joint Bayesian/frequentist paradigm. On average this optimal region can be much smaller than the corresponding Bayesian highest posterior density region. The new estimation procedure is appealing when one deals with data generated under a highly parallel structure, for example, data from a trial with a large number of clinical centers involved or genome-wide …


Probabilistic Methodology For Record Linkage Determining Robustness Of Weights, Krista Peine Jensen Jul 2004

Probabilistic Methodology For Record Linkage Determining Robustness Of Weights, Krista Peine Jensen

Theses and Dissertations

Record linkage is the process that joins separately recorded pieces of information for a particular individual from one or more sources. To facilitate record linkage, a reliable computer based approach is ideal. In genealogical research computerized record linkage is useful in combing information for an individual across multiple censuses.

In creating a computerized method for linking censuse records it needs to be determined if weights calculated from one geographical area, can be used to link records from another geographical area. Research performed by Marcie Francis calculates field weights using census records from 1910 and 1920 for Ascension Parish Louisiana. These …


Hiring Practices For Graphic Designers In Utah County, Utah, Landon T. Densley Jul 2004

Hiring Practices For Graphic Designers In Utah County, Utah, Landon T. Densley

Theses and Dissertations

The purpose of this study was to show how hiring standards of evidence for graphic designers in Utah County compared with the national standards of evidence. The four major national standards of evidence for hiring graphic designers, identified by American Institute of Graphic Arts (AIGA) and Goldfarb, in order of importance are portfolio, recommendations, personality, and education. The data from this study revealed that Utah County employer's standards of evidence matched up closely to national standards of evidence, but the order of importance was slightly different because personality was ranked ahead of recommendations and education.


A Note On Empirical Likelihood Inference Of Residual Life Regression, Ying Qing Chen, Yichuan Zhao Jul 2004

A Note On Empirical Likelihood Inference Of Residual Life Regression, Ying Qing Chen, Yichuan Zhao

U.C. Berkeley Division of Biostatistics Working Paper Series

Mean residual life function, or life expectancy, is an important function to characterize distribution of residual life. The proportional mean residual life model by Oakes and Dasu (1990) is a regression tool to study the association between life expectancy and its associated covariates. Although semiparametric inference procedures have been proposed in the literature, the accuracy of such procedures may be low when the censoring proportion is relatively large. In this paper, the semiparametric inference procedures are studied with an empirical likelihood ratio method. An empirical likelihood confidence region is constructed for the regression parameters. The proposed method is further compared …


Large Prandtl Number Behavior Of The Boussinesq System Of Rayleigh-Bénard Convection, Xiaoming Wang Jul 2004

Large Prandtl Number Behavior Of The Boussinesq System Of Rayleigh-Bénard Convection, Xiaoming Wang

Mathematics and Statistics Faculty Research & Creative Works

We establish the validity of the infinite Prandtl number model as an approximation of the Boussinesq system at large Prandtl number on finite and infinite time interval, as well as in some statistical sense. © 2004 Elsevier Ltd. All rights reserved.


Semiparametric Quantitative-Trait-Locus Mapping: I. On Functional Growth Curves, Ying Qing Chen, Rongling Wu Jul 2004

Semiparametric Quantitative-Trait-Locus Mapping: I. On Functional Growth Curves, Ying Qing Chen, Rongling Wu

U.C. Berkeley Division of Biostatistics Working Paper Series

The genetic study of certain quantitative traits in growth curves as a function of time has recently been of major scientific interest to explore the developmental evolution processes of biological subjects. Various parametric approaches in the statistical literature have been proposed to study the quantitative-trait-loci (QTL) mapping of the growth curves as multivariate outcomes. In this article, we view the growth curves as functional quantitative traits and propose some semiparametric models to relax the strong parametric assumptions which may not be always practical in reality. Appropriate inference procedures are developed to estimate the parameters of interest which characterise the possible …


Semiparametric Quantitative-Trait-Locus Mapping: Ii. On Censored Age-At-Onset, Ying Qing Chen, Chengcheng Hu, Rongling Wu Jul 2004

Semiparametric Quantitative-Trait-Locus Mapping: Ii. On Censored Age-At-Onset, Ying Qing Chen, Chengcheng Hu, Rongling Wu

U.C. Berkeley Division of Biostatistics Working Paper Series

In genetic studies, the variation in genotypes may not only affect different inheritance patterns in qualitative traits, but may also affect the age-at-onset as quantitative trait. In this article, we use standard cross designs, such as backcross or F2, to propose some hazard regression models, namely, the additive hazards model in quantitative trait loci mapping for age-at-onset, although the developed method can be extended to more complex designs. With additive invariance of the additive hazards models in mixture probabilities, we develop flexible semiparametric methodologies in interval regression mapping without heavy computing burden. A recently developed multiple comparison procedures is adapted …


The Morphology Of Steve, Eugenie C. Scott, Nicholas J. Matzke, Glenn Branch, Steven Mccullagh Jul 2004

The Morphology Of Steve, Eugenie C. Scott, Nicholas J. Matzke, Glenn Branch, Steven Mccullagh

Faculty Articles

This report is part of Project Steve. Project Steve is, among other things, the first scientific analysis of the sex, geographic location, and body size of scientists named Steve. We performed this research for the best of all reasons: we discovered that we had lots of data. No scientist can resist the opportunity to analyze data, regardless of where that data came from or why it was gathered.


Oscillation Of Symplectic Dynamic Systems, Martin Bohner, Ondřej Došlý Jul 2004

Oscillation Of Symplectic Dynamic Systems, Martin Bohner, Ondřej Došlý

Mathematics and Statistics Faculty Research & Creative Works

We investigate oscillatory properties of a perturbed symplectic dynamic system on a time scale that is unbounded above. the unperturbed system is supposed to be Non oscillatory, and we give conditions on the perturbation matrix, which guarantee that the perturbed system becomes oscillatory. Examples illustrating the general results are given as well. © Australian Mathematical Society 2004.


Implicit Racial Attitudes Of Death Penalty Lawyers, Theodore Eisenberg, Sheri Lynn Johnson Jul 2004

Implicit Racial Attitudes Of Death Penalty Lawyers, Theodore Eisenberg, Sheri Lynn Johnson

Cornell Law Faculty Publications

Defense attorneys commonly suspect that the defendant's race plays a role in prosecutors' decisions to seek the death penalty, especially when the victim of the crime was white. When the defendant is convicted of the crime and sentenced to death, it is equally common for such attorneys to question the racial attitudes of the jury. These suspicions are not merely partisan conjectures; ample historical, statistical, and anecdotal evidence supports the inference that race matters in capital cases. Even the General Accounting Office of the United States concludes as much. Despite McCleskey v. Kemp, in which the United States Supreme Court …


Quantification And Visualization Of Ld Patterns And Identification Of Haplotype Blocks, Yan Wang, Sandrine Dudoit Jun 2004

Quantification And Visualization Of Ld Patterns And Identification Of Haplotype Blocks, Yan Wang, Sandrine Dudoit

U.C. Berkeley Division of Biostatistics Working Paper Series

Classical measures of linkage disequilibrium (LD) between two loci, based only on the joint distribution of alleles at these loci, present noisy patterns. In this paper, we propose a new distance-based LD measure, R, which takes into account multilocus haplotypes around the two loci in order to exploit information from neighboring loci. The LD measure R yields a matrix of pairwise distances between markers, based on the correlation between the lengths of shared haplotypes among chromosomes around these markers. Data analysis demonstrates that visualization of LD patterns through the R matrix reveals more deterministic patterns, with much less noise, than …


Quantitative Methods For Tracking Cognitive Change 3 Years After Coronary Artery Bypass Surgery, Sarah Barry, Scott L. Zeger, Ola A. Selnes, Maura A. Grega, Louis M. Borowicz, Jr., Guy M. Mckhann Jun 2004

Quantitative Methods For Tracking Cognitive Change 3 Years After Coronary Artery Bypass Surgery, Sarah Barry, Scott L. Zeger, Ola A. Selnes, Maura A. Grega, Louis M. Borowicz, Jr., Guy M. Mckhann

Johns Hopkins University, Dept. of Biostatistics Working Papers

Background: The analysis and interpretation of change in cognitive function test scores after Coronary Artery Bypass Grafting (CABG). Longitudinal studies with multiple outcomes present considerable statistical challenges. Application of hierarchical linear statistical models can estimate the effects of a surgical intervention on the time course of multiple biomarkers.

Methods: We use an "analyze then summarize" approach whereby we estimate the intervention effects separately for each cognitive test and then pool them, taking appropriate account of their statistical correlations. The model accounts for dropouts at follow-up, the chance of which may be related to past cognitive score, by implicitly imputing the …


New Estimating Methods For Surrogate Outcome Data, Bin Nan Jun 2004

New Estimating Methods For Surrogate Outcome Data, Bin Nan

The University of Michigan Department of Biostatistics Working Paper Series

Surrogate outcome data arise frequently in medical research. The true outcomes of interest are expensive or hard to ascertain, but measurements of surrogate outcomes (or more generally speaking, the correlates of the true outcomes) are usually available. In this paper we assume that the conditional expectation of the true outcome given covariates is known up to a finite dimensional parameter. When the true outcome is missing at random, the e±cient score function for the parameter in the conditional mean model has a simple form, which is similar to the generalized estimating functions. There is no integral equation involved as in …


Differential Expression With The Bioconductor Project, Anja Von Heydebreck, Wolfgang Huber, Robert Gentleman Jun 2004

Differential Expression With The Bioconductor Project, Anja Von Heydebreck, Wolfgang Huber, Robert Gentleman

Bioconductor Project Working Papers

A basic, yet challenging task in the analysis of microarray gene expression data is the identification of changes in gene expression that are associated with particular biological conditions. We discuss different approaches to this task and illustrate how they can be applied using software from the Bioconductor Project. A central problem is the high dimensionality of gene expression space, which prohibits a comprehensive statistical analysis without focusing on particular aspects of the joint distribution of the genes expression levels. Possible strategies are to do univariate gene-by-gene analysis, and to perform data-driven nonspecific filtering of genes before the actual statistical analysis. …


Asymptotic Results For Simultaneous Group Sequential Analysis Of Rank-Based And Weighted Kaplan-Meier Tests With Paired Survival Data In The Presence Of Censoring. Technical Report, Adin-Cristian Andrei, Susan Murray Jun 2004

Asymptotic Results For Simultaneous Group Sequential Analysis Of Rank-Based And Weighted Kaplan-Meier Tests With Paired Survival Data In The Presence Of Censoring. Technical Report, Adin-Cristian Andrei, Susan Murray

The University of Michigan Department of Biostatistics Working Paper Series

This research sequentially monitors paired survival differences using a new class of non-parametric tests based on functionals of standardized paired weighted log-rank (PWLR) and standardized paired weighted Kaplan-Meier (PWKM) tests. During a trial these tests may alternately assume the role of the more extreme statistic. By monitoring PEMAX, the maximum between the absolute values of the standardized PWLR and PWKM, one combines advantages of rank-based and non rank-based paired testing paradigms. Simulations show that monitoring treatment differences using PEMAX maintains type I error and is nearly as powerful as using the more advantageous of the two tests, in proportional hazards …


Bayesian Geostatistical Design, Peter J. Diggle, Soren Lophaven Jun 2004

Bayesian Geostatistical Design, Peter J. Diggle, Soren Lophaven

Johns Hopkins University, Dept. of Biostatistics Working Papers

This paper describes the use of model-based geostatistics for choosing the optimal set of sampling locations, collectively called the design, for a geostatistical analysis. Two types of design situations are considered. These are retrospective design, which concerns the addition of sampling locations to, or deletion of locations from, an existing design, and prospective design, which consists of choosing optimal positions for a new set of sampling locations. We propose a Bayesian design criterion which focuses on the goal of efficient spatial prediction whilst allowing for the fact that model parameter values are unknown. The results show that in this situation …


Nonparametric Methods For Analyzing Replication Origins In Genomewide Data, Debashis Ghosh Jun 2004

Nonparametric Methods For Analyzing Replication Origins In Genomewide Data, Debashis Ghosh

The University of Michigan Department of Biostatistics Working Paper Series

Due to the advent of high-throughput genomic technology, it has become possible to globally monitor cellular activities on a genomewide basis. With these new methods, scientists can begin to address important biological questions. One such question involves the identification of replication origins, which are regions in chromosomes where DNA replication is initiated. In addition, one hypothesis regarding replication origins is that their locations are non-random throughout the genome. In this article, we develop methods for identification of and cluster inference regarding replication origins involving genomewide expression data. We compare several nonparametric regression methods for the identification of replication origin locations. …


Mean Response Models Of Repeated Measurements In Presence Of Varying Effectiveness Onset, Ying Qing Chen, Su-Chun Cheng Jun 2004

Mean Response Models Of Repeated Measurements In Presence Of Varying Effectiveness Onset, Ying Qing Chen, Su-Chun Cheng

U.C. Berkeley Division of Biostatistics Working Paper Series

Repeated measurements are often collected over time to evaluate treatment efficacy in clinical trials. Most of the statistical models of the repeated measurements have been focusing on their mean response as function of time. These models usually assume that the treatment has persistent effect of constant additivity or multiplicity on the mean response functions throughout the observation period of time. In reality, however, such assumption may be confounded by the potential existence of the so-called effectiveness action onset, although they are often unobserved or difficult to obtain. Instead of including nonparametric time-varying coefficients in the mean response models, we propose …


Semiparametric Methods For Identification Of Tumor Progression Genes From Microarray Data, Debashis Ghosh, Arul Chinnaiyan Jun 2004

Semiparametric Methods For Identification Of Tumor Progression Genes From Microarray Data, Debashis Ghosh, Arul Chinnaiyan

The University of Michigan Department of Biostatistics Working Paper Series

The use of microarray data has become quite commonplace in medical and scientific experiments. We focus here on microarray data generated from cancer studies. It is potentially important for the discovery of biomarkers to identify genes whose expression levels correlate with tumor progression. In this article, we develop statistical procedures for the identification of such genes, which we term tumor progression genes. Two methods are considered in this paper. The first is use of a proportional odds procedure, combined with false discovery rate estimation techniques to adjust for the multiple testing problem. The second method is based on order-restricted estimation …


The False Discovery Rate: A Variable Selection Perspective, Debashis Ghosh, Wei Chen, Trivellore E. Raghuanthan Jun 2004

The False Discovery Rate: A Variable Selection Perspective, Debashis Ghosh, Wei Chen, Trivellore E. Raghuanthan

The University of Michigan Department of Biostatistics Working Paper Series

In many scientific and medical settings, large-scale experiments are generating large quantities of data that lead to inferential problems involving multiple hypotheses. This has led to recent tremendous interest in statistical methods regarding the false discovery rate (FDR). Several authors have studied the properties involving FDR in a univariate mixture model setting. In this article, we turn the problem on its side; in this manuscript, we show that FDR is a by-product of Bayesian analysis of variable selection problem for a hierarchical linear regression model. This equivalence gives many Bayesian insights as to why FDR is a natural quantity to …


Multiple Testing Methods For Chip-Chip High Density Oligonucleotide Array Data, Sunduz Keles, Mark J. Van Der Laan, Sandrine Dudoit, Simon E. Cawley Jun 2004

Multiple Testing Methods For Chip-Chip High Density Oligonucleotide Array Data, Sunduz Keles, Mark J. Van Der Laan, Sandrine Dudoit, Simon E. Cawley

U.C. Berkeley Division of Biostatistics Working Paper Series

Cawley et al. (2004) have recently mapped the locations of binding sites for three transcription factors along human chromosomes 21 and 22 using ChIP-Chip experiments. ChIP-Chip experiments are a new approach to the genome-wide identification of transcription factor binding sites and consist of chromatin (Ch) immunoprecipitation (IP) of transcription factor-bound genomic DNA followed by high density oligonucleotide hybridization (Chip) of the IP-enriched DNA. We investigate the ChIP-Chip data structure and propose methods for inferring the location of transcription factor binding sites from these data. The proposed methods involve testing for each probe whether it is part of a bound sequence …


Combining Predictors For Classification Using The Area Under The Roc Curve, Margaret S. Pepe, Tianxi Cai, Zheng Zhang Jun 2004

Combining Predictors For Classification Using The Area Under The Roc Curve, Margaret S. Pepe, Tianxi Cai, Zheng Zhang

UW Biostatistics Working Paper Series

We compare simple logistic regression with an alternative robust procedure for constructing linear predictors to be used for the two state classification task. Theoritical advantages of the robust procedure over logistic regression are: (i) although it assumes a generalized linear model for the dichotomous outcome variable, it does not require specification of the link function; (ii) it accommodates case-control designs even when the model is not logistic; and (iii) it yields sensible results even when the generalized linear model assumption fails to hold. Surprisingly, we find that the linear predictor derived from the logistic regression likelihood is very robust in …