Open Access. Powered by Scholars. Published by Universities.®

2008

Discipline
Institution
Keyword
Publication
Publication Type

Articles 1 - 9 of 9

Full-Text Articles in Categorical Data Analysis

A Small Sample Correction For Estimating Attributable Risk In Case-Control Studies, Daniel B. Rubin Dec 2008

A Small Sample Correction For Estimating Attributable Risk In Case-Control Studies, Daniel B. Rubin

U.C. Berkeley Division of Biostatistics Working Paper Series

The attributable risk, often called the population attributable risk, is in many epidemiological contexts a more relevant measure of exposure-disease association than the excess risk, relative risk, or odds ratio. When estimating attributable risk with case-control data and a rare disease, we present a simple correction to the standard approach making it essentially unbiased, and also less noisy. As with analogous corrections given in Jewell (1986) for other measures of association, the adjustment often won't make a substantial difference unless the sample size is very small or point estimates are desired within fine strata, but we discuss the possible utility …


Multilevel Latent Class Models With Dirichlet Mixing Distribution, Chongzhi Di, Karen Bandeen-Roche Oct 2008

Multilevel Latent Class Models With Dirichlet Mixing Distribution, Chongzhi Di, Karen Bandeen-Roche

Johns Hopkins University, Dept. of Biostatistics Working Papers

Latent class analysis (LCA) and latent class regression (LCR) are widely used for modeling multivariate categorical outcomes in social sciences and biomedical studies. Standard analyses assume data of different respondents to be mutually independent, excluding application of the methods to familial and other designs in which participants are clustered. In this paper, we develop multilevel latent class model, in which subpopulation mixing probabilities are treated as random effects that vary among clusters according to a common Dirichlet distribution. We apply the Expectation-Maximization (EM) algorithm for model fitting by maximum likelihood (ML). This approach works well, but is computationally intensive when …


Measurement Error Caused By Spatial Misalignment In Environmental Epidemiology, Alexandros Gryparis, Christopher J. Paciorek, Ariana Zeka, Joel Schwartz, Brent A. Coull Sep 2008

Measurement Error Caused By Spatial Misalignment In Environmental Epidemiology, Alexandros Gryparis, Christopher J. Paciorek, Ariana Zeka, Joel Schwartz, Brent A. Coull

Harvard University Biostatistics Working Paper Series

No abstract provided.


A Study In Rule-Specific Issue Categorization For E-Rulemaking, Claire Cardie, Cynthia R. Farina, Adil Aijaz, Matt Rawding, Stephen Purpura May 2008

A Study In Rule-Specific Issue Categorization For E-Rulemaking, Claire Cardie, Cynthia R. Farina, Adil Aijaz, Matt Rawding, Stephen Purpura

Cornell Law Faculty Publications

We address the e-rulemaking problem of categorizing public comments according to the issues that they address. In contrast to previous text categorization research in e-rulemaking, and in an attempt to more closely duplicate the comment analysis process in federal agencies, we employ a set of rule-specific categories, each of which corresponds to a significant issue raised in the comments. We describe the creation of a corpus to support this text categorization task and report interannotator agreement results for a group of six annotators. We outline those features of the task and of the e-rulemaking context that engender both a non-traditional …


Changes In Poverty And Educational Attainment, 2000 To 2007 Poverty Rates Increasing For Those With College Education, Too, Mark Salling Jan 2008

Changes In Poverty And Educational Attainment, 2000 To 2007 Poverty Rates Increasing For Those With College Education, Too, Mark Salling

All Maxine Goodman Levin School of Urban Affairs Publications

No abstract provided.


Hispanics And Asians Increase In Numbers In Cuyahoga County An Analysis Of 2007 County Population Estimates, Mark Salling Jan 2008

Hispanics And Asians Increase In Numbers In Cuyahoga County An Analysis Of 2007 County Population Estimates, Mark Salling

All Maxine Goodman Levin School of Urban Affairs Publications

No abstract provided.


The Cleveland-Akron-Elyria Region Doing Well: More Persons Attending College And Getting Degrees, 2000 To 2007, Mark Salling Jan 2008

The Cleveland-Akron-Elyria Region Doing Well: More Persons Attending College And Getting Degrees, 2000 To 2007, Mark Salling

All Maxine Goodman Levin School of Urban Affairs Publications

Discussions of economic development and job availability in northeast Ohio often lament the unavailability of a qualified workforce in some sectors. Workforce training and attracting more educated population to the region are sited as important, even critical, objectives for the region. While a more detailed study of the regions’ workforce by The Center for Community Solutions is nearing completion, the release of new data by the Census Bureau provides some enlightening observations about college enrollments and educational attainment in the region.


Ohio Continues To Lag In Population Growth And Comments On Prospects For The Future An Analysis Of 2007 State Population Estimates, Mark Salling Jan 2008

Ohio Continues To Lag In Population Growth And Comments On Prospects For The Future An Analysis Of 2007 State Population Estimates, Mark Salling

All Maxine Goodman Levin School of Urban Affairs Publications

No abstract provided.


Data Mining Methods For Malware Detection, Muazzam Siddiqui Jan 2008

Data Mining Methods For Malware Detection, Muazzam Siddiqui

Electronic Theses and Dissertations

This research investigates the use of data mining methods for malware (malicious programs) detection and proposed a framework as an alternative to the traditional signature detection methods. The traditional approaches using signatures to detect malicious programs fails for the new and unknown malwares case, where signatures are not available. We present a data mining framework to detect malicious programs. We collected, analyzed and processed several thousand malicious and clean programs to find out the best features and build models that can classify a given program into a malware or a clean class. Our research is closely related to information retrieval …