Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Medicine and Health Sciences (230)
- Biostatistics (214)
- Public Health (175)
- Social and Behavioral Sciences (125)
- Epidemiology (123)
-
- Applied Statistics (119)
- Life Sciences (100)
- Data Science (74)
- Mathematics (68)
- Health Services Research (65)
- Medical Specialties (65)
- Statistical Models (57)
- Statistical Methodology (55)
- Education (50)
- Public Affairs, Public Policy and Public Administration (44)
- Clinical Trials (43)
- Applied Mathematics (42)
- Engineering (41)
- Medical Sciences (40)
- Public Health Education and Promotion (40)
- Environmental Public Health (38)
- Computer Sciences (36)
- Women's Health (36)
- Health Policy (35)
- Occupational Health and Industrial Hygiene (34)
- Nutrition (33)
- Probability (31)
- Health and Medical Administration (30)
- Institution
-
- University of South Carolina (66)
- University of Kentucky (50)
- Universitas Indonesia (33)
- Roseman University of Health Sciences (31)
- Wayne State University (27)
-
- Kennesaw State University (25)
- Chulalongkorn University (21)
- Missouri University of Science and Technology (20)
- University of South Florida (15)
- Illinois State University (12)
- Old Dominion University (12)
- Prairie View A&M University (12)
- Georgia Southern University (11)
- Northern Illinois University (11)
- University of Arkansas, Fayetteville (11)
- University of Nebraska - Lincoln (11)
- Air Force Institute of Technology (10)
- University of Nevada, Las Vegas (10)
- Virginia Commonwealth University (10)
- Claremont Colleges (9)
- University of Texas at El Paso (9)
- Utah State University (9)
- City University of New York (CUNY) (8)
- University of Denver (8)
- Marquette University (7)
- Smith College (7)
- University of New Mexico (7)
- SIT Graduate Institute/SIT Study Abroad (6)
- University of Louisville (6)
- Western University (6)
- Keyword
-
- Statistics (22)
- COVID-19 (21)
- Machine learning (16)
- Dietary inflammatory index (11)
- Inflammation (9)
-
- Obesity (8)
- Bayesian (7)
- Forecasting (7)
- Machine Learning (7)
- Morgridge College of Education (7)
- Research Methods and Information Science (7)
- Research Methods and Statistics (7)
- Simulation (7)
- United States (7)
- Diabetes (6)
- Diet (6)
- Humans (6)
- Maximum likelihood estimation (6)
- Regression (6)
- Risk (6)
- Causal inference (5)
- Deep learning (5)
- HIV (5)
- Mathematics (5)
- Mental health (5)
- Metabolic syndrome (5)
- Missing data (5)
- Pregnancy (5)
- Age (4)
- Appalachia (4)
- Publication
-
- Faculty Publications (53)
- Theses and Dissertations (36)
- Kesmas (33)
- Annual Research Symposium (31)
- Symposium of Student Scholars (24)
-
- Electronic Theses and Dissertations (23)
- Journal of Modern Applied Statistical Methods (23)
- Chulalongkorn University Theses and Dissertations (Chula ETD) (21)
- Mathematics and Statistics Faculty Research & Creative Works (14)
- Applications and Applied Mathematics: An International Journal (AAM) (12)
- Graduate Research Theses & Dissertations (11)
- Biostatistics Faculty Publications (10)
- Biostatistics, Epidemiology & Environmental Health Sciences: Faculty Publications (10)
- Annual Symposium on Biomathematics and Ecology Education and Research (9)
- Open Access Theses & Dissertations (9)
- Epidemiology and Environmental Health Faculty Publications (8)
- USF Tampa Graduate Theses and Dissertations (8)
- Department of Statistics: Faculty Publications (7)
- Mathematics & Statistics ETDs (7)
- Statistical and Data Sciences: Faculty Publications (7)
- Dissertations (6)
- Epidemiology and Biostatistics Publications (6)
- Graduate Theses and Dissertations (6)
- Independent Study Project (ISP) Collection (6)
- Numeracy (6)
- Dissertations, Master's Theses and Master's Reports (5)
- Mathematical and Statistical Science Faculty Research and Publications (5)
- Theses and Dissertations--Statistics (5)
- All Graduate Theses and Dissertations, Spring 1920 to Summer 2023 (4)
- CMC Senior Theses (4)
- Publication Type
- File Type
Articles 571 - 600 of 662
Full-Text Articles in Statistics and Probability
Dimension Reduction Techniques In Regression, Pei Wang
Dimension Reduction Techniques In Regression, Pei Wang
Theses and Dissertations--Statistics
Because of the advances of modern technology, the size of the collected data nowadays is larger and the structure is more complex. To deal with such kinds of data, sufficient dimension reduction (SDR) and reduced rank (RR) regression are two powerful tools. This dissertation focuses on these two tools and it is composed of three projects. In the first project, we introduce a new SDR method through a novel approach of feature filter to recover the central mean subspace exhaustively along with a method to determine the dimension, two variable selection methods, and extensions to multivariate response and large p …
Novel Nonparametric Testing Approaches For Multivariate Growth Curve Data: Finite-Sample, Resampling And Rank-Based Methods, Ting Zeng
Theses and Dissertations--Statistics
Multivariate growth curve data naturally arise in various fields, for example, biomedical science, public health, agriculture, social science and so on. For data of this type, the classical approach is to conduct multivariate analysis of variance (MANOVA) based on Wilks' Lambda and other multivariate statistics, which require the assumptions of multivariate normality and homogeneity of within-cell covariance matrices. However, data being analyzed nowadays show marked departure from multivariate normal distribution and homoscedasticity. In this dissertation, we investigate nonparametric testing approaches for multivariate growth curve data from three aspects, i.e., finite-sample, resampling and rank-based methods.
The first project proposes an approximate …
Novel Methods For Characterizing Conditional Quantiles In Zero-Inflated Count Regression Models, Xuan Shi
Novel Methods For Characterizing Conditional Quantiles In Zero-Inflated Count Regression Models, Xuan Shi
Theses and Dissertations--Statistics
Despite its popularity in diverse disciplines, quantile regression methods are primarily designed for the continuous response setting and cannot be directly applied to the discrete (or count) response setting. There can also be challenges when modeling count responses, such as the presence of excess zero counts, formally known as zero-inflation. To address the aforementioned challenges, we propose a comprehensive model-aware strategy that synthesizes quantile regression methods with estimation of zero-inflated count regression models. Various competing computational routines are examined, while residual analysis and model selection procedures are included to validate our method. The performance of these methods is characterized through …
Estimating And Testing Treatment Effects With Misclassified Multivariate Data, Zi Ye
Estimating And Testing Treatment Effects With Misclassified Multivariate Data, Zi Ye
Theses and Dissertations--Statistics
Clinical trials are often used to assess drug efficacy and safety. Participants are sometimes pre-stratified into different groups by diagnostic tools. However, these diagnostic tools are fallible. The traditional method ignores this problem and assumes the diagnostic devices are perfect. This assumption will lead to inefficient and biased estimators. In this era of personalized medicine and measurement-based care, the issues of bias and efficiency are of paramount importance. Despite the prominence, only few researches evaluated the treatment effect in the presence of misclassifications in some special cases and most others focus on assessing the accuracy of the diagnostic devices. In …
Development And Properties Of The Roc-Abc Bayes Factor For The Quantification Of The Weight Of Forensic Evidence, Jessie Hendricks
Development And Properties Of The Roc-Abc Bayes Factor For The Quantification Of The Weight Of Forensic Evidence, Jessie Hendricks
Electronic Theses and Dissertations
Many scholars have proposed the use of a Bayes factor to quantify the weight of forensic evidence. However, due to the complex and high-dimensional nature of pattern evidence, likelihood functions are intractable and thus, Bayes factors cannot be assigned using traditional methods. Approximate Bayesian Computation (ABC) model selection algorithms provide likelihood-free methods to assign Bayes factors. ABC Bayes factors leverage the use of the scoring functions commonly used in recent years in forensic statistics in a rigorous statistical manner. However, traditional methods for assigning ABC Bayes factors are subject of several criticisms. In this dissertation, one of the main criticisms …
Uncovering Object Categories In Infant Views, Naiti S. Bhatt
Uncovering Object Categories In Infant Views, Naiti S. Bhatt
Scripps Senior Theses
While adults recognize objects in a near-instant, infants must learn how to categorize the objects in their visual environments. Recent work has shown that egocentric head-mounted camera videos contain rich data that illuminate the infant experience (Clerkin et al., 2017; Franchak et al., 2011; Yoshida & Smith, 2008). While past work has focused on the social information in view, in this work, we aim to characterize the objects in infants’ at-home visual environments by modifying modern computer vision models for the infant view. To do so, we collected manual annotations of objects that infants seemed to be interacting within a …
A Gender And Race Theoretical And Probabilistic Analysis Of The Recent Title Ix Policy Changes, Jordan Wellington
A Gender And Race Theoretical And Probabilistic Analysis Of The Recent Title Ix Policy Changes, Jordan Wellington
Scripps Senior Theses
On May 6th, 2020, after extensive public comment and review, the Department of Education published the final rule for the new Title IX regulations, which took effect in schools on August 14th. Title IX is the nearly fifty year old piece of the Education Amendments that prohibits sexual discrimination in federally funded schools. Several of these changes, such as the inclusion of live hearings and cross examination of witnesses, have been widely criticized by victims’ rights advocates for potentially retraumatizing victims of sexual assault and discouraging students from pursuing a Title IX claim. While the impact of the new regulations …
Information Prioritization: A Comparison Between Utility Maximizers And Probability Matchers, Yusuf Ismaeel
Information Prioritization: A Comparison Between Utility Maximizers And Probability Matchers, Yusuf Ismaeel
CMC Senior Theses
This thesis examines the differences between probability matchers and utility maximizers in their preferences for information sources in a lab environment. In this paper, we consider the best source of information to be the most connected one. We conducted several linear probability model type regressions along with logit regressions. Furthermore, we also attempted to control and fix any potential misclassifications in classifying the cognitive strategy by using instrumental variables. The results show that utility maximizers will almost always choose the most informed node. Probability matchers, on the other hand, do not exhibit such a behavior as the probability matching strategy …
Innovative Statistical Models In Cancer Immunotherapy Trial Design, Jing Wei
Innovative Statistical Models In Cancer Immunotherapy Trial Design, Jing Wei
Theses and Dissertations--Statistics
A challenge arising in cancer immunotherapy trial design is the presence of non-proportional hazards (NPH) patterns in survival curves. We considered three different NPH patterns caused by delayed treatment effect, cure rate and responder rate of treatment group in this dissertation. These three NPH patterns would violate the proportional hazard model assumption and ignoring any of them in an immunotherapy trial design will result in substantial loss of statistical power.
In this dissertation, four models to deal with NPH patterns are discussed. First, a piecewise proportional hazards model is proposed to incorporate delayed treatment effect into the trial design consideration. …
Immigration Offenses Throughout Federal Sentencing: An Analysis Of The Impact Of Political Affiliation Among Districts, Robin Hood
All Master's Theses
Immigration has remained one of the most controversial political debates throughout the United States. Research has yet to fully examine the effects of political affiliation of federal districts on sentencing outcomes for specific immigration offenses. To fill the gaps in research, this study compares political affiliation of federal districts among immigration offenses to determine variations in sentencing outcomes. Data included Presidential and House of Representative votes for the 2016 election and Monitoring of Federal Sentencing for the fiscal years of 2015-2016. Analysis includes case processing/legal variables, defendant characteristics, and political affiliation. To analyze political affiliation, a binary logistic regression was …
การเปรียบเทียบประสิทธิภาพของวิธีทดแทนค่าสูญหายในข้อมูลพหุระดับ: การประยุกต์ใช้กับการวิเคราะห์ความเหลื่อมล้ำทางการศึกษา, นวลรัตน์ ฉิมสุด
การเปรียบเทียบประสิทธิภาพของวิธีทดแทนค่าสูญหายในข้อมูลพหุระดับ: การประยุกต์ใช้กับการวิเคราะห์ความเหลื่อมล้ำทางการศึกษา, นวลรัตน์ ฉิมสุด
Chulalongkorn University Theses and Dissertations (Chula ETD)
การวิจัยครั้งนี้มีวัตถุประสงค์เพื่อ (1) เพื่อเปรียบเทียบประสิทธิภาพของวิธีทดแทนค่าข้อมูลสูญหาย 3 วิธี ได้แก่วิธี MI-FCS, วิธี RF และวิธี Opt.impute ซึ่งประกอบด้วย วิธี Opt.knn , Opt.tree, วิธี Opt.svm, และวิธี Opt.cv โดยใช้การจำลองข้อมูลและนำผลที่ได้มาประยุกต์ใช้กับข้อมูลจริง (2) เพื่อวิเคราะห์ความเหลื่อมล้ำทางการศึกษา ด้วยโมเดลพหุระดับโดยใช้ข้อมูลที่มีการทดแทนค่าสูญหาย และเปรียบเทียบผลที่ได้ กับการวิเคราะห์ความเหลื่อมล้ำทางการศึกษาที่ไม่ได้ทดแทนค่าสูญหาย ผลการวิจัยพบว่า (1) จากการพิจารณาผลการเปรียบเทียบประสิทธิภาพของวิธีทดแทนค่าสูญหายโดยใช้การจำลองข้อมูลในภาพรวม จะพบว่าส่วนใหญ่วิธีทดแทนค่าสูญหาย Otp.impute มีแนวโน้มให้ประสิทธิภาพสูงที่สุด รองลงมาคือ วิธีทดแทนค่าสูญหาย RF และวิธีทดแทนค่าสูญหาย MI – FCS ตามลำดับ (2) ผู้วิจัยรวบรวมข้อมูลทุติยภูมิของนักเรียนชั้นมัธยมศึกษาปีที่ 3 จากสถาบันทดสอบทางการศึกษาแห่งชาติ (สทศ.) ปีการศึกษา 2563 จำนวน 2,109 โรงเรียนที่อยู่ในสังกัดสำนักเขตพื้นที่การศึกษามัธยมศึกษา(สพม.) นำวิธีทดแทนค่าสูญหายที่ได้จากการจำลองข้อมูลมาประยุกต์ใช้กับข้อมูลทุติยภูมิดังกล่าว ผลการวิจัย จะพบว่าสัดส่วนของนักเรียนที่ครอบครัวขาดแคลนทุนทรัพย์และไม่ได้พักอาศัยอยู่กับบิดามารดาระดับโรงเรียน ส่งผลกระทบต่อผลสัมฤทธิ์ ทางการเรียนของนักเรียนระดับโรงเรียน อย่างมีนัยสำคัญทางสถิติ โดยผลกระทบที่เกิดขึ้นสะท้อนให้เห็นถึงความเหลื่อมล้ำทางการศึกษา และเมื่อเปรียบเทียบผลที่ได้กับการวิเคราะห์ความเหลื่อมล้ำทางการศึกษาที่ไม่ได้ทดแทนค่าสูญหาย แสดงให้เห็นว่าหากนำข้อมูลวิเคราะห์ผลการวิจัยโดยไม่คำนึงถึงค่าสูญหาย หรือตัดค่าสูญหายทิ้ง อาจจะส่งผลกระทบต่อการประมาณค่าพารามิเตอร์ที่แท้จริง อย่างมีนัยสำคัญทางสถิติ หรือไม่สามารถอนุมานไปสู่ประชากรได้อย่างถูกต้องและแม่นยำ
การเปรียบเทียบประสิทธิภาพของโมเดลการถดถอยเชิงลำดับชั้นที่มีอัตสหสัมพันธ์เชิงพื้นที่และโมเดลการถดถอยพหุระดับสำหรับการทำนายความอยู่ดีมีสุขของนักเรียน, ประภาพรรณ ยดย้อย
การเปรียบเทียบประสิทธิภาพของโมเดลการถดถอยเชิงลำดับชั้นที่มีอัตสหสัมพันธ์เชิงพื้นที่และโมเดลการถดถอยพหุระดับสำหรับการทำนายความอยู่ดีมีสุขของนักเรียน, ประภาพรรณ ยดย้อย
Chulalongkorn University Theses and Dissertations (Chula ETD)
ความอยู่ดีมีสุขของนักเรียนเป็นสิ่งสำคัญทางการศึกษาเชิงบวกและโรงเรียนมีบทบาทสำคัญในการสร้างเสริมให้นักเรียนทุกคนมีความอยู่ดีมีสุข การวิจัยครั้งนี้มีวัตถุประสงค์ 2 ประการ คือ (1) เพื่อวิเคราะห์ลักษณะความอยู่ดีมีสุขของนักเรียน บรรยากาศโรงเรียน และความร่วมมือระหว่างโรงเรียนจำแนกตามภูมิหลังและพื้นที่ (2) เพื่อเปรียบเทียบและวิเคราะห์ปัจจัยเชิงสาเหตุของความอยู่ดีมีสุขของนักเรียนระหว่างโมเดลการถดถอยเชิงลำดับชั้นที่มีอัตสหสัมพันธ์เชิงพื้นที่ (Hierarchical Spatial Autoregressive Model: HSAR) กับโมเดลการถดถอยพหุระดับ (Multilevel Regression Model: MLM) ด้วยวิธีการประมาณค่าแบบเบย์ (Bayesian estimation) และใช้อัลกอรึทึมการสุ่มตัวอย่างด้วยลูกโซ่มาร์คอฟมอนติคาร์โล (Markov Chain Monte Carlo) โดยใช้ข้อมูลจริงจากนักเรียน 1,981 คน และคุณครู 282 คน ของโรงเรียนในจังหวัดเชียงใหม่จำนวน 55 โรงเรียน ด้วยวิธีการสุ่มตัวอย่างแบบหลายขั้นตอน มีตัวแปรทำนายสำคัญ คือ บรรยากาศโรงเรียน และความร่วมมือระหว่างโรงเรียนซึ่งมีปฏิสัมพันธ์ข้ามระดับ (cross-level interaction term) ของความร่วมมือระหว่างโรงเรียนกับบรรยากาศโรงเรียนโดยความร่วมมือระหว่างโรงเรียนเป็นตัวแปรปรับ (moderator) และมีผลสัมฤทธิ์ทางการเรียนเป็นตัวแปรควบคุม (covariate) ผลการวิจัยพบว่า โมเดลทั้งสองมีประสิทธิภาพในการทำนายความอยู่ดีมีสุขของนักเรียนใกล้เคียงกัน (R2 MLM = 0.534, R2 HSAR = 0.529, LLMLM = -2039.6, LLHSAR = -2389.75, DICMLM = 4151.91, DICHSAR = 4955.43) แต่ให้สารสนเทศในมุมมองที่แตกต่างกัน โดยโมเดล HSAR จะให้รายละเอียดได้มากกว่าโดยเฉพาะการแสดงให้เห็นถึงอิทธิพลของความสัมพันธ์เชิงพื้นที่อย่างมีนัยสำคัญ (Lambda = 0.70 , SE = 0.30) ในขณะที่โมเดล MLM ไม่สามารถให้ผลวิเคราะห์ส่วนนี้ได้อีกทั้งยังตรวจพบอัตสหสัมพันธ์เชิงพื้นที่ในเศษเหลือของโมเดล MLM (Moran’s I = 0.09, p-value = 0.031) ซึ่งเป็นการละเมิดข้อตกลงเบื้องต้นของการวิเคราะห์ถดถอยอีกด้วย โมเดล HSAR จึงเป็นโมเดลที่เหมาะสมในการอธิบายปัจจัยเชิงสาเหตุของความอยู่ดีมีสุขของนักเรียนมากกว่า ผลการวิเคราะห์จากโมเดล HSAR …
Integrating Snp Data And Imputation Methods Into The Dna Methylation Analysis Framework, Yuqing Su
Integrating Snp Data And Imputation Methods Into The Dna Methylation Analysis Framework, Yuqing Su
Doctoral Dissertations
"DNA methylation is a widely studied epigenetic modification that can influence the expression and regulation of functional genes, especially those related to aging, cancer and other diseases. The common goal of methylation studies is to find differences in methylation levels between samples collected under different conditions. Differences can be detected at the site level, but regulated methylation targets are most commonly clustered into short regions. Thus, identifying differentially methylated regions (DMRs) between different groups is of prime interest. Despite advanced technology that enables measuring methylation genome-wide, misinterpretations in the readings can arise due to the existence of single nucleotide polymorphisms …
The Causes And Control Measures Of Extended Spectrum Beta-Lactamase Producing Enterobacteriaceae In Long-Term Care Facilities, Ismaila Olatunji Sule
The Causes And Control Measures Of Extended Spectrum Beta-Lactamase Producing Enterobacteriaceae In Long-Term Care Facilities, Ismaila Olatunji Sule
Walden Dissertations and Doctoral Studies
Due to extended-spectrum beta-lactamase-producing Enterobacteriaceae (ESBL-PE), infections among residents are increasing in long-term care facilities (LTCFs), resulting in high rate of morbidity and healthcare costs. ESBL-PE resists empirical antibiotics and reduces treatment options, and a designated infection control team is unavailable to prevent the prevalence of the disease. Ecological theory guided this study. A systematic review and meta-analysis were conducted to characterize the causes of ESBL-PE and evaluate the infection control strategies within LTCFs. Multiple regression analysis (MRA) was included as supplementary statistical analysis to identify relationships between LTCFs, geographical locations, infection control measures (ICMs), and ESBL-PE. A systematic search …
Automatic Hierarchy Expansion For Improved Structure And Chord Evaluation, Katherine M. Kinnaird, Brian Mcfee
Automatic Hierarchy Expansion For Improved Structure And Chord Evaluation, Katherine M. Kinnaird, Brian Mcfee
Statistical and Data Sciences: Faculty Publications
No abstract provided.
Predicting The Winning Percentage Of Limited-Overs Cricket Using The Pythagorean Formula, Hasika K. W. Senevirathne, Ananda B.W. Manage
Predicting The Winning Percentage Of Limited-Overs Cricket Using The Pythagorean Formula, Hasika K. W. Senevirathne, Ananda B.W. Manage
Mathematics & Statistics Faculty Publications
The Pythagorean Win-Loss formula can be effectively used to estimate winning percentages for sporting events. This formula was initially developed by baseball statistician Bill James and later was extended by other researchers to sports such as football, basketball, and ice hockey. Although one can calculate actual winning percentages based on the outcomes of played games, that approach does not take into account the margin of victory. The key benefit of the Pythagorean formula is its utilization of actual average runs scored and actual average runs allowed. This article presents the application of the Pythagorean Win-Loss formula to two different types …
Spatio-Temporal Modelling Of Tick Life-Stage Count Data With Spatially Varying Coefficients, Thabo Lephoto, Henry Mwambi, Oliver Bodhlyera, Holly Gaff
Spatio-Temporal Modelling Of Tick Life-Stage Count Data With Spatially Varying Coefficients, Thabo Lephoto, Henry Mwambi, Oliver Bodhlyera, Holly Gaff
Biological Sciences Faculty Publications
There is a vast amount of geo-referenced data in many fields of study including ecological studies. Geo-referencing is usually by point referencing; that is, latitudes and longitudes or by areal referencing, which includes districts, counties, states, provinces and other administrative units. The availability of large geo-referenced datasets for modelling has necessitated the development and application of spatial statistical methods. However, spatial varying coefficients models exploring the abundance of tick counts remain limited. In this study we used data that was collected and prepared by researchers in the Department of Biological Sciences from the Old Dominion University, Virginia, USA. We modelled …
Classification Of Chess Games: An Exploration Of Classifiers For Anomaly Detection In Chess, Masudul Hoque
Classification Of Chess Games: An Exploration Of Classifiers For Anomaly Detection In Chess, Masudul Hoque
All Graduate Theses, Dissertations, and Other Capstone Projects
Chess is a strategy board game with its inception dating back to the 15th century. The Covid-19 pandemic has led to a chess boom online with 95,853,038 chess games being played during January, 2021 on lichess.com. Along with the chess boom, instances of cheating have also become more rampant. Classifications have been used for anomaly detection in different fields and thus it is a natural idea to develop classifiers to detect cheating in chess. However, there are no specific examples of this, and it is difficult to obtain data where cheating has occurred. So, in this paper, we develop 4 …
Comparing Various Robust Estimation Techniques In Regression Analysis, Tracy S. Morrison
Comparing Various Robust Estimation Techniques In Regression Analysis, Tracy S. Morrison
All Graduate Theses, Dissertations, and Other Capstone Projects
In regression analysis, the use of the ordinary least squares (OLS) method is inadvisable when dealing with outlier or extreme observations. As a result, we require a method of robust estimation in which the estimation value is not significantly affected by outlier or extreme observations. Four methods of estimation will be compared in this paper in order to determine the best estimation: the M estimation method, the Least Trimmed Square Estimator, the S-estimation method, and the MM estimation method in robust regression. We discover that the best method is the MM-estimation method in this study. The M-estimation method is an …
Design Project: Smart Headband, John Michel, Jack Durkin, Noah Lewis
Design Project: Smart Headband, John Michel, Jack Durkin, Noah Lewis
Williams Honors College, Honors Research Projects
Concussion in sports is a prevalent medical issue. It can be difficult for medical professionals to diagnose concussions. With the fast pace nature of many sports, and the damaging effects of concussions, it is important that any concussion risks are assessed immediately. There is a growing trend of wearable technology that collects data such as steps and provides the wearer with in-depth information regarding their performance. The Smart Headband project created a wearable that can record impact data and provide the wearer with a detailed analysis on their risk of sustaining a concussion. The Smart Headband uses accelerometers and gyroscopes …
A Review Of Sample Size And Design Efficacy In Crossover Design In Peer-Reviewed Psychology Research, Kyle Moxley
A Review Of Sample Size And Design Efficacy In Crossover Design In Peer-Reviewed Psychology Research, Kyle Moxley
Wayne State University Dissertations
A REVIEW OF SAMPLE SIZE AND DESIGN EFFICACY IN CROSSOVER DESIGN IN PEER-REVIEWED PSYCHOLOGY RESEARCHby KYLE C. MOXLEY November 2021 Advisor: Dr. Shlomo S. Sawilowsky Major: Education Evaluation and Research Degree: Doctor of Philosophy The present study seeks to investigate the efficacy of crossover research designs, and the application of crossover designs, in the field of behavioral sciences. Under ideal conditions, crossover designs are assumed to be more efficacious than parallel studies in that participants are given both treatments. However, the presence of carryover effects from treatments may influence outcomes (Jones & Kenward, 2014). To prevent carryover effects, researchers frequently …
The Data Science Corps Wrangle-Analyze- Visualize Program: Building Data Acumen For Undergraduate Students, Nicholas J. Horton, Benjamin Baumer, Andrew Zieffler, Valerie Barr
The Data Science Corps Wrangle-Analyze- Visualize Program: Building Data Acumen For Undergraduate Students, Nicholas J. Horton, Benjamin Baumer, Andrew Zieffler, Valerie Barr
Statistical and Data Sciences: Faculty Publications
We congratulate Kolaczyk, Wright, and Yajima on their innovative statistics practicum that places “practice” at the center of data science education (Kolaczyk et al., 2021, this issue). Their year-long practicum course focuses on the data science life cycle with engagement with external partners and university consulting projects. We agree that training postgraduates in practice needs to be foregrounded in the curriculum in order for students to develop necessary depth in data science practice.
Toward Uncharted Territory Of Cellular Heterogeneity: Advances And Applications Of Single-Cell Rna-Seq, Brandon Lieberman, Meena Kusi, Chia Nung Hung, Chih Wei Chou, Ning He, Yen Yi Ho, Josephine A. Taverna, Tim H.M. Huang, Chun Liang Chen
Toward Uncharted Territory Of Cellular Heterogeneity: Advances And Applications Of Single-Cell Rna-Seq, Brandon Lieberman, Meena Kusi, Chia Nung Hung, Chih Wei Chou, Ning He, Yen Yi Ho, Josephine A. Taverna, Tim H.M. Huang, Chun Liang Chen
Faculty Publications
Among single-cell analysis technologies, single-cell RNA-seq (scRNA-seq) has been one of the front runners in technical inventions. Since its induction, scRNA-seq has been well received and undergone many fast-paced technical improvements in cDNA synthesis and amplification, processing and alignment of next generation sequencing reads, differentially expressed gene calling, cell clustering, subpopulation identification, and developmental trajectory prediction. scRNA-seq has been exponentially applied to study global transcriptional profiles in all cell types in humans and animal models, healthy or with diseases, including cancer. Accumulative novel subtypes and rare subpopulations have been discovered as potential underlying mechanisms of stochasticity, differentiation, proliferation, tumorigenesis, and …
Nutritional Approach For Increasing Public Health During Pandemic Of Covid-19: A Comprehensive Review Of Antiviral Nutrients And Nutraceuticals, Vahideh Ebrahimzadeh-Attari, Ghodratollah Panahi, James R. Hébert Scd, Alireza Ostadrahimi, Maryam Saghafi-Asl, Neda Lotfi-Yaghin, Behzad Baradaran
Nutritional Approach For Increasing Public Health During Pandemic Of Covid-19: A Comprehensive Review Of Antiviral Nutrients And Nutraceuticals, Vahideh Ebrahimzadeh-Attari, Ghodratollah Panahi, James R. Hébert Scd, Alireza Ostadrahimi, Maryam Saghafi-Asl, Neda Lotfi-Yaghin, Behzad Baradaran
Faculty Publications
Background: The novel coronavirus (COVID-19) is considered as the most life-threatening pandemic disease during the last decade. The individual nutritional status, though usually ignored in the management of COVID-19, plays a critical role in the immune function and pathogenesis of infection. Accordingly, the present review article aimed to report the effects of nutrients and nutraceuticals on respiratory viral infections including COVID-19, with a focus on their mechanisms of action.
Methods: Studies were identified via systematic searches of the databases including PubMed/ MEDLINE, ScienceDirect, Scopus, and Google Scholar from 2000 until April 2020, using keywords. All relevant clinical and experimental studies …
The Need To Incorporate Communities In Compartmental Models, Michael J. Kane, Owais Gilani
The Need To Incorporate Communities In Compartmental Models, Michael J. Kane, Owais Gilani
Faculty Journal Articles
Tian et al. provide a framework for assessing population- level interventions of disease outbreaks through the construction of counterfactuals in a large-scale, natural experiment assessing the efficacy of mild, but early interventions compared to delayed interventions. The technique is applied to the recent SARS-CoV-2 outbreak with the population of Shenzhen, China acting as the mild-but-early treatment group and a combination of several US counties resembling Shenzhen but enacting a delayed intervention acting as the control. To help further the development of this framework and identify an avenue for further enhancement, we focus on the use and potential limitations of compartmental …
การเปรียบเทียบประสิทธิภาพของการประมาณค่าของพารามิเตอร์ด้วยวิธีลาสโซและวิธีการคัดเลือกชุดข้อมูลย่อยที่ดีที่สุดในการวิเคราะห์การถดถอยเชิงเส้นสำหรับข้อมูลที่มีมิติสูง, วรัญญา บุตรบุรี
Chulalongkorn University Theses and Dissertations (Chula ETD)
งานวิจัยครั้งนี้มีวัตถุประสงค์เพื่อเปรียบเทียบประสิทธิภาพของวิธีการประมาณค่าพารามิเตอร์สำหรับข้อมูลที่มีมิติสูงด้วยทั้งหมด 5 วิธี ได้แก่ วิธี L0Learn, L0L2Learn, L1, A-L1 และวิธี A-L1L2 โดยการเปรียบเทียบประสิทธิภาพจะเปรียบเทียบใน 2 ด้าน คือ 1) เปรียบเทียบประสิทธิภาพด้านการพยากรณ์ ซึ่งวัดจากค่าคลาดเคลื่อนการทำนาย (MSE) และ 2) ความถูกต้องในการคัดเลือกตัวแปรอิสระเข้าสู่ตัวแบบ ซึ่งพิจารณาจากของค่า Precision Recall และค่า AUC ข้อมูลที่มีมิติสูงที่ใช้ในการศึกษาครั้งนี้ได้จากการจำลอง โดยกำหนดให้ในแต่ละชุดข้อมูลประกอบด้วยจำนวนค่าสังเกต 100 ค่าสังเกต (n = 100) และมีตัวแปรอิสระจำนวน 100 ตัว (p = 1000) โดยตัวแปรอิสระมีการแจกแจงแบบปรกติหลายตัวแปรซึ่งมีความสัมพันธ์กันแบบยกกำลัง (Exponential Correlation) 3 ระดับคือ 0, 0.5 และ 0.9 ค่าความคลาดเคลื่อนสุ่มขึ้นอยู่กับอัตราส่วนสัญญาณต่อสัญญาณรบกวน (SNR) ซึ่งมี 6 ระดับคือ 0.1, 0.5, 1, 5, 10, และ 20 โดยจำลองข้อมูลจำนวน 100 ชุดในแต่ละสถานการณ์ จากการวัดประสิทธิภาพจากค่าเฉลี่ยของข้อมูลทั้ง 100 ชุด ผลการเปรียบเทียบประสิทธิภาพด้านการพยากรณ์พบว่า เมื่อข้อมูลมีค่า SNR ต่ำและตัวแปรอิสระมีความสัมพันธ์กันน้อยถึงปานกลาง วิธี L1 จะมีประสิทธิภาพสูงที่สุด ตามด้วยวิธี L0L2Leran วิธี L0Learn วิธี A-L1L2 และวิธี A-L1 ตามลำดับ แต่เมื่อข้อมูลมีค่า SNR เพิ่มสูงขึ้นและในขณะเดียวกันตัวแปรอิสระมีความสัมพันธ์กันมากขึ้นวิธี A-L1 และวิธี A-L1L2 จะมีประสิทธิภาพสูงที่สุด ตามด้วยวิธี L1 วิธี L0L2Leran วิธี L0Learn ตามลำดับ ส่วนผลการเปรียบเทียบประสิทธิภาพด้านการคัดเลือกตัวแปรเข้าสู่ตัวแบบ เมื่อพิจารณาจากค่าเฉลี่ยของค่า Precision …
Feature Investigation For Stock Returns Prediction Using Xgboost And Deep Learning Sentiment Classification, Seungho (Samuel) Lee
Feature Investigation For Stock Returns Prediction Using Xgboost And Deep Learning Sentiment Classification, Seungho (Samuel) Lee
CMC Senior Theses
This paper attempts to quantify predictive power of social media sentiment and financial data in stock prediction by utilizing a comprehensive set of stock-related fundamental and technical variables and social media sentiments. For conducting sentiment analysis, this study employs a pretrained finBERT model that provides three different sentiment classifications and respective softmax scores. Hence, the significance of these variables is evaluated with XGBoost regression and Shapley Additive exPlanations (SHAP) frameworks. Through investigating feature importance, this study finds that statistical properties of sentiment variables provide a stronger predictive power than a weighted sentiment score and that it is possible to quantify …
Using Twitter Api To Solve The Goat Debate: Michael Jordan Vs. Lebron James, Jordan Trey Leonard
Using Twitter Api To Solve The Goat Debate: Michael Jordan Vs. Lebron James, Jordan Trey Leonard
CMC Senior Theses
Using a Twitter API, I gather and analyze tweets by performing sentiment analysis to solve the GOAT debate among professional athletes with the primary focus on comparing Michael Jordan and LeBron James. Athletes from the National Football League (NFL), the National Basketball Association (NBA), Major League Baseball (MLB), and the National Collegiate Athletic Association (NCAA) Division 1 Men's and Women's Basketball were selected to compare how sentiment polarity varies across sports. Sentiment polarity is measured by labeling text as "positive", "neutral", or "negative" which allows us to determine which athlete/sport is highly favored among the Twitter community when it comes …
An Evaluation Of Knot Placement Strategies For Spline Regression, William Klein
An Evaluation Of Knot Placement Strategies For Spline Regression, William Klein
CMC Senior Theses
Regression splines have an established value for producing quality fit at a relatively low-degree polynomial. This paper explores the implications of adopting new methods for knot selection in tandem with established methodology from the current literature. Structural features of generated datasets, as well as residuals collected from sequential iterative models are used to augment the equidistant knot selection process. From analyzing a simulated dataset and an application onto the Racial Animus dataset, I find that a B-spline basis paired with equally-spaced knots remains the best choice when data are evenly distributed, even when structural features of a dataset are known …
A Class Of Copula-Based Bivariate Poisson Time Series Models With Applications, Mohammed Alqawba, Dimuthu Fernando, Norou Diawara
A Class Of Copula-Based Bivariate Poisson Time Series Models With Applications, Mohammed Alqawba, Dimuthu Fernando, Norou Diawara
Mathematics & Statistics Faculty Publications
A class of bivariate integer-valued time series models was constructed via copula theory. Each series follows a Markov chain with the serial dependence captured using copula-based transition probabilities from the Poisson and the zero-inflated Poisson (ZIP) margins. The copula theory was also used again to capture the dependence between the two series using either the bivariate Gaussian or “t-copula” functions. Such a method provides a flexible dependence structure that allows for positive and negative correlation, as well. In addition, the use of a copula permits applying different margins with a complicated structure such as the ZIP distribution. Likelihood-based inference was …