Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Applied Statistics (2918)
- Social and Behavioral Sciences (2693)
- Medicine and Health Sciences (2514)
- Biostatistics (2512)
- Mathematics (2035)
-
- Statistical Theory (1633)
- Public Health (1563)
- Statistical Methodology (1562)
- Statistical Models (1309)
- Life Sciences (1285)
- Engineering (1179)
- Epidemiology (1097)
- Computer Sciences (1061)
- Applied Mathematics (986)
- Civil and Environmental Engineering (666)
- Materials Science and Engineering (597)
- Transportation Engineering (588)
- Other Civil and Environmental Engineering (579)
- Construction Engineering and Management (575)
- Structural Materials (575)
- Public Affairs, Public Policy and Public Administration (556)
- Medical Specialties (554)
- Data Science (550)
- Education (523)
- Multivariate Analysis (523)
- Health Services Research (502)
- Design of Experiments and Sample Surveys (500)
- Probability (493)
- Institution
-
- Wayne State University (1162)
- COBRA (1108)
- Changsha University of Science and Technology (570)
- Missouri University of Science and Technology (537)
- University of Kentucky (401)
-
- Marquette University (385)
- University of South Carolina (376)
- Universitas Indonesia (369)
- Loma Linda University (326)
- Utah State University (317)
- University of Nebraska - Lincoln (254)
- University of Nevada, Las Vegas (246)
- Old Dominion University (228)
- Wright State University (206)
- University of New Mexico (202)
- Air Force Institute of Technology (181)
- California Polytechnic State University, San Luis Obispo (180)
- University of South Florida (176)
- Virginia Commonwealth University (176)
- Prairie View A&M University (161)
- Himmelfarb Health Sciences Library, The George Washington University (159)
- Roseman University of Health Sciences (155)
- Brigham Young University (144)
- Georgia Southern University (140)
- City University of New York (CUNY) (138)
- University of Texas at El Paso (130)
- Claremont Colleges (120)
- Southern Methodist University (119)
- Western Michigan University (119)
- University of Arkansas, Fayetteville (113)
- Keyword
-
- Statistics (412)
- Machine learning (153)
- Humans (146)
- Simulation (110)
- Machine Learning (101)
-
- Bayesian (96)
- Female (91)
- Regression (91)
- COVID-19 (81)
- Male (81)
- Classification (76)
- Probability (70)
- Logistic regression (69)
- Mathematics (65)
- Reliability (63)
- Students (63)
- Survival analysis (63)
- Prediction (62)
- Missing data (61)
- Bootstrap (59)
- Epidemiology (59)
- Estimation (56)
- Empirical legal studies (55)
- Road engineering (53)
- Teachers (53)
- Education (52)
- Longitudinal data (52)
- Time series (51)
- Bias (50)
- Modeling (50)
- Publication Year
- Publication
-
- Journal of Modern Applied Statistical Methods (1093)
- Journal of China & Foreign Highway (570)
- Theses and Dissertations (565)
- Mathematics and Statistics Faculty Research & Creative Works (434)
- Kesmas (352)
-
- Loma Linda University Electronic Theses, Dissertations & Projects (326)
- Mathematics, Statistics and Computer Science Faculty Research and Publications (319)
- Faculty Publications (279)
- Electronic Theses and Dissertations (261)
- U.C. Berkeley Division of Biostatistics Working Paper Series (242)
- UW Biostatistics Working Paper Series (215)
- Harvard University Biostatistics Working Paper Series (212)
- Johns Hopkins University, Dept. of Biostatistics Working Papers (178)
- Department of Statistics: Faculty Publications (162)
- Applications and Applied Mathematics: An International Journal (AAM) (161)
- Mathematics & Statistics ETDs (159)
- Annual Research Symposium (155)
- Mathematics and Statistics Faculty Publications (152)
- USF Tampa Graduate Theses and Dissertations (136)
- All Graduate Theses and Dissertations, Spring 1920 to Summer 2023 (123)
- Dissertations (121)
- Open Access Theses & Dissertations (121)
- The University of Michigan Department of Biostatistics Working Paper Series (111)
- All Graduate Plan B and other Reports, Spring 1920 to Spring 2023 (110)
- Statistics (107)
- Epidemiology Faculty Publications (105)
- Chulalongkorn University Theses and Dissertations (Chula ETD) (99)
- International Conference on Gambling & Risk Taking (94)
- COBRA Preprint Series (91)
- Graduate Theses and Dissertations (84)
- Publication Type
Articles 2581 - 2610 of 12804
Full-Text Articles in Statistics and Probability
Wilcoxon-Mann-Whitney Effects For Clustered Data: Informative Cluster Size, Changrui Liu
Wilcoxon-Mann-Whitney Effects For Clustered Data: Informative Cluster Size, Changrui Liu
Theses and Dissertations--Statistics
In recent research, there has been a growing interest in understanding the impact of informative cluster size (ICS) on statistical inference for clustered data. In the non-parametric context, the problem for testing equality of distribution functions has been the main consideration. We are aiming to develop inferential procedures for the Wilcoxon-Mann-Whitney effect, also known as the non-parametric relative effect, involving two or more groups. Computationally, results from both simulated and real-world data have shown promising results that our proposed tests effectively account for ICS and they particularly outperform other methods in the literature designed for ignorable cluster sizes. The applications …
An Assessment Of "Long-Thin" Airline Routes: Network Structure And Emissions Implications For Environmental Policy, Porter Burns
An Assessment Of "Long-Thin" Airline Routes: Network Structure And Emissions Implications For Environmental Policy, Porter Burns
All Master's Theses
The purpose of this research was to define, map, and quantify the network and environmental implications of “long-thin” routes (LTRs) – a route structure that has been discussed in the aviation industry but not formally studied in literature. LTRs were defined through the use of global OAG scheduling data from 1998 to 2018 to identify trends in air traffic growth and network dynamics. Flights were separated into seven aircraft class sizes (e.g., 75–150 seats, 150–225 seats) to measure LTRs at multiple scales. Routes were considered “long” if the stage length was at or above the 75th percentile in each …
Recurrent Event Data Analysis With Mismeasured Covariates, Ravinath Alahakoon Mudiyanselage
Recurrent Event Data Analysis With Mismeasured Covariates, Ravinath Alahakoon Mudiyanselage
Doctoral Dissertations
"Consider a study with n units wherein every unit is monitored for the occurrence of an event that can recur with random end of monitoring. At each recurrence, p concomitant variables associated to the event recurrence are recorded with q (q ≤ p) collected with errors. Of interest in this dissertation is the estimation of the regression parameters of event time regression models accounting for the covariates. To circumvent the problem of bias and consistency associated with model's parameter estimation in the presence of measurement errors, we propose inference for corrected estimating functions with well-behaved roots under additive measurement errors …
Expectile Neural Networks For Genetic Data Analysis Of Complex Diseases, Jinghang Lin, Xiaoran Tong, Chenxi Li, Qing Lu
Expectile Neural Networks For Genetic Data Analysis Of Complex Diseases, Jinghang Lin, Xiaoran Tong, Chenxi Li, Qing Lu
Biostatistics Faculty Publications
The genetic etiologies of common diseases are highly complex and heterogeneous. Classic methods, such as linear regression, have successfully identified numerous variants associated with complex diseases. Nonetheless, for most diseases, the identified variants only account for a small proportion of heritability. Challenges remain to discover additional variants contributing to complex diseases. Expectile regression is a generalization of linear regression and provides complete information on the conditional distribution of a phenotype of interest. While expectile regression has many nice properties, it has rarely been used in genetic research. In this paper, we develop an expectile neural network (ENN) method for genetic …
Joint Probability Analysis Of Extreme Precipitation And Water Level For Chicago, Illinois, Anna Li Holey
Joint Probability Analysis Of Extreme Precipitation And Water Level For Chicago, Illinois, Anna Li Holey
Dissertations, Master's Theses and Master's Reports
A compound flooding event occurs when there is a combination of two or more extreme factors that happen simultaneously or in quick succession and can lead to flooding. In the Great Lakes region, it is common for a compound flooding event to occur with a high lake water level and heavy rainfall. With the potential of increasing water levels and an increase in precipitation under climate change, the Great Lakes coastal regions could be at risk for more frequent and severe flooding. The City of Chicago which is located on Lake Michigan has a high population and dense infrastructure and …
Evaluation Of Edison's Data Science Competency Framework Through A Comparative Literature Analysis, Karl R. B. Schmitt, Linda Clark, Katherine M. Kinnaird, Ruth E. H. Wertz, Björn Sandstede
Evaluation Of Edison's Data Science Competency Framework Through A Comparative Literature Analysis, Karl R. B. Schmitt, Linda Clark, Katherine M. Kinnaird, Ruth E. H. Wertz, Björn Sandstede
Statistical and Data Sciences: Faculty Publications
During the emergence of Data Science as a distinct discipline, discussions of what exactly constitutes Data Science have been a source of contention, with no clear resolution. These disagreements have been exacerbated by the lack of a clear single disciplinary 'parent.' Many early efforts at defining curricula and courses exist, with the EDISON Project's Data Science Framework (EDISON-DSF) from the European Union being the most complete. The EDISON-DSF includes both a Data Science Body of Knowledge (DS-BoK) and Competency Framework (CF-DS). This paper takes a critical look at how EDISON's CF-DS compares to recent work and other published curricular or …
Deep Learning-Based Technique For The Perception Of The Cervical Cancer, Aya Haraz, Hossam El-Din Moustafa, Abeer Twakol Khaleel, Ahmed H. Eltanboly
Deep Learning-Based Technique For The Perception Of The Cervical Cancer, Aya Haraz, Hossam El-Din Moustafa, Abeer Twakol Khaleel, Ahmed H. Eltanboly
Mansoura Engineering Journal
In third-world countries, cervical cancer is the most prevalent and leading cause of death. It is affected by a variety of factors, including smoking, poor nutritional status, immunological inadequacy, and prolonged use of contraception. The Pap smear test, which is intended to prevent cervical cancer, finds preneoplastic changes in cervical epithelial cells. This study framework classified cervical cancer cells from Pap smears into five specified cell types using machine learning-based classification algorithms. The SIPaKMeD database is used in this investigation. This public dataset, which was manually cropped from 966 cluster cell images taken from Pap smear slides, has 4045 isolated …
Novel Feature Evaluation In Ultra-High Dimensional Right-Censored Data, With Applications To Head And Neck Cancer, Atika Farzana Urmi
Novel Feature Evaluation In Ultra-High Dimensional Right-Censored Data, With Applications To Head And Neck Cancer, Atika Farzana Urmi
Graduate Research Posters
Background: Head and neck cancer is the 6th most common cancer worldwide with an expected 1.08 million new cases each year. Such cancer data are ultra-high dimensional with thousands of clinical features and gene expressions, making it challenging for the traditional analytical tools to extract the potential biomarker for the cancer survival and control false discoveries. In addition, presence of heavy censoring can affect the screening procedures based on Kaplan-Meier (K-M) survival estimates.
Aim: To propose a model free, ultra-high dimensional feature screening method with two-dimensional survival outcome allowing false discovery rate (FDR) control.
Method: 516 primary tumor patients with …
Automated Machine Learning: Intellient Binning Data Preparation And Regularized Regression Classfier, Jianbin Zhu
Automated Machine Learning: Intellient Binning Data Preparation And Regularized Regression Classfier, Jianbin Zhu
Electronic Theses and Dissertations, 2020-2023
Automated machine learning (AutoML) has become a new trend which is the process of automating the complete pipeline from the raw dataset to the development of machine learning model. It not only can relief data scientists' works but also allows non-experts to finish the jobs without solid knowledge and understanding of statistical inference and machine learning. One limitation of AutoML framework is the data quality differs significantly batch by batch. Consequently, fitted model quality for some batches of data can be very poor due to distribution shift for some numerical predictors. In this dissertation, we develop an intelligent binning to …
Meta-Analysis Of Scent Detection Canines And Potential Factors Influencing Their Success Rates, Molly Marie Jaskinia
Meta-Analysis Of Scent Detection Canines And Potential Factors Influencing Their Success Rates, Molly Marie Jaskinia
Graduate Student Theses, Dissertations, & Professional Papers
Objective: This is a meta-analysis focused on the success rates of scent detection canines and potential factors that could influence their accuracy. A series of statistical analyses were conducted to determine if certain demographic factors, such as the dog’s gender, age, and breed, have an effect on a scent dog’s accuracy during a search. Or if more circumstantial factors, like the dog’s level of experience in scent work, the type of target scent, and their handler’s awareness of the target’s location, affect the outcome of the search.
Materials and Methods: A dataset was created from 37 different articles consisting of …
Applications Of Transfer Learning From Malicious To Vulnerable Binaries, Sean Patrick Mcnulty
Applications Of Transfer Learning From Malicious To Vulnerable Binaries, Sean Patrick Mcnulty
Graduate Student Theses, Dissertations, & Professional Papers
Malware detection and vulnerability detection are important cybersecurity tasks. Previous research has successfully applied a variety of machine learning methods to both. However, despite their potential synergies, previous research has yet to unite these two tasks. Given the recent success of transfer learning in many domains, such as language modeling and image recognition, this thesis investigated the use of transfer learning to improve vulnerability detection. Specifically, we pre-trained a series of models to detect malicious binaries and used the weights from those models to kickstart the detection of vulnerable binaries. In our study, we also investigated five different data representations …
Poisson Regression Model With Application To Wastewater Surveillance Under A Threshold Linear Mixed Model For Covid-19 Sensitivity Rates, Norou Diawara, Hueiwang Anna Jeng, Kyle Curtis, Raul Gonzalez, Nancy Welch, Cynthia Jackson, Rekha Singh, David Jurgens, Sasanka Adikari, Omotomilola Jegede
Poisson Regression Model With Application To Wastewater Surveillance Under A Threshold Linear Mixed Model For Covid-19 Sensitivity Rates, Norou Diawara, Hueiwang Anna Jeng, Kyle Curtis, Raul Gonzalez, Nancy Welch, Cynthia Jackson, Rekha Singh, David Jurgens, Sasanka Adikari, Omotomilola Jegede
Mathematics & Statistics Faculty Publications
A Threshold Linear Mixed Model (TLMM) has been developed to identify specific thresholds based on wastewater SARS-CoV-2 viral concentrations, which reflect COVID-19 cases. The thresholds can guide decisions regarding public health responses and prevention measures. To assess the practical application of TLMM, a simple simulation was conducted using a sample size of 100 and 500 replications. The simulation allowed for comparing parameter estimators by assessing bias and standard deviation and the root of the mean square error. The model and estimation procedures were applied to reported wastewater and clinic data to test its application for real-world scenarios. Our results demonstrated …
Utilizing Markov Chains To Estimate Allele Progression Through Generations, Ronit Gandhi
Utilizing Markov Chains To Estimate Allele Progression Through Generations, Ronit Gandhi
Honors Program: Senior Projects (Public)
All populations display patterns in allele frequencies over time. Some alleles cease to exist, while some grow to become the norm. These frequencies can shift or stay constant based on the conditions the population lives in. If in Hardy-Weinberg equilibrium, the allele frequencies stay constant. Most populations, however, have bias from environmental factors, sexual preferences, other organisms, etc. We propose a stochastic Markov chain model to study allele progression across generations. In such a model, the allele frequencies in the next generation depend only on the frequencies in the current one.
We use this model to track a recessive allele …
Evaluating Ai Sentiment Analysis, Aakriti Shah
Evaluating Ai Sentiment Analysis, Aakriti Shah
Honors Program Theses
This paper presents a comparative analysis of human and AI performance on a sentiment analysis task involving the coding of qualitative data from community program transcripts. The results demonstrate promising but imperfect agreement between two AI models, Claude and Bing, versus three human annotators and one expert annotator using the Community Capitals framework categories. While both models achieved fair alignment with human judgment, confusion patterns emerged involving metaphorical language and text overlapping multiple categories. The findings provide a case study for benchmarking conversational AI systems against human baselines to reveal limitations and target improvements. Key gaps center around distinguishing between …
A Qualitative Analysis Of Construct Measurement Techniques Used In Industrial/Organizational Research, Benjamin Michael, Andrea F. Snell, Katie Rosneck
A Qualitative Analysis Of Construct Measurement Techniques Used In Industrial/Organizational Research, Benjamin Michael, Andrea F. Snell, Katie Rosneck
Williams Honors College, Honors Research Projects
This project aims to challenge the appropriateness of the methodological strategies and tools utilized within psychological research. We will look at the types of statistical modeling used and the context in which they are used, such as measurement modeling, confirmatory factor analysis, and bifactor analysis within survey development, as well as the use of psychological constructs such as extraversion and leadership. The objective of this research is to search for and recognize patterns from the content of some of the top journal articles in the field of industrial and organizational psychology. The information gained from analyzing the content of the …
Applications For Functional Data Analysis, Kacy D. Kane
Applications For Functional Data Analysis, Kacy D. Kane
Graduate Research Theses & Dissertations
Functional Data Analysis is often used in the study of data that exists over a continuum, such as time. There are two datasets that will be considered here. For the first study we have a dataset on the efficacy of a lobectomy in reduction or elimination of epileptic seizures in patients. After an initial analysis of the dataset from a multinomial model perspective, we found that there were outliers in our dataset. From there, we considered a Multinomial Mixture Model to aid in the detection of outliers. In our second dataset we are considering a social robotics dataset where the …
Macroeconomic Factors Influencing Foreign Direct Investment In Some Selected Countries In Africa, Richard Essel Mensah
Macroeconomic Factors Influencing Foreign Direct Investment In Some Selected Countries In Africa, Richard Essel Mensah
Graduate Research Theses & Dissertations
This paper investigates the possible factors that influence foreign direct investment inflow rate to Africa after controlling for other macroeconomic factors. Using the heterogenous Toeplitz mixed method on a sample of 23 countries from 1998 – 2020, we find evidence of the statistical significance of a relationship between the amount of trade done in Africa and the FDI inflow rate in Africa. We also find a statistical relationship between the labor force participation rate and the FDI inflow rate to Africa. Although the Fixed effect and GLM method did not find the relationship between LFP rate and FDI inflow to …
Impacts Of Covid-19 On Industrial Growth In The United States, Emily G. Warthman, Charles J. Landis
Impacts Of Covid-19 On Industrial Growth In The United States, Emily G. Warthman, Charles J. Landis
Williams Honors College, Honors Research Projects
COVID-19 has caused massive ramifications on all parts of life in the world and industry growth/decline is not immune to it. This report will analyze nine different industries’ profit and revenue from quarterly data during the years 2009-2022. Forecast models will be generated using various methods and different techniques of validating to predict the values from Q2 2020- Q4 2022 based on historical data. After which, a comparison will be conducted between those predicted values to the actual average revenue and profit generated by order of greatest error percentage made. Thorough research will then be completed to determine if there …
Modeling The Bidirectional Relationship Between Shared-Patient Physician Networks And Patient Longitudinal Treatment Patterns: Application To Physician Risky-Prescribing, Xin Ran
Dartmouth College Ph.D Dissertations
Risky-prescribing is a pressing public health concern in the United States. Opioids, benzodiazepines, and non-benzodiazepine sedative-hypnotics (sedative-hypnotics) are three commonly-prescribed but potentially risky drug groups, prescribed alone or in combination. Physician shared-patient networks provide a unique perspective in studying physician network characteristics and structures, as well as their association with the delivery of health care. Understanding how physician shared-patient networks are related to their prescribing may inform network-based interventions targeting risky-prescribing, which is yet to be fully studied.
We investigated patient receipt of risky prescriptions and physician risky-prescribing intensity through the scope of shared-patient networks. We used retrospective Medicare insurance …
การเปรียบเทียบวิธีการใส่ค่าสูญหาย ในการวิเคราะห์การถดถอยโลจิสติก เมื่อตัวแปรตามมีการสูญหายแบบนอนอิกนอร์เรเบิล, อภิชาติ ฉัตรเรืองเลิศ
การเปรียบเทียบวิธีการใส่ค่าสูญหาย ในการวิเคราะห์การถดถอยโลจิสติก เมื่อตัวแปรตามมีการสูญหายแบบนอนอิกนอร์เรเบิล, อภิชาติ ฉัตรเรืองเลิศ
Chulalongkorn University Theses and Dissertations (Chula ETD)
งานวิจัยนี้มีวัตถุประสงค์เพื่อเปรียบเทียบวิธีการใส่ค่าสูญหาย ในการวิเคราะห์การถดถอยโลจิสติก เมื่อตัวแปรตามมีการสูญหายแบบนอนอิกนอร์เรเบิล วิธีการที่ใช้ศึกษา คือ วิธี Complete Case Analysis (CC) วิธี Mode Imputation (MODE) วิธี Expectation Maximization Algorithm (EM) วิธี Multiple Imputation (MI) วิธี Hard Cutoff Augmentation (HARDCUT) วิธี Parceling Augmentation (PARCELING) และวิธี Fuzzy Augmentation (FUZZY) งานวิจัยนี้ใช้การจำลองข้อมูลในการศึกษาตามขนาดของตัวอย่าง ร้อยละของการสูญหายของข้อมูล และระดับของการสูญหายแบบนอนอิกนอร์เรเบิล การจำลองข้อมูลในแต่ละสถานการณ์จะกระทำ 5,000 รอบ โดยมีเกณฑ์ที่ใช้เปรียบเทียบประสิทธิภาพของวิธีการใส่ค่าสูญหาย ได้แก่ ค่าเฉลี่ยของค่าเฉลี่ยความคลาดเคลื่อนกำลังสอง (Average Mean Squared Error: AMSE) ของค่าประมาณความน่าจะเป็นของการเกิดเหตุการณ์ที่สนใจ (P(Y = 1)) และค่าประสิทธิภาพสัมพัทธ์ (Relative Efficiency: RE) จากผลการทดลองสรุปได้ว่า ค่า AMSE จะลดลงเมื่อขนาดของตัวอย่างใหญ่ขึ้น และจะมีค่ามากขึ้นเมื่อร้อยละของการสูญหายของข้อมูลเพิ่มขึ้น เมื่อพิจารณาผลของระดับของการสูญหายแบบนอนอิกนอร์เรเบิลต่อค่า AMSE พบว่ามีเพียง AMSE ของวิธี MODE เท่านั้นที่มีแนวโน้มเพิ่มขึ้น เมื่อระดับของการสูญหายแบบนอนอิกนอร์เรเบิลเพิ่มขึ้น และเมื่อพิจารณาค่า RE โดยเปรียบเทียบ AMSE ของวิธี CC กับวิธีการใส่ค่าสูญหายวิธีอื่น พบว่า วิธี EM และวิธี FUZZY ให้ค่า AMSE เท่ากับ AMSE ของวิธี CC ในขณะที่ AMSE ของวิธีอื่น ๆ มีค่าน้อยกว่า AMSE ของวิธี CC
Exploring Information Leakage In Historical Stock Market Data, Edison Hua
Exploring Information Leakage In Historical Stock Market Data, Edison Hua
Dissertations and Theses
Information leakage is a major concern for traders who want to execute large orders without affecting the market price. In this paper, we explore the sources and effects of information leakage in historical stock market data using various methods and metrics. We first define information leakage as a pattern caused by a trader that would otherwise not occur without the trader’s activity. Using historical data, the direct impact of a potential large trade cannot be measured, but we consider a minimal impact large trade to be one that minimizes changes to the established trading data. We then analyze how information …
Shallow Water Coral Distribution And Its Response To Climate Change, Amaury De Jesus
Shallow Water Coral Distribution And Its Response To Climate Change, Amaury De Jesus
Dissertations and Theses
Shallow water corals are one of the main reef-building organisms that secrete carbonates as their skeletons, and therefore, are one of the major sinks of CO2 in the ocean. These reef builders are also very crucial to marine environments and human society. As the global energy demand continues rising, fossil fuel burning increases at a faster pace despite the increase in energy supply using clean and renewable energy. The increase of CO2 in the atmosphere has been shown to exacerbate global warming and may cause ocean acidification, threatening the habitat of shallow-water corals. Many recent observations show alarming signs of …
Application Of Sentiment Analysis And Machine Learning Techniques To Predict Daily Cryptocurrency Price Returns, Edward Wu
CMC Senior Theses
This paper examines the effects of social media sentiment relating to Bitcoin on the daily price returns of Bitcoin and other popular cryptocurrencies by utilizing sentiment analysis and machine learning techniques to predict daily price returns. Many investors think that social media sentiment affects cryptocurrency prices. However, the results of this paper find that social media sentiment relating to Bitcoin does not add significant predictive value to forecasting daily price returns for each of the six cryptocurrencies used for analysis and that machine learning models that do not assume linearity between the current day price return and previous daily price …
Modeling Growth And Stress Factors For Converted Silvopasture Systems In The Missouri Ozarks, Bailee N. Suedmeyer
Modeling Growth And Stress Factors For Converted Silvopasture Systems In The Missouri Ozarks, Bailee N. Suedmeyer
Graduate Theses/Dissertations
Silvopasture systems are becoming increasingly popular among sustainable agriculture ranchers, due to the increase in knowledge of benefits to the cattle and ability to grow cool season grasses beneath the canopy. This project focuses on the forest crop aspect of silvopasture systems from monitoring of the health of the trees over time to recommendations for thinning management to keep it functioning as viable silvopasture. The study site consists of five acres of upland hardwood forest area in Southern Missouri with 18 monumented fixed area plots. Arial and ground data was collected at each plot throughout the growing season, along with …
The Influence Of Instrumental Sources Of Variance On Mass Spectral Comparison Algorithms, Isabel Cristina Galvez Valencia
The Influence Of Instrumental Sources Of Variance On Mass Spectral Comparison Algorithms, Isabel Cristina Galvez Valencia
Graduate Theses, Dissertations, and Problem Reports (ETD)
Current search algorithms for the identification of substances based only on their electron ionization mass spectra provide the correct compound as their top result approximately 80% of the time. One contributing factor to the ~20% deviation in the first-hit recognition rate is that traditional algorithms work by comparing the unknown spectrum to an ‘ideal’ or consensus spectrum of each reference compound. The inclusion of replicate reference spectra in a database has been shown to improve the probability of ranking the correct identity in the number one position, but the variance in ion abundances caused by different conditions or different instruments …
Investigating Collaborative Explainable Ai (Cxai)/Social Forum As An Explainable Ai (Xai) Method In Autonomous Driving (Ad), Tauseef Ibne Mamun
Investigating Collaborative Explainable Ai (Cxai)/Social Forum As An Explainable Ai (Xai) Method In Autonomous Driving (Ad), Tauseef Ibne Mamun
Dissertations, Master's Theses and Master's Reports
Explainable AI (XAI) systems primarily focus on algorithms, integrating additional information into AI decisions and classifications to enhance user or developer comprehension of the system's behavior. These systems often incorporate untested concepts of explainability, lacking grounding in the cognitive and educational psychology literature (S. T. Mueller et al., 2021). Consequently, their effectiveness may be limited, as they may address problems that real users don't encounter or provide information that users do not seek.
In contrast, an alternative approach called Collaborative XAI (CXAI), as proposed by S. Mueller et al (2021), emphasizes generating explanations without relying solely on algorithms. CXAI centers …
Additive P-Value Combination Test, Xing Ling
Additive P-Value Combination Test, Xing Ling
Dissertations, Master's Theses and Master's Reports
This dissertation includes four Chapters. A brief description of each chapter is organized as follows.
In Chapter 1, some developments on multiple hypotheses tests are introduced. Some preliminaries about the definition and the assumption are included.
In Chapter 2, a Stable Combination Test is proposed to combine $p$-values from multiple hypotheses tests. We show the proposed method controls the family-wise error rate at the target level and maintains asymptotically optimal power even when the elementary p-values from the individual hypotheses are dependent.
In Chapter 3, a deeper dig into the additive p-value combination test is performed. A common idea behind …
Variability In Causal Effects On A Binary Outcome And Noncompliance In A Multisite Randomized Trial, Xinxin Sun
Variability In Causal Effects On A Binary Outcome And Noncompliance In A Multisite Randomized Trial, Xinxin Sun
Theses and Dissertations
Noncompliance to treatment assignment is widespread in randomized trials and presents challenges in causal inference. In the presence of noncompliance, the most commonly estimated effect of treatment assignment, also known as intent-to-treat (ITT) effect, is biased. Of interest in this setting is the complier average causal effect (CACE), the ITT effect among compliers. Further complication arises when the outcome variable is partially observed.
My research focuses on estimating the distribution of a site-specific CACE in a multisite randomized controlled trial (MRCT) by maximum likelihood (ML). Assuming compliance missing at random (MAR). We express the likelihood as an integral with respect …
Reassessing Replication: Addressing The Replication Crisis From A Statistical Perspective, Alicia Richards Phd
Reassessing Replication: Addressing The Replication Crisis From A Statistical Perspective, Alicia Richards Phd
Theses and Dissertations
In 2015, Open Science Framework directly replicated 100 psychology studies and found astonishingly low replication rates. Since, researchers have suggested factors that may have influenced the low rates, including the metrics used to assess replications. The definitions used to decide whether a replication study was successful all suffer from flaws. Therefore, we propose a new metric for assessing replication that can estimate the likelihood a study successfully replicated rather than forcing a binary choice and accounts for study design limitations.
Using equivalence study techniques, we first propose a new metric to assess replication, defining a successful replication as one where …
Early Termination In Phase Ii Clinical Trials: Admissible Designs Using Decreasingly Informative Priors, Chen Wang
Theses and Dissertations
In Phase II clinical trials, Thall and Simon’s Bayesian posterior probability design is commonly implemented to allow for an early termination to determine whether a new treatment warrants further investigation in a larger-scale Phase III trial; this in turn requires a pre-selected prior distribution based on known clinical opinion or historical information. Moreover, this Bayesian approach can result in an issue of inflating type I error rate by monitoring interim data to inform early termination decisions. Alternatively, a Bayesian approach with the decreasingly informative prior (DIP), which is an informative yet skeptical prior, can be implemented to overcome the contentious …