Open Access. Powered by Scholars. Published by Universities.®

Statistics and Probability Commons™

Open Access. Powered by Scholars. Published by Universities.®

Discipline
Institution
Keyword
Publication Year
Publication
Publication Type
File Type

Articles 1711 - 1740 of 12804

Full-Text Articles in Statistics and Probability

ประสิทธิภาพการพยากรณ์ของการวิเคราะห์เชิงฟังก์ชันและการเรียนรู้เชิงลึก, บุณฑริกา พรหมสถิตย์ Jan 2024

ประสิทธิภาพการพยากรณ์ของการวิเคราะห์เชิงฟังก์ชันและการเรียนรู้เชิงลึก, บุณฑริกา พรหมสถิตย์

Chulalongkorn University Theses and Dissertations (Chula ETD)

การวิจัยนี้เปรียบเทียบประสิทธิภาพการพยากรณ์ของตัวแบบ Functional Principal Component Regression (FPCR) กับตัวแบบการเรียนรู้เชิงลึก (Recursive Neural Network: RNN, Long Short-term Memory: LSTM, Gated Recurrent Unit: GRU) ภายใต้สถานการณ์ที่มีความผันผวนแตกต่างกัน โดยใช้ชุดข้อมูลที่มีความผันผวนต่ำ (อุณหภูมิเฉลี่ยรายวัน), ปานกลาง (ปริมาณ PM 2.5 รายชั่วโมง) และสูง (อัตราการแลกเปลี่ยน Bitcoin รายนาที) ศึกษาการพยากรณ์ระยะสั้น ระยะกลาง และระยะยาว โดยใช้ตัวชี้วัด Mean Squared Error (MSE) และ Mean Integrated Squared Error (MISE) ผลการศึกษาพบว่า สำหรับข้อมูลที่มีความผันผวนต่ำ FPCR ให้ผลลัพธ์ที่แม่นยำกว่าตัวแบบการเรียนรู้เชิงลึก โดยเฉพาะในการพยากรณ์ระยะกลางและระยะยาว ในทางกลับกัน สำหรับข้อมูลที่มีความผันผวนสูง FPCR เหนือกว่าการเรียนรู้เชิงลึกเฉพาะการพยากรณ์ระยะกลางเท่านั้น สำหรับข้อมูลที่มีความผันผวนปานกลางและสูง ขนาดของชุดข้อมูลฝึกไม่มีผลกระทบอย่างชัดเจนต่อประสิทธิภาพของตัวแบบทั้งสอง อย่างไรก็ตาม ในกรณีของข้อมูลที่มีความผันผวนต่ำ เมื่อมีชุดข้อมูลขนาดใหญ่ ตัวแบบการเรียนรู้เชิงลึกให้ผลลัพธ์ที่แม่นยำกว่า FPCR ในขณะที่ FPCR มีความแม่นยำสูงกว่าหากใช้ข้อมูลจำนวนน้อยและทำการพยากรณ์ในระยะกลางถึงระยะยาว นอกจากนี้ ในการพยากรณ์ระยะสั้น FPCR มักให้ผลลัพธ์ที่ด้อยกว่าตัวแบบการเรียนรู้เชิงลึกในทุกกรณี


การเปรียบเทียบวิธีการใส่ค่าสูญหายสำหรับอนุกรมเวลาเชิงพหุ กรณีศึกษาดัชนีราคากลุ่มอุตสาหกรรมตลาดหลักทรัพย์แห่งประเทศไทย, พงษ์พล ยิ่งประทานพร Jan 2024

การเปรียบเทียบวิธีการใส่ค่าสูญหายสำหรับอนุกรมเวลาเชิงพหุ กรณีศึกษาดัชนีราคากลุ่มอุตสาหกรรมตลาดหลักทรัพย์แห่งประเทศไทย, พงษ์พล ยิ่งประทานพร

Chulalongkorn University Theses and Dissertations (Chula ETD)

การศึกษานี้มีวัตถุประสงค์เพื่อเปรียบเทียบวิธีการใส่ค่าสูญหายสำหรับอนุกรมเวลาเชิงพหุ และประเมินผลเพื่อเลือกวิธีการใส่ค่าสูญหายที่เหมาะสมที่สุดสำหรับอนุกรมเวลาเชิงพหุ โดยใช้ข้อมูลทุติยภูมิดัชนีราคากลุ่มอุตสาหกรรมของตลาดหลักทรัพย์แห่งประเทศไทย 8 กลุ่มอุตสาหกรรม จากฐานข้อมูล SETSMART ตั้งแต่วันที่ 1 มกราคม พ.ศ. 2547 ถึง 1 มกราคม พ.ศ. 2567 รวมทั้งสิ้น 4877 วัน ซึ่งได้มีการกำหนดรูปแบบการสูญหายออกเป็น 3 รูปแบบ ได้แก่ การสูญหายรูปแบบสุ่ม การสูญหายรูปแบบช่วงตามลำดับ และการสูญหายรูปแบบบล็อก และกำหนดสัดส่วนการสูญหายของข้อมูลที่ร้อยละ 5 10 20 30 40 และ 50 ตามลำดับ โดยจำแนกวิธีการใส่ค่าสูญหายออกเป็น 3 กลุ่ม ได้แก่ การใส่ค่าสูญหายด้วยวิธีการเชิงสถิติ ประกอบไปด้วย ค่าเฉลี่ย ค่ามัธยฐาน ข้อมูลสุดท้ายก่อนการสูญหาย (LOCF) ข้อมูลล่าสุดหลังการสูญหาย (NOCB) และการประมาณค่าช่วงเส้นตรง (Linear Interpolation) การใส่ค่าสูญหายด้วยวิธีการเรียนรู้ของเครื่อง ประกอบไปด้วย ค่าคาดหวังสูงที่สุด (EM) การใส่ค่าสูญหายด้วยการทดแทนแบบพหุคูณด้วยสมการลูกโซ่ (MICE) เพื่อนบ้านใกล้เคียงที่สุด (KNN) และป่าสุ่ม (Random Forest) และการใส่ค่าสูญหายด้วยวิธีการเรียนรู้เชิงลึก ประกอบไปด้วย GP-VAE USGAN และ SAITS นอกจากนี้ผู้วิจัยใช้ค่ารากที่สองของค่าความคลาดเคลื่อนกำลังสองโดยเฉลี่ย (RMSE) ค่าความคลาดเคลื่อนสัมบูรณ์โดยเฉลี่ย (MAE) และค่าร้อยละความคลาดเคลื่อนสัมบูรณ์โดยเฉลี่ย (MAPE) ในการวัดประสิทธิภาพการใส่ค่าสูญหาย ผลการศึกษาพบว่าที่รูปแบบการสูญหายทั้ง 3 รูปแบบ และสัดส่วนการสูญหายที่น้อยกว่าร้อยละ 50 การใสค่าสูญหายด้วยการประมาณค่าช่วงเส้นตรง (Linear Interpolation) มีประสิทธิภาพสูงที่สุด ในขณะที่สัดส่วนการสูญหายร้อยละ 50 ของรูปแบบการสูญหายทั้ง 3 รูปแบบ การใส่ค่าสูญหายด้วยวิธีป่าสุ่ม (Random Forest) มีประสิทธิภาพสูงที่สุด


ประสิทธิภาพของแบบจำลองแบบผสมของการเรียนรู้เชิงลึกสำหรับการพยากรณ์ราคาหุ้น, กิตติคุณ ทัดประดิษฐ Jan 2024

ประสิทธิภาพของแบบจำลองแบบผสมของการเรียนรู้เชิงลึกสำหรับการพยากรณ์ราคาหุ้น, กิตติคุณ ทัดประดิษฐ

Chulalongkorn University Theses and Dissertations (Chula ETD)

งานวิจัยนี้มีวัตถุประสงค์เพื่อวิเคราะห์และเปรียบเทียบประสิทธิภาพและระยะเวลาที่ใช้ของแบบจำลองการเรียนรู้เชิงลึกแบบผสม (Hybrid Deep Learning Models) ได้แก่ RNN, LSTM และ GRU ในรูปแบบการใช้โครงสร้างแบบผสม (Stacked layer) และเครือข่ายประสาทเทียมแบบซ้อนกัน (Cascaded neural network) ในการพยากรณ์ราคาปิดหุ้น รวมถึงศึกษาผลกระทบของการสลับลำดับของแบบจำลองภายใน Hybrid model เพื่อพิจารณาความแตกต่างของประสิทธิภาพการพยากรณ์ ทั้งในระยะสั้น 7 วัน และ ระยะยาว 30 วัน โดยมีการใช้ข้อมูลราคาหุ้นทั้งหมด 5 อุตสาหกรรม เลือกกลุ่มอุตสาหกรรมละ 3 หุ้นตามระดับความผันผวนของราคาหุ้นเมื่อเทียบกับตลาด (Beta) รวมทั้งหมด 15 ชุดข้อมูล ผลการศึกษาพบว่า ในภาพรวมการพยากรณ์ระยะสั้น 7 วันและระยะยาว 30 วัน การสร้างแบบจำลองผสมด้วยวิธี Cascaded neural network มีประสิทธิภาพที่ไม่แตกต่างกับวิธี Stacked layers อย่างมีนัยสำคัญ แต่ถ้าหากพิจารณาด้วยระยะเวลาที่ใช้ในการสร้างแบบจำลองจะพบว่า การสร้างแบบจำลองผสมด้วยวิธี Stacked layers ใช้เวลาสร้างแบบจำลองน้อยกว่าวิธี Cascaded neural network อย่างมีนัยสำคัญ ซึ่งจะช่วยประหยัดทรัพยากรในการสร้างแบบจำลองในการพยากรณ์เป็นอย่างมาก นอกจากนี้การสลับลำดับของแบบจำลองที่ใช้ในการผสมแบบจำลอง ด้วยวิธี Cascaded neural network และ Stacked layers ในการพยากรณ์ระยะสั้น 7 วันและระยะยาว 30 วัน ไม่ส่งผลต่อประสิทธิภาพของ Hybrid Deep Learning Models อย่างมีนัยสำคัญ จึงสามารถสรุปได้ว่าการสร้าง Hybrid Deep Learning Models ด้วยวิธี Stacked layers ให้ประสิทธิภาพและความคุ้มค่าที่มากกว่าการผสมด้วยวิธี Cascaded neural network ทั้งในระยะสั้นและระยะยาว


การประเมินประสิทธิภาพของการเรียนรู้เชิงรุกโดยวิธีการเลือกแบบละโมบและวิธีการสุ่มตัวอย่างของทอมป์สันกับข้อมูลรูปแบบข้อความ, อานนท์ พรหมจรรย์ Jan 2024

การประเมินประสิทธิภาพของการเรียนรู้เชิงรุกโดยวิธีการเลือกแบบละโมบและวิธีการสุ่มตัวอย่างของทอมป์สันกับข้อมูลรูปแบบข้อความ, อานนท์ พรหมจรรย์

Chulalongkorn University Theses and Dissertations (Chula ETD)

การวิจัยนี้มุ่งเน้นการประเมินประสิทธิภาพของวิธีการเรียนรู้เชิงรุกสำหรับการคัดเลือกข้อมูลที่มีประโยชน์สูงสุดในการทำป้ายกำกับ โดยเปรียบเทียบวิธีการเลือกข้อมูลแบบสุ่ม วิธีการเลือกข้อมูลโดยวิธีละโมบ และวิธีการสุ่มตัวอย่างของทอมป์สันด้วยการประมาณค่าแบบลาพลาซ ผ่านการทดลอง 100 รอบในการเลือกทวีตเกี่ยวกับการท่องเที่ยวในกรุงเทพมหานครเพื่อฝึกโมเดลการถดถอยโลจิสติก ผลการทดลองพบว่าวิธีการเลือกข้อมูลโดยวิธีละโมบให้ประสิทธิภาพสูงสุดตลอดการทดลอง เนื่องจากสามารถปรับปรุงโมเดลได้อย่างรวดเร็ว แต่ประสิทธิภาพลดลงในช่วงท้ายเมื่อจำนวนทวีตที่มีประโยชน์ลดลง ขณะที่วิธีการสุ่มตัวอย่างของทอมป์สันด้วยการประมาณค่าแบบลาพลาซใช้เวลาในการคัดเลือกข้อมูลมากที่สุดและมีประสิทธิภาพต่ำกว่าในช่วงแรก อย่างไรก็ตาม เมื่อจำนวนรอบการทดลองเพิ่มขึ้น ความแม่นยำของโมเดลก็ค่อย ๆ ดีขึ้นเมื่อเทียบกับช่วงต้น ส่วนวิธีการเลือกข้อมูลแบบสุ่มใช้เวลาน้อยที่สุด แต่ไม่มีการเรียนรู้หรือปรับปรุงโมเดล ทำให้ประสิทธิภาพไม่ดีขึ้น จากผลการทดลองสามารถสรุปได้ว่าวิธีการเลือกข้อมูลโดยวิธีละโมบเป็นทางเลือกที่มีประสิทธิภาพสูงในสถานการณ์ที่ต้องการการเรียนรู้ที่รวดเร็ว ในขณะที่วิธีการสุ่มตัวอย่างของทอมป์สันด้วยการประมาณค่าแบบลาพลาซยังคงต้องมีการศึกษาเพิ่มเติมเกี่ยวกับศักยภาพในการเรียนรู้ในระยะยาว งานวิจัยนี้สามารถนำไปประยุกต์ใช้กับการวิเคราะห์ข้อมูลข้อความ เช่น การจำแนกประเภทความรู้สึกของผู้ใช้โซเชียลมีเดีย หรือการประเมินความคิดเห็นของลูกค้าในอุตสาหกรรมต่างๆ


The Influence Of Echo Chamber On Thailand's 2023 Election, Isariyaporn Sukcharoenchaikul Jan 2024

The Influence Of Echo Chamber On Thailand's 2023 Election, Isariyaporn Sukcharoenchaikul

Chulalongkorn University Theses and Dissertations (Chula ETD)

This research develops visual methods to explore the echo chamber effect, employing analytical approaches to investigate its dynamics through a case study on the 2023 General Election in Thailand. The study utilizes visualization techniques such as node-link diagrams, t-SNE projections, and heatmaps to analyze homophilic relationships, clustering tendencies, and polarization within online communities. To minimize inaccuracies and biases, network graphs are constructed based on contextual analysis of user-generated content, rather than relying on predefined relationship definitions (e.g., friendships, followers, retweets) or users' interpretations. The research uses the Echo Chamber Score (ECS) alongside visualizations to explore echo chambers, reveal significant variations …


A Comparative Study Of Early Fusion And Multimodal Siamese Neural Network Using Image And Text Data In Food Classification, Kanokporn Sintarasirikulchai Jan 2024

A Comparative Study Of Early Fusion And Multimodal Siamese Neural Network Using Image And Text Data In Food Classification, Kanokporn Sintarasirikulchai

Chulalongkorn University Theses and Dissertations (Chula ETD)

The economic development under capitalism has significantly transformed people's lifestyles, resulting in a fast-paced daily life. This shift has increased the consumption of convenient food options, leading to a preference for fast food, which is often high in carbohydrates and fats. Consequently, there has been a rise in obesity and related health issues, highlighting the importance of monitoring food intake. Automated systems utilizing artificial intelligence (AI) have emerged as potent tools for providing personalized dietary advice and monitoring. With the growing volume of food-related content on social media, including images and accompanying text, leveraging multimodal data has become essential for …


Extremal Graphs For Widom–Rowlinson Colorings In K-Chromatic Graphs, John Engbers, Aysel Erey Jan 2024

Extremal Graphs For Widom–Rowlinson Colorings In K-Chromatic Graphs, John Engbers, Aysel Erey

Mathematical and Statistical Science Faculty Research and Publications

The Widom–Rowlinson graph, HWR , is the fully looped path on three vertices. Let hom(G,HWR) be the number of graph homomorphisms from G to HWR or, equivalently, the number of HWR-colorings of G. We investigate extremal graphs for hom(G,HWR) for G in the family of k-chromatic graphs subject to various connectivity requirements. In particular, we determine the graphs G maximizing hom(G,HWR) in the families of n-vertex k-chromatic graphs, n-vertex connected k-chromatic graphs, n-vertex k-chromatic graphs with c components, n …


Integrating Machine Learning With Cure Models And Associated Inference, Wisdom Aselisewine Jan 2024

Integrating Machine Learning With Cure Models And Associated Inference, Wisdom Aselisewine

Mathematics Dissertations - Archive

Recent advancements in medical treatments have significantly enhanced the rates of recovery for numerous chronic illnesses. This progress has sparked growing interest in developing suitable statistical models capable of handling survival data that includes substantial cure fractions. The mixture cure model finds extensive application in analyzing survival data when there exists a cured subgroup. Standard logistic regression-based approaches for modeling the incidence part of the mixture cure model may suffer from poor predictive accuracy, especially in the presence of high dimensional covariates and/or non-linear covariate effects. To overcome this limitation, we propose the integration of distinct machine learning algorithms with …


The Performance Of Marginal Modeling Methods For Rare Events With Application To Opioid Overdose Mortality And Morbidity, Shawn Nigam Jan 2024

The Performance Of Marginal Modeling Methods For Rare Events With Application To Opioid Overdose Mortality And Morbidity, Shawn Nigam

Theses and Dissertations--Epidemiology and Biostatistics

Opioid misuse is a nationwide epidemic, with Kentucky having one of the highest opioid overdose-related fatality rates across all US states. These rates have increased significantly over the past decade, with particularly large increases during the COVID-19 pandemic. This dissertation aims to study the behavior of these increases and the methods for the marginal modeling of count outcomes related to opioid overdose.

Opioid overdose-related fatality rates in Kentucky increased significantly during the COVID-19 pandemic. In this chapter, we characterize the changes in opioid overdose fatality rates in Kentucky and identify associations between potential factors and fatality rates. County-level opioid overdose …


Variable Selection For High-Dimensional Data With Interaction Effects: Methods, Applications, And Inferences, Leiyue Li Jan 2024

Variable Selection For High-Dimensional Data With Interaction Effects: Methods, Applications, And Inferences, Leiyue Li

Theses and Dissertations--Statistics

For high-dimensional data where the number of variables greatly exceeds the number of observations, selecting important variables while maintaining the required heredity conditions can be challenging. This dissertation is structured into three interconnected parts. In the first part, we propose a variable selection method by implementing a well-known optimization technique, the Genetic Algorithm. An R package was developed to simplify the implementation and usage of the proposed method. We then propose another variable selection method by extending the study from the Genetic Algorithm to a different but related optimization technique, Simulated Annealing. We consider three different hierarchical structures in both …


High-Dimensional Tests And Projection Methods: Subvector Analysis And Matrix Variate Data, Shouryya Mitra Jan 2024

High-Dimensional Tests And Projection Methods: Subvector Analysis And Matrix Variate Data, Shouryya Mitra

Theses and Dissertations--Statistics

In this dissertation, we study projection-based methods to testing problems for high-dimensional data. We also investigate inferential methods for matrix variate data with block compound symmetry (BCS) covariance structure. The research reported in dissertation consists of three projects.

The first project addresses power loss and ill-conditioned error covariance estimates commonly faced by multivariate tests in high-dimensional settings. To overcome these challenges, previous approaches avoided correlations in constructing test statistics, but this required strong assumptions about covariance matrices and dependence structures. More recently, some methods have incorporated correlations by employing random projection into a lower-dimensional space. We develop a unified framework …


Estimation And Testing Of Nonparametric Effects Of Incomplete Multivariate And Repeated Measures Data, Swetalina Maity Jan 2024

Estimation And Testing Of Nonparametric Effects Of Incomplete Multivariate And Repeated Measures Data, Swetalina Maity

Theses and Dissertations--Statistics

In this dissertation, we investigate three distinct but interrelated problems in analyzing repeated measure data with missing values.

The first project is on semiparametric methods for analyzing repeated measures designs with missing values. A closed-form estimator for the parameters and their asymptotic covariance are derived using block partitioning based on the missing data pattern. We also derive the asymptotic distribution of the estimators and construct test statistics to test hypothesis formulated as linear contrasts of the mean vector.

In the second project, we focus on purely nonparametric methods. In this setup, nonparametric treatment effects are defined as functionals of the …


From Non-Parametric Methods To Self-Supervised Learning: Applications In Edge Detection And Image Denoising, Jiacheng Xu Jan 2024

From Non-Parametric Methods To Self-Supervised Learning: Applications In Edge Detection And Image Denoising, Jiacheng Xu

Theses and Dissertations--Statistics

This dissertation explores advanced methodologies for edge detection and image denoising through the application of both traditional non-parametric methods and modern self-supervised deep learning techniques. Beginning with non-parametric approaches, we refine surface fitting and jump detection criteria to enhance the detection of discontinuous regression surfaces in grayscale images. These foundational techniques are extended to color images, with analyses across RGB and CIELAB color spaces to improve edge detection accuracy. We then introduce a self-supervised neural network model that integrates Masked Modeling into the Bi-Directional Cascade Network (BDCN) framework. This approach shows the potential of reducing the dependency on annotated data …


Patterns Into Pathways For Improving Safety Culture: Refined Latent Class Analysis Informs Tailored Decision Support For South Carolina Dss Safety Culture Improvements, Michaela Voit Jan 2024

Patterns Into Pathways For Improving Safety Culture: Refined Latent Class Analysis Informs Tailored Decision Support For South Carolina Dss Safety Culture Improvements, Michaela Voit

Theses and Dissertations--Public Health (M.P.H. & Dr.P.H.)

The rising prevalence of exposures to adverse childhood experiences (ACEs) demands a coordinated public health response, as a significant body of research details the cumulative impact of ACEs on chronic morbidities contributing to reduced life expectancy. Child welfare workers (CWW) are embedded in this public health effort, tasked with preventing and mitigating the impacts of ACEs through family and prevention services. The National Partnership for Child Safety (NPCS) may improve the wellbeing of CWWs and the effectiveness of Child Welfare (CW) services by improving the quality of safety culture within CW organizations. To inform NPCS quality improvement efforts, our project …


Imputation Strategies For Different Categories Of Missing Data, Karthik Chalumuri Jan 2024

Imputation Strategies For Different Categories Of Missing Data, Karthik Chalumuri

Honors Theses and Capstones

Addressing missing data in research is crucial for ensuring the reliability and validity of study findings, yet it remains a significant challenge. This study investigates the impact of missing data on research outcomes and explores the underutilization of existing tools for managing missingness, potentially leading to gaps in critical information with tangible implications for decision-making processes (Dziura et al.).

Focusing on the different categories of missing data—Missing Completely At Random (MCAR), Missing At Random (MAR), and Missing Not At Random (MNAR)—this research examines various imputation strategies tailored to each category. Specifically, we compare the efficacy of several model-based imputation methods, …


A Case Study On Variations In Network Structure And Cross- Sector Alignment In Two Local Systems Serving Pregnant And Parenting Women In Recovery, Liza M. Creel, Yana Feygin, Madeline Shipley, Deborah Winders Davis, Tiffany Cole Hall, Chaly Downs, Stephanie Hoskins, Natalie Pasquenza, Scott D. Duncan Jan 2024

A Case Study On Variations In Network Structure And Cross- Sector Alignment In Two Local Systems Serving Pregnant And Parenting Women In Recovery, Liza M. Creel, Yana Feygin, Madeline Shipley, Deborah Winders Davis, Tiffany Cole Hall, Chaly Downs, Stephanie Hoskins, Natalie Pasquenza, Scott D. Duncan

Biostatistics Faculty Publications

Objective: To describe network structure and alignment across organizations in healthcare, public health, and social services sectors that serve pregnant and parenting women with substance use disorder (SUD) in an urban and a rural community.

Data Sources and Study Settings: Two community networks, one urban and one rural with each including a residential substance use treatment program, in Kentucky during 2021.

Study Design: Social network analysis measured system collaboration and cross-sector alignment between healthcare, public health, and social services organizations, applying the Framework for Aligning Sectors. To understand the alignment and structure of each network, we measured network density overall …


Higher First 30-Day Dose Of Buprenorphine For Opioid Use Disorder Treatment Is Associated With Decreased Mortality, Feitong Lei, Michelle R. Lofwall, Jana Mcanich, Reuben Adatorwovor, Emily Slade, Patricia R. Freeman, Daniela Moga, Nabarun Dasgupta, Sharon L. Walsh, Rachel Vickers-Smith, Svetla Slavova Jan 2024

Higher First 30-Day Dose Of Buprenorphine For Opioid Use Disorder Treatment Is Associated With Decreased Mortality, Feitong Lei, Michelle R. Lofwall, Jana Mcanich, Reuben Adatorwovor, Emily Slade, Patricia R. Freeman, Daniela Moga, Nabarun Dasgupta, Sharon L. Walsh, Rachel Vickers-Smith, Svetla Slavova

Biostatistics Faculty Publications

Objective: Buprenorphine is a medication for opioid use disorder that reduces mortality. This study aims to investigate the less well-understood relationship between the dose in the early stages of treatment and the subsequent risk of death.

Methods: We used Kentucky prescription monitoring data to identify adult Kentucky residents initiating transmucosal buprenorphine medication for opioid use disorder (January 2017 to November 2019). Average daily buprenorphine dose for days covered in the first 30 days of treatment was categorized as ≤8 mg, >8 to ≤16 mg, and >16 mg. Patients were followed for 365 days after the first 30 days of buprenorphine …


Synergetic Effects Of Democracy And Economic Development On Income Inequality And The Role Of Relative Political Capacity, Nazif Sali Jan 2024

Synergetic Effects Of Democracy And Economic Development On Income Inequality And The Role Of Relative Political Capacity, Nazif Sali

CGU Theses & Dissertations

Diving into the complex dynamics of income inequality, this dissertation studies the multifaceted relationships between democracy, economic development, and inequality, placing a spotlight on the mediating influence of Relative Political Capacity (RPC). Uncovering the pivotal role of RPC as democracies advance and economies progress, this research navigates the contours of income inequality. By dissecting distinct sub-samples of OECD and non-OECD countries, nuanced insights surface. Employing robust methodologies such as ordinary least squares regressions and fixed effects analysis, I argue that targeted policies addressing inequality can not only foster inclusive economic growth but also fortify the foundations of democratic institutions. As …


Intelligent Capabilities Of Traditional Knowledge Organization Methods, Xinning Su Jan 2024

Intelligent Capabilities Of Traditional Knowledge Organization Methods, Xinning Su

Journal of Scientific Information Research

[Purpose/significance]By analyzing the system and rules of traditional knowledge organization methods, the intelligent capabilities of traditional knowledge organization methods are refined and integrated into artificial intelligence(AI) technology, to enhance the precision and efficiency of AI in information processing. [Method/process]This paper reviews the development of knowledge organization and analyses the inherit structure and mechanisms of traditional knowledge organization methods. [Result/conclusion]Research suggests that over centuries of development and evolution, knowledge organization has gained the ability to reflect knowledge systems and disciplinary systems across different disciplines from diverse perspectives, establish semantic relations from diverse knowledge associations, and associate and integrate knowledge of different …


A Quantitative Analysis Of Seaplane Accidents From 1982-2021, David C. Ison Jan 2024

A Quantitative Analysis Of Seaplane Accidents From 1982-2021, David C. Ison

International Journal of Aviation, Aeronautics, and Aerospace

This study aimed to assess and analyze all historical National Transportation Safety Board accident reports since 1982. For analysis, reports were bisected into seaplane (float, amphibian, and hull) and non-seaplane groups. Findings showed that there is a deficiency in the level of available detail on the seaplane fleet and cadre of seaplane pilots in the U.S. During the most recent ten years of complete data (2012-2021) showed a negative trend in all accidents and fatal accidents, although only the latter being statistically convincing. During this timeframe, seaplane accident pilots had significantly higher total time and age than other groups (non-seaplane …


Bar-Code Variable: A Novel Approach To Efficiently Find Interaction Effects, Lee Sak Park Jan 2024

Bar-Code Variable: A Novel Approach To Efficiently Find Interaction Effects, Lee Sak Park

Theses and Dissertations--Statistics

This paper introduces the bar-code variable, a novel method for processing a sequence of binary explanatory variables efficiently in the linear regression modeling framework. Represented as an integer or a sequence of bits, the bar-code variable captures infor- mation on original binary variables and their potential interaction effects. Utilizing the bar-code variable, the study explores streamlined feature selection in linear re- gression modeling with binary explanatory variables. The paper demonstrates how the bar-code variable, through re-parameterization, facilitates the transition from cell means estimates, µ̂, in the cell-means ANOVA model to coefficient estimates, β̂, in the linear regression model, and vice …


On Generative Models And Joint Architectures For Document-Level Relation Extraction, Aviv Brokman Jan 2024

On Generative Models And Joint Architectures For Document-Level Relation Extraction, Aviv Brokman

Theses and Dissertations--Statistics

Biomedical text is being generated at a high rate in scientific literature publications and electronic health records. Within these documents lies a wealth of potentially useful information in biomedicine. Relation extraction (RE), the process of automating the identification of structured relationships between entities within text, represents a highly sought-after goal in biomedical informatics, offering the potential to unlock deeper insights and connections from this vast corpus of data. In this dissertation, we tackle this problem with a variety of approaches.

We review the recent history of the field of document-level RE. Several themes emerge. First, graph neural networks dominate the …


Difs And Bayescluster: Novel Methods For Single_Cell Rna Sequencing Analysis, Kun Liu Jan 2024

Difs And Bayescluster: Novel Methods For Single_Cell Rna Sequencing Analysis, Kun Liu

Theses and Dissertations--Statistics

Single-cell RNA sequencing (scRNA-seq) has transformed our understanding of cellular heterogeneity and gene expression dynamics. Despite its potential, the inherent noise and sparsity of scRNA-seq data pose significant challenges in clustering cells into biologically meaningful groups. This dissertation addresses these challenges through two novel methodologies aimed at enhancing the accuracy and robustness of scRNA-seq data analysis.

First, we introduce the Differential Feature Selection (DIFS) framework, designed to improve the identification of differential features in scRNA-seq data. DIFS employs a two-stage marker identification process. In the first stage, a modified Dip Test is used to filter and identify genes with significant …


A Case Report On A Women’S Residential Substance Use Program In A Rural And Urban Setting, Deborah Winders Davis, Yana Feygin, Madeline Shipley, Tiffany Cole Hall, Chaly Downs, Stephanie Hoskins, Natalie Pasquenza, Scott D. Duncan, Liza M. Creel Jan 2024

A Case Report On A Women’S Residential Substance Use Program In A Rural And Urban Setting, Deborah Winders Davis, Yana Feygin, Madeline Shipley, Tiffany Cole Hall, Chaly Downs, Stephanie Hoskins, Natalie Pasquenza, Scott D. Duncan, Liza M. Creel

Biostatistics Faculty Publications

Purpose To describe program characteristics and outcomes of a residential substance use recovery program serving pregnant and parenting women in a rural and urban location.

Description This assessment of administrative records from April 1, 2020 through March 31, 2022, included women in a rural (n = 140) and urban (n = 321) county in Kentucky.

Assessment This retrospective case study used descriptive and non-parametric analyses to assess the population and examine differences between locations, race, and ethnicity for women served. Logistic regression tested predictors of goal achievement by community. Of 461 women served, 65 (14.1%) delivered a baby while in …


Health Care For People Who Are Incarcerated: Teaching Third-Year Medical Students About Rights, Challenges, And Avenues Of Advocacy, Anna-Maria South, Kelsey N. Karnik, Sara Hieneman, Anthony A. Mangino, Michelle R. Lofwall Jan 2024

Health Care For People Who Are Incarcerated: Teaching Third-Year Medical Students About Rights, Challenges, And Avenues Of Advocacy, Anna-Maria South, Kelsey N. Karnik, Sara Hieneman, Anthony A. Mangino, Michelle R. Lofwall

Biostatistics Faculty Publications

Introduction: Incarcerated patients are a vulnerable patient population with unique barriers to health care that physicians in every specialty encounter. Current medical school curricula lack universal education on health care for incarcerated people.

Methods: We developed an interactive workshop to provide third-year medical students at the University of Kentucky with information about delivering care outside of dedicated carceral settings to individuals who are incarcerated. The workshop included education on the demographic characteristics and medical conditions present in these populations along with understanding incarcerated persons’ rights to health care and how to interact with them and the associated jail/prison workforce often …


Contrastive Learning, With Application To Forensic Identification Of Source, Cole Ryan Patten Jan 2024

Contrastive Learning, With Application To Forensic Identification Of Source, Cole Ryan Patten

Electronic Theses and Dissertations

Forensic identification of source problems often fall under the category of verification problems, where recent advances in deep learning have been made by contrastive learning methods. Many forensic identification of source problems deal with a scarcity of data, an issue addressed by few-shot learning. In this work, we make specific what makes a neural network a contrastive network. We then consider the use of contrastive neural networks for few-shot learning classification problems and compare them to other statistical and deep learning methods. Our findings indicate similar performance between models trained by contrastive loss and models trained by cross-entropy loss. We …


Generating Neutrosophic Random Variables Based Gamma Distribution, Maissam Ahmad Jdid, Florentin Smarandache, Khalifa Al Shaqsi Jan 2024

Generating Neutrosophic Random Variables Based Gamma Distribution, Maissam Ahmad Jdid, Florentin Smarandache, Khalifa Al Shaqsi

Branch Mathematics and Statistics Faculty and Staff Publications

In practical life, we encounter many systems that cannot be studied directly, either due to their high cost or because some of these systems cannot be studied directly. Therefore, we resort to the simulation method, which depends on applying the study to systems similar to real ones and then projecting these results if they are suitable for the real system. The simulation process requires a good understanding of probability distributions and the methods used to transform random numbers that follow a regular distribution in the field [0,1] into random variables that follow them, so that we can achieve the greatest …


Row-Column Designs: A Novel Approach For Analyzing Imprecise And Uncertain Observations, Abdulrahman Alaita, Muhammad Aslam, Florentin Smarandache Jan 2024

Row-Column Designs: A Novel Approach For Analyzing Imprecise And Uncertain Observations, Abdulrahman Alaita, Muhammad Aslam, Florentin Smarandache

Branch Mathematics and Statistics Faculty and Staff Publications

Classical row-column designs cannot be applied when the underlying data set contains some imprecise, uncertain, or undetermined observations. In this paper, we discuss row-column design under a neutrosophic statistical framework. A significant contribution of our study is to propose a novel approach to analyzing row-column designs using neutrosophic data. This approach involves calculating the neutrosophic analysis of variance (NANOVA) table for the proposed design and using it to derive the FN -test in an uncertain environment. Two numerical examples have been used to assess the proposed design’s performance. Results from the study indicated that a row column design under …


Examining Information Systems Use To Facilitate The Workplace Accommodation Process, Shiya Cao Jan 2024

Examining Information Systems Use To Facilitate The Workplace Accommodation Process, Shiya Cao

Statistical and Data Sciences: Faculty Publications

BACKGROUND: The workplace accommodation process is often affected by ineffective and inefficient communications and information exchanges among disabled employees and other stakeholders. Information systems (IS) can play a key role in facilitating a more effective and efficient accommodation process since IS has been shown to facilitate business processes and effect positive organizational changes.

OBJECTIVE: Since there is little to no research that exists on IS use to facilitate the workplace accommodation process, this paper, as a critical first step, examines how IS have been used in the accommodation process.

METHODS: Thirty-six interviews were conducted with disabled employees from various organizations. …


Questions (And Answers) For Incorporating Nontraditional Grading In Your Statistics Courses, Brenna Curley Jan 2024

Questions (And Answers) For Incorporating Nontraditional Grading In Your Statistics Courses, Brenna Curley

Statistical and Data Sciences: Faculty Publications

Nontraditional grading methods have recently become more common, and as with any large pedagogical shift, there are a number of questions to consider when applying a new grading scheme to a course. This article summarizes four types of nontraditional grading and shares experiences from the authors who have applied them to a variety of courses in statistics. This article is structured as a set of questions and answers, seeking to address many of the concerns and considerations that one may face as they transition a course’s grading structure. Supplementary materials for this article are available online.