Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Medicine and Health Sciences (172)
- Biostatistics (160)
- Public Health (143)
- Applied Statistics (113)
- Epidemiology (112)
-
- Social and Behavioral Sciences (108)
- Mathematics (79)
- Life Sciences (76)
- Statistical Models (67)
- Data Science (65)
- Applied Mathematics (54)
- Computer Sciences (51)
- Statistical Methodology (46)
- Health Services Research (45)
- Other Statistics and Probability (41)
- Public Affairs, Public Policy and Public Administration (38)
- Engineering (32)
- Public Health Education and Promotion (32)
- Environmental Public Health (31)
- Health Policy (31)
- Nutrition (31)
- Probability (31)
- Women's Health (31)
- Business (30)
- Occupational Health and Industrial Hygiene (30)
- Categorical Data Analysis (29)
- Education (29)
- Clinical Trials (25)
- Institution
-
- University of South Carolina (62)
- Universitas Indonesia (31)
- Missouri University of Science and Technology (26)
- Chulalongkorn University (19)
- University of Nebraska - Lincoln (19)
-
- University of South Florida (18)
- University of New Mexico (17)
- Roseman University of Health Sciences (16)
- Utah State University (16)
- University of Arkansas, Fayetteville (15)
- University of Kentucky (15)
- Air Force Institute of Technology (14)
- Georgia Southern University (14)
- Clemson University (11)
- Prairie View A&M University (11)
- Virginia Commonwealth University (11)
- Central Bank of Nigeria (10)
- City University of New York (CUNY) (9)
- Louisiana State University (9)
- Southern Methodist University (9)
- University of Nevada, Las Vegas (9)
- Bethel University (8)
- Smith College (8)
- University of Denver (8)
- Old Dominion University (7)
- Northern Illinois University (6)
- University of Mississippi (6)
- Washington University in St. Louis (6)
- Wayne State University (6)
- DePauw University (5)
- Keyword
-
- COVID-19 (28)
- Statistics (20)
- Machine learning (17)
- Machine Learning (9)
- Dietary inflammatory index (8)
-
- Mortality (8)
- Psychology (8)
- Risk (8)
- Inflammation (7)
- Data science (6)
- Morgridge College of Education (6)
- Regression (6)
- Research Methods and Information Science (6)
- Research Methods and Statistics (6)
- Women (6)
- Classification (5)
- Deep Learning (5)
- Epidemiology (5)
- Exercise (5)
- Nutrition (5)
- Obesity (5)
- Pregnancy (5)
- Survival analysis (5)
- Biomarkers (4)
- Deep learning (4)
- Forecasting (4)
- HIV (4)
- Health (4)
- Humans (4)
- Mathematics (4)
- Publication
-
- Faculty Publications (57)
- Theses and Dissertations (32)
- Kesmas (31)
- Mathematics and Statistics Faculty Research & Creative Works (21)
- Chulalongkorn University Theses and Dissertations (Chula ETD) (19)
-
- Department of Statistics: Faculty Publications (17)
- Annual Research Symposium (16)
- Electronic Theses and Dissertations (15)
- Mathematics & Statistics ETDs (15)
- USF Tampa Graduate Theses and Dissertations (14)
- Applications and Applied Mathematics: An International Journal (AAM) (11)
- All Dissertations (10)
- Biostatistics, Epidemiology & Environmental Health Sciences: Faculty Publications (10)
- CBN Journal of Applied Statistics (JAS) (10)
- Graduate Theses and Dissertations (10)
- All Graduate Theses and Dissertations, Spring 1920 to Summer 2023 (8)
- Psychology Student Works (8)
- LSU Doctoral Dissertations (7)
- SMU Data Science Review (7)
- Conference on Applied Statistics in Agriculture and Natural Resources (6)
- Statistical and Data Sciences: Faculty Publications (6)
- UNLV Theses, Dissertations, Professional Papers, and Capstones (6)
- Arts & Sciences Graduate Student Theses and Dissertations (5)
- Dissertations and Theses (Open Access) (5)
- Harrisburg University Research Symposium: Highlighting Research, Innovation, & Creativity (5)
- Journal of Modern Applied Statistical Methods (5)
- Legacy Theses & Dissertations (2009 - 2024) (5)
- Open Access Theses & Dissertations (5)
- Publications (5)
- Research outputs 2022 to 2026 (5)
- Publication Type
- File Type
Articles 571 - 595 of 595
Full-Text Articles in Statistics and Probability
การเปรียบเทียบสถาปัตยกรรมโครงข่ายประสาทคอนโวลูชัน 3 มิติ โดยการจำแนกโรคหลอดเลือดสมองจากภาพการฉายรังสีเอกซเรย์สมอง, ชานนท์ วรโชติสืบตระกูล
การเปรียบเทียบสถาปัตยกรรมโครงข่ายประสาทคอนโวลูชัน 3 มิติ โดยการจำแนกโรคหลอดเลือดสมองจากภาพการฉายรังสีเอกซเรย์สมอง, ชานนท์ วรโชติสืบตระกูล
Chulalongkorn University Theses and Dissertations (Chula ETD)
แบบจำลองโครงข่ายคอนโวลูชัน หรือ ซีเอ็นเอ็น (Convolutional Neural Networks หรือ CNN) ได้รับการนำมาใช้กันอย่างแพร่หลายในการจำแนกภาพ โดยเฉพาะในทางการแพทย์ ซึ่งปกติการจำแนกภาพทางการแพทย์นิยมใช้โครงข่ายคอนโวลูชั่น 2 มิติ แต่เนื่องจากข้อมูลภาพบางประเภท เช่น ภาพการฉายรังสีเอกซเรย์สมองมีลักษณะมองภาพ 3 มิติ ให้เป็นภาพ 2 มิติ ดังนั้นในงานวิจัยนี้จึงมีแนวคิดในการใช้โครงข่ายคอนโวลูชัน 3 มิติมาใช้ในการจำแนกภาพเพื่อนำเอาจุดเด่นจากความสามารถในการดึงคุณลักษณะความสัมพันธ์ในชั้นความลึกที่เพิ่มเข้ามาซึ่งมีความแตกต่างจากรูปแบบ 2 มิติ เพื่อเพิ่มประสิทธิภาพให้แบบจำลองสามารถดึงคุณลักษณะสำคัญของภาพให้มีความหลากหลายมากขึ้น งานวิจัยนี้มีวัตถุประสงค์เพื่อเปรียบเทียบประสิทธิภาพโครงข่ายคอนโวลูชัน 3 มิติ ร่วมกับแบบจำลองที่ถูกฝึกมาเรียบร้อยแล้ว (pre-trained model) 4 แบบจำลอง ประกอบไปด้วย อเล็กซ์เน็ต (Alexnet) วีจีจี-16 (Vgg-16) กูเกิลเน็ต (Googlenet) และเรสเน็ต (Resnet) เพื่อจำแนกข้อมูลภาพผู้ป่วยที่เป็นโรคหลอดเลือดสมอง และผู้ป่วยที่มีสุขภาพปกติ จากภาพฉายรังสีเอกซเรย์สมอง (CT-Scan) จากฐานข้อมูลเว็บไซด์ Kaggle ชุดข้อมูลประกอบด้วยภาพผู้ป่วยที่เป็นโรคหลอดเลือดสมอง 950 ภาพ จาก 40 คน และภาพผู้ป่วยสุขภาพปกติ 1551 ภาพ จาก 82 คน ซึ่งงานวิจัยนี้มีการปรับรายละเอียดโดยการนำจุดเด่นของแต่ละแบบจำลองมาใช้ และเพิ่มชั้นความลึกที่เป็นจุดเด่นของการค้นหาคุณลักษณะสำคัญของรูปแบบ 3 มิติ ร่วมกับการประมวลผลภาพล่วงหน้า (Image Preprocessing) และการทำการเพิ่มจำนวนข้อมูล (Data augmentation) เพื่อเพิ่มประสิทธิภาพของแบบจำลอง จากนั้นเพื่อไม่ให้การทดลองโน้มเอียงต่อแต่ละแบบจำลอง มีการนำเทคนิค K-Fold Cross validation (K=5) มาเพื่อแก้ปัญหาในงานวิจัยชิ้นนี้ ในส่วนของการวัดประสิทธิภาพผลการทดลองใช้ Confusion matrix เป็นเครื่องมือในการประเมินประสิทธิภาพของแบบจำลอง ซึ่งพบว่าสมรรถนะแบบจำลองโครงข่ายคอนโวลูชันกูเกิลเน็ต 3 มิติ ให้ผลลัพธ์ที่ดีที่สุด โดยผลการทดสอบการจำแนกภาพผู้ป่วยที่เป็นโรคหลอดเลือดสมองจากภาพฉายรังสีเอกซเรย์ ให้ค่าความแม่นยำ ความเที่ยงตรง ค่าความครบถ้วน และ F1-Score ที่ 92.00% 94.01% 83.96% และ 88.70% …
การเรียนรู้แบบรวมกลุ่มด้วยตัวแบบที่แตกต่างกันแบบขนานสำหรับข้อมูลไม่สมดุล กรณีศึกษาข้อมูลเครดิตเยอรมัน, ศศิวิมล ศรีโรจน์
การเรียนรู้แบบรวมกลุ่มด้วยตัวแบบที่แตกต่างกันแบบขนานสำหรับข้อมูลไม่สมดุล กรณีศึกษาข้อมูลเครดิตเยอรมัน, ศศิวิมล ศรีโรจน์
Chulalongkorn University Theses and Dissertations (Chula ETD)
งานวิจัยนี้มีวัตถุประสงค์เพื่อสร้างตัวแบบการเรียนรู้แบบรวมกลุ่มด้วยตัวแบบที่แตกต่างกันแบบขนาน (Bagging Heterogeneous Ensemble) และหาวิธีการลดมิติข้อมูลและวิธีการสุ่มตัวอย่างซ้ำที่เหมาะสมกับข้อมูลเครดิตเยอรมันที่มีอัตราส่วนความไม่สมดุลแตกต่างกัน 3 ค่าคือ 2.3, 10 และ 14 โดยวัดประสิทธิภาพด้วยตัวชี้วัด Accuracy, The area under the curve, F1-score, Precision, Brier score และ Kolmogorov-Smirnov และทดสอบทางสถิติเพื่อแสดงว่าประสิทธิภาพของตัวแบบมีความแตกต่างกัน ที่ระดับนัยสำคัญ 0.05 ผลการศึกษาพบว่าข้อมูลเครดิตเยอรมันที่มีอัตราส่วนความไม่สมดุลต่ำ (IR = 2.3) ตัวแบบ Logistic Regression ที่ใช้เทคนิค Linear Discriminant Analysis (LDA) และ Systematic Minority Over-Sampling Technique (SM) จะมีประสิทธิภาพเฉลี่ยดีที่สุดในการจำแนกประเภท ในส่วนของอัตราส่วนความไม่สมดุลกลาง (IR = 10) และ อัตราส่วนความไม่สมดุลสูง (IR = 14) วิธีการลดมิติข้อมูลและการสุ่มตัวอย่างซ้ำที่มีประสิทธิภาพคือ Linear Discriminant Analysis (LDA), Random Under-Sampling (RUS) และ Linear Discriminant Analysis (LDA), Borderline SMOTE (BSM) ตามลำดับ โดยที่การเรียนรู้แบบรวมกลุ่มด้วยตัวแบบที่แตกต่างกันแบบขนานมีประสิทธิภาพเฉลี่ยดีที่สุด ทั้งในกรณีที่มีและไม่มีวิธีการลดมิติข้อมูลและสุ่มตัวอย่างซ้ำของอัตราส่วนความไม่สมดุลกลางและสูง
ประสิทธิภาพของวิธีการจัดการข้อมูลไม่สมดุลสำหรับการจำแนกกลุ่มภายใต้เงื่อนไขที่แตกต่างกัน, กาญธนา ลออสิริกุล
ประสิทธิภาพของวิธีการจัดการข้อมูลไม่สมดุลสำหรับการจำแนกกลุ่มภายใต้เงื่อนไขที่แตกต่างกัน, กาญธนา ลออสิริกุล
Chulalongkorn University Theses and Dissertations (Chula ETD)
การวิจัยนี้มีจุดประสงค์เพื่อศึกษาปฏิสัมพันธ์ของวิธีการปรับสมดุลข้อมูลกับเงื่อนไขด้านขนาดตัวอย่าง เทคนิคการจำแนกข้อมูล จำนวนตัวแปรระหว่างกลุ่มตัวแปรจัดประเภทต่อกลุ่มตัวแปรต่อเนื่อง อัตราออด และร้อยละของจำนวนข้อมูลกลุ่มหลักต่อข้อมูลกลุ่มรองที่มีต่อประสิทธิภาพของการจำแนกกลุ่ม การปรับสมดุลของข้อมูลแบ่งออกเป็น 3 วิธี ได้แก่ (1) ไม่ปรับสมดุล (2) วิธี random oversampling และ (3) วิธีผสมผสานระหว่างรูปแบบสุ่มเกินและสุ่มลด (hybrid) โดยใช้แพคเกจ ROSE ส่วนเงื่อนไขด้านขนาดตัวอย่างแบ่งออกเป็น ขนาดตัวอย่างเท่ากับ 100 300 และ 500 หน่วย ด้านเทคนิคการจำแนกข้อมูล แบ่งออกเป็น 4 วิธี ได้แก่ (1) เคเนียร์เรสเนเบอร์ (2) การถดถอยโลจิสติก (3) แรนดอมฟอร์เรส และ (4) ซัพพอร์ตเวกเตอร์แมชชีน ตัวแปรจากการจำลองแบ่งออกเป็นตัวแปรตามซึ่งจำลองด้วยการถดถอยโลจิสติก ส่วนตัวแปรอิสระในการจำลองข้อมูลครั้งนี้จะกำหนดให้ใช้ตัวแปรอิสระจำลองทั้งหมด 8 ตัว โดยกำหนดให้มีจำนวนตัวแปรระหว่างกลุ่มตัวแปรจัดประเภทต่อกลุ่มตัวแปรต่อเนื่อง 3 กรณี คือ 4:4 5:3 และ 6:2 ในขณะที่ระดับของอัตราออด จะสุ่มค่าจากช่วง [1,2) หรือ [2,3) และร้อยละของข้อมูลระหว่างข้อมูลกลุ่มหลักต่อข้อมูลกลุ่มรอง แบ่งออกเป็น 2 กรณี ได้แก่ 60:40 และ 70:30 พิจารณาเกณฑ์ประสิทธิภาพของข้อมูลด้วยตัวชี้วัดความถูกต้องในการจำแนก ความไว และความจำเพาะ การจำลองแต่ละสถานการณ์จะทำซ้ำสถานการณ์ละ 500 รอบ การวิเคราะห์ปฏิสัมพันธ์ระหว่างวิธีการปรับสมดุลข้อมูลกับเงื่อนไขต่าง ๆ ใช้การวิเคราะห์ความแปรปรวนพหุคูณหลายทาง (n-way MANOVA) ผลการวิจัยพบว่า วิธีการปรับสมดุลข้อมูลมีปฏิสัมพันธ์แบบสองทางกับเงื่อนไขด้านขนาดตัวอย่าง ร้อยละของข้อมูลระหว่างข้อมูลกลุ่มหลักต่อข้อมูลกลุ่มรอง อัตราออด และเทคนิคการจำแนกข้อมูล และพบปฏิสัมพันธ์แบบสามทางกับเงื่อนไขต่อไปนี้ (1) ขนาดตัวอย่างและจำนวนตัวแปรระหว่างกลุ่มตัวแปรจัดประเภทต่อกลุ่มตัวแปรต่อเนื่อง (2) ขนาดตัวอย่างและเทคนิคการจำแนกข้อมูล และ (3) ร้อยละของข้อมูลระหว่างข้อมูลกลุ่มหลักต่อข้อมูลกลุ่มรอง และเทคนิคการจำแนกข้อมูล ดังนั้นนักวิเคราะห์ข้อมูลควรเลือกใช้วิธีการปรับสมดุลข้อมูลโดยพิจารณาให้เหมาะสมกับสภาพของข้อมูลที่ใช้ในการวิเคราะห์
สมรรถนะดิจิทัลขององค์กรทหาร: การวิเคราะห์องค์ประกอบเชิงสำรวจและเชิงยืนยันพหุระดับ, รัมณรา สมประสงค์
สมรรถนะดิจิทัลขององค์กรทหาร: การวิเคราะห์องค์ประกอบเชิงสำรวจและเชิงยืนยันพหุระดับ, รัมณรา สมประสงค์
Chulalongkorn University Theses and Dissertations (Chula ETD)
สมรรถนะดิจิทัลขององค์กรทหารในปัจจุบันมีความสำคัญต่อการปฏิบัติงานในยุคของการเปลี่ยนแปลงทางดิจิทัลที่เกิดขึ้นอย่างรวดเร็ว งานวิจัยนี้เป็นงานวิจัยเชิงบรรยาย มีวัตถุประสงค์ดังนี้ 1) เพื่อสังเคราะห์ตัวชี้วัดสมรรถนะดิจิทัลขององค์กรทหาร 2) เพื่อสำรวจองค์ประกอบพหุระดับสมรรถนะดิจิทัลขององค์กรทหาร 3) เพื่อตรวจสอบความสอดคล้องเชิงประจักษ์ขององค์ประกอบพหุระดับสมรรถนะดิจิทัลขององค์กรทหาร ตัวอย่างวิจัย เป็นบุคลากรระดับปฏิบัติการในองค์กรทหารสังกัดกระทรวงกลาโหม 50 หน่วยงาน จำนวน 860 คน สำหรับใช้ในการวิเคราะห์องค์ประกอบเชิงสำรวจพหุระดับ และจำนวน 863 คน สำหรับใช้ในการวิเคราะห์องค์ประกอบเชิงยืนยันพหุระดับ เครื่องมือที่ใช้ในการวิจัยเพื่อการวิเคราะห์องค์ประกอบพหุระดับ คือแบบวัดสมรรถนะดิจิทัลขององค์กรทหาร ประกอบไปด้วย 2 ตอน คือ ข้อมูลพื้นฐานของกำลังพลผู้ตอบแบบสอบถาม และแบบวัดสมรรถนะดิจิทัลขององค์กรทหาร จำนวน 69 ข้อคำถาม วิเคราะห์ข้อมูลด้วยโปรแกรม IBM SPSS Statistics 22 และ MPlus6 ผลการวิจัยพบว่า (1) ตัวชี้วัดสมรรถนะดิจิทัลขององค์กรทหาร ประกอบไปด้วย 16 ตัวชี้วัด ได้แก่ 1) การวางแผนการใช้งานอุปกรณ์เทคโนโลยีดิจิทัลแบบบูรณาการ 2) การสืบค้นข้อมูลทางดิจิทัล 3) การประเมินความน่าเชื่อถือของข้อมูล 4) การใช้งานเทคโนโลยีเบื้องต้น 5) การแก้ปัญหาจากการใช้เทคโนโลยีดิจิทัล 6) การใช้อินทราเน็ตขององค์กร 7) การรักษาความลับในโลกไซเบอร์ 8) การจัดการไฟล์ดิจิทัลทางการทหาร 9) การเข้าถึงไฟล์ดิจิทัลในกรณีปฏิบัติงานนอกสถานที่ 10) การจัดการฐานข้อมูลทางการทหาร 11) การใช้สื่อดิจิทัลทางไกลเพื่อการสื่อสารทางการทหาร 12) การนำเสนอข้อมูลทางทหารในรูปแบบดิจิทัล 13) การสร้างสิ่งแวดล้อมทางดิจิทัลเพื่อการทำงาน 14) การตระหนักถึงความปลอดภัยบนโลกไซเบอร์ 15) การรักษามารยาทในสังคมดิจิทัล 16) เจตคติต่อการใช้เทคโนโลยีดิจิทัลในองค์กร (2) องค์ประกอบเชิงสำรวจพหุระดับสมรรถนะดิจิทัลขององค์กรทหาร มีจำนวน 3 โมเดล คือ 1) องค์ประกอบระดับระดับบุคคล 4 องค์ประกอบ ระดับองค์กร 1 องค์ประกอบ 2) องค์ประกอบระดับบุคคล 4 องค์ประกอบ ระดับองค์กร 2 องค์ประกอบ 3) องค์ประกอบระดับบุคคล …
การเปรียบเทียบวิธีการใส่ค่าสูญหาย ในตัวแบบการถดถอยเชิงเส้นพหุที่ตัวแปรอิสระมีการสูญหายแบบนอนอิกนอร์เรเบิลที่สัมพันธ์กัน, ศุภสันติ์ ดีมาก
การเปรียบเทียบวิธีการใส่ค่าสูญหาย ในตัวแบบการถดถอยเชิงเส้นพหุที่ตัวแปรอิสระมีการสูญหายแบบนอนอิกนอร์เรเบิลที่สัมพันธ์กัน, ศุภสันติ์ ดีมาก
Chulalongkorn University Theses and Dissertations (Chula ETD)
งานวิจัยนี้มีจุดประสงค์เพื่อศึกษาและเปรียบเทียบวิธีการประมาณสูญหายในตัวแบบการถดถอยเชิงเส้นพหุคูณ ที่ตัวแปรอิสระมีการสูญหายแบบนอนอิกนอร์เรเบิลที่มีความสัมพันธ์กัน ในการศึกษานี้มีวิธีการที่ถูกพัฒนาขึ้นคือ Expected Regression Imputation (ERI) และ Conditional Expected Regression Imputation (CERI) โดยจะเปรียบเทียบประสิทธิภาพวิธีการที่พัฒนาขึ้นมากับอีก 3 วิธีการ ได้แก่ วิธี K-Nearest Neighbor Imputation (KNN), วิธี Expectation Maximization Algorithm (EM) และ วิธี Predictive Mean Matching Imputation (PMM) ) การศึกษานี้ได้ควบคุมปัจจัยความแปรปรวนของตัวแปรอิสระ, ความสัมพันธ์ของตัวแปรอิสระ, ส่วนเบี่ยงเบนมาตรฐานค่าความคลาดเคลื่อน, ร้อยละการสูญหายและระดับ Nonignorability โดยวิธีการที่ให้ค่าเฉลี่ยของค่าเฉลี่ยความคลาดเคลื่อนกำลังสอง (Average mean square error) น้อยที่สุดจะเป็นวิธีการที่มีประสิทธิภาพสูงที่สุด ผลการวิจัยพบว่า เมื่อข้อมูลมีการกระจายตัวสูงและกลางวิธี KNN มีประสิทธิภาพสูงสุดในทุกกรณีที่ศึกษา แต่ถ้าข้อมูลกระจายตัวต่ำ วิธี KNN จะดีเมื่อกรณีตัวแปรมีความสัมพันธ์กันสูงและร้อยละการสูญหายต่ำ วิธี EM จะประสิทธิภาพสูงเมื่อร้อยละการสูญหายสูงในทุกระดับความสัมพันธ์ วิธี ERI จะประสิทธิภาพสูงเมื่อตัวแปรมีความสัมพันธ์เชิงบวกในระดับกลางลงไปในเกือบทุกกรณีที่ศึกษา วิธี CERI จะประสิทธิภาพสูงเมื่อตัวแปรมีความสัมพันธ์เชิงลบในระดับกลางลงไปและร้อยละการสูญหายต่ำ
An Analysis On Trends Of Research Topics In Civic Education Using Dynamic Topic Model, Poon Thongsai
An Analysis On Trends Of Research Topics In Civic Education Using Dynamic Topic Model, Poon Thongsai
Chulalongkorn University Theses and Dissertations (Chula ETD)
The aim of this thesis is to study the trend of civic and citizenship education research from 2000 to 2020 and the influence the regional background of researches has on the research discussion. Relevant data is collected from ERIC and SCOPUS database. This includes abstracts, published year, regional background of researchers, and author h-index. The keywords used are “civic education” or “citizenship education” or “civics”. There are 4917 papers extracted in total. Upon doing further preparation, 4854 articles are prepared for analysis. We apply Structural Topic model (STM) technique to the abstracts with covariates including the published year and the …
A Comparison Of Imbalanced Data Handling Methods For Pre-Trained Model In Multi-Label Classification Of Stack Overflow, Arisa Umparat
A Comparison Of Imbalanced Data Handling Methods For Pre-Trained Model In Multi-Label Classification Of Stack Overflow, Arisa Umparat
Chulalongkorn University Theses and Dissertations (Chula ETD)
Tag classification is essential in Stack Overflow. Instead of combining through pages or replies of irrelevant information, users can easily and quickly pinpoint relevant posts and answers using tags. Since User-submitted posts can have multiple tags, classifying tags in Stack Overflow can be challenging. This results in an imbalance problem between labels in the whole labelset. Pretrained deep learning models with small datasets can improve tag classification accuracy. Common multi-label resampling techniques with machine learning classifiers can also fix this issue. Still, few studies have explored which resampling technique can improve the performance of pre-trained deep models for predicting tags. …
An Application Of Reinforcement Learning To Credit Scoring Based On The Logistic Bandit Framework, Kantapong Visantavarakul
An Application Of Reinforcement Learning To Credit Scoring Based On The Logistic Bandit Framework, Kantapong Visantavarakul
Chulalongkorn University Theses and Dissertations (Chula ETD)
This study applies reinforcement learning to credit scoring by using the logistic bandit framework. The credit scoring and the credit underwriting are modeled into a single sequential decision problem where the credit underwriter takes a sequence of actions over an indefinite number of time steps. The traditional credit scoring approach considers the model construction separately from the underwriting process. This approach is identified as a greedy algorithm in the reinforcement learning literature, which is commonly believed to be inferior to an efficient reinforcement learning approach such as Thompson sampling. This is true under the simple setting, i.e., granting credit to …
Multi-Output Learning For Predicting Evaluation And Reopening Of Github Pull Requests On Open-Source Projects, Peerachai Banyongrakkul
Multi-Output Learning For Predicting Evaluation And Reopening Of Github Pull Requests On Open-Source Projects, Peerachai Banyongrakkul
Chulalongkorn University Theses and Dissertations (Chula ETD)
GitHub's pull-based development model is widely used by software development teams to manage software complexity. Contributors create pull requests for merging changes into the main codebase, and integrators review these requests to maintain quality and stability. However, a high volume of pull requests can overburden integrators, causing feedback delays. Previous studies have used machine learning and statistical techniques with tabular data as features, but these may lose meaningful information. Additionally, acceptance and latency may not be sufficient for the pull request evaluation. Moreover, reopened pull requests can add maintenance costs and burden already-busy developers. This thesis proposes a novel multi-output …
Spatio-Temporal Copula-Based Graph Neural Networks For Traffic Forecasting, Pitikorn Khlaisamniang
Spatio-Temporal Copula-Based Graph Neural Networks For Traffic Forecasting, Pitikorn Khlaisamniang
Chulalongkorn University Theses and Dissertations (Chula ETD)
Modern cities heavily rely on complex transportation, making accurate traffic speed prediction crucial for traffic management authorities. Classical methods, including statistical techniques and traditional machine learning techniques, fail to capture complex relationships, while deep learning approaches may have weaknesses such as error accumulation, difficulty in handling long sequences, and overlooking spatial correlations. Graph neural networks (GNNs) have shown promise in extracting spatial features from non-Euclidean graph structures, but they usually initialize the adjacency matrix based on distance and may fail to detect hidden statistical correlations. The choice of correlation measure can have a significant impact on the resulting adjacency matrix …
Energy Integrated Ratio Analysis Of The Anomalous Precession Frequency In The Fermilab Muon G-2 Experiment, Ritwika Chakraborty
Energy Integrated Ratio Analysis Of The Anomalous Precession Frequency In The Fermilab Muon G-2 Experiment, Ritwika Chakraborty
Theses and Dissertations--Physics and Astronomy
The muon’s anomalous magnetic moment, aμ, provides a unique way for probing physics beyond the standard model experimentally as it gathers contributions from all the known and unknown forces and particles in nature. The theoretical prediction of aμ has been in greater than 3 σ tension with the experimental measurement since the results of the Muon g-2 Experiment at the Brookhaven National Laboratory (E-821) were published in the early 2000s with a precision of 540 ppb. To settle this tension, the new Fermilab Muon g - 2 Experiment (E-989) is currently taking data with the aim of …
Estimating Weighted Panel Sizes For Primary Care Providers: An Assessment Of Clustering And Novel Methods Of Panel Size Estimation On Electronic Medical Records, Martin A. Lavallee
Estimating Weighted Panel Sizes For Primary Care Providers: An Assessment Of Clustering And Novel Methods Of Panel Size Estimation On Electronic Medical Records, Martin A. Lavallee
Theses and Dissertations
Primary Care is on the frontlines of healthcare, thus they see the most diverse set of patients. In order to achieve high functioning primary care, a practice must establish empanelment, the pairing of patients to providers. Enumeration of empanelment, or estimating panel sizes, helps ensure that the demands of the patients demand the supply of providers and optimize the balance of primary care resources to improve quality of care. Further we can adjust panel sizes by using patient-level data on healthcare utilization and complexity extracted from the electronic medial record to determine the amount of care or burden of work …
Improving College Students’ Views And Beliefs Relative To Mathematics: A Systematic Literature Review Followed By A Multiple Case Mixed Methods Exploration Of The Experiences That Underpin Community College Students’ Attitudes, Self-Efficacy, And Values In Mathematics, Marquita H. Sea
Theses and Dissertations
Mathematics is particularly important due to its relevance in our daily lives. It is a general requirement throughout schooling. Unfortunately, many students openly declare negative views/beliefs regarding math in their personal and academic lives. These in turn, negatively influence students’ achievement related behaviors and outcomes. First, a systematic literature review was conducted to determine what types of studies/initiatives have aimed to enhance students’ views/beliefs relative to mathematics, including domain general and specific perceptions of math as well as their judgements of who is successful in mathematics and if they themselves can be successful. Specifically, the review centered on the components …
Estimating The Statistics Of Operational Loss Through The Analyzation Of A Time Series, Maurice L. Brown
Estimating The Statistics Of Operational Loss Through The Analyzation Of A Time Series, Maurice L. Brown
Theses and Dissertations
In the world of finance, appropriately understanding risk is key to success or failure because it is a fundamental driver for institutional behavior. Here we focus on risk as it relates to the operations of financial institutions, namely operational risk. Quantifying operational risk begins with data in the form of a time series of realized losses, which can occur for a number of reasons, can vary over different time intervals, and can pose a challenge that is exacerbated by having to account for both frequency and severity of losses. We introduce a stochastic point process model for the frequency distribution …
Role Of Inhibition And Spiking Variability In Ortho- And Retronasal Olfactory Processing, Michelle F. Craft
Role Of Inhibition And Spiking Variability In Ortho- And Retronasal Olfactory Processing, Michelle F. Craft
Theses and Dissertations
Odor perception is the impetus for important animal behaviors, most pertinently for feeding, but also for mating and communication. There are two predominate modes of odor processing: odors pass through the front of nose (ortho) while inhaling and sniffing, or through the rear (retro) during exhalation and while eating and drinking. Despite the importance of olfaction for an animal’s well-being and specifically that ortho and retro naturally occur, it is unknown whether the modality (ortho versus retro) is transmitted to cortical brain regions, which could significantly instruct how odors are processed. Prior imaging studies show different …
M-Cubes: An Efficient And Portable Implementation Of Multi-Dimensional Integration For Gpus, Ioannis Sakiotis, Kamesh Arumugam, Marc Paterno, Desh Ranjan, Balŝa Terzić, Mohammad Zubair
M-Cubes: An Efficient And Portable Implementation Of Multi-Dimensional Integration For Gpus, Ioannis Sakiotis, Kamesh Arumugam, Marc Paterno, Desh Ranjan, Balŝa Terzić, Mohammad Zubair
Computer Science Faculty Publications
The task of multi-dimensional numerical integration is frequently encountered in physics and other scientific fields, e.g., in modeling the effects of systematic uncertainties in physical systems and in Bayesian parameter estimation. Multi-dimensional integration is often time-prohibitive on CPUs. Efficient implementation on many-core architectures is challenging as the workload across the integration space cannot be predicted a priori. We propose m-Cubes, a novel implementation of the well-known Vegas algorithm for execution on GPUs. Vegas transforms integration variables followed by calculation of a Monte Carlo integral estimate using adaptive partitioning of the resulting space. mCubes improves performance on GPUs by maintaining relatively …
Information Entropic Content Of Astrophysical Spectra: Applications To Cosmology And Astrobiology, Sara Vannah
Information Entropic Content Of Astrophysical Spectra: Applications To Cosmology And Astrobiology, Sara Vannah
Dartmouth College Ph.D Dissertations
Astrophysics faces two critical challenges: the difficulty of observing very distant targets and the difficulty of interpreting science in diverse and often extreme environments that have not been replicated on Earth. In this thesis, we discuss two types of spectra — one from early universe cosmology and one from astrobiology — where improvements in telescope technology are just ushering in a wave of precise observations, addressing the first challenge. This accelerates the need for a solution to the second challenge. Traditional methods for analyzing these two spectra rely heavily on unsettled science, biasing results to match the input assumptions. In …
Lake Huron Shoreline Analysis, Shubham Satish Nandanwar
Lake Huron Shoreline Analysis, Shubham Satish Nandanwar
Theses and Dissertations (Comprehensive)
Lake Huron is a popular tourist destination and is home to several businesses and residents. Since the shoreline is dynamic and is subject to change over the years due to several factors such as a change in water level, soil type, human encroachment, etc., these locations tend to encounter floods due to increased water levels and wind speed. This causes erosion and loss to the properties along the shoreline.
This study is based on two areas of interest named Pinery Provincial Park and Sauble Beach which are located on the shoreline of Lake Huron where Pinery Provincial Park is a …
Exploring Cyberterrorism, Topic Models And Social Networks Of Jihadists Dark Web Forums: A Computational Social Science Approach, Vivian Fiona Guetler
Exploring Cyberterrorism, Topic Models And Social Networks Of Jihadists Dark Web Forums: A Computational Social Science Approach, Vivian Fiona Guetler
Graduate Theses, Dissertations, and Problem Reports (ETD)
This three-article dissertation focuses on cyber-related topics on terrorist groups, specifically Jihadists’ use of technology, the application of natural language processing, and social networks in analyzing text data derived from terrorists' Dark Web forums. The first article explores cybercrime and cyberterrorism. As technology progresses, it facilitates new forms of behavior, including tech-related crimes known as cybercrime and cyberterrorism. In this article, I provide an analysis of the problems of cybercrime and cyberterrorism within the field of criminology by reviewing existing literature focusing on (a) the issues in defining terrorism, cybercrime, and cyberterrorism, (b) ways that cybercriminals commit a crime in …
A Monte Carlo Simulation Of Rat Choice Behavior With Interdependent Outcomes, Michelle A. Frankot
A Monte Carlo Simulation Of Rat Choice Behavior With Interdependent Outcomes, Michelle A. Frankot
Graduate Theses, Dissertations, and Problem Reports (ETD)
Preclinical behavioral neuroscience often uses choice paradigms to capture psychiatric symptoms. In particular, the subfield of operant research produces nested datasets with many discrete choices in a session. The standard analytic practice is to aggregate choice into a continuous variable and analyze using ANOVA or linear regression. However, choice data often have multiple interdependent outcomes of interest, violating an assumption of general linear models. The aim of the current study was to quantify the accuracy of linear mixed-effects regression (LMER) for analyzing data from a 4-choice operant task called the Rodent Gambling Task (RGT), which measures decision-making in the context …
Searching For Anomalous Extensive Air Showers Using The Pierre Auger Observatory Fluorescence Detector, Andrew Puyleart
Searching For Anomalous Extensive Air Showers Using The Pierre Auger Observatory Fluorescence Detector, Andrew Puyleart
Dissertations, Master's Theses and Master's Reports
Anomalous extensive air showers have yet to be detected by cosmic ray observatories. Fluorescence detectors provide a way to view the air showers created by cosmic rays with primary energies reaching up to hundreds of EeV . The resulting air showers produced by these highly energetic collisions can contain features that deviate from average air showers. Detection of these anomalous events may provide information into unknown regions of particle physics, and place constraints on cross-sectional interaction lengths of protons. In this dissertation, I propose measurements of extensive air shower profiles that are used in a machine learning pipeline to distinguish …
Modeling Digital Camera Monitoring Count Data With Intermittent Zeros For Short-Term Prediction, Eben Afrifa-Yamoah, Ute Mueller
Modeling Digital Camera Monitoring Count Data With Intermittent Zeros For Short-Term Prediction, Eben Afrifa-Yamoah, Ute Mueller
Research outputs 2022 to 2026
Digital camera monitoring has revolutionised survey designs in many fields, as an important source of information. The extended sampling coverage offered by this monitoring scheme makes it preferable compared to other traditional methods of survey. However, data obtained from digital camera monitoring are often highly variable, and characterized by sparse periods of zero counts, interspersed with missing observations due to outages. In practice, missing data of relatively shorter duration are mostly observed and are often imputed using interpolation techniques, ignoring long-term trends leading to inherent estimation biases. In this study, we investigated time series forecasting methods that adequately handle intermittency …
The Three-Way Interplay Among Early Life Exposures, The Gut Microbiome, And Outcomes In Infancy, Yuka Moroishi
The Three-Way Interplay Among Early Life Exposures, The Gut Microbiome, And Outcomes In Infancy, Yuka Moroishi
Dartmouth College Ph.D Dissertations
The bidirectional relationship between the gut microbiome and immune system plays an important role in host immune status: the immune system provides the gut microbiome the optimal environment to thrive in, and the gut microbiome helps regulate the immune system. This relationship is especially important in infants, whose immune system is still premature and rely on innate immunity.
We investigated the three-way interplay among early-life exposures, the developing gut microbiome, and outcomes in infancy from the general population in New Hampshire, US. We used prospective cohort data from the New Hampshire Birth Cohort study to 1) determine whether timing of …
Optimizing Model Observer Performance In Learning-Based Ct Reconstruction, Gregory Ongie, Emil Y. Sidky, Ingrid S. Reiser, Xiaochuan Pan
Optimizing Model Observer Performance In Learning-Based Ct Reconstruction, Gregory Ongie, Emil Y. Sidky, Ingrid S. Reiser, Xiaochuan Pan
Mathematical and Statistical Science Faculty Research and Publications
Deep neural networks used for reconstructing sparse-view CT data are typically trained by minimizing a pixel- wise mean-squared error or similar loss function over a set of training images. However, networks trained with such losses are prone to wipe out small, low-contrast features that are critical for screening and diagnosis. To remedy this issue, we introduce a novel training loss inspired by the model observer framework to enhance the detectability of weak signals in the reconstructions. We evaluate our approach on the reconstruction of synthetic sparse-view breast CT data, and demonstrate an improvement in signal detectability with the proposed loss.
A Participatory Group Process To Collect And Disseminate Covid-19 Needs Assessment Data, Areebah Ahmed
A Participatory Group Process To Collect And Disseminate Covid-19 Needs Assessment Data, Areebah Ahmed
Undergraduate Research Posters
The Richmond, VA COVID-19 Needs Assessment Survey (RVA CoNA) was created in March 2020 to identify behaviors and needs related to COVID-19 in Richmond area adults ages 18 and over. Results are being used to inform support, strategic efforts, and educational outreach of local community organizations. The purpose of this study is to (1) summarize the process used to develop the RVA CoNA, (2) summarize preliminary survey results from a second phase of data collection as well as initial feedback from community partners, and (3) summarize initial conclusions and results dissemination strategies.Community partners and researchers at Virginia Commonwealth University jointly …