Open Access. Powered by Scholars. Published by Universities.®

Statistics and Probability Commons™

Open Access. Powered by Scholars. Published by Universities.®

2022

Discipline
Institution
Keyword
Publication
Publication Type
File Type

Articles 571 - 595 of 595

Full-Text Articles in Statistics and Probability

การเปรียบเทียบสถาปัตยกรรมโครงข่ายประสาทคอนโวลูชัน 3 มิติ โดยการจำแนกโรคหลอดเลือดสมองจากภาพการฉายรังสีเอกซเรย์สมอง, ชานนท์ วรโชติสืบตระกูล Jan 2022

การเปรียบเทียบสถาปัตยกรรมโครงข่ายประสาทคอนโวลูชัน 3 มิติ โดยการจำแนกโรคหลอดเลือดสมองจากภาพการฉายรังสีเอกซเรย์สมอง, ชานนท์ วรโชติสืบตระกูล

Chulalongkorn University Theses and Dissertations (Chula ETD)

แบบจำลองโครงข่ายคอนโวลูชัน หรือ ซีเอ็นเอ็น (Convolutional Neural Networks หรือ CNN) ได้รับการนำมาใช้กันอย่างแพร่หลายในการจำแนกภาพ โดยเฉพาะในทางการแพทย์ ซึ่งปกติการจำแนกภาพทางการแพทย์นิยมใช้โครงข่ายคอนโวลูชั่น 2 มิติ แต่เนื่องจากข้อมูลภาพบางประเภท เช่น ภาพการฉายรังสีเอกซเรย์สมองมีลักษณะมองภาพ 3 มิติ ให้เป็นภาพ 2 มิติ ดังนั้นในงานวิจัยนี้จึงมีแนวคิดในการใช้โครงข่ายคอนโวลูชัน 3 มิติมาใช้ในการจำแนกภาพเพื่อนำเอาจุดเด่นจากความสามารถในการดึงคุณลักษณะความสัมพันธ์ในชั้นความลึกที่เพิ่มเข้ามาซึ่งมีความแตกต่างจากรูปแบบ 2 มิติ เพื่อเพิ่มประสิทธิภาพให้แบบจำลองสามารถดึงคุณลักษณะสำคัญของภาพให้มีความหลากหลายมากขึ้น งานวิจัยนี้มีวัตถุประสงค์เพื่อเปรียบเทียบประสิทธิภาพโครงข่ายคอนโวลูชัน 3 มิติ ร่วมกับแบบจำลองที่ถูกฝึกมาเรียบร้อยแล้ว (pre-trained model) 4 แบบจำลอง ประกอบไปด้วย อเล็กซ์เน็ต (Alexnet) วีจีจี-16 (Vgg-16) กูเกิลเน็ต (Googlenet) และเรสเน็ต (Resnet) เพื่อจำแนกข้อมูลภาพผู้ป่วยที่เป็นโรคหลอดเลือดสมอง และผู้ป่วยที่มีสุขภาพปกติ จากภาพฉายรังสีเอกซเรย์สมอง (CT-Scan) จากฐานข้อมูลเว็บไซด์ Kaggle ชุดข้อมูลประกอบด้วยภาพผู้ป่วยที่เป็นโรคหลอดเลือดสมอง 950 ภาพ จาก 40 คน และภาพผู้ป่วยสุขภาพปกติ 1551 ภาพ จาก 82 คน ซึ่งงานวิจัยนี้มีการปรับรายละเอียดโดยการนำจุดเด่นของแต่ละแบบจำลองมาใช้ และเพิ่มชั้นความลึกที่เป็นจุดเด่นของการค้นหาคุณลักษณะสำคัญของรูปแบบ 3 มิติ ร่วมกับการประมวลผลภาพล่วงหน้า (Image Preprocessing) และการทำการเพิ่มจำนวนข้อมูล (Data augmentation) เพื่อเพิ่มประสิทธิภาพของแบบจำลอง จากนั้นเพื่อไม่ให้การทดลองโน้มเอียงต่อแต่ละแบบจำลอง มีการนำเทคนิค K-Fold Cross validation (K=5) มาเพื่อแก้ปัญหาในงานวิจัยชิ้นนี้ ในส่วนของการวัดประสิทธิภาพผลการทดลองใช้ Confusion matrix เป็นเครื่องมือในการประเมินประสิทธิภาพของแบบจำลอง ซึ่งพบว่าสมรรถนะแบบจำลองโครงข่ายคอนโวลูชันกูเกิลเน็ต 3 มิติ ให้ผลลัพธ์ที่ดีที่สุด โดยผลการทดสอบการจำแนกภาพผู้ป่วยที่เป็นโรคหลอดเลือดสมองจากภาพฉายรังสีเอกซเรย์ ให้ค่าความแม่นยำ ความเที่ยงตรง ค่าความครบถ้วน และ F1-Score ที่ 92.00% 94.01% 83.96% และ 88.70% …


การเรียนรู้แบบรวมกลุ่มด้วยตัวแบบที่แตกต่างกันแบบขนานสำหรับข้อมูลไม่สมดุล กรณีศึกษาข้อมูลเครดิตเยอรมัน, ศศิวิมล ศรีโรจน์ Jan 2022

การเรียนรู้แบบรวมกลุ่มด้วยตัวแบบที่แตกต่างกันแบบขนานสำหรับข้อมูลไม่สมดุล กรณีศึกษาข้อมูลเครดิตเยอรมัน, ศศิวิมล ศรีโรจน์

Chulalongkorn University Theses and Dissertations (Chula ETD)

งานวิจัยนี้มีวัตถุประสงค์เพื่อสร้างตัวแบบการเรียนรู้แบบรวมกลุ่มด้วยตัวแบบที่แตกต่างกันแบบขนาน (Bagging Heterogeneous Ensemble) และหาวิธีการลดมิติข้อมูลและวิธีการสุ่มตัวอย่างซ้ำที่เหมาะสมกับข้อมูลเครดิตเยอรมันที่มีอัตราส่วนความไม่สมดุลแตกต่างกัน 3 ค่าคือ 2.3, 10 และ 14 โดยวัดประสิทธิภาพด้วยตัวชี้วัด Accuracy, The area under the curve, F1-score, Precision, Brier score และ Kolmogorov-Smirnov และทดสอบทางสถิติเพื่อแสดงว่าประสิทธิภาพของตัวแบบมีความแตกต่างกัน ที่ระดับนัยสำคัญ 0.05 ผลการศึกษาพบว่าข้อมูลเครดิตเยอรมันที่มีอัตราส่วนความไม่สมดุลต่ำ (IR = 2.3) ตัวแบบ Logistic Regression ที่ใช้เทคนิค Linear Discriminant Analysis (LDA) และ Systematic Minority Over-Sampling Technique (SM) จะมีประสิทธิภาพเฉลี่ยดีที่สุดในการจำแนกประเภท ในส่วนของอัตราส่วนความไม่สมดุลกลาง (IR = 10) และ อัตราส่วนความไม่สมดุลสูง (IR = 14) วิธีการลดมิติข้อมูลและการสุ่มตัวอย่างซ้ำที่มีประสิทธิภาพคือ Linear Discriminant Analysis (LDA), Random Under-Sampling (RUS) และ Linear Discriminant Analysis (LDA), Borderline SMOTE (BSM) ตามลำดับ โดยที่การเรียนรู้แบบรวมกลุ่มด้วยตัวแบบที่แตกต่างกันแบบขนานมีประสิทธิภาพเฉลี่ยดีที่สุด ทั้งในกรณีที่มีและไม่มีวิธีการลดมิติข้อมูลและสุ่มตัวอย่างซ้ำของอัตราส่วนความไม่สมดุลกลางและสูง


ประสิทธิภาพของวิธีการจัดการข้อมูลไม่สมดุลสำหรับการจำแนกกลุ่มภายใต้เงื่อนไขที่แตกต่างกัน, กาญธนา ลออสิริกุล Jan 2022

ประสิทธิภาพของวิธีการจัดการข้อมูลไม่สมดุลสำหรับการจำแนกกลุ่มภายใต้เงื่อนไขที่แตกต่างกัน, กาญธนา ลออสิริกุล

Chulalongkorn University Theses and Dissertations (Chula ETD)

การวิจัยนี้มีจุดประสงค์เพื่อศึกษาปฏิสัมพันธ์ของวิธีการปรับสมดุลข้อมูลกับเงื่อนไขด้านขนาดตัวอย่าง เทคนิคการจำแนกข้อมูล จำนวนตัวแปรระหว่างกลุ่มตัวแปรจัดประเภทต่อกลุ่มตัวแปรต่อเนื่อง อัตราออด และร้อยละของจำนวนข้อมูลกลุ่มหลักต่อข้อมูลกลุ่มรองที่มีต่อประสิทธิภาพของการจำแนกกลุ่ม การปรับสมดุลของข้อมูลแบ่งออกเป็น 3 วิธี ได้แก่ (1) ไม่ปรับสมดุล (2) วิธี random oversampling และ (3) วิธีผสมผสานระหว่างรูปแบบสุ่มเกินและสุ่มลด (hybrid) โดยใช้แพคเกจ ROSE ส่วนเงื่อนไขด้านขนาดตัวอย่างแบ่งออกเป็น ขนาดตัวอย่างเท่ากับ 100 300 และ 500 หน่วย ด้านเทคนิคการจำแนกข้อมูล แบ่งออกเป็น 4 วิธี ได้แก่ (1) เคเนียร์เรสเนเบอร์ (2) การถดถอยโลจิสติก (3) แรนดอมฟอร์เรส และ (4) ซัพพอร์ตเวกเตอร์แมชชีน ตัวแปรจากการจำลองแบ่งออกเป็นตัวแปรตามซึ่งจำลองด้วยการถดถอยโลจิสติก ส่วนตัวแปรอิสระในการจำลองข้อมูลครั้งนี้จะกำหนดให้ใช้ตัวแปรอิสระจำลองทั้งหมด 8 ตัว โดยกำหนดให้มีจำนวนตัวแปรระหว่างกลุ่มตัวแปรจัดประเภทต่อกลุ่มตัวแปรต่อเนื่อง 3 กรณี คือ 4:4 5:3 และ 6:2 ในขณะที่ระดับของอัตราออด จะสุ่มค่าจากช่วง [1,2) หรือ [2,3) และร้อยละของข้อมูลระหว่างข้อมูลกลุ่มหลักต่อข้อมูลกลุ่มรอง แบ่งออกเป็น 2 กรณี ได้แก่ 60:40 และ 70:30 พิจารณาเกณฑ์ประสิทธิภาพของข้อมูลด้วยตัวชี้วัดความถูกต้องในการจำแนก ความไว และความจำเพาะ การจำลองแต่ละสถานการณ์จะทำซ้ำสถานการณ์ละ 500 รอบ การวิเคราะห์ปฏิสัมพันธ์ระหว่างวิธีการปรับสมดุลข้อมูลกับเงื่อนไขต่าง ๆ ใช้การวิเคราะห์ความแปรปรวนพหุคูณหลายทาง (n-way MANOVA) ผลการวิจัยพบว่า วิธีการปรับสมดุลข้อมูลมีปฏิสัมพันธ์แบบสองทางกับเงื่อนไขด้านขนาดตัวอย่าง ร้อยละของข้อมูลระหว่างข้อมูลกลุ่มหลักต่อข้อมูลกลุ่มรอง อัตราออด และเทคนิคการจำแนกข้อมูล และพบปฏิสัมพันธ์แบบสามทางกับเงื่อนไขต่อไปนี้ (1) ขนาดตัวอย่างและจำนวนตัวแปรระหว่างกลุ่มตัวแปรจัดประเภทต่อกลุ่มตัวแปรต่อเนื่อง (2) ขนาดตัวอย่างและเทคนิคการจำแนกข้อมูล และ (3) ร้อยละของข้อมูลระหว่างข้อมูลกลุ่มหลักต่อข้อมูลกลุ่มรอง และเทคนิคการจำแนกข้อมูล ดังนั้นนักวิเคราะห์ข้อมูลควรเลือกใช้วิธีการปรับสมดุลข้อมูลโดยพิจารณาให้เหมาะสมกับสภาพของข้อมูลที่ใช้ในการวิเคราะห์


สมรรถนะดิจิทัลขององค์กรทหาร: การวิเคราะห์องค์ประกอบเชิงสำรวจและเชิงยืนยันพหุระดับ, รัมณรา สมประสงค์ Jan 2022

สมรรถนะดิจิทัลขององค์กรทหาร: การวิเคราะห์องค์ประกอบเชิงสำรวจและเชิงยืนยันพหุระดับ, รัมณรา สมประสงค์

Chulalongkorn University Theses and Dissertations (Chula ETD)

สมรรถนะดิจิทัลขององค์กรทหารในปัจจุบันมีความสำคัญต่อการปฏิบัติงานในยุคของการเปลี่ยนแปลงทางดิจิทัลที่เกิดขึ้นอย่างรวดเร็ว งานวิจัยนี้เป็นงานวิจัยเชิงบรรยาย มีวัตถุประสงค์ดังนี้ 1) เพื่อสังเคราะห์ตัวชี้วัดสมรรถนะดิจิทัลขององค์กรทหาร 2) เพื่อสำรวจองค์ประกอบพหุระดับสมรรถนะดิจิทัลขององค์กรทหาร 3) เพื่อตรวจสอบความสอดคล้องเชิงประจักษ์ขององค์ประกอบพหุระดับสมรรถนะดิจิทัลขององค์กรทหาร ตัวอย่างวิจัย เป็นบุคลากรระดับปฏิบัติการในองค์กรทหารสังกัดกระทรวงกลาโหม 50 หน่วยงาน จำนวน 860 คน สำหรับใช้ในการวิเคราะห์องค์ประกอบเชิงสำรวจพหุระดับ และจำนวน 863 คน สำหรับใช้ในการวิเคราะห์องค์ประกอบเชิงยืนยันพหุระดับ เครื่องมือที่ใช้ในการวิจัยเพื่อการวิเคราะห์องค์ประกอบพหุระดับ คือแบบวัดสมรรถนะดิจิทัลขององค์กรทหาร ประกอบไปด้วย 2 ตอน คือ ข้อมูลพื้นฐานของกำลังพลผู้ตอบแบบสอบถาม และแบบวัดสมรรถนะดิจิทัลขององค์กรทหาร จำนวน 69 ข้อคำถาม วิเคราะห์ข้อมูลด้วยโปรแกรม IBM SPSS Statistics 22 และ MPlus6 ผลการวิจัยพบว่า (1) ตัวชี้วัดสมรรถนะดิจิทัลขององค์กรทหาร ประกอบไปด้วย 16 ตัวชี้วัด ได้แก่ 1) การวางแผนการใช้งานอุปกรณ์เทคโนโลยีดิจิทัลแบบบูรณาการ 2) การสืบค้นข้อมูลทางดิจิทัล 3) การประเมินความน่าเชื่อถือของข้อมูล 4) การใช้งานเทคโนโลยีเบื้องต้น 5) การแก้ปัญหาจากการใช้เทคโนโลยีดิจิทัล 6) การใช้อินทราเน็ตขององค์กร 7) การรักษาความลับในโลกไซเบอร์ 8) การจัดการไฟล์ดิจิทัลทางการทหาร 9) การเข้าถึงไฟล์ดิจิทัลในกรณีปฏิบัติงานนอกสถานที่ 10) การจัดการฐานข้อมูลทางการทหาร 11) การใช้สื่อดิจิทัลทางไกลเพื่อการสื่อสารทางการทหาร 12) การนำเสนอข้อมูลทางทหารในรูปแบบดิจิทัล 13) การสร้างสิ่งแวดล้อมทางดิจิทัลเพื่อการทำงาน 14) การตระหนักถึงความปลอดภัยบนโลกไซเบอร์ 15) การรักษามารยาทในสังคมดิจิทัล 16) เจตคติต่อการใช้เทคโนโลยีดิจิทัลในองค์กร (2) องค์ประกอบเชิงสำรวจพหุระดับสมรรถนะดิจิทัลขององค์กรทหาร มีจำนวน 3 โมเดล คือ 1) องค์ประกอบระดับระดับบุคคล 4 องค์ประกอบ ระดับองค์กร 1 องค์ประกอบ 2) องค์ประกอบระดับบุคคล 4 องค์ประกอบ ระดับองค์กร 2 องค์ประกอบ 3) องค์ประกอบระดับบุคคล …


การเปรียบเทียบวิธีการใส่ค่าสูญหาย ในตัวแบบการถดถอยเชิงเส้นพหุที่ตัวแปรอิสระมีการสูญหายแบบนอนอิกนอร์เรเบิลที่สัมพันธ์กัน, ศุภสันติ์ ดีมาก Jan 2022

การเปรียบเทียบวิธีการใส่ค่าสูญหาย ในตัวแบบการถดถอยเชิงเส้นพหุที่ตัวแปรอิสระมีการสูญหายแบบนอนอิกนอร์เรเบิลที่สัมพันธ์กัน, ศุภสันติ์ ดีมาก

Chulalongkorn University Theses and Dissertations (Chula ETD)

งานวิจัยนี้มีจุดประสงค์เพื่อศึกษาและเปรียบเทียบวิธีการประมาณสูญหายในตัวแบบการถดถอยเชิงเส้นพหุคูณ ที่ตัวแปรอิสระมีการสูญหายแบบนอนอิกนอร์เรเบิลที่มีความสัมพันธ์กัน ในการศึกษานี้มีวิธีการที่ถูกพัฒนาขึ้นคือ Expected Regression Imputation (ERI) และ Conditional Expected Regression Imputation (CERI) โดยจะเปรียบเทียบประสิทธิภาพวิธีการที่พัฒนาขึ้นมากับอีก 3 วิธีการ ได้แก่ วิธี K-Nearest Neighbor Imputation (KNN), วิธี Expectation Maximization Algorithm (EM) และ วิธี Predictive Mean Matching Imputation (PMM) ) การศึกษานี้ได้ควบคุมปัจจัยความแปรปรวนของตัวแปรอิสระ, ความสัมพันธ์ของตัวแปรอิสระ, ส่วนเบี่ยงเบนมาตรฐานค่าความคลาดเคลื่อน, ร้อยละการสูญหายและระดับ Nonignorability โดยวิธีการที่ให้ค่าเฉลี่ยของค่าเฉลี่ยความคลาดเคลื่อนกำลังสอง (Average mean square error) น้อยที่สุดจะเป็นวิธีการที่มีประสิทธิภาพสูงที่สุด ผลการวิจัยพบว่า เมื่อข้อมูลมีการกระจายตัวสูงและกลางวิธี KNN มีประสิทธิภาพสูงสุดในทุกกรณีที่ศึกษา แต่ถ้าข้อมูลกระจายตัวต่ำ วิธี KNN จะดีเมื่อกรณีตัวแปรมีความสัมพันธ์กันสูงและร้อยละการสูญหายต่ำ วิธี EM จะประสิทธิภาพสูงเมื่อร้อยละการสูญหายสูงในทุกระดับความสัมพันธ์ วิธี ERI จะประสิทธิภาพสูงเมื่อตัวแปรมีความสัมพันธ์เชิงบวกในระดับกลางลงไปในเกือบทุกกรณีที่ศึกษา วิธี CERI จะประสิทธิภาพสูงเมื่อตัวแปรมีความสัมพันธ์เชิงลบในระดับกลางลงไปและร้อยละการสูญหายต่ำ


An Analysis On Trends Of Research Topics In Civic Education Using Dynamic Topic Model, Poon Thongsai Jan 2022

An Analysis On Trends Of Research Topics In Civic Education Using Dynamic Topic Model, Poon Thongsai

Chulalongkorn University Theses and Dissertations (Chula ETD)

The aim of this thesis is to study the trend of civic and citizenship education research from 2000 to 2020 and the influence the regional background of researches has on the research discussion. Relevant data is collected from ERIC and SCOPUS database. This includes abstracts, published year, regional background of researchers, and author h-index. The keywords used are “civic education” or “citizenship education” or “civics”. There are 4917 papers extracted in total. Upon doing further preparation, 4854 articles are prepared for analysis. We apply Structural Topic model (STM) technique to the abstracts with covariates including the published year and the …


A Comparison Of Imbalanced Data Handling Methods For Pre-Trained Model In Multi-Label Classification Of Stack Overflow, Arisa Umparat Jan 2022

A Comparison Of Imbalanced Data Handling Methods For Pre-Trained Model In Multi-Label Classification Of Stack Overflow, Arisa Umparat

Chulalongkorn University Theses and Dissertations (Chula ETD)

Tag classification is essential in Stack Overflow. Instead of combining through pages or replies of irrelevant information, users can easily and quickly pinpoint relevant posts and answers using tags. Since User-submitted posts can have multiple tags, classifying tags in Stack Overflow can be challenging. This results in an imbalance problem between labels in the whole labelset. Pretrained deep learning models with small datasets can improve tag classification accuracy. Common multi-label resampling techniques with machine learning classifiers can also fix this issue. Still, few studies have explored which resampling technique can improve the performance of pre-trained deep models for predicting tags. …


An Application Of Reinforcement Learning To Credit Scoring Based On The Logistic Bandit Framework, Kantapong Visantavarakul Jan 2022

An Application Of Reinforcement Learning To Credit Scoring Based On The Logistic Bandit Framework, Kantapong Visantavarakul

Chulalongkorn University Theses and Dissertations (Chula ETD)

This study applies reinforcement learning to credit scoring by using the logistic bandit framework. The credit scoring and the credit underwriting are modeled into a single sequential decision problem where the credit underwriter takes a sequence of actions over an indefinite number of time steps. The traditional credit scoring approach considers the model construction separately from the underwriting process. This approach is identified as a greedy algorithm in the reinforcement learning literature, which is commonly believed to be inferior to an efficient reinforcement learning approach such as Thompson sampling. This is true under the simple setting, i.e., granting credit to …


Multi-Output Learning For Predicting Evaluation And Reopening Of Github Pull Requests On Open-Source Projects, Peerachai Banyongrakkul Jan 2022

Multi-Output Learning For Predicting Evaluation And Reopening Of Github Pull Requests On Open-Source Projects, Peerachai Banyongrakkul

Chulalongkorn University Theses and Dissertations (Chula ETD)

GitHub's pull-based development model is widely used by software development teams to manage software complexity. Contributors create pull requests for merging changes into the main codebase, and integrators review these requests to maintain quality and stability. However, a high volume of pull requests can overburden integrators, causing feedback delays. Previous studies have used machine learning and statistical techniques with tabular data as features, but these may lose meaningful information. Additionally, acceptance and latency may not be sufficient for the pull request evaluation. Moreover, reopened pull requests can add maintenance costs and burden already-busy developers. This thesis proposes a novel multi-output …


Spatio-Temporal Copula-Based Graph Neural Networks For Traffic Forecasting, Pitikorn Khlaisamniang Jan 2022

Spatio-Temporal Copula-Based Graph Neural Networks For Traffic Forecasting, Pitikorn Khlaisamniang

Chulalongkorn University Theses and Dissertations (Chula ETD)

Modern cities heavily rely on complex transportation, making accurate traffic speed prediction crucial for traffic management authorities. Classical methods, including statistical techniques and traditional machine learning techniques, fail to capture complex relationships, while deep learning approaches may have weaknesses such as error accumulation, difficulty in handling long sequences, and overlooking spatial correlations. Graph neural networks (GNNs) have shown promise in extracting spatial features from non-Euclidean graph structures, but they usually initialize the adjacency matrix based on distance and may fail to detect hidden statistical correlations. The choice of correlation measure can have a significant impact on the resulting adjacency matrix …


Energy Integrated Ratio Analysis Of The Anomalous Precession Frequency In The Fermilab Muon G-2 Experiment, Ritwika Chakraborty Jan 2022

Energy Integrated Ratio Analysis Of The Anomalous Precession Frequency In The Fermilab Muon G-2 Experiment, Ritwika Chakraborty

Theses and Dissertations--Physics and Astronomy

The muon’s anomalous magnetic moment, aμ, provides a unique way for probing physics beyond the standard model experimentally as it gathers contributions from all the known and unknown forces and particles in nature. The theoretical prediction of aμ has been in greater than 3 σ tension with the experimental measurement since the results of the Muon g-2 Experiment at the Brookhaven National Laboratory (E-821) were published in the early 2000s with a precision of 540 ppb. To settle this tension, the new Fermilab Muon g - 2 Experiment (E-989) is currently taking data with the aim of …


Estimating Weighted Panel Sizes For Primary Care Providers: An Assessment Of Clustering And Novel Methods Of Panel Size Estimation On Electronic Medical Records, Martin A. Lavallee Jan 2022

Estimating Weighted Panel Sizes For Primary Care Providers: An Assessment Of Clustering And Novel Methods Of Panel Size Estimation On Electronic Medical Records, Martin A. Lavallee

Theses and Dissertations

Primary Care is on the frontlines of healthcare, thus they see the most diverse set of patients. In order to achieve high functioning primary care, a practice must establish empanelment, the pairing of patients to providers. Enumeration of empanelment, or estimating panel sizes, helps ensure that the demands of the patients demand the supply of providers and optimize the balance of primary care resources to improve quality of care. Further we can adjust panel sizes by using patient-level data on healthcare utilization and complexity extracted from the electronic medial record to determine the amount of care or burden of work …


Improving College Students’ Views And Beliefs Relative To Mathematics: A Systematic Literature Review Followed By A Multiple Case Mixed Methods Exploration Of The Experiences That Underpin Community College Students’ Attitudes, Self-Efficacy, And Values In Mathematics, Marquita H. Sea Jan 2022

Improving College Students’ Views And Beliefs Relative To Mathematics: A Systematic Literature Review Followed By A Multiple Case Mixed Methods Exploration Of The Experiences That Underpin Community College Students’ Attitudes, Self-Efficacy, And Values In Mathematics, Marquita H. Sea

Theses and Dissertations

Mathematics is particularly important due to its relevance in our daily lives. It is a general requirement throughout schooling. Unfortunately, many students openly declare negative views/beliefs regarding math in their personal and academic lives. These in turn, negatively influence students’ achievement related behaviors and outcomes. First, a systematic literature review was conducted to determine what types of studies/initiatives have aimed to enhance students’ views/beliefs relative to mathematics, including domain general and specific perceptions of math as well as their judgements of who is successful in mathematics and if they themselves can be successful. Specifically, the review centered on the components …


Estimating The Statistics Of Operational Loss Through The Analyzation Of A Time Series, Maurice L. Brown Jan 2022

Estimating The Statistics Of Operational Loss Through The Analyzation Of A Time Series, Maurice L. Brown

Theses and Dissertations

In the world of finance, appropriately understanding risk is key to success or failure because it is a fundamental driver for institutional behavior. Here we focus on risk as it relates to the operations of financial institutions, namely operational risk. Quantifying operational risk begins with data in the form of a time series of realized losses, which can occur for a number of reasons, can vary over different time intervals, and can pose a challenge that is exacerbated by having to account for both frequency and severity of losses. We introduce a stochastic point process model for the frequency distribution …


Role Of Inhibition And Spiking Variability In Ortho- And Retronasal Olfactory Processing, Michelle F. Craft Jan 2022

Role Of Inhibition And Spiking Variability In Ortho- And Retronasal Olfactory Processing, Michelle F. Craft

Theses and Dissertations

Odor perception is the impetus for important animal behaviors, most pertinently for feeding, but also for mating and communication. There are two predominate modes of odor processing: odors pass through the front of nose (ortho) while inhaling and sniffing, or through the rear (retro) during exhalation and while eating and drinking. Despite the importance of olfaction for an animal’s well-being and specifically that ortho and retro naturally occur, it is unknown whether the modality (ortho versus retro) is transmitted to cortical brain regions, which could significantly instruct how odors are processed. Prior imaging studies show different …


M-Cubes: An Efficient And Portable Implementation Of Multi-Dimensional Integration For Gpus, Ioannis Sakiotis, Kamesh Arumugam, Marc Paterno, Desh Ranjan, Balŝa Terzić, Mohammad Zubair Jan 2022

M-Cubes: An Efficient And Portable Implementation Of Multi-Dimensional Integration For Gpus, Ioannis Sakiotis, Kamesh Arumugam, Marc Paterno, Desh Ranjan, Balŝa Terzić, Mohammad Zubair

Computer Science Faculty Publications

The task of multi-dimensional numerical integration is frequently encountered in physics and other scientific fields, e.g., in modeling the effects of systematic uncertainties in physical systems and in Bayesian parameter estimation. Multi-dimensional integration is often time-prohibitive on CPUs. Efficient implementation on many-core architectures is challenging as the workload across the integration space cannot be predicted a priori. We propose m-Cubes, a novel implementation of the well-known Vegas algorithm for execution on GPUs. Vegas transforms integration variables followed by calculation of a Monte Carlo integral estimate using adaptive partitioning of the resulting space. mCubes improves performance on GPUs by maintaining relatively …


Information Entropic Content Of Astrophysical Spectra: Applications To Cosmology And Astrobiology, Sara Vannah Jan 2022

Information Entropic Content Of Astrophysical Spectra: Applications To Cosmology And Astrobiology, Sara Vannah

Dartmouth College Ph.D Dissertations

Astrophysics faces two critical challenges: the difficulty of observing very distant targets and the difficulty of interpreting science in diverse and often extreme environments that have not been replicated on Earth. In this thesis, we discuss two types of spectra — one from early universe cosmology and one from astrobiology — where improvements in telescope technology are just ushering in a wave of precise observations, addressing the first challenge. This accelerates the need for a solution to the second challenge. Traditional methods for analyzing these two spectra rely heavily on unsettled science, biasing results to match the input assumptions. In …


Lake Huron Shoreline Analysis, Shubham Satish Nandanwar Jan 2022

Lake Huron Shoreline Analysis, Shubham Satish Nandanwar

Theses and Dissertations (Comprehensive)

Lake Huron is a popular tourist destination and is home to several businesses and residents. Since the shoreline is dynamic and is subject to change over the years due to several factors such as a change in water level, soil type, human encroachment, etc., these locations tend to encounter floods due to increased water levels and wind speed. This causes erosion and loss to the properties along the shoreline.

This study is based on two areas of interest named Pinery Provincial Park and Sauble Beach which are located on the shoreline of Lake Huron where Pinery Provincial Park is a …


Exploring Cyberterrorism, Topic Models And Social Networks Of Jihadists Dark Web Forums: A Computational Social Science Approach, Vivian Fiona Guetler Jan 2022

Exploring Cyberterrorism, Topic Models And Social Networks Of Jihadists Dark Web Forums: A Computational Social Science Approach, Vivian Fiona Guetler

Graduate Theses, Dissertations, and Problem Reports (ETD)

This three-article dissertation focuses on cyber-related topics on terrorist groups, specifically Jihadists’ use of technology, the application of natural language processing, and social networks in analyzing text data derived from terrorists' Dark Web forums. The first article explores cybercrime and cyberterrorism. As technology progresses, it facilitates new forms of behavior, including tech-related crimes known as cybercrime and cyberterrorism. In this article, I provide an analysis of the problems of cybercrime and cyberterrorism within the field of criminology by reviewing existing literature focusing on (a) the issues in defining terrorism, cybercrime, and cyberterrorism, (b) ways that cybercriminals commit a crime in …


A Monte Carlo Simulation Of Rat Choice Behavior With Interdependent Outcomes, Michelle A. Frankot Jan 2022

A Monte Carlo Simulation Of Rat Choice Behavior With Interdependent Outcomes, Michelle A. Frankot

Graduate Theses, Dissertations, and Problem Reports (ETD)

Preclinical behavioral neuroscience often uses choice paradigms to capture psychiatric symptoms. In particular, the subfield of operant research produces nested datasets with many discrete choices in a session. The standard analytic practice is to aggregate choice into a continuous variable and analyze using ANOVA or linear regression. However, choice data often have multiple interdependent outcomes of interest, violating an assumption of general linear models. The aim of the current study was to quantify the accuracy of linear mixed-effects regression (LMER) for analyzing data from a 4-choice operant task called the Rodent Gambling Task (RGT), which measures decision-making in the context …


Searching For Anomalous Extensive Air Showers Using The Pierre Auger Observatory Fluorescence Detector, Andrew Puyleart Jan 2022

Searching For Anomalous Extensive Air Showers Using The Pierre Auger Observatory Fluorescence Detector, Andrew Puyleart

Dissertations, Master's Theses and Master's Reports

Anomalous extensive air showers have yet to be detected by cosmic ray observatories. Fluorescence detectors provide a way to view the air showers created by cosmic rays with primary energies reaching up to hundreds of EeV . The resulting air showers produced by these highly energetic collisions can contain features that deviate from average air showers. Detection of these anomalous events may provide information into unknown regions of particle physics, and place constraints on cross-sectional interaction lengths of protons. In this dissertation, I propose measurements of extensive air shower profiles that are used in a machine learning pipeline to distinguish …


Modeling Digital Camera Monitoring Count Data With Intermittent Zeros For Short-Term Prediction, Eben Afrifa-Yamoah, Ute Mueller Jan 2022

Modeling Digital Camera Monitoring Count Data With Intermittent Zeros For Short-Term Prediction, Eben Afrifa-Yamoah, Ute Mueller

Research outputs 2022 to 2026

Digital camera monitoring has revolutionised survey designs in many fields, as an important source of information. The extended sampling coverage offered by this monitoring scheme makes it preferable compared to other traditional methods of survey. However, data obtained from digital camera monitoring are often highly variable, and characterized by sparse periods of zero counts, interspersed with missing observations due to outages. In practice, missing data of relatively shorter duration are mostly observed and are often imputed using interpolation techniques, ignoring long-term trends leading to inherent estimation biases. In this study, we investigated time series forecasting methods that adequately handle intermittency …


The Three-Way Interplay Among Early Life Exposures, The Gut Microbiome, And Outcomes In Infancy, Yuka Moroishi Jan 2022

The Three-Way Interplay Among Early Life Exposures, The Gut Microbiome, And Outcomes In Infancy, Yuka Moroishi

Dartmouth College Ph.D Dissertations

The bidirectional relationship between the gut microbiome and immune system plays an important role in host immune status: the immune system provides the gut microbiome the optimal environment to thrive in, and the gut microbiome helps regulate the immune system. This relationship is especially important in infants, whose immune system is still premature and rely on innate immunity.

We investigated the three-way interplay among early-life exposures, the developing gut microbiome, and outcomes in infancy from the general population in New Hampshire, US. We used prospective cohort data from the New Hampshire Birth Cohort study to 1) determine whether timing of …


Optimizing Model Observer Performance In Learning-Based Ct Reconstruction, Gregory Ongie, Emil Y. Sidky, Ingrid S. Reiser, Xiaochuan Pan Jan 2022

Optimizing Model Observer Performance In Learning-Based Ct Reconstruction, Gregory Ongie, Emil Y. Sidky, Ingrid S. Reiser, Xiaochuan Pan

Mathematical and Statistical Science Faculty Research and Publications

Deep neural networks used for reconstructing sparse-view CT data are typically trained by minimizing a pixel- wise mean-squared error or similar loss function over a set of training images. However, networks trained with such losses are prone to wipe out small, low-contrast features that are critical for screening and diagnosis. To remedy this issue, we introduce a novel training loss inspired by the model observer framework to enhance the detectability of weak signals in the reconstructions. We evaluate our approach on the reconstruction of synthetic sparse-view breast CT data, and demonstrate an improvement in signal detectability with the proposed loss.


A Participatory Group Process To Collect And Disseminate Covid-19 Needs Assessment Data, Areebah Ahmed Jan 2022

A Participatory Group Process To Collect And Disseminate Covid-19 Needs Assessment Data, Areebah Ahmed

Undergraduate Research Posters

The Richmond, VA COVID-19 Needs Assessment Survey (RVA CoNA) was created in March 2020 to identify behaviors and needs related to COVID-19 in Richmond area adults ages 18 and over. Results are being used to inform support, strategic efforts, and educational outreach of local community organizations. The purpose of this study is to (1) summarize the process used to develop the RVA CoNA, (2) summarize preliminary survey results from a second phase of data collection as well as initial feedback from community partners, and (3) summarize initial conclusions and results dissemination strategies.Community partners and researchers at Virginia Commonwealth University jointly …