Open Access. Powered by Scholars. Published by Universities.®
- Institution
-
- Singapore Management University (3555)
- Wright State University (631)
- Walden University (447)
- New Jersey Institute of Technology (143)
- University of Malaya (130)
-
- University of Nebraska at Omaha (119)
- Old Dominion University (108)
- California State University, San Bernardino (100)
- San Jose State University (89)
- University of Dayton (82)
- City University of New York (CUNY) (70)
- University of Dar es Salaam (63)
- Air Force Institute of Technology (61)
- University of Nebraska - Lincoln (60)
- University of South Florida (56)
- Kennesaw State University (54)
- Nova Southeastern University (52)
- Technological University Dublin (51)
- University of Arkansas, Fayetteville (46)
- Dakota State University (43)
- Claremont Colleges (42)
- California Polytechnic State University, San Luis Obispo (41)
- Institute of Business Administration (38)
- Western Kentucky University (36)
- Purdue University (35)
- Ateneo de Manila University (34)
- Governors State University (34)
- Portland State University (34)
- University of Arkansas Little Rock (33)
- University of Nevada, Las Vegas (32)
- Keyword
-
- Machine learning (122)
- Information technology (91)
- Data mining (90)
- Social media (83)
- Machine Learning (64)
-
- Cybersecurity (63)
- Deep learning (60)
- Twitter (60)
- Artificial intelligence (58)
- Semantic Web (53)
- Online learning (51)
- Databases (46)
- Cloud computing (45)
- Information Technology (45)
- Information retrieval (45)
- Classification (43)
- Database (42)
- Blockchain (41)
- Natural language processing (41)
- Ontology (41)
- Big data (40)
- Security (39)
- Technology (39)
- Computer science (38)
- Privacy (38)
- Algorithms (37)
- Clustering (37)
- Deep Learning (37)
- Information systems (37)
- Management (37)
- Publication Year
- Publication
-
- Research Collection School Of Computing and Information Systems (3436)
- Kno.e.sis Publications (540)
- Walden Dissertations and Doctoral Studies (447)
- Theses and Dissertations (129)
- Student Works (2000-2009) (120)
-
- Dissertations (113)
- Computer Science Faculty Publications (95)
- Computer Science and Engineering Faculty Publications (91)
- Theses Digitization Project (86)
- Master's Projects (68)
- Information Systems and Quantitative Analysis Faculty Proceedings & Presentations (64)
- Tanzania Journal of Engineering and Technology (TJET) (60)
- Dissertations and Theses Collection (Open Access) (58)
- USF Tampa Graduate Theses and Dissertations (51)
- Theses (48)
- CCAC Theses and Dissertations (43)
- Information Systems and Quantitative Analysis Faculty Publications (41)
- CGU Faculty Publications and Research (37)
- International Conference on Information and Communication Technologies (36)
- Open Educational Resources (35)
- Graduate Theses and Dissertations (34)
- Department of Information Systems & Computer Science Faculty Publications (33)
- All Capstone Projects (32)
- Masters Theses & Doctoral Dissertations (32)
- Conference papers (28)
- All Maxine Goodman Levin School of Urban Affairs Publications (27)
- UBT International Conference (23)
- Electronic Theses and Dissertations (22)
- Faculty Articles (22)
- Master's Theses (22)
- Publication Type
- File Type
Articles 211 - 240 of 7250
Full-Text Articles in Computer Sciences
Dual-Target Disjointed Cross-Domain Recommendation Mediated Via Latent User Preferences, Dinh Hieu Do, Hady Wirawan Lauw
Dual-Target Disjointed Cross-Domain Recommendation Mediated Via Latent User Preferences, Dinh Hieu Do, Hady Wirawan Lauw
Research Collection School Of Computing and Information Systems
Users often navigate multiple platforms online, each characterized by its own set of scarce data. Recommender systems face a significant challenge in such fragmented environments. This paper proposes a novel approach to enhance recommendation systems by leveraging connections across distinct yet conceptually similar datasets from multiple platforms. We introduce a unique scenario of dual-target overlapping-free cross-platform recommendation, presenting a bridging mechanism to mutually improve across platforms and learn latent user preferences. Our approach addresses the data sparsity prevalent in each platform and enhances recommendation quality by harnessing redundant, rich, and similar domain data. Experiments validate the effectiveness of our method, …
An Efficient Diffusion-Based Non-Autoregressive Solver For Traveling Salesman Problem, Mingzhao Wang, You Zhou, Zhiguang Cao, Yubin Xiao, Xuan Wu, Wei Pang, Yuan Jiang, Hui Yang, Peng Zhao, Yuanshu Li
An Efficient Diffusion-Based Non-Autoregressive Solver For Traveling Salesman Problem, Mingzhao Wang, You Zhou, Zhiguang Cao, Yubin Xiao, Xuan Wu, Wei Pang, Yuan Jiang, Hui Yang, Peng Zhao, Yuanshu Li
Research Collection School Of Computing and Information Systems
Recent advances in neural models have shown considerable promise in solving Traveling Salesman Problems (TSPs) without relying on much hand-crafted engineering. However, while non-autoregressive (NAR) approaches benefit from faster inference through parallelism, they typically deliver solutions of inferior quality compared to autoregressive ones. To enhance the solution quality while maintaining fast inference, we propose DEITSP, a diffusion model with efficient iterations tailored for TSP that operates in a NAR manner. Firstly, we introduce a one-step diffusion model that integrates the controlled discrete noise addition process with self-consistency enhancement, enabling optimal solution prediction through simultaneous denoising of multiple solutions. Secondly, we …
Sparse-To-Dense: A Free Lunch For Lossless Acceleration Of Video Understanding In Llms, Xuan Zhang, Cunxiao Du, Sicheng Yu, Jiawei Wu, Fengzhuo Zhang, Wei Gao, Qian Liu
Sparse-To-Dense: A Free Lunch For Lossless Acceleration Of Video Understanding In Llms, Xuan Zhang, Cunxiao Du, Sicheng Yu, Jiawei Wu, Fengzhuo Zhang, Wei Gao, Qian Liu
Research Collection School Of Computing and Information Systems
Due to the auto-regressive nature of current video large language models (Video-LLMs), the inference latency increases as the input sequence length grows, posing challenges for the efficient processing of video sequences that are usually very long. We observe that during decoding, the attention scores of most tokens in Video-LLMs tend to be sparse and concentrated, with only certain tokens requiring comprehensive full attention. Based on this insight, we introduce Sparse-to-Dense (StD), a novel decoding strategy that integrates two distinct modules: one leveraging sparse top-K attention and the other employing dense full attention. These modules collaborate to accelerate Video-LLMs without loss. …
Towards Metrology 4.0 In Developing Countries’ Manufacturing Industries, Jailos Nzumile
Towards Metrology 4.0 In Developing Countries’ Manufacturing Industries, Jailos Nzumile
Tanzania Journal of Engineering and Technology (TJET)
A systematic literature review was conducted to unveil the status of the digital transformation of metrology in developing countries, as they are lagging in utilising fourth industrial revolution (IR4.0) technologies to transform manufacturing industries. A PRISMA technique was employed using various keywords to identify, screen, and select the relevant literature. Forty publications were selected for the review, mainly discussing IR 4.0 technologies in metrological operations. The results indicate that the digital transformation of metrology has yet to be initiated in developing countries. However, the employment of IR4.0 technologies in advancing metrological operations in manufacturing industries is mostly discussed in the …
Towards Future Sustainable Infrastructure: The Role Of Technical Audit In Tanzania’S Public Works, George C. Haule
Towards Future Sustainable Infrastructure: The Role Of Technical Audit In Tanzania’S Public Works, George C. Haule
Tanzania Journal of Engineering and Technology (TJET)
This study aimed to investigate the vital role and impact of technical audits in promoting sustainable infrastructure development in Tanzania. The role and effects of technical audits in long-term infrastructure development were studied using a mixed-methods approach with both quantitative and qualitative parts. Data were collected through analysis of technical audit documentation, a semi-structured questionnaire, and stakeholder interviews. The study revealed the various dimensions of infrastructure investment projects, including initiation and planning, design, procurement of contractors and consultants, contract management, environment, health, and safety. The technical audit findings reported weaknesses or non-performance issues in infrastructure planning at the national level …
Assessment Of Digital Solutions For Conformity Assessment Of Legally Controlled Measuring Instruments In Tanzania, Faraja Nyoni
Assessment Of Digital Solutions For Conformity Assessment Of Legally Controlled Measuring Instruments In Tanzania, Faraja Nyoni
Tanzania Journal of Engineering and Technology (TJET)
The advent of state-of-the-art digital technologies since 2011 has led to the digital transformation of legal metrology practices to ensure the trustworthiness of software-controlled measuring instruments globally. Despite the digital transformation in legal metrological practices, the conformity assessment of legally controlled measuring instruments is manually done (i.e., paper-based) in Tanzania. The paper-based conformity assessment of legally controlled measuring instruments is prone to error and lacks efficiency and effectiveness. This study aimed to assess digital solutions for improving conformity assessment through a comprehensive survey conducted across various regions in Tanzania, targeting a stratified sample of 51 respondents from organizations involved in …
Adapting Large Language Models For Parameter-Efficient Log Anomaly Detection, Ying Fu Lim, Jiawen Zhu, Guansong Pang
Adapting Large Language Models For Parameter-Efficient Log Anomaly Detection, Ying Fu Lim, Jiawen Zhu, Guansong Pang
Research Collection School Of Computing and Information Systems
Log Anomaly Detection (LAD) seeks to identify atypical patterns in log data that are crucial to assessing the security and condition of systems. Although Large Language Models (LLMs) have shown tremendous success in various fields, the use of LLMs in enabling the detection of log anomalies is largely unexplored. This work aims to fill this gap. Due to the prohibitive costs involved in fully fine-tuning LLMs,we explore the use of parameter-efficient fine-tuning techniques (PEFTs) for adapting LLMs to LAD.To have an in-depth exploration of the potential of LLM-driven LAD, we present a comprehensive investigation of leveraging two of the most …
A Decision Support System For Conference Session Selection Using Natural Language Processing, Tillman E. Erb
A Decision Support System For Conference Session Selection Using Natural Language Processing, Tillman E. Erb
Master's Theses
Conference attendees are faced with selecting from hundreds to thousands of presentations and sessions in pursuit of new findings and methods relevant to their area of interest, an overwhelming amount of information from which to clearly make a decision. To address this, we developed a decision support system leveraging natural language processing (NLP) techniques such as semantic matching. By creating and matching embeddings of conference presentation abstracts and titles, the application provides improved query matching compared to keyword searching. We introduce Session Scout, a novel conference decision support system built upon a semantic retrieval framework. Session Scout is designed to …
Alayadb: The Data Foundation For Efficient And Effective Long-Context Llm Inference, Yangshen Deng, Zhengxin You, Long Xiang, Qilong Li, Peiqi Yuan, Zhaoyang Hong, Yitao Zheng, Wanting Li, Runzhong Li, Haotian Liu, Kyriakos Mouratidis, Man Lung Yiu, Huan Li, Qiaomu Shen, Rui Mao, Bo Tang
Alayadb: The Data Foundation For Efficient And Effective Long-Context Llm Inference, Yangshen Deng, Zhengxin You, Long Xiang, Qilong Li, Peiqi Yuan, Zhaoyang Hong, Yitao Zheng, Wanting Li, Runzhong Li, Haotian Liu, Kyriakos Mouratidis, Man Lung Yiu, Huan Li, Qiaomu Shen, Rui Mao, Bo Tang
Research Collection School Of Computing and Information Systems
AlayaDB is a cutting-edge vector database system natively architected for efficient and effective long-context inference for Large Language Models (LLMs) at AlayaDB AI. Specifically, it decouples the KV cache and attention computation from the LLM inference systems, and encapsulates them into a novel vector database system. For the Model as a Service providers (MaaS), AlayaDB consumes fewer hardware resources and offers higher generation quality for various workloads with different kinds of Service Level Objectives (SLOs), when compared with the existing alternative solutions (e.g., KV cache disaggregation, retrieval-based sparse attention). The crux of AlayaDB is that it abstracts the attention computation …
Rich Models And Methods For On-Demand Same Day Deliveries, Zhiqin Zhang
Rich Models And Methods For On-Demand Same Day Deliveries, Zhiqin Zhang
Dissertations and Theses Collection (Open Access)
Same-day delivery has brought numerous conveniences to people’s lives, but it has also presented challenges in terms of service management. To effectively optimize on-demand same-day delivery operations within urban logistics, intelligent decision-making strategies capable of adapting to rapidly changing circumstances are essential. Employing effective decisionmaking strategies that account for order allocation, route planning, courier scheduling, and other relevant factors, is pivotal in advancing logistics operations, enhancing efficiency, customer satisfaction, and resource utilization in the context of dynamic same-day delivery problems.
The focus of this thesis revolves around different emerging challenges presented by on-demand same-day delivery problems, with a particular emphasis …
Lessons Learned From Sandboxing, Piloting And Policy Experimentation With Ai And Other Digital Initiatives: Part 1, Summary Report, Steven M. Miller
Lessons Learned From Sandboxing, Piloting And Policy Experimentation With Ai And Other Digital Initiatives: Part 1, Summary Report, Steven M. Miller
Research Collection School Of Computing and Information Systems
This report, "Lessons Learned from Sandboxing, Piloting and Policy Experimentation with AI and Other Digital Initiatives," captures insights and experiences from project experts involved in recent digital innovation initiatives with the governments of Bangladesh, Maldives, and Kazakhstan, and from project experts actively involved with the use of AI for delivering government digital services in the EU, New Zealand, Rwanda, Singapore, United States, and Uzbekistan. The ten in-depth interview write-ups produced from these nine different country settings provide a small but highly informative sample of rich descriptions of some of the important realities, approaches, nuances, issues and challenges related to testing …
Hd-Epic: A Highly-Detailed Egocentric Video Dataset, Toby Perrett, Ahmad Darkhalil, Saptarshi Sinha, Omar Emara, Sam Pollard, Kranti Kumar Parida, Kaiting Liu, Prajwal Gatti, Siddhant Bansal, Kevin Flanagan, Jacob Chalk, Zhifan Zhu, Rhodri Guerrier, Fahd Abdelazim, Bin Zhu, Davide Moltisanti, Michael Wray, Hazel Doughty, Dima Damen
Hd-Epic: A Highly-Detailed Egocentric Video Dataset, Toby Perrett, Ahmad Darkhalil, Saptarshi Sinha, Omar Emara, Sam Pollard, Kranti Kumar Parida, Kaiting Liu, Prajwal Gatti, Siddhant Bansal, Kevin Flanagan, Jacob Chalk, Zhifan Zhu, Rhodri Guerrier, Fahd Abdelazim, Bin Zhu, Davide Moltisanti, Michael Wray, Hazel Doughty, Dima Damen
Research Collection School Of Computing and Information Systems
We present a validation dataset of newly-collected kitchenbased egocentric videos, manually annotated with highly detailed and interconnected ground-truth labels covering: recipe steps, fine-grained actions, ingredients with nutritional values, moving objects, and audio annotations. Importantly, all annotations are grounded in 3D through digital twinning of the scene, fixtures, object locations, and primed with gaze. Footage is collected from unscripted recordings in diverse home environments, making HDEPIC the first dataset collected in-the-wild but with detailed annotations matching those in controlled lab environments. We show the potential of our highly-detailed annotations through a challenging VQA benchmark of 26K questions assessing the capability to …
Predicting Consumers’ Purchase Intention Of Browsed Products: A Study Based On Eye‑Tracking, Feiyan Jia, Choon Ling Sia, Yani Shi, Fiona Fui-Hoon Nah, Keng Siau
Predicting Consumers’ Purchase Intention Of Browsed Products: A Study Based On Eye‑Tracking, Feiyan Jia, Choon Ling Sia, Yani Shi, Fiona Fui-Hoon Nah, Keng Siau
Research Collection School Of Computing and Information Systems
Predicting consumers’ purchase intention of browsed products enables sellers to implement nuanced promotion strategies to stimulate purchase. But how can we predict consumers’ purchase intention of browsed products? Our research demonstrates that consumers’ eye movement data collected when they browse products can serve this aim. We train and test the prediction model using logistic regression and random forest algorithms. Using data collected in a laboratory experiment, our empirical results show that both algorithms perform much better than a random guess, and the logistic regression performs slightly better than the random forest. Our findings imply that eye movement data enable sellers …
Lessons Learned From Sandboxing, Piloting And Policy Experimentation With Ai And Other Digital Initiatives: Part 2, Ten In-Depth Interviews, Steven M. Miller
Lessons Learned From Sandboxing, Piloting And Policy Experimentation With Ai And Other Digital Initiatives: Part 2, Ten In-Depth Interviews, Steven M. Miller
Research Collection School Of Computing and Information Systems
This report, "Lessons Learned from Sandboxing, Piloting and Policy Experimentation with AI and Other Digital Initiatives," captures insights and experiences from project experts involved in recent digital innovation initiatives with the governments of Bangladesh, Maldives, and Kazakhstan, and from project experts actively involved with the use of AI for delivering government digital services in the EU, New Zealand, Rwanda, Singapore, United States, and Uzbekistan. The ten in-depth interview write-ups produced from these nine different country settings provide a small but highly informative sample of rich descriptions of some of the important realities, approaches, nuances, issues and challenges related to testing …
Less Is More: On The Importance Of Data Quality For Unit Test Generation, Junwei Zhang, Xing Hu, Shan Gao, Xin Xia, David Lo, Shanping Li
Less Is More: On The Importance Of Data Quality For Unit Test Generation, Junwei Zhang, Xing Hu, Shan Gao, Xin Xia, David Lo, Shanping Li
Research Collection School Of Computing and Information Systems
Unit testing is crucial for software development and maintenance. Effective unit testing ensures and improves software quality, but writing unit tests is time-consuming and labor-intensive. Recent studies have proposed deep learning (DL) techniques or large language models (LLMs) to automate unit test generation. These models are usually trained or fine-tuned on large-scale datasets. Despite growing awareness of the importance of data quality, there has been limited research on the quality of datasets used for test generation. To bridge this gap, we systematically examine the impact of noise on the performance of learning-based test generation models. We first apply the open …
Large Language Models For Logical Fallacy Detection, Nicole Anne Hui-Ying Teo, Donghao Huang, Erik Cambria, Zhaoxia Wang
Large Language Models For Logical Fallacy Detection, Nicole Anne Hui-Ying Teo, Donghao Huang, Erik Cambria, Zhaoxia Wang
Research Collection School Of Computing and Information Systems
Identifying logical fallacies is essential for maintaining log-ical reasoning and reducing false information in a variety of domains, such as the media, law, and education. We present an extensive study on the use of large language models (LLMs) for logical fallacy detection and provide a comparative overview of model performance across various fallacy classes. We evaluate the logical fallacy detection capabilities of multiple state-of-the-art models (LLaMA, Qwen, Gemma, Phi) utilizing accuracy, precision, recall, and F1-score as assessment measures. Accord-ing to our findings, our models do well on simple fallacies like “circular reasoning,” but they have trouble with more interpretive reasoning …
Human-Computer Interaction And Artificial Intelligence For Ageing Population, Keng Siau, Hailiang Wang, Fiona Fui-Hoon Nah, Runyu Wang, Ruitong Che, Can Liu
Human-Computer Interaction And Artificial Intelligence For Ageing Population, Keng Siau, Hailiang Wang, Fiona Fui-Hoon Nah, Runyu Wang, Ruitong Che, Can Liu
Research Collection School Of Computing and Information Systems
As the global population ages rapidly, the field of human-computer interaction (HCI) is in urgent need of innovation, redesign, and reengineering to meet the evolving needs of older adults. The older demographic faces a range of challenges—including physical limitations, cognitive decline, reduced social in-tegration, and varying levels of technological literacy—that can hinder effective engagement with digital technologies. In response to these challenges, research-ers and designers are using inclusive and adaptive approaches to enhance acces-sibility, usability, and emotional well-being. This paper reviews key design prin-ciples in HCI for the ageing population and discusses how artificial intelligence (AI) tools, such as voice …
Unlocking The Power Of Socio-Knowledge Association For Enterprise Risk Identification In Stock Market, Zhenghao Liu, Keng Siau, Shaochen Yang, Feicheng Ma
Unlocking The Power Of Socio-Knowledge Association For Enterprise Risk Identification In Stock Market, Zhenghao Liu, Keng Siau, Shaochen Yang, Feicheng Ma
Research Collection School Of Computing and Information Systems
Potential risk signals reflected in supply chain and equity connections between enterprises and social connections between investors are becoming crucial to identifying enterprise risks in addition to basic financial indicators. Traditional risk management systems face challenges in adapting to these complexities, highlighting the need for a proactive paradigm shift in risk management. Leveraging graph models such as social networks and knowledge graphs offers a promising approach to identifying and managing potential associated risks effectively. To bridge existing research gaps, a novel risk identification framework driven by social-knowledge graphs has been proposed, integrating graph deep learning and reinforcement learning techniques guided …
Fact-Checking As A Multi-Step Process: From Ambiguity Resolution To Claim Validation, Wenbo Wang
Fact-Checking As A Multi-Step Process: From Ambiguity Resolution To Claim Validation, Wenbo Wang
Dissertations
The spread of misinformation and disinformation has become a major concern, particularly with the rise of social media as a primary source of information for many people. Fact-checking—the process of verifying claims against credible evidence—has emerged as a critical safeguard against misinformation. Yet, the task is fraught with challenges: claims are often ambiguous, context-dependent, or composed of multiple intertwined assertions, while automated systems struggle to replicate the nuanced reasoning of human experts. This dissertation addresses these challenges by reimagining fact-checking as a multi-step, knowledge-guided process that systematically resolves ambiguity, decomposes complexity, and validates claims through structured reasoning. Additionally, the proposed …
From Neural Networks To Large Language Models: Innovations In Financial Ai, Mathematical Reasoning, And Structured Data Representation, Junyi Ye
Dissertations
This dissertation explores the evolution and application of artificial intelligence techniques across three critical domains: financial modeling, mathematical reasoning, and structured data analysis. The dissertation presents seven research projects that chart a progression from specialized neural architectures to sophisticated large language models (LLMs), contributing novel methodologies and frameworks at each stage.
In the financial domain, the research first introduces TS-Mixer, a MLP-based architecture for time-series forecasting that captures both feature relationships and temporal dependencies through a simple yet effective design, outperforming more complex models in S&P500 index prediction. The dissertation then presents DySTAGE, a dynamic graph representation learning framework that …
Towards Explainable Ai On Graph Neural Networks: Xaig, Jiaxing Zhang
Towards Explainable Ai On Graph Neural Networks: Xaig, Jiaxing Zhang
Dissertations
In the evolving landscape of artificial intelligence (AI), Graph Neural Networks (GNNs) have garnered growing prominence for their adeptness in processing graph-structured data. Despite this, the interpretability of their predictions often remains elusive. The demand for transparency and explainability in complex prediction models has reached unprecedented levels. To address this, post-hoc instance-level explanation techniques have emerged, aiming to unveil the rationale behind GNN predictions. These techniques endeavor to unearth substructures that elucidate the predictive behavior of trained GNNs.
This dissertation embarks on an exploration of Explainable AI (XAI) technologies within the realm of GNNs. Amid the challenges posed by the …
Blockchain-Enabled Master Data Management, Shakhawat Hossain
Blockchain-Enabled Master Data Management, Shakhawat Hossain
Theses and Dissertations
Master Data Management (MDM) is essential for maintaining data quality, accuracy, consistency, and governance within organizations. However, traditional centralized MDM systems continue to face challenges related to data integrity, security, and scalability. This research presents a blockchain-enabled MDM framework designed to overcome these limitations by leveraging blockchain’s decentralized, immutable, and secure architecture. The study aims to identify and address the shortcomings of conventional MDM practices, examine the applicability of blockchain technology in enhancing these systems, and develop a functional prototype to validate the proposed model. The framework incorporates decentralized review mechanisms that improve auditability and ensure trusted data verification by …
Selected Artificial Intelligence Provisions In U.S. Fiscal Year 2025 National Defense Authorization Act, Bert Chapman
Selected Artificial Intelligence Provisions In U.S. Fiscal Year 2025 National Defense Authorization Act, Bert Chapman
Libraries Faculty and Staff Presentations
The 2025 Fiscal Year National Defense Authorization Act contains multiple provisions relating to artificial intelligence (AI). These congressionally mandated provisions direct various sections of the Department of Defense (DOD) and individual U.S. armed service branches to execute congressional intent for AI policymaking. Examples of such intent include identifying and planning DOD's AI workforce, demonstrating AI biotechnology applications for national security, improving the human usability of AI systems, and establishing an AI security center. This presentation will note that reports on these initiatives must be prepared for relevant congressional oversight committees, and, in many cases, are in many cases, publicly released …
A Computational Method For Detecting Compound Promiscuity In Early-Stage Pharmaceutical Discovery, John Allen Ringer
A Computational Method For Detecting Compound Promiscuity In Early-Stage Pharmaceutical Discovery, John Allen Ringer
Computer Science ETDs
Modern drug discovery and chemical biology research relies heavily on analyzing bioassay data. One of the many challenges in bioassay data analysis is identifying false trails, i.e., chemical compounds which initially appear to have desirable activity but are found to be problematic upon further investigation. Badapple (the BioAssay-Data Associative Promiscuity Pattern Learning Engine) was created over ten years ago to help researchers identify promiscuous compounds and thus avoid a common source of these false trails. Through an effort involving software engineering, cheminformatics, and biomedical data science we have developed Badapple 2.0, which incorporates updated assay records and expanded data semantics. …
From Data To Action: An Adaptable Crosstabs Template For Participatory Survey Data Analysis, Natalia Pinzon, Vikram Koundinya, William O'R Dowling, Ryan Galt
From Data To Action: An Adaptable Crosstabs Template For Participatory Survey Data Analysis, Natalia Pinzon, Vikram Koundinya, William O'R Dowling, Ryan Galt
Journal of Extension
We present a practical and accessible template for quantitative survey data analysis designed for non-academic researchers in order to facilitate engagement from community collaborators. The template, created in Google Sheets, is mainly for computing cross-tabulations, but it also displays frequency distributions and p-values for determining statistical significance. The template allows collaborators to record their observations and questions, promoting an efficient yet interactive review process and fostering a democratic analysis environment. Based on our own experience using this template for a data party, we highlight its effectiveness in promoting collaborative data interpretation, decision-making, and the actionable use of survey findings.
Full-Stack Web Applications: Industry Standard Frameworks, Libraries & Technologies, Yassine Chahid, Patrick Slattery
Full-Stack Web Applications: Industry Standard Frameworks, Libraries & Technologies, Yassine Chahid, Patrick Slattery
Publications and Research
This research explores emerging full-stack web development technologies across front-end, back-end, and DevSecOps domains. It evaluates modern tools including Django, React, and TypeScript—focusing on their key features such as compile-time error checking—through to the development of a web application. By examining documentation for the frameworks Node.js, Next.js, Tailwind CSS, and others, along with the deployment tools Docker and Git for version/release control, the study analyzes how these innovations speed up development, improve existing practices, and have often replaced older technologies. Cloud solutions for tasks such as authentication and deployment will also be evaluated, along with various web-application technology stacks and …
Computer Vision In Soccer: Yolov11 Analytics Engine For Quantifying Game Strategy, Connor S. Maurer
Computer Vision In Soccer: Yolov11 Analytics Engine For Quantifying Game Strategy, Connor S. Maurer
Data Science Undergraduate Honors Theses
Single-shot object detection capabilities significantly reduce computational overhead for real-time computer vision in sports analytics at 60 FPS. YOLO11’s lightweight CNN gives promising accuracy while meeting the low-latency demand of dynamic soccer matches. As data-driven approaches take over the sport of soccer, efficient player tracking systems become critical for informing coach’s strategies. I prototype the ETL (Extract, Transform, Load) process of data collected from a single- shot detection program and evaluate its viability for estimating player fatigue. YOLO11 detects players, the ball, and other characteristics, with the output transformed by homography to estimate the positions in the real world. These …
Learning Behaviors In Physics-Informed Deep Learning, Alex Glover
Learning Behaviors In Physics-Informed Deep Learning, Alex Glover
Electronic Theses and Dissertations
Physics-informed deep learning is a methodology in artificial intelligence aimed at combating the large training data requirement and the barrier of domain awareness that deep learning architectures commonly face in applications. Stochastic modeling integrated into the predictive models provides that domain knowledge. Variations of the Intelligent Driving Model impact the learning behaviors of the joint-training architecture. This thesis examines the effect of substituting the standard linear Intelligent Driving Model with a modified nonlinear version, as applied to real human driving behavior on the I-80 interstate. The experimentation also critically evaluates the complications that impede the viability of this architecture in …
Towards Reliable Ml: Data Attribution And Adversarial Robustness, Xiaosen Zheng
Towards Reliable Ml: Data Attribution And Adversarial Robustness, Xiaosen Zheng
Dissertations and Theses Collection (Open Access)
Modern machine learning (ML) models achieve remarkable success, but face critical reliability challenges. This thesis advances two pillars of reliable ML systems: interpretability through data attribution and robustness against adversarial threats.
In the first part, we develop novel data attribution methods to elucidate the data-model relationship. We establish the critical role of memorization in model generalization through token-level influence analysis, extend sample-level attribution to diffusion models with effective approximation techniques, and introduce REGMIX, a group-level approach that predicts data mixture performance using small-scale experiments. These contributions provide practitioners with scalable tools to audit training data impacts across modalities.
The second …
Analytics Insights From Text: Machine Learning, Ai, And Sentiment Analysis On Beige Books, Charlie Smith
Analytics Insights From Text: Machine Learning, Ai, And Sentiment Analysis On Beige Books, Charlie Smith
Graduate Theses and Dissertations (2019 - present)
Business analytics is about drawing actionable insights from data. These distinct but connected essays represent a novel approach to explore how natural language processing (NLP) advances and machine learning can transform unstructured text data into actionable conclusions. Essay 1 provides a broad framework. Essay 2 strengthens the sentiment analysis with the most recent artificial intelligence methodologies for capturing nuanced sentiment in complex texts. Essay 3 applies those insights to forecast recessions using topics that can be readily interpreted and applied.
The research demonstrates how these methodologies can be applied to enhance understanding of the same dataset, Beige Books. Published by …