Open Access. Powered by Scholars. Published by Universities.®

Databases and Information Systems Commons™

Open Access. Powered by Scholars. Published by Universities.®

7,250 Full-Text Articles 10,408 Authors 4,901,411 Downloads 214 Institutions

All Articles in Databases and Information Systems

Faceted Search

7,250 full-text articles. Page 12 of 268.

Lessons Learned From Sandboxing, Piloting And Policy Experimentation With Ai And Other Digital Initiatives: Part 1, Summary Report, Steven M. Miller 2025 Singapore Management University

Lessons Learned From Sandboxing, Piloting And Policy Experimentation With Ai And Other Digital Initiatives: Part 1, Summary Report, Steven M. Miller

Research Collection School Of Computing and Information Systems

This report, "Lessons Learned from Sandboxing, Piloting and Policy Experimentation with AI and Other Digital Initiatives," captures insights and experiences from project experts involved in recent digital innovation initiatives with the governments of Bangladesh, Maldives, and Kazakhstan, and from project experts actively involved with the use of AI for delivering government digital services in the EU, New Zealand, Rwanda, Singapore, United States, and Uzbekistan. The ten in-depth interview write-ups produced from these nine different country settings provide a small but highly informative sample of rich descriptions of some of the important realities, approaches, nuances, issues and challenges related to testing …


Hd-Epic: A Highly-Detailed Egocentric Video Dataset, Toby PERRETT, Ahmad DARKHALIL, Saptarshi SINHA, Omar EMARA, Sam POLLARD, Kranti Kumar PARIDA, Kaiting LIU, Prajwal GATTI, Siddhant BANSAL, Kevin FLANAGAN, Jacob CHALK, Zhifan ZHU, Rhodri GUERRIER, Fahd ABDELAZIM, Bin ZHU, Davide MOLTISANTI, Michael WRAY, Hazel DOUGHTY, Dima DAMEN 2025 Singapore Management University

Hd-Epic: A Highly-Detailed Egocentric Video Dataset, Toby Perrett, Ahmad Darkhalil, Saptarshi Sinha, Omar Emara, Sam Pollard, Kranti Kumar Parida, Kaiting Liu, Prajwal Gatti, Siddhant Bansal, Kevin Flanagan, Jacob Chalk, Zhifan Zhu, Rhodri Guerrier, Fahd Abdelazim, Bin Zhu, Davide Moltisanti, Michael Wray, Hazel Doughty, Dima Damen

Research Collection School Of Computing and Information Systems

We present a validation dataset of newly-collected kitchenbased egocentric videos, manually annotated with highly detailed and interconnected ground-truth labels covering: recipe steps, fine-grained actions, ingredients with nutritional values, moving objects, and audio annotations. Importantly, all annotations are grounded in 3D through digital twinning of the scene, fixtures, object locations, and primed with gaze. Footage is collected from unscripted recordings in diverse home environments, making HDEPIC the first dataset collected in-the-wild but with detailed annotations matching those in controlled lab environments. We show the potential of our highly-detailed annotations through a challenging VQA benchmark of 26K questions assessing the capability to …


Predicting Consumers’ Purchase Intention Of Browsed Products: A Study Based On Eye‑Tracking, Feiyan JIA, Choon Ling SIA, Yani SHI, Fiona Fui-hoon NAH, Keng SIAU 2025 Singapore Management University

Predicting Consumers’ Purchase Intention Of Browsed Products: A Study Based On Eye‑Tracking, Feiyan Jia, Choon Ling Sia, Yani Shi, Fiona Fui-Hoon Nah, Keng Siau

Research Collection School Of Computing and Information Systems

Predicting consumers’ purchase intention of browsed products enables sellers to implement nuanced promotion strategies to stimulate purchase. But how can we predict consumers’ purchase intention of browsed products? Our research demonstrates that consumers’ eye movement data collected when they browse products can serve this aim. We train and test the prediction model using logistic regression and random forest algorithms. Using data collected in a laboratory experiment, our empirical results show that both algorithms perform much better than a random guess, and the logistic regression performs slightly better than the random forest. Our findings imply that eye movement data enable sellers …


Lessons Learned From Sandboxing, Piloting And Policy Experimentation With Ai And Other Digital Initiatives: Part 2, Ten In-Depth Interviews, Steven M. Miller 2025 Singapore Management University

Lessons Learned From Sandboxing, Piloting And Policy Experimentation With Ai And Other Digital Initiatives: Part 2, Ten In-Depth Interviews, Steven M. Miller

Research Collection School Of Computing and Information Systems

This report, "Lessons Learned from Sandboxing, Piloting and Policy Experimentation with AI and Other Digital Initiatives," captures insights and experiences from project experts involved in recent digital innovation initiatives with the governments of Bangladesh, Maldives, and Kazakhstan, and from project experts actively involved with the use of AI for delivering government digital services in the EU, New Zealand, Rwanda, Singapore, United States, and Uzbekistan. The ten in-depth interview write-ups produced from these nine different country settings provide a small but highly informative sample of rich descriptions of some of the important realities, approaches, nuances, issues and challenges related to testing …


Less Is More: On The Importance Of Data Quality For Unit Test Generation, Junwei ZHANG, Xing HU, Shan GAO, Xin XIA, David LO, Shanping LI 2025 Singapore Management University

Less Is More: On The Importance Of Data Quality For Unit Test Generation, Junwei Zhang, Xing Hu, Shan Gao, Xin Xia, David Lo, Shanping Li

Research Collection School Of Computing and Information Systems

Unit testing is crucial for software development and maintenance. Effective unit testing ensures and improves software quality, but writing unit tests is time-consuming and labor-intensive. Recent studies have proposed deep learning (DL) techniques or large language models (LLMs) to automate unit test generation. These models are usually trained or fine-tuned on large-scale datasets. Despite growing awareness of the importance of data quality, there has been limited research on the quality of datasets used for test generation. To bridge this gap, we systematically examine the impact of noise on the performance of learning-based test generation models. We first apply the open …


Large Language Models For Logical Fallacy Detection, NICOLE ANNE HUI-YING TEO, Donghao HUANG, Erik CAMBRIA, Zhaoxia WANG 2025 Singapore Management University

Large Language Models For Logical Fallacy Detection, Nicole Anne Hui-Ying Teo, Donghao Huang, Erik Cambria, Zhaoxia Wang

Research Collection School Of Computing and Information Systems

Identifying logical fallacies is essential for maintaining log-ical reasoning and reducing false information in a variety of domains, such as the media, law, and education. We present an extensive study on the use of large language models (LLMs) for logical fallacy detection and provide a comparative overview of model performance across various fallacy classes. We evaluate the logical fallacy detection capabilities of multiple state-of-the-art models (LLaMA, Qwen, Gemma, Phi) utilizing accuracy, precision, recall, and F1-score as assessment measures. Accord-ing to our findings, our models do well on simple fallacies like “circular reasoning,” but they have trouble with more interpretive reasoning …


Human-Computer Interaction And Artificial Intelligence For Ageing Population, Keng SIAU, Hailiang WANG, Fiona Fui-hoon NAH, Runyu WANG, Ruitong CHE, Can LIU 2025 Singapore Management University

Human-Computer Interaction And Artificial Intelligence For Ageing Population, Keng Siau, Hailiang Wang, Fiona Fui-Hoon Nah, Runyu Wang, Ruitong Che, Can Liu

Research Collection School Of Computing and Information Systems

As the global population ages rapidly, the field of human-computer interaction (HCI) is in urgent need of innovation, redesign, and reengineering to meet the evolving needs of older adults. The older demographic faces a range of challenges—including physical limitations, cognitive decline, reduced social in-tegration, and varying levels of technological literacy—that can hinder effective engagement with digital technologies. In response to these challenges, research-ers and designers are using inclusive and adaptive approaches to enhance acces-sibility, usability, and emotional well-being. This paper reviews key design prin-ciples in HCI for the ageing population and discusses how artificial intelligence (AI) tools, such as voice …


Unlocking The Power Of Socio-Knowledge Association For Enterprise Risk Identification In Stock Market, Zhenghao LIU, Keng SIAU, Shaochen YANG, Feicheng MA 2025 Singapore Management University

Unlocking The Power Of Socio-Knowledge Association For Enterprise Risk Identification In Stock Market, Zhenghao Liu, Keng Siau, Shaochen Yang, Feicheng Ma

Research Collection School Of Computing and Information Systems

Potential risk signals reflected in supply chain and equity connections between enterprises and social connections between investors are becoming crucial to identifying enterprise risks in addition to basic financial indicators. Traditional risk management systems face challenges in adapting to these complexities, highlighting the need for a proactive paradigm shift in risk management. Leveraging graph models such as social networks and knowledge graphs offers a promising approach to identifying and managing potential associated risks effectively. To bridge existing research gaps, a novel risk identification framework driven by social-knowledge graphs has been proposed, integrating graph deep learning and reinforcement learning techniques guided …


Fact-Checking As A Multi-Step Process: From Ambiguity Resolution To Claim Validation, Wenbo Wang 2025 New Jersey Institute of Technology

Fact-Checking As A Multi-Step Process: From Ambiguity Resolution To Claim Validation, Wenbo Wang

Dissertations

The spread of misinformation and disinformation has become a major concern, particularly with the rise of social media as a primary source of information for many people. Fact-checking—the process of verifying claims against credible evidence—has emerged as a critical safeguard against misinformation. Yet, the task is fraught with challenges: claims are often ambiguous, context-dependent, or composed of multiple intertwined assertions, while automated systems struggle to replicate the nuanced reasoning of human experts. This dissertation addresses these challenges by reimagining fact-checking as a multi-step, knowledge-guided process that systematically resolves ambiguity, decomposes complexity, and validates claims through structured reasoning. Additionally, the proposed …


From Neural Networks To Large Language Models: Innovations In Financial Ai, Mathematical Reasoning, And Structured Data Representation, Junyi Ye 2025 New Jersey Institute of Technology

From Neural Networks To Large Language Models: Innovations In Financial Ai, Mathematical Reasoning, And Structured Data Representation, Junyi Ye

Dissertations

This dissertation explores the evolution and application of artificial intelligence techniques across three critical domains: financial modeling, mathematical reasoning, and structured data analysis. The dissertation presents seven research projects that chart a progression from specialized neural architectures to sophisticated large language models (LLMs), contributing novel methodologies and frameworks at each stage.

In the financial domain, the research first introduces TS-Mixer, a MLP-based architecture for time-series forecasting that captures both feature relationships and temporal dependencies through a simple yet effective design, outperforming more complex models in S&P500 index prediction. The dissertation then presents DySTAGE, a dynamic graph representation learning framework that …


Towards Explainable Ai On Graph Neural Networks: Xaig, Jiaxing Zhang 2025 New Jersey Institute of Technology

Towards Explainable Ai On Graph Neural Networks: Xaig, Jiaxing Zhang

Dissertations

In the evolving landscape of artificial intelligence (AI), Graph Neural Networks (GNNs) have garnered growing prominence for their adeptness in processing graph-structured data. Despite this, the interpretability of their predictions often remains elusive. The demand for transparency and explainability in complex prediction models has reached unprecedented levels. To address this, post-hoc instance-level explanation techniques have emerged, aiming to unveil the rationale behind GNN predictions. These techniques endeavor to unearth substructures that elucidate the predictive behavior of trained GNNs.

This dissertation embarks on an exploration of Explainable AI (XAI) technologies within the realm of GNNs. Amid the challenges posed by the …


Blockchain-Enabled Master Data Management, Shakhawat Hossain 2025 University of Arkansas Little Rock

Blockchain-Enabled Master Data Management, Shakhawat Hossain

Theses and Dissertations

Master Data Management (MDM) is essential for maintaining data quality, accuracy, consistency, and governance within organizations. However, traditional centralized MDM systems continue to face challenges related to data integrity, security, and scalability. This research presents a blockchain-enabled MDM framework designed to overcome these limitations by leveraging blockchain’s decentralized, immutable, and secure architecture. The study aims to identify and address the shortcomings of conventional MDM practices, examine the applicability of blockchain technology in enhancing these systems, and develop a functional prototype to validate the proposed model. The framework incorporates decentralized review mechanisms that improve auditability and ensure trusted data verification by …


Selected Artificial Intelligence Provisions In U.S. Fiscal Year 2025 National Defense Authorization Act, Bert Chapman 2025 Purdue University

Selected Artificial Intelligence Provisions In U.S. Fiscal Year 2025 National Defense Authorization Act, Bert Chapman

Libraries Faculty and Staff Presentations

The 2025 Fiscal Year National Defense Authorization Act contains multiple provisions relating to artificial intelligence (AI). These congressionally mandated provisions direct various sections of the Department of Defense (DOD) and individual U.S. armed service branches to execute congressional intent for AI policymaking. Examples of such intent include identifying and planning DOD's AI workforce, demonstrating AI biotechnology applications for national security, improving the human usability of AI systems, and establishing an AI security center. This presentation will note that reports on these initiatives must be prepared for relevant congressional oversight committees, and, in many cases, are in many cases, publicly released …


A Computational Method For Detecting Compound Promiscuity In Early-Stage Pharmaceutical Discovery, John Allen Ringer 2025 University of New Mexico - Main Campus

A Computational Method For Detecting Compound Promiscuity In Early-Stage Pharmaceutical Discovery, John Allen Ringer

Computer Science ETDs

Modern drug discovery and chemical biology research relies heavily on analyzing bioassay data. One of the many challenges in bioassay data analysis is identifying false trails, i.e., chemical compounds which initially appear to have desirable activity but are found to be problematic upon further investigation. Badapple (the BioAssay-Data Associative Promiscuity Pattern Learning Engine) was created over ten years ago to help researchers identify promiscuous compounds and thus avoid a common source of these false trails. Through an effort involving software engineering, cheminformatics, and biomedical data science we have developed Badapple 2.0, which incorporates updated assay records and expanded data semantics. …


From Data To Action: An Adaptable Crosstabs Template For Participatory Survey Data Analysis, Natalia Pinzon, Vikram Koundinya, William O'R Dowling, Ryan Galt 2025 University of California, Davis

From Data To Action: An Adaptable Crosstabs Template For Participatory Survey Data Analysis, Natalia Pinzon, Vikram Koundinya, William O'R Dowling, Ryan Galt

Journal of Extension

We present a practical and accessible template for quantitative survey data analysis designed for non-academic researchers in order to facilitate engagement from community collaborators. The template, created in Google Sheets, is mainly for computing cross-tabulations, but it also displays frequency distributions and p-values for determining statistical significance. The template allows collaborators to record their observations and questions, promoting an efficient yet interactive review process and fostering a democratic analysis environment. Based on our own experience using this template for a data party, we highlight its effectiveness in promoting collaborative data interpretation, decision-making, and the actionable use of survey findings.


Full-Stack Web Applications: Industry Standard Frameworks, Libraries & Technologies, Yassine Chahid, Patrick Slattery 2025 CUNY New York City College of Technology

Full-Stack Web Applications: Industry Standard Frameworks, Libraries & Technologies, Yassine Chahid, Patrick Slattery

Publications and Research

This research explores emerging full-stack web development technologies across front-end, back-end, and DevSecOps domains. It evaluates modern tools including Django, React, and TypeScript—focusing on their key features such as compile-time error checking—through to the development of a web application. By examining documentation for the frameworks Node.js, Next.js, Tailwind CSS, and others, along with the deployment tools Docker and Git for version/release control, the study analyzes how these innovations speed up development, improve existing practices, and have often replaced older technologies. Cloud solutions for tasks such as authentication and deployment will also be evaluated, along with various web-application technology stacks and …


Computer Vision In Soccer: Yolov11 Analytics Engine For Quantifying Game Strategy, Connor S. Maurer 2025 University of Arkansas, Fayetteville

Computer Vision In Soccer: Yolov11 Analytics Engine For Quantifying Game Strategy, Connor S. Maurer

Data Science Undergraduate Honors Theses

Single-shot object detection capabilities significantly reduce computational overhead for real-time computer vision in sports analytics at 60 FPS. YOLO11’s lightweight CNN gives promising accuracy while meeting the low-latency demand of dynamic soccer matches. As data-driven approaches take over the sport of soccer, efficient player tracking systems become critical for informing coach’s strategies. I prototype the ETL (Extract, Transform, Load) process of data collected from a single- shot detection program and evaluate its viability for estimating player fatigue. YOLO11 detects players, the ball, and other characteristics, with the output transformed by homography to estimate the positions in the real world. These …


Learning Behaviors In Physics-Informed Deep Learning, Alex Glover 2025 East Tennessee State University

Learning Behaviors In Physics-Informed Deep Learning, Alex Glover

Electronic Theses and Dissertations

Physics-informed deep learning is a methodology in artificial intelligence aimed at combating the large training data requirement and the barrier of domain awareness that deep learning architectures commonly face in applications. Stochastic modeling integrated into the predictive models provides that domain knowledge. Variations of the Intelligent Driving Model impact the learning behaviors of the joint-training architecture. This thesis examines the effect of substituting the standard linear Intelligent Driving Model with a modified nonlinear version, as applied to real human driving behavior on the I-80 interstate. The experimentation also critically evaluates the complications that impede the viability of this architecture in …


Towards Reliable Ml: Data Attribution And Adversarial Robustness, Xiaosen ZHENG 2025 Singapore Management University

Towards Reliable Ml: Data Attribution And Adversarial Robustness, Xiaosen Zheng

Dissertations and Theses Collection (Open Access)

Modern machine learning (ML) models achieve remarkable success, but face critical reliability challenges. This thesis advances two pillars of reliable ML systems: interpretability through data attribution and robustness against adversarial threats.

In the first part, we develop novel data attribution methods to elucidate the data-model relationship. We establish the critical role of memorization in model generalization through token-level influence analysis, extend sample-level attribution to diffusion models with effective approximation techniques, and introduce REGMIX, a group-level approach that predicts data mixture performance using small-scale experiments. These contributions provide practitioners with scalable tools to audit training data impacts across modalities.

The second …


Analytics Insights From Text: Machine Learning, Ai, And Sentiment Analysis On Beige Books, Charlie Smith 2025 University of South Alabama

Analytics Insights From Text: Machine Learning, Ai, And Sentiment Analysis On Beige Books, Charlie Smith

Graduate Theses and Dissertations (2019 - present)

Business analytics is about drawing actionable insights from data. These distinct but connected essays represent a novel approach to explore how natural language processing (NLP) advances and machine learning can transform unstructured text data into actionable conclusions. Essay 1 provides a broad framework. Essay 2 strengthens the sentiment analysis with the most recent artificial intelligence methodologies for capturing nuanced sentiment in complex texts. Essay 3 applies those insights to forecast recessions using topics that can be readily interpreted and applied.

The research demonstrates how these methodologies can be applied to enhance understanding of the same dataset, Beige Books. Published by …


Digital Commons powered by bepress