Open Access. Powered by Scholars. Published by Universities.®

Artificial Intelligence and Robotics Commons™

Open Access. Powered by Scholars. Published by Universities.®

11,189 Full-Text Articles 24,569 Authors 5,758,021 Downloads 274 Institutions

All Articles in Artificial Intelligence and Robotics

Faceted Search

11,189 full-text articles. Page 54 of 542.

Coresets For Clustering Under Stochastic Noise, Lingxiao HUANG, Zhize LI, Nisheeth K. VISHNOI, Runkai YANG, Haoyu ZHAO 2025 Singapore Management University

Coresets For Clustering Under Stochastic Noise, Lingxiao Huang, Zhize Li, Nisheeth K. Vishnoi, Runkai Yang, Haoyu Zhao

Research Collection School Of Computing and Information Systems

We study the problem of constructing coresets for $(k, z)$-clustering when the input dataset is corrupted by stochastic noise drawn from a known distribution. In this setting, evaluating the quality of a coreset is inherently challenging, as the true underlying dataset is unobserved. To address this, we investigate coreset construction using surrogate error metrics that are tractable and provably related to the true clustering cost. We analyze a traditional metric from prior work and introduce a new error metric that more closely aligns with the true cost. Although our metric is defined independently of the noise distribution, it enables approximation …


Generalization Bounds For Rank‑Sparse Neural Networks, Antoine LEDENT, Rodrigo ALVES, Yunwen LEI 2025 Singapore Management University

Generalization Bounds For Rank‑Sparse Neural Networks, Antoine Ledent, Rodrigo Alves, Yunwen Lei

Research Collection School Of Computing and Information Systems

It has been recently observed in much of the literature that neural networks exhibit a bottleneck rank property: for larger depths, the activation and weights of neural networks trained with gradient-based methods tend to be of approximately low rank. In fact, the rank of the activations of each layer converges to a fixed value referred to as the “bottleneck rank”, which is the minimum rank required to represent the training data. This perspective is in line with the observation that regularizing linear networks (without activations) with weight decay is equivalent to minimizing the Schatten p quasi norm of the neural …


Sempo: Lightweight Foundation Models For Time Series Forecasting, Hui HE, Kun YI, Yuanchi MA, Qi ZHANG, Zhengdong NIU, Guansong PANG 2025 Singapore Management University

Sempo: Lightweight Foundation Models For Time Series Forecasting, Hui He, Kun Yi, Yuanchi Ma, Qi Zhang, Zhengdong Niu, Guansong Pang

Research Collection School Of Computing and Information Systems

The recent boom of large pre-trained models witnesses remarkable success in developing foundation models (FMs) for time series forecasting. Despite impressive performance across diverse downstream forecasting tasks, existing time series FMs possess massive network architectures and require substantial pre-training on large-scale datasets, which significantly hinders their deployment in resource-constrained environments. In response to this growing tension between versatility and affordability, we propose SEMPO, a novel lightweight foundation model that requires pretraining on relatively small-scale data, yet exhibits strong general time series forecasting. Concretely, SEMPO comprises two key modules: 1) energy-aware SpEctral decomposition module, that substantially improves the utilization of pre-training …


Bias Testing And Mitigation In Llm-Based Code Generation, Dong HUANG, Jie M. ZHANG, Qingwen BU, Xiaofei XIE, Junjie CHEN, Heming CUI 2025 Singapore Management University

Bias Testing And Mitigation In Llm-Based Code Generation, Dong Huang, Jie M. Zhang, Qingwen Bu, Xiaofei Xie, Junjie Chen, Heming Cui

Research Collection School Of Computing and Information Systems

As the adoption of LLMs becomes more widespread in software coding ecosystems, a pressing issue has emerged: does the generated code contain social bias and unfairness, such as those related to age, gender, and race? This issue concerns the integrity, fairness, and ethical foundation of software applications that depend on the code generated by these models but are underexplored in the literature. This paper presents a novel bias testing framework that is specifically designed for code generation tasks. Based on this framework, we conduct an extensive empirical study on the biases in code generated by five widely studied LLMs (i.e., …


Kpiroot+: An Efficient Integrated Framework For Anomaly Detection And Root Cause Analysis In Large-Scale Cloud Systems, Wenwei GU, Renyi ZHONG, Guangba YU, Xinying SUN, Jinyang LIU, Yintong HUO, Zhuangbin CHEN, Jianping ZHANG, Jiazhen GU, Yongqiang YANG, Michael R. LYU 2025 Singapore Management University

Kpiroot+: An Efficient Integrated Framework For Anomaly Detection And Root Cause Analysis In Large-Scale Cloud Systems, Wenwei Gu, Renyi Zhong, Guangba Yu, Xinying Sun, Jinyang Liu, Yintong Huo, Zhuangbin Chen, Jianping Zhang, Jiazhen Gu, Yongqiang Yang, Michael R. Lyu

Research Collection School Of Computing and Information Systems

To ensure the reliability of cloud systems, their runtime status reflecting the service quality is periodically monitored with monitoring metrics, i.e., KPIs (key performance indicators). When performance issues happen, root cause localization pinpoints the specific KPIs that are responsible for the degradation of overall service quality, facilitating prompt problem diagnosis and resolution. To this end, existing methods generally locate root-cause KPIs by identifying the KPIs that exhibit a similar anomalous trend to the overall service performance. While straightforward, solely relying on the similarity calculation may be ineffective when dealing with cloud systems with complicated interdependent services. Recent deep learning-based methods …


Genscore: Agent-Based Short-Answer Question Generation And Scoring In Software Engineering Courses, Nguyen Binh Duong TA, Lwin Khin SHAR 2025 Singapore Management University

Genscore: Agent-Based Short-Answer Question Generation And Scoring In Software Engineering Courses, Nguyen Binh Duong Ta, Lwin Khin Shar

Research Collection School Of Computing and Information Systems

Short-answer questions are commonly used in educational assessments, as they are often viewed as a more effective way than multiple-choice questions to determine whether students have achieved the intended learning outcomes. However, manually creating appropriate questions targeting different cognitive levels such as those defined by the Bloom’s Taxonomy, and grading text answers from students are not trivial tasks for instructors. Existing work on auto-question generation and scoring in computing education typically targets coding-based questions. However, in software engineering courses, assessments can extend beyond coding to understanding of processes, DevOps methodologies, system design, etc. This work aims to address the dual …


Reliable-Data-Split (Rds): Maximizing Model Potential With Reinforced Selection Strategy, Hoang D. Nguyen, Xuan-Son Vu, Quoc Tuan TRUONG, Duc-Trong Le 2025 Singapore Management University

Reliable-Data-Split (Rds): Maximizing Model Potential With Reinforced Selection Strategy, Hoang D. Nguyen, Xuan-Son Vu, Quoc Tuan Truong, Duc-Trong Le

Research Collection School Of Computing and Information Systems

The nexus between data characteristics and parametric models is fundamental for developing effective and reliable artificial intelligence (AI) systems. Mismatches in data properties for model development may lead to deleterious effects on AI model performance in machine learning practice. This paper proposes a Reliable Data Split (RDS) procedure to learn how to select data points that will generalise the target domain adequately by employing prior knowledge of the data generative process. We introduce a reinforced selection strategy using deep reinforcement learning with diverse black box predictors in maximising ensemble rewards as the proxy of model performance potential while maintaining an …


Examining The Roles Of Embodiment And Theory Of Mind In Shaping User Perceptions Of Llm-Driven Conversational Agents, Elizabeth A. Schlesener 2025 Clemson University

Examining The Roles Of Embodiment And Theory Of Mind In Shaping User Perceptions Of Llm-Driven Conversational Agents, Elizabeth A. Schlesener

All Dissertations

Large Language Models (LLMs) have advanced conversational agents, enabling natural, human-like interactions in domains such as education, programming, and workplace collaboration. Yet, user distrust persists over privacy, accuracy, and bias. As developers work to mitigate these issues and human-AI collaboration expands, reinforcing trust in LLM-driven systems is essential. To address this problem, this dissertation explores the role of anthropomorphic form in LLM-driven conversational agents and its impact on user perception.

According to the familiarity thesis, humans attribute human-like characteristics to nonhuman entities — a process known as anthropomorphism — to better comprehend unfamiliar phenomena, based on the assumption that they …


Ai In Consideration Of Her: Accounting For Gendered Workplace Dynamics In The Design And Evaluation Of Human-Centered Ai Integration In Everyday Workplaces, Kelsea S. Schulenberg 2025 Clemson University

Ai In Consideration Of Her: Accounting For Gendered Workplace Dynamics In The Design And Evaluation Of Human-Centered Ai Integration In Everyday Workplaces, Kelsea S. Schulenberg

All Dissertations

Rapid advancements in the technical capabilities and availability of generative Artificial Intelligence (AI) systems, such as OpenAI's ChatGPT, have drawn widespread attention to the opportunities and challenges associated with AI integration into everyday workplaces (i.e., office-type work). Following calls for organizations to consider the ethical and workplace-specific impacts of generative AI's use before integrating it into the workplace, this dissertation addresses three critical gaps in Human-Centered Computing (HCC) and AI workplace integration research. First, this dissertation unpacks the underdeveloped links between women's representation - or lack thereof - in AI-related fields and how their experiences with gendered workplace dynamics in …


Early Conceptual Sketches Of Blended Reality And The Precursor To The Bbs Quad (2022), David Smith 2025 CUNY New York City College of Technology

Early Conceptual Sketches Of Blended Reality And The Precursor To The Bbs Quad (2022), David Smith

Publications and Research

This document contains two original hand-drawn conceptual sketches created in early 2022, representing the earliest visual formulations of what would later evolve into the Balanced Blended Space (BBS) framework. The drawings predate my first conversations with ChatGPT and were produced as part of my independent sabbatical research into blended environments, mediated performance, and human–machine interaction.

The first drawing examines human–computational mediation, perception, and internal mapping. The second sketch—later referred to informally as the “BBS Quad”—extends this idea by reconciling cognition–computation symmetry with physical–virtual spatial relationships. Published together, these images document the conceptual foundations of the BBS framework prior to its …


Accelerating Relationship Discovery In Chronic Lower Back Pain Through Knowledge Graph And Ontology Enhanced Large Language Models, Damon Lin 2025 California Polytechnic State University, San Luis Obispo

Accelerating Relationship Discovery In Chronic Lower Back Pain Through Knowledge Graph And Ontology Enhanced Large Language Models, Damon Lin

Master's Theses

Chronic lower back pain (cLBP) is a widespread public health burden linked to anxiety, depression, and opioid addiction. Interventions aimed at treating cLBP have shown minimal improvements in pain outcomes, leading researchers to reexamine our understanding of cLBP through constructing a causal model. However, constructing causal models through Randomized Controlled Trials are often unfeasible, and relying on domain expertise requires extensive and time-consuming research, posing a serious bottleneck for designing effective treatments. To accelerate this process, we apply Knowledge Graphs, Ontologies, and Large Language Models (LLMs) to aid researchers in determining possible causal relationships. First, we demonstrate how LLMs can …


Generalized Detection Of Animal Behavior Using Accelerometers, Alexander J. Arrieta 2025 California Polytechnic State University, San Luis Obispo

Generalized Detection Of Animal Behavior Using Accelerometers, Alexander J. Arrieta

Master's Theses

Animal mounted sensors are becoming increasingly used to passively monitor both domestic and wild animals. Advances in lightweight accelerometer and GPS technology have allowed many animals to be fitted with high accuracy sensors for extended periods of time. This leads to new opportunities to study animal behavior without direct observation. However, interpreting the raw data is difficult due to the high volume and missing context of the information. Machine learning techniques excel at extracting information from raw data streams and are excellent candidates for processing the sensor data. However, due to large variance in how different animals execute the same …


Synthetic Dataset For Understanding Negation In Text-Guided Image Editing, Nhat-Tan Bui 2025 University of Arkansas-Fayetteville

Synthetic Dataset For Understanding Negation In Text-Guided Image Editing, Nhat-Tan Bui

Graduate Theses and Dissertations

Negation is a fundamental linguistic concept used by humans to convey information that they do not desire. Despite this, minimal research has focused on negation within text-guided image editing. This lack of research means that vision-language models (VLMs) for image editing may struggle to understand negation, implying that they struggle to provide accurate results. One barrier to achieving human-level intelligence is the lack of a standard collection by which research into negation can be evaluated. This thesis presents the first large-scale dataset, Negative Instruction (NeIn), for studying negation within instruction-based image editing. Our dataset comprises 366,957 quintuplets, i.e., source image, …


Large Language Models (Llms) For Clinical Note Generation: International Classification Of Disease (Icd) Code, Knowledge Graph (Kg) And Prompt Evaluation, Ivan P. Makohon 2025 Old Dominion University

Large Language Models (Llms) For Clinical Note Generation: International Classification Of Disease (Icd) Code, Knowledge Graph (Kg) And Prompt Evaluation, Ivan P. Makohon

Computer Science Theses & Dissertations

In the past decade, a surge in the amount of electronic health record (EHR) data in the United States occurred, driven by a favorable policy environment created by the Health Information Technology for Economic and Clinical Health (HITECH) Act of 2009 and the 21st Century Cures Act of 2016. Clinical notes for patients’ assessments, diagnoses, and treatments are captured in these EHRs in free-form text by physicians, who spend a considerable amount of time entering them. Manually writing these notes is time-consuming, increasing patient waiting times and potentially delaying diagnoses. Large language models (LLMs), such as GPT-4o, possess the ability …


Copyright Ownership And Duration Of Ai-Authored Works, Cheng Lim SAW 2025 Singapore Management University

Copyright Ownership And Duration Of Ai-Authored Works, Cheng Lim Saw

Research Collection Yong Pung How School Of Law

On the assumption that Parliament has endorsed the notion of AI authorship and the prospect that copyright may well subsist in works created autonomously by the AI itself, this essay further explores allied issues surrounding the ownership and duration of copyright in AI-authored works.


Navigating Ai-Nature Frictions: Autonomous Vehicle Testing And Nature-Based Constraints, Prerona DAS, Orlando WOODS, Lily KONG 2025 Singapore Management University

Navigating Ai-Nature Frictions: Autonomous Vehicle Testing And Nature-Based Constraints, Prerona Das, Orlando Woods, Lily Kong

Research Collection College of Integrative Studies

In cities, the application of Artificial Intelligence (AI) is being directed towards transforming different aspects of urban life. These applications take material form in urban spaces, with autonomous vehicles (AVs) providing a prominent example. AI systems rely on large volumes of data on their surroundings to refine the algorithms and enhance the accuracy of prediction for operational efficiency and safety. However, such algorithmic learning and execution can present challenges when dealing with the unpredictable, complex, and dynamic aspects of urban spaces. Nature is a paradigmatic example of such unpredictability, because natural phenomena usually defy consistent patterns and precise data-based modelling. …


Mando-Llm: Heterogeneous Graph Transformers With Large Language Models For Smart Contract Vulnerability Detection, Nhat Minh NGUYEN, Huu Hoang NGUYEN, Long LE THANH, Zahra AHMADI, Thanh Nam DOAN, Daoyuan WU, Lingxiao JIANG 2025 Singapore Management University

Mando-Llm: Heterogeneous Graph Transformers With Large Language Models For Smart Contract Vulnerability Detection, Nhat Minh Nguyen, Huu Hoang Nguyen, Long Le Thanh, Zahra Ahmadi, Thanh Nam Doan, Daoyuan Wu, Lingxiao Jiang

Research Collection School Of Computing and Information Systems

Detecting vulnerabilities in smart contracts is vital for the security and reliability of decentralized apps. To facilitate vulnerability detection, contract codes, including bug patterns, are represented as heterogeneous graphs with various nodes and edges, like control-flow and function-call graphs. However, existing graph learning techniques struggle with large, complex graphs. This paper presents MANDO-LLM, a novel framework that combines heterogeneous graph transformers (HGTs) with large language models (LLMs) for detecting vulnerabilities in smart contracts represented as heterogeneous contract graphs built upon control-flow and call graphs. MANDO-LLM uses LLMs to capture code features from control-flow and call data, customizes HGTs to learn …


Agentguard: An Active Threat Discovery System For Package Confusion Using Multi-Agent Collaboration, Wei MA, Yu LI, Zhi CHEN, Ye LIU, Lingxiao JIANG, Qiang HU, Junyi TAO 2025 Singapore Management University

Agentguard: An Active Threat Discovery System For Package Confusion Using Multi-Agent Collaboration, Wei Ma, Yu Li, Zhi Chen, Ye Liu, Lingxiao Jiang, Qiang Hu, Junyi Tao

Research Collection School Of Computing and Information Systems

The proliferation of open-source software (OSS) has made software supply chains prime targets for attacks like Package Confusion, where adversaries publish malicious packages with names deceptively similar to legitimate ones. Existing detection methods often rely on simple lexical similarity or passive analysis of known package pairs, struggle with high false positive rates (FPR), fail to proactively identify emerging threats, and are vulnerable to adversarial evasion. To overcome these limitations, we introduce AgentGuard, a novel framework for proactive, single-input package confusion detection. AgentGuard employs a multi-agent architecture that autonomously discovers potential confusion targets using fine-tuned word embedding model to hybird semantic …


Registration Is A Powerful Rotation-Invariance Learner For 3d Anomaly Detection, Yuyang YU, Zhengwei CHEN, Xuemiao XU, Lei ZHANG, Haoxin YANG, Yongwei NIE, Shengfeng HE 2025 Singapore Management University

Registration Is A Powerful Rotation-Invariance Learner For 3d Anomaly Detection, Yuyang Yu, Zhengwei Chen, Xuemiao Xu, Lei Zhang, Haoxin Yang, Yongwei Nie, Shengfeng He

Research Collection School Of Computing and Information Systems

3D anomaly detection in point-cloud data is critical for industrial quality control, aiming to identify structural defects with high reliability. However, current memory bank-based methods often suffer from inconsistent feature transformations and limited discriminative capacity, particularly in capturing local geometric details and achieving rotation invariance. These limitations become more pronounced when registration fails, leading to unreliable detection results. We argue that point-cloud registration plays an essential role not only in aligning geometric structures but also in guiding feature extraction toward rotation-invariant and locally discriminative representations. To this end, we propose a registration-induced, rotation-invariant feature extraction framework that integrates the objectives …


Safe-Sora: Safe Text-To-Video Generation Via Graphical Watermarking, Zihan SU, Xuerui QIU, Hongbin XU, Tangyu JIANG, Jun-hao ZHUANG, Chun YUAN, Ming LI, Shengfeng HE, Fei YU 2025 Singapore Management University

Safe-Sora: Safe Text-To-Video Generation Via Graphical Watermarking, Zihan Su, Xuerui Qiu, Hongbin Xu, Tangyu Jiang, Jun-Hao Zhuang, Chun Yuan, Ming Li, Shengfeng He, Fei Yu

Research Collection School Of Computing and Information Systems

The explosive growth of generative video models has amplified the demand for reliable copyright preservation of AI-generated content. Despite its popularity in image synthesis, invisible generative watermarking remains largely underexplored in video generation. To address this gap, we propose Safe-Sora, the first framework to embed graphical watermarks directly into the video generation process. Motivated by the observation that watermarking performance is closely tied to the visual similarity between the watermark and cover content, we introduce a hierarchical coarse-to-fine adaptive matching mechanism. Specifically, the watermark image is divided into patches, each assigned to the most visually similar video frame, and further …


Digital Commons powered by bepress