Open Access. Powered by Scholars. Published by Universities.®

Digital Commons Network™

Open Access. Powered by Scholars. Published by Universities.®

Singapore Management University

Discipline
Keyword
Publication Year
Publication
Publication Type
File Type

Articles 121 - 150 of 10460

Full-Text Articles in Entire DC Network

Group Conversational Agents: A Review Of Designs That Support And Shape Group Interaction, Shunyi Yeo, Tianyi Zhang, Scott Bateman, Gary Hsieh, Young-Ho Kim, Simon Tangi Perrault, Jiannan Li, Anthony Tang Jun 2026

Group Conversational Agents: A Review Of Designs That Support And Shape Group Interaction, Shunyi Yeo, Tianyi Zhang, Scott Bateman, Gary Hsieh, Young-Ho Kim, Simon Tangi Perrault, Jiannan Li, Anthony Tang

Research Collection School Of Computing and Information Systems

Conversational agents that participate in or mediate group interaction introduce challenges that extend beyond supporting individual users, raising new questions about how agents participate in and influence groups. To characterise this emerging design space, we present a systematic review of 53 peer-reviewed studies on group conversational agents (GCAs). We analyse how GCAs intervene in group-level processes, including participation regulation, conflict mediation, task alignment, and execution support. Using concepts from group research as an analytic lens, we organise prior GCA work around recurring group interactional challenges (orientation, conflict, alignment, and execution), and examine the roles agents are designed to play in …


The Stars Align: Modeling User Rating Calibration With Sparse Semantic Review Features, Rodrigo Alves, Antoine Ledent Jun 2026

The Stars Align: Modeling User Rating Calibration With Sparse Semantic Review Features, Rodrigo Alves, Antoine Ledent

Research Collection School Of Computing and Information Systems

User ratings are often treated as comparable across users, although identical scores may reflect different experiences. We study whether ratings can be viewed as user-specific discretizations of a shared semantic continuum derived from review text. Our method maps reviews into sparse semantic features with a sparse autoencoder and learns user-specific filters for each rating level. On Amazon Electronics, the learned embeddings align along a shared low-dimensional rating axis. Users differ mainly in how they anchor and partition this continuum, while preserving its overall ordinal structure. These findings support a semantic view of calibration beyond scalar bias correction.


Hide-And-Sweep: Detecting Concealed Cameras Via Led Illumination Sweeps, Jonghyuk Yun, Jaeyoung Moon, Yunseo Park, Sean Rui Xiang Tan, Byunghyun Kim, Rajesh Krishna Balan, Jun Han Jun 2026

Hide-And-Sweep: Detecting Concealed Cameras Via Led Illumination Sweeps, Jonghyuk Yun, Jaeyoung Moon, Yunseo Park, Sean Rui Xiang Tan, Byunghyun Kim, Rajesh Krishna Balan, Jun Han

Research Collection School Of Computing and Information Systems

Hidden cameras have increasingly infiltrated hotel and Airbnb rooms, posing serious privacy risks. Detecting such cameras is challenging because they are visually inconspicuous and often embedded inside everyday objects. Even worse, existing handheld detectors are manual and also rely on single-angle illumination and hence suffer from high false-positive rates. We present SweepLED (pronounced "sweepled")1, a practical hidden camera detection system that operates on a commodity smartphone augmented with an unobtrusive LED-embedded case. SweepLED performs LED sweeping - a controlled sequence of multi-angle illumination - while the user simply holds the phone still by hand, enabling the camera to capture how …


Rc-Nf: Robot-Conditioned Normalizing Flow For Real-Time Anomaly Detection In Robotic Manipulation, Shijie Zhou, Bin Zhu, Jiarui Yang, Xiangyu Zhao, Jingjing Chen, Yu-Gang Jiang Jun 2026

Rc-Nf: Robot-Conditioned Normalizing Flow For Real-Time Anomaly Detection In Robotic Manipulation, Shijie Zhou, Bin Zhu, Jiarui Yang, Xiangyu Zhao, Jingjing Chen, Yu-Gang Jiang

Research Collection School Of Computing and Information Systems

Recent advances in Vision-Language-Action (VLA) models have enabled robots to execute increasingly complex tasks. However, VLA models trained through imitation learning struggle to operate reliably in dynamic environments and often fail under Out-of-Distribution (OOD) conditions. To address this issue, we propose Robot-Conditioned Normalizing Flow(RC-NF), a real-time monitoring model for robotic anomaly detection and intervention that ensures the robot's state and the object's motion trajectory align with the task. RC-NF decouples the processing of task-aware robot and object states within the normalizing flow. It requires only positive samples for unsupervised training and calculates accurate robotic anomaly scores during inference through the …


Long-Term Mine Planning: A Survey Of Classical, Hybrid And Artificial Intelligence-Based Methods, Nurul Asyikeen Azhar, Aldy Gunawan, Shih-Fen Cheng, Erwin Leonardi Jun 2026

Long-Term Mine Planning: A Survey Of Classical, Hybrid And Artificial Intelligence-Based Methods, Nurul Asyikeen Azhar, Aldy Gunawan, Shih-Fen Cheng, Erwin Leonardi

Research Collection School Of Computing and Information Systems

The aim of long-term mine planning (LTMP) is two-fold: to maximize the net present value of profits (NPV) and determine how ores are sequentially processed over the lifetime. This scheduling task is computationally complex as it is rife with variables, constraints, periods, uncertainties, and unique operations. In this paper, we present trends in the literature in the recent decade. One trend is the shift from deterministic toward stochastic problems as they reflect real-world complexities. A complexity of growing concern is also in sustainable mine planning. Another trend is the shift from traditional operational research solutions — relying on exact or …


Towards Auto-Evaluation For Large Language Models, Jiahao Ying Jun 2026

Towards Auto-Evaluation For Large Language Models, Jiahao Ying

Dissertations and Theses Collection (Open Access)

The rapid advancement of large language models (LLMs) has created an urgent need for evaluation methodologies that are timely, scalable, reliable, and informative. Conventional evaluation benchmarks, although essential for measuring model capabilities and guiding model development, are often constructed and maintained through labor-intensive human annotation. As LLMs continue to improve through increases in model scale, training data, and computational resources, static benchmarks may quickly lose discriminative power. Moreover, the growing use of large and diverse training corpora increases the risk of benchmark leakage, which can inflate evaluation results and obscure the true capabilities of models. These challenges call for a …


Public Acceptance Of Meat-Reduction Policies Across Cultures: Comparing Singapore And Switzerland, Bianca Wassmann, Shu Tian Ng, Mark Chong, Angela K. Y. Leung, Michael Siegrist Jun 2026

Public Acceptance Of Meat-Reduction Policies Across Cultures: Comparing Singapore And Switzerland, Bianca Wassmann, Shu Tian Ng, Mark Chong, Angela K. Y. Leung, Michael Siegrist

Research Collection Lee Kong Chian School Of Business

Purpose – As meat-reduction policies are discussed across the globe, many are met with public resistance. Cultural values may help explain this pushback, yet their role in shaping support for food policy remains poorly understood. This study is the first to apply cultural cognition theory to food policy by examining how cultural worldviews shape the acceptance of meat-reduction interventions in Singapore and Switzerland—two economically developed countries with contrasting cultural profiles. Design/methodology/approach – In an online survey, participants (Singapore: n = 357; Switzerland: n = 495) rated their acceptance of 11 meat-reduction interventions (e.g. taxes, subsidies, labelling). We then analysed to …


History To Future: Evolving Agent With Experience And Thought For Zero-Shot Vision-And-Language Navigation, Guangzhao Dai, Shuo Wang, Zihan Wang, Guo-Sen Xie, Yang Yang, Jinshan Pan, Qianru Sun, Xiangbo Shu Jun 2026

History To Future: Evolving Agent With Experience And Thought For Zero-Shot Vision-And-Language Navigation, Guangzhao Dai, Shuo Wang, Zihan Wang, Guo-Sen Xie, Yang Yang, Jinshan Pan, Qianru Sun, Xiangbo Shu

Research Collection School Of Computing and Information Systems

Vision-and-Language Navigation in Continuous Environment (VLN-CE) requires an agent to follow language instructions to navigate the target destination. With the advancement of large language models (LLMs), recent efforts have explored adapting them for zero-shot VLN-CE, offering a promising solution in addressing the drawbacks of poor generalization in the training-based paradigm. However, existing LLM-based works primarily perform naive reasoning for decision-making and lack feedback, e.g., reviewing historical errors and predicting future potentials. Consequently, it may suffer from continuous failure for those initial error tasks. In this paper, we rethink LLM-based zero-shot VLN-CE and propose a new paradigm, named EvoNav, to improve …


Task Complexity Matters: An Empirical Study Of Reasoning In Llms For Sentiment Analysis, Donghao Huang, Zhaoxia Wang Jun 2026

Task Complexity Matters: An Empirical Study Of Reasoning In Llms For Sentiment Analysis, Donghao Huang, Zhaoxia Wang

Research Collection School Of Computing and Information Systems

Large language models (LLMs) with reasoning capabilities have fueled a compelling narrative that reasoning universally improves performance across language tasks. We test this claim through a comprehensive evaluation of 504 configurations across seven model families—including adaptive, conditional, and reinforcement learning-based reasoning architectures—on sentiment analysis datasets of varying granularity (binary, five-class, and 27-class emotion). Our findings reveal that reasoning effectiveness is strongly task-dependent, challenging prevailing assumptions: (1) Reasoning shows task-complexity dependence—binary classification degrades up to -19.9 F1% points (pp), while 27-class emotion recognition gains up to  +16.0 pp; (2) Distilled reasoning variants underperform base models by 3–18 pp on simpler tasks, …


Cfalr: Collaborative Filtering-Augmented Large Language Model For Personalized Fashion Outfit Recommendation, Yujuan Ding, Junrong Liao, Yunshan Ma, Yi Bin, Wenqi Fan, Tat-Seng Chua, Qing Li Jun 2026

Cfalr: Collaborative Filtering-Augmented Large Language Model For Personalized Fashion Outfit Recommendation, Yujuan Ding, Junrong Liao, Yunshan Ma, Yi Bin, Wenqi Fan, Tat-Seng Chua, Qing Li

Research Collection School Of Computing and Information Systems

Personalized outfit recommendation poses a significant challenge in e-commerce and social media platforms, requiring systems that balance user preferences with aesthetic compatibility. Collaborative filtering (CF) provides a traditional solution for this, but it struggles with data-sparse scenarios and complex user-item-outfit relationships. Meanwhile, existing template-based approaches are constrained by rigid pre-designed structures. To bridge these research gaps, we introduce CFALR (Collaborative Filtering-Augmented Large Language Model for Recommendation), a novel framework that synergizes collaborative filtering with large language models for personalized outfit recommendation. Specifically, CFALR describes user-outfit interactions in natural language and leverages LLMs to capture fashion semantics while employing CF-enhanced embeddings …


To Wait Or To Transfer? A Three-Level Optimization Framework For Intermodal Transfer Coordination In First Train Timetabling And Bus Bridging Services Management, Hao Li, Liujiang Kang, Norman Weik, Huijun Sun, Qingying Lai, Zhiguang Cao Jun 2026

To Wait Or To Transfer? A Three-Level Optimization Framework For Intermodal Transfer Coordination In First Train Timetabling And Bus Bridging Services Management, Hao Li, Liujiang Kang, Norman Weik, Huijun Sun, Qingying Lai, Zhiguang Cao

Research Collection School Of Computing and Information Systems

This study addresses the integrated optimization of the first train timetabling and bus bridging service design (FTT-BBSD) for morning transfer challenges, two critical but interdependent passenger services in the public transit system. In contrast to most existing studies and conventional approaches, this study explicitly models the influence of passenger path choices and transfer mode selections on FTT-BBSD. Through a novel dual-level network representation that integrates subway and bus systems, we formulate the FTT-BBSD problem as a mixed-integer nonlinear programming model. The model simultaneously determines subway and bus timetables and bridging line deployment to minimize total travel time for all first …


A Pruning-Based Question-Answering For Interactive Video Search: A Simple Baseline, Yu Tong Cheng, Phuong Anh Nguyen, Chong-Wah Ngo Jun 2026

A Pruning-Based Question-Answering For Interactive Video Search: A Simple Baseline, Yu Tong Cheng, Phuong Anh Nguyen, Chong-Wah Ngo

Research Collection School Of Computing and Information Systems

There are various factors affecting the performance of video search. An imprecise query will enlarge search space and reduce the discriminative power of ranking functions. This problem is further exacerbated by the presence of numerous visually or semantically similar videos in large datasets. Consequently, users need to painstakingly browse through many highly similar candidates to locate the search target, leading to increased cognitive load and inefficient searching. Ideally, engaging users through interactive questioning to resolve uncertainties in the search process is an effective strategy for progressively narrowing down the search space. However, despite rapid advances in deep learning, generating informative …


Interfold: Learning Interpretable Diffusion Manifolds Beyond Binary Samples, Alexander Vincent Lewi, Rainer Tan, Shengfeng He Jun 2026

Interfold: Learning Interpretable Diffusion Manifolds Beyond Binary Samples, Alexander Vincent Lewi, Rainer Tan, Shengfeng He

Research Collection School Of Computing and Information Systems

We propose InterFold, a framework for learning and applying interpretable semantic manifolds in latent diffusion models, without requiring binary or paired supervision. Existing methods for semantic editing either rely on limited paired data or uncover only coarse, unsupervised directions that fail to capture user-specific, fine-grained attributes. InterFold addresses these limitations by learning a target attribute manifold in the H-space of diffusion models using only a set of positive, unlabeled examples. To edit a new image, InterFold projects its H-space representation toward this learned manifold through test-time optimization, enabling precise, identity-preserving modifications of complex, non-binary concepts. To make these edits effective …


Vehicle-Based Multi-Services For Future Smart Cities, Hao Sun, Jinhua Zhao, Hai Yang, Shenhao Wang, Hamsa Balakrishnan, Thomas W. Malone, Hai Wang Jun 2026

Vehicle-Based Multi-Services For Future Smart Cities, Hao Sun, Jinhua Zhao, Hai Yang, Shenhao Wang, Hamsa Balakrishnan, Thomas W. Malone, Hai Wang

Research Collection School Of Computing and Information Systems

Vehicles are crucial for sustaining socioeconomic activity and improving quality of life in modern cities by offering diverse services. These include passenger mobility, goods delivery, information acquisition, and acting as mobile servers such as food trucks and mobile lockers. At the same time, they also contribute to traffic congestion and air pollution. This tension fosters the rise of urban resource-conserving and sustainable service solutions. In this article, we introduce the concept of “Vehicle-Based Multi-Services” (VeMuS), in which a single vehicle offers multiple services simultaneously. Drawing on practical use cases, we examine service classification and integration for vehicles and the potential …


Adaptive Outlier Detection Over Data Stream, Rui Zhu, Mingyuan Jiang, Xiaochun Yang, Baihua Zheng, Bin Wang, Tao Qiu Jun 2026

Adaptive Outlier Detection Over Data Stream, Rui Zhu, Mingyuan Jiang, Xiaochun Yang, Baihua Zheng, Bin Wang, Tao Qiu

Research Collection School Of Computing and Information Systems

Continuous distance-based outlier detection in streaming data poses significant challenges and has a wide range of practical applications. Traditional threshold-based methods perform well under stable streaming conditions, where fixed parameters remain effective. However, they often struggle with dynamic data distributions and high stream speeds, leading to suboptimal performance, limited control over the number of returned outliers, and failure to meet real-time detection requirements. To address these issues, this paper introduces a novel Recall and Proportion-Aware Outlier Detection (RPA-OD) query. In RPA-OD, ρ defines a distance relaxation that enables real-time outlier detection. Specifically, objects with fewer than k neighbors within the …


A Novel Hierarchical Multi-Agent System For Payments Using Llms, Donghao Huang, Joon Kiat Chua, Zhaoxia Wang Jun 2026

A Novel Hierarchical Multi-Agent System For Payments Using Llms, Donghao Huang, Joon Kiat Chua, Zhaoxia Wang

Research Collection School Of Computing and Information Systems

Large language model (LLM) agents, such as OpenAI’s Operator and Claude’s Computer Use, can automate workflows but unable to handle payment tasks. Existing agentic solutions have gained significant attention; however, even the latest approaches face challenges in implementing end-to-end agentic payment workflows. To address this gap, this research proposes the Hierarchical Multi-Agent System for Payments (HMASP), which provides an end-to-end agentic method for completing payment workflows. The proposed HMASP leverages either open-weight or proprietary LLMs and employs a modular architecture consisting of the Conversational Payment Agent (CPA - first agent level), Supervisor agents (second agent level), Routing agents (third agent …


“Alexa, Do Not Say That In Front Of My Boss!” A Cross-Cultural Comparison Of User And Ai Preferences For Privacy-Aware Smart Speaker Interactions Across Contexts, Lynne Warin, Emily Aurelia, Anthony Tang, Emily Aurelia, Delphine Reinhardt Jun 2026

“Alexa, Do Not Say That In Front Of My Boss!” A Cross-Cultural Comparison Of User And Ai Preferences For Privacy-Aware Smart Speaker Interactions Across Contexts, Lynne Warin, Emily Aurelia, Anthony Tang, Emily Aurelia, Delphine Reinhardt

Research Collection School Of Computing and Information Systems

Due to their limited ability to reason about the social context in which they are used, smart speakers pose significant privacy risks by responding in ways that may violate people's implicit social boundaries. We conducted a cross-cultural vignette study (N = 944) in Germany and Singapore to investigate how situational factors—specifically social context (bystander relationships and closeness), physical context (location), and interaction context (topic and deceptive intent)—regulate user preferences for smart speaker responses. Our results demonstrate that these factors are superior predictors of response preferences than dispositional user traits (i.e., intrinsic personal traits). We identify two distinct social dynamics: a …


Potential Recyclable Materials In Buildings: A Framework For Greenhouse Gas Emissions Assessment Of Residential Buildings In Singapore, Pradeep Alva, Riccardo Talami, Wanyu Pei, Goran Sibenik, Martin Mosteiro-Romero, Clayton Miller, Rudi Stouffs Jun 2026

Potential Recyclable Materials In Buildings: A Framework For Greenhouse Gas Emissions Assessment Of Residential Buildings In Singapore, Pradeep Alva, Riccardo Talami, Wanyu Pei, Goran Sibenik, Martin Mosteiro-Romero, Clayton Miller, Rudi Stouffs

Research Collection College of Integrative Studies

As countries aim to reduce resource consumption and greenhouse gas (GHG) emissions, Whole Life Carbon Assessment (WLCA) has become a vital method for quantifying embodied and operational GHG emissions. However, few studies have conducted WLCA on an urban scale, often addressing operational or embodied GHG emissions in isolation without considering their cumulative impact. This study introduces a city-wide WLCA framework to assess the potential recyclable materials of urban building stock, using Singapore as a case study with 5915 public residential buildings. Upfront GHG emissions are calculated from material intensity and building information, while operational emissions are based on energy use …


Scaling Up Multi-Agent Reinforcement Learning For Large Agent Teams And Long-Horizon Tasks: A Survey, Minghong Geng, Shubham Pateria, Budhitama Subagdja, Ah-Hwee Tan Jun 2026

Scaling Up Multi-Agent Reinforcement Learning For Large Agent Teams And Long-Horizon Tasks: A Survey, Minghong Geng, Shubham Pateria, Budhitama Subagdja, Ah-Hwee Tan

Research Collection School Of Computing and Information Systems

Multi-agent reinforcement learning (MARL) empowers multiple autonomous agents to acquire effective policies for collaborative problem-solving. Over the last decade, MARL has seen significant advancements, with numerous algorithms achieving impressive performance across various benchmarks and real-world applications. Nevertheless, the scalability of multi-agent systems, in terms of the number of agents and the length of the task horizon, remains a critical consideration for applying MARL methods to complex problem-solving. Given that a dedicated review of the existing approaches and challenges in scaling up multi-agent systems remains largely absent, this survey aims to bridge this gap by delivering a comprehensive review of MARL …


Rode: Linear Rectified Mixture Of Diverse Experts For Food Large Multi-Modal Models, Pengkun Jiao, Xinlan Wu, Bin Zhu, Jingjing Chen, Chong-Wah Ngo, Yu-Gang Jun 2026

Rode: Linear Rectified Mixture Of Diverse Experts For Food Large Multi-Modal Models, Pengkun Jiao, Xinlan Wu, Bin Zhu, Jingjing Chen, Chong-Wah Ngo, Yu-Gang

Research Collection School Of Computing and Information Systems

Large Multi-modal Models (LMMs) have significantly advanced a variety of vision-language tasks. The scalability and availability of high-quality training data play a pivotal role in the success of LMMs. In the realm of food, while comprehensive food datasets such as Recipe1M offer an abundance of ingredient and recipe information, they often fall short of providing ample data for nutritional analysis. The Recipe1M+ dataset, despite offering a subset for nutritional evaluation, is limited in the scale and accuracy of nutrition information. To bridge this gap, we introduce Uni-Food, a unified food dataset that comprises over 100,000 images with various food labels, …


Benchmarking Gaslighting Negation Attacks Against Multimodal Large Language Models, Bin Zhu, Yinxuan Gui, Huiyan Qi, Jingjing Chen, Chong-Wah Ngo, Ee-Peng Lim Jun 2026

Benchmarking Gaslighting Negation Attacks Against Multimodal Large Language Models, Bin Zhu, Yinxuan Gui, Huiyan Qi, Jingjing Chen, Chong-Wah Ngo, Ee-Peng Lim

Research Collection School Of Computing and Information Systems

Multimodal Large Language Models (MLLMs) have exhibited remarkable advancements in integrating different modalities, excelling in complex understanding and generation tasks. Despite their success, MLLMs remain vulnerable to conversational adversarial inputs. In this paper, we systematically study gaslighting negation attacks—a phenomenon where models, despite initially providing correct answers, are persuaded by user-provided negations to reverse their outputs, often fabricating justifications. We conduct extensive evaluations of state-of-the-art MLLMs across diverse benchmarks and observe substantial performance drops when negation is introduced. Notably, we introduce the first benchmark GaslightingBench, specifically designed to evaluate the vulnerability of MLLMs to negation arguments. GaslightingBench consists of multiple-choice …


Not Too Early, Not All At Once: Design Tensions In Ai-Mediated Self-Disclosure In Online Dating, Pei-Hua Tsai, Tianyi Zhang, Emran Bin Elias Poh, Anthony Tang, Yung-Ju Chang Jun 2026

Not Too Early, Not All At Once: Design Tensions In Ai-Mediated Self-Disclosure In Online Dating, Pei-Hua Tsai, Tianyi Zhang, Emran Bin Elias Poh, Anthony Tang, Yung-Ju Chang

Research Collection School Of Computing and Information Systems

Online dating relies on self-disclosure, yet initial conversations are fragile: users must navigate uncertainty around timing, boundaries, and reciprocity with little shared context. While advances in AI raise the possibility of mediating disclosure, how such support might reshape the experience of early-stage relational disclosure remains underexplored. We conducted 29 semi-structured interviews to examine how daters envision AI-mediated self-disclosure in online dating. Our findings surface recurring design tensions rather than simple opportunities or risks. Participants welcomed guidance that could pace disclosure, support reflection, and reduce social awkwardness, but stressed preserving agency and authorship. They valued interpretive assistance for sense-making of ambiguous …


Language Embeddings Meet Shallow Autoencoders, Rodrigo Alves, Vojtěch Vančura, Pavel Kordík, Antoine Ledent Jun 2026

Language Embeddings Meet Shallow Autoencoders, Rodrigo Alves, Vojtěch Vančura, Pavel Kordík, Antoine Ledent

Research Collection School Of Computing and Information Systems

Shallow autoencoders are appealing recommenders due to their simplicity, scalability, and competitive retrieval quality, but they struggle in strict cold-start settings where new items have no interactions. We propose an inductive shallow autoencoder that leverages item side information (language embeddings) by fixing the decoder to item features and learning only an encoder in the same semantic space. To prevent trivial self-reconstruction without enforcing a hard zero diagonal, we introduce diagonal gating: a leave-one-item-out objective that blocks the self-copy shortcut only for the item being updated while retaining context from the rest of the user history. An alternating-style optimization trains the …


Sok: Understanding Zkvm: From Research To Practice, Guomin Yang, Yunbo Yang, Yuejia Cheng, Haibo Tang, Bingsheng Zhang, Kui Ren Jun 2026

Sok: Understanding Zkvm: From Research To Practice, Guomin Yang, Yunbo Yang, Yuejia Cheng, Haibo Tang, Bingsheng Zhang, Kui Ren

Research Collection School Of Computing and Information Systems

Zero-knowledge virtual machine (zkVM) is a powerful infrastructure for proving the correctness of a program execution with a succinct proof, attracting significant interest from researchers, developers, and users. It has been widely used in applications such as blockchain rollups, privacy-preserving machine learning, and off-chain computation. As the field grows, a wide range of zkVMs have been proposed. However, they adopt different choices in instruction formats, trace layouts, and proving backends, which results in a highly heterogeneous design landscape and makes it difficult to understand the relations among these systems.To bridge this gap, we provide a comprehensive study of zkVMs that …


Anatomical Domain Shifts: Test-Time Heterogeneous Adaptation For 3d Human Pose Prediction, Qiongjie Cui, Pan Zhou, Jingjing Chen, Na Zhao Jun 2026

Anatomical Domain Shifts: Test-Time Heterogeneous Adaptation For 3d Human Pose Prediction, Qiongjie Cui, Pan Zhou, Jingjing Chen, Na Zhao

Research Collection School Of Computing and Information Systems

The research frontier in human pose prediction (HPP) is advancing toward continual test-time adaptation (TTA), where models must self-adapt to dynamic test distributions. To date, the homeostatic continual TTA remains the sole viable solution, which isolates the model parameters and update domain-sensitive ones. Despite mitigating full-body domain gaps, human anatomical heterogeneity (domain shifts often localize to specific regions) is ignored. This anatomical-agnostic approach forces uniform parameter adaptation across kinematically distinct segments, causing: over-adaptation of stable regions and under-adaptation of shift-prone articulations. To address it, we introduce TT-HA, a novel Test-Time Heterogeneous Adaptation that implicitly estimates domain changes for anatomical segments, …


Enhancing Pointing Gestures Of Non-Hmd Users In Asymmetric Collocated Mixed Reality Collaboration, Nam-Dang Vo, Van-Vinh Thai, Anthony Tang, Khanh-Duy Le Jun 2026

Enhancing Pointing Gestures Of Non-Hmd Users In Asymmetric Collocated Mixed Reality Collaboration, Nam-Dang Vo, Van-Vinh Thai, Anthony Tang, Khanh-Duy Le

Research Collection School Of Computing and Information Systems

A common collocated group setting in mixed-reality (MR) collaboration is a person wearing a MR headset (HMD user) and presenting MR contents to audiences who are not provided with such specialized devices (Non-HMD users). In this setting, while Non-HMD users can view the MR environment shown on a large physical display, it still remains challenging for the HMD user to interpret their pointing gesture when they spatially refer to objects in the MR environment. To address this, we designed and evaluated two pointing techniques—SCREEN and SCREEN+SPACE—that support Non-HMD users in referring to MR content. Screen pointing allows users to refer …


Happycal: Designing Text And Image-Based Supports For Savouring Positive Work Experiences, Molly Stewart, Minghao Cai, Anthony Tang, Sam Liu, Chris Mosunic, Sowmya Somanath Jun 2026

Happycal: Designing Text And Image-Based Supports For Savouring Positive Work Experiences, Molly Stewart, Minghao Cai, Anthony Tang, Sam Liu, Chris Mosunic, Sowmya Somanath

Research Collection School Of Computing and Information Systems

Savouring positive work experiences can promote positive affect and well-being at work, yet there is limited guidance on how digital applications can support workers to engage in savouring. We developed HappyCal, a work-focused savouring application offering two forms of savouring support: text-based, a common modality in workplace reflection tools, and images, a largely unexplored approach in work-related savouring. We conducted an exploratory qualitative study where participants (N=36) used HappyCal over five days and engaged in savouring through either a text-only modality (n=17) or text input paired with image output (n=19). We found that (1) participants in both groups reported heightened …


How Do Machine Learning Models Change?, Joel Castaño, Rafael Cabañas, Antonio Salmerón, David Lo, Silverio Martínez-Fernández Jun 2026

How Do Machine Learning Models Change?, Joel Castaño, Rafael Cabañas, Antonio Salmerón, David Lo, Silverio Martínez-Fernández

Research Collection School Of Computing and Information Systems

The proliferation of Machine Learning (ML) models and their open source implementations has transformed AI research and applications. Platforms like Hugging Face (HF) enable this evolving ecosystem, yet a large-scale longitudinal study of how these models change is lacking. This study addresses this gap by analyzing over 680,000 commits from 100,000 models and 2,251 releases from 202 of these models on HF using repository mining and longitudinal methods. We apply an extended ML change taxonomy to classify commits and use Bayesian networks to model temporal patterns in commit and release activities. Our findings show that commit activities align with established …


Ai In Healthcare: Regulatory Guidelines And Judge-Made Negligence Principles For Ai Implementers, Gary K. Y. Chan Jun 2026

Ai In Healthcare: Regulatory Guidelines And Judge-Made Negligence Principles For Ai Implementers, Gary K. Y. Chan

Research Collection Yong Pung How School Of Law

The use of artificial intelligence (AI) in healthcare may, notwithstanding its potential benefits, result in harm to patients from allegedly negligent acts or omissions by hospitals and medical doctors. In such circumstances, how should the principles in the tort of negligence (duty of care, breach, causation, remoteness of damage, and defences) respond to AI innovations in healthcare? In particular, how may the standard of care expected of hospitals and medical doctors be informed by regulatory guidelines? We refer to case law precedents and regulatory guidelines on the roles and responsibilities of doctors and hospitals as AI implementers. Importantly, they prompt …


Sequential Robustness In Adversarial Reinforcement Learning, Roman Lok-Ming Belaire May 2026

Sequential Robustness In Adversarial Reinforcement Learning, Roman Lok-Ming Belaire

Dissertations and Theses Collection (Open Access)

My goal is to build autonomous systems that expand the reach of human capability in challenging domains such as undersea and space exploration, disaster response, and large-scale infrastructure. In everyday settings, these systems will increasingly appear in safety-critical applications such as autonomous driving, robotics, and industrial manufacturing. A central requirement for these systems is the ability to operate reliably under uncertainty, particularly when the environment behaves in unanticipated ways.

The robust handling of unforeseen environment dynamics is therefore a technical cornerstone of autonomous decision-making; Adversarial attacks provide a useful and principled lens through which to study this problem. Adversarial \textit{robustness}, …