Open Access. Powered by Scholars. Published by Universities.®

Digital Commons Network™

Open Access. Powered by Scholars. Published by Universities.®

Singapore Management University

Discipline
Keyword
Publication Year
Publication
Publication Type
File Type

Articles 151 - 180 of 10460

Full-Text Articles in Entire DC Network

Perspectives On Interpretability For Neural Text Representations, Jia Peng Lim May 2026

Perspectives On Interpretability For Neural Text Representations, Jia Peng Lim

Dissertations and Theses Collection (Open Access)

In this dissertation, we investigate interpretability in the three elements of learning neural text representations: inputs, passed into models, to produce probabilistic outputs. We emphasise perspectives as we present alternative novel methods to mine and organise meaning in this work.

Models. We initiate our investigation by examining Neural Topic Models (NTM), proposing an alternate angle of interpreting its word-topic distribution, producing better topic representations for interpretation. Our method maps the problem of finding these better interpretations to classical NP-hard graph problems, enabling examination of topic distributions in a composite manner. Next, we apply our previous findings to extract interpretations from …


Func: Reducing The Impact Of Android Framework Evolution On Malware Detection, Hailong Yu, Tiantian Wang, Lwin Khin Shar, Hanmeng Li, David Lo May 2026

Func: Reducing The Impact Of Android Framework Evolution On Malware Detection, Hailong Yu, Tiantian Wang, Lwin Khin Shar, Hanmeng Li, David Lo

Research Collection School Of Computing and Information Systems

Android malware detection approaches commonly use APIs and permissions as features for classifying malware. However, since the release of the first Android operating system in 2008, the Android framework has undergone numerous version updates. The evolution of the Android framework over time has led to changes in APIs and permissions, including deprecations and replacements. These changes can result in inaccurate characterization of Android malware, thereby affecting performance of malware detectors. There is a lack of methods to mitigate the impact of Android framework evolution on malware detection. To fill this gap, we conduct a systematic study of the impact of …


Benchmarking Gaslighting Attacks Against Speech Large Language Models, Jinyang Wu, Bin Zhu, Xiandong Zou, Qiquan Zhang May 2026

Benchmarking Gaslighting Attacks Against Speech Large Language Models, Jinyang Wu, Bin Zhu, Xiandong Zou, Qiquan Zhang

PhD Student’s Publications Collection

As Speech Large Language Models (Speech LLMs) become increasingly integrated into voice-based applications, ensuring their robustness against manipulative or adversarial input becomes critical. Although prior work has studied adversarial attacks in text-based LLMs and vision-language models, the unique cognitive and perceptual challenges of speech-based interaction remain underexplored. In contrast, speech presents inherent ambiguity, continuity, and perceptual diversity, which make adversarial attacks more difficult to detect. In this paper, we introduce gaslighting attacks, strategically crafted prompts designed to mislead, override, or distort model reasoning as a means to evaluate the vulnerability of Speech LLMs. Specifically, we construct five manipulation strategies: Anger, …


Teacher-Student Diffusion Model For Text-Driven 3d Hand Motion Generation, Ching Lam Cheng, Bin Zhu, Shengfeng He May 2026

Teacher-Student Diffusion Model For Text-Driven 3d Hand Motion Generation, Ching Lam Cheng, Bin Zhu, Shengfeng He

PhD Student’s Publications Collection

Generating realistic 3D hand motion from natural language is vital for VR, robotics, and human-computer interaction. Existing methods either focus on full-body motion, overlooking detailed hand gestures, or require explicit 3D object meshes, limiting generality. We propose TSHaMo, a model-agnostic teacher-student diffusion framework for text-driven hand motion generation. The student model learns to synthesize motions from text alone, while the teacher leverages auxiliary signals (e.g., MANO parameters) to provide structured guidance during training. A co-training strategy enables the student to benefit from the teacher’s intermediate predictions while remaining text-only at inference. Evaluated using two diffusion backbones on GRAB and H2O, …


Digital Grief Technology To Support Bereavement: A Systematic Review Of Potential Benefits And Risks, Xun Ci Soh, Adalia Yin Hui Goh, Paye Shin Koh, Andree Hartanto May 2026

Digital Grief Technology To Support Bereavement: A Systematic Review Of Potential Benefits And Risks, Xun Ci Soh, Adalia Yin Hui Goh, Paye Shin Koh, Andree Hartanto

Research Collection School of Social Sciences

Grief is a universal and inevitable experience. However, the way we support the bereaved is changing, especially in the digital era. This systematic review examines the potential benefits and risks associated with various digital grief technologies, including online grief support groups, generative AI chatbots, online memorials, online therapy interventions, virtual reality, and digitally reproduced visuals or audio of the deceased. A systematic search was conducted in seven databases, and 30 articles were included in the final review. Findings indicate that digital grief technologies offer several benefits, such as reductions in grief and depressive symptoms, enhanced social support, greater accessibility, and …


Understanding Critical Thinking In Generative Artificial Intelligence Use: Development, Validation, And Correlates Of The Critical Thinking In Ai Use Scale, Gabriel R. Lau, Wei Yan Low, Louis Tay, Ysabel Thereze Ang Guevarra, Dragon Gašević, Andree Hartanto May 2026

Understanding Critical Thinking In Generative Artificial Intelligence Use: Development, Validation, And Correlates Of The Critical Thinking In Ai Use Scale, Gabriel R. Lau, Wei Yan Low, Louis Tay, Ysabel Thereze Ang Guevarra, Dragon Gašević, Andree Hartanto

Research Collection School of Social Sciences

Generative AI tools are increasingly embedded in everyday work and learning, yet their fluency, opacity, and propensity to hallucinate mean that users must critically evaluate AI outputs rather than accept them at face value. The present research conceptualises critical thinking in AI use as a dispositional tendency to verify the source and content of AI-generated information, to understand how models work and where they fail, and to reflect on the broader implications of relying on AI. Across six studies ( N = 1341), we developed and validated the 13-item critical thinking in AI use scale and mapped its nomological network. …


Uncertainty, Information, And Major Choice: Evidence From Chinese Gaokao, Sikun Dou May 2026

Uncertainty, Information, And Major Choice: Evidence From Chinese Gaokao, Sikun Dou

Dissertations and Theses Collection (Open Access)

An unresolved background risk makes people more averse to other, independent  risks—most of all when a high-stakes, irreversible choice must be made under that risk. College major choice is such a setting. In China, students are admitted directly into a major they cannot easily change, and in the early 2000s they had to commit before learning their exam scores, bearing a field’s academic risk on top of an unresolved admission risk. Because STEM fields carry more academic risk and women are on average more risk-averse, this burden falls hardest on women’s entry into STEM. Exploiting a reform that moved application …


Sevoauth: Secure Voiceprint Authentication With Hash-Based Feature Transformation, Rui Zhang, Zheng Yan, Robert H. Deng May 2026

Sevoauth: Secure Voiceprint Authentication With Hash-Based Feature Transformation, Rui Zhang, Zheng Yan, Robert H. Deng

Research Collection School Of Computing and Information Systems

While voiceprint authentication offers convenient user authentication and access control through voice feature recognition, a critical research gap remains: existing voiceprint authentication systems fail to simultaneously achieve sound security against replay, spoofing, and adversarial attacks, preserve voice privacy leakage, and satisfy usability demand. Previous efforts have struggled to balance these issues comprehensively. To bridge this gap, we present SeVoAuth, a cloud-based Voiceprint Authentication as a Service (VAaaS) system designed to provide privacy preservation, robust security, and enhanced usability. SeVoAuth stores a synthesized voiceprint of a user in the cloud during user registration, thereby safeguarding the privacy of the real voiceprint …


Synthesis And Evaluation Of Long-Term History-Aware Medical Dialogue, Hebin Hu, Renke Dai, Ah-Hwee Tan, Yilin Kang May 2026

Synthesis And Evaluation Of Long-Term History-Aware Medical Dialogue, Hebin Hu, Renke Dai, Ah-Hwee Tan, Yilin Kang

Research Collection School Of Computing and Information Systems

An effective healthcare agent must be able to recall and reason over a patient’s longitudinal medical history. However, the absence of datasets with realistic long-term dialogue timelines limits systematic evaluation. Real clinical text is constrained by privacy and ethics, while existing benchmarks focus on isolated interactions, failing to capture cross-session reasoning. We introduce a framework for synthesizing high-quality, long-term medical dialogues with LLMs. Our approach entails a knowledge-guided decomposition into three stages: constructing synthetic patient profiles with diverse disease and complication trajectories, generating multiturn dialogues per encounter, and integrating them into a coherent longitudinal history dataset, MediLongChat. We establish three …


Open Source Software Development Tool Installation: Challenges And Strategies For Novice Developers, Larissa Salerno, Christoph Treude, Patanamon Thongtanunam May 2026

Open Source Software Development Tool Installation: Challenges And Strategies For Novice Developers, Larissa Salerno, Christoph Treude, Patanamon Thongtanunam

Research Collection School Of Computing and Information Systems

As the world of technology advances, so do the tools that software developers use to create new programs. In recent years, software development tools have become more popular, allowing developers to work more efficiently and produce higher-quality software. Still, installing such tools can be challenging for novice developers at the early stage of their careers, as they may face issues such as compatibility problems (e.g., with operating systems) and unclear instructions. Therefore, this work aims to investigate the challenges novice developers face when installing software development tools and the strategies they employ to overcome them. To investigate these, we conducted …


Generation Of Elaborated, Targeted And Effective Feedback For Novice Programmers Using Llm, Hua Leong Fwa May 2026

Generation Of Elaborated, Targeted And Effective Feedback For Novice Programmers Using Llm, Hua Leong Fwa

Research Collection School Of Computing and Information Systems

Programming errors and misconceptions are pervasive in novice programmers which causes difficulty in the learning of computer programming. Large Language Models (LLMs), with their ability to comprehend and generate programming codes have shown promising results in the automatic identification of errors. This can potentially benefit student programmers by providing them with timely formative feedback at efficiencies and scale that were not attainable previously. In this study, we leveraged an LLM - OpenAI o4-mini for the generation of elaborated, targeted feedback for novice programmers across PHP and JavaScript exercises. We contend that the feedback needs to be effective and targeted other …


Fighting Against Recruitment Scams: Theory-Driven Supervised Learning And Empirical Analysis For Digital Fraudulent Recruitment Posting Behavior, Tom (Tianteng) Wang, David (Jingjun) Xu, Keng Siau, Zhongju (John) Zhang May 2026

Fighting Against Recruitment Scams: Theory-Driven Supervised Learning And Empirical Analysis For Digital Fraudulent Recruitment Posting Behavior, Tom (Tianteng) Wang, David (Jingjun) Xu, Keng Siau, Zhongju (John) Zhang

Research Collection School Of Computing and Information Systems

The number of recruitment postings on digital recruitment hiring platforms has increased since the COVID-19 pandemic. However, the weak surveillance and operations of these platforms, combined with the fact that most job seekers have relatively low vigilance and a strong desire for recruitment offers, enable scammers to easily deceive job seekers for their money and confidential information. In this work, we combine prevailing text mining techniques (i.e., ChatGPT with prompting engineering and supervised machine learning) with interpersonal deception theory (IDT) from social science to design an interpretable IT system to predict fraudulent recruitment postings on digital recruitment-hiring platforms. We compare …


Securing Cloud-Native Systems: From Vulnerability Analysis To External And Insider Threat Detection, Jiongchi Yu May 2026

Securing Cloud-Native Systems: From Vulnerability Analysis To External And Insider Threat Detection, Jiongchi Yu

Dissertations and Theses Collection (Open Access)

Cloud-native systems have become the backbone of modern software infrastructure. However, their dynamic resource orchestration and complex configurability introduce a large attack surface and intricate security challenges. Adversaries can externally exploit vulnerabilities in cloud components or perform insider movement within cloud environments to launch attacks. As these systems increasingly support critical services, security breaches can lead to severe operational and economic consequences.

Despite extensive efforts in vulnerability detection and attack monitoring, existing approaches struggle to remain effective in cloud-native environments characterized by rapid evolution and inherent heterogeneity. In particular, they exhibit three fundamental limitations: (1) Insufficient understanding of defect patterns …


Leading Ai Adoption In Organizations: Introducing A Behavioral Human-Centered Approach, Shane Schweitzer, Devesh Narayan, Jack Mcguire, David De Cremer May 2026

Leading Ai Adoption In Organizations: Introducing A Behavioral Human-Centered Approach, Shane Schweitzer, Devesh Narayan, Jack Mcguire, David De Cremer

Research Collection Lee Kong Chian School Of Business

Initiatives to implement AI technologies in organizations fail at an alarming rate. We argue that leading the adoption of AI is not a simple engineering exercise but rather represents a behavioral exercise where change management principles—the process by which organizations plan, implement, and embed changes in practices—are employed. However, many AI initiatives in business focus predominantly on the AI systems themselves, assuming humans will fall in line. To solve this, we integrate ideas from change management with scholarship on human-centered artificial intelligence to offer a behavioral approach that accounts for the impact of AI adoption on humans at all stages …


Public Pool Usage As Adaptation Against Urban Heat, Stefan Borsky, Eric Fesselmeyer May 2026

Public Pool Usage As Adaptation Against Urban Heat, Stefan Borsky, Eric Fesselmeyer

Research Collection College of Integrative Studies

This paper examines the relationship between urban heat and outdoor public pool usage. Using attendance data from all 53 outdoor public pools in New York City, we analyze nonlinear effects of heat on pool usage across socioeconomic contexts. Pool attendance rises sharply with heat, especially in low-income neighborhoods where alternative coping options are likely limited. We also find that public pools reduce heat-related emergency medical service calls. Our findings highlight the need for equitable investment in blue infrastructure to enhance urban climate resilience and demonstrate how this type of adaptive infrastructure can play a critical role in managing urban heat.


Selective Concolic Testing, Guofeng Zhang, Zhenbang Chen, Ziqi Shuai, Jun Sun, Weijiang Hong, Yufeng Zhang, Ji Wang, Yang Liu May 2026

Selective Concolic Testing, Guofeng Zhang, Zhenbang Chen, Ziqi Shuai, Jun Sun, Weijiang Hong, Yufeng Zhang, Ji Wang, Yang Liu

Research Collection School Of Computing and Information Systems

The principled combination of symbolic execution and random testing lacks a formal foundation, especially in deciding which inputs to symbolize. We propose selective concolic testing, a cost-aware framework that formulates this choice as an optimized policy problem of a MDP (Markov Decision Process). We model program exploration over a finite control-flow graph, where MDP states represent covered statements, actions partition path constraints into symbolic and random fragments, rewards reflect coverage gain, and costs account for SMT solving effort and sampling inefficiency. Our framework yields the first formal characterization of selective symbolization as policy synthesis in a probabilistic system. We prove …


Detecting Doubt In Reflective Learning: A Learning Analytics Study With Large And Small Language Models, Eng Lieh Ouh, Kar Way Tan, Siaw Ling Lo, Yuhao Zhang May 2026

Detecting Doubt In Reflective Learning: A Learning Analytics Study With Large And Small Language Models, Eng Lieh Ouh, Kar Way Tan, Siaw Ling Lo, Yuhao Zhang

Research Collection School Of Computing and Information Systems

Reflective learning enhances understanding, especially when instructors promptly address difficulties raised in student reflections. Automated doubt detection can reduce time for instructors, yet existing classification approaches take substantial time for manual annotation and model training. This paper investigates whether large and small language models (LLMs, SLMs) can automate doubt detection without time-consuming training. Using a dataset of anonymized student reflections, we evaluate zeroshot, few-shot prompting, and multi-step reasoning against prior supervised classification baselines. We show that LLMs (GPT-4o, Claude-4, Gemini-2.5) surpass earlier F1 scores without prompting, while prompting further improves their performance. However, using proprietary LLMs can raise cost and …


Gencode: A Generic Data Augmentation Framework For Boosting Deep Learning-Based Code Understanding, Zeming Dong, Qiang Hu, Xiaofei Xie, Maxime Cordy, Mike Papadakis, Yves Le Traon, Jianjun Zhao May 2026

Gencode: A Generic Data Augmentation Framework For Boosting Deep Learning-Based Code Understanding, Zeming Dong, Qiang Hu, Xiaofei Xie, Maxime Cordy, Mike Papadakis, Yves Le Traon, Jianjun Zhao

Research Collection School Of Computing and Information Systems

Pre-trained code models lead the era of code intelligence, with multiple models designed with impressive performance. However, one important problem, data augmentation for code data that automatically helps developers prepare training data lacks study in this field. In this paper, we introduce a generic data augmentation framework, GenCode, to enhance the training of code understanding models. Simply speaking, GenCode follows a generation-and-selection paradigm to prepare useful training code data. Specifically, it employs code augmentation techniques to generate new code candidates first and then identifies important ones as the training data by influence scores. To evaluate the effectiveness of GenCode, we …


Market Reactions To Deceptive Language In Fake News: Implications From Language Expectancy Theory And Transfer Learning, Ka Chung Ng, Ping Fan Ke, Ping Fan, Mike So, Tam, Kar Yan May 2026

Market Reactions To Deceptive Language In Fake News: Implications From Language Expectancy Theory And Transfer Learning, Ka Chung Ng, Ping Fan Ke, Ping Fan, Mike So, Tam, Kar Yan

Research Collection School Of Computing and Information Systems

The advent of generative artificial intelligence (AI) has heightened the proliferation of fake news. A key challenge is the limited real-world data to investigate the societal impact of fake news produced by generative AI. In this paper, we examine stock market reactions to financial news articles that exhibit stylometric similarity to human-crafted and AI-crafted fake financial news. Grounded in language expectancy theory, we employ a style-based transfer learning model, pre-trained to recognizing deceptive language employed in various types of fake news intricacies. We then apply this model to a comprehensive dataset of financial news, assigning a “veracity style score” to …


Quantitative Bounds On Resource Usage Of Probabilistic Programs, Krishnendu Chatterjee, Amir Kafshdar Goharshady, Tobias Meggendorfer, Dorde Zikelic May 2026

Quantitative Bounds On Resource Usage Of Probabilistic Programs, Krishnendu Chatterjee, Amir Kafshdar Goharshady, Tobias Meggendorfer, Dorde Zikelic

Research Collection School Of Computing and Information Systems

Cost analysis, also known as resource usage analysis, is the task of finding bounds on the total cost of a program and is a well-studied problem in static analysis. In this work, we consider two classical quantitative problems in cost analysis for probabilistic programs. The first problem is to find a bound on the expected total cost of the program. This is a natural measure for the resource usage of the program and can also be directly applied to average-case runtime analysis. The second problem asks for a tail bound, i.e. ‍given a threshold t the goal is to find …


Brief Virtual Reality And Mixed Reality Mindfulness Breathing Exercise For Emotional Well-Being And Cognitive Functions In University Students: Within-Subjects Experimental Design Study, Zoey Khai Yee Eun, Charmaine Jiali Koh, Hwajin Yang, Adalia Yin Hui Goh, Meilan Hu, K Tennakoon Appuhamillage Sandeeshwara Kasturiratna, Andree Hartanto May 2026

Brief Virtual Reality And Mixed Reality Mindfulness Breathing Exercise For Emotional Well-Being And Cognitive Functions In University Students: Within-Subjects Experimental Design Study, Zoey Khai Yee Eun, Charmaine Jiali Koh, Hwajin Yang, Adalia Yin Hui Goh, Meilan Hu, K Tennakoon Appuhamillage Sandeeshwara Kasturiratna, Andree Hartanto

Research Collection School of Social Sciences

Background: Mindfulness has been shown to enhance emotional well-being and cognitive performance, yet much of this evidence stems from interventions requiring prolonged practice, making them time-consuming and less accessible. Recent studies suggest that brief mindfulness sessions may also yield positive outcomes, but the effectiveness of such interventions in virtual reality (VR) and mixed reality (MR) remains underexplored. Objective: This study investigates the effects of brief mindfulness breathing exercises delivered through VR and MR on attentional and emotional restoration and self-control capacity. Methods: Using a within-subjects experimental design, 102 undergraduate participants (n=83, 81.4% female; mean age 20.87, SD 1.89) completed a …


Mease: Multi-Agent Episodic Action Sequence Explanation, Phyo Wai Khaing, Minghong Geng, Shubham Pateria, Budhitama Subagdja, Ah-Hwee Tan May 2026

Mease: Multi-Agent Episodic Action Sequence Explanation, Phyo Wai Khaing, Minghong Geng, Shubham Pateria, Budhitama Subagdja, Ah-Hwee Tan

Research Collection School Of Computing and Information Systems

Multi-agent reinforcement learning (MARL) achieves remarkable performance in complex coordination tasks, yet interpreting the emergent behaviors of trained agents remains a fundamental challenge. Most current explainability methods focus on individual agent decisions, overlooking the critical interplay of joint strategiesand temporal coordination patterns that define successful multi-agent policies. We present MEASE (Multi-agent Episodic Action Sequence Explanation), a novel explainable MARL (XMARL) framework that explains trained MARL policies as human-interpretable emergent cooperative joint behaviors. MEASE employs a cognition-inspired episodic memory model to learn spatio-temporal multi-agent interaction patterns, coupled with abstraction algorithms that identify significant cooperative agent behaviors. We evaluate MEASE on diverse …


Collaborative Practices And Tool Utilization In Software Development Projects: A Student Perspective, Yi Meng Lau, Muhammad Syahmi Bin Abbas, Lingxiao Jiang May 2026

Collaborative Practices And Tool Utilization In Software Development Projects: A Student Perspective, Yi Meng Lau, Muhammad Syahmi Bin Abbas, Lingxiao Jiang

Research Collection School Of Computing and Information Systems

Software development is a collaborative activity that depends on effective teamwork, shared understanding, and coordinated use of development practices and tools. While these aspects are well studied in professional environments, they are less frequently examined within software engineering education. This study investigates how students collaborate in group projects, focusing on collaborative practices, tool usage, and their perceptions of software quality. We conducted a quantitative post-project survey with 143 second-year undergraduate students enrolled in a software development course. The results show that students actively share information and often establish team norms to support coordination and collaboration. However, students face challenges in …


Causality-Driven Test Case Minimisation For Cyber-Physical Systems, Michael Foster, Christopher M. Poskitt, Nicholas R. Latimer, Neil Walkinshaw, Richard Somers, Robert M. Hierons May 2026

Causality-Driven Test Case Minimisation For Cyber-Physical Systems, Michael Foster, Christopher M. Poskitt, Nicholas R. Latimer, Neil Walkinshaw, Richard Somers, Robert M. Hierons

Research Collection School Of Computing and Information Systems

Cyber-physical systems allow digital control systems to interact with the physical world using sensors and actuators. They are increasingly being used to automate critical infrastructure, where software faults can have dire consequences. Due to the complex nature and unpredictability of these systems, their resilience is often tested using a technique called fuzzing, which generates quasi-random sequences of sensor and actuator manipulations with the goal of forcing a system into unsafe states. However, there is currently no way of determining which manipulations of a test case cause a failure without systematically removing each one and re-running the test, which can be …


Natural Adversaries: Fuzzing Autonomous Vehicles With Realistic Roadside Object Placements, Yang Sun, Haoyu Wang, Christopher M. Poskitt, Jun Sun May 2026

Natural Adversaries: Fuzzing Autonomous Vehicles With Realistic Roadside Object Placements, Yang Sun, Haoyu Wang, Christopher M. Poskitt, Jun Sun

Research Collection School Of Computing and Information Systems

The emergence of Autonomous Vehicles (AVs) has spurred research into testing the resilience of their perception systems, i.e., ensuring that they are not susceptible to critical misjudgements. It is important that these systems are tested not only with respect to other vehicles on the road, but also with respect to objects placed on the roadside. Trash bins, billboards, and greenery are examples of such objects, typically positioned according to guidelines developed for the human visual system, which may not align perfectly with the needs of AVs. Existing tests, however, usually focus on adversarial objects with conspicuous shapes or patches, which …


Enhancing Action And Ingredient Modeling For Semantically Grounded Recipe Generation, Guoshan Liu, Bin Zhu, Yian Li, Jingjing Chen, Chong-Wah Ngo, Yu-Gang Jiang May 2026

Enhancing Action And Ingredient Modeling For Semantically Grounded Recipe Generation, Guoshan Liu, Bin Zhu, Yian Li, Jingjing Chen, Chong-Wah Ngo, Yu-Gang Jiang

Research Collection School Of Computing and Information Systems

Recent advances in Multimodal Large Language Models (MLMMs) have enabled recipe generation from food images, yet outputs often contain semantically incorrect actions or ingredients despite high lexical scores (e.g., BLEU, ROUGE). To address this gap, we propose a semantically grounded framework that predicts and validates actions and ingredients as internal context for instruction generation. Our two-stage pipeline combines supervised fine-tuning (SFT) with reinforcement fine-tuning (RFT): SFT builds foundational accuracy using an Action-Reasoning dataset and ingredient corpus, while RFT employs frequency-aware rewards to improve long-tail action prediction and ingredient generalization. A Semantic Confidence Scoring and Rectification (SCSR) module further filters and …


Be Responsible In Your Answers! Monitoring Out-Of-Domain Behaviors In Domain-Specific Llms, Boquan Li, Chenzhe Lou, Zhe Ren, Peixin Zhang, Zirui Fu, Jun Sun, Yaowen Zheng Apr 2026

Be Responsible In Your Answers! Monitoring Out-Of-Domain Behaviors In Domain-Specific Llms, Boquan Li, Chenzhe Lou, Zhe Ren, Peixin Zhang, Zirui Fu, Jun Sun, Yaowen Zheng

Research Collection School Of Computing and Information Systems

Large Language Models (LLMs) have accelerated the rapid development of chatbot web applications in various domains, such as coding, biomedicine and psychology. Compared to general LLMs like ChatGPT, domain-specific LLMs require a greater sense of responsibility. For instance, if a programming LLM casually answers medical or psychological questions, it not only misleads the public but also poses legal risks. This highlights new demands for monitoring and preventing such irresponsible behaviors. Existing efforts attempt to monitor LLMs from multiple aspects, such as lying, jailbreaks, and toxic content, while overlooking out-of-domain behaviors. In this work, we propose an innovative LLM domain monitoring …


Architecture-Agnostic Test-Time Adaptation Via Backprop-Free Embedding Alignment, Xiao Ma, Young D. Kwon, Pan Zhou, Dong Ma Apr 2026

Architecture-Agnostic Test-Time Adaptation Via Backprop-Free Embedding Alignment, Xiao Ma, Young D. Kwon, Pan Zhou, Dong Ma

PhD Student’s Publications Collection

Test-Time Adaptation (TTA) adapts a deployed model during online inference to mitigate the impact of domain shift. While achieving strong accuracy, most existing methods rely on backpropagation, which is memory and computation intensive, making them unsuitable for resource-constrained devices. Recent attempts to reduce this overhead often suffer from high latency or are tied to specific architectures such as ViT-only or CNN-only. In this work, we revisit domain shift from an embedding perspective. Our analysis reveals that domain shift induces three distinct structural changes in the embedding space: translation (mean shift), scaling (variance shift), and rotation (covariance shift). Based on this …


Scalable Multi-Task Low-Rank Model Adaptation, Zichen Tian, Antoine Ledent, Qianru Sun Apr 2026

Scalable Multi-Task Low-Rank Model Adaptation, Zichen Tian, Antoine Ledent, Qianru Sun

PhD Student’s Publications Collection

Scaling multi-task low-rank adaptation (LoRA) to a large number of tasks induces catastrophic performance degradation, such as an accuracy drop from 88.2% to 2.0% on DOTA when scaling from 5 to 15 tasks. This failure is due to parameter and representation misalignment. We find that existing solutions, like regularization and dynamic routing, fail at scale because they are constrained by a fundamental trade-off: strengthening regularization to reduce inter-task conflict inadvertently suppresses the essential feature discrimination required for effective routing. In this work, we identify two root causes for this trade-off. First, uniform regularization disrupts inter-task knowledge sharing: shared underlying knowledge …


Penforge: On-The-Fly Expert Agent Construction For Automated Penetration Testing, Huihui Huang, Jieke Shi, Junkai Chen, Ting Zhang, Yikun Li, Chengran Yang, Eng Lieh Ouh, Lwin Khin Shar, David Lo Apr 2026

Penforge: On-The-Fly Expert Agent Construction For Automated Penetration Testing, Huihui Huang, Jieke Shi, Junkai Chen, Ting Zhang, Yikun Li, Chengran Yang, Eng Lieh Ouh, Lwin Khin Shar, David Lo

Research Collection School Of Computing and Information Systems

Penetration testing is essential for identifying vulnerabilities in web applications before real adversaries can exploit them. Recent work has explored automating this process with Large Language Model (LLM)-powered agents, but existing approaches either rely on a single generic agent that struggles in complex scenarios or narrowly specialized agents that cannot adapt to diverse vulnerability types. We therefore introduce PenForge, a framework that dynamically constructs expert agents during testing rather than relying on those prepared beforehand. By integrating automated reconnaissance of potential attack surfaces with agents instantiated on the fly for context-aware exploitation, PenForge achieves a 30.0% exploit success rate (12/40) …