Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Physical Sciences and Mathematics (9350)
- Computer Sciences (9042)
- Databases and Information Systems (3565)
- Software Engineering (2211)
- Artificial Intelligence and Robotics (1910)
-
- Social and Behavioral Sciences (1628)
- Business (1305)
- Information Security (1111)
- Numerical Analysis and Scientific Computing (1060)
- Graphics and Human Computer Interfaces (950)
- Engineering (935)
- Theory and Algorithms (513)
- Computer Engineering (469)
- Economics (425)
- Operations Research, Systems Engineering and Industrial Engineering (425)
- Programming Languages and Compilers (413)
- Communication (348)
- OS and Networks (346)
- International and Area Studies (336)
- Asian Studies (327)
- Public Affairs, Public Policy and Public Administration (309)
- Education (293)
- Social Media (271)
- Finance and Financial Management (248)
- Environmental Sciences (240)
- Medicine and Health Sciences (228)
- Transportation (219)
- Econometrics (214)
- Technology and Innovation (201)
- Management Information Systems (187)
- Keyword
-
- Machine learning (148)
- Deep learning (130)
- Artificial intelligence (127)
- Singapore (125)
- Social media (82)
-
- Reinforcement learning (74)
- Data mining (70)
- Privacy (67)
- Security (62)
- Sustainability (61)
- Cloud computing (60)
- Deep Learning (58)
- Optimization (56)
- Empirical study (55)
- Software engineering (55)
- Blockchain (52)
- Online learning (52)
- Visualization (52)
- Natural language processing (51)
- Neural networks (50)
- Training (50)
- Anomaly detection (49)
- Twitter (49)
- Large Language Models (48)
- Task analysis (48)
- Collaboration (47)
- Machine Learning (47)
- Feature extraction (45)
- Algorithms (44)
- Semantics (44)
- Publication Year
- Publication
-
- Research Collection School Of Computing and Information Systems (8553)
- Research Collection Lee Kong Chian School Of Business (428)
- Research Collection School Of Economics (340)
- Dissertations and Theses Collection (Open Access) (272)
- Research Collection School of Social Sciences (193)
-
- Research Collection College of Integrative Studies (138)
- Research Collection Yong Pung How School Of Law (110)
- Research Collection School Of Accountancy (65)
- Asian Management Insights (63)
- Perspectives@SMU (49)
- Dissertations and Theses Collection (20)
- Research Collection Library (19)
- SMU Press Releases and News (18)
- Knowledge@SMU (17)
- FORCE 2026 (14)
- Sim Kee Boon Institute for Financial Economics (14)
- Research@SMU: Connecting the Dots (13)
- Social Space (12)
- MITB Thought Leadership Series (11)
- Research Collection School of Computing and Information Systems (11)
- Oral History Collection (9)
- Report to Stakeholders (9)
- PhD Student’s Publications Collection (8)
- LARC Research Publications (7)
- Research Collection School of Economics (7)
- Student Publications (7)
- Centre for Computational Law (2022-2025) (6)
- CCX Research (5)
- SMU Research Data (4)
- 2024 AI for Research Week (3)
- Publication Type
- File Type
Articles 151 - 180 of 10460
Full-Text Articles in Entire DC Network
Perspectives On Interpretability For Neural Text Representations, Jia Peng Lim
Perspectives On Interpretability For Neural Text Representations, Jia Peng Lim
Dissertations and Theses Collection (Open Access)
In this dissertation, we investigate interpretability in the three elements of learning neural text representations: inputs, passed into models, to produce probabilistic outputs. We emphasise perspectives as we present alternative novel methods to mine and organise meaning in this work.
Models. We initiate our investigation by examining Neural Topic Models (NTM), proposing an alternate angle of interpreting its word-topic distribution, producing better topic representations for interpretation. Our method maps the problem of finding these better interpretations to classical NP-hard graph problems, enabling examination of topic distributions in a composite manner. Next, we apply our previous findings to extract interpretations from …
Func: Reducing The Impact Of Android Framework Evolution On Malware Detection, Hailong Yu, Tiantian Wang, Lwin Khin Shar, Hanmeng Li, David Lo
Func: Reducing The Impact Of Android Framework Evolution On Malware Detection, Hailong Yu, Tiantian Wang, Lwin Khin Shar, Hanmeng Li, David Lo
Research Collection School Of Computing and Information Systems
Android malware detection approaches commonly use APIs and permissions as features for classifying malware. However, since the release of the first Android operating system in 2008, the Android framework has undergone numerous version updates. The evolution of the Android framework over time has led to changes in APIs and permissions, including deprecations and replacements. These changes can result in inaccurate characterization of Android malware, thereby affecting performance of malware detectors. There is a lack of methods to mitigate the impact of Android framework evolution on malware detection. To fill this gap, we conduct a systematic study of the impact of …
Benchmarking Gaslighting Attacks Against Speech Large Language Models, Jinyang Wu, Bin Zhu, Xiandong Zou, Qiquan Zhang
Benchmarking Gaslighting Attacks Against Speech Large Language Models, Jinyang Wu, Bin Zhu, Xiandong Zou, Qiquan Zhang
PhD Student’s Publications Collection
As Speech Large Language Models (Speech LLMs) become increasingly integrated into voice-based applications, ensuring their robustness against manipulative or adversarial input becomes critical. Although prior work has studied adversarial attacks in text-based LLMs and vision-language models, the unique cognitive and perceptual challenges of speech-based interaction remain underexplored. In contrast, speech presents inherent ambiguity, continuity, and perceptual diversity, which make adversarial attacks more difficult to detect. In this paper, we introduce gaslighting attacks, strategically crafted prompts designed to mislead, override, or distort model reasoning as a means to evaluate the vulnerability of Speech LLMs. Specifically, we construct five manipulation strategies: Anger, …
Teacher-Student Diffusion Model For Text-Driven 3d Hand Motion Generation, Ching Lam Cheng, Bin Zhu, Shengfeng He
Teacher-Student Diffusion Model For Text-Driven 3d Hand Motion Generation, Ching Lam Cheng, Bin Zhu, Shengfeng He
PhD Student’s Publications Collection
Generating realistic 3D hand motion from natural language is vital for VR, robotics, and human-computer interaction. Existing methods either focus on full-body motion, overlooking detailed hand gestures, or require explicit 3D object meshes, limiting generality. We propose TSHaMo, a model-agnostic teacher-student diffusion framework for text-driven hand motion generation. The student model learns to synthesize motions from text alone, while the teacher leverages auxiliary signals (e.g., MANO parameters) to provide structured guidance during training. A co-training strategy enables the student to benefit from the teacher’s intermediate predictions while remaining text-only at inference. Evaluated using two diffusion backbones on GRAB and H2O, …
Digital Grief Technology To Support Bereavement: A Systematic Review Of Potential Benefits And Risks, Xun Ci Soh, Adalia Yin Hui Goh, Paye Shin Koh, Andree Hartanto
Digital Grief Technology To Support Bereavement: A Systematic Review Of Potential Benefits And Risks, Xun Ci Soh, Adalia Yin Hui Goh, Paye Shin Koh, Andree Hartanto
Research Collection School of Social Sciences
Grief is a universal and inevitable experience. However, the way we support the bereaved is changing, especially in the digital era. This systematic review examines the potential benefits and risks associated with various digital grief technologies, including online grief support groups, generative AI chatbots, online memorials, online therapy interventions, virtual reality, and digitally reproduced visuals or audio of the deceased. A systematic search was conducted in seven databases, and 30 articles were included in the final review. Findings indicate that digital grief technologies offer several benefits, such as reductions in grief and depressive symptoms, enhanced social support, greater accessibility, and …
Understanding Critical Thinking In Generative Artificial Intelligence Use: Development, Validation, And Correlates Of The Critical Thinking In Ai Use Scale, Gabriel R. Lau, Wei Yan Low, Louis Tay, Ysabel Thereze Ang Guevarra, Dragon Gašević, Andree Hartanto
Understanding Critical Thinking In Generative Artificial Intelligence Use: Development, Validation, And Correlates Of The Critical Thinking In Ai Use Scale, Gabriel R. Lau, Wei Yan Low, Louis Tay, Ysabel Thereze Ang Guevarra, Dragon Gašević, Andree Hartanto
Research Collection School of Social Sciences
Generative AI tools are increasingly embedded in everyday work and learning, yet their fluency, opacity, and propensity to hallucinate mean that users must critically evaluate AI outputs rather than accept them at face value. The present research conceptualises critical thinking in AI use as a dispositional tendency to verify the source and content of AI-generated information, to understand how models work and where they fail, and to reflect on the broader implications of relying on AI. Across six studies ( N = 1341), we developed and validated the 13-item critical thinking in AI use scale and mapped its nomological network. …
Uncertainty, Information, And Major Choice: Evidence From Chinese Gaokao, Sikun Dou
Uncertainty, Information, And Major Choice: Evidence From Chinese Gaokao, Sikun Dou
Dissertations and Theses Collection (Open Access)
An unresolved background risk makes people more averse to other, independent risks—most of all when a high-stakes, irreversible choice must be made under that risk. College major choice is such a setting. In China, students are admitted directly into a major they cannot easily change, and in the early 2000s they had to commit before learning their exam scores, bearing a field’s academic risk on top of an unresolved admission risk. Because STEM fields carry more academic risk and women are on average more risk-averse, this burden falls hardest on women’s entry into STEM. Exploiting a reform that moved application …
Sevoauth: Secure Voiceprint Authentication With Hash-Based Feature Transformation, Rui Zhang, Zheng Yan, Robert H. Deng
Sevoauth: Secure Voiceprint Authentication With Hash-Based Feature Transformation, Rui Zhang, Zheng Yan, Robert H. Deng
Research Collection School Of Computing and Information Systems
While voiceprint authentication offers convenient user authentication and access control through voice feature recognition, a critical research gap remains: existing voiceprint authentication systems fail to simultaneously achieve sound security against replay, spoofing, and adversarial attacks, preserve voice privacy leakage, and satisfy usability demand. Previous efforts have struggled to balance these issues comprehensively. To bridge this gap, we present SeVoAuth, a cloud-based Voiceprint Authentication as a Service (VAaaS) system designed to provide privacy preservation, robust security, and enhanced usability. SeVoAuth stores a synthesized voiceprint of a user in the cloud during user registration, thereby safeguarding the privacy of the real voiceprint …
Synthesis And Evaluation Of Long-Term History-Aware Medical Dialogue, Hebin Hu, Renke Dai, Ah-Hwee Tan, Yilin Kang
Synthesis And Evaluation Of Long-Term History-Aware Medical Dialogue, Hebin Hu, Renke Dai, Ah-Hwee Tan, Yilin Kang
Research Collection School Of Computing and Information Systems
An effective healthcare agent must be able to recall and reason over a patient’s longitudinal medical history. However, the absence of datasets with realistic long-term dialogue timelines limits systematic evaluation. Real clinical text is constrained by privacy and ethics, while existing benchmarks focus on isolated interactions, failing to capture cross-session reasoning. We introduce a framework for synthesizing high-quality, long-term medical dialogues with LLMs. Our approach entails a knowledge-guided decomposition into three stages: constructing synthetic patient profiles with diverse disease and complication trajectories, generating multiturn dialogues per encounter, and integrating them into a coherent longitudinal history dataset, MediLongChat. We establish three …
Open Source Software Development Tool Installation: Challenges And Strategies For Novice Developers, Larissa Salerno, Christoph Treude, Patanamon Thongtanunam
Open Source Software Development Tool Installation: Challenges And Strategies For Novice Developers, Larissa Salerno, Christoph Treude, Patanamon Thongtanunam
Research Collection School Of Computing and Information Systems
As the world of technology advances, so do the tools that software developers use to create new programs. In recent years, software development tools have become more popular, allowing developers to work more efficiently and produce higher-quality software. Still, installing such tools can be challenging for novice developers at the early stage of their careers, as they may face issues such as compatibility problems (e.g., with operating systems) and unclear instructions. Therefore, this work aims to investigate the challenges novice developers face when installing software development tools and the strategies they employ to overcome them. To investigate these, we conducted …
Generation Of Elaborated, Targeted And Effective Feedback For Novice Programmers Using Llm, Hua Leong Fwa
Generation Of Elaborated, Targeted And Effective Feedback For Novice Programmers Using Llm, Hua Leong Fwa
Research Collection School Of Computing and Information Systems
Programming errors and misconceptions are pervasive in novice programmers which causes difficulty in the learning of computer programming. Large Language Models (LLMs), with their ability to comprehend and generate programming codes have shown promising results in the automatic identification of errors. This can potentially benefit student programmers by providing them with timely formative feedback at efficiencies and scale that were not attainable previously. In this study, we leveraged an LLM - OpenAI o4-mini for the generation of elaborated, targeted feedback for novice programmers across PHP and JavaScript exercises. We contend that the feedback needs to be effective and targeted other …
Fighting Against Recruitment Scams: Theory-Driven Supervised Learning And Empirical Analysis For Digital Fraudulent Recruitment Posting Behavior, Tom (Tianteng) Wang, David (Jingjun) Xu, Keng Siau, Zhongju (John) Zhang
Fighting Against Recruitment Scams: Theory-Driven Supervised Learning And Empirical Analysis For Digital Fraudulent Recruitment Posting Behavior, Tom (Tianteng) Wang, David (Jingjun) Xu, Keng Siau, Zhongju (John) Zhang
Research Collection School Of Computing and Information Systems
The number of recruitment postings on digital recruitment hiring platforms has increased since the COVID-19 pandemic. However, the weak surveillance and operations of these platforms, combined with the fact that most job seekers have relatively low vigilance and a strong desire for recruitment offers, enable scammers to easily deceive job seekers for their money and confidential information. In this work, we combine prevailing text mining techniques (i.e., ChatGPT with prompting engineering and supervised machine learning) with interpersonal deception theory (IDT) from social science to design an interpretable IT system to predict fraudulent recruitment postings on digital recruitment-hiring platforms. We compare …
Securing Cloud-Native Systems: From Vulnerability Analysis To External And Insider Threat Detection, Jiongchi Yu
Securing Cloud-Native Systems: From Vulnerability Analysis To External And Insider Threat Detection, Jiongchi Yu
Dissertations and Theses Collection (Open Access)
Cloud-native systems have become the backbone of modern software infrastructure. However, their dynamic resource orchestration and complex configurability introduce a large attack surface and intricate security challenges. Adversaries can externally exploit vulnerabilities in cloud components or perform insider movement within cloud environments to launch attacks. As these systems increasingly support critical services, security breaches can lead to severe operational and economic consequences.
Despite extensive efforts in vulnerability detection and attack monitoring, existing approaches struggle to remain effective in cloud-native environments characterized by rapid evolution and inherent heterogeneity. In particular, they exhibit three fundamental limitations: (1) Insufficient understanding of defect patterns …
Leading Ai Adoption In Organizations: Introducing A Behavioral Human-Centered Approach, Shane Schweitzer, Devesh Narayan, Jack Mcguire, David De Cremer
Leading Ai Adoption In Organizations: Introducing A Behavioral Human-Centered Approach, Shane Schweitzer, Devesh Narayan, Jack Mcguire, David De Cremer
Research Collection Lee Kong Chian School Of Business
Initiatives to implement AI technologies in organizations fail at an alarming rate. We argue that leading the adoption of AI is not a simple engineering exercise but rather represents a behavioral exercise where change management principles—the process by which organizations plan, implement, and embed changes in practices—are employed. However, many AI initiatives in business focus predominantly on the AI systems themselves, assuming humans will fall in line. To solve this, we integrate ideas from change management with scholarship on human-centered artificial intelligence to offer a behavioral approach that accounts for the impact of AI adoption on humans at all stages …
Public Pool Usage As Adaptation Against Urban Heat, Stefan Borsky, Eric Fesselmeyer
Public Pool Usage As Adaptation Against Urban Heat, Stefan Borsky, Eric Fesselmeyer
Research Collection College of Integrative Studies
This paper examines the relationship between urban heat and outdoor public pool usage. Using attendance data from all 53 outdoor public pools in New York City, we analyze nonlinear effects of heat on pool usage across socioeconomic contexts. Pool attendance rises sharply with heat, especially in low-income neighborhoods where alternative coping options are likely limited. We also find that public pools reduce heat-related emergency medical service calls. Our findings highlight the need for equitable investment in blue infrastructure to enhance urban climate resilience and demonstrate how this type of adaptive infrastructure can play a critical role in managing urban heat.
Selective Concolic Testing, Guofeng Zhang, Zhenbang Chen, Ziqi Shuai, Jun Sun, Weijiang Hong, Yufeng Zhang, Ji Wang, Yang Liu
Selective Concolic Testing, Guofeng Zhang, Zhenbang Chen, Ziqi Shuai, Jun Sun, Weijiang Hong, Yufeng Zhang, Ji Wang, Yang Liu
Research Collection School Of Computing and Information Systems
The principled combination of symbolic execution and random testing lacks a formal foundation, especially in deciding which inputs to symbolize. We propose selective concolic testing, a cost-aware framework that formulates this choice as an optimized policy problem of a MDP (Markov Decision Process). We model program exploration over a finite control-flow graph, where MDP states represent covered statements, actions partition path constraints into symbolic and random fragments, rewards reflect coverage gain, and costs account for SMT solving effort and sampling inefficiency. Our framework yields the first formal characterization of selective symbolization as policy synthesis in a probabilistic system. We prove …
Detecting Doubt In Reflective Learning: A Learning Analytics Study With Large And Small Language Models, Eng Lieh Ouh, Kar Way Tan, Siaw Ling Lo, Yuhao Zhang
Detecting Doubt In Reflective Learning: A Learning Analytics Study With Large And Small Language Models, Eng Lieh Ouh, Kar Way Tan, Siaw Ling Lo, Yuhao Zhang
Research Collection School Of Computing and Information Systems
Reflective learning enhances understanding, especially when instructors promptly address difficulties raised in student reflections. Automated doubt detection can reduce time for instructors, yet existing classification approaches take substantial time for manual annotation and model training. This paper investigates whether large and small language models (LLMs, SLMs) can automate doubt detection without time-consuming training. Using a dataset of anonymized student reflections, we evaluate zeroshot, few-shot prompting, and multi-step reasoning against prior supervised classification baselines. We show that LLMs (GPT-4o, Claude-4, Gemini-2.5) surpass earlier F1 scores without prompting, while prompting further improves their performance. However, using proprietary LLMs can raise cost and …
Gencode: A Generic Data Augmentation Framework For Boosting Deep Learning-Based Code Understanding, Zeming Dong, Qiang Hu, Xiaofei Xie, Maxime Cordy, Mike Papadakis, Yves Le Traon, Jianjun Zhao
Gencode: A Generic Data Augmentation Framework For Boosting Deep Learning-Based Code Understanding, Zeming Dong, Qiang Hu, Xiaofei Xie, Maxime Cordy, Mike Papadakis, Yves Le Traon, Jianjun Zhao
Research Collection School Of Computing and Information Systems
Pre-trained code models lead the era of code intelligence, with multiple models designed with impressive performance. However, one important problem, data augmentation for code data that automatically helps developers prepare training data lacks study in this field. In this paper, we introduce a generic data augmentation framework, GenCode, to enhance the training of code understanding models. Simply speaking, GenCode follows a generation-and-selection paradigm to prepare useful training code data. Specifically, it employs code augmentation techniques to generate new code candidates first and then identifies important ones as the training data by influence scores. To evaluate the effectiveness of GenCode, we …
Market Reactions To Deceptive Language In Fake News: Implications From Language Expectancy Theory And Transfer Learning, Ka Chung Ng, Ping Fan Ke, Ping Fan, Mike So, Tam, Kar Yan
Market Reactions To Deceptive Language In Fake News: Implications From Language Expectancy Theory And Transfer Learning, Ka Chung Ng, Ping Fan Ke, Ping Fan, Mike So, Tam, Kar Yan
Research Collection School Of Computing and Information Systems
The advent of generative artificial intelligence (AI) has heightened the proliferation of fake news. A key challenge is the limited real-world data to investigate the societal impact of fake news produced by generative AI. In this paper, we examine stock market reactions to financial news articles that exhibit stylometric similarity to human-crafted and AI-crafted fake financial news. Grounded in language expectancy theory, we employ a style-based transfer learning model, pre-trained to recognizing deceptive language employed in various types of fake news intricacies. We then apply this model to a comprehensive dataset of financial news, assigning a “veracity style score” to …
Quantitative Bounds On Resource Usage Of Probabilistic Programs, Krishnendu Chatterjee, Amir Kafshdar Goharshady, Tobias Meggendorfer, Dorde Zikelic
Quantitative Bounds On Resource Usage Of Probabilistic Programs, Krishnendu Chatterjee, Amir Kafshdar Goharshady, Tobias Meggendorfer, Dorde Zikelic
Research Collection School Of Computing and Information Systems
Cost analysis, also known as resource usage analysis, is the task of finding bounds on the total cost of a program and is a well-studied problem in static analysis. In this work, we consider two classical quantitative problems in cost analysis for probabilistic programs. The first problem is to find a bound on the expected total cost of the program. This is a natural measure for the resource usage of the program and can also be directly applied to average-case runtime analysis. The second problem asks for a tail bound, i.e. given a threshold t the goal is to find …
Brief Virtual Reality And Mixed Reality Mindfulness Breathing Exercise For Emotional Well-Being And Cognitive Functions In University Students: Within-Subjects Experimental Design Study, Zoey Khai Yee Eun, Charmaine Jiali Koh, Hwajin Yang, Adalia Yin Hui Goh, Meilan Hu, K Tennakoon Appuhamillage Sandeeshwara Kasturiratna, Andree Hartanto
Brief Virtual Reality And Mixed Reality Mindfulness Breathing Exercise For Emotional Well-Being And Cognitive Functions In University Students: Within-Subjects Experimental Design Study, Zoey Khai Yee Eun, Charmaine Jiali Koh, Hwajin Yang, Adalia Yin Hui Goh, Meilan Hu, K Tennakoon Appuhamillage Sandeeshwara Kasturiratna, Andree Hartanto
Research Collection School of Social Sciences
Background: Mindfulness has been shown to enhance emotional well-being and cognitive performance, yet much of this evidence stems from interventions requiring prolonged practice, making them time-consuming and less accessible. Recent studies suggest that brief mindfulness sessions may also yield positive outcomes, but the effectiveness of such interventions in virtual reality (VR) and mixed reality (MR) remains underexplored. Objective: This study investigates the effects of brief mindfulness breathing exercises delivered through VR and MR on attentional and emotional restoration and self-control capacity. Methods: Using a within-subjects experimental design, 102 undergraduate participants (n=83, 81.4% female; mean age 20.87, SD 1.89) completed a …
Mease: Multi-Agent Episodic Action Sequence Explanation, Phyo Wai Khaing, Minghong Geng, Shubham Pateria, Budhitama Subagdja, Ah-Hwee Tan
Mease: Multi-Agent Episodic Action Sequence Explanation, Phyo Wai Khaing, Minghong Geng, Shubham Pateria, Budhitama Subagdja, Ah-Hwee Tan
Research Collection School Of Computing and Information Systems
Multi-agent reinforcement learning (MARL) achieves remarkable performance in complex coordination tasks, yet interpreting the emergent behaviors of trained agents remains a fundamental challenge. Most current explainability methods focus on individual agent decisions, overlooking the critical interplay of joint strategiesand temporal coordination patterns that define successful multi-agent policies. We present MEASE (Multi-agent Episodic Action Sequence Explanation), a novel explainable MARL (XMARL) framework that explains trained MARL policies as human-interpretable emergent cooperative joint behaviors. MEASE employs a cognition-inspired episodic memory model to learn spatio-temporal multi-agent interaction patterns, coupled with abstraction algorithms that identify significant cooperative agent behaviors. We evaluate MEASE on diverse …
Collaborative Practices And Tool Utilization In Software Development Projects: A Student Perspective, Yi Meng Lau, Muhammad Syahmi Bin Abbas, Lingxiao Jiang
Collaborative Practices And Tool Utilization In Software Development Projects: A Student Perspective, Yi Meng Lau, Muhammad Syahmi Bin Abbas, Lingxiao Jiang
Research Collection School Of Computing and Information Systems
Software development is a collaborative activity that depends on effective teamwork, shared understanding, and coordinated use of development practices and tools. While these aspects are well studied in professional environments, they are less frequently examined within software engineering education. This study investigates how students collaborate in group projects, focusing on collaborative practices, tool usage, and their perceptions of software quality. We conducted a quantitative post-project survey with 143 second-year undergraduate students enrolled in a software development course. The results show that students actively share information and often establish team norms to support coordination and collaboration. However, students face challenges in …
Causality-Driven Test Case Minimisation For Cyber-Physical Systems, Michael Foster, Christopher M. Poskitt, Nicholas R. Latimer, Neil Walkinshaw, Richard Somers, Robert M. Hierons
Causality-Driven Test Case Minimisation For Cyber-Physical Systems, Michael Foster, Christopher M. Poskitt, Nicholas R. Latimer, Neil Walkinshaw, Richard Somers, Robert M. Hierons
Research Collection School Of Computing and Information Systems
Cyber-physical systems allow digital control systems to interact with the physical world using sensors and actuators. They are increasingly being used to automate critical infrastructure, where software faults can have dire consequences. Due to the complex nature and unpredictability of these systems, their resilience is often tested using a technique called fuzzing, which generates quasi-random sequences of sensor and actuator manipulations with the goal of forcing a system into unsafe states. However, there is currently no way of determining which manipulations of a test case cause a failure without systematically removing each one and re-running the test, which can be …
Natural Adversaries: Fuzzing Autonomous Vehicles With Realistic Roadside Object Placements, Yang Sun, Haoyu Wang, Christopher M. Poskitt, Jun Sun
Natural Adversaries: Fuzzing Autonomous Vehicles With Realistic Roadside Object Placements, Yang Sun, Haoyu Wang, Christopher M. Poskitt, Jun Sun
Research Collection School Of Computing and Information Systems
The emergence of Autonomous Vehicles (AVs) has spurred research into testing the resilience of their perception systems, i.e., ensuring that they are not susceptible to critical misjudgements. It is important that these systems are tested not only with respect to other vehicles on the road, but also with respect to objects placed on the roadside. Trash bins, billboards, and greenery are examples of such objects, typically positioned according to guidelines developed for the human visual system, which may not align perfectly with the needs of AVs. Existing tests, however, usually focus on adversarial objects with conspicuous shapes or patches, which …
Enhancing Action And Ingredient Modeling For Semantically Grounded Recipe Generation, Guoshan Liu, Bin Zhu, Yian Li, Jingjing Chen, Chong-Wah Ngo, Yu-Gang Jiang
Enhancing Action And Ingredient Modeling For Semantically Grounded Recipe Generation, Guoshan Liu, Bin Zhu, Yian Li, Jingjing Chen, Chong-Wah Ngo, Yu-Gang Jiang
Research Collection School Of Computing and Information Systems
Recent advances in Multimodal Large Language Models (MLMMs) have enabled recipe generation from food images, yet outputs often contain semantically incorrect actions or ingredients despite high lexical scores (e.g., BLEU, ROUGE). To address this gap, we propose a semantically grounded framework that predicts and validates actions and ingredients as internal context for instruction generation. Our two-stage pipeline combines supervised fine-tuning (SFT) with reinforcement fine-tuning (RFT): SFT builds foundational accuracy using an Action-Reasoning dataset and ingredient corpus, while RFT employs frequency-aware rewards to improve long-tail action prediction and ingredient generalization. A Semantic Confidence Scoring and Rectification (SCSR) module further filters and …
Be Responsible In Your Answers! Monitoring Out-Of-Domain Behaviors In Domain-Specific Llms, Boquan Li, Chenzhe Lou, Zhe Ren, Peixin Zhang, Zirui Fu, Jun Sun, Yaowen Zheng
Be Responsible In Your Answers! Monitoring Out-Of-Domain Behaviors In Domain-Specific Llms, Boquan Li, Chenzhe Lou, Zhe Ren, Peixin Zhang, Zirui Fu, Jun Sun, Yaowen Zheng
Research Collection School Of Computing and Information Systems
Large Language Models (LLMs) have accelerated the rapid development of chatbot web applications in various domains, such as coding, biomedicine and psychology. Compared to general LLMs like ChatGPT, domain-specific LLMs require a greater sense of responsibility. For instance, if a programming LLM casually answers medical or psychological questions, it not only misleads the public but also poses legal risks. This highlights new demands for monitoring and preventing such irresponsible behaviors. Existing efforts attempt to monitor LLMs from multiple aspects, such as lying, jailbreaks, and toxic content, while overlooking out-of-domain behaviors. In this work, we propose an innovative LLM domain monitoring …
Architecture-Agnostic Test-Time Adaptation Via Backprop-Free Embedding Alignment, Xiao Ma, Young D. Kwon, Pan Zhou, Dong Ma
Architecture-Agnostic Test-Time Adaptation Via Backprop-Free Embedding Alignment, Xiao Ma, Young D. Kwon, Pan Zhou, Dong Ma
PhD Student’s Publications Collection
Test-Time Adaptation (TTA) adapts a deployed model during online inference to mitigate the impact of domain shift. While achieving strong accuracy, most existing methods rely on backpropagation, which is memory and computation intensive, making them unsuitable for resource-constrained devices. Recent attempts to reduce this overhead often suffer from high latency or are tied to specific architectures such as ViT-only or CNN-only. In this work, we revisit domain shift from an embedding perspective. Our analysis reveals that domain shift induces three distinct structural changes in the embedding space: translation (mean shift), scaling (variance shift), and rotation (covariance shift). Based on this …
Scalable Multi-Task Low-Rank Model Adaptation, Zichen Tian, Antoine Ledent, Qianru Sun
Scalable Multi-Task Low-Rank Model Adaptation, Zichen Tian, Antoine Ledent, Qianru Sun
PhD Student’s Publications Collection
Scaling multi-task low-rank adaptation (LoRA) to a large number of tasks induces catastrophic performance degradation, such as an accuracy drop from 88.2% to 2.0% on DOTA when scaling from 5 to 15 tasks. This failure is due to parameter and representation misalignment. We find that existing solutions, like regularization and dynamic routing, fail at scale because they are constrained by a fundamental trade-off: strengthening regularization to reduce inter-task conflict inadvertently suppresses the essential feature discrimination required for effective routing. In this work, we identify two root causes for this trade-off. First, uniform regularization disrupts inter-task knowledge sharing: shared underlying knowledge …
Penforge: On-The-Fly Expert Agent Construction For Automated Penetration Testing, Huihui Huang, Jieke Shi, Junkai Chen, Ting Zhang, Yikun Li, Chengran Yang, Eng Lieh Ouh, Lwin Khin Shar, David Lo
Penforge: On-The-Fly Expert Agent Construction For Automated Penetration Testing, Huihui Huang, Jieke Shi, Junkai Chen, Ting Zhang, Yikun Li, Chengran Yang, Eng Lieh Ouh, Lwin Khin Shar, David Lo
Research Collection School Of Computing and Information Systems
Penetration testing is essential for identifying vulnerabilities in web applications before real adversaries can exploit them. Recent work has explored automating this process with Large Language Model (LLM)-powered agents, but existing approaches either rely on a single generic agent that struggles in complex scenarios or narrowly specialized agents that cannot adapt to diverse vulnerability types. We therefore introduce PenForge, a framework that dynamically constructs expert agents during testing rather than relying on those prepared beforehand. By integrating automated reconnaissance of potential attack surfaces with agents instantiated on the fly for context-aware exploitation, PenForge achieves a 30.0% exploit success rate (12/40) …