Fact-Checker: A Web Application For Leveraging Large Language Models For Fact-Checking Youtube Videos,
2025
California State University, San Bernardino
Fact-Checker: A Web Application For Leveraging Large Language Models For Fact-Checking Youtube Videos, Andrew R. Craig
Electronic Theses, Projects, and Dissertations
Fact-Checker is a web application that allows users to fact-check YouTube videos. It feeds YouTube’s closed captioning transcript to a large language model (LLM) to extract claims. It then uses multiple LLMs, such as Gemini, Llama, and Claude, to verify these claims. The modular design makes it easy to change to a different LLM or model if needed. The application is built using Python for access to Application Programming Interfaces (APIs) and Streamlit as the front-end framework. The utilization of Docker and Dockerfiles enables easy distribution and deployment. It enables the application to be deployed on almost any hardware platform …
Reimagining Education With Ai,
2025
SKEMA Business School
Reimagining Education With Ai, Margherita Pagani, Steven M. Miller, Jerry Wind
Research Collection School Of Computing and Information Systems
This chapter examines AI’s transformative potential in education, focusing on Generative AI (GenAI) and Large Language Models (LLMs) while at the same time emphasizing the importance of grounding and guiding AI efforts with learning science and education research findings. It synthesizes analyses and expert recommendations, highlighting opportunities like personalized learning and enhanced teacher productivity, alongside challenges such as over-reliance on AI. Practical steps for instructors include adopting a question-first approach, utilizing AI for personalized feedback, designing AI-enhanced learning experiences, fostering critical thinking, and ensuring ethical AI use. The chapter concludes with strategic recommendations for leveraging AI to sustainably improve educational …
Dreamanime: Learning Style-Identity Textual Disentanglement For Anime And Beyond,
2025
Singapore Management University
Dreamanime: Learning Style-Identity Textual Disentanglement For Anime And Beyond, Chenshu Xu, Yangyang Xu, Huaidong Zhang, Xuemiao Xu, Shengfeng He
Research Collection School Of Computing and Information Systems
Text-to-image generation models have significantly broadened the horizons of creative expression through the power of natural language. However, navigating these models to generate unique concepts, alter their appearance, or reimagine them in unfamiliar roles presents an intricate challenge. For instance, how can we exploit language-guided models to transpose an anime character into a different art style, or envision a beloved character in a radically different setting or role? This paper unveils a novel approach named DreamAnime, designed to provide this level of creative freedom. Using a minimal set of 2-3 images of a user-specified concept such as an anime character …
L3net: Localized And Layered Reparameterization For Incremental Learning,
2025
Singapore Management University
L3net: Localized And Layered Reparameterization For Incremental Learning, Xuandi Luo, Huaidong Zhang, Yi Xie, Hongrui Zhang, Xuemiao Xu, Shengfeng He
Research Collection School Of Computing and Information Systems
Model-based class incremental learning (CIL) methods aim to address the challenge of catastrophic forgetting by retaining certain parameters and expanding the model architecture. However, retaining too many parameters can lead to an overly complex model, increasing inference overhead. Additionally, compressing these parameters to reduce the model size can result in performance degradation. To tackle these challenges, we propose a novel three-stage CIL framework called Localized and Layered Reparameterization for Incremental Learning (L3Net). The rationale behind our approach is to balance model complexity and performance by selectively expanding and optimizing critical components. Specifically, the framework introduces a Localized Dual-path Expansion structure, …
Solving Two-Stage Stochastic Integer Programs Via Representation Learning,
2025
Singapore Management University
Solving Two-Stage Stochastic Integer Programs Via Representation Learning, Yaoxin Wu, Zhiguang Cao, Wen Song, Yingqian Zhang
Research Collection School Of Computing and Information Systems
Solving stochastic integer programs (SIPs) is extremely intractable due to the high computational complexity. To solve two-stage SIPs efficiently, we propose a conditional variational autoencoder (CVAE) for scenario representation learning. A graph convolutional network (GCN) based VAE embeds scenarios into a low-dimensional latent space, conditioned on the deterministic context of each instance. With the latent representations of stochastic scenarios, we perform two auxiliary tasks: objective prediction and scenario contrast, which predict scenario objective values and the similarities between them, respectively. These tasks further integrate objective information into the representations through gradient backpropagation. Experiments show that the learned scenario representations can …
How To Enable Effective Cooperation Between Humans And Nlp Models: A Survey Of Principles, Formalizations, And Beyond,
2025
Singapore Management University
How To Enable Effective Cooperation Between Humans And Nlp Models: A Survey Of Principles, Formalizations, And Beyond, Chen Huang, Yang Deng, Wenqiang Lei, Jiancheng Lv, Tat-Seng Chua, Jimmy Huang
Research Collection School Of Computing and Information Systems
With the advancement of large language models (LLMs), intelligent models have evolved from mere tools to autonomous agents with their own goals and strategies for cooperating with humans. This evolution has birthed a novel paradigm in NLP, i.e., human-model cooperation, that has yielded remarkable progress in numerous NLP tasks in recent years. In this paper, we take the first step to present a thorough review of human-model cooperation, exploring its principles, formalizations, and open challenges. In particular, we introduce a new taxonomy that provides a unified perspective to summarize existing approaches. Also, we discuss potential frontier areas and their corresponding …
Evowiki: Evaluating Llms On Evolving Knowledge,
2025
Singapore Management University
Evowiki: Evaluating Llms On Evolving Knowledge, Wei Tang, Yixin Cao, Yang Deng, Jiahao Ying, Bo Wang, Yizhe Yang, Yuyue Zhao, Qi Zhang, Xuanjing Huang, Yu-Gang Jiang, Yong Liao
Research Collection School Of Computing and Information Systems
Knowledge utilization is a critical aspect of LLMs, and understanding how they adapt to evolving knowledge is essential for their effective deployment. However, existing benchmarks are predominantly static, failing to capture the evolving nature of LLMs and knowledge, leading to inaccuracies and vulnerabilities such as contamination. In this paper, we introduce EvoWiki, an evolving dataset designed to reflect knowledge evolution by categorizing information into stable, evolved, and uncharted states. EvoWiki is fully auto-updated, enabling precise evaluation of continuously changing knowledge and newly released LLMs. Through experiments with Retrieval-Augmented Generation (RAG) and Continual Learning (CL), we evaluate how effectively LLMs adapt …
Knowledge Boundary Of Large Language Models: A Survey,
2025
Singapore Management University
Knowledge Boundary Of Large Language Models: A Survey, Moxin Li, Yong Zhao, Wenxuan Zhang, Shuaiyi Li, Wenya Xie, See-Kiong Ng, Tat-Seng Chua, Yang Deng
Research Collection School Of Computing and Information Systems
Although large language models (LLMs) store vast amount of knowledge in their parameters, they still have limitations in the memorization and utilization of certain knowledge, leading to undesired behaviors such as generating untruthful and inaccurate responses. This highlights the critical need to understand the knowledge boundary of LLMs, a concept that remains inadequately defined in existing research. In this survey, we propose a comprehensive definition of the LLM knowledge boundary and introduce a formalized taxonomy categorizing knowledge into four distinct types. Using this foundation, we systematically review the field through three key lenses: the motivation for studying LLM knowledge boundaries, …
Beware Of Your Po! Measuring And Mitigating Ai Safety Risks In Role-Play Fine-Tuning Of Llms,
2025
Singapore Management University
Beware Of Your Po! Measuring And Mitigating Ai Safety Risks In Role-Play Fine-Tuning Of Llms, Weixiang Zhao, Yulin Hu, Yang Deng, Jiahe Guo, Xingyu Sui, Xinyang Han, An Zhang, Yanyan Zhao, Bing Qin, Tat-Seng Chua, Ting Liu
Research Collection School Of Computing and Information Systems
Although large language models (LLMs) store vast amount of knowledge in their parameters, they still have limitations in the memorization and utilization of certain knowledge, leading to undesired behaviors such as generating untruthful and inaccurate responses. This highlights the critical need to understand the knowledge boundary of LLMs, a concept that remains inadequately defined in existing research. In this survey, we propose a comprehensive definition of the LLM knowledge boundary and introduce a formalized taxonomy categorizing knowledge into four distinct types. Using this foundation, we systematically review the field through three key lenses: the motivation for studying LLM knowledge boundaries, …
Browsing Like Human: A Multimodal Web Agent With Experiential Fast-And-Slow Thinking,
2025
Singapore Management University
Browsing Like Human: A Multimodal Web Agent With Experiential Fast-And-Slow Thinking, Haohao Luo, Jiayi Kuang, Wei Liu, Ying Shen, Jian Luan, Yang Deng
Research Collection School Of Computing and Information Systems
Automating web navigation which aims to build a web agent that follows user instructions to complete tasks like booking flights by interacting with websites, has received increasing attention due to its practical value. Although existing web agents are mostly equipped with visual perception, planning, and memory abilities, their reasoning process are still deviate from human cognition. In this work, we study the human thought pattern to empower agent with more human-like abilities in web navigation. To tackle this problem, we propose a novel multimodal web agent framework called WebExperT, which is designed to emulate the human planning process of “thinking …
Mpo: Multilingual Safety Alignment Via Reward Gap Optimization,
2025
Singapore Management University
Mpo: Multilingual Safety Alignment Via Reward Gap Optimization, Weixiang Zhao, Yulin Hu, Yang Deng, Tongtong Wu, Wenxuan Zhang, Jiahe Guo, An Zhang, Yanyan Zhao, Bing Qin, Tat-Seng Chua, Ting Liu
Research Collection School Of Computing and Information Systems
Large language models (LLMs) have become increasingly central to AI applications worldwide, necessitating robust multilingual safety alignment to ensure secure deployment across diverse linguistic contexts. Existing preference learning methods for safety alignment, such as RLHF and DPO, are primarily monolingual and struggle with noisy multilingual data. To address these limitations, we introduce Multilingual reward gaP Optimization (MPO), a novel approach that leverages the well-aligned safety capabilities of the dominant language (e.g., English) to improve safety alignment across multiple languages. MPO directly minimizes the reward gap difference between the dominant language and target languages, effectively transferring safety capabilities while preserving the …
Think Both Ways: Teacher-Student Bidirectional Reasoning Enhances Mcq Generation And Distractor Quality,
2025
Singapore Management University
Think Both Ways: Teacher-Student Bidirectional Reasoning Enhances Mcq Generation And Distractor Quality, Yimiao Qiu, Yang Deng, Quanming Yao, Zhimeng Zhang, Zhiang Dong, Chang Yao, Jingyuan Chen
Research Collection School Of Computing and Information Systems
Generating high-quality Multiple Choice Questions (MCQs) remains challenging for educational tools due to the need for contextual relevance and plausible distractors. Existing methods still struggle with these dual requirements, leading to questions that lack depth and distractors that are either too obvious or irrelevant. In this paper, we propose BiFlow, a novel framework that integrates bidirectional reasoning perspectives: teacher reasoning generates contextually relevant questions and plausible distractors, while student reasoning evaluates question clarity and the misleading nature of the distractors. To further enhance reasoning, we introduce PathFinder, a mechanism that employs breadth-first search and Chainof-Thought (CoT) strategies to explore diverse …
Coleclip: Open-Domain Continual Learning Via Joint Task Prompt And Vocabulary Learning,
2025
Singapore Management University
Coleclip: Open-Domain Continual Learning Via Joint Task Prompt And Vocabulary Learning, Yukun Li, Guansong Pang, Wei Suo, Chenchen Chen, Yuling Xi, Lingqiao Liu, Hao Chen, Guoqiang Liang, Peng Wang
Research Collection School Of Computing and Information Systems
This article investigates the problem of continual learning (CL) of vision-language models (VLMs) in open domains, where models are required to perform continual updating and inference on a stream of datasets from diverse seen and unseen domains with novel classes. Such a capability is crucial for various applications in open environments, e.g., AI assistants, autonomous driving systems, and robotics. Current CL studies mostly focus on closed-set scenarios in a single domain with known classes. Large pretrained VLMs such as CLIP have showcased exceptional zero-shot recognition capabilities, and several recent studies have leveraged the unique characteristics of VLMs to mitigate catastrophic …
Bhvit: Binarized Hybrid Vision Transformer,
2025
Singapore Management University
Bhvit: Binarized Hybrid Vision Transformer, Tian Gao, Yu Zhang, Zhiyuan Zhang, Huajun Liu, Kaijie Yin, Chengzhong Xu, Hui Kong
Research Collection School Of Computing and Information Systems
Model binarization has made significant progress in enabling real-time and energy-efficient computation for con-volutional neural networks (CNN), offering a potential solution to the deployment challenges faced by Vision Transformers (ViTs) on edge devices. However, due to the structural differences between CNN and Transformer architectures, simply applying binary CNN strategies to the ViT models will lead to a significant performance drop. To tackle this challenge, we propose BHViT, a binarization-friendly hybrid ViT architecture and its full binarization model with the guidance of three important observations. Initially, BHViT utilizes the local information interaction and hierarchical feature aggregation technique from coarse to fine …
Other Orienteering Problem Variants,
2025
Singapore Management University
Other Orienteering Problem Variants, Pieter Vansteenwegen, Aldy Gunawan
Research Collection School Of Computing and Information Systems
In this chapter, different variants of routing problems with profits will be discussed. Based on what is available in literature, mostly variants of the orienteering problem will be discussed. A first variant considers capacity constraints, since these appear frequently in many practical applications. Next, multi-objective orienteering problems, explicitly considering different types of profits separately are discussed. Time-dependent and stochastic travel times are also relevant for most practical applications. These are considered together with time-dependent and stochastic profits. More and more routing problems are considered together with inventory management. For routing problems with profits, this leads to the inventory orienteering problem, …
Optimal Transport Alignment Of User Preferences From Ratings And Texts,
2025
Singapore Management University
Optimal Transport Alignment Of User Preferences From Ratings And Texts, Nhu Thuat Tran, Hady Wirawan Lauw
Research Collection School Of Computing and Information Systems
Modeling hidden factors driving user preferences is crucial for recommendation yet challenging due to sparse rating data. While aligning preference factors from ratings and texts, as a solution, shows improvements, existing methods impose restrictive one-to-one factor correspondences and underutilize cross-modal interest signals. We propose an optimal transport (OT) approach to address these gaps. By modeling rating- and text-based preference factors as distributions, we compute an OT plan that captures their probabilistic relationships. This plan serves dual roles: 1) to regularize cross-modal preference factors without rigid correspondence assumptions, and 2) to blend preference signals across modalities through barycentric mapping. Experiments on …
Analysis Of Extended Producer Responsibility In E-Waste Management: Policy Drivers And Challenges In Singapore,
2025
Singapore Management University
Analysis Of Extended Producer Responsibility In E-Waste Management: Policy Drivers And Challenges In Singapore, Aldy Gunawan, Aidan Marc Wong, Tasaporn Visawameteekul, Minh Phuong Huynh, Linh Chi Tran
Research Collection School Of Computing and Information Systems
This paper examines the role of the Extended Producer Responsibility (EPR) scheme in electronic waste (e-waste) management in Singapore. It investigates the policy drivers and challenges of e-waste management, using data from an online survey to explore the attitudes and behaviors of young consumers, with a particular focus on youth. We employ the Theory of Reasoned Action (TRA) and the Theory of Planned Behavior (TPB) frameworks to develop a model that examines the relationships among attitudes, perceived norms, awareness, and perceived convenience in relation to EPR awareness and perception. The findings highlight the need for customized policies tailored to different …
A Review: The Beauty Of Serendipity Between Integrated Circuit Security And Artificial Intelligence,
2025
Singapore Management University
A Review: The Beauty Of Serendipity Between Integrated Circuit Security And Artificial Intelligence, Chen Dong, Decheng Qiu, Bolun Li, Yang Yang, Chenxi Lyu, Dong Cheng, Hao Zhang, Zhenyi. Chen
Research Collection School Of Computing and Information Systems
Integrated circuits are the core of a cyber-physical system, where tens of billions of components are integrated into a tiny silicon chip to conduct complex functions. To maximize utilities, the design and manufacturing life cycle of integrated circuits rely on numerous untrustworthy third parties, forming a global supply chain model. At the same time, this model produces unpredictable and catastrophic issues, threatening the security of individuals and countries. As for guaranteeing the security of ultra-highly integrated chips, detecting slight abnormalities caused by malicious behavior in the current and voltage is challenging, as is achieving computability within a reasonable time and …
Explainable Multimodal Sentiment Analysis Of Social Media Visual Content For Child Safety,
2025
Singapore Management University
Explainable Multimodal Sentiment Analysis Of Social Media Visual Content For Child Safety, Yee Sen Tan, Zhaoxia Wang
Research Collection School Of Computing and Information Systems
Ensuring the safety and well-being of children is increasingly important, especially in a world where visual content is pervasive. This paper proposes a novel multimodal, multilingual, and multiclass sentiment analysis method for social media content, aimed at improving content moderation for child safety. Our approach integrates textual, visual, and audio data from videos, categorizing sentiment into four levels: positive, slightly negative, negative, and strongly negative, enabling granular detection of harmful content. To enhance explainability and trust, we also leverage interpretable mechanisms to analyze the contributions of each modality. Evaluation of our method demonstrates strong generalization across diverse video types, and …
Rl4co: An Extensive Reinforcement Learning For Combinatorial Optimization Benchmark,
2025
Singapore Management University
Rl4co: An Extensive Reinforcement Learning For Combinatorial Optimization Benchmark, Federico Berto, Et. Al
Research Collection School Of Computing and Information Systems
Combinatorial optimization (CO) is fundamental to several real-world applications, from logistics and scheduling to hardware design and resource allocation. Deep reinforcement learning (RL) has recently shown significant benefits in solving CO problems, reducing reliance on domain expertise and improving computational efficiency. However, the absence of a unified benchmarking framework leads to inconsistent evaluations, limits reproducibility, and increases engineering overhead, raising barriers to adoption for new researchers. To address these challenges, we introduce RL4CO, a unified and extensive benchmark with in-depth library coverage of 27 CO problem environments and 23 state-of-the-art baselines. Built on efficient software libraries and best practices in …
