Open Access. Powered by Scholars. Published by Universities.®
Artificial Intelligence and Robotics Commons™
Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Engineering (5396)
- Computer Engineering (4371)
- Operations Research, Systems Engineering and Industrial Engineering (4244)
- Numerical Analysis and Scientific Computing (4163)
- Systems Science (3895)
-
- Social and Behavioral Sciences (985)
- Databases and Information Systems (629)
- Medicine and Health Sciences (606)
- Data Science (528)
- Theory and Algorithms (487)
- Business (447)
- Graphics and Human Computer Interfaces (414)
- Electrical and Computer Engineering (403)
- Education (383)
- Software Engineering (374)
- Arts and Humanities (338)
- Public Affairs, Public Policy and Public Administration (298)
- Information Security (293)
- Other Computer Sciences (275)
- Life Sciences (270)
- Law (241)
- Statistics and Probability (204)
- Medical Specialties (180)
- Library and Information Science (165)
- Psychology (157)
- Robotics (157)
- Programming Languages and Compilers (155)
- Institution
-
- China Simulation Federation (3880)
- Singapore Management University (1897)
- Old Dominion University (644)
- San Jose State University (277)
- MBZUAI (233)
-
- City University of New York (CUNY) (184)
- Technological University Dublin (157)
- Air Force Institute of Technology (137)
- Chapman University (125)
- California Polytechnic State University, San Luis Obispo (116)
- Chinese Academy of Sciences (113)
- University of Arkansas, Fayetteville (103)
- Lindenwood University (97)
- Edith Cowan University (92)
- Embry-Riddle Aeronautical University (92)
- University of Nebraska - Lincoln (78)
- University of Kentucky (76)
- University of South Florida (71)
- Clemson University (63)
- University of Nevada, Las Vegas (63)
- Dartmouth College (62)
- University of Denver (59)
- University of Michigan Law School (57)
- Utah State University (57)
- The Texas Medical Center Library (54)
- Thomas Jefferson University (54)
- New Jersey Institute of Technology (53)
- University of Malaya (50)
- Purdue University (48)
- Missouri University of Science and Technology (47)
- Keyword
-
- Artificial intelligence (780)
- Machine learning (685)
- Deep learning (437)
- Artificial Intelligence (359)
- Machine Learning (359)
-
- AI (240)
- Deep Learning (202)
- Simulation (160)
- Computer vision (158)
- Reinforcement learning (140)
- Generative AI (135)
- Neural networks (129)
- Large language models (109)
- Natural language processing (108)
- Robotics (97)
- Natural Language Processing (91)
- ChatGPT (89)
- Path planning (89)
- Optimization (82)
- Large Language Models (78)
- Computer Vision (76)
- Classification (71)
- Neural network (67)
- Neural Networks (65)
- Virtual reality (64)
- Reinforcement Learning (63)
- Computer Science (59)
- Cybersecurity (59)
- Genetic algorithm (58)
- Algorithms (57)
- Publication Year
- Publication
-
- Journal of System Simulation (3880)
- Research Collection School Of Computing and Information Systems (1664)
- Master's Projects (248)
- Theses and Dissertations (183)
- Computer Science Faculty Publications (126)
-
- Bulletin of Chinese Academy of Sciences (Chinese Version) (113)
- Faculty Scholarship (108)
- Publications and Research (99)
- Computer Vision Faculty Publications (98)
- Master's Theses (96)
- Conference papers (92)
- Electrical & Computer Engineering Faculty Publications (90)
- Machine Learning Faculty Publications (86)
- Electronic Theses and Dissertations (85)
- Faculty Publications (77)
- Dissertations (70)
- Research outputs 2022 to 2026 (64)
- USF Tampa Graduate Theses and Dissertations (59)
- Dissertations and Theses Collection (Open Access) (57)
- Articles (54)
- Dissertations, Theses, and Capstone Projects (53)
- Theses and Dissertations--Computer Science (48)
- Natural Language Processing Faculty Publications (46)
- Teaching and Generative AI: Pedagogical Possibilities and Productive Tensions (46)
- Graduate Theses and Dissertations (45)
- Open Access Theses & Dissertations (42)
- Graduate Theses, Dissertations, and Problem Reports (ETD) (40)
- Theses (40)
- Electrical & Computer Engineering Theses & Dissertations (39)
- Publications (39)
- Publication Type
- File Type
Articles 1381 - 1410 of 11189
Full-Text Articles in Artificial Intelligence and Robotics
Advancing Fishery Dependent And Independent Habitat Assessments Using Automated Image Analysis: A Fisheries Management Agency Case Study, Scott Evans, Bronson Philippa, Carlo Mattone, Nick Konzewitsch, Renae Hovey, Marcus Sheaves, Gary A. Kendrick, Lynda M. Bellchambers
Advancing Fishery Dependent And Independent Habitat Assessments Using Automated Image Analysis: A Fisheries Management Agency Case Study, Scott Evans, Bronson Philippa, Carlo Mattone, Nick Konzewitsch, Renae Hovey, Marcus Sheaves, Gary A. Kendrick, Lynda M. Bellchambers
Fisheries Research Articles
Advances in artificial intelligence and machine learning have revolutionised data analysis, including in the field of marine and fisheries sciences. However, many fisheries agencies manage sensitive or proprietary data that cannot be shared externally, which can limit the adoption of externally hosted artificial intelligence platforms. In this study, we develop and evaluate two residual network-based automatic image annotation models to process fishery specific habitat data to support ecosystem-based fisheries management in the Exmouth Gulf Prawn Managed Fishery in Western Australia. Using an extensive dataset of 13,128 manually annotated benthic habitat images, we train a grid-based annotation model and an image-level …
Fact-Checker: A Web Application For Leveraging Large Language Models For Fact-Checking Youtube Videos, Andrew R. Craig
Fact-Checker: A Web Application For Leveraging Large Language Models For Fact-Checking Youtube Videos, Andrew R. Craig
Electronic Theses, Projects, and Dissertations
Fact-Checker is a web application that allows users to fact-check YouTube videos. It feeds YouTube’s closed captioning transcript to a large language model (LLM) to extract claims. It then uses multiple LLMs, such as Gemini, Llama, and Claude, to verify these claims. The modular design makes it easy to change to a different LLM or model if needed. The application is built using Python for access to Application Programming Interfaces (APIs) and Streamlit as the front-end framework. The utilization of Docker and Dockerfiles enables easy distribution and deployment. It enables the application to be deployed on almost any hardware platform …
Ai-Enhanced Structured Literacy Intervention For Secondary Students: A Case Study Of Science Of Reading, Jennifer Bird
Ai-Enhanced Structured Literacy Intervention For Secondary Students: A Case Study Of Science Of Reading, Jennifer Bird
Teaching & Learning Faculty Publications
This study examines the effectiveness of Lexia PowerUp, an AI-powered literacy program, for sixth-grade students requiring Tier 3 reading intervention. Seven sixth-grade students (six boys, one girl; five African American, two Caucasian; all qualifying for free/reduced lunch) participated in a six-month intervention combining 50 minutes of daily small-group instruction with individualized Lexia PowerUp usage. Researchers measured progress through Achieve 3000 Lexile assessments and Lexia PowerUp performance data across three skill strands: Word Study, Grammar, and Comprehension. All participants demonstrated Lexile level improvements from beginning-of-year to mid-year assessments, though students remained below sixth-grade benchmarks (925-1070L). Analysis of Lexia PowerUp progression showed …
Towards Securing Ai Systems: Investigating Threats In Multimodal Autonomous Driving & Rag Systems, Saket Sanjeev Chaturvedi
Towards Securing Ai Systems: Investigating Threats In Multimodal Autonomous Driving & Rag Systems, Saket Sanjeev Chaturvedi
All Dissertations
Artificial Intelligence (AI) systems have become central to high-stakes applications such as autonomous driving and language-based decision support. As their deployment accelerates, ensuring the security and trustworthiness of these systems becomes paramount. Among the most stealthy and potent threats are backdoor attacks, where models behave as expected under normal conditions but exhibit malicious behavior when triggered by specific inputs, either digital or physical.
This thesis investigates novel backdoor and adversarial vulnerabilities across two emerging classes of AI architectures: (1) multimodal 3D object detection systems that fuse LiDAR and camera data, and (2) Retrieval-Augmented Generation (RAG) systems that pair large language …
Wildfires Classification In Canadian Boreal Forest: A Comparative Study Of Logistic Regression And Xgboost Models, Brandon Tran, Elijah James Duran, Mike Luu, Hesham Morgan, Surendra Maharjan, Wenzhao Li, Hesham El-Askary
Wildfires Classification In Canadian Boreal Forest: A Comparative Study Of Logistic Regression And Xgboost Models, Brandon Tran, Elijah James Duran, Mike Luu, Hesham Morgan, Surendra Maharjan, Wenzhao Li, Hesham El-Askary
Mathematics, Physics, and Computer Science Faculty Articles and Research
In recent years, Canada has faced a growing number of wildfires. These events have devastated ecosystems, displaced communities, and posed severe health risks. To minimize the damage caused by such disasters, this study aims to develop an early warning system that predicts wildfire occurrences. Two machine learning models for binary classification of wildfire occurrence in Canadian wild forests, Logistic regression and XGBoost, will be compared and evaluated. The models are used to predict the likelihood of wildfire events based on various environmental and climatic factors. The models are evaluated using a 70-30 split validation approach and their performance is assessed …
Characterization Of Search Spaces And Effects On Machine Learning, Leo Ghelarducci
Characterization Of Search Spaces And Effects On Machine Learning, Leo Ghelarducci
Doctoral Dissertations and Master's Theses
The present status of the field of Machine Learning (ML) focuses on optimization of popular models. Rarely are the effects of the problem characteristics upon the solution algorithm studied. There exists no standard for knowing when to apply ML algorithms to a given problem or how to estimate the effectiveness of results. Focusing on the search space of problems, a rigorous study was conducted to generate an in-depth understanding of the impact of search space characteristics to the performance of a ML algorithm, specifically a Genetic Algorithm (GA). The effects of specific problem characteristics, represented via solution space characteristics, on …
Faced With Genai, Educators’ Engagement Capacity Matters More Than Ever, Thomas Menkhoff
Faced With Genai, Educators’ Engagement Capacity Matters More Than Ever, Thomas Menkhoff
Research Collection Lee Kong Chian School Of Business
In a commentary, SMU Professor of Organisational Behaviour & Human Resources (Education) Thomas Menkhoff stressed the need for educators to upskill so they can guide students in using generative artificial intelligence (GenAI) responsibly, rather than dismissing it. He argued that universities should move beyond prohibition and invest in AI literacy to safeguard academic integrity. Prof Menkhoff mentioned that combining the use of GenAI tools with effective prompting and Socratic questioning transforms students’ use of technology from passive consumption to active, reflective and critical engagement. To achieve this, he said that schools must set clear guidelines and design AI-compatible assessments that …
Intersecting Realities And Evolving Landscapes: Mapping Generative Ai Within The Framework Of Digital Rhetoric, Joshua Troy Nieubuurt
Intersecting Realities And Evolving Landscapes: Mapping Generative Ai Within The Framework Of Digital Rhetoric, Joshua Troy Nieubuurt
English Theses & Dissertations
The increased usage of [Generative] AI technologies (GenAI) in the 21st century has called into the question the rhetorical agency of these digital things. [Gen]AI has historically been framed within a Heideggerian “readiness-to-hand” dynamic in which it has been unilaterally conceived as a tool to be used by humans. This dissertation proposes that the GenAI assemblage is capable of being a co-actor in rhetorical spaces. To provide evidence for this stance This dissertation utilizes Actor Network Theory to map the actants within a GenAI assemblage. In doing so it allows for an understanding of the stakeholders (both human and non-human) …
Human Activity Recognition And Identification Driven Automated Deep Learning For Time-Series Classification, Justin Alan Gamble
Human Activity Recognition And Identification Driven Automated Deep Learning For Time-Series Classification, Justin Alan Gamble
Engineering Management & Systems Engineering Theses & Dissertations
The growing emphasis on Digital Engineering (DE) within the U.S. Department of Defense (DoD) demands advanced methods for leveraging vast time-series data generated by sensor-rich environments. Deep learning models offer promising solutions for complex timeseries classification tasks, however their design and optimization remain highly resource intensive, requiring specialized expertise. This dissertation addresses this challenge by developing and evaluating an Automated Machine Learning (AutoML) framework specifically tailored for the time-series classification task of Human Activity Recognition and Identification (HARI).
A systematic investigation was conducted using the Design Science Research Methodology (DSRM) comparing traditional search strategies of grid search and random search …
Unfolding Particle Detector Effects And Solving Qcd Inverse Problem With Generative Ai, Tareq Saeed Alghamdi
Unfolding Particle Detector Effects And Solving Qcd Inverse Problem With Generative Ai, Tareq Saeed Alghamdi
Computer Science Theses & Dissertations
Advancements in artificial intelligence (AI) have revolutionized high-energy physics by enabling generative models to address key detector-related Challenges. This work explores the generative model to mitigate smearing, acceptance, and inefficiency in particle detectors, enhancing experimental precision.
We present a generative model-based framework to model and correct detector distortions. Using the Jefferson Lab CLAS g11 experiment as a case study, our approach successfully unfolds detector effects in multi-particle final states while preserving multidimensional correlations despite complex reaction mechanisms. A key focus is addressing the acceptance problem—accurately modeling detector acceptance without computationally expensive simulations. By training generative model-based framework on simulated detector …
Input Structure Based Optimization For Privacy Preserving Ai Systems, Feng Yizhou
Input Structure Based Optimization For Privacy Preserving Ai Systems, Feng Yizhou
Electrical & Computer Engineering Theses & Dissertations
As Artificial Intelligence (AI) systems become increasingly integrated into critical domains, ensuring privacy-preserving model design and system deployment has become a pressing priority. Safeguarding both sensitive user data and proprietary model parameters is critical throughout the AI model and system, from data acquisition and pre-processing to model inference and deployment. However, existing privacy-preserving frameworks face several limitations, including fragmented data ownership, incomplete protection across system stages, substantial computational overhead, and poor scalability to modern architectures such as large language models. This dissertation explores a unifying optimization strategy centered on input structure design to address these challenges. The core idea is …
Service With A Smile Or Salesperson Mirroring? Understanding The Flow Of Emotional Contagion In Sales Encounters, Vinh Quoc Trong Luong
Service With A Smile Or Salesperson Mirroring? Understanding The Flow Of Emotional Contagion In Sales Encounters, Vinh Quoc Trong Luong
Theses and Dissertations in Business Administration
This study examines the directionality of emotional contagion in sales interactions, addressing a critical gap in understanding whether emotions flow primarily from the salesperson to the customer, from the customer to the salesperson, or bidirectionally. While prior research emphasizes customer-driven emotional flow or bidirectional alignment, this study challenges these assumptions by employing categorical Cross-Recurrence Quantification Analysis (CRQA) to assess temporal emotional synchronization in sales dialogues. Leveraging automated sentiment analysis and multi-agent AI evaluation for performance metrics, the research analyzes 166 sales interactions to quantify emotional influence dynamics. Results reveal that salespeople predominantly lead emotional exchanges, exhibiting stronger and more stable …
Unsupervised Deep Learning For Video Restoration, Mary Damilola Aiyetigbo
Unsupervised Deep Learning For Video Restoration, Mary Damilola Aiyetigbo
All Dissertations
In today's digital era, visual data is vital across several domains such as medical diagnostics, scientific imaging, surveillance, and entertainment. However, video data often suffers from degradations like noise, blur, compression artifacts, and low resolution, which degrade quality and downstream usability. Video restoration aims to recover clean, high-fidelity video from such corrupted inputs. Unlike static images, video restoration must maintain temporal consistency across frames, making it a significantly more complex problem. While supervised deep learning methods have achieved state-of-the-art results, they typically require large datasets of paired noisy-clean video datasets that are scarce or impractical to obtain in real-world settings …
Trust In Healthcare Ai Can’T Just Be Designed – It Must Be Felt By Clinicians And Patients, Adriana Banozic-Tang, Heng Wang
Trust In Healthcare Ai Can’T Just Be Designed – It Must Be Felt By Clinicians And Patients, Adriana Banozic-Tang, Heng Wang
Research Collection Yong Pung How School Of Law
Trust in healthcare AI currently over-relies on system design, not lived medical realities.Continuous feedback loops are necessary to embed trust in healthcare AI that is responsive to clinician and patient needs.Initiatives in South-East Asia show how trust in technology can be extended from policy to practice.
Ai In The Judiciary: The Singapore Case, Nydia Remolina Leon
Ai In The Judiciary: The Singapore Case, Nydia Remolina Leon
Research Collection Yong Pung How School Of Law
This paper examines the integration of Artificial Intelligence (AI) within the judicial system of Singapore. Singapore's judiciary has embraced AI not as a tool for adjudication, but as an augmentative instrument for legal research, procedural efficiency, and access to justice. It provides a detailed account of AI use cases in the courts, including case summarization, evidence review, assistance for selfrepresented litigants, and tools like the Divorce Assets Informative Division Estimator. The discussion then turns to the legal profession, exploring how law firms in Singapore are adopting AI technologies. The paper also addresses how AI implementation in the judicial system is …
Roadside Asset Extraction From Mobile Lidar Point Cloud, Yushin Ahn, Riadh Munjy, Stephen Choi
Roadside Asset Extraction From Mobile Lidar Point Cloud, Yushin Ahn, Riadh Munjy, Stephen Choi
Mineta Transportation Institute
Mobile LiDAR systems are powerful tools that help us map roads and their surroundings in 3D with great speed and precision. The data provided by these systems support urban planning efforts, digital mapping, transportation infrastructure maintenance, and more. This report presents a comprehensive workflow for roadside asset extraction using Mobile Terrestrial Laser Scanning (MTLS) data, focusing on road lane detection, cross-section slope analysis, and point cloud classification. Roadside asset extraction is the identification and classification of roadside features like signs and poles. The dataset, acquired using a high-resolution mobile LiDAR system, contains over 5.7 billion points (pieces of data) across …
How To Do Things With Little Talking Tubes: Nonideal Speech Acts In The Digital Age, Anthony Holdier
How To Do Things With Little Talking Tubes: Nonideal Speech Acts In The Digital Age, Anthony Holdier
Graduate Theses and Dissertations
In this work, I develop a view about what it means to share a social or conversational context with others, as well as what is normatively entailed by doing so. Working from a broadly Austinian perspective about the moral foundations of language use and how we use words to position ourselves within social space, as well as from a generally Stalnakerian social ontology (demarcating groups by dint of aligned or overlapping sets of commitments), I present three papers demonstrating how nonidealized, ordinary language philosophy can make sense of complex, real-world phenomena. Through analyses of heretical utterances, context collapse, and chatbot …
Use Of A Preliminary Artificial Intelligence-Based Laryngeal Cancer Screening Framework For Low-Resource Settings: Development And Validation Study, Shao Wei Sean Lam, Min Hun Lee, Michael Dorosan, Samuel Altonji, Hiang Khoon Tan, Walter T. Lee
Use Of A Preliminary Artificial Intelligence-Based Laryngeal Cancer Screening Framework For Low-Resource Settings: Development And Validation Study, Shao Wei Sean Lam, Min Hun Lee, Michael Dorosan, Samuel Altonji, Hiang Khoon Tan, Walter T. Lee
Research Collection School Of Computing and Information Systems
Background: Early-stage diagnosis of laryngeal cancer significantly improves patient survival and quality of life. However, the scarcity of specialists in low-resource settings hinders the timely review of flexible nasopharyngoscopy (FNS) videos, which are essential for accurate triage of at-risk patients.Objective: We introduce a preliminary AI-based screening framework to address this challenge for the triaging of at-risk patients in low-resource settings. This formative research addresses multiple challenges common in high-dimensional FNS videos: (1) selecting clear, informative images; (2) deriving regions within frames that show an anatomical landmark of interest; and (3) classifying patients into referral grades based on the FNS video …
Ai-Assisted Triage And Decision Support Of Head And Neck Cancer Screening And Diagnosis In Low-Resourced Settings, Min Hun Lee, Sean Shao Wei Lam, Shaun Xin Hong Liew, Michael Dorosan, Nicholas Graves, Jonas Karlström, Hiang Khoon Tan, Walter Tsong Lee
Ai-Assisted Triage And Decision Support Of Head And Neck Cancer Screening And Diagnosis In Low-Resourced Settings, Min Hun Lee, Sean Shao Wei Lam, Shaun Xin Hong Liew, Michael Dorosan, Nicholas Graves, Jonas Karlström, Hiang Khoon Tan, Walter Tsong Lee
Research Collection School Of Computing and Information Systems
The mortality burden of head and neck cancer (HNC) is increasing globally and disproportionately affects people in low-and middle-income countries with limited medical workforce. To address this issue, artificial intelligence (AI) algorithms are increasingly being explored to process medical imaging data, demonstrating competitive performance. However, the clinical adoption of AI remains challenging as clinicians struggle to understand how complex AI works and trust it to use in practice. In addition, AI may not perform well on varying data qualities of endoscopy videos for HNC screening and diagnosis from multiple sites.In this project, our international and interdisciplinary team will collaborate with …
Faithfulrag: Fact-Level Conflict Modeling For Context-Faithful Retrieval-Augmented Generation, Qinggang Zhang, Zhishang Xiang, Yilin Xiao, Le Wang, Junhui Li, Xinrun Wang, Jinsong Su
Faithfulrag: Fact-Level Conflict Modeling For Context-Faithful Retrieval-Augmented Generation, Qinggang Zhang, Zhishang Xiang, Yilin Xiao, Le Wang, Junhui Li, Xinrun Wang, Jinsong Su
Research Collection School Of Computing and Information Systems
Large language models (LLMs) augmented with retrieval systems have demonstrated significant potential in handling knowledge-intensive tasks. However, these models often struggle with unfaithfulness issues, generating outputs that either ignore the retrieved context or inconsistently blend it with the LLM’s parametric knowledge. This issue is particularly severe in cases of knowledge conflict, where the retrieved context conflicts with the model’s parametric knowledge. While existing faithful RAG approaches enforce strict context adherence through well-designed prompts or modified decoding strategies, our analysis reveals a critical limitation: they achieve faithfulness by forcibly suppressing the model’s parametric knowledge, which undermines the model’s internal knowledge structure …
Colloquial Singaporean English Style Transfer With Fine-Grained Explainable Control, Jinggui Liang, Dung Vo, Yap Hong Xian, Hai Leong Chieu, Kian Ming A. Chai, Jing Jiang, Lizi Liao
Colloquial Singaporean English Style Transfer With Fine-Grained Explainable Control, Jinggui Liang, Dung Vo, Yap Hong Xian, Hai Leong Chieu, Kian Ming A. Chai, Jing Jiang, Lizi Liao
Research Collection School Of Computing and Information Systems
Colloquial Singaporean English (Singlish) is an informal English marked by a unique blend of languages reflecting Singapore’s multicultural identity. Style transfer between Singlish and Standard (formal) English is vital for various applications, yet existing methods often lack explainability and fine-grained control. To fill this gap, we contribute in two key ways. First, we construct a large, high-quality dataset of formal and informal sentences, annotated across six linguistic aspects—Syntax, Lexical Borrowing, Pragmatics, Prosody/Phonology, Emoticons/Punctuation, and Code-Switching—with detailed explanations. Starting with manually annotated cases, we scaled the dataset to 140K with ensured quality. Second, inspired by the “Society of Mind” theory, we …
Debate, Reflect, And Distill: Multi-Agent Feedback With Tree-Structured Preference Optimization For Efficient Language Model Enhancement, Xiaofeng Zhou, Heyan Huang, Lizi Liao
Debate, Reflect, And Distill: Multi-Agent Feedback With Tree-Structured Preference Optimization For Efficient Language Model Enhancement, Xiaofeng Zhou, Heyan Huang, Lizi Liao
Research Collection School Of Computing and Information Systems
Large Language Models (LLMs) continue to set new standards in knowledge-intensive and complex reasoning tasks, yet their high computational demands limit widespread adoption. While distilling large models into smaller ones offers a sustainable solution, current techniques—such as static knowledge distillation, resource-intensive reinforcement learning from human feedback, or limited self-reflection—struggle to yield substantial and lasting performance gains. In this paper, we present a novel Debate and Reflect (D&R) framework that orchestrates multi-turn debates between smaller models and stronger teacher models, eliciting actionable feedback (e.g., error analysis, corrective strategies) to guide student models. Further, we introduce Tree-structured Direct Preference Optimization (T-DPO) to …
Memotune: A Measure And Moment-Driven Fine-Tuning Framework For Quantized Large Language Models, Yun Zhang, Xue Geng, Lizi Liao, Jintong Sun, Minghe Yu, Ge Yu
Memotune: A Measure And Moment-Driven Fine-Tuning Framework For Quantized Large Language Models, Yun Zhang, Xue Geng, Lizi Liao, Jintong Sun, Minghe Yu, Ge Yu
Research Collection School Of Computing and Information Systems
Quantizing large language models (LLMs) is essential for reducing memory and computational costs in natural language processing. Existing methods combine quantization with parameter-efficient fine-tuning but often fail to meet practical performance requirements. This paper introduces MeMoTune, a novel fine-tuning framework for quantized LLMs. By employing a measure and moment approach within a low-rank approximation framework in probability measure space, MeMoTune optimizes the objective function for superior fine-tuning results. The update process is further refined through scaled gradient, enhancing convergence efficiency and noise robustness. Experiments on tasks like text generation, summarization, and understanding show MeMoTune significantly outperforms state-of-the-art methods, e.g. fine-tuning …
R2dqg: A Quality Meets Diversity Framework For Question Generation Over Knowledge Bases, Yimeng Ren, Yanhua Yu, Lizi Liao, Yuhu Shang, Kangkang Lu, Mingliang Yan
R2dqg: A Quality Meets Diversity Framework For Question Generation Over Knowledge Bases, Yimeng Ren, Yanhua Yu, Lizi Liao, Yuhu Shang, Kangkang Lu, Mingliang Yan
Research Collection School Of Computing and Information Systems
The task of Knowledge-Based Question Generation (KBQG) involves generating natural language questions from structured knowledge sources, posing unique challenges in balancing linguistic diversity and semantic relevance. Existing models often focus on maximizing surface-level similarity to ground-truth questions, neglecting the need for diverse syntactic forms and leading to semantic drift during generation. To overcome these challenges, we propose Refine-Reinforced Diverse Question Generation (R2DQG), a two-phase framework leveraging a generation-then-refinement paradigm. The Generator first constructs a diverse set of expressive templates using dependency parse tree similarity, capturing a wide range of syntactic patterns and styles. These templates guide the creation of question …
Consistent Client Simulation For Motivational Interviewing-Based Counseling, Yizhe Yang, Palakorn Achananuparp, Heyan Huang, Jing Jiang, Nicholas Gabriel Lim, Cameron Shi Ern Tan, Phey Ling Kit, Jenny Xiuhui Giam, John Pinto, Ee-Peng Lim
Consistent Client Simulation For Motivational Interviewing-Based Counseling, Yizhe Yang, Palakorn Achananuparp, Heyan Huang, Jing Jiang, Nicholas Gabriel Lim, Cameron Shi Ern Tan, Phey Ling Kit, Jenny Xiuhui Giam, John Pinto, Ee-Peng Lim
Research Collection School Of Computing and Information Systems
Simulating human clients in mental health counseling is crucial for training and evaluating counselors (both human or simulated) in a scalable manner. Nevertheless, past research on client simulation did not focus on complex conversation tasks such as mental health counseling. In these tasks, the challenge is to ensure that the client’s actions (i.e., interactions with the counselor) are consistent with with its stipulated profiles and negative behavior settings. In this paper, we propose a novel framework that supports consistent client simulation for mental health counseling. Our framework tracks the mental state of a simulated client, controls its state transitions, and …
Reimagining Education With Ai, Margherita Pagani, Steven M. Miller, Jerry Wind
Reimagining Education With Ai, Margherita Pagani, Steven M. Miller, Jerry Wind
Research Collection School Of Computing and Information Systems
This chapter examines AI’s transformative potential in education, focusing on Generative AI (GenAI) and Large Language Models (LLMs) while at the same time emphasizing the importance of grounding and guiding AI efforts with learning science and education research findings. It synthesizes analyses and expert recommendations, highlighting opportunities like personalized learning and enhanced teacher productivity, alongside challenges such as over-reliance on AI. Practical steps for instructors include adopting a question-first approach, utilizing AI for personalized feedback, designing AI-enhanced learning experiences, fostering critical thinking, and ensuring ethical AI use. The chapter concludes with strategic recommendations for leveraging AI to sustainably improve educational …
Dreamanime: Learning Style-Identity Textual Disentanglement For Anime And Beyond, Chenshu Xu, Yangyang Xu, Huaidong Zhang, Xuemiao Xu, Shengfeng He
Dreamanime: Learning Style-Identity Textual Disentanglement For Anime And Beyond, Chenshu Xu, Yangyang Xu, Huaidong Zhang, Xuemiao Xu, Shengfeng He
Research Collection School Of Computing and Information Systems
Text-to-image generation models have significantly broadened the horizons of creative expression through the power of natural language. However, navigating these models to generate unique concepts, alter their appearance, or reimagine them in unfamiliar roles presents an intricate challenge. For instance, how can we exploit language-guided models to transpose an anime character into a different art style, or envision a beloved character in a radically different setting or role? This paper unveils a novel approach named DreamAnime, designed to provide this level of creative freedom. Using a minimal set of 2-3 images of a user-specified concept such as an anime character …
L3net: Localized And Layered Reparameterization For Incremental Learning, Xuandi Luo, Huaidong Zhang, Yi Xie, Hongrui Zhang, Xuemiao Xu, Shengfeng He
L3net: Localized And Layered Reparameterization For Incremental Learning, Xuandi Luo, Huaidong Zhang, Yi Xie, Hongrui Zhang, Xuemiao Xu, Shengfeng He
Research Collection School Of Computing and Information Systems
Model-based class incremental learning (CIL) methods aim to address the challenge of catastrophic forgetting by retaining certain parameters and expanding the model architecture. However, retaining too many parameters can lead to an overly complex model, increasing inference overhead. Additionally, compressing these parameters to reduce the model size can result in performance degradation. To tackle these challenges, we propose a novel three-stage CIL framework called Localized and Layered Reparameterization for Incremental Learning (L3Net). The rationale behind our approach is to balance model complexity and performance by selectively expanding and optimizing critical components. Specifically, the framework introduces a Localized Dual-path Expansion structure, …
Solving Two-Stage Stochastic Integer Programs Via Representation Learning, Yaoxin Wu, Zhiguang Cao, Wen Song, Yingqian Zhang
Solving Two-Stage Stochastic Integer Programs Via Representation Learning, Yaoxin Wu, Zhiguang Cao, Wen Song, Yingqian Zhang
Research Collection School Of Computing and Information Systems
Solving stochastic integer programs (SIPs) is extremely intractable due to the high computational complexity. To solve two-stage SIPs efficiently, we propose a conditional variational autoencoder (CVAE) for scenario representation learning. A graph convolutional network (GCN) based VAE embeds scenarios into a low-dimensional latent space, conditioned on the deterministic context of each instance. With the latent representations of stochastic scenarios, we perform two auxiliary tasks: objective prediction and scenario contrast, which predict scenario objective values and the similarities between them, respectively. These tasks further integrate objective information into the representations through gradient backpropagation. Experiments show that the learned scenario representations can …
How To Enable Effective Cooperation Between Humans And Nlp Models: A Survey Of Principles, Formalizations, And Beyond, Chen Huang, Yang Deng, Wenqiang Lei, Jiancheng Lv, Tat-Seng Chua, Jimmy Huang
How To Enable Effective Cooperation Between Humans And Nlp Models: A Survey Of Principles, Formalizations, And Beyond, Chen Huang, Yang Deng, Wenqiang Lei, Jiancheng Lv, Tat-Seng Chua, Jimmy Huang
Research Collection School Of Computing and Information Systems
With the advancement of large language models (LLMs), intelligent models have evolved from mere tools to autonomous agents with their own goals and strategies for cooperating with humans. This evolution has birthed a novel paradigm in NLP, i.e., human-model cooperation, that has yielded remarkable progress in numerous NLP tasks in recent years. In this paper, we take the first step to present a thorough review of human-model cooperation, exploring its principles, formalizations, and open challenges. In particular, we introduce a new taxonomy that provides a unified perspective to summarize existing approaches. Also, we discuss potential frontier areas and their corresponding …