Advancing Telehealth Through Artificial Intelligence: Incorporating Emotional Intelligence And Addressing Cybersecurity Challenges,
2024
California State University - San Bernardino
Advancing Telehealth Through Artificial Intelligence: Incorporating Emotional Intelligence And Addressing Cybersecurity Challenges, Mahima Rajendra Pulgaonkar
Electronic Theses, Projects, and Dissertations
This culminating experience project explores the integration of Emotional Artificial Intelligence (Emotional AI) into telehealth systems, addressing the dual challenges of enhancing patient care and mitigating cybersecurity risks. The research questions are: (Q1) How can Emotionally Intelligent AI improve telehealth systems' ability to recognize and respond to mental health symptoms? and (Q2) What are the specific cybersecurity challenges associated with AI in telehealth and how can they be mitigated? The findings for each question are: Q1: Emotionally Intelligent AI can significantly enhance telehealth by providing personalized, empathetic interactions that improve patient engagement, adherence to treatment plans, and early detection of …
Querymate: A Custom Llm Powered By Llamacpp,
2024
CUNY New York City College of Technology
Querymate: A Custom Llm Powered By Llamacpp, Pegah Khosravi
Open Educational Resources
No abstract provided.
Establishing The Importance Of Co-Creation And Self-Efficacy In Creative Collaboration With Artificial Intelligence,
2024
Singapore Management University
Establishing The Importance Of Co-Creation And Self-Efficacy In Creative Collaboration With Artificial Intelligence, Jack Mcguire, David De Cremer, Tim Van De Cruys
Research Collection Lee Kong Chian School Of Business
The emergence of generative AI technologies has led to an increasing number of people collaborating with AI to produce creative works. Across two experimental studies, in which we carefully designed and programmed state-of-the-art human–AI interfaces, we examine how the design of generative AI systems influences human creativity (poetry writing). First, we find that people were most creative when writing a poem on their own, compared to first receiving a poem generated by an AI system and using sophisticated tools to edit it (Study 1). Following this, we demonstrate that this creativity deficit dissipates when people co-create with—not edit—AI and establish …
We Train Ai, Why Not Humans, Too? An Exploration Of Human-Ai Team Training For Future Workplace Viability,
2024
Clemson University
We Train Ai, Why Not Humans, Too? An Exploration Of Human-Ai Team Training For Future Workplace Viability, Caitlin M. Lancaster
All Dissertations
The integration of Artificial Intelligence (AI) in the workforce is transforming team dynamics, leading to the emergence of Human-AI Teams (HATs). These teams offer opportunities to capitalize on human strengths with AI's prowess, offering significant opportunities for innovation and efficiency. Effective HAT functioning requires aligning human expectations with AI capabilities and bridging knowledge gaps between teammates. Despite this potential, key integration challenges remain, such as developing shared mental models, addressing skill limitations, and overcoming negative AI perceptions. Existing training efforts often apply human-human teaming principles directly to HATs, overlooking AI's role as a teammate and limiting the development of HAT-specific …
Ee-Lce: An Event Extraction Framework Based On Llm-Generated Cot Explanation,
2024
Singapore Management University
Ee-Lce: An Event Extraction Framework Based On Llm-Generated Cot Explanation, Yanhua Yu, Yuanlong Wang, Yunshan Ma, Jie Li, Kangkang Lu, Zhiyong Huang, Tat-Seng Chua
Research Collection School Of Computing and Information Systems
Generative models have been widely used in event extraction. However, the interpretability of event extraction has not been fully investigated. In this paper, we propose an Event Extraction framework based on LLM-generated CoT Explanation EE-LCE, which can generate chain-of-thought-style (CoT-style) explanations for events. To this end, we provide each sample of event datasets with an explanation of the reasoning process using a large language model (LLM) GPT-3.5, and fine-tune the Flan-T5 lightweight language model (LM) supervised by the augmented dataset, enhancing both interpretability and performance of the event extraction. Moreover, we use a prefix tree (trie) to normalize the decoding …
Tackling Stackelberg Network Interdiction Against A Boundedly Rational Adversary,
2024
Singapore Management University
Tackling Stackelberg Network Interdiction Against A Boundedly Rational Adversary, Tien Mai, Avinandan Bose, Arunesh Sinha, Thanh Nguyen, Ayushman Kumar Singh
Research Collection School Of Computing and Information Systems
This work studies Stackelberg network interdiction games --- an important class of games in which a defender first allocates (randomized) defense resources to a set of critical nodes on a graph while an adversary chooses its path to attack these nodes accordingly. We consider a boundedly rational adversary in which the adversary's response model is based on a dynamic form of classic logit-based (quantal response) discrete choice models. The resulting optimization is non-convex and additionally, involves complex terms that sum over exponentially many paths. We tackle these computational challenges by presenting new efficient algorithms with solution guarantees. First, we present …
Analyzing Temporal Complex Events With Large Language Models? A Benchmark Towards Temporal, Long Context Understanding,
2024
Singapore Management University
Analyzing Temporal Complex Events With Large Language Models? A Benchmark Towards Temporal, Long Context Understanding, Zhihan Zhang, Yixin Cao, Chenchen Ye, Ma. Yunshan, Lizi Liao, Tat-Seng Chua
Research Collection School Of Computing and Information Systems
The digital landscape is rapidly evolving with an ever-increasing volume of online news, emphasizing the need for swift and precise analysis of complex events.We refer to the complex events composed of many news articles over an extended period as Temporal Complex Event (TCE). This paper proposes a novel approach using Large Language Models (LLMs) to systematically extract and analyze the event chain within TCE, characterized by their key points and timestamps. We establish a benchmark, named TCELongBench, to evaluate the proficiency of LLMs in handling temporal dynamics and understanding extensive text. This benchmark encompasses three distinct tasks - reading comprehension, …
A Survey On Neural Question Generation : Methods, Applications, And Prospects,
2024
Singapore Management University
A Survey On Neural Question Generation : Methods, Applications, And Prospects, Shasha Guo, Lizi Liao, Cuiping Li, Tat-Seng Chua
Research Collection School Of Computing and Information Systems
In this survey, we present a detailed examination of the advancements in Neural Question Generation (NQG), a field leveraging neural network techniques to generate relevant questions from diverse inputs like knowledge bases, texts, and images. The survey begins with an overview of NQG’s background, encompassing the task’s problem formulation, prevalent benchmark datasets, established evaluation metrics, and notable applications. It then methodically classifies NQG approaches into three predominant categories: structured NQG, which utilizes organized data sources, unstructured NQG, focusing on more loosely structured inputs like texts or visual content, and hybrid NQG, drawing on diverse input modalities. This classification is followed …
Empathyear : An Open-Source Avatar Multimodal Empathetic Chatbot,
2024
Singapore Management University
Empathyear : An Open-Source Avatar Multimodal Empathetic Chatbot, Hao Fei, Han Zhang, Bin Wang, Lizi Liao, Qian Liu, Erik Cambria
Research Collection School Of Computing and Information Systems
This paper introduces EmpathyEar, a pioneering open-source, avatar-based multimodal empathetic chatbot, to fill the gap in traditional text-only empathetic response generation (ERG) systems. Leveraging the advancements of a large language model, combined with multimodal encoders and generators, EmpathyEar supports user inputs in any combination of text, sound, and vision, and produces multimodal empathetic responses, offering users, not just textual responses but also digital avatars with talking faces and synchronized speeches. A series of emotion-aware instruction-tuning is performed for comprehensive emotional understanding and generation capabilities. In this way, EmpathyEar provides users with responses that achieve a deeper emotional resonance, closely emulating …
Planning Like Human : A Dual-Process Framework For Dialogue Planning,
2024
Singapore Management University
Planning Like Human : A Dual-Process Framework For Dialogue Planning, Tao He, Lizi Liao, Yixin Cao, Yuanxing Liu, Ming Liu, Zerui Chen, Bing Qin
Research Collection School Of Computing and Information Systems
In proactive dialogue, the challenge lies not just in generating responses but in steering conversations toward predetermined goals, a task where Large Language Models (LLMs) typically struggle due to their reactive nature. Traditional approaches to enhance dialogue planning in LLMs, ranging from elaborate prompt engineering to the integration of policy networks, either face efficiency issues or deliver suboptimal performance. Inspired by the dual-process theory in psychology, which identifies two distinct modes of thinking—intuitive (fast) and analytical (slow), we propose the Dual-Process Dialogue Planning (DPDP) framework. DPDP embodies this theory through two complementary planning systems: an instinctive policy model for familiar …
Synergizing Large Language Models And Pre-Trained Smaller Models For Conversational Intent Discovery,
2024
Singapore Management University
Synergizing Large Language Models And Pre-Trained Smaller Models For Conversational Intent Discovery, Jinggui Liang, Lizi Liao, Hao Fei, Jing Jiang
Research Collection School Of Computing and Information Systems
In Conversational Intent Discovery (CID), Small Language Models (SLMs) struggle with overfitting to familiar intents and fail to label newly discovered ones. This issue stems from their limited grasp of semantic nuances and their intrinsically discriminative framework. Therefore, we propose Synergizing Large Language Models (LLMs) with pre-trained SLMs for CID (SynCID). It harnesses the profound semantic comprehension of LLMs alongside the operational agility of SLMs. By utilizing LLMs to refine both utterances and existing intent labels, SynCID significantly enhances the semantic depth, subsequently realigning these enriched descriptors within the SLMs’ feature space to correct cluster distortion and promote robust learning …
Unifying Global-Local Representations In Salient Object Detection With Transformers,
2024
Singapore Management University
Unifying Global-Local Representations In Salient Object Detection With Transformers, Sucheng Ren, Nanxuan Zhao, Qiang Wen, Guoqiang Han, Shengfeng He
Research Collection School Of Computing and Information Systems
The fully convolutional network (FCN) has dominated salient object detection for a long period. However, the locality of CNN requires the model deep enough to have a global receptive field and such a deep model always leads to the loss of local details. In this paper, we introduce a new attention-based encoder, vision transformer, into salient object detection to ensure the globalization of the representations from shallow to deep layers. With the global view in very shallow layers, the transformer encoder preserves more local representations to recover the spatial details in final saliency maps. Besides, as each layer can capture …
Speaker Verification In Agent-Generated Conversations,
2024
Singapore Management University
Speaker Verification In Agent-Generated Conversations, Yizhe Yang, Palakorn Achananuparp, Heyan Huang, Jing Jiang, Ee-Peng Lim
Research Collection School Of Computing and Information Systems
The recent success of large language models (LLMs) has attracted widespread interest to develop role-playing conversational agents personalized to the characteristics and styles of different speakers to enhance their abilities to perform both general and special purpose dialogue tasks. However, the ability to personalize the generated utterances to speakers, whether conducted by human or LLM, has not been well studied. To bridge this gap, our study introduces a novel evaluation challenge: speaker verification in agent-generated conversations, which aimed to verify whether two sets of utterances originate from the same speaker. To this end, we assemble a large dataset collection encompassing …
Self-Adaptive Psro : Towards An Automatic Population-Based Game Solver,
2024
Singapore Management University
Self-Adaptive Psro : Towards An Automatic Population-Based Game Solver, Pengdeng Li, Shuxin Li, Chang Yang, Xinrun Wang, Xiao Huang, Hau Chan, Bo An
Research Collection School Of Computing and Information Systems
Policy-Space Response Oracles (PSRO) as a general algorithmic framework has achieved state-of-the-art performance in learning equilibrium policies of two-player zero-sum games. However, the hand-crafted hyperparameter value selection in most of the existing works requires extensive domain knowledge, forming the main barrier to applying PSRO to different games. In this work, we make the first attempt to investigate the possibility of self-adaptively determining the optimal hyperparameter values in the PSRO framework. Our contributions are three-fold: (1) Using several hyperparameters, we propose a parametric PSRO that unifies the gradient descent ascent (GDA) and different PSRO variants. (2) We propose the self-adaptive PSRO …
A Multimodal Foundation Agent For Financial Trading : Tool-Augmented, Diversified, And Generalist,
2024
Singapore Management University
A Multimodal Foundation Agent For Financial Trading : Tool-Augmented, Diversified, And Generalist, Wentao Zhang, Lingxuan Zhao, Haochong Xia, Shuo Sun, Jiaze Sun, Molei Qin, Xinyi Li, Yuqing Zhao, Yilei Zhao, Xinyu Cai, Longtao Zheng, Xinrun Wang, Bo An
Research Collection School Of Computing and Information Systems
Financial trading is a crucial component of the markets, informed by a multimodal information landscape encompassing news, prices, and Kline charts, and encompasses diverse tasks such as quantitative trading and high-frequency trading with various assets. While advanced AI techniques like deep learning and reinforcement learning are extensively utilized in finance, their application in financial trading tasks often faces challenges due to inadequate handling of multimodal data and limited generalizability across various tasks. To address these challenges, we present FinAgent, a multimodal foundational agent with tool augmentation for financial trading. FinAgent's market intelligence module processes a diverse range of data-numerical, textual, …
Macrohft : Memory Augmented Context-Aware Reinforcement Learning On High Frequency Trading,
2024
Singapore Management University
Macrohft : Memory Augmented Context-Aware Reinforcement Learning On High Frequency Trading, Chuqiao Zong, Chaojie Wang, Molei Qin, Lei Feng, Xinrun Wang, Xinrun Wang
Research Collection School Of Computing and Information Systems
High-frequency trading (HFT) that executes algorithmic trading in short time scales, has recently occupied the majority of cryptocurrency market. Besides traditional quantitative trading methods, reinforcement learning (RL) has become another appealing approach for HFT due to its terrific ability of handling high-dimensional financial data and solving sophisticated sequential decision-making problems, e.g., hierarchical reinforcement learning (HRL) has shown its promising performance on second-level HFT by training a router to select only one sub-agent from the agent pool to execute the current transaction. However, existing RL methods for HFT still have some defects: 1) standard RL-based trading agents suffer from the overfitting …
Topic Modeling On Document Networks With Dirichlet Optimal Transport Barycenter,
2024
Singapore Management University
Topic Modeling On Document Networks With Dirichlet Optimal Transport Barycenter, Ce Zhang, Hady Wirawan Lauw
Research Collection School Of Computing and Information Systems
Text documents are often interconnected in a network structure, e.g., academic papers via citations, Web pages via hyperlinks. On the one hand, though Graph Neural Networks (GNNs) have shown promising ability to derive effective embeddings for such networked documents, they do not assume a latent topic structure and result in uninterpretable embeddings. On the other hand, topic models can infer semantically interpretable topic distributions for documents by associating each topic with a group of understandable key words. However, most topic models mainly focus on plain text within documents and fail to leverage network structure across documents. Network connectivity reveals topic …
Reinforcement Nash Equilibrium Solver,
2024
Singapore Management University
Reinforcement Nash Equilibrium Solver, Xinrun Wang, Chang Yang, Shuxin Li, Pengdeng Li, Xiao Huang, Hau Chan, Bo An
Research Collection School Of Computing and Information Systems
Nash Equilibrium (NE) is the canonical solution concept of game theory, which provides an elegant tool to understand the rationalities. Though mixed strategy NE exists in any game with finite players and actions, computing NE in two- or multi-player general-sum games is PPAD-Complete. Various alternative solutions, e.g., Correlated Equilibrium (CE), and learning methods, e.g., fictitious play (FP), are proposed to approximate NE. For convenience, we call these methods as "inexact solvers", or "solvers" for short. However, the alternative solutions differ from NE and the learning methods generally fail to converge to NE. Therefore, in this work, we propose REinforcement Nash …
Exploring A Multimodal Fusion-Based Deep Learning Network For Detecting Facial Palsy,
2024
Singapore Management University
Exploring A Multimodal Fusion-Based Deep Learning Network For Detecting Facial Palsy, Heng Yim Nicole Oo, Min Hun Lee, J. H. Lim
Research Collection School Of Computing and Information Systems
Algorithmic detection of facial palsy offers the potential to improve current practices, which usually involve labor-intensive and subjective assessment by clinicians. In this paper, we present a multimodal fusion-based deep learning model that utilizes unstructured data (i.e. an image frame with facial line segments) and structured data (i.e. features of facial expressions) to detect facial palsy. We then contribute to a study to analyze the effect of different data modalities and the benefits of a multimodal fusion-based approach using videos of 21 facial palsy patients. Our experimental results show that among various data modalities (i.e. unstructured data - RGB images …
Towards Gradient-Based Time-Series Explanations Through A Spatiotemporal Attention Network,
2024
Singapore Management University
Towards Gradient-Based Time-Series Explanations Through A Spatiotemporal Attention Network, Min Hun Lee
Research Collection School Of Computing and Information Systems
In this paper, we explore the feasibility of using a transformer-based, spatiotemporal attention network (STAN) for gradient-based time-series explanations. First, we trained the STAN model for video classifications using the global and local views of data and weakly supervised labels on time-series data (i.e. the type of an activity). We then leveraged a gradient-based XAI technique (e.g. saliency map) to identify salient frames of time-series data. According to the experiments using the datasets of four medically relevant activities, the STAN model demonstrated its potential to identify important frames of videos.
