Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Artificial Intelligence and Robotics (108)
- Engineering (68)
- Computer Engineering (46)
- Databases and Information Systems (41)
- Social and Behavioral Sciences (37)
-
- Electrical and Computer Engineering (28)
- Data Science (24)
- Numerical Analysis and Scientific Computing (23)
- Other Computer Sciences (22)
- Software Engineering (18)
- Medicine and Health Sciences (16)
- Business (14)
- Communication (14)
- Information Security (14)
- Arts and Humanities (12)
- Linguistics (9)
- Social Media (9)
- Theory and Algorithms (9)
- Law (8)
- Library and Information Science (8)
- Computational Linguistics (6)
- Education (6)
- Operations Research, Systems Engineering and Industrial Engineering (6)
- Bioinformatics (5)
- Graphics and Human Computer Interfaces (5)
- Life Sciences (5)
- Public Affairs, Public Policy and Public Administration (5)
- Scholarly Publishing (5)
- Institution
-
- Singapore Management University (47)
- Old Dominion University (30)
- TÜBİTAK (17)
- Technological University Dublin (13)
- Wright State University (9)
-
- Dartmouth College (7)
- New Jersey Institute of Technology (7)
- Boise State University (6)
- City University of New York (CUNY) (6)
- United Arab Emirates University (6)
- Virginia Commonwealth University (6)
- Zayed University (6)
- Brigham Young University (5)
- Loyola University Chicago (5)
- California Polytechnic State University, San Luis Obispo (4)
- San Jose State University (4)
- University of Arkansas Little Rock (4)
- University of Arkansas, Fayetteville (4)
- University of Denver (4)
- Western Michigan University (4)
- Air Force Institute of Technology (3)
- Edith Cowan University (3)
- Mississippi State University (3)
- Missouri University of Science and Technology (3)
- Montclair State University (3)
- Purdue University (3)
- Southern Methodist University (3)
- The Texas Medical Center Library (3)
- University of Kentucky (3)
- University of Nebraska at Omaha (3)
- Publication Year
- Publication
-
- Research Collection School Of Computing and Information Systems (37)
- Theses and Dissertations (18)
- Turkish Journal of Electrical Engineering and Computer Sciences (17)
- Dissertations (15)
- Electronic Theses and Dissertations (10)
-
- Computer Science Faculty Publications (8)
- All Works (6)
- Browse all Theses and Dissertations (6)
- Conference papers (6)
- Master's Theses (6)
- VMASC Publications (6)
- Dissertations and Theses Collection (Open Access) (5)
- Boise State University Theses and Dissertations (4)
- Computer Science Faculty Proceedings & Presentations (3)
- Computer Science Senior Theses (3)
- Computer Science Theses & Dissertations (3)
- Computer Science: Faculty Publications and Other Works (3)
- Dissertations, Theses, and Capstone Projects (3)
- Doctoral Dissertations (3)
- Engineering Management & Systems Engineering Faculty Publications (3)
- Master's Projects (3)
- SMU Data Science Review (3)
- Theses and Dissertations--Computer Science (3)
- Thesis/ Dissertation Defenses (3)
- UNLV Theses, Dissertations, Professional Papers, and Capstones (3)
- Articles (2)
- CGU Faculty Publications and Research (2)
- Computer Science Faculty Research & Creative Works (2)
- Computer Science and Engineering Faculty Publications (2)
- Computer Science and Engineering Theses - Archive (2)
- Publication Type
Articles 31 - 60 of 293
Full-Text Articles in Computer Sciences
Fact-Checking As A Multi-Step Process: From Ambiguity Resolution To Claim Validation, Wenbo Wang
Fact-Checking As A Multi-Step Process: From Ambiguity Resolution To Claim Validation, Wenbo Wang
Dissertations
The spread of misinformation and disinformation has become a major concern, particularly with the rise of social media as a primary source of information for many people. Fact-checking—the process of verifying claims against credible evidence—has emerged as a critical safeguard against misinformation. Yet, the task is fraught with challenges: claims are often ambiguous, context-dependent, or composed of multiple intertwined assertions, while automated systems struggle to replicate the nuanced reasoning of human experts. This dissertation addresses these challenges by reimagining fact-checking as a multi-step, knowledge-guided process that systematically resolves ambiguity, decomposes complexity, and validates claims through structured reasoning. Additionally, the proposed …
Analyzing Unmanned Aircraft System (Uas) Incidents From Nasa Asrs Data Using Unsupervised Machine Learning, Kacey Haws
Analyzing Unmanned Aircraft System (Uas) Incidents From Nasa Asrs Data Using Unsupervised Machine Learning, Kacey Haws
Electrical Engineering and Computer Science Undergraduate Honors Theses
The NASA Aviation Safety Reporting System (ASRS) assembles voluntarily submitted aviation safety incident reports in their database to act on the information provided. This database allows the government, companies, and citizens to submit incident or situational reports to its database to discern recurring issues in the National Aviation System (NAS) so that the proper officials can act [1]. The narratives provided in these reports are text-based, resulting in large amounts of data to process. Previous work in the University of Arkansas Aerospace Systems Engineering and Transportation Laboratory (ASYST) lab involved parsing unmanned aircraft system (UAS) incident reports manually. While these …
Cross-Dataset Fairness Evaluation Of Transformer-Based Sentiment Models, Sara Zuiran
Cross-Dataset Fairness Evaluation Of Transformer-Based Sentiment Models, Sara Zuiran
Theses and Dissertations
With the growing exploration of Natural Language Processing (NLP) systems in decision-making environments, it is essential to evaluate technical and ethical aspects of the dataset and the NLP model to improve fairness. To assess fairness, the thesis examines demographic imbalances in sentiment classification models by evaluating transformer-based models fine-tuned on the Stanford Sentiment Treebank version 2 dataset (SST-2) against the demographically annotated Comprehensive Assessment of Language Model dataset (CALM). This work identifies performance disparities in sentiment prediction across demographic groups by examining sensitive attributes such as gender and race. The study evaluates both the RoBERTa and MentalBERT transformer models using …
Ai Model For Predicting Asthma Prognosis In Children, Elham Sagheb, Chung-Il Wi, Katherine S King, Bhavani Singh Agnikula Kshatriya, Euijung Ryu, Hongfang Liu, Miguel A Park, Hee Yun Seol, Shauna M Overgaard, Deepak K Sharma, Young J Juhn, Sunghwan Sohn
Ai Model For Predicting Asthma Prognosis In Children, Elham Sagheb, Chung-Il Wi, Katherine S King, Bhavani Singh Agnikula Kshatriya, Euijung Ryu, Hongfang Liu, Miguel A Park, Hee Yun Seol, Shauna M Overgaard, Deepak K Sharma, Young J Juhn, Sunghwan Sohn
Faculty, Staff and Student Publications
BACKGROUND: Childhood asthma often continues into adulthood, but some children experience remission. Utilizing electronic health records (EHRs) to predict asthma prognosis can aid health care providers and patients in developing effective prioritized care plans.
OBJECTIVE: We aimed to develop artificial intelligence (AI) models using various clinical variables extracted from EHRs to predict childhood asthma prognosis (remission vs no remission) in different age groups.
METHODS: We developed AI models utilizing patients' EHRs during the first 6, 9, or 12 years of their lives to predict their asthma prognosis status at ages 6 to 9, 9 to 12, or 12 to 15 …
Modeling Language And Vision At Human Scales, Clayton Fields
Modeling Language And Vision At Human Scales, Clayton Fields
Boise State University Theses and Dissertations
The impressive results that have recently been achieved in natural language processing and artificial intelligence have been primarily driven by the introduction of the transformer deep learning architecture, increasingly large models with many parameters and using enormous datasets. The size of models and their training data requirements present costly demands that freeze many researchers out of training with cutting edge models. Beyond these practical implications, current methods learn from text alone, without the rich array of sensory information that human beings use in learning language. This means that language models are often incapable of reasoning about the concrete world that …
Creating Talking Points For Client Advisers At Banks To Promote Sustainable Investing, Wewe Zi Yi, Pradeep Varakantham, Alan Megargel
Creating Talking Points For Client Advisers At Banks To Promote Sustainable Investing, Wewe Zi Yi, Pradeep Varakantham, Alan Megargel
Research Collection School Of Computing and Information Systems
Environmental, social and governance (ESG) factors have become key nonfinancial factors for investors to evaluate companies with respect to understanding material risks and growth opportunities. While not mandatory, companies are providing ESG reports that outline progress in different ESG metrics (six broad metrics and 15 specific ones). Client advisers (CAs) read these reports to identify key metrics of interest to investors. Given the number of companies and investment products, however, it is not feasible for CAs to read all the reports, which can sometimes run into tens or hundreds of pages). The authors have developed multiple frameworks building on leading …
Ai Foundations And Applications: Summary Of A Panel Discussion At Loyola University Chicago, George K. Thiruvathukal, Dmitry Dligach, Shilpika, Michael B. Burns, Joseph Vukov, Fraser Turner, Mary Usher
Ai Foundations And Applications: Summary Of A Panel Discussion At Loyola University Chicago, George K. Thiruvathukal, Dmitry Dligach, Shilpika, Michael B. Burns, Joseph Vukov, Fraser Turner, Mary Usher
Computer Science: Faculty Publications and Other Works
This document summarizes the panel discussion titled "AI Foundations and Applications," held at Loyola University Chicago as part of the "Forum on Global Affairs: Artificial Intelligence in a Globalized World" series. The panel brought together interdisciplinary experts to discuss the foundational aspects of artificial intelligence (AI), its applications, ethical considerations, and implications for education and society.
Leveraging Large Language Models For Knowledge-Free Weak Supervision In Clinical Natural Language Processing, Enshuo Hsu, Kirk Roberts
Leveraging Large Language Models For Knowledge-Free Weak Supervision In Clinical Natural Language Processing, Enshuo Hsu, Kirk Roberts
Faculty, Staff and Student Publications
The performance of deep learning-based natural language processing systems is based on large amounts of labeled training data which, in the clinical domain, are not easily available or affordable. Weak supervision and in-context learning offer partial solutions to this issue, particularly using large language models (LLMs), but their performance still trails traditional supervised methods with moderate amounts of gold-standard data. In particular, inferencing with LLMs is computationally heavy. We propose an approach leveraging fine-tuning LLMs and weak supervision with virtually no domain knowledge that still achieves consistently dominant performance. Using a prompt-based approach, the LLM is used to generate weakly-labeled …
If You Were A Sesame Street Character, Which One Would You Be? Natural Language Processing And Personality With Big Bird And Friends, Joseph Uran Meyer
If You Were A Sesame Street Character, Which One Would You Be? Natural Language Processing And Personality With Big Bird And Friends, Joseph Uran Meyer
Doctoral Dissertations
This paper examined and compared several natural language processing and machine learning techniques in predicting self-reported Big Five personality traits from text responses. The models were validated on the open-source 2019 SIOP Machine Learning Competition dataset (N = 1,689). The techniques evaluated included bag-of-words, Empath dictionary, LSTM networks, fine-tuning Transformer models, and stacked generalization. Results indicated that the present study’s models had lower error in four of the five constructs analyzed. Limitations of the study include use of an MTurk sample and small sample size. Future research should explore similar techniques on larger applicant samples. Practical implications and contributions to …
Exploring The Trade-Offs: Unified Large Language Models Vs Local Fine-Tuned Models For Highly-Specific Radiology Nli Task, Zihao Wu, Lu Zhang, Chao Cao, Xiaowei Yu, Zhengliang Liu, Lin Zhao, Yiwei Li, Haixing Dai, Chong Ma, Gang Li, Wei Liu, Quanzheng Li, Dinggang Shen, Xiang Li, Dajiang Zhu, Tianming Liu
Exploring The Trade-Offs: Unified Large Language Models Vs Local Fine-Tuned Models For Highly-Specific Radiology Nli Task, Zihao Wu, Lu Zhang, Chao Cao, Xiaowei Yu, Zhengliang Liu, Lin Zhao, Yiwei Li, Haixing Dai, Chong Ma, Gang Li, Wei Liu, Quanzheng Li, Dinggang Shen, Xiang Li, Dajiang Zhu, Tianming Liu
Computer Science Faculty Research & Creative Works
Recently, ChatGPT and GPT-4 have emerged and gained immense global attention due to their unparalleled performance in language processing. Despite demonstrating impressive capability in various open-domain tasks, their adequacy in highly specific fields like radiology remains untested. Radiology presents unique linguistic phenomena distinct from open-domain data due to its specificity and complexity. Assessing the performance of large language models (LLMs) in such specific domains is crucial not only for a thorough evaluation of their overall performance but also for providing valuable insights into future model design directions: whether model design should be generic or domain specific. To this end, in …
Opinion Mining On Offshore Wind Energy For Environmental Engineering, Isabele Bittencourt, Aparna S. Varde, Pankaj Lal
Opinion Mining On Offshore Wind Energy For Environmental Engineering, Isabele Bittencourt, Aparna S. Varde, Pankaj Lal
School of Computing Faculty Scholarship and Creative Works
Renewable energy sources are vital to help mitigate the effects of climate change, and reducing the carbon dioxide emissions of fossil fuels, e.g. the state of New Jersey has a goal of producing 100% clean energy by 2050. However, the plans for offshore wind energy by the shore of the state still brings much controversy between residents due to the wind farms’ impact on wildlife, coastline, and the people’s view from the beaches. In this context, we perform sentiment analysis on social media data to investigate people’s opinions and concerns regarding offshore wind energy. We adapt 3 machine learning models, …
A Spine-Specific Lexicon For The Sentiment Analysis Of Interviews With Adult Spinal Deformity Patients Correlate With Sf-36, Sf-36, And Odi Scores: A Pilot Study Of 25 Patients, Ross Gore, Michael M. Safaee, Christopher J. Lynch, Christopher P. Ames
A Spine-Specific Lexicon For The Sentiment Analysis Of Interviews With Adult Spinal Deformity Patients Correlate With Sf-36, Sf-36, And Odi Scores: A Pilot Study Of 25 Patients, Ross Gore, Michael M. Safaee, Christopher J. Lynch, Christopher P. Ames
VMASC Publications
Classic health-related quality of life (HRQOL) metrics are cumbersome, time-intensive, and subject to biases based on the patient’s native language, educational level, and cultural values. Natural language processing (NLP) converts text into quantitative metrics. Sentiment analysis enables subject matter experts to construct domain-specific lexicons that assign a value of either negative (−1) or positive (1) to certain words. The growth of telehealth provides opportunities to apply sentiment analysis to transcripts of adult spinal deformity patients’ visits to derive a novel and less biased HRQOL metric. In this study, we demonstrate the feasibility of constructing a spine-specific lexicon for sentiment analysis …
An Integrated Machine Learning Approach For Identifying Emergency Rescue Messages On Social Media During Natural Disasters, Wael Khallouli, Jiang Li, Jingwei Huang, Ghaith Rabadi, Samuel Kovacic
An Integrated Machine Learning Approach For Identifying Emergency Rescue Messages On Social Media During Natural Disasters, Wael Khallouli, Jiang Li, Jingwei Huang, Ghaith Rabadi, Samuel Kovacic
School of Cybersecurity Faculty Publications
During large-scale disasters, emergency call centers are often overwhelmed by the large volume of rescue requests and calls for help. Consequently, people are turning to social media platforms to seek assistance. Rescue information posted on these platforms is extremely valuable for first responders to make informed rescue decisions. Therefore, the automatic identification of these requests from the vast amount of data posted on social media during crises is critical yet challenging. This work presents our ongoing research on applying deep learning techniques to extract actionable rescue information from social media during crises. We proposed a novel deep learning model that …
From Cyclones To Cybersecurity: A Call For Convergence In Risk And Crisis Communications Research, Ann Marie Reinhold, Ross J. Gore, Barry Ezell, Clemente I. Izurieta, Elizabeth A. Shanahan
From Cyclones To Cybersecurity: A Call For Convergence In Risk And Crisis Communications Research, Ann Marie Reinhold, Ross J. Gore, Barry Ezell, Clemente I. Izurieta, Elizabeth A. Shanahan
VMASC Publications
Effective risk and crisis communication can improve health and safety and reduce harmful effects of hazards and disasters. A robust body of literature investigates mechanisms for improving risk and crisis communication. While effective risk and crisis communication strategies are equally desired across different hazard types (e.g., natural hazards, cyber security), the extent to which risk and crisis communication experts utilize the “lessons learned” from scientific domains outside their own is suspect. Therefore, we hypothesized that risk and crisis communication research is siloed according to academic disciplines at the detriment to the advancement of the field of risk communications research writ …
Ai-Generated Messaging For Life Events Using Structured Prompts: A Comparative Study Of Gpt With Human Experts And Machine Learning, Christopher Lynch, Erik Jensen, Ross Gore, Virginia Zamponi, Kevin O'Brien, Brandon Feldhaus, Katherine Smith, Joseph Martínez, Madison H. Munro, Timur E. Ozkose, Tugce B. Gundogdu, Ann Marie Reinhold, Hamdi Kavak, Barry Ezell
Ai-Generated Messaging For Life Events Using Structured Prompts: A Comparative Study Of Gpt With Human Experts And Machine Learning, Christopher Lynch, Erik Jensen, Ross Gore, Virginia Zamponi, Kevin O'Brien, Brandon Feldhaus, Katherine Smith, Joseph Martínez, Madison H. Munro, Timur E. Ozkose, Tugce B. Gundogdu, Ann Marie Reinhold, Hamdi Kavak, Barry Ezell
VMASC Publications
Large Language Models (LLMs) play an increasingly integrated and pivotal role in generating diverse types of texts, such as social media messages, emails, narratives, and technical reports, among other textual communication forms. As AI-generated messaging filters into human communication, a systematic exploration of their effectiveness for mimicking human-like communication of life events is needed. In this study, we employ a zero-shot structured narrative prompt to generate 24,000 life event messages for birth, death, hiring, and firing events using OpenAI's GPT-4. From this dataset, we manually classify 2880 messages and evaluate their validity in conveying these life events through the form …
Toward Embodied Navigation Through Vision And Language, Muraleekrishna Gopinathan
Toward Embodied Navigation Through Vision And Language, Muraleekrishna Gopinathan
Theses: Doctorates and Masters
Embodied AI is a challenging but exciting field in which a robot learns to interact with human-living spaces to perform various tasks. This thesis studies the embodied navigation problem in which a robotic agent navigates in a previously unseen indoor environment based on a challenging task. In particular, the Vision-and-Language Navigation (VLN) task requires a robot to navigate based on a descriptive human-language instruction. This thesis aims to improve VLN agents on four key aspects - their understanding of the environment, training via additional data, correcting navigational errors, and predicting the layout of the environment for better planning.
First, we …
Enhancing Cybersecurity Through Autonomous Knowledge Graph Construction By Integrating Heterogeneous Data Sources, Hatoon Alharbi, Ali Hur, Hasan Alkahtani, Hafiz Farooq Ahmad
Enhancing Cybersecurity Through Autonomous Knowledge Graph Construction By Integrating Heterogeneous Data Sources, Hatoon Alharbi, Ali Hur, Hasan Alkahtani, Hafiz Farooq Ahmad
Research outputs 2022 to 2026
Cybersecurity plays a critical role in today’s modern human society, and leveraging knowledge graphs can enhance cybersecurity and privacy in the cyberspace. By harnessing the heterogeneous and vast amount of information on potential attacks, organizations can improve their ability to proactively detect and mitigate any threat or damage to their online valuable resources. Integrating critical cyberattack information into a knowledge graph offers a significant boost to cybersecurity, safeguarding cyberspace from malicious activities. This information can be obtained from structured and unstructured data, with a particular focus on extracting valuable insights from unstructured text through natural language processing (NLP). By storing …
Optimizing Ai Language Models: A Study Of Chatgpt-4 Vs. Chatgpt-4o, Md Nurul Absar Siddiky, Muhammad Enayetur Rahman, Md Fayaz Bin Hossen, Muhammad Rezaur Rahman, Md. Shahadat Jaman
Optimizing Ai Language Models: A Study Of Chatgpt-4 Vs. Chatgpt-4o, Md Nurul Absar Siddiky, Muhammad Enayetur Rahman, Md Fayaz Bin Hossen, Muhammad Rezaur Rahman, Md. Shahadat Jaman
Electrical & Computer Engineering Faculty Publications
This paper presents a comparative analysis of OpenAI's GPT-4 and its optimized variant, GPT-4o, focusing on their architectural differences, performance, and real-world applications. GPT-4, built upon the Transformer architecture, has set new standards in natural language processing (NLP) with its capacity to generate coherent and contextually relevant text across a wide range of tasks. However, its computational demands, requiring substantial hardware resources, make it less accessible for smaller organizations and real-time applications. In contrast, GPT-4o addresses these challenges by incorporating optimizations such as model compression, parameter pruning, and memory-efficient computation, allowing it to deliver similar performance with significantly lower computational …
A Review On Knowledge And Information Extraction From Pdf Documents And Storage Approaches, Salvador D. Atagong, Henri Tonnang, Kennedy Senagi, Mark Wamalwa, Komi M. Agboka, John Odindi
A Review On Knowledge And Information Extraction From Pdf Documents And Storage Approaches, Salvador D. Atagong, Henri Tonnang, Kennedy Senagi, Mark Wamalwa, Komi M. Agboka, John Odindi
All Peer-Reviewed Publications
Introduction: Automating the extraction of information from Portable Document Format (PDF) documents represents a major advancement in information extraction, with applications in various domains such as healthcare, law, or biochemistry. However, existing solutions face challenges related to accuracy, domain adaptability, and implementation complexity. Methods: A systematic review of the literature was conducted using the Preferred Reporting Items for Systematic Reviews and Meta-Analyses (PRISMA) methodology to examine approaches and trends in PDF information extraction and storage approaches. Results: The review revealed three dominant methodological categories: rule-based systems, statistical learning models, and neural network-based approaches. Key limitations include the rigidity of rule-based …
An Analytical Review Of Preprocessing Techniques In Bengali Natural Language Processing, Sovon Chakraborty, Protiva Das, Shakib Mahmud Dipto, Md Aktaruzzaman Pramanik, Jannatun Noor
An Analytical Review Of Preprocessing Techniques In Bengali Natural Language Processing, Sovon Chakraborty, Protiva Das, Shakib Mahmud Dipto, Md Aktaruzzaman Pramanik, Jannatun Noor
Computer Science Faculty Publications
Research in Bengali Natural Language Processing (BNLP) is rapidly expanding. Despite being one of the most widely spoken languages in the world, BNLP research remains insufficient, particularly in Bengali speech recognition. The languages rich morphology, agglutinative structure, and diverse dialects make text and speech processing especially challenging. However, these challenges can be addressed with effective preprocessing techniques. Various organizations in Bangladesh and West Bengal are integrating Natural Language Processing (NLP) into their services, but without a thorough understanding of preprocessing, these implementations remain incomplete. Applying proper preprocessing techniques to the Bengali language will serve as a foundation for developing robust …
From Philosophy To Nlu: Evolving Definitions With Research Hypotheses, Jian Wu, Sarah Rajtmajer
From Philosophy To Nlu: Evolving Definitions With Research Hypotheses, Jian Wu, Sarah Rajtmajer
Computer Science Faculty Publications
Over the past decades, alongside advancements in natural language processing, significant attention has been paid to training models to automatically extract, understand, test, and generate hypotheses in open and scientific domains. However, interpretations of the term hypothesis for various natural language understanding (NLU) tasks have migrated from traditional definitions in the natural, social, and formal sciences. Even within NLU, we observe differences defining hypotheses across literature. In this paper, we overview and delineate various definitions of hypothesis. Especially, we discern the nuances of definitions across recently published NLU tasks. We highlight the importance of well-structured and well-defined hypotheses, particularly as …
From Philosophy To Nlu: Evolving Definitions Of Research Hypotheses, Jian Wu, Sarah Rajtmajer
From Philosophy To Nlu: Evolving Definitions Of Research Hypotheses, Jian Wu, Sarah Rajtmajer
Computer Science Faculty Publications
Over the past decades, alongside advancements in natural language processing, significant attention has been paid to training models to automatically extract, understand, test, and generate hypotheses in open and scientific domains. However, interpretations of the term hypothesis for various natural language understanding (NLU) tasks have migrated from traditional definitions in the natural, social, and formal sciences. Even within NLU, we observe differences defining hypotheses across literature. In this paper, we overview and delineate various definitions of hypothesis. Especially, we discern the nuances of definitions across recently published NLU tasks. We highlight the importance of well-structured and well-defined hypotheses, particularly as …
Pushing The Boundaries Of Large Language Models: Innovations And Limitations In Nlp, Finance, And Mathematics, A M Muntasir Rahman
Pushing The Boundaries Of Large Language Models: Innovations And Limitations In Nlp, Finance, And Mathematics, A M Muntasir Rahman
Dissertations
Large Language Models (LLMs) have emerged as transformative tools across a spectrum of domains, yet their practical deployment reveals a blend of remarkable potential and notable limitations. This research explores innovative methodologies to extend the capabilities of LLMs while addressing critical challenges in their evaluation and application. By leveraging rule-based approaches, the in-context learning capabilities of LLMs, and human-in-the-loop validation across three focused studies, this research introduces robust strategies for dataset synthesis, model enhancement, and model assessment in three distinct domains: natural language processing, financial sentiment analysis, and mathematical reasoning
The first study proposes an efficient data augmentation framework, EASE, …
Sd-Weat: Towards Robustly Measuring Bias In Input Embeddings For Artificial Intelligence Language Models, Magnus Gray
Sd-Weat: Towards Robustly Measuring Bias In Input Embeddings For Artificial Intelligence Language Models, Magnus Gray
Theses and Dissertations
Artificial intelligence (AI) is rapidly transforming industries and markets, from healthcare to entertainment, revolutionizing decision-making processes. However, as AI grow more influential, they also risk amplifying existing biases, potentially leading to harmful consequences. Recent advancements in large language models (LLMs), such as GPT-4 and Llama, have heightened concerns about bias in natural language processing (NLP) tasks, driving the need for robust methods to detect and mitigate bias. Current approaches, such as the Word Embedding Association Test (WEAT) and its sentence-level extension the Sentence Encoder Association Test (SEAT) often fall short in capturing the nuances of biases in the input embeddings …
Lova3 : Learning To Visual Question Answering, Asking And Assessment, Henry Hengyuan Zhao, Pan Zhou, Difei Gao, Bai Shou, Mike Zheng Shou
Lova3 : Learning To Visual Question Answering, Asking And Assessment, Henry Hengyuan Zhao, Pan Zhou, Difei Gao, Bai Shou, Mike Zheng Shou
Research Collection School Of Computing and Information Systems
Question answering, asking, and assessment are three innate human traits crucial for understanding the world and acquiring knowledge. By enhancing these capabilities, humans can more effectively utilize data, leading to better comprehension and learning outcomes. Current Multimodal Large Language Models (MLLMs) primarily focus on question answering, often neglecting the full potential of questioning and assessment skills. Inspired by the human learning mechanism, we introduce LOVA3 , an innovative framework named “Learning tO Visual question Answering, Asking and Assessment,” designed to equip MLLMs with these additional capabilities. Our approach involves the creation of two supplementary training tasks GenQA and EvalQA, aiming …
Self-Supervised Fine-Tuning For Neural Expert Finding, Budhitama Subagdja, Dan Sanchari, Ah-Hwee Tan
Self-Supervised Fine-Tuning For Neural Expert Finding, Budhitama Subagdja, Dan Sanchari, Ah-Hwee Tan
Research Collection School Of Computing and Information Systems
Expert finding systems allow ones to find individuals who have expertise in specific fields or domains. Traditional expert finding are mostly based on topic modeling or keyword search methods that are limited in their capability to encode contextual knowledge from natural language. To address the limitation, this paper presents Neural Expert Finder (NEF), a novel method that takes a transfer learning approach based on transformer encoder networks to leverage the rich seman-tic and syntactic patterns of language encoded in pre-trained language models (PLMs). We propose a self-supervised learning approach utilizing contrastive training using both positive and automatically generated negative samples …
Converting Vocal Performances Into Sheet Music Leveraging Large Language Models, Jinjing Jiang, Nicole Teo, Haibo Pen, Seng-Beng Ho, Zhaoxia Wang
Converting Vocal Performances Into Sheet Music Leveraging Large Language Models, Jinjing Jiang, Nicole Teo, Haibo Pen, Seng-Beng Ho, Zhaoxia Wang
Research Collection School Of Computing and Information Systems
Advanced natural language processing (NLP) models are increasingly applied in music composition and performance, particularly for generating vocal melodies and simulating singing voices. While NLP techniques have been effective in analyzing vocal performance data to assess quality and style, the automatic transcription of vocal performances into sheet music remains a significant challenge. Manual transcription tools often fall short due to the intricate dynamics of vocal expression. This study tackles the automation of vocal performance transcription into sheet music using innovative techniques, including large language models (LLMs). We propose a method to translate vocal audio input into display-ready sheet music effectively. …
Tackling Toxicity And Harassment In Online Environments Through The Use Of Artificial Intelligence, Heba Saleous
Tackling Toxicity And Harassment In Online Environments Through The Use Of Artificial Intelligence, Heba Saleous
Thesis/ Dissertation Defenses
With the increase in popularity of online communities, such as social media platforms, online games, and chatroom servers, there is a need to improve chat and content moderation. Platforms have reported an increase in the prevalence of toxic behavior and hate speech. Meanwhile, moderators are reporting difficulties in keeping up with the amount of data to check as well and the type of content they are exposed to, which further harms their own mental health. The main objective of this work is to address the challenges that exist within online communities with the rising prevalence of hate speech. Additionally, some …
Dc-Instruct : An Effective Framework For Generative Multi-Intent Spoken Language Understanding, Bowen Xing, Lizi Liao, Minlie Huang
Dc-Instruct : An Effective Framework For Generative Multi-Intent Spoken Language Understanding, Bowen Xing, Lizi Liao, Minlie Huang
Research Collection School Of Computing and Information Systems
In the realm of multi-intent spoken language understanding, recent advancements have leveraged the potential of prompt learning frameworks. However, critical gaps exist in these frameworks: the lack of explicit modeling of dual-task dependencies and the oversight of task-specific semantic differences among utterances. To address these shortcomings, we propose DC-Instruct, a novel generative framework based on Dual-task Inter-dependent Instructions (DII) and Supervised Contrastive Instructions (SCI). Specifically, DII guides large language models (LLMs) to generate labels for one task based on the other task’s labels, thereby explicitly capturing dual-task inter-dependencies. Moreover, SCI leverages utterance semantics differences by guiding LLMs to determine whether …
A Comprehensive Survey On Relation Extraction: Recent Advances And New Frontiers, Xiaoyan Zhao, Yang Deng, Min Yang, Lingzhi Wang, Rui Zhang, Hong Cheng, Wai Lam, Ying Shen, Ruifeng Xu
A Comprehensive Survey On Relation Extraction: Recent Advances And New Frontiers, Xiaoyan Zhao, Yang Deng, Min Yang, Lingzhi Wang, Rui Zhang, Hong Cheng, Wai Lam, Ying Shen, Ruifeng Xu
Research Collection School Of Computing and Information Systems
Relation extraction (RE) involves identifying the relations between entities from underlying content. RE serves as the foundation for many natural language processing (NLP) and information retrieval applications, such as knowledge graph completion and question answering. In recent years, deep neural networks have dominated the field of RE and made noticeable progress. Subsequently, the large pre-trained language models (PLMs) have taken the state-of-the-art RE to a new level. This survey provides a comprehensive review of existing deep learning techniques for RE. First, we introduce RE resources, including datasets and evaluation metrics. Second, we propose a new taxonomy to categorize existing works …