Open Access. Powered by Scholars. Published by Universities.®

Computer Sciences Commons

Open Access. Powered by Scholars. Published by Universities.®

Natural language processing

Discipline
Institution
Publication Year
Publication
Publication Type
File Type

Articles 31 - 60 of 293

Full-Text Articles in Computer Sciences

Fact-Checking As A Multi-Step Process: From Ambiguity Resolution To Claim Validation, Wenbo Wang May 2025

Fact-Checking As A Multi-Step Process: From Ambiguity Resolution To Claim Validation, Wenbo Wang

Dissertations

The spread of misinformation and disinformation has become a major concern, particularly with the rise of social media as a primary source of information for many people. Fact-checking—the process of verifying claims against credible evidence—has emerged as a critical safeguard against misinformation. Yet, the task is fraught with challenges: claims are often ambiguous, context-dependent, or composed of multiple intertwined assertions, while automated systems struggle to replicate the nuanced reasoning of human experts. This dissertation addresses these challenges by reimagining fact-checking as a multi-step, knowledge-guided process that systematically resolves ambiguity, decomposes complexity, and validates claims through structured reasoning. Additionally, the proposed …


Analyzing Unmanned Aircraft System (Uas) Incidents From Nasa Asrs Data Using Unsupervised Machine Learning, Kacey Haws May 2025

Analyzing Unmanned Aircraft System (Uas) Incidents From Nasa Asrs Data Using Unsupervised Machine Learning, Kacey Haws

Electrical Engineering and Computer Science Undergraduate Honors Theses

The NASA Aviation Safety Reporting System (ASRS) assembles voluntarily submitted aviation safety incident reports in their database to act on the information provided. This database allows the government, companies, and citizens to submit incident or situational reports to its database to discern recurring issues in the National Aviation System (NAS) so that the proper officials can act [1]. The narratives provided in these reports are text-based, resulting in large amounts of data to process. Previous work in the University of Arkansas Aerospace Systems Engineering and Transportation Laboratory (ASYST) lab involved parsing unmanned aircraft system (UAS) incident reports manually. While these …


Cross-Dataset Fairness Evaluation Of Transformer-Based Sentiment Models, Sara Zuiran May 2025

Cross-Dataset Fairness Evaluation Of Transformer-Based Sentiment Models, Sara Zuiran

Theses and Dissertations

With the growing exploration of Natural Language Processing (NLP) systems in decision-making environments, it is essential to evaluate technical and ethical aspects of the dataset and the NLP model to improve fairness. To assess fairness, the thesis examines demographic imbalances in sentiment classification models by evaluating transformer-based models fine-tuned on the Stanford Sentiment Treebank version 2 dataset (SST-2) against the demographically annotated Comprehensive Assessment of Language Model dataset (CALM). This work identifies performance disparities in sentiment prediction across demographic groups by examining sensitive attributes such as gender and race. The study evaluates both the RoBERTa and MentalBERT transformer models using …


Ai Model For Predicting Asthma Prognosis In Children, Elham Sagheb, Chung-Il Wi, Katherine S King, Bhavani Singh Agnikula Kshatriya, Euijung Ryu, Hongfang Liu, Miguel A Park, Hee Yun Seol, Shauna M Overgaard, Deepak K Sharma, Young J Juhn, Sunghwan Sohn May 2025

Ai Model For Predicting Asthma Prognosis In Children, Elham Sagheb, Chung-Il Wi, Katherine S King, Bhavani Singh Agnikula Kshatriya, Euijung Ryu, Hongfang Liu, Miguel A Park, Hee Yun Seol, Shauna M Overgaard, Deepak K Sharma, Young J Juhn, Sunghwan Sohn

Faculty, Staff and Student Publications

BACKGROUND: Childhood asthma often continues into adulthood, but some children experience remission. Utilizing electronic health records (EHRs) to predict asthma prognosis can aid health care providers and patients in developing effective prioritized care plans.

OBJECTIVE: We aimed to develop artificial intelligence (AI) models using various clinical variables extracted from EHRs to predict childhood asthma prognosis (remission vs no remission) in different age groups.

METHODS: We developed AI models utilizing patients' EHRs during the first 6, 9, or 12 years of their lives to predict their asthma prognosis status at ages 6 to 9, 9 to 12, or 12 to 15 …


Modeling Language And Vision At Human Scales, Clayton Fields May 2025

Modeling Language And Vision At Human Scales, Clayton Fields

Boise State University Theses and Dissertations

The impressive results that have recently been achieved in natural language processing and artificial intelligence have been primarily driven by the introduction of the transformer deep learning architecture, increasingly large models with many parameters and using enormous datasets. The size of models and their training data requirements present costly demands that freeze many researchers out of training with cutting edge models. Beyond these practical implications, current methods learn from text alone, without the rich array of sensory information that human beings use in learning language. This means that language models are often incapable of reasoning about the concrete world that …


Creating Talking Points For Client Advisers At Banks To Promote Sustainable Investing, Wewe Zi Yi, Pradeep Varakantham, Alan Megargel May 2025

Creating Talking Points For Client Advisers At Banks To Promote Sustainable Investing, Wewe Zi Yi, Pradeep Varakantham, Alan Megargel

Research Collection School Of Computing and Information Systems

Environmental, social and governance (ESG) factors have become key nonfinancial factors for investors to evaluate companies with respect to understanding material risks and growth opportunities. While not mandatory, companies are providing ESG reports that outline progress in different ESG metrics (six broad metrics and 15 specific ones). Client advisers (CAs) read these reports to identify key metrics of interest to investors. Given the number of companies and investment products, however, it is not feasible for CAs to read all the reports, which can sometimes run into tens or hundreds of pages). The authors have developed multiple frameworks building on leading …


Ai Foundations And Applications: Summary Of A Panel Discussion At Loyola University Chicago, George K. Thiruvathukal, Dmitry Dligach, Shilpika, Michael B. Burns, Joseph Vukov, Fraser Turner, Mary Usher Apr 2025

Ai Foundations And Applications: Summary Of A Panel Discussion At Loyola University Chicago, George K. Thiruvathukal, Dmitry Dligach, Shilpika, Michael B. Burns, Joseph Vukov, Fraser Turner, Mary Usher

Computer Science: Faculty Publications and Other Works

This document summarizes the panel discussion titled "AI Foundations and Applications," held at Loyola University Chicago as part of the "Forum on Global Affairs: Artificial Intelligence in a Globalized World" series. The panel brought together interdisciplinary experts to discuss the foundational aspects of artificial intelligence (AI), its applications, ethical considerations, and implications for education and society.


Leveraging Large Language Models For Knowledge-Free Weak Supervision In Clinical Natural Language Processing, Enshuo Hsu, Kirk Roberts Mar 2025

Leveraging Large Language Models For Knowledge-Free Weak Supervision In Clinical Natural Language Processing, Enshuo Hsu, Kirk Roberts

Faculty, Staff and Student Publications

The performance of deep learning-based natural language processing systems is based on large amounts of labeled training data which, in the clinical domain, are not easily available or affordable. Weak supervision and in-context learning offer partial solutions to this issue, particularly using large language models (LLMs), but their performance still trails traditional supervised methods with moderate amounts of gold-standard data. In particular, inferencing with LLMs is computationally heavy. We propose an approach leveraging fine-tuning LLMs and weak supervision with virtually no domain knowledge that still achieves consistently dominant performance. Using a prompt-based approach, the LLM is used to generate weakly-labeled …


If You Were A Sesame Street Character, Which One Would You Be? Natural Language Processing And Personality With Big Bird And Friends, Joseph Uran Meyer Mar 2025

If You Were A Sesame Street Character, Which One Would You Be? Natural Language Processing And Personality With Big Bird And Friends, Joseph Uran Meyer

Doctoral Dissertations

This paper examined and compared several natural language processing and machine learning techniques in predicting self-reported Big Five personality traits from text responses. The models were validated on the open-source 2019 SIOP Machine Learning Competition dataset (N = 1,689). The techniques evaluated included bag-of-words, Empath dictionary, LSTM networks, fine-tuning Transformer models, and stacked generalization. Results indicated that the present study’s models had lower error in four of the five constructs analyzed. Limitations of the study include use of an MTurk sample and small sample size. Future research should explore similar techniques on larger applicant samples. Practical implications and contributions to …


Exploring The Trade-Offs: Unified Large Language Models Vs Local Fine-Tuned Models For Highly-Specific Radiology Nli Task, Zihao Wu, Lu Zhang, Chao Cao, Xiaowei Yu, Zhengliang Liu, Lin Zhao, Yiwei Li, Haixing Dai, Chong Ma, Gang Li, Wei Liu, Quanzheng Li, Dinggang Shen, Xiang Li, Dajiang Zhu, Tianming Liu Jan 2025

Exploring The Trade-Offs: Unified Large Language Models Vs Local Fine-Tuned Models For Highly-Specific Radiology Nli Task, Zihao Wu, Lu Zhang, Chao Cao, Xiaowei Yu, Zhengliang Liu, Lin Zhao, Yiwei Li, Haixing Dai, Chong Ma, Gang Li, Wei Liu, Quanzheng Li, Dinggang Shen, Xiang Li, Dajiang Zhu, Tianming Liu

Computer Science Faculty Research & Creative Works

Recently, ChatGPT and GPT-4 have emerged and gained immense global attention due to their unparalleled performance in language processing. Despite demonstrating impressive capability in various open-domain tasks, their adequacy in highly specific fields like radiology remains untested. Radiology presents unique linguistic phenomena distinct from open-domain data due to its specificity and complexity. Assessing the performance of large language models (LLMs) in such specific domains is crucial not only for a thorough evaluation of their overall performance but also for providing valuable insights into future model design directions: whether model design should be generic or domain specific. To this end, in …


Opinion Mining On Offshore Wind Energy For Environmental Engineering, Isabele Bittencourt, Aparna S. Varde, Pankaj Lal Jan 2025

Opinion Mining On Offshore Wind Energy For Environmental Engineering, Isabele Bittencourt, Aparna S. Varde, Pankaj Lal

School of Computing Faculty Scholarship and Creative Works

Renewable energy sources are vital to help mitigate the effects of climate change, and reducing the carbon dioxide emissions of fossil fuels, e.g. the state of New Jersey has a goal of producing 100% clean energy by 2050. However, the plans for offshore wind energy by the shore of the state still brings much controversy between residents due to the wind farms’ impact on wildlife, coastline, and the people’s view from the beaches. In this context, we perform sentiment analysis on social media data to investigate people’s opinions and concerns regarding offshore wind energy. We adapt 3 machine learning models, …


A Spine-Specific Lexicon For The Sentiment Analysis Of Interviews With Adult Spinal Deformity Patients Correlate With Sf-36, Sf-36, And Odi Scores: A Pilot Study Of 25 Patients, Ross Gore, Michael M. Safaee, Christopher J. Lynch, Christopher P. Ames Jan 2025

A Spine-Specific Lexicon For The Sentiment Analysis Of Interviews With Adult Spinal Deformity Patients Correlate With Sf-36, Sf-36, And Odi Scores: A Pilot Study Of 25 Patients, Ross Gore, Michael M. Safaee, Christopher J. Lynch, Christopher P. Ames

VMASC Publications

Classic health-related quality of life (HRQOL) metrics are cumbersome, time-intensive, and subject to biases based on the patient’s native language, educational level, and cultural values. Natural language processing (NLP) converts text into quantitative metrics. Sentiment analysis enables subject matter experts to construct domain-specific lexicons that assign a value of either negative (−1) or positive (1) to certain words. The growth of telehealth provides opportunities to apply sentiment analysis to transcripts of adult spinal deformity patients’ visits to derive a novel and less biased HRQOL metric. In this study, we demonstrate the feasibility of constructing a spine-specific lexicon for sentiment analysis …


An Integrated Machine Learning Approach For Identifying Emergency Rescue Messages On Social Media During Natural Disasters, Wael Khallouli, Jiang Li, Jingwei Huang, Ghaith Rabadi, Samuel Kovacic Jan 2025

An Integrated Machine Learning Approach For Identifying Emergency Rescue Messages On Social Media During Natural Disasters, Wael Khallouli, Jiang Li, Jingwei Huang, Ghaith Rabadi, Samuel Kovacic

School of Cybersecurity Faculty Publications

During large-scale disasters, emergency call centers are often overwhelmed by the large volume of rescue requests and calls for help. Consequently, people are turning to social media platforms to seek assistance. Rescue information posted on these platforms is extremely valuable for first responders to make informed rescue decisions. Therefore, the automatic identification of these requests from the vast amount of data posted on social media during crises is critical yet challenging. This work presents our ongoing research on applying deep learning techniques to extract actionable rescue information from social media during crises. We proposed a novel deep learning model that …


From Cyclones To Cybersecurity: A Call For Convergence In Risk And Crisis Communications Research, Ann Marie Reinhold, Ross J. Gore, Barry Ezell, Clemente I. Izurieta, Elizabeth A. Shanahan Jan 2025

From Cyclones To Cybersecurity: A Call For Convergence In Risk And Crisis Communications Research, Ann Marie Reinhold, Ross J. Gore, Barry Ezell, Clemente I. Izurieta, Elizabeth A. Shanahan

VMASC Publications

Effective risk and crisis communication can improve health and safety and reduce harmful effects of hazards and disasters. A robust body of literature investigates mechanisms for improving risk and crisis communication. While effective risk and crisis communication strategies are equally desired across different hazard types (e.g., natural hazards, cyber security), the extent to which risk and crisis communication experts utilize the “lessons learned” from scientific domains outside their own is suspect. Therefore, we hypothesized that risk and crisis communication research is siloed according to academic disciplines at the detriment to the advancement of the field of risk communications research writ …


Ai-Generated Messaging For Life Events Using Structured Prompts: A Comparative Study Of Gpt With Human Experts And Machine Learning, Christopher Lynch, Erik Jensen, Ross Gore, Virginia Zamponi, Kevin O'Brien, Brandon Feldhaus, Katherine Smith, Joseph Martínez, Madison H. Munro, Timur E. Ozkose, Tugce B. Gundogdu, Ann Marie Reinhold, Hamdi Kavak, Barry Ezell Jan 2025

Ai-Generated Messaging For Life Events Using Structured Prompts: A Comparative Study Of Gpt With Human Experts And Machine Learning, Christopher Lynch, Erik Jensen, Ross Gore, Virginia Zamponi, Kevin O'Brien, Brandon Feldhaus, Katherine Smith, Joseph Martínez, Madison H. Munro, Timur E. Ozkose, Tugce B. Gundogdu, Ann Marie Reinhold, Hamdi Kavak, Barry Ezell

VMASC Publications

Large Language Models (LLMs) play an increasingly integrated and pivotal role in generating diverse types of texts, such as social media messages, emails, narratives, and technical reports, among other textual communication forms. As AI-generated messaging filters into human communication, a systematic exploration of their effectiveness for mimicking human-like communication of life events is needed. In this study, we employ a zero-shot structured narrative prompt to generate 24,000 life event messages for birth, death, hiring, and firing events using OpenAI's GPT-4. From this dataset, we manually classify 2880 messages and evaluate their validity in conveying these life events through the form …


Toward Embodied Navigation Through Vision And Language, Muraleekrishna Gopinathan Jan 2025

Toward Embodied Navigation Through Vision And Language, Muraleekrishna Gopinathan

Theses: Doctorates and Masters

Embodied AI is a challenging but exciting field in which a robot learns to interact with human-living spaces to perform various tasks. This thesis studies the embodied navigation problem in which a robotic agent navigates in a previously unseen indoor environment based on a challenging task. In particular, the Vision-and-Language Navigation (VLN) task requires a robot to navigate based on a descriptive human-language instruction. This thesis aims to improve VLN agents on four key aspects - their understanding of the environment, training via additional data, correcting navigational errors, and predicting the layout of the environment for better planning.

First, we …


Enhancing Cybersecurity Through Autonomous Knowledge Graph Construction By Integrating Heterogeneous Data Sources, Hatoon Alharbi, Ali Hur, Hasan Alkahtani, Hafiz Farooq Ahmad Jan 2025

Enhancing Cybersecurity Through Autonomous Knowledge Graph Construction By Integrating Heterogeneous Data Sources, Hatoon Alharbi, Ali Hur, Hasan Alkahtani, Hafiz Farooq Ahmad

Research outputs 2022 to 2026

Cybersecurity plays a critical role in today’s modern human society, and leveraging knowledge graphs can enhance cybersecurity and privacy in the cyberspace. By harnessing the heterogeneous and vast amount of information on potential attacks, organizations can improve their ability to proactively detect and mitigate any threat or damage to their online valuable resources. Integrating critical cyberattack information into a knowledge graph offers a significant boost to cybersecurity, safeguarding cyberspace from malicious activities. This information can be obtained from structured and unstructured data, with a particular focus on extracting valuable insights from unstructured text through natural language processing (NLP). By storing …


Optimizing Ai Language Models: A Study Of Chatgpt-4 Vs. Chatgpt-4o, Md Nurul Absar Siddiky, Muhammad Enayetur Rahman, Md Fayaz Bin Hossen, Muhammad Rezaur Rahman, Md. Shahadat Jaman Jan 2025

Optimizing Ai Language Models: A Study Of Chatgpt-4 Vs. Chatgpt-4o, Md Nurul Absar Siddiky, Muhammad Enayetur Rahman, Md Fayaz Bin Hossen, Muhammad Rezaur Rahman, Md. Shahadat Jaman

Electrical & Computer Engineering Faculty Publications

This paper presents a comparative analysis of OpenAI's GPT-4 and its optimized variant, GPT-4o, focusing on their architectural differences, performance, and real-world applications. GPT-4, built upon the Transformer architecture, has set new standards in natural language processing (NLP) with its capacity to generate coherent and contextually relevant text across a wide range of tasks. However, its computational demands, requiring substantial hardware resources, make it less accessible for smaller organizations and real-time applications. In contrast, GPT-4o addresses these challenges by incorporating optimizations such as model compression, parameter pruning, and memory-efficient computation, allowing it to deliver similar performance with significantly lower computational …


A Review On Knowledge And Information Extraction From Pdf Documents And Storage Approaches, Salvador D. Atagong, Henri Tonnang, Kennedy Senagi, Mark Wamalwa, Komi M. Agboka, John Odindi Jan 2025

A Review On Knowledge And Information Extraction From Pdf Documents And Storage Approaches, Salvador D. Atagong, Henri Tonnang, Kennedy Senagi, Mark Wamalwa, Komi M. Agboka, John Odindi

All Peer-Reviewed Publications

Introduction: Automating the extraction of information from Portable Document Format (PDF) documents represents a major advancement in information extraction, with applications in various domains such as healthcare, law, or biochemistry. However, existing solutions face challenges related to accuracy, domain adaptability, and implementation complexity. Methods: A systematic review of the literature was conducted using the Preferred Reporting Items for Systematic Reviews and Meta-Analyses (PRISMA) methodology to examine approaches and trends in PDF information extraction and storage approaches. Results: The review revealed three dominant methodological categories: rule-based systems, statistical learning models, and neural network-based approaches. Key limitations include the rigidity of rule-based …


An Analytical Review Of Preprocessing Techniques In Bengali Natural Language Processing, Sovon Chakraborty, Protiva Das, Shakib Mahmud Dipto, Md Aktaruzzaman Pramanik, Jannatun Noor Jan 2025

An Analytical Review Of Preprocessing Techniques In Bengali Natural Language Processing, Sovon Chakraborty, Protiva Das, Shakib Mahmud Dipto, Md Aktaruzzaman Pramanik, Jannatun Noor

Computer Science Faculty Publications

Research in Bengali Natural Language Processing (BNLP) is rapidly expanding. Despite being one of the most widely spoken languages in the world, BNLP research remains insufficient, particularly in Bengali speech recognition. The languages rich morphology, agglutinative structure, and diverse dialects make text and speech processing especially challenging. However, these challenges can be addressed with effective preprocessing techniques. Various organizations in Bangladesh and West Bengal are integrating Natural Language Processing (NLP) into their services, but without a thorough understanding of preprocessing, these implementations remain incomplete. Applying proper preprocessing techniques to the Bengali language will serve as a foundation for developing robust …


From Philosophy To Nlu: Evolving Definitions With Research Hypotheses, Jian Wu, Sarah Rajtmajer Jan 2025

From Philosophy To Nlu: Evolving Definitions With Research Hypotheses, Jian Wu, Sarah Rajtmajer

Computer Science Faculty Publications

Over the past decades, alongside advancements in natural language processing, significant attention has been paid to training models to automatically extract, understand, test, and generate hypotheses in open and scientific domains. However, interpretations of the term hypothesis for various natural language understanding (NLU) tasks have migrated from traditional definitions in the natural, social, and formal sciences. Even within NLU, we observe differences defining hypotheses across literature. In this paper, we overview and delineate various definitions of hypothesis. Especially, we discern the nuances of definitions across recently published NLU tasks. We highlight the importance of well-structured and well-defined hypotheses, particularly as …


From Philosophy To Nlu: Evolving Definitions Of Research Hypotheses, Jian Wu, Sarah Rajtmajer Jan 2025

From Philosophy To Nlu: Evolving Definitions Of Research Hypotheses, Jian Wu, Sarah Rajtmajer

Computer Science Faculty Publications

Over the past decades, alongside advancements in natural language processing, significant attention has been paid to training models to automatically extract, understand, test, and generate hypotheses in open and scientific domains. However, interpretations of the term hypothesis for various natural language understanding (NLU) tasks have migrated from traditional definitions in the natural, social, and formal sciences. Even within NLU, we observe differences defining hypotheses across literature. In this paper, we overview and delineate various definitions of hypothesis. Especially, we discern the nuances of definitions across recently published NLU tasks. We highlight the importance of well-structured and well-defined hypotheses, particularly as …


Pushing The Boundaries Of Large Language Models: Innovations And Limitations In Nlp, Finance, And Mathematics, A M Muntasir Rahman Dec 2024

Pushing The Boundaries Of Large Language Models: Innovations And Limitations In Nlp, Finance, And Mathematics, A M Muntasir Rahman

Dissertations

Large Language Models (LLMs) have emerged as transformative tools across a spectrum of domains, yet their practical deployment reveals a blend of remarkable potential and notable limitations. This research explores innovative methodologies to extend the capabilities of LLMs while addressing critical challenges in their evaluation and application. By leveraging rule-based approaches, the in-context learning capabilities of LLMs, and human-in-the-loop validation across three focused studies, this research introduces robust strategies for dataset synthesis, model enhancement, and model assessment in three distinct domains: natural language processing, financial sentiment analysis, and mathematical reasoning

The first study proposes an efficient data augmentation framework, EASE, …


Sd-Weat: Towards Robustly Measuring Bias In Input Embeddings For Artificial Intelligence Language Models, Magnus Gray Dec 2024

Sd-Weat: Towards Robustly Measuring Bias In Input Embeddings For Artificial Intelligence Language Models, Magnus Gray

Theses and Dissertations

Artificial intelligence (AI) is rapidly transforming industries and markets, from healthcare to entertainment, revolutionizing decision-making processes. However, as AI grow more influential, they also risk amplifying existing biases, potentially leading to harmful consequences. Recent advancements in large language models (LLMs), such as GPT-4 and Llama, have heightened concerns about bias in natural language processing (NLP) tasks, driving the need for robust methods to detect and mitigate bias. Current approaches, such as the Word Embedding Association Test (WEAT) and its sentence-level extension the Sentence Encoder Association Test (SEAT) often fall short in capturing the nuances of biases in the input embeddings …


Lova3 : Learning To Visual Question Answering, Asking And Assessment, Henry Hengyuan Zhao, Pan Zhou, Difei Gao, Bai Shou, Mike Zheng Shou Dec 2024

Lova3 : Learning To Visual Question Answering, Asking And Assessment, Henry Hengyuan Zhao, Pan Zhou, Difei Gao, Bai Shou, Mike Zheng Shou

Research Collection School Of Computing and Information Systems

Question answering, asking, and assessment are three innate human traits crucial for understanding the world and acquiring knowledge. By enhancing these capabilities, humans can more effectively utilize data, leading to better comprehension and learning outcomes. Current Multimodal Large Language Models (MLLMs) primarily focus on question answering, often neglecting the full potential of questioning and assessment skills. Inspired by the human learning mechanism, we introduce LOVA3 , an innovative framework named “Learning tO Visual question Answering, Asking and Assessment,” designed to equip MLLMs with these additional capabilities. Our approach involves the creation of two supplementary training tasks GenQA and EvalQA, aiming …


Self-Supervised Fine-Tuning For Neural Expert Finding, Budhitama Subagdja, Dan Sanchari, Ah-Hwee Tan Dec 2024

Self-Supervised Fine-Tuning For Neural Expert Finding, Budhitama Subagdja, Dan Sanchari, Ah-Hwee Tan

Research Collection School Of Computing and Information Systems

Expert finding systems allow ones to find individuals who have expertise in specific fields or domains. Traditional expert finding are mostly based on topic modeling or keyword search methods that are limited in their capability to encode contextual knowledge from natural language. To address the limitation, this paper presents Neural Expert Finder (NEF), a novel method that takes a transfer learning approach based on transformer encoder networks to leverage the rich seman-tic and syntactic patterns of language encoded in pre-trained language models (PLMs). We propose a self-supervised learning approach utilizing contrastive training using both positive and automatically generated negative samples …


Converting Vocal Performances Into Sheet Music Leveraging Large Language Models, Jinjing Jiang, Nicole Teo, Haibo Pen, Seng-Beng Ho, Zhaoxia Wang Dec 2024

Converting Vocal Performances Into Sheet Music Leveraging Large Language Models, Jinjing Jiang, Nicole Teo, Haibo Pen, Seng-Beng Ho, Zhaoxia Wang

Research Collection School Of Computing and Information Systems

Advanced natural language processing (NLP) models are increasingly applied in music composition and performance, particularly for generating vocal melodies and simulating singing voices. While NLP techniques have been effective in analyzing vocal performance data to assess quality and style, the automatic transcription of vocal performances into sheet music remains a significant challenge. Manual transcription tools often fall short due to the intricate dynamics of vocal expression. This study tackles the automation of vocal performance transcription into sheet music using innovative techniques, including large language models (LLMs). We propose a method to translate vocal audio input into display-ready sheet music effectively. …


Tackling Toxicity And Harassment In Online Environments Through The Use Of Artificial Intelligence, Heba Saleous Nov 2024

Tackling Toxicity And Harassment In Online Environments Through The Use Of Artificial Intelligence, Heba Saleous

Thesis/ Dissertation Defenses

With the increase in popularity of online communities, such as social media platforms, online games, and chatroom servers, there is a need to improve chat and content moderation. Platforms have reported an increase in the prevalence of toxic behavior and hate speech. Meanwhile, moderators are reporting difficulties in keeping up with the amount of data to check as well and the type of content they are exposed to, which further harms their own mental health. The main objective of this work is to address the challenges that exist within online communities with the rising prevalence of hate speech. Additionally, some …


Dc-Instruct : An Effective Framework For Generative Multi-Intent Spoken Language Understanding, Bowen Xing, Lizi Liao, Minlie Huang Nov 2024

Dc-Instruct : An Effective Framework For Generative Multi-Intent Spoken Language Understanding, Bowen Xing, Lizi Liao, Minlie Huang

Research Collection School Of Computing and Information Systems

In the realm of multi-intent spoken language understanding, recent advancements have leveraged the potential of prompt learning frameworks. However, critical gaps exist in these frameworks: the lack of explicit modeling of dual-task dependencies and the oversight of task-specific semantic differences among utterances. To address these shortcomings, we propose DC-Instruct, a novel generative framework based on Dual-task Inter-dependent Instructions (DII) and Supervised Contrastive Instructions (SCI). Specifically, DII guides large language models (LLMs) to generate labels for one task based on the other task’s labels, thereby explicitly capturing dual-task inter-dependencies. Moreover, SCI leverages utterance semantics differences by guiding LLMs to determine whether …


A Comprehensive Survey On Relation Extraction: Recent Advances And New Frontiers, Xiaoyan Zhao, Yang Deng, Min Yang, Lingzhi Wang, Rui Zhang, Hong Cheng, Wai Lam, Ying Shen, Ruifeng Xu Nov 2024

A Comprehensive Survey On Relation Extraction: Recent Advances And New Frontiers, Xiaoyan Zhao, Yang Deng, Min Yang, Lingzhi Wang, Rui Zhang, Hong Cheng, Wai Lam, Ying Shen, Ruifeng Xu

Research Collection School Of Computing and Information Systems

Relation extraction (RE) involves identifying the relations between entities from underlying content. RE serves as the foundation for many natural language processing (NLP) and information retrieval applications, such as knowledge graph completion and question answering. In recent years, deep neural networks have dominated the field of RE and made noticeable progress. Subsequently, the large pre-trained language models (PLMs) have taken the state-of-the-art RE to a new level. This survey provides a comprehensive review of existing deep learning techniques for RE. First, we introduce RE resources, including datasets and evaluation metrics. Second, we propose a new taxonomy to categorize existing works …