Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Social and Behavioral Sciences (3)
- Artificial Intelligence and Robotics (2)
- Computational Linguistics (2)
- Computer Sciences (2)
- Linguistics (2)
-
- Medicine and Health Sciences (2)
- Physical Sciences and Mathematics (2)
- Robotics (2)
- Science and Technology Studies (2)
- Analytical, Diagnostic and Therapeutic Techniques and Equipment (1)
- Anatomy (1)
- Applied Behavior Analysis (1)
- Applied Mathematics (1)
- Archival Science (1)
- Business (1)
- Business Administration, Management, and Operations (1)
- Business Intelligence (1)
- Cataloging and Metadata (1)
- Chemistry (1)
- Cognition and Perception (1)
- Cognitive Psychology (1)
- Collection Development and Management (1)
- Communication (1)
- Communication Sciences and Disorders (1)
- Communication Technology and New Media (1)
- Community-Based Learning (1)
- Community-Based Research (1)
- Institution
- Publication
- Publication Type
Articles 1 - 15 of 15
Full-Text Articles in Other Computer Engineering
Augmenting Machine Learning Technique Through Natural Language, Tasmia Tasrin
Augmenting Machine Learning Technique Through Natural Language, Tasmia Tasrin
Theses and Dissertations--Computer Science
While artificial intelligence (AI) and machine learning (ML) have proven effective at addressing many of the challenges that we face in our everyday lives, there are many situations in which these methods struggle. Examples include environments where AI or ML systems must perform complex behaviors or those where rewards are difficult to calculate. To address this limitation, interactive machine learning (IML) techniques have been introduced, which incorporate machine-understandable human feedback into traditional ML approaches. This feedback is often given as a discrete, positive or negative numeric value. This feedback is typically provided as often as possible to convey a dense …
Advanced Knowledge Extraction With Biomedical Data Using Llms, Akshat Krishna
Advanced Knowledge Extraction With Biomedical Data Using Llms, Akshat Krishna
Master's Projects
The rapid growth of biomedical research has led to an overwhelming volume of unstructured textual data in the scientific literature. This has necessitated the development of an automated approach for knowledge extraction and integration. In
this project, we present a comprehensive pipeline for constructing a unified biomed- ical knowledge graph by combining two well-known datasets: CHEMPROT [1],
which captures chemical–protein interactions, and EU-ADR [2], which annotates drug–gene–disease relationships. In order to identify important biomedical entities and interactions from CHEMPROT dataset, we perform Named Entity Recognition (NER) and relation Extraction (RE) using state-of-the-art biomedical models like BioBERT [3], BioGPT [4] and …
Advancing Phishing Protection: Employing Sophisticated Methods For Precise Url Evaluation, Abbhinav Reddie Nomuiia
Advancing Phishing Protection: Employing Sophisticated Methods For Precise Url Evaluation, Abbhinav Reddie Nomuiia
Master's Projects
In short, the incidence of phishing - the illegal act of people pretending to be well-known companies to secure personal information - has skyrocketed in the past few years. In 2022 alone, 300,000 consumers in the United States were captured by scammers using phishing techniques, losing in all over $50 million. In the span of two weeks, over 510 million attempts occurred in a variety of sectors, particularly instant messaging platforms, package delivery businesses, and digital currency trading. Since most businesses have recognized that they are prone to these exposure cases, there has been a sixty percent increase in businesses …
Opinion Graphs Construction For Reviews Using Transfer Learning And Large Language Models, Yichen Lin
Opinion Graphs Construction For Reviews Using Transfer Learning And Large Language Models, Yichen Lin
Master's Projects
With the rapid development of the Internet, reading online reviews before making a purchase, booking a hotel, or making a restaurant reservation has become a part of daily life. Customers often consider reviews as crucial supplementary information before making decisions on how to spend their money. However, reading many reviews to gain helpful information takes time and effort. This project proposes a new method OpinionGraphGenerator that aims to create opinion graphs from hotel reviews to reduce the high volume of text in reviews while preserving essential insights. In an opinion graph, vertices are semantically similar opinions, where each opinion consists …
Hindi Image Captioning Using Indictrans2 And Encoder-Decoder Architecture, Anahita Vayalombrone Dinesh
Hindi Image Captioning Using Indictrans2 And Encoder-Decoder Architecture, Anahita Vayalombrone Dinesh
Master's Projects
One of the most prominent tasks that lie on the conjunction of Natural Language Processing (NLP) and computer vision, is image captioning. Image captioning is the generative task of achieving textual descriptions from images. Its application finds use in many real-world scenarios like aiding the visually impaired, editing applications, recommendation systems, and medical imaging. This research focus lies in Hindi image captioning, the official language of India, as it has not been explored as far as its need. Several challenges such as the lack of substantial Hindi text data for training models, the need for human annotators to verify the …
A Rule-Based Hybrid Translation Model For Context-Aware Speech To Indian Sign Language Tasks, Dania Jaison
A Rule-Based Hybrid Translation Model For Context-Aware Speech To Indian Sign Language Tasks, Dania Jaison
Master's Projects
In India, one of the most significant tools to communicate with the Deaf and Hard of Hearing (DHH) communities is Indian Sign Language(ISL). The issue of a scarcity of computational resources for ISL is even more pronounced when it comes to the translation of written or spoken English into ISL. This paper proposes a rule-based model for translation where the spoken english sentences are translated into ISL by focusing on the specific syntactic and grammatical differences between these two languages. ISL uses a simplified syntax unlike the nuanced sentence structures in English — we omit most of the auxiliary verbs …
Leveraging Large Language Models For Transforming Student Information Into Actionable Data, Sree Hari Karri
Leveraging Large Language Models For Transforming Student Information Into Actionable Data, Sree Hari Karri
Master's Projects
Admission season places significant demands on university committees, necessitating the review of vast arrays of documents to assess students’ competence. This project advances the development of an automated system designed to streamline this process by evaluating application materials such as Letters of Recommendation (LoRs), Statements of Purpose (SoPs), and resumes. Utilizing a variety of advanced Natural Language Processing (NLP) techniques, the system compares the performance of several Large Language Model (LLM) approaches. It also experiments with different data handling strategies, including the use of vector stores versus traditional context-based processing, to optimize model efficiency and accuracy. Special attention is given …
Chatgpt As Metamorphosis Designer For The Future Of Artificial Intelligence (Ai): A Conceptual Investigation, Amarjit Kumar Singh (Library Assistant), Dr. Pankaj Mathur (Deputy Librarian)
Chatgpt As Metamorphosis Designer For The Future Of Artificial Intelligence (Ai): A Conceptual Investigation, Amarjit Kumar Singh (Library Assistant), Dr. Pankaj Mathur (Deputy Librarian)
Library Philosophy and Practice (e-journal)
Abstract
Purpose: The purpose of this research paper is to explore ChatGPT’s potential as an innovative designer tool for the future development of artificial intelligence. Specifically, this conceptual investigation aims to analyze ChatGPT’s capabilities as a tool for designing and developing near about human intelligent systems for futuristic used and developed in the field of Artificial Intelligence (AI). Also with the helps of this paper, researchers are analyzed the strengths and weaknesses of ChatGPT as a tool, and identify possible areas for improvement in its development and implementation. This investigation focused on the various features and functions of ChatGPT that …
Improving Relation Extraction From Unstructured Genealogical Texts Using Fine-Tuned Transformers, Carloangello Parrolivelli
Improving Relation Extraction From Unstructured Genealogical Texts Using Fine-Tuned Transformers, Carloangello Parrolivelli
Master's Theses
Though exploring one’s family lineage through genealogical family trees can be insightful to developing one’s identity, this knowledge is typically held behind closed doors by private companies or require expensive technologies, such as DNA testing, to uncover. With the ever-booming explosion of data on the world wide web, many unstructured text documents, both old and new, are being discovered, written, and processed which contain rich genealogical information. With access to this immense amount of data, however, entails a costly process whereby people, typically volunteers, have to read large amounts of text to find relationships between people. This delays having genealogical …
Data Science Methods For Standardization, Safety, And Quality Assurance In Radiation Oncology, Khajamoinuddin Syed
Data Science Methods For Standardization, Safety, And Quality Assurance In Radiation Oncology, Khajamoinuddin Syed
Theses and Dissertations
Radiation oncology is the field of medicine that deals with treating cancer patients through ionizing radiation. The clinical modality or technique used to treat the cancer patients in the radiation oncology domain is referred to as radiation therapy. Radiation therapy aims to deliver precisely measured dose irradiation to a defined tumor volume (target) with as minimal damage as possible to surrounding healthy tissue (organs-at-risk), resulting in eradication of the tumor, high quality of life, and prolongation of survival. A typical radiotherapy process requires the use of different clinical systems at various stages of the workflow. The data generated in these …
Finding Truth In Fake News: Reverse Plagiarism And Other Models Of Classification, Matthew Przybyla, David Tran, Amber Whelpley, Daniel W. Engels
Finding Truth In Fake News: Reverse Plagiarism And Other Models Of Classification, Matthew Przybyla, David Tran, Amber Whelpley, Daniel W. Engels
SMU Data Science Review
As the digital age creates new ways of spreading news, fake stories are propagated to widen audiences. A majority of people obtain both fake and truthful news without knowing which is which. There is not currently a reliable and efficient method to identify “fake news”. Several ways of detecting fake news have been produced, but the various algorithms have low accuracy of detection and the definition of what makes a news item ‘fake’ remains unclear. In this paper, we propose a new method of detecting on of fake news through comparison to other news items on the same topic, as …
Natural Language Processing Based Generator Of Testing Instruments, Qianqian Wang
Natural Language Processing Based Generator Of Testing Instruments, Qianqian Wang
Electronic Theses, Projects, and Dissertations
Natural Language Processing (NLP) is the field of study that focuses on the interactions between human language and computers. By “natural language” we mean a language that is used for everyday communication by humans. Different from programming languages, natural languages are hard to be defined with accurate rules. NLP is developing rapidly and it has been widely used in different industries. Technologies based on NLP are becoming increasingly widespread, for example, Siri or Alexa are intelligent personal assistants using NLP build in an algorithm to communicate with people. “Natural Language Processing Based Generator of Testing Instruments” is a stand-alone program …
An Empirical Study Of Semantic Similarity In Wordnet And Word2vec, Abram Handler
An Empirical Study Of Semantic Similarity In Wordnet And Word2vec, Abram Handler
LSU New Orleans Theses and Dissertations
This thesis performs an empirical analysis of Word2Vec by comparing its output to WordNet, a well-known, human-curated lexical database. It finds that Word2Vec tends to uncover more of certain types of semantic relations than others -- with Word2Vec returning more hypernyms, synonomyns and hyponyms than hyponyms or holonyms. It also shows the probability that neighbors separated by a given cosine distance in Word2Vec are semantically related in WordNet. This result both adds to our understanding of the still-unknown Word2Vec and helps to benchmark new semantic tools built from word vectors.
Tspoons: Tracking Salience Profiles Of Online News Stories, Kimberly Laurel Paterson
Tspoons: Tracking Salience Profiles Of Online News Stories, Kimberly Laurel Paterson
Master's Theses
News space is a relatively nebulous term that describes the general discourse concerning events that affect the populace. Past research has focused on qualitatively analyzing news space in an attempt to answer big questions about how the populace relates to the news and how they respond to it. We want to ask when do stories begin? What stories stand out among the noise? In order to answer the big questions about news space, we need to track the course of individual stories in the news. By analyzing the specific articles that comprise stories, we can synthesize the information gained from …
A System For Natural Language Unmarked Clausal Transformations In Text-To-Text Applications, Daniel Miller
A System For Natural Language Unmarked Clausal Transformations In Text-To-Text Applications, Daniel Miller
Master's Theses
A system is proposed which separates clauses from complex sentences into simpler stand-alone sentences. This is useful as an initial step on raw text, where the resulting processed text may be fed into text-to-text applications such as Automatic Summarization, Question Answering, and Machine Translation, where complex sentences are difficult to process. Grammatical natural language transformations provide a possible method to simplify complex sentences to enhance the results of text-to-text applications. Using shallow parsing, this system improves the performance of existing systems to identify and separate marked and unmarked embedded clauses in complex sentence structure resulting in syntactically simplified source for …