Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Physical Sciences and Mathematics (98)
- Computer Sciences (92)
- Arts and Humanities (47)
- Artificial Intelligence and Robotics (43)
- Discourse and Text Linguistics (26)
-
- Communication (25)
- Engineering (23)
- Semantics and Pragmatics (23)
- Psychology (22)
- Library and Information Science (20)
- Computer Engineering (19)
- Applied Linguistics (18)
- Psycholinguistics and Neurolinguistics (17)
- Phonetics and Phonology (15)
- First and Second Language Acquisition (13)
- Language Description and Documentation (13)
- Communication Technology and New Media (12)
- Anthropological Linguistics and Sociolinguistics (11)
- Cognition and Perception (11)
- Data Science (11)
- Medicine and Health Sciences (11)
- Databases and Information Systems (10)
- Other Computer Sciences (10)
- Syntax (10)
- Digital Humanities (9)
- Electrical and Computer Engineering (9)
- Social Media (9)
- Institution
-
- City University of New York (CUNY) (67)
- Technological University Dublin (30)
- Montclair State University (18)
- University of Kentucky (16)
- Brigham Young University (10)
-
- Chapman University (7)
- Dartmouth College (5)
- University of Nebraska - Lincoln (5)
- Western University (5)
- East Tennessee State University (4)
- University of Malaya (4)
- Air Force Institute of Technology (3)
- California Polytechnic State University, San Luis Obispo (3)
- Hunan Provincial Institute of Scientific and Technology Information (3)
- Minnesota State University, Mankato (3)
- Old Dominion University (3)
- Portland State University (3)
- San Jose State University (3)
- University of Central Florida (3)
- Binghamton University (2)
- Boise State University (2)
- Claremont Colleges (2)
- New Jersey Institute of Technology (2)
- Purdue University (2)
- The University of Akron (2)
- University of Louisville (2)
- Ursinus College (2)
- Valparaiso University (2)
- Bellarmine University (1)
- Bowling Green State University (1)
- Keyword
-
- Natural Language Processing (19)
- Computational linguistics (18)
- Natural language processing (16)
- Machine learning (12)
- NLP (10)
-
- Computational Linguistics (9)
- Machine Learning (7)
- Corpus linguistics (6)
- Deep learning (6)
- Dialogue (6)
- WordNet (6)
- BERT (5)
- Linguistics (5)
- Prepositions (5)
- Social media (5)
- Text classification (5)
- AI (4)
- Artificial intelligence (4)
- Digital humanities (4)
- Information retrieval (4)
- Natural language processing (Computer science) (4)
- Prosody (4)
- Situated Dialog (4)
- Sociolinguistics (4)
- Spatial Language (4)
- Spatial Templates (4)
- Twitter (4)
- Artificial Intelligence (3)
- Authorship attribution (3)
- Automatic Speech Recognition (3)
- Publication Year
- Publication
-
- Dissertations, Theses, and Capstone Projects (57)
- Conference papers (18)
- Department of Computer Science Faculty Scholarship and Creative Works (13)
- Faculty Publications (11)
- Theses and Dissertations--Linguistics (10)
-
- Publications and Research (9)
- Articles (6)
- Communication Sciences and Disorders Faculty Articles and Research (4)
- Conference Papers (4)
- Department of Linguistics Faculty Scholarship and Creative Works (4)
- Theses and Dissertations (4)
- Data and Test Instruments (3)
- Dissertations (3)
- Electronic Theses and Dissertations (3)
- Journal of Scientific Information Research (3)
- Student Works (2020-2029) (3)
- CGU Faculty Publications and Research (2)
- Commonwealth Computational Summit (2)
- Computer Science Summer Fellows (2)
- Electrical & Computer Engineering Theses & Dissertations (2)
- Electronic Literature Organization Conference 2020 (2)
- Journal of Tolkien Research (2)
- LING 590/Internet Language (2)
- Library Philosophy and Practice (e-journal) (2)
- Master's Theses (2)
- Masters Theses (2)
- Northeast Journal of Complex Systems (NEJCS) (2)
- Other Resources (2)
- Proceedings from the Document Academy (2)
- School of Computing: Conference and Workshop Papers (2)
- Publication Type
- File Type
Articles 31 - 60 of 250
Full-Text Articles in Computational Linguistics
Expanding The Corpus Of Vocalized Hebrew Text: Compiling An Unvocalized Text Corpus And Building An Online Interface For Vocalization Annotation, Rachel Shanblatt Bloch
Expanding The Corpus Of Vocalized Hebrew Text: Compiling An Unvocalized Text Corpus And Building An Online Interface For Vocalization Annotation, Rachel Shanblatt Bloch
Dissertations, Theses, and Capstone Projects
Written modern Hebrew presents a unique challenge for training computational models for language processing because modern Hebrew text often lacks vocalization. The lack of available vocalized Hebrew data can lead to ambiguity in training these models and generally hinders work on natural language processing problems. The goal of this project is to contribute to the collection of vocalized Hebrew text by collecting and preprocessing a large corpus of unvocalized Hebrew text and building an online annotation tool. The annotation tool allows people to upload unvocalized Hebrew text, to annotate by adding Hebrew vocalization, and to download comma-separated values files of …
Gpt Assisted Annotation Of Rhetorical And Linguistic Features For Interpretable Propaganda Technique Detection In News Text., Kyle Hamilton, Bojan Bozic, Luca Longo
Gpt Assisted Annotation Of Rhetorical And Linguistic Features For Interpretable Propaganda Technique Detection In News Text., Kyle Hamilton, Bojan Bozic, Luca Longo
Articles
While the use of machine learning for the detection of propaganda techniques in text has garnered considerable attention, most approaches focus on "black-box'' solutions with opaque inner workings. Interpretable approaches provide a solution, however, they depend on careful feature engineering and costly expert annotated data. Additionally, language features specific to propagandistic text are generally the focus of rhetoricians or linguists, and there is no data set labeled with such features suitable for machine learning. This study codifies 22 rhetorical and linguistic features identified in literature related to the language of persuasion for the purpose of annotating an existing data set …
Exploring Asynchronous Pronunciation Training Through Context-Aware Pronunciation Applications, Claire L. Schweikert
Exploring Asynchronous Pronunciation Training Through Context-Aware Pronunciation Applications, Claire L. Schweikert
Theses/Capstones/Creative Projects
This paper provides a survey of various research articles on context-aware asynchronous pronunciation training applications. First, a set of seven articles is reviewed and summarized. Next, they are synthesized over the three main topics of 1) automated speech recognition, 2) non-native speaker considerations in language learning, and 3) future directions for research and development within computer-assisted pronunciation training (CAPT). Research in the areas of acoustic and pronunciation modeling (both implicit and explicit), pedagogical considerations for CAPT application design, Goodness of Pronunciation algorithm scoring, accent recognition and neutralization, and more are discussed.
Bugsy The Bee's Big Adventure, Jonah L., Zachery Irvin, Jake Wildstrom, Zachariah Robinson, Hilaria Cruz
Bugsy The Bee's Big Adventure, Jonah L., Zachery Irvin, Jake Wildstrom, Zachariah Robinson, Hilaria Cruz
LING 590/Internet Language
Bugsy the bee goes on a quest for pollen.
Skyler's Lunch, Noah Sherman, Autumn Boone, Hilaria Cruz
Skyler's Lunch, Noah Sherman, Autumn Boone, Hilaria Cruz
LING 590/Internet Language
Our class was studying the use of emojis across different platforms and wanted to explore how stories using emojis could impact young readers. Here, we try to translate the story of Skyler into emoji, providing translations along the way. We replace words completely with emoji, represent phrases with a few emoji, and use additional emoji to make sense of the content, including punctuation. In this book, we explore the character of Skyler, who is a picky eater. But they learn to eat the nutritious food that is good for them. In the end, they even get a reward!
Research On The Application Of Part-Of-Speech Tagging Of Ancient Books Under The Domain Large Language Model, Danhao Zhu, Zhao Zhixiao, Die Hu, Wenhua Zhao
Research On The Application Of Part-Of-Speech Tagging Of Ancient Books Under The Domain Large Language Model, Danhao Zhu, Zhao Zhixiao, Die Hu, Wenhua Zhao
Journal of Scientific Information Research
[Purpose/significance]The development of the large language model has brought new ideas for ancient text mining, and combining the large language model with the digitisation and intelligence of ancient books is a necessary path for the work of ancient books in the new era. [Methods/process]This paper uses the lexically annotated corpus of Zuozhuan to construct a batch of high-quality lexically annotated instruction data through data cleaning and preprocessing, on the basis of which 500, 1 000, 2 000, and 5 000 pieces of data are used to fine-tune the instructions of the large language model, and the performance test is carried …
Retórica Intercultural En El Discurso Académico Universitario: Las Funciones Retóricas De La Citación En Los Trabajos De Fin De Máster Escritos En Español Y En Inglés Por Hablantes Nativos Y No Nativos, David Sanchez-Jimenez
Retórica Intercultural En El Discurso Académico Universitario: Las Funciones Retóricas De La Citación En Los Trabajos De Fin De Máster Escritos En Español Y En Inglés Por Hablantes Nativos Y No Nativos, David Sanchez-Jimenez
Publications and Research
This research derives from the interest in learning the cultural differences in citation practices in the academic genre of Master's thesis of native Spanish (Ee), non-native Filipino writers of Spanish (Fe), native Filipino writers of English (Fi), and American writers of English. A total of thirty-two (32) master´s theses – eight (8) for each group – were analyzed. A quantitative and qualitative methodology was used to study this phenomenon based on the computerized textual analysis of the rhetorical function of citations arranged in typological classification that modified the outline proposed by Petrić in his 2007 article. The results obtained from …
Consonant (De)Gradation In Ingrian?, Andrea M. Harrison
Consonant (De)Gradation In Ingrian?, Andrea M. Harrison
Dissertations, Theses, and Capstone Projects
This paper will present a dual method toward data enrichment for low-resource languages. Using Yoyodyne -- a Fairseq-inspired neural library for small-vocabulary sequence-to-sequence generation -- a morphological generation task was tested across labeled data encompassing multiple stages of enrichment for the low-resource language Ingrian. Due to limitations in the available data for Ingrian, weighted finite-state transducers (WFSTs) were used to generate an expanded vocabulary via HFST's toolkit for Uralic languages, and GiellaLT, a source for FST-driven lexica for low-resource languages. Further stages of experimentation used labeled data from related, higher-resource languages (Finnish, Estonian) to encourage cross-lingual transfer in the interest …
How Do We Learn What We Cannot Say?, Daniel Yakubov
How Do We Learn What We Cannot Say?, Daniel Yakubov
Dissertations, Theses, and Capstone Projects
The contributions of this thesis are two-fold. First, this thesis presents UDTube, an easily usable software developed to perform morphological analysis in a multi-task fashion. This work shows the strong performance of UDTube versus the current state-of-the-art, UDPipe, across eight languages, primarily in the annotation of morphological features. The second contribution of this thesis is a exploration into the study of defectivity. UDTube is used to annotate a large amount of data in Greek and Russian which is ultimately used to investigate the plausibility of Indirect Negative Evidence (INE), a popular approach to the acquisition of morphological defectivity. The reported …
The Ring Cycle: Journeying Through The Language Of Tolkien’S Third Age With Corpus Linguistics, Michael Livesey
The Ring Cycle: Journeying Through The Language Of Tolkien’S Third Age With Corpus Linguistics, Michael Livesey
Journal of Tolkien Research
This article explores the journey taken by the One Ring across J.R.R. Tolkien’s Third Age writings. It employs a digital humanities approach to analyse linguistic patterns in Tolkien’s use of the word ring, across The Hobbit and The Lord of the Rings. Specifically, the article employs corpus linguistic methods to track shifts in the quantities and qualities of the Ring’s appearance across these texts. It uses techniques of keyness and collocation analysis to trace transformations in these quantities/qualities, including: a) the Ring’s transition from a central to a peripheral place in the Third Age’s narrative arc; and b) …
Using Ai For Qualitative Labeling: Consistency And Comparisons, James Temple
Using Ai For Qualitative Labeling: Consistency And Comparisons, James Temple
Honors Program Theses
This paper continues research that evaluates the capacity of artificial intelligence (AI) to perform qualitative coding tasks. The previous study found that AI models lacked consistency with themselves and did not agree with human coded data. Since that study, AI’s general level of intelligence has increased. Hence, this study re-evaluates how well the newest set of AI models (Claude 3 and Gemini) can perform qualitative coding tasks. When tested, the new AI models perform about the same or better than previous models depending on the metric tested. While Gemini and Claude 3 do not agree with human output any more …
Linguistic Inquiry And Word Count (Liwc) For Text Analysis, Isabella Lenzini, Destiny Fore Msw, Anna W. Wright Phd
Linguistic Inquiry And Word Count (Liwc) For Text Analysis, Isabella Lenzini, Destiny Fore Msw, Anna W. Wright Phd
IRBEH/Spit for Science Publications and Presentations
No abstract provided.
Streamlining Public Engagement In Transportation Projects Using Text Analytics, Alireza Shamshiri
Streamlining Public Engagement In Transportation Projects Using Text Analytics, Alireza Shamshiri
Civil Engineering Dissertations - Archive
Infrastructure projects impact a broad range of stakeholders, particularly local communities, whose engagement is critical for successful outcomes. Despite the importance of public engagement in these projects, traditional methods of capturing and analyzing public opinion often fail to fully represent the diverse, genuine perspectives involved. This has led to conflicts between community members and project sponsors. On the other hand, despite advancements in text analytics, including natural language processing (NLP) and its subfields such as topic modeling, sentiment analysis, and neural networks, their functionalities and effectiveness in analyzing public opinion in the domain of infrastructure projects have not been fully …
A Computational Investigation Of English Spelling, John Winstead
A Computational Investigation Of English Spelling, John Winstead
Theses and Dissertations--Linguistics
This thesis examines the predictability and regularity of English orthography through computational methods. The primary objective is to use n-gram models to predict missing letters in English words by exploiting contextual information from adjacent letters. The study evaluates the impact of dataset size, word length, letter position, and vowel presence on the predictive accuracy of these models, uncovering patterns and structures inherent to English spelling.
The research utilizes a range of datasets, including the Carnegie Mellon University Pronouncing Dictionary, the Brown Corpus, the Corpus of Late Modern English Texts, the Lampeter Corpus of Early Modern English Tracts, and the Open …
A Computer-Assisted Approach To Lexical Borrowing In Northeast Caucasian Languages, Bonnie Eleanor Wren-Hardin
A Computer-Assisted Approach To Lexical Borrowing In Northeast Caucasian Languages, Bonnie Eleanor Wren-Hardin
Theses and Dissertations--Linguistics
The disambiguation of loanwords and cognates can be a challenge, especially in areas where there has been intense language contact over an extended period of time, when the contact is between genetically related languages, and when the number of languages involved is large Over the past several decades, more and more computational approaches to automatic cognate and borrowing detection have been created in an attempt to ease the load of examining hundreds to thousands of individual lexemes, as well as determine language family relationships with allegedly greater accuracy. While these methods are not perfect and cannot replace the knowledge or …
Guilty Machines: On Ab-Sens In The Age Of Ai, Dylan Lackey, Katherine Weinschenk
Guilty Machines: On Ab-Sens In The Age Of Ai, Dylan Lackey, Katherine Weinschenk
Critical Humanities
For Lacan, guilt arises in the sublimation of ab-sens (non-sense) into the symbolic comprehension of sen-absexe (sense without sex, sense in the deficiency of sexual relation), or in the maturation of language to sensibility through the effacement of sex. Though, as Slavoj Žižek himself points out in a recent article regarding ChatGPT, the split subject always misapprehends the true reason for guilt’s manifestation, such guilt at best provides a sort of evidence for the inclusion of the subject in the order of language, acting as a necessary, even enjoyable mark of the subject’s coherence (or, more importantly, the subject’s separation …
Executive Order On The Safe, Secure, And Trustworthy Development And Use Of Artificial Intelligence, Joseph R. Biden
Executive Order On The Safe, Secure, And Trustworthy Development And Use Of Artificial Intelligence, Joseph R. Biden
Copyright, Fair Use, Scholarly Communication, etc.
Section 1. Purpose. Artificial intelligence (AI) holds extraordinary potential for both promise and peril. Responsible AI use has the potential to help solve urgent challenges while making our world more prosperous, productive, innovative, and secure. At the same time, irresponsible use could exacerbate societal harms such as fraud, discrimination, bias, and disinformation; displace and disempower workers; stifle competition; and pose risks to national security. Harnessing AI for good and realizing its myriad benefits requires mitigating its substantial risks. This endeavor demands a society-wide effort that includes government, the private sector, academia, and civil society.
My Administration places the highest urgency …
Towards Interpretable Machine Reading Comprehension With Mixed Effects Regression And Exploratory Prompt Analysis, Luca Del Signore
Towards Interpretable Machine Reading Comprehension With Mixed Effects Regression And Exploratory Prompt Analysis, Luca Del Signore
Dissertations, Theses, and Capstone Projects
We investigate the properties of natural language prompts that determine their difficulty in machine reading comprehension tasks. While much work has been done benchmarking language model performance at the task level, there is considerably less literature focused on how individual task items can contribute to interpretable evaluations of natural language understanding. Such work is essential to deepening our understanding of language models and ensuring their responsible use as a key tool in human machine communication. We perform an in depth mixed effects analysis on the behavior of three major generative language models, comparing their performance on a large reading comprehension …
A Computational Analysis Of Volodymyr Zelenskyy's Public Diplomacy Discourse In Times Of Crisis, Amber Brittain-Hale, Amber Brittain-Hale
A Computational Analysis Of Volodymyr Zelenskyy's Public Diplomacy Discourse In Times Of Crisis, Amber Brittain-Hale, Amber Brittain-Hale
Education Division Scholarship
In this study, we delve into the public diplomacy discourse of Ukrainian President Volodymyr Zelenskyy during the ongoing crisis of the Russo-Ukrainian War. We aim to conduct a computational analysis of Zelenskyy's English, Russian, and Ukrainian speeches, exploring the linguistic patterns and code-switching employed in his discourse. The study period encompasses Russia’s build-up to and full-scale invasion of Ukraine from May 2019 to May 30, 2023. This time frame is crucial as it captures the dynamic development of the crisis and the expansion of Zelenskyy's presidency, providing a unique context for analyzing his public diplomacy efforts. By utilizing Linguistic Inquiry …
Ideology Prediction From Scarce And Biased Supervision: Learn To Disregard The “What” And Focus On The “How”!, Chen Chen, Dylan Walker, Venkatesh Saligrama
Ideology Prediction From Scarce And Biased Supervision: Learn To Disregard The “What” And Focus On The “How”!, Chen Chen, Dylan Walker, Venkatesh Saligrama
Business Faculty Articles and Research
We propose a novel supervised learning approach for political ideology prediction (PIP) that is capable of predicting out-of-distribution inputs. This problem is motivated by the fact that manual data-labeling is expensive, while self-reported labels are often scarce and exhibit significant selection bias. We propose a novel statistical model that decomposes the document embeddings into a linear superposition of two vectors; a latent neutral context vector independent of ideology, and a latent position vector aligned with ideology. We train an end-to-end model that has intermediate contextual and positional vectors as outputs. At deployment time, our model predicts labels for input documents …
Destined Failure, Chengjun Pan
Destined Failure, Chengjun Pan
Masters Theses
I attempt to examine the complex structure of human communication, explaining why it is bound to fail. By reproducing experienceable phenomena, I demonstrate how they can expose communication structure and reveal the limitations of our perception and symbolization.I divide the process of communication into six stages: input, detection, symbolization, dictionary, interpretation, and output. In this thesis, I examine the flaws and challenges that arise in the first five stages. I argue that reception acts as a filter and that understanding relies on a symbolic system that is full of redundancies. Therefore, every interpretation is destined to be a deviation.
The Sociolinguistics Of Code-Switching In Hong Kong’S Digital Landscape: A Mixed-Methods Exploration Of Cantonese-English Alternation Patterns On Whatsapp, Wilkinson Daniel Wong Gonzales, Yuen Man Tsang
The Sociolinguistics Of Code-Switching In Hong Kong’S Digital Landscape: A Mixed-Methods Exploration Of Cantonese-English Alternation Patterns On Whatsapp, Wilkinson Daniel Wong Gonzales, Yuen Man Tsang
Journal of English and Applied Linguistics
This paper examines the prevalence of Cantonese-English code mixing in Hong Kong through an under-researched digital medium. Prior research on this code-alternation practice has often been limited to exploring either the social or linguistic constraints of code-switching in spoken or written communication. Our study takes a holistic approach to analyzing code-switching in a hybrid medium that exhibits features of both spoken and written discourse. We specifically analyze the code-switching patterns of 24 undergraduates from a Hong Kong university on WhatsApp and examine how both social and linguistic factors potentially constrain these patterns. Utilizing a self-compiled sociolinguistic corpus as well as …
Neural Network Vs. Rule-Based G2p: A Hybrid Approach To Stress Prediction And Related Vowel Reduction In Bulgarian, Maria Karamihaylova
Neural Network Vs. Rule-Based G2p: A Hybrid Approach To Stress Prediction And Related Vowel Reduction In Bulgarian, Maria Karamihaylova
Dissertations, Theses, and Capstone Projects
An effective grapheme-to-phoneme (G2P) conversion system is a critical element of speech synthesis. Rule-based systems were an early method for G2P conversion. In recent years, machine learning tools have been shown to outperform rule-based approaches in G2P tasks. We investigate neural network sequence-to-sequence modeling for the prediction of syllable stress and resulting vowel reductions in the Bulgarian language. We then develop a hybrid G2P approach which combines manually written grapheme-to-phoneme mapping rules with neural network-enabled syllable stress predictions by inserting stress markers in the predicted stress position of the transcription produced by the rule-based finite-state transducer. Finally, we apply vowel …
Evaluating Neural Networks As Cognitive Models For Learning Quasi-Regularities In Language, Xiaomeng Ma
Evaluating Neural Networks As Cognitive Models For Learning Quasi-Regularities In Language, Xiaomeng Ma
Dissertations, Theses, and Capstone Projects
Many aspects of language can be categorized as quasi-regular: the relationship between the inputs and outputs is systematic but allows many exceptions. Common domains that contain quasi-regularity include morphological inflection and grapheme-phoneme mapping. How humans process quasi-regularity has been debated for decades. This thesis implemented modern neural network models, transformer models, on two tasks: English past tense inflection and Chinese character naming, to investigate how transformer models perform quasi-regularity tasks. This thesis focuses on investigating to what extent the models' performances can represent human behavior. The results show that the transformers' performance is very similar to human behavior in many …
Topics For He But Not For She: Quantifying And Classifying Gender Bias In The Media, Tyler J. Lanni
Topics For He But Not For She: Quantifying And Classifying Gender Bias In The Media, Tyler J. Lanni
Dissertations, Theses, and Capstone Projects
In this study, we used computational techniques to analyze the language used in news articles to describe female and male politicians. Our corpus included 370 subtexts for male candidates and 374 subtexts for female candidates, gathered through the New York Times API. We conducted two experiments: an LDA topic analysis to explore the data, and a logistic regression to classify the subtexts as either male or female. Our analysis revealed some noteworthy findings that suggest the possibility of developing a gender bias classifier in the future. However, to create a more robust understanding of bias, additional research and data are …
Ai Approaches To Understand Human Deceptions, Perceptions, And Perspectives In Social Media, Chih-Yuan Li
Ai Approaches To Understand Human Deceptions, Perceptions, And Perspectives In Social Media, Chih-Yuan Li
Dissertations
Social media platforms have created virtual space for sharing user generated information, connecting, and interacting among users. However, there are research and societal challenges: 1) The users are generating and sharing the disinformation 2) It is difficult to understand citizens' perceptions or opinions expressed on wide variety of topics; and 3) There are overloaded information and echo chamber problems without overall understanding of the different perspectives taken by different people or groups.
This dissertation addresses these three research challenges with advanced AI and Machine Learning approaches. To address the fake news, as deceptions on the facts, this dissertation presents Machine …
Predicting High-Cap Tech Stock Polarity: A Combined Approach Using Support Vector Machines And Bidirectional Encoders From Transformers, Ian L. Grisham
Predicting High-Cap Tech Stock Polarity: A Combined Approach Using Support Vector Machines And Bidirectional Encoders From Transformers, Ian L. Grisham
Electronic Theses and Dissertations
The abundance, accessibility, and scale of data have engendered an era where machine learning can quickly and accurately solve complex problems, identify complicated patterns, and uncover intricate trends. One research area where many have applied these techniques is the stock market. Yet, financial domains are influenced by many factors and are notoriously difficult to predict due to their volatile and multivariate behavior. However, the literature indicates that public sentiment data may exhibit significant predictive qualities and improve a model’s ability to predict intricate trends. In this study, momentum SVM classification accuracy was compared between datasets that did and did not …
Improving Sign Recognition With Phonology, Lee Kezar, Jesse Thomason, Zed Sevcikova Sehyr
Improving Sign Recognition With Phonology, Lee Kezar, Jesse Thomason, Zed Sevcikova Sehyr
Communication Sciences and Disorders Faculty Articles and Research
We use insights from research on American Sign Language (ASL) phonology to train models for isolated sign language recognition (ISLR), a step towards automatic sign language understanding. Our key insight is to explicitly recognize the role of phonology in sign production to achieve more accurate ISLR than existing work which does not consider sign language phonology. We train ISLR models that take in pose estimations of a signer producing a single sign to predict not only the sign but additionally its phonological characteristics, such as the handshape. These auxiliary predictions lead to a nearly 9% absolute gain in sign recognition …
Content-Based Unsupervised Fake News Detection On Ukraine-Russia War, Yucheol Shin, Yvan Sojdehei, Limin Zheng, Brad Blanchard
Content-Based Unsupervised Fake News Detection On Ukraine-Russia War, Yucheol Shin, Yvan Sojdehei, Limin Zheng, Brad Blanchard
SMU Data Science Review
The Ukrainian-Russian war has garnered significant attention worldwide, with fake news obstructing the formation of public opinion and disseminating false information. This scholarly paper explores the use of unsupervised learning methods and the Bidirectional Encoder Representations from Transformers (BERT) to detect fake news in news articles from various sources. BERT topic modeling is applied to cluster news articles by their respective topics, followed by summarization to measure the similarity scores. The hypothesis posits that topics with larger variances are more likely to contain fake news. The proposed method was evaluated using a dataset of approximately 1000 labeled news articles related …
Research On Topic Discovery And Evolution Trend Based On Temporal Keyword Characteristics Analysis, Shuqing Li, Juntao Zhu, Wan Wang
Research On Topic Discovery And Evolution Trend Based On Temporal Keyword Characteristics Analysis, Shuqing Li, Juntao Zhu, Wan Wang
Journal of Scientific Information Research
[Purpose/significance]Excavating the research topics in a large number of articles, sorting out the evolution context and correlation of the research topics, predicting the frontier hot spots of the topics can be helpful to enhance the scientificity and vividness of the evolution results.[Method/precess]This paper puts forward the concept of time series influence factor as an important feature in keyword extraction, uses the method of time window to mine and identify topics by using topic model, and makes visual analysis. By applying time series model in the field of deep learning, the purpose of predicting topic popularity is achieved.[Result/concluson]It is verified that …