An Analytical Review Of Preprocessing Techniques In Bengali Natural Language Processing,
2025
United International University
An Analytical Review Of Preprocessing Techniques In Bengali Natural Language Processing, Sovon Chakraborty, Protiva Das, Shakib Mahmud Dipto, Md Aktaruzzaman Pramanik, Jannatun Noor
Computer Science Faculty Publications
Research in Bengali Natural Language Processing (BNLP) is rapidly expanding. Despite being one of the most widely spoken languages in the world, BNLP research remains insufficient, particularly in Bengali speech recognition. The languages rich morphology, agglutinative structure, and diverse dialects make text and speech processing especially challenging. However, these challenges can be addressed with effective preprocessing techniques. Various organizations in Bangladesh and West Bengal are integrating Natural Language Processing (NLP) into their services, but without a thorough understanding of preprocessing, these implementations remain incomplete. Applying proper preprocessing techniques to the Bengali language will serve as a foundation for developing robust …
Can Llms Beat Humans On Discerning Human-Written And Llm-Generated Science News,
2025
Old Dominion University
Can Llms Beat Humans On Discerning Human-Written And Llm-Generated Science News, Dominik Soós, Meng Jiang, Jian Wu
Computer Science Faculty Publications
Science news is increasingly important in connecting scientists and the public by sharing discoveries and innovations. With the rise of large language models (LLMs), there is potential to automate science news creation, but concerns exist about the quality of LLM-generated news versus human-written news. This paper explores whether LLMs can outperform humans in distinguishing between human-written and LLM-generated news. Inspired by the Chain-of-Thought prompting method, we designed a simple yet effective variant called Guided Few-shot (GFS), which encodes the characteristics of news of two types with examples. Our experiments indicated that GFS with just a single example effectively boosted the …
Deepssetracer 2.0: Improved Deep Learning Model Performance For Protein Secondary Structure Segmentation From Cryo-Em Maps,
2025
Old Dominion University
Deepssetracer 2.0: Improved Deep Learning Model Performance For Protein Secondary Structure Segmentation From Cryo-Em Maps, Bryan Hawickhorst, Thu Nguyen, Willy Wriggers, Jiangwen Sun, Jing He
Computer Science Faculty Publications
DeepSSETracer is a method for segmenting protein secondary structure from medium-resolution (5-10Å) cryogenic electron microscopy (cryo-EM) density maps. We conducted experiments and ablation studies to examine the effects of normalization methods, max-pooling, activation functions, and loss calculation region on DeepSSETracer. By combining multiple technical improvements, the performance of the new version, DeepSSETracer 2.0, was significantly enhanced compared to DeepSSETracer 1.1. On a set of 77 test cases, the weighted average per-voxel F1 score increased from 62.1% to 70.3% for helix detection, and from 47.8% to 62.5% for β-sheet detection. While each of the five modifications in the network enhanced the …
Benchmarking And Improving Foundation Model Dietary Estimates From Meal Images,
2025
Old Dominion University
Benchmarking And Improving Foundation Model Dietary Estimates From Meal Images, Yongcheng Mu, Jiangwen Sun, Jing He
Computer Science Faculty Publications
Accurate quantifying dietary contents, such as calories, proteins, carbohydrates, and fats, from an image of a meal plate is vital for managing diabetes. Recently, Large Multimodal Models (LMMs) have excelled in complex vision-language tasks due to their use of very large, highly diverse data. This study benchmarked the use of seven LMMs that include full and lightweight models of GPT, Gemini, and Llama for nutrition estimation based on Google's Nutrition5k dataset and our own phone-collected DonateAndLearn dataset. We analyzed the performance of LMMs and the RGB-D fusion model, in which the RGB-D model was specifically trained using Nutrition5k data. On …
Understanding Pii Leakage In Large Language Models: A Systematic Survey,
2025
Zhejiang University
Understanding Pii Leakage In Large Language Models: A Systematic Survey, Shuai Cheng, Zhao Li, Shu Meng, Mengxia Ren, Haitao Xu, Shuai Hao, Chuan Yue, Fang Zhang
Computer Science Faculty Publications
Large Language Models (LLMs) have demonstrated exceptional success across a variety of tasks, particularly in natural language processing, leading to their growing integration into numerous facets of daily life. However, this widespread deployment has raised substantial privacy concerns, especially regarding personally identifiable information (PII), which can be directly associated with specific individuals. The leakage of such information presents significant real-world privacy threats. In this paper, we conduct a systematic investigation into existing research on PII leakage in LLMs, encompassing commonly utilized PII datasets, evaluation metrics, and current studies on both PII leakage attacks and defensive strategies. Finally, we identify unresolved …
The Reliability Response To Patent Law’S Ai Challenges,
2025
Duke Law School
The Reliability Response To Patent Law’S Ai Challenges, Arti K. Rai
Faculty Scholarship
Pervasive AI use adds newfound importance to longstanding debates over patent timing and reliability. Patent claims on speculative ideas generated by AI, or even the infusion of speculative AI-generated ideas into the public domain, may defeat patent incentives for more careful research. Although challenges that AI use poses for patent validity requirements like human inventorship and nonobviousness have received more attention, reliability is equally important.
Indeed, as this Article argues, the issues are linked. If requirements for inventorship and nonobviousness were adjusted to emphasize reliability, a human role could be preserved, and AI use would not necessarily threaten patents. Currently, …
Multi-Lingual And Cross-Domain Frontiers In Machine-Generated Content Detection,
2025
Wilfrid Laurier University
Multi-Lingual And Cross-Domain Frontiers In Machine-Generated Content Detection, Gurunameh Singh Chhatwal
Theses and Dissertations (Comprehensive)
The rapid advancement of generative artificial intelligence, particularly Large Language Models (LLMs) such as GPT-4 and their multilingual capabilities, has significantly blurred the distinction between human-authored and machine-generated content. This technological evolution introduces critical challenges concerning the detection and attribution of textual authenticity and authorship, exacerbating societal issues like misinformation proliferation and compromising academic and professional integrity. Traditional detection methodologies, predominantly monolingual and heuristic-based, have demonstrated inadequate generalizability and efficacy against the sophisticated, multilingual capabilities of contemporary generative models.
This thesis addresses two major problems arising from these advancements. Firstly, it introduces novel multilingual detection methodologies explicitly designed to differentiate …
A New Deepfake Detection Method With No-Reference Image Quality Assessment To Resist Image Degradation,
2025
Old Dominion University
A New Deepfake Detection Method With No-Reference Image Quality Assessment To Resist Image Degradation, Jiajun Jiang, Wen-Chao Yang, Chung-Hao Chen, Timothy Young
Electrical & Computer Engineering Faculty Publications
Deepfake technology, which utilizes advanced AI models such as Generative Adversarial Networks (GANs), has led to the proliferation of highly convincing manipulated media, posing significant challenges for detection. Existing detection methods often struggle with the low-quality or compressed press, which is prevalent on social media platforms. This paper proposes a novel Deepfake detection framework that leverages No-Reference Image Quality Assessment (NRIQA) techniques, specifically, BRISQUE, NIQE, and PIQUE, to extract quality-related features from facial images. These features are then classified using a Support Vector Machine (SVM) with various kernel functions. We evaluate our method under both intra-dataset and cross-dataset settings. For …
Key Brain Region Identification In Obesity Prediction With Structural Mri And Probabilistic Uncertainty Aware Model,
2025
Old Dominion University
Key Brain Region Identification In Obesity Prediction With Structural Mri And Probabilistic Uncertainty Aware Model, Walia Farzana, Megan A. Witherow, Ahmed Temtam, Liangsuo Ma, Melanie Bean, F. Gerry Moeller, K. M. Iftekharuddin
Electrical & Computer Engineering Faculty Publications
Objectives/Goals: Predictive performance alone may not determine a model’s clinical utility. Neurobiological changes in obesity alter brain structures, but traditional voxel-based morphometry is limited to group-level analysis. We propose a probabilistic model with uncertainty heatmaps to improve interpretability and personalized prediction. Methods/Study Population: The data for this study are sourced from the Human Connectome Project (HCP), with approval from the Washington University in St. Louis Institutional Review Board. We preprocessed raw T1-weighted structural MRI scans from 525 patients using an automated pipeline. The dataset is divided into training (357 cases), calibration (63 cases), and testing (105 cases). Our probabilistic model …
A Fast Framework For Generating Radioactive Mixture Spectra And Its Application To Remote High-Performance Mixture Identification,
2025
Applied Research LLC
A Fast Framework For Generating Radioactive Mixture Spectra And Its Application To Remote High-Performance Mixture Identification, Chiman Kwan, Bulent Ayhan, Adam Stavola, Kazi Aminul Islam, Hongfang Zhang, Jiang Li
Electrical & Computer Engineering Faculty Publications
Remote detection of radioactive materials in mixtures using handheld or portal detectors remains a challenge because of factors such as low concentration, environmental interference, sensor noise, and other complications. This work introduces a fast framework for generating realistic mixture spectra. Moreover, we present mixture isotope identification using data generated by the fast framework. Researchers have examined a range of conventional and recent algorithms within the fields of machine learning and deep learning. An application to uranium enrichment-level prediction has been included. Extensive simulation experiments validated the efficacy of the proposed framework.
Adaptive Fusion Neural Networks For Sparse-Angle X-Ray 3d Reconstruction,
2025
Guangzhou Huashang College
Adaptive Fusion Neural Networks For Sparse-Angle X-Ray 3d Reconstruction, Shaoyong Hong, Bo Yang, Yan Chen, Hao Quan, Shan Liu, Minyi Tang, Jiawei Tian
Electrical & Computer Engineering Faculty Publications
3D medical image reconstruction has significantly enhanced diagnostic accuracy, yet the reliance on densely sampled projection data remains a major limitation in clinical practice. Sparse-angle X-ray imaging, though safer and faster, poses challenges for accurate volumetric reconstruction due to limited spatial information. This study proposes a 3D reconstruction neural network based on adaptive weight fusion (AdapFusionNet) to achieve high-quality 3D medical image reconstruction from sparse-angle X-ray images. To address the issue of spatial inconsistency in multi-angle image reconstruction, an innovative adaptive fusion module was designed to score initial reconstruction results during the inference stage and perform weighted fusion, thereby improving …
The Effect Of Facial Phenotypes On Differential Performance Of Facial Recognition,
2025
West Virginia University
The Effect Of Facial Phenotypes On Differential Performance Of Facial Recognition, Evan R. Garrett
Graduate Theses, Dissertations, and Problem Reports (ETD)
Facial recognition technology is utilized in many facets of life. As the use has become more widespread these systems have improved in reliability and performance approaching the level of human accuracy. With these improvements the problem of bias still remains as a persistent problem. Efforts have been made to minimize the bias prevalent in the systems via studies into various demographic factors, creating training datasets that have a more uniform distribution of subjects, and other methods. As facial recognition is one of the most utilized forms of biometric recognition it is vital to analyze potential causes of bias to help …
Embodied Ai For Challenging Rearrangement Tasks In The Context Of Service And Assistive Robots,
2025
Edith Cowan University
Embodied Ai For Challenging Rearrangement Tasks In The Context Of Service And Assistive Robots, Mariia Khan
Theses: Doctorates and Masters
Embodied AI explores intelligent agents that learn through interaction with their environment, aiming to replicate human-like learning processes. Achieving this requires agents capable of understanding a scene via various sensors, reasoning about their actions, and reacting accordingly. These abilities are necessary for service domestic robots to assist humans in their day-to-day activities. Embodied AI tasks can include but are not limited to: visual exploration, visual navigation, instruction following and embodied question answering, which typically consider static (unchanging) environments, where objects do not move over time. This thesis addresses one of the most challenging Embodied AI tasks – visual room rearrangement, …
The Future Of Ai Regulation In Drug Development: A Comparative Analysis,
2025
Duke Law School
The Future Of Ai Regulation In Drug Development: A Comparative Analysis, Gabriela Lenarczyk, Timo Minssen, Nicholson Price, Arti Rai
Faculty Scholarship
As artificial intelligence (AI) transforms drug development, regulatory frameworks are evolving to oversee its implementation, particularly at the US Food and Drug Administration (FDA) and the European Medicines Agency (EMA). This paper makes three contributions to understanding emerging regulatory approaches. First, we offer a comparative analysis of how these agencies have responded to AI-driven advances, incorporating new US executive orders and the European Union (EU)’s AI Act. Second, we propose a novel analytical framework to understand regulatory divergence: the FDA’s flexible, dialog-driven model contrasts with the EMA’s structured, risk-tiered approach, reflecting broader institutional and political-economic differences. While the former encourages …
Information Retrieval In The Age Of Generative Ai: A Mismatch That Matters,
2025
Duke Law School
Information Retrieval In The Age Of Generative Ai: A Mismatch That Matters, Alex Zhang
Faculty Scholarship
This short piece explores a widespread and yet underexamined or even overlooked misconception, that is, large language models (LLMs) function like traditional legal research databases. They do not. As a matter of fact, information retrieval from databases functions very differently from LLMs in terms of inputs, retrieval processes, and outputs. These differences have significant implications for transparency, traceability, and overall effectiveness in AI-driven legal research. Without intentional oversight and adaption, these changes could profoundly affect how we develop research skills and a cumulative knowledge base, both of which are essential skills for lifelong learning in the legal field.
This article …
Towards Human Explainable Digital Forensics: Generating Human Interpretable Evidence For Semantic Understanding In Manipulated Images And Text,
2025
University at Albany, State University of New York
Towards Human Explainable Digital Forensics: Generating Human Interpretable Evidence For Semantic Understanding In Manipulated Images And Text, Yuwei Chen
Electronic Theses & Dissertations (2024 - present)
Detecting and characterizing manipulations in digital media continues to pose a significant challenge within the field of digital forensics. Despite notable advancements, the discipline often remains in a reactive stance against emerging threats. Current state-of-the-art methods, typically evaluated within academic settings, fails to mirror the complexities of real-world disinformation scenarios. These methods generally prioritize high performance based on quantitative metrics, yet they demonstrate a considerable dependency on training data and lack adaptability to new novel attack signatures. With the rapid evolution of attack methodologies, the dependency on highly accurate models that do not generalize or adapt well to unseen threats …
Artificial Intelligence And Procedural Due Process,
2025
Duke Law School
Artificial Intelligence And Procedural Due Process, Brandon L. Garrett
Faculty Scholarship
Artificial intelligence (AI) violates procedural due process rights if the government uses it to deprive people of life, liberty, and property without adequate notice or an opportunity to be heard. A wide range of government agencies deploy AI systems, including in courts, law enforcement, public benefits administration, and national security. If the government refuses to disclose the reasons why it denied a person bail, public benefits, or immigration status, serious due process concerns arise. If the government delegates such tasks to an AI system, the due process analysis does not change. One asks whether a person received adequate notice and …
Feel Bad To Discard A Fashion Product: How Ai Designers Influence Individuals' Sustainable Consumption,
2025
Chungnam National University
Feel Bad To Discard A Fashion Product: How Ai Designers Influence Individuals' Sustainable Consumption, Ha Kyung Lee, Dooyoung Choi
Educational Leadership & Workforce Development Faculty Publications
This study explores how AI technology in fashion design influences consumers' sustainable consumption behaviors, focusing on emotional attachment to products. By comparing AI-generated and human-designed fashion items, the study examines how designer type impacts negative emotions about discarding products, mediated by emotional attachment. Results from two experimental studies reveal that designer type significantly affects negative emotions toward discarding human-designed items, but emotional attachment was not influenced by designer type in the first study. This lack of difference may be due to personal characteristics that moderate the effect. The second study found that individuals who perceive AI as human-like form stronger …
Towards Dynamic Learner State: Orchestrating Ai Agents And Workplace Performance Via The Model Context Protocol,
2025
Texas A&M University
Towards Dynamic Learner State: Orchestrating Ai Agents And Workplace Performance Via The Model Context Protocol, Mohan Yang, Nolan Lovett, Belle Li, Zhen Hou
Educational Leadership & Workforce Development Faculty Publications
Current learning and development approaches often struggle to capture dynamic individual capabilities, particularly the skills they acquire informally every day on the job. This dynamic creates a significant gap between what traditional models think people know and their actual performance, leading to an incomplete and often outdated understanding of how ready the workforce truly is, which can hinder organizational adaptability in rapidly evolving environments. This paper proposes a novel dynamic learner-state ecosystem—an AI-driven solution designed to bridge this gap. Our approach leverages specialized AI agents, orchestrated via the Model Context Protocol (MCP), to continuously track and evolve an individual’s multi-dimensional …
Exploring The Impact Of Value Co-Creation Through Ai-Driven Chatbbots On Customer Repeat Purchases,
2025
Old Dominion University
Exploring The Impact Of Value Co-Creation Through Ai-Driven Chatbbots On Customer Repeat Purchases, Dooyoung Choi, Jaeha Lee
Educational Leadership & Workforce Development Faculty Publications
Drawing on the Stimulus-Organism-Response (S-O-R) framework, this study explores how perceived value co-creation during chatbot interactions influences customer repeat purchase intentions through cognitive, emotional, and social responses to chatbots. A survey of 220 participants revealed that perceived value co-creation significantly affected repeat purchase intentions, with cognitive evaluations, emotional reactions, and social value serving as key mediators. However, the direct effect of value co-creation on purchase intentions was not significant. The findings suggest that while value co-creation enhances consumer engagement, repeat purchases occur only when consumers experience positive cognitive, emotional, and social outcomes. Therefore, it is crucial for retailers to incorporate …
