Open Access. Powered by Scholars. Published by Universities.®

2026

Discipline
Institution
Keyword
Publication
Publication Type
File Type

Articles 91 - 111 of 111

Full-Text Articles in Graphics and Human Computer Interfaces

Error-Driven Density Control For Compact Gaussian Splatting Under Sparse Supervision, Abdelrhman Elrawy Jan 2026

Error-Driven Density Control For Compact Gaussian Splatting Under Sparse Supervision, Abdelrhman Elrawy

Theses and Dissertations (Comprehensive)

This thesis studies efficiency and stability challenges in Gaussian-splatting-based reconstruction under sparse supervision. In few-shot novel view synthesis, standard 3D Gaussian Splatting (3DGS) can overfit the limited training views and grow an unnecessarily large number of primitives due to limitations in its Adaptive Density Control (ADC) mechanism. This thesis introduces an error-driven reformulation of ADC that triggers densification using opacity gradients as a lightweight proxy for rendering error, and shows that such aggressive densification must be paired with delayed and conservative pruning to prevent destructive create--destroy cycles. When combined with depth-based geometric regularization, the resulting framework produces substantially more compact …


Improving Medical Diagnostics With Vision-Language Models: Convex Hull-Based Uncertainty Analysis, Ferhat Ozgur Catak, Murat Kuzlu, Taylor Patrick, Michel Audette Jan 2026

Improving Medical Diagnostics With Vision-Language Models: Convex Hull-Based Uncertainty Analysis, Ferhat Ozgur Catak, Murat Kuzlu, Taylor Patrick, Michel Audette

Engineering Technology Faculty Publications

In recent years, vision-language models (VLMs) have been applied to various fields, including healthcare, education, finance, and manufacturing, with remarkable performance. However, concerns remain regarding VLMs' consistency and uncertainty, particularly in critical applications such as healthcare, which demand a high level of trust and reliability. This paper proposes a novel approach to evaluate uncertainty in VLMs' responses using a convex hull approach on a healthcare application for visual question answering (VQA). For any VLM, temperature refers to a sampling parameter used in probabilistic generation, which controls the randomness of the model's output. The LLM-CXR model is selected as the medical …


A Web-Based Wizard-Of-Oz Platform For Collaborative And Reproducible Human-Robot Interaction Research, Sean O'Connor Jan 2026

A Web-Based Wizard-Of-Oz Platform For Collaborative And Reproducible Human-Robot Interaction Research, Sean O'Connor

Honors Theses

The Wizard-of-Oz (WoZ) technique is widely used in Human-Robot Interaction (HRI) research, but two persistent problems limit its effectiveness: existing tools impose technical barriers that exclude non-engineering domain experts (the Accessibility Problem), and the fragmented landscape of robot-specific implementations makes interaction scripts difficult to port across platforms (the Reproducibility Problem- concerning execution consistency and portability, not third-party replication). Through a literature review, I identified three design principles to address both: a hierarchical specification model, an event-driven execution model, and a plugin architecture that decouples experiment logic from robot-specific implementations. I realized these principles in HRIStudio, an open-source, web-based platform providing …


Geovig And Purevig: Geometry-Aware Architectures For Efficient Computer Vision, Omar Ismail Jan 2026

Geovig And Purevig: Geometry-Aware Architectures For Efficient Computer Vision, Omar Ismail

Theses and Dissertations (Comprehensive)

Deploying deep learning models for medical image analysis on mobile devices requires a balance between inference latency, memory footprint, and delineating anatomical boundaries with high accuracy. While Convolutional Neural Networks (CNNs) and mobile Vision Transformers (ViTs) offer efficiency, they often struggle to model the irregular, non-local geometric structures inherent in biological tissues without incurring prohibitive computational costs. In this thesis, we introduce GeoViG (Geometric Vision Graph), an architecture that bridges the gap between efficient grid-based processing and explicit Geometric Deep Learning. GeoViG introduces a novel transition from high-resolution pixel grids to low-resolution dynamic graphs via a SpreadEdgePool operator, a geometry-aware …


Visual Analytics For Interpretable Quantum Computing, Shaolun Ruan Jan 2026

Visual Analytics For Interpretable Quantum Computing, Shaolun Ruan

Dissertations and Theses Collection (Open Access)

Quantum computing has entered a stage of increasing practicality. Many quantum hardware vendors such as IBM, Rigetti, Honeywell, and IonQ now enable experiments on real devices in the Noisy Intermediate-Scale Quantum (NISQ) era. These platforms show computational advantages in domains such as optimization, machine learning, and materials science. However, they remain limited by hardware noise and the absence of human-interpretable information. Existing visual metaphors, such as the Bloch Sphere for single-qubit states or circuit schematics for algorithm design, struggle to convey multi-qubit entanglement or measurement probabilities in ways accessible to human reasoning. Likewise, the rise of variational quantum circuits and …


Eeg And Imu Gait Signal Processing: A Comparative Assessment Of The "Reza" Exponential Filter And Classic Filters, Reza Pousti, Daniel M. Russell, Derek C. Monroe, Christopher K. Rhea Jan 2026

Eeg And Imu Gait Signal Processing: A Comparative Assessment Of The "Reza" Exponential Filter And Classic Filters, Reza Pousti, Daniel M. Russell, Derek C. Monroe, Christopher K. Rhea

Rehabilitation Sciences Faculty Publications

Noise degrades both EEG and gait signals, and classical IIR filters (Butterworth, Chebyshev, elliptic) involve trade-offs between passband flatness, ripple, and roll-off. This study compared a novel exponential "Reza" filter with these designs for neural and locomotor data. We analyzed an open-source mobile brain-body imaging dataset with EEG and gait data from 49 healthy adults (EEG: 256-channel, 512 Hz; IMUs: six APDM Opals, 128 Hz). EEG channels were grand-averaged and band-pass filtered at 0.5-50 Hz, while IMU axes were averaged and band-pass filtered at 0.5-5 Hz. The outcomes were signal-to-noise ratio SNR (dB) and band-integrated Welch PSD (EEG:0.5-50 Hz; IMU:0.5-5 …


Examining Inclusive Computing Education For Blind Students In India, Akshay Kolgar Nayak, Yash Prakash, Md Javedul Ferdous, Sampath Jayarathna, Hae-Na Lee, Vikas Ashok Jan 2026

Examining Inclusive Computing Education For Blind Students In India, Akshay Kolgar Nayak, Yash Prakash, Md Javedul Ferdous, Sampath Jayarathna, Hae-Na Lee, Vikas Ashok

Computer Science Faculty Publications

The growing demand for computer professionals, driven by the expanding Information Technology industry, has led to numerous inclusive computing education efforts. These efforts have even included blind or visually-impaired (BVI) students, who are being increasingly encouraged to pursue education and a career in computing, despite the visually-oriented nature of the discipline. Extant literature has predominantly focused on identifying and addressing the accessibility barriers faced by BVI students to promote more inclusive learning environments. While few studies have also investigated the accessibility of computing education from the perspectives of BVI learners and instructors, these have been primarily situated in the Global …


Finding The Signal In The Noise: An Exploratory Study On Assessing The Effectiveness Of Ai And Accessibility Forums For Blind Users' Support Needs, Satwik Ram Kodandaram, Jiawei Zhou, Xiaojun Bi, Iv Ramakrishnan, Vikas Ashok Jan 2026

Finding The Signal In The Noise: An Exploratory Study On Assessing The Effectiveness Of Ai And Accessibility Forums For Blind Users' Support Needs, Satwik Ram Kodandaram, Jiawei Zhou, Xiaojun Bi, Iv Ramakrishnan, Vikas Ashok

Computer Science Faculty Publications

Accessibility forums and, more recently, generative AI tools have become vital resources for blind users seeking solutions to computer-interaction issues and learning about new assistive technologies, screen reader features, tutorials, and software updates. Understanding user experiences with these resources is essential for identifying and addressing persistent support gaps. Towards this, we interviewed 14 blind users who regularly engage with forums and GenAI tools. Findings revealed that forums often overwhelm users with multiple overlapping topics, redundant or irrelevant content, and fragmented responses that must be mentally pieced together, increasing cognitive load. GenAI tools, while offering more direct assistance, introduce new barriers …


A Comparative Analysis Of Explainable Ai (Xai) Techniques For Transparent And Reliable Image Classification, Sovon Chakraborty, Shakib Mahmud Dipto, Kevin R. Pilkiewicz, Michael L. Mayo, Pratip Rana Jan 2026

A Comparative Analysis Of Explainable Ai (Xai) Techniques For Transparent And Reliable Image Classification, Sovon Chakraborty, Shakib Mahmud Dipto, Kevin R. Pilkiewicz, Michael L. Mayo, Pratip Rana

Computer Science Faculty Publications

Evaluating the trustworthiness of black-box machine learning models remains a significant methodological challenge. Their lack of transparency and interpretability limits applicability, because stakeholders often seek transparency before trusting the results of black-box machine learning models. Explainable AI (XAI) methods provide for human-understandable justifications and informed decision-making of these black-box architectures. Therefore, it is imperative to select the proper XAI model tailored to specific tasks. In this research, we focus on examining four XAI techniques: PEEK, LRP, GRAD-CAM, and LIME to understand how they perform against each other for image classification tasks. We evaluate the performance, robustness, generalizability, noise stability, and …


Susceptibility To High-Fidelity Misinformation: An Eye-Tracking Analysis, Yasasi Abeysinghe, Gavindya Jayawardena, Enkelejda Kasneci, Sampath Jayarathna Jan 2026

Susceptibility To High-Fidelity Misinformation: An Eye-Tracking Analysis, Yasasi Abeysinghe, Gavindya Jayawardena, Enkelejda Kasneci, Sampath Jayarathna

Computer Science Faculty Publications

With the rise of online misinformation and AI-generated text, understanding human perception of news truthfulness is critical. In this study, we examine visual attention and cognitive processing using eye-tracking measures as individuals read fake and real news articles sharing nearly identical structure and imagery, differing only in subtle textual changes. Using the public FakeNewsPerception dataset, we analyze advanced gaze measures, including scanpaths, AOI transitions, and luminance-corrected pupil measures, beyond basic gaze features, in relation to news truthfulness and perceived believability. Results show that, given the high fidelity of the fake news, readers exhibited comparable visual scanning patterns, attention allocation across …


Cognitive Prosthetic: An Ai-Enabled Multimodal System For Episodic Recall In Knowledge Work, Lawrence Obiuwevwi, Krzystof J. Rechowicz, Vikas Ashok, Sachin Shetty, Sampath Jayarathna Jan 2026

Cognitive Prosthetic: An Ai-Enabled Multimodal System For Episodic Recall In Knowledge Work, Lawrence Obiuwevwi, Krzystof J. Rechowicz, Vikas Ashok, Sachin Shetty, Sampath Jayarathna

Computer Science Faculty Publications

Modern knowledge workplaces increasingly strain human episodic memory as individuals navigate fragmented attention, overlapping meetings, and multimodal information streams. Existing workplace tools provide partial support through note-taking or analytics but rarely integrate cognitive, physiological, and attentional context into retrievable memory representations. This paper presents the Cognitive Prosthetic Multimodal System (CPMS)—an AI-enabled proof-of-concept designed to support episodic recall in knowledge work through structured episodic capture and natural language retrieval. CPMS synchronizes speech transcripts, physiological signals, and gaze behavior into temporally aligned, JSON-based episodic records processed locally for privacy. Beyond data logging, the system includes a web-based retrieval interface that allows users …


Generalized Visual Relation Detection With Diffusion Models, Kaifeng Gao, Siqi Chen, Hanwang Zhang, Jun Xiao, Yueting Zhuang, Qianru Sun Jan 2026

Generalized Visual Relation Detection With Diffusion Models, Kaifeng Gao, Siqi Chen, Hanwang Zhang, Jun Xiao, Yueting Zhuang, Qianru Sun

Research Collection School Of Computing and Information Systems

Visual relation detection (VRD) aims to identify relationships (or interactions) between object pairs in an image. Although recent VRD models have achieved impressive performance, they are all restricted to pre-defined relation categories, while failing to consider the semantic ambiguity characteristic of visual relations. Unlike objects, the appearance of visual relations is always subtle and can be described by multiple predicate words from different perspectives, e.g., “ride” can be depicted as “race” and “sit on”, from the sports and spatial position views, respectively. To this end, we propose to model visual relations as continuous embeddings, and design diffusion models to achieve …


Trajlens: Visual Analysis For Constructing Cell Developmental Trajectories In Cross-Sample Exploration, Qipeng Wang, Shaolun Ruan, Rui Sheng, Yong Wang, Min Zhu, Huamin Qu Jan 2026

Trajlens: Visual Analysis For Constructing Cell Developmental Trajectories In Cross-Sample Exploration, Qipeng Wang, Shaolun Ruan, Rui Sheng, Yong Wang, Min Zhu, Huamin Qu

Research Collection School Of Computing and Information Systems

Constructing cell developmental trajectories is a critical task in single-cell RNA sequencing (scRNA-seq) analysis, enabling the inference of potential cellular progression paths. However, current automated methods are limited to establishing cell developmental trajectories within individual samples, necessitating biologists to manually link cells across samples to construct complete cross-sample evolutionary trajectories that consider cellular spatial dynamics. This process demands substantial human effort due to the complex spatial correspondence between each pair of samples. To address this challenge, we first proposed a GNN-based model to predict cross-sample cell developmental trajectories. We then developed TrajLens, a visual analytics system that supports biologists in …


Qualitative Study For Llm-Assisted Design Study Process: Strategies, Challenges, And Roles, Shaolun Ruan, Rui Sheng, Xiaolin Wen, Jiachen Wang, Tianyi Zhang, Yong Wang, Tim Dwyer, Jiannan Li Jan 2026

Qualitative Study For Llm-Assisted Design Study Process: Strategies, Challenges, And Roles, Shaolun Ruan, Rui Sheng, Xiaolin Wen, Jiachen Wang, Tianyi Zhang, Yong Wang, Tim Dwyer, Jiannan Li

Research Collection School Of Computing and Information Systems

Design studies aim to develop visualization solutions for real-world problems across various application domains. Recently, the emergence of large language models (LLMs) has introduced new opportunities to enhance the design study process, providing capabilities such as creative problem-solving, data handling, and insightful analysis. However, despite their growing popularity, there remains a lack of systematic understanding of how LLMs can effectively assist researchers in visualization-specific design studies. In this paper, we conducted a rnulti-stage qualitative study to fill this gap, which involved 30 design study researchers from diverse backgrounds and expertise levels. Through in-depth interviews and carefully-designed questionnaires, we investigated strategies …


Cross-Modal Proxy Evolving For Ood Detection With Vision-Language Models, Hao Tang, Yu Liu, Shuanglin Yan, Fei Shen, Shengfeng He, Jing Qin Jan 2026

Cross-Modal Proxy Evolving For Ood Detection With Vision-Language Models, Hao Tang, Yu Liu, Shuanglin Yan, Fei Shen, Shengfeng He, Jing Qin

Research Collection School Of Computing and Information Systems

Reliable zero-shot detection of out-of-distribution (OOD) inputs is critical for deploying vision-language models in open-world settings. However, the lack of labeled negatives in zero-shot OOD detection necessitates proxy signals that remain effective under distribution shift. Existing negative-label methods rely on a fixed set of textual proxies, which (i) sparsely sample the semantic space beyond in-distribution (ID) classes and (ii) remain static while only visual features drift, leading to cross-modal misalignment and unstable predictions. In this paper, we propose CoEvo, a training- and annotation-free test-time framework that performs bidirectional, sample-conditioned adaptation of both textual and visual proxies. Specifically, CoEvo introduces a …


Seeing The Invisible Load: Xr+ Multimodal Sensing For Cognitive Ergonomics In Industrial Training, Jessica M. Johnson, Andwele Grant Jan 2026

Seeing The Invisible Load: Xr+ Multimodal Sensing For Cognitive Ergonomics In Industrial Training, Jessica M. Johnson, Andwele Grant

Virginia Digital Maritime Center (VDMC) Faculty Publications

Extended reality (XR) technologies are increasingly positioned as disruptive Industry 5.0 tools for human-centric industrial training and intelligent human–system integration. Coupled with multimodal sensing (eye tracking, EEG, HRV, GSR, and other physiological signals), XR environments promise to make otherwise invisible cognitive demands observable, especially for novice trainees entering complex industrial settings. Yet the evidence base is fragmented: (1) there is no quantitative synthesis of the cognitive ergonomics benefits of XR plus sensing; (2) little is known about which XR–sensor configurations yield the strongest effects; (3) prior reviews rarely focus on industrial and manufacturing tasks; (4) multimodal signals are used predominantly …


Gui Test Migration Via Abstraction And Concretization, Yakun Zhang, Chen Liu, Xiaofei Xie, Yun Lin, Jin Song Dong, Dan Hao, Lu Zhang Jan 2026

Gui Test Migration Via Abstraction And Concretization, Yakun Zhang, Chen Liu, Xiaofei Xie, Yun Lin, Jin Song Dong, Dan Hao, Lu Zhang

Research Collection School Of Computing and Information Systems

GUI test migration aims to produce test cases with events and assertions to test specific functionalities of a target app. Existing migration approaches typically focus on the widget-mapping paradigm that maps widgets from source apps to target apps. However, since different apps may implement the same functionality in different ways, direct mapping may result in incomplete or buggy test cases, thus significantly impacting the effectiveness of testing the target functionality and the practical applicability of migration approaches.In this article, we propose a new migration paradigm (i.e., the abstraction-concretization paradigm) that first abstracts the test logic for the target functionality and …


Swimming In Uncertainty: Filling Data Gaps And Providing An Educational Platform For Beach Water Quality At Tybee Island, Georgia, Lukas Roberson Jan 2026

Swimming In Uncertainty: Filling Data Gaps And Providing An Educational Platform For Beach Water Quality At Tybee Island, Georgia, Lukas Roberson

College of Graduate Studies: Theses & Dissertations

@font-face {font-family:"Cambria Math"; panose-1:2 4 5 3 5 4 6 3 2 4; mso-font-charset:0; mso-generic-font-family:roman; mso-font-pitch:variable; mso-font-signature:-536870145 1107305727 0 0 415 0;}p.MsoNormal, li.MsoNormal, div.MsoNormal {mso-style-unhide:no; mso-style-qformat:yes; mso-style-parent:""; margin:0in; mso-pagination:widow-orphan; font-size:12.0pt; font-family:"Times New Roman",serif; mso-fareast-font-family:"Times New Roman";}.MsoChpDefault {mso-style-type:export-only; mso-default-props:yes; mso-font-kerning:0pt; mso-ligatures:none;}div.WordSection1 {page:WordSection1;}

Swimming in beaches water contaminated with high levels of bacteria can make you sick. Current monitoring at the public beaches on Tybee Island consists of weekly monitoring and enumeration of fecal indicator bacteria that takes 24 hours for results. If the number of bacteria exceed regulatory limits, a public health advisory is issued, and affected waters are retested until …


Purified Zero-Shot Sketch-Based Image Retrieval, Yang Zhou, Jingru Yang, Jin Wang, Kaixiang Huang, Guodong Lu, Shengfeng He Jan 2026

Purified Zero-Shot Sketch-Based Image Retrieval, Yang Zhou, Jingru Yang, Jin Wang, Kaixiang Huang, Guodong Lu, Shengfeng He

Research Collection School Of Computing and Information Systems

Sketches, as a new solution in multimedia systems that can replace natural language, are characterized by sparse visual cues such as simple strokes that differ significantly from natural images containing complex elements such as background, foreground, and texture. This misalignment poses substantial challenges for zero-shot sketch-based image retrieval (ZS-SBIR). Prior approaches match sketches to full images and tend to overlook redundant elements in natural images, leading to model distraction and semantic ambiguity. To address this issue, we introduce a distraction-agnostic framework, purified cross-domain matching (PuXIM), which operates on a straightforward principle: masking and matching. We devise a visual-cross-linguistic (VxL) sampler …


Voxvista: Enhancing Screen Reading Experience For Online User Comments, Yash Prakash, Akshay Kolgar Nayak, Mohammed Shoaib Alyaan, Sampath Jayarathna, Hae-Na Lee, Vikas Ashok Jan 2026

Voxvista: Enhancing Screen Reading Experience For Online User Comments, Yash Prakash, Akshay Kolgar Nayak, Mohammed Shoaib Alyaan, Sampath Jayarathna, Hae-Na Lee, Vikas Ashok

Computer Science Faculty Publications

Online discussions have become integral to how people exchange ideas, form opinions, and participate in collective deliberation. While sighted users can comfortably engage with online discussions, blind users who are dependent on screen readers are forced to listen to long threads narrated in a single, monotonic voice that lacks prosodic variation, rhythm, or emotion. This robotic auditory experience not only deteriorates the user engagement with the content but also increases cognitive strain, by making it difficult to remain attentive and discern meaning beyond literal words. In an interview study, most blind participants reported that monotonous narration hindered their ability to …


Gaze Transition Entropy And Automation Trust In Multitasking Workspace, Yusuke Yamani, Austin Jackson, Tetsuya Sato, Feyishola Ashimi, Michael S. Politowicz, Eric T. Chancey, Makoto Itoh Jan 2026

Gaze Transition Entropy And Automation Trust In Multitasking Workspace, Yusuke Yamani, Austin Jackson, Tetsuya Sato, Feyishola Ashimi, Michael S. Politowicz, Eric T. Chancey, Makoto Itoh

Psychology Faculty Publications

Safe flight operation requires visual scanning across multiple displays in a cockpit, which collectively represent the state of the aircraft and supporting automation. Trust is a crucial factor that drives human-automation interaction, and recent work has suggested a relationship between an operator's visual attention and automation trust. One index that captures predictability of eye movements between different areas of interest is gaze transition entropy. The current work reanalyzed data from Sato et al., which examined eye movement patterns and trust in automation associated with the system monitoring task of the Multi-Attribute Task Battery. Results showed credible positive correlations between the …