Open Access. Powered by Scholars. Published by Universities.®

Graphics and Human Computer Interfaces Commons

Open Access. Powered by Scholars. Published by Universities.®

2,362 Full-Text Articles 4,415 Authors 1,168,477 Downloads 165 Institutions

All Articles in Graphics and Human Computer Interfaces

Faceted Search

2,362 full-text articles. Page 6 of 101.

Gaze Transition Entropy And Automation Trust In Multitasking Workspace, Yusuke Yamani, Austin Jackson, Tetsuya Sato, Feyishola Ashimi, Michael S. Politowicz, Eric T. Chancey, Makoto Itoh 2026 Old Dominion University

Gaze Transition Entropy And Automation Trust In Multitasking Workspace, Yusuke Yamani, Austin Jackson, Tetsuya Sato, Feyishola Ashimi, Michael S. Politowicz, Eric T. Chancey, Makoto Itoh

Psychology Faculty Publications

Safe flight operation requires visual scanning across multiple displays in a cockpit, which collectively represent the state of the aircraft and supporting automation. Trust is a crucial factor that drives human-automation interaction, and recent work has suggested a relationship between an operator's visual attention and automation trust. One index that captures predictability of eye movements between different areas of interest is gaze transition entropy. The current work reanalyzed data from Sato et al., which examined eye movement patterns and trust in automation associated with the system monitoring task of the Multi-Attribute Task Battery. Results showed credible positive correlations between the …


Cognitive Prosthetic: An Ai-Enabled Multimodal System For Episodic Recall In Knowledge Work, Lawrence Obiuwevwi, Krzystof J. Rechowicz, Vikas Ashok, Sachin Shetty, Sampath Jayarathna 2026 Old Dominion University

Cognitive Prosthetic: An Ai-Enabled Multimodal System For Episodic Recall In Knowledge Work, Lawrence Obiuwevwi, Krzystof J. Rechowicz, Vikas Ashok, Sachin Shetty, Sampath Jayarathna

Computer Science Faculty Publications

Modern knowledge workplaces increasingly strain human episodic memory as individuals navigate fragmented attention, overlapping meetings, and multimodal information streams. Existing workplace tools provide partial support through note-taking or analytics but rarely integrate cognitive, physiological, and attentional context into retrievable memory representations. This paper presents the Cognitive Prosthetic Multimodal System (CPMS)—an AI-enabled proof-of-concept designed to support episodic recall in knowledge work through structured episodic capture and natural language retrieval. CPMS synchronizes speech transcripts, physiological signals, and gaze behavior into temporally aligned, JSON-based episodic records processed locally for privacy. Beyond data logging, the system includes a web-based retrieval interface that allows users …


Generalized Visual Relation Detection With Diffusion Models, Kaifeng GAO, Siqi CHEN, Hanwang ZHANG, Jun XIAO, Yueting ZHUANG, Qianru SUN 2026 Singapore Management University

Generalized Visual Relation Detection With Diffusion Models, Kaifeng Gao, Siqi Chen, Hanwang Zhang, Jun Xiao, Yueting Zhuang, Qianru Sun

Research Collection School Of Computing and Information Systems

Visual relation detection (VRD) aims to identify relationships (or interactions) between object pairs in an image. Although recent VRD models have achieved impressive performance, they are all restricted to pre-defined relation categories, while failing to consider the semantic ambiguity characteristic of visual relations. Unlike objects, the appearance of visual relations is always subtle and can be described by multiple predicate words from different perspectives, e.g., “ride” can be depicted as “race” and “sit on”, from the sports and spatial position views, respectively. To this end, we propose to model visual relations as continuous embeddings, and design diffusion models to achieve …


Trajlens: Visual Analysis For Constructing Cell Developmental Trajectories In Cross-Sample Exploration, Qipeng WANG, Shaolun RUAN, Rui SHENG, Yong WANG, Min ZHU, Huamin QU 2026 Singapore Management University

Trajlens: Visual Analysis For Constructing Cell Developmental Trajectories In Cross-Sample Exploration, Qipeng Wang, Shaolun Ruan, Rui Sheng, Yong Wang, Min Zhu, Huamin Qu

Research Collection School Of Computing and Information Systems

Constructing cell developmental trajectories is a critical task in single-cell RNA sequencing (scRNA-seq) analysis, enabling the inference of potential cellular progression paths. However, current automated methods are limited to establishing cell developmental trajectories within individual samples, necessitating biologists to manually link cells across samples to construct complete cross-sample evolutionary trajectories that consider cellular spatial dynamics. This process demands substantial human effort due to the complex spatial correspondence between each pair of samples. To address this challenge, we first proposed a GNN-based model to predict cross-sample cell developmental trajectories. We then developed TrajLens, a visual analytics system that supports biologists in …


Qualitative Study For Llm-Assisted Design Study Process: Strategies, Challenges, And Roles, Shaolun RUAN, Rui SHENG, Xiaolin WEN, Jiachen WANG, Tianyi ZHANG, Yong WANG, Tim DWYER, Jiannan LI 2026 Singapore Management University

Qualitative Study For Llm-Assisted Design Study Process: Strategies, Challenges, And Roles, Shaolun Ruan, Rui Sheng, Xiaolin Wen, Jiachen Wang, Tianyi Zhang, Yong Wang, Tim Dwyer, Jiannan Li

Research Collection School Of Computing and Information Systems

Design studies aim to develop visualization solutions for real-world problems across various application domains. Recently, the emergence of large language models (LLMs) has introduced new opportunities to enhance the design study process, providing capabilities such as creative problem-solving, data handling, and insightful analysis. However, despite their growing popularity, there remains a lack of systematic understanding of how LLMs can effectively assist researchers in visualization-specific design studies. In this paper, we conducted a rnulti-stage qualitative study to fill this gap, which involved 30 design study researchers from diverse backgrounds and expertise levels. Through in-depth interviews and carefully-designed questionnaires, we investigated strategies …


Cross-Modal Proxy Evolving For Ood Detection With Vision-Language Models, Hao Tang, Yu LIU, Shuanglin YAN, Fei Shen, Shengfeng HE, Jing Qin 2026 Singapore Management University

Cross-Modal Proxy Evolving For Ood Detection With Vision-Language Models, Hao Tang, Yu Liu, Shuanglin Yan, Fei Shen, Shengfeng He, Jing Qin

Research Collection School Of Computing and Information Systems

Reliable zero-shot detection of out-of-distribution (OOD) inputs is critical for deploying vision-language models in open-world settings. However, the lack of labeled negatives in zero-shot OOD detection necessitates proxy signals that remain effective under distribution shift. Existing negative-label methods rely on a fixed set of textual proxies, which (i) sparsely sample the semantic space beyond in-distribution (ID) classes and (ii) remain static while only visual features drift, leading to cross-modal misalignment and unstable predictions. In this paper, we propose CoEvo, a training- and annotation-free test-time framework that performs bidirectional, sample-conditioned adaptation of both textual and visual proxies. Specifically, CoEvo introduces a …


Seeing The Invisible Load: Xr+ Multimodal Sensing For Cognitive Ergonomics In Industrial Training, Jessica M. Johnson, Andwele Grant 2026 Old Dominion University

Seeing The Invisible Load: Xr+ Multimodal Sensing For Cognitive Ergonomics In Industrial Training, Jessica M. Johnson, Andwele Grant

Virginia Digital Maritime Center (VDMC) Faculty Publications

Extended reality (XR) technologies are increasingly positioned as disruptive Industry 5.0 tools for human-centric industrial training and intelligent human–system integration. Coupled with multimodal sensing (eye tracking, EEG, HRV, GSR, and other physiological signals), XR environments promise to make otherwise invisible cognitive demands observable, especially for novice trainees entering complex industrial settings. Yet the evidence base is fragmented: (1) there is no quantitative synthesis of the cognitive ergonomics benefits of XR plus sensing; (2) little is known about which XR–sensor configurations yield the strongest effects; (3) prior reviews rarely focus on industrial and manufacturing tasks; (4) multimodal signals are used predominantly …


Gui Test Migration Via Abstraction And Concretization, Yakun ZHANG, Chen LIU, Xiaofei XIE, Yun LIN, Jin Song DONG, Dan HAO, Lu ZHANG 2026 Peking University

Gui Test Migration Via Abstraction And Concretization, Yakun Zhang, Chen Liu, Xiaofei Xie, Yun Lin, Jin Song Dong, Dan Hao, Lu Zhang

Research Collection School Of Computing and Information Systems

GUI test migration aims to produce test cases with events and assertions to test specific functionalities of a target app. Existing migration approaches typically focus on the widget-mapping paradigm that maps widgets from source apps to target apps. However, since different apps may implement the same functionality in different ways, direct mapping may result in incomplete or buggy test cases, thus significantly impacting the effectiveness of testing the target functionality and the practical applicability of migration approaches.In this article, we propose a new migration paradigm (i.e., the abstraction-concretization paradigm) that first abstracts the test logic for the target functionality and …


Swimming In Uncertainty: Filling Data Gaps And Providing An Educational Platform For Beach Water Quality At Tybee Island, Georgia, Lukas Roberson 2026 Georgia Southern University

Swimming In Uncertainty: Filling Data Gaps And Providing An Educational Platform For Beach Water Quality At Tybee Island, Georgia, Lukas Roberson

College of Graduate Studies: Theses & Dissertations

@font-face {font-family:"Cambria Math"; panose-1:2 4 5 3 5 4 6 3 2 4; mso-font-charset:0; mso-generic-font-family:roman; mso-font-pitch:variable; mso-font-signature:-536870145 1107305727 0 0 415 0;}p.MsoNormal, li.MsoNormal, div.MsoNormal {mso-style-unhide:no; mso-style-qformat:yes; mso-style-parent:""; margin:0in; mso-pagination:widow-orphan; font-size:12.0pt; font-family:"Times New Roman",serif; mso-fareast-font-family:"Times New Roman";}.MsoChpDefault {mso-style-type:export-only; mso-default-props:yes; mso-font-kerning:0pt; mso-ligatures:none;}div.WordSection1 {page:WordSection1;}

Swimming in beaches water contaminated with high levels of bacteria can make you sick. Current monitoring at the public beaches on Tybee Island consists of weekly monitoring and enumeration of fecal indicator bacteria that takes 24 hours for results. If the number of bacteria exceed regulatory limits, a public health advisory is issued, and affected waters are retested until …


Purified Zero-Shot Sketch-Based Image Retrieval, Yang ZHOU, Jingru YANG, Jin WANG, Kaixiang HUANG, Guodong LU, Shengfeng HE 2026 Singapore Management University

Purified Zero-Shot Sketch-Based Image Retrieval, Yang Zhou, Jingru Yang, Jin Wang, Kaixiang Huang, Guodong Lu, Shengfeng He

Research Collection School Of Computing and Information Systems

Sketches, as a new solution in multimedia systems that can replace natural language, are characterized by sparse visual cues such as simple strokes that differ significantly from natural images containing complex elements such as background, foreground, and texture. This misalignment poses substantial challenges for zero-shot sketch-based image retrieval (ZS-SBIR). Prior approaches match sketches to full images and tend to overlook redundant elements in natural images, leading to model distraction and semantic ambiguity. To address this issue, we introduce a distraction-agnostic framework, purified cross-domain matching (PuXIM), which operates on a straightforward principle: masking and matching. We devise a visual-cross-linguistic (VxL) sampler …


Voxvista: Enhancing Screen Reading Experience For Online User Comments, Yash Prakash, Akshay Kolgar Nayak, Mohammed Shoaib Alyaan, Sampath Jayarathna, Hae-Na Lee, Vikas Ashok 2026 Old Dominion University

Voxvista: Enhancing Screen Reading Experience For Online User Comments, Yash Prakash, Akshay Kolgar Nayak, Mohammed Shoaib Alyaan, Sampath Jayarathna, Hae-Na Lee, Vikas Ashok

Computer Science Faculty Publications

Online discussions have become integral to how people exchange ideas, form opinions, and participate in collective deliberation. While sighted users can comfortably engage with online discussions, blind users who are dependent on screen readers are forced to listen to long threads narrated in a single, monotonic voice that lacks prosodic variation, rhythm, or emotion. This robotic auditory experience not only deteriorates the user engagement with the content but also increases cognitive strain, by making it difficult to remain attentive and discern meaning beyond literal words. In an interview study, most blind participants reported that monotonous narration hindered their ability to …


Tests Without Borders: A Global Approach To Measuring Visualization Literacy, Olivia A. Guess 2025 Washington University in St. Louis

Tests Without Borders: A Global Approach To Measuring Visualization Literacy, Olivia A. Guess

McKelvey School of Engineering Graduate Student Theses & Dissertations

Visualization literacy assessments shape how we understand people's ability to interpret data, yet most existing instruments embed Western datasets and assumptions that limit their relevance for global audiences. This thesis argues that because data is personal, assessments must also be culturally grounded. We introduce a unified framework for adapting the Mini-VLAT into 22 regionally responsive short-form assessments, each retaining the structure of the original test while incorporating datasets and scenarios tailored to specific regions around the world. To demonstrate how such adaptations can be customized and validated, we present a detailed case study of a Ghana-adapted Mini-VLAT, developed in collaboration …


Developing An Ai-Assisted Grading System Using Large Language Models, Andrei Modiga 2025 Southern Adventist University

Developing An Ai-Assisted Grading System Using Large Language Models, Andrei Modiga

MS in Computer Science Project Reports

We present a grading system that accelerates evaluation of open-ended student work across scanned and digital workflows. The system crops answer regions from PDFs, assigns submissions via OCR on identity regions only, and groups answers by visual semantics using a vision LLM. Instructors review and edit groups, apply rubric items once per group, and export grades from an on-screen table. The solution integrates Ghostscript rasterization, PdfPig page orchestration, SkiaSharp region extraction, Tesseract identity OCR, and GPT-4o Vision for grouping. We detail the architecture, token-budgeted batching strategy, and persistence design, then describe testing results for grouping quality, time-on-task, and usability. The …


My First Conversation With Chatgpt (February 22, 2023): Origins Of A Generative Dialogue, David Smith 2025 CUNY New York City College of Technology

My First Conversation With Chatgpt (February 22, 2023): Origins Of A Generative Dialogue, David Smith

Publications and Research

This working paper presents the first recorded interaction between the author and the generative AI system ChatGPT, written on February 22, 2023 during the initial weeks of a faculty sabbatical in Boston. The document preserves a complete and unedited transcript of an exploratory conversation conducted without predetermined research aims, marking the author’s first encounter with a large-language-model conversational interface. Although the exchange includes creative experimentation—including musical and poetic prompts—the discussion remains informal and wide-ranging, and no theoretical framework is articulated at this stage. Rather, this transcript is published as primary-source material documenting the moment of discovery and experimentation that precedes …


Decoding The Chameleon Game, Tri Dang '25, Hieu Tran, Brian T. Howard, Sutthirut Charoenphon, Dat Nguyen '25 2025 DePauw University

Decoding The Chameleon Game, Tri Dang '25, Hieu Tran, Brian T. Howard, Sutthirut Charoenphon, Dat Nguyen '25

Student Research

The Chameleon game is a challenging word association activity where players are given a secret word and must respond with words relevant to that secret word. It requires strategic thinking and deduction. The Chameleon must cleverly guess the secret keyword in this game while avoiding suspicion. Our research aims to create an advanced artificial intelligence (AI) model that can play the Chameleon game from both perspectives: as the Chameleon and as a Human. This AI is designed to guess secret keywords based on the information provided by the players, choose the best strategies to avoid detection as the Chameleon, identify …


Developing Accessible Narrative-Based Stem Learning Software For K-6 Braille Display Users, Dylan Ravel, Daniel Tsivkovski, Brandon Foley, Maryam Etezad, Franceli Cibrian, Ariel Han, Rajeev Joshi 2025 Chapman University

Developing Accessible Narrative-Based Stem Learning Software For K-6 Braille Display Users, Dylan Ravel, Daniel Tsivkovski, Brandon Foley, Maryam Etezad, Franceli Cibrian, Ariel Han, Rajeev Joshi

Student Scholar Symposium Abstracts and Posters

This research develops a free, accessible web application that enables K-6 students who are blind or visually impaired (BVI) to learn STEM concepts using refreshable braille displays. Currently, most online learning tools are not designed for BVI students, creating a significant educational barrier.

The application interfaces with commercial braille displays and uses narrative-based learning to make STEM content approachable and engaging. By presenting material as interactive stories, students can connect with concepts while developing braille reading skills. The curriculum design prioritizes accessibility through the Accessible Rich Internet Applications (ARIA) standards and screen reader support.

The goal is to provide BVI …


Visionglow: Evaluating Minimal-Disruption Smart-Home Control In Apple Vision Pro, Hongxiao Zheng 2025 Dartmouth College

Visionglow: Evaluating Minimal-Disruption Smart-Home Control In Apple Vision Pro, Hongxiao Zheng

Dartmouth College Master’s Theses

Smart-home control in mixed-reality environments like Apple Vision Pro often relies on disruptive, application-based paradigms, such as using a smartphone or a windowed virtual interface. These methods create a “mode switch” that imposes cognitive load and pulls users from their primary tasks. We present VisionGlow, a minimal-disruption spatial interaction technique for Vision Pro. VisionGlow represents devices as spatially-anchored “orbs.” To control a device, the user looks at its orb and performs a pinch gesture, which invokes a compact, contextual control panel. We conducted a within-subjects study (N=18) comparing VisionGlow against two baselines: the standard Apple Home app on a smartphone …


The Future Is Now: Empowering Society Through Ai Literacy, Jason S. Wrench, Sanae Elmoudden 2025 SUNY New Paltz

The Future Is Now: Empowering Society Through Ai Literacy, Jason S. Wrench, Sanae Elmoudden

Milne Open Textbooks

Artificial Intelligence (AI) is no longer a futuristic concept—it is the reality of the present. From the algorithms shaping our social media feeds to the generative tools transforming our workplaces, AI has permeated every aspect of modern life. The Future is Now moves beyond the hype to provide a comprehensive roadmap for understanding, navigating, and shaping this technological revolution.

Demystifying the Machine

This textbook serves as a user-friendly guide to the “black box” of AI. It breaks down complex technical concepts—from machine learning and neural networks to large language models—making them accessible to students across all disciplines. By establishing a …


Digital Reflections: Evaluating Body Dissatisfaction In Xr Through Eye- And Body-Tracked Virtual Humans, Deyrel Diaz 2025 Clemson University

Digital Reflections: Evaluating Body Dissatisfaction In Xr Through Eye- And Body-Tracked Virtual Humans, Deyrel Diaz

All Dissertations

In an era where digital and physical realities increasingly intertwine, the perception of body image is undergoing a significant transformation. Traditional understandings of body dissatisfaction, long studied in relation to psychological distress and eating disorders, are now being reshaped by technologies such as Virtual Reality (VR), Augmented Reality (AR), and Artificially Intelligent (AI)- generated media. These technologies have introduced novel ways of experiencing and interacting with the human form, raising critical questions about their impact on self-perception and internalization of beauty standards.

As virtual representations become more prevalent in entertainment, social media, and interactive platforms, it is becoming more crucial …


Instance-Level Video Depth In Groups Beyond Occlusions, Yuan LIANG, Yang ZHOU, Ziming SUN, Tianyi XIANG, Guiqing LI, Shengfeng HE 2025 Singapore Management University

Instance-Level Video Depth In Groups Beyond Occlusions, Yuan Liang, Yang Zhou, Ziming Sun, Tianyi Xiang, Guiqing Li, Shengfeng He

Research Collection School Of Computing and Information Systems

Depth estimation in dynamic, multi-object scenes remains a major challenge, especially under severe occlusions. Existing monocular models, including foundation models, struggle with instance-wise depth consistency due to their reliance on global regression. We tackle this problem from two key aspects: data and methodology. First, we introduce the Group Instance Depth (GID) dataset, the first large-scale video depth dataset with instance-level annotations, featuring 101,500 frames from real-world activity scenes. GID bridges the gap between synthetic and real-world depth data by providing high-fidelity depth supervision for multi-object interactions. Second, we propose InstanceDepth, the first occlusion-aware depth estimation framework for multi-object environments. Our …


Digital Commons powered by bepress