Open Access. Powered by Scholars. Published by Universities.®
Graphics and Human Computer Interfaces Commons™
Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Databases and Information Systems (473)
- Artificial Intelligence and Robotics (400)
- Engineering (335)
- Software Engineering (307)
- Social and Behavioral Sciences (264)
-
- Other Computer Sciences (241)
- Computer Engineering (178)
- Arts and Humanities (142)
- Theory and Algorithms (133)
- Education (126)
- OS and Networks (97)
- Medicine and Health Sciences (93)
- Numerical Analysis and Scientific Computing (93)
- Systems Architecture (81)
- Art and Design (79)
- Business (78)
- Programming Languages and Compilers (78)
- Psychology (77)
- Communication (74)
- Electrical and Computer Engineering (73)
- Data Storage Systems (71)
- Information Security (69)
- Life Sciences (58)
- Data Science (51)
- Educational Technology (48)
- Library and Information Science (40)
- Communication Technology and New Media (39)
- Institution
-
- Singapore Management University (938)
- University of Dayton (114)
- Air Force Institute of Technology (98)
- Old Dominion University (97)
- California Polytechnic State University, San Luis Obispo (96)
-
- University of Arkansas, Fayetteville (89)
- University of Nebraska - Lincoln (51)
- City University of New York (CUNY) (48)
- Technological University Dublin (48)
- University of Malaya (42)
- San Jose State University (37)
- Dartmouth College (34)
- Embry-Riddle Aeronautical University (24)
- Clemson University (23)
- Purdue University (23)
- Rochester Institute of Technology (23)
- The University of Akron (22)
- Chapman University (20)
- Edith Cowan University (20)
- University of Kentucky (18)
- Michigan Technological University (16)
- University of Central Florida (15)
- Southern Adventist University (13)
- California State University, San Bernardino (12)
- Kennesaw State University (12)
- St. Mary's University (12)
- Nova Southeastern University (11)
- University of Minnesota Morris Digital Well (11)
- Louisiana State University (10)
- University of Nevada, Las Vegas (10)
- Keyword
-
- Virtual reality (62)
- Visualization (46)
- Computer graphics (38)
- Computer vision (37)
- Accessibility (36)
-
- Human-computer interaction (35)
- Augmented reality (33)
- Usability (31)
- Machine learning (29)
- Computer Science (25)
- Data visualization (25)
- Machine Learning (25)
- Artificial intelligence (24)
- Deep learning (24)
- Virtual Reality (23)
- HCI (22)
- Computer science (20)
- Eye tracking (20)
- Human computer interaction (20)
- User experience (20)
- Design (19)
- Deep Learning (16)
- Education (16)
- Feature extraction (15)
- Graph Neural Networks (15)
- Graphics (15)
- VR (15)
- Gamification (14)
- Image processing (14)
- Applied sciences (13)
- Publication Year
- Publication
-
- Research Collection School Of Computing and Information Systems (912)
- Computer Science Faculty Publications (135)
- Theses and Dissertations (98)
- Master's Theses (50)
- Graduate Theses and Dissertations (43)
-
- Student Works (2000-2009) (33)
- Computer Science and Computer Engineering Undergraduate Honors Theses (31)
- 3-D Printed Model Structural Files (29)
- Publications and Research (28)
- Dartmouth College Master’s Theses (24)
- Williams Honors College, Honors Research Projects (22)
- Master's Projects (20)
- Conference papers (19)
- Frameless (19)
- Computer Science and Software Engineering (18)
- All Dissertations (17)
- Dissertations and Theses Collection (Open Access) (16)
- Dissertations, Master's Theses and Master's Reports (16)
- H-Workload 2017: Models and Applications (Works in Progress) (15)
- Computer Engineering (14)
- Electronic Theses and Dissertations (14)
- Theses : Honours (14)
- Honors Theses (13)
- MAICS: The Modern Artificial Intelligence and Cognitive Science Conference (12)
- AFIT Patents (11)
- CCAC Theses and Dissertations (11)
- Engineering Faculty Articles and Research (11)
- Scholarly Horizons: University of Minnesota, Morris Undergraduate Journal (10)
- Inquiry: The University of Arkansas Undergraduate Research Journal (9)
- Publications (9)
- Publication Type
- File Type
Articles 91 - 120 of 2362
Full-Text Articles in Graphics and Human Computer Interfaces
Error-Driven Density Control For Compact Gaussian Splatting Under Sparse Supervision, Abdelrhman Elrawy
Error-Driven Density Control For Compact Gaussian Splatting Under Sparse Supervision, Abdelrhman Elrawy
Theses and Dissertations (Comprehensive)
This thesis studies efficiency and stability challenges in Gaussian-splatting-based reconstruction under sparse supervision. In few-shot novel view synthesis, standard 3D Gaussian Splatting (3DGS) can overfit the limited training views and grow an unnecessarily large number of primitives due to limitations in its Adaptive Density Control (ADC) mechanism. This thesis introduces an error-driven reformulation of ADC that triggers densification using opacity gradients as a lightweight proxy for rendering error, and shows that such aggressive densification must be paired with delayed and conservative pruning to prevent destructive create--destroy cycles. When combined with depth-based geometric regularization, the resulting framework produces substantially more compact …
Improving Medical Diagnostics With Vision-Language Models: Convex Hull-Based Uncertainty Analysis, Ferhat Ozgur Catak, Murat Kuzlu, Taylor Patrick, Michel Audette
Improving Medical Diagnostics With Vision-Language Models: Convex Hull-Based Uncertainty Analysis, Ferhat Ozgur Catak, Murat Kuzlu, Taylor Patrick, Michel Audette
Engineering Technology Faculty Publications
In recent years, vision-language models (VLMs) have been applied to various fields, including healthcare, education, finance, and manufacturing, with remarkable performance. However, concerns remain regarding VLMs' consistency and uncertainty, particularly in critical applications such as healthcare, which demand a high level of trust and reliability. This paper proposes a novel approach to evaluate uncertainty in VLMs' responses using a convex hull approach on a healthcare application for visual question answering (VQA). For any VLM, temperature refers to a sampling parameter used in probabilistic generation, which controls the randomness of the model's output. The LLM-CXR model is selected as the medical …
A Web-Based Wizard-Of-Oz Platform For Collaborative And Reproducible Human-Robot Interaction Research, Sean O'Connor
A Web-Based Wizard-Of-Oz Platform For Collaborative And Reproducible Human-Robot Interaction Research, Sean O'Connor
Honors Theses
The Wizard-of-Oz (WoZ) technique is widely used in Human-Robot Interaction (HRI) research, but two persistent problems limit its effectiveness: existing tools impose technical barriers that exclude non-engineering domain experts (the Accessibility Problem), and the fragmented landscape of robot-specific implementations makes interaction scripts difficult to port across platforms (the Reproducibility Problem- concerning execution consistency and portability, not third-party replication). Through a literature review, I identified three design principles to address both: a hierarchical specification model, an event-driven execution model, and a plugin architecture that decouples experiment logic from robot-specific implementations. I realized these principles in HRIStudio, an open-source, web-based platform providing …
Geovig And Purevig: Geometry-Aware Architectures For Efficient Computer Vision, Omar Ismail
Geovig And Purevig: Geometry-Aware Architectures For Efficient Computer Vision, Omar Ismail
Theses and Dissertations (Comprehensive)
Deploying deep learning models for medical image analysis on mobile devices requires a balance between inference latency, memory footprint, and delineating anatomical boundaries with high accuracy. While Convolutional Neural Networks (CNNs) and mobile Vision Transformers (ViTs) offer efficiency, they often struggle to model the irregular, non-local geometric structures inherent in biological tissues without incurring prohibitive computational costs. In this thesis, we introduce GeoViG (Geometric Vision Graph), an architecture that bridges the gap between efficient grid-based processing and explicit Geometric Deep Learning. GeoViG introduces a novel transition from high-resolution pixel grids to low-resolution dynamic graphs via a SpreadEdgePool operator, a geometry-aware …
Visual Analytics For Interpretable Quantum Computing, Shaolun Ruan
Visual Analytics For Interpretable Quantum Computing, Shaolun Ruan
Dissertations and Theses Collection (Open Access)
Quantum computing has entered a stage of increasing practicality. Many quantum hardware vendors such as IBM, Rigetti, Honeywell, and IonQ now enable experiments on real devices in the Noisy Intermediate-Scale Quantum (NISQ) era. These platforms show computational advantages in domains such as optimization, machine learning, and materials science. However, they remain limited by hardware noise and the absence of human-interpretable information. Existing visual metaphors, such as the Bloch Sphere for single-qubit states or circuit schematics for algorithm design, struggle to convey multi-qubit entanglement or measurement probabilities in ways accessible to human reasoning. Likewise, the rise of variational quantum circuits and …
Eeg And Imu Gait Signal Processing: A Comparative Assessment Of The "Reza" Exponential Filter And Classic Filters, Reza Pousti, Daniel M. Russell, Derek C. Monroe, Christopher K. Rhea
Eeg And Imu Gait Signal Processing: A Comparative Assessment Of The "Reza" Exponential Filter And Classic Filters, Reza Pousti, Daniel M. Russell, Derek C. Monroe, Christopher K. Rhea
Rehabilitation Sciences Faculty Publications
Noise degrades both EEG and gait signals, and classical IIR filters (Butterworth, Chebyshev, elliptic) involve trade-offs between passband flatness, ripple, and roll-off. This study compared a novel exponential "Reza" filter with these designs for neural and locomotor data. We analyzed an open-source mobile brain-body imaging dataset with EEG and gait data from 49 healthy adults (EEG: 256-channel, 512 Hz; IMUs: six APDM Opals, 128 Hz). EEG channels were grand-averaged and band-pass filtered at 0.5-50 Hz, while IMU axes were averaged and band-pass filtered at 0.5-5 Hz. The outcomes were signal-to-noise ratio SNR (dB) and band-integrated Welch PSD (EEG:0.5-50 Hz; IMU:0.5-5 …
Examining Inclusive Computing Education For Blind Students In India, Akshay Kolgar Nayak, Yash Prakash, Md Javedul Ferdous, Sampath Jayarathna, Hae-Na Lee, Vikas Ashok
Examining Inclusive Computing Education For Blind Students In India, Akshay Kolgar Nayak, Yash Prakash, Md Javedul Ferdous, Sampath Jayarathna, Hae-Na Lee, Vikas Ashok
Computer Science Faculty Publications
The growing demand for computer professionals, driven by the expanding Information Technology industry, has led to numerous inclusive computing education efforts. These efforts have even included blind or visually-impaired (BVI) students, who are being increasingly encouraged to pursue education and a career in computing, despite the visually-oriented nature of the discipline. Extant literature has predominantly focused on identifying and addressing the accessibility barriers faced by BVI students to promote more inclusive learning environments. While few studies have also investigated the accessibility of computing education from the perspectives of BVI learners and instructors, these have been primarily situated in the Global …
Finding The Signal In The Noise: An Exploratory Study On Assessing The Effectiveness Of Ai And Accessibility Forums For Blind Users' Support Needs, Satwik Ram Kodandaram, Jiawei Zhou, Xiaojun Bi, Iv Ramakrishnan, Vikas Ashok
Finding The Signal In The Noise: An Exploratory Study On Assessing The Effectiveness Of Ai And Accessibility Forums For Blind Users' Support Needs, Satwik Ram Kodandaram, Jiawei Zhou, Xiaojun Bi, Iv Ramakrishnan, Vikas Ashok
Computer Science Faculty Publications
Accessibility forums and, more recently, generative AI tools have become vital resources for blind users seeking solutions to computer-interaction issues and learning about new assistive technologies, screen reader features, tutorials, and software updates. Understanding user experiences with these resources is essential for identifying and addressing persistent support gaps. Towards this, we interviewed 14 blind users who regularly engage with forums and GenAI tools. Findings revealed that forums often overwhelm users with multiple overlapping topics, redundant or irrelevant content, and fragmented responses that must be mentally pieced together, increasing cognitive load. GenAI tools, while offering more direct assistance, introduce new barriers …
A Comparative Analysis Of Explainable Ai (Xai) Techniques For Transparent And Reliable Image Classification, Sovon Chakraborty, Shakib Mahmud Dipto, Kevin R. Pilkiewicz, Michael L. Mayo, Pratip Rana
A Comparative Analysis Of Explainable Ai (Xai) Techniques For Transparent And Reliable Image Classification, Sovon Chakraborty, Shakib Mahmud Dipto, Kevin R. Pilkiewicz, Michael L. Mayo, Pratip Rana
Computer Science Faculty Publications
Evaluating the trustworthiness of black-box machine learning models remains a significant methodological challenge. Their lack of transparency and interpretability limits applicability, because stakeholders often seek transparency before trusting the results of black-box machine learning models. Explainable AI (XAI) methods provide for human-understandable justifications and informed decision-making of these black-box architectures. Therefore, it is imperative to select the proper XAI model tailored to specific tasks. In this research, we focus on examining four XAI techniques: PEEK, LRP, GRAD-CAM, and LIME to understand how they perform against each other for image classification tasks. We evaluate the performance, robustness, generalizability, noise stability, and …
Susceptibility To High-Fidelity Misinformation: An Eye-Tracking Analysis, Yasasi Abeysinghe, Gavindya Jayawardena, Enkelejda Kasneci, Sampath Jayarathna
Susceptibility To High-Fidelity Misinformation: An Eye-Tracking Analysis, Yasasi Abeysinghe, Gavindya Jayawardena, Enkelejda Kasneci, Sampath Jayarathna
Computer Science Faculty Publications
With the rise of online misinformation and AI-generated text, understanding human perception of news truthfulness is critical. In this study, we examine visual attention and cognitive processing using eye-tracking measures as individuals read fake and real news articles sharing nearly identical structure and imagery, differing only in subtle textual changes. Using the public FakeNewsPerception dataset, we analyze advanced gaze measures, including scanpaths, AOI transitions, and luminance-corrected pupil measures, beyond basic gaze features, in relation to news truthfulness and perceived believability. Results show that, given the high fidelity of the fake news, readers exhibited comparable visual scanning patterns, attention allocation across …
Cognitive Prosthetic: An Ai-Enabled Multimodal System For Episodic Recall In Knowledge Work, Lawrence Obiuwevwi, Krzystof J. Rechowicz, Vikas Ashok, Sachin Shetty, Sampath Jayarathna
Cognitive Prosthetic: An Ai-Enabled Multimodal System For Episodic Recall In Knowledge Work, Lawrence Obiuwevwi, Krzystof J. Rechowicz, Vikas Ashok, Sachin Shetty, Sampath Jayarathna
Computer Science Faculty Publications
Modern knowledge workplaces increasingly strain human episodic memory as individuals navigate fragmented attention, overlapping meetings, and multimodal information streams. Existing workplace tools provide partial support through note-taking or analytics but rarely integrate cognitive, physiological, and attentional context into retrievable memory representations. This paper presents the Cognitive Prosthetic Multimodal System (CPMS)—an AI-enabled proof-of-concept designed to support episodic recall in knowledge work through structured episodic capture and natural language retrieval. CPMS synchronizes speech transcripts, physiological signals, and gaze behavior into temporally aligned, JSON-based episodic records processed locally for privacy. Beyond data logging, the system includes a web-based retrieval interface that allows users …
Generalized Visual Relation Detection With Diffusion Models, Kaifeng Gao, Siqi Chen, Hanwang Zhang, Jun Xiao, Yueting Zhuang, Qianru Sun
Generalized Visual Relation Detection With Diffusion Models, Kaifeng Gao, Siqi Chen, Hanwang Zhang, Jun Xiao, Yueting Zhuang, Qianru Sun
Research Collection School Of Computing and Information Systems
Visual relation detection (VRD) aims to identify relationships (or interactions) between object pairs in an image. Although recent VRD models have achieved impressive performance, they are all restricted to pre-defined relation categories, while failing to consider the semantic ambiguity characteristic of visual relations. Unlike objects, the appearance of visual relations is always subtle and can be described by multiple predicate words from different perspectives, e.g., “ride” can be depicted as “race” and “sit on”, from the sports and spatial position views, respectively. To this end, we propose to model visual relations as continuous embeddings, and design diffusion models to achieve …
Trajlens: Visual Analysis For Constructing Cell Developmental Trajectories In Cross-Sample Exploration, Qipeng Wang, Shaolun Ruan, Rui Sheng, Yong Wang, Min Zhu, Huamin Qu
Trajlens: Visual Analysis For Constructing Cell Developmental Trajectories In Cross-Sample Exploration, Qipeng Wang, Shaolun Ruan, Rui Sheng, Yong Wang, Min Zhu, Huamin Qu
Research Collection School Of Computing and Information Systems
Constructing cell developmental trajectories is a critical task in single-cell RNA sequencing (scRNA-seq) analysis, enabling the inference of potential cellular progression paths. However, current automated methods are limited to establishing cell developmental trajectories within individual samples, necessitating biologists to manually link cells across samples to construct complete cross-sample evolutionary trajectories that consider cellular spatial dynamics. This process demands substantial human effort due to the complex spatial correspondence between each pair of samples. To address this challenge, we first proposed a GNN-based model to predict cross-sample cell developmental trajectories. We then developed TrajLens, a visual analytics system that supports biologists in …
Qualitative Study For Llm-Assisted Design Study Process: Strategies, Challenges, And Roles, Shaolun Ruan, Rui Sheng, Xiaolin Wen, Jiachen Wang, Tianyi Zhang, Yong Wang, Tim Dwyer, Jiannan Li
Qualitative Study For Llm-Assisted Design Study Process: Strategies, Challenges, And Roles, Shaolun Ruan, Rui Sheng, Xiaolin Wen, Jiachen Wang, Tianyi Zhang, Yong Wang, Tim Dwyer, Jiannan Li
Research Collection School Of Computing and Information Systems
Design studies aim to develop visualization solutions for real-world problems across various application domains. Recently, the emergence of large language models (LLMs) has introduced new opportunities to enhance the design study process, providing capabilities such as creative problem-solving, data handling, and insightful analysis. However, despite their growing popularity, there remains a lack of systematic understanding of how LLMs can effectively assist researchers in visualization-specific design studies. In this paper, we conducted a rnulti-stage qualitative study to fill this gap, which involved 30 design study researchers from diverse backgrounds and expertise levels. Through in-depth interviews and carefully-designed questionnaires, we investigated strategies …
Cross-Modal Proxy Evolving For Ood Detection With Vision-Language Models, Hao Tang, Yu Liu, Shuanglin Yan, Fei Shen, Shengfeng He, Jing Qin
Cross-Modal Proxy Evolving For Ood Detection With Vision-Language Models, Hao Tang, Yu Liu, Shuanglin Yan, Fei Shen, Shengfeng He, Jing Qin
Research Collection School Of Computing and Information Systems
Reliable zero-shot detection of out-of-distribution (OOD) inputs is critical for deploying vision-language models in open-world settings. However, the lack of labeled negatives in zero-shot OOD detection necessitates proxy signals that remain effective under distribution shift. Existing negative-label methods rely on a fixed set of textual proxies, which (i) sparsely sample the semantic space beyond in-distribution (ID) classes and (ii) remain static while only visual features drift, leading to cross-modal misalignment and unstable predictions. In this paper, we propose CoEvo, a training- and annotation-free test-time framework that performs bidirectional, sample-conditioned adaptation of both textual and visual proxies. Specifically, CoEvo introduces a …
Seeing The Invisible Load: Xr+ Multimodal Sensing For Cognitive Ergonomics In Industrial Training, Jessica M. Johnson, Andwele Grant
Seeing The Invisible Load: Xr+ Multimodal Sensing For Cognitive Ergonomics In Industrial Training, Jessica M. Johnson, Andwele Grant
Virginia Digital Maritime Center (VDMC) Faculty Publications
Extended reality (XR) technologies are increasingly positioned as disruptive Industry 5.0 tools for human-centric industrial training and intelligent human–system integration. Coupled with multimodal sensing (eye tracking, EEG, HRV, GSR, and other physiological signals), XR environments promise to make otherwise invisible cognitive demands observable, especially for novice trainees entering complex industrial settings. Yet the evidence base is fragmented: (1) there is no quantitative synthesis of the cognitive ergonomics benefits of XR plus sensing; (2) little is known about which XR–sensor configurations yield the strongest effects; (3) prior reviews rarely focus on industrial and manufacturing tasks; (4) multimodal signals are used predominantly …
Gui Test Migration Via Abstraction And Concretization, Yakun Zhang, Chen Liu, Xiaofei Xie, Yun Lin, Jin Song Dong, Dan Hao, Lu Zhang
Gui Test Migration Via Abstraction And Concretization, Yakun Zhang, Chen Liu, Xiaofei Xie, Yun Lin, Jin Song Dong, Dan Hao, Lu Zhang
Research Collection School Of Computing and Information Systems
GUI test migration aims to produce test cases with events and assertions to test specific functionalities of a target app. Existing migration approaches typically focus on the widget-mapping paradigm that maps widgets from source apps to target apps. However, since different apps may implement the same functionality in different ways, direct mapping may result in incomplete or buggy test cases, thus significantly impacting the effectiveness of testing the target functionality and the practical applicability of migration approaches.In this article, we propose a new migration paradigm (i.e., the abstraction-concretization paradigm) that first abstracts the test logic for the target functionality and …
Swimming In Uncertainty: Filling Data Gaps And Providing An Educational Platform For Beach Water Quality At Tybee Island, Georgia, Lukas Roberson
Swimming In Uncertainty: Filling Data Gaps And Providing An Educational Platform For Beach Water Quality At Tybee Island, Georgia, Lukas Roberson
College of Graduate Studies: Theses & Dissertations
@font-face {font-family:"Cambria Math"; panose-1:2 4 5 3 5 4 6 3 2 4; mso-font-charset:0; mso-generic-font-family:roman; mso-font-pitch:variable; mso-font-signature:-536870145 1107305727 0 0 415 0;}p.MsoNormal, li.MsoNormal, div.MsoNormal {mso-style-unhide:no; mso-style-qformat:yes; mso-style-parent:""; margin:0in; mso-pagination:widow-orphan; font-size:12.0pt; font-family:"Times New Roman",serif; mso-fareast-font-family:"Times New Roman";}.MsoChpDefault {mso-style-type:export-only; mso-default-props:yes; mso-font-kerning:0pt; mso-ligatures:none;}div.WordSection1 {page:WordSection1;}
Swimming in beaches water contaminated with high levels of bacteria can make you sick. Current monitoring at the public beaches on Tybee Island consists of weekly monitoring and enumeration of fecal indicator bacteria that takes 24 hours for results. If the number of bacteria exceed regulatory limits, a public health advisory is issued, and affected waters are retested until …
Purified Zero-Shot Sketch-Based Image Retrieval, Yang Zhou, Jingru Yang, Jin Wang, Kaixiang Huang, Guodong Lu, Shengfeng He
Purified Zero-Shot Sketch-Based Image Retrieval, Yang Zhou, Jingru Yang, Jin Wang, Kaixiang Huang, Guodong Lu, Shengfeng He
Research Collection School Of Computing and Information Systems
Sketches, as a new solution in multimedia systems that can replace natural language, are characterized by sparse visual cues such as simple strokes that differ significantly from natural images containing complex elements such as background, foreground, and texture. This misalignment poses substantial challenges for zero-shot sketch-based image retrieval (ZS-SBIR). Prior approaches match sketches to full images and tend to overlook redundant elements in natural images, leading to model distraction and semantic ambiguity. To address this issue, we introduce a distraction-agnostic framework, purified cross-domain matching (PuXIM), which operates on a straightforward principle: masking and matching. We devise a visual-cross-linguistic (VxL) sampler …
Voxvista: Enhancing Screen Reading Experience For Online User Comments, Yash Prakash, Akshay Kolgar Nayak, Mohammed Shoaib Alyaan, Sampath Jayarathna, Hae-Na Lee, Vikas Ashok
Voxvista: Enhancing Screen Reading Experience For Online User Comments, Yash Prakash, Akshay Kolgar Nayak, Mohammed Shoaib Alyaan, Sampath Jayarathna, Hae-Na Lee, Vikas Ashok
Computer Science Faculty Publications
Online discussions have become integral to how people exchange ideas, form opinions, and participate in collective deliberation. While sighted users can comfortably engage with online discussions, blind users who are dependent on screen readers are forced to listen to long threads narrated in a single, monotonic voice that lacks prosodic variation, rhythm, or emotion. This robotic auditory experience not only deteriorates the user engagement with the content but also increases cognitive strain, by making it difficult to remain attentive and discern meaning beyond literal words. In an interview study, most blind participants reported that monotonous narration hindered their ability to …
Gaze Transition Entropy And Automation Trust In Multitasking Workspace, Yusuke Yamani, Austin Jackson, Tetsuya Sato, Feyishola Ashimi, Michael S. Politowicz, Eric T. Chancey, Makoto Itoh
Gaze Transition Entropy And Automation Trust In Multitasking Workspace, Yusuke Yamani, Austin Jackson, Tetsuya Sato, Feyishola Ashimi, Michael S. Politowicz, Eric T. Chancey, Makoto Itoh
Psychology Faculty Publications
Safe flight operation requires visual scanning across multiple displays in a cockpit, which collectively represent the state of the aircraft and supporting automation. Trust is a crucial factor that drives human-automation interaction, and recent work has suggested a relationship between an operator's visual attention and automation trust. One index that captures predictability of eye movements between different areas of interest is gaze transition entropy. The current work reanalyzed data from Sato et al., which examined eye movement patterns and trust in automation associated with the system monitoring task of the Multi-Attribute Task Battery. Results showed credible positive correlations between the …
Tests Without Borders: A Global Approach To Measuring Visualization Literacy, Olivia A. Guess
Tests Without Borders: A Global Approach To Measuring Visualization Literacy, Olivia A. Guess
McKelvey School of Engineering Graduate Student Theses & Dissertations
Visualization literacy assessments shape how we understand people's ability to interpret data, yet most existing instruments embed Western datasets and assumptions that limit their relevance for global audiences. This thesis argues that because data is personal, assessments must also be culturally grounded. We introduce a unified framework for adapting the Mini-VLAT into 22 regionally responsive short-form assessments, each retaining the structure of the original test while incorporating datasets and scenarios tailored to specific regions around the world. To demonstrate how such adaptations can be customized and validated, we present a detailed case study of a Ghana-adapted Mini-VLAT, developed in collaboration …
Developing An Ai-Assisted Grading System Using Large Language Models, Andrei Modiga
Developing An Ai-Assisted Grading System Using Large Language Models, Andrei Modiga
MS in Computer Science Project Reports
We present a grading system that accelerates evaluation of open-ended student work across scanned and digital workflows. The system crops answer regions from PDFs, assigns submissions via OCR on identity regions only, and groups answers by visual semantics using a vision LLM. Instructors review and edit groups, apply rubric items once per group, and export grades from an on-screen table. The solution integrates Ghostscript rasterization, PdfPig page orchestration, SkiaSharp region extraction, Tesseract identity OCR, and GPT-4o Vision for grouping. We detail the architecture, token-budgeted batching strategy, and persistence design, then describe testing results for grouping quality, time-on-task, and usability. The …
My First Conversation With Chatgpt (February 22, 2023): Origins Of A Generative Dialogue, David Smith
My First Conversation With Chatgpt (February 22, 2023): Origins Of A Generative Dialogue, David Smith
Publications and Research
This working paper presents the first recorded interaction between the author and the generative AI system ChatGPT, written on February 22, 2023 during the initial weeks of a faculty sabbatical in Boston. The document preserves a complete and unedited transcript of an exploratory conversation conducted without predetermined research aims, marking the author’s first encounter with a large-language-model conversational interface. Although the exchange includes creative experimentation—including musical and poetic prompts—the discussion remains informal and wide-ranging, and no theoretical framework is articulated at this stage. Rather, this transcript is published as primary-source material documenting the moment of discovery and experimentation that precedes …
Decoding The Chameleon Game, Tri Dang '25, Hieu Tran, Brian T. Howard, Sutthirut Charoenphon, Dat Nguyen '25
Decoding The Chameleon Game, Tri Dang '25, Hieu Tran, Brian T. Howard, Sutthirut Charoenphon, Dat Nguyen '25
Student Research
The Chameleon game is a challenging word association activity where players are given a secret word and must respond with words relevant to that secret word. It requires strategic thinking and deduction. The Chameleon must cleverly guess the secret keyword in this game while avoiding suspicion. Our research aims to create an advanced artificial intelligence (AI) model that can play the Chameleon game from both perspectives: as the Chameleon and as a Human. This AI is designed to guess secret keywords based on the information provided by the players, choose the best strategies to avoid detection as the Chameleon, identify …
Developing Accessible Narrative-Based Stem Learning Software For K-6 Braille Display Users, Dylan Ravel, Daniel Tsivkovski, Brandon Foley, Maryam Etezad, Franceli Cibrian, Ariel Han, Rajeev Joshi
Developing Accessible Narrative-Based Stem Learning Software For K-6 Braille Display Users, Dylan Ravel, Daniel Tsivkovski, Brandon Foley, Maryam Etezad, Franceli Cibrian, Ariel Han, Rajeev Joshi
Student Scholar Symposium Abstracts and Posters
This research develops a free, accessible web application that enables K-6 students who are blind or visually impaired (BVI) to learn STEM concepts using refreshable braille displays. Currently, most online learning tools are not designed for BVI students, creating a significant educational barrier.
The application interfaces with commercial braille displays and uses narrative-based learning to make STEM content approachable and engaging. By presenting material as interactive stories, students can connect with concepts while developing braille reading skills. The curriculum design prioritizes accessibility through the Accessible Rich Internet Applications (ARIA) standards and screen reader support.
The goal is to provide BVI …
Visionglow: Evaluating Minimal-Disruption Smart-Home Control In Apple Vision Pro, Hongxiao Zheng
Visionglow: Evaluating Minimal-Disruption Smart-Home Control In Apple Vision Pro, Hongxiao Zheng
Dartmouth College Master’s Theses
Smart-home control in mixed-reality environments like Apple Vision Pro often relies on disruptive, application-based paradigms, such as using a smartphone or a windowed virtual interface. These methods create a “mode switch” that imposes cognitive load and pulls users from their primary tasks. We present VisionGlow, a minimal-disruption spatial interaction technique for Vision Pro. VisionGlow represents devices as spatially-anchored “orbs.” To control a device, the user looks at its orb and performs a pinch gesture, which invokes a compact, contextual control panel. We conducted a within-subjects study (N=18) comparing VisionGlow against two baselines: the standard Apple Home app on a smartphone …
The Future Is Now: Empowering Society Through Ai Literacy, Jason S. Wrench, Sanae Elmoudden
The Future Is Now: Empowering Society Through Ai Literacy, Jason S. Wrench, Sanae Elmoudden
Milne Open Textbooks
Artificial Intelligence (AI) is no longer a futuristic concept—it is the reality of the present. From the algorithms shaping our social media feeds to the generative tools transforming our workplaces, AI has permeated every aspect of modern life. The Future is Now moves beyond the hype to provide a comprehensive roadmap for understanding, navigating, and shaping this technological revolution.
Demystifying the Machine
This textbook serves as a user-friendly guide to the “black box” of AI. It breaks down complex technical concepts—from machine learning and neural networks to large language models—making them accessible to students across all disciplines. By establishing a …
Digital Reflections: Evaluating Body Dissatisfaction In Xr Through Eye- And Body-Tracked Virtual Humans, Deyrel Diaz
Digital Reflections: Evaluating Body Dissatisfaction In Xr Through Eye- And Body-Tracked Virtual Humans, Deyrel Diaz
All Dissertations
In an era where digital and physical realities increasingly intertwine, the perception of body image is undergoing a significant transformation. Traditional understandings of body dissatisfaction, long studied in relation to psychological distress and eating disorders, are now being reshaped by technologies such as Virtual Reality (VR), Augmented Reality (AR), and Artificially Intelligent (AI)- generated media. These technologies have introduced novel ways of experiencing and interacting with the human form, raising critical questions about their impact on self-perception and internalization of beauty standards.
As virtual representations become more prevalent in entertainment, social media, and interactive platforms, it is becoming more crucial …
Instance-Level Video Depth In Groups Beyond Occlusions, Yuan Liang, Yang Zhou, Ziming Sun, Tianyi Xiang, Guiqing Li, Shengfeng He
Instance-Level Video Depth In Groups Beyond Occlusions, Yuan Liang, Yang Zhou, Ziming Sun, Tianyi Xiang, Guiqing Li, Shengfeng He
Research Collection School Of Computing and Information Systems
Depth estimation in dynamic, multi-object scenes remains a major challenge, especially under severe occlusions. Existing monocular models, including foundation models, struggle with instance-wise depth consistency due to their reliance on global regression. We tackle this problem from two key aspects: data and methodology. First, we introduce the Group Instance Depth (GID) dataset, the first large-scale video depth dataset with instance-level annotations, featuring 101,500 frames from real-world activity scenes. GID bridges the gap between synthetic and real-world depth data by providing high-fidelity depth supervision for multi-object interactions. Second, we propose InstanceDepth, the first occlusion-aware depth estimation framework for multi-object environments. Our …