Open Access. Powered by Scholars. Published by Universities.®

Graphics and Human Computer Interfaces Commons™

Open Access. Powered by Scholars. Published by Universities.®

2,378 Full-Text Articles 4,459 Authors 1,194,747 Downloads 165 Institutions

All Articles in Graphics and Human Computer Interfaces

Faceted Search

2,378 full-text articles. Page 26 of 101.

Bias-Aware Gaze Uniformity Assessment In Group Images, Omkar Kulkarni 2024 University at Albany, State University of New York

Bias-Aware Gaze Uniformity Assessment In Group Images, Omkar Kulkarni

Electronic Theses & Dissertations (2024 - present)

Today, more than 5 billion photos are captured every day, with smartphones generating over 94\% of these images. However, despite advancements in technology, achieving aesthetically pleasing group photos remains challenging, especially when it comes to aligning the direction of everyone’s gaze. While current methods focus on facial features, they often fail to ensure consistent gaze direction. The introduction of the iPhone's Live mode, which captures a 1.5-second video snippet along with still images, complicates the selection of the best key photo due to its subjective nature and a lack of publicly available data, especially during the pandemic.

To address these …


Garbage In ≠ Garbage Out: Exploring Gan Resilience To Image Training Set Degradations, Nicholas Crino, Bruce A. Cox, Nathan B. Gaw 2024 Air Force Institute of Technology

Garbage In ≠ Garbage Out: Exploring Gan Resilience To Image Training Set Degradations, Nicholas Crino, Bruce A. Cox, Nathan B. Gaw

Faculty Publications

Generative Adversarial Networks (GANs) have received immense attention in recent years due to their ability to capture complex, high-dimensional data distributions without the need for extensive labeling. Since their conception in 2014, a wide array of GAN variants have been proposed featuring alternative architectures, optimizers, and loss functions with the goal of improving performance and training stability. This manuscript focuses on quantifying the resilience of a GAN architecture to specific modes of image degradation. We conduct systematic experimentation to empirically determine the effects of 10 fundamental image degradation modes, applied to the training image dataset, on the Fréchet inception distance …


Trust: The Feature That Vending Machines And Atms Share, But Simplygo Lacks, Sun Sun LIM 2024 Singapore Management University

Trust: The Feature That Vending Machines And Atms Share, But Simplygo Lacks, Sun Sun Lim

Research Collection College of Integrative Studies

The article discussed the intricacies of trust in the SimplyGo debacle and highlighted how the design of physical interfaces like vending machines and ATMs and digital interfaces from apps like Grab, Parking.sg and ShopBack have critical features to instil trust. People need to be reassured that their transactions have proceeded as they should, and thay have not been short-changed.


Breaking Down Computer Networking Instructional Videos: Automatic Summarization With Video Attributes And Language Models, Totok Sukardiyono, Muhammad Irfan Luthfi, Nisa Dwi Septiyanti 2023 Department of Electronics and Informatics Engineering Education, Faculty of Engineering, Universitas Negeri Yogyakarta, Indonesia

Breaking Down Computer Networking Instructional Videos: Automatic Summarization With Video Attributes And Language Models, Totok Sukardiyono, Muhammad Irfan Luthfi, Nisa Dwi Septiyanti

Elinvo (Electronics, Informatics, and Vocational Education)

Instructional videos have become a popular tool for teaching complex topics in computer networking. However, these videos can often be lengthy and time-consuming, making it difficult for learners to obtain the key information they need. In this study, we propose an approach that leverages automatic summarization and language models to generate concise and informative summaries of instructional videos. To enhance the performance of the summarization algorithm, we also incorporate video attributes that provide contextual information about the video content. Using a dataset of computer networking tutorials, we evaluate the effectiveness of the proposed method and show that it significantly improves …


Differences In Software Usability Level Based On User Background, Abdur Rohman Sholeh, Agung Fatwanto 2023 UIN Sunan Kalijaga Yogyakarta, Indonesia

Differences In Software Usability Level Based On User Background, Abdur Rohman Sholeh, Agung Fatwanto

Elinvo (Electronics, Informatics, and Vocational Education)

The development of software must consider usability as one of its key success indicators. Relatively few studies discuss the factors influencing usability, including users' backgrounds. The purpose of this research was to investigate the impact of user background, specifically gender, class (year of college admission), and frequency of use on the rating of usability. This research utilized a descriptive quantitative method with instruments: the usability matrix of the Computer System Usability Questionnaire (CSUQ), the System Usability Scale (SUS), the Usability Metric for User Experience (UMUX), and the Net Promoter Score (NPS). The research object was the Learning Management System (LMS) …


Usability Of Mobile Application For Implementing Genetic Counselling Intervention Among Thalassemia Patients And Caregivers: A Case Study Of Cyber Gen, Henri Setiawan, Nur Hidayat, Atun Farihatun, Marlina Indriastuti, Rudi Kurniawan, Andan Firmansyah, Esti Andarini, Yudisa Diaz Lutfi Sandi 2023 STIKes Muhammadiyah Ciamis, Indonesia

Usability Of Mobile Application For Implementing Genetic Counselling Intervention Among Thalassemia Patients And Caregivers: A Case Study Of Cyber Gen, Henri Setiawan, Nur Hidayat, Atun Farihatun, Marlina Indriastuti, Rudi Kurniawan, Andan Firmansyah, Esti Andarini, Yudisa Diaz Lutfi Sandi

Elinvo (Electronics, Informatics, and Vocational Education)

Utilization of communication and information technology has been widely used in the health sector, especially nursing. As one of the nursing interventions for thalassemia patients and caregivers, genetic counseling is not only done face to face but can use android-based telenursing facilities through the complete features available in the Cyber Gen application. This study aims to measure the usability level of Cyber Gen application as an indirect genetic counseling medium for thalassemia patients. This application was developed with four main services: basic information about diseases, consultation rooms, social support, and direct surveys. This application is built using the Flutter Framework, …


Reducing Food Scarcity: The Benefits Of Urban Farming, S.A. Claudell, Emilio Mejia 2023 Brigham Young University

Reducing Food Scarcity: The Benefits Of Urban Farming, S.A. Claudell, Emilio Mejia

Journal of Nonprofit Innovation

Urban farming can enhance the lives of communities and help reduce food scarcity. This paper presents a conceptual prototype of an efficient urban farming community that can be scaled for a single apartment building or an entire community across all global geoeconomics regions, including densely populated cities and rural, developing towns and communities. When deployed in coordination with smart crop choices, local farm support, and efficient transportation then the result isn’t just sustainability, but also increasing fresh produce accessibility, optimizing nutritional value, eliminating the use of ‘forever chemicals’, reducing transportation costs, and fostering global environmental benefits.

Imagine Doris, who is …


Deep Learning Image Analysis To Isolate And Characterize Different Stages Of S-Phase In Human Cells, Kevin A. Boyd, Rudranil Mitra, John Santerre, Christopher L. Sansam 2023 SMU

Deep Learning Image Analysis To Isolate And Characterize Different Stages Of S-Phase In Human Cells, Kevin A. Boyd, Rudranil Mitra, John Santerre, Christopher L. Sansam

SMU Data Science Review

Abstract. This research used deep learning for image analysis by isolating and characterizing distinct DNA replication patterns in human cells. By leveraging high-resolution microscopy images of multiple cells stained with 5-Ethynyl-2′-deoxyuridine (EdU), a replication marker, this analysis utilized Convolutional Neural Networks (CNNs) to perform image segmentation and to provide robust and reliable classification results. First multiple cells in a field of focus were identified using a pretrained CNN called Cellpose. After identifying the location of each cell in the image a python script was created to crop out each cell into individual .tif files. After careful annotation, a CNN was …


Adaptable Object And Animation System For Game Development, Isaiah Turner 2023 Western Kentucky University

Adaptable Object And Animation System For Game Development, Isaiah Turner

Masters Theses & Specialist Projects

In contemporary times, video games have swiftly evolved into a prominent medium, excelling in both entertainment and narrative delivery, positioning themselves as significant rivals to traditional forms such as film and theater. The burgeoning popularity of gaming has led to a surge in aspiring game developers seeking to craft their own creations, driven by both commercial aspirations and personal passion. However, a common challenge faced by these individuals involves the considerable time investment required to acquire essential skills and establish a foundational framework for their projects. Accessible game development engines that offer a diverse range of fundamental features play a …


Mermaid: A Dataset And Framework For Multimodal Meme Semantic Understanding, Shaun TOH, Adriel KUEK, Wen Haw CHONG, Roy Ka Wei LEE 2023 Singapore Management University

Mermaid: A Dataset And Framework For Multimodal Meme Semantic Understanding, Shaun Toh, Adriel Kuek, Wen Haw Chong, Roy Ka Wei Lee

Research Collection School Of Computing and Information Systems

Memes are widely used to convey cultural and societal issues and have a significant impact on public opinion. However, little work has been done on understanding and explaining the semantics expressed in multimodal memes. To fill this research gap, we introduce MERMAID, a dataset consisting of 3,633 memes annotated with their entities and relations, and propose a novel MERF pipeline that extracts entities and their relationships in memes. Our framework combines state-of-the-art techniques from natural language processing and computer vision to extract text and image features and infer relationships between entities in memes. We evaluate the proposed framework on a …


Depwignn: A Depth-Wise Graph Neural Network For Multi-Hop Spatial Reasoning In Text, Shuaiyi LI, Yang DENG, Wai LAM 2023 Singapore Management University

Depwignn: A Depth-Wise Graph Neural Network For Multi-Hop Spatial Reasoning In Text, Shuaiyi Li, Yang Deng, Wai Lam

Research Collection School Of Computing and Information Systems

Spatial reasoning in text plays a crucial role in various real-world applications. Existing approaches for spatial reasoning typically infer spatial relations from pure text, which overlook the gap between natural language and symbolic structures. Graph neural networks (GNNs) have showcased exceptional proficiency in inducing and aggregating symbolic structures. However, classical GNNs face challenges in handling multi-hop spatial reasoning due to the over-smoothing issue, i.e., the performance decreases substantially as the number of graph layers increases. To cope with these challenges, we propose a novel Depth-Wise Graph Neural Network (DepWiGNN). Specifically, we design a novel node memory scheme and aggregate the …


Unifying Text, Tables, And Images For Multimodal Question Answering, Haohao LUO, Ying SHEN, Yang DENG 2023 Singapore Management University

Unifying Text, Tables, And Images For Multimodal Question Answering, Haohao Luo, Ying Shen, Yang Deng

Research Collection School Of Computing and Information Systems

Multimodal question answering (MMQA), which aims to derive the answer from multiple knowledge modalities (e.g., text, tables, and images), has received increasing attention due to its board applications. Current approaches to MMQA often rely on single-modal or bi-modal QA models, which limits their ability to effectively integrate information across all modalities and leverage the power of pre-trained language models. To address these limitations, we propose a novel framework called UniMMQA, which unifies three different input modalities into a text-to-text format by employing position-enhanced table linearization and diversified image captioning techniques. Additionally, we enhance cross-modal reasoning by incorporating a multimodal rationale …


Self-Supervised Pseudo Multi-Class Pre-Training For Unsupervised Anomaly Detection And Segmentation In Medical Images, Yu TIAN, Fengbei LIU, Guansong PANG, Yuanhong CHEN, Yuyuan LIU, Johan W. VERJANS, Rajvinder SINGH, Gustavo CARNEIRO 2023 Singapore Management University

Self-Supervised Pseudo Multi-Class Pre-Training For Unsupervised Anomaly Detection And Segmentation In Medical Images, Yu Tian, Fengbei Liu, Guansong Pang, Yuanhong Chen, Yuyuan Liu, Johan W. Verjans, Rajvinder Singh, Gustavo Carneiro

Research Collection School Of Computing and Information Systems

Unsupervised anomaly detection (UAD) methods are trained with normal (or healthy) images only, but during testing, they are able to classify normal and abnormal (or disease) images. UAD is an important medical image analysis (MIA) method to be applied in disease screening problems because the training sets available for those problems usually contain only normal images. However, the exclusive reliance on normal images may result in the learning of ineffective low-dimensional image representations that are not sensitive enough to detect and segment unseen abnormal lesions of varying size, appearance, and shape. Pre-training UAD methods with self-supervised learning, based on computer …


Graph Contrastive Learning With Stable And Scalable Spectral Encoding, Deyu BO, Yuan FANG, Yang LIU, Chuan SHI 2023 Singapore Management University

Graph Contrastive Learning With Stable And Scalable Spectral Encoding, Deyu Bo, Yuan Fang, Yang Liu, Chuan Shi

Research Collection School Of Computing and Information Systems

Graph contrastive learning (GCL) aims to learn representations by capturing the agreements between different graph views. Traditional GCL methods generate views in the spatial domain, but it has been recently discovered that the spectral domain also plays a vital role in complementing spatial views. However, existing spectral-based graph views either ignore the eigenvectors that encode valuable positional information, or suffer from high complexity when trying to address the instability of spectral features. To tackle these challenges, we first design an informative, stable, and scalable spectral encoder, termed EigenMLP, to learn effective representations from the spectral features. Theoretically, EigenMLP is invariant …


Video Sentiment Analysis For Child Safety, Yee Sen TAN, Nicole Anne Huiying TEO, Ezekiel En Zhe GHE, Jolie Zhi Yi FONG, Zhaoxia WANG 2023 Singapore Management University

Video Sentiment Analysis For Child Safety, Yee Sen Tan, Nicole Anne Huiying Teo, Ezekiel En Zhe Ghe, Jolie Zhi Yi Fong, Zhaoxia Wang

Research Collection School Of Computing and Information Systems

The proliferation of online video content underscores the critical need for effective sentiment analysis, particularly in safeguarding children from potentially harmful material. This research addresses this concern by presenting a multimodal analysis method for assessing video sentiment, categorizing it as either positive (child-friendly) or negative (potentially harmful). This method leverages three key components: text analysis, facial expression analysis, and audio analysis, including music mood analysis, resulting in a comprehensive sentiment assessment. Our evaluation results validate the effectiveness of this approach, making significant contributions to the field of video sentiment analysis and bolstering child safety measures. This research serves as a …


Mrim: Lightweight Saliency-Based Mixed-Resolution Imaging For Low-Power Pervasive Vision, Jiyan WU, Vithurson SUBASHARAN, Minh Anh Tuan TRAN, Kasun Pramuditha GAMLATH, Archan MISRA 2023 Singapore Management University

Mrim: Lightweight Saliency-Based Mixed-Resolution Imaging For Low-Power Pervasive Vision, Jiyan Wu, Vithurson Subasharan, Minh Anh Tuan Tran, Kasun Pramuditha Gamlath, Archan Misra

Research Collection School Of Computing and Information Systems

While many pervasive computing applications increasingly utilize real-time context extracted from a vision sensing infrastructure, the high energy overhead of DNN-based vision sensing pipelines remains a challenge for sustainable in-the-wild deployment. One common approach to reducing such energy overheads is the capture and transmission of lower-resolution images to an edge node (where the DNN inferencing task is executed), but this results in an accuracy-vs-energy tradeoff, as the DNN inference accuracy typically degrades with a drop in resolution. In this work, we introduce MRIM, a simple but effective framework to tackle this tradeoff. Under MRIM, the vision sensor platform first executes …


Enhanced Privacy-Enabled Face Recognition Using Κ-Identity Optimization, Ryan Karl 2023 University of Nebraska-Lincoln

Enhanced Privacy-Enabled Face Recognition Using Κ-Identity Optimization, Ryan Karl

Department of Electrical and Computer Engineering: Dissertations, Theses, and Student Research

Facial recognition is becoming more and more prevalent in the daily lives of the common person. Law enforcement utilizes facial recognition to find and track suspects. The newest smartphones have the ability to unlock using the user's face. Some door locks utilize facial recognition to allow correct users to enter restricted spaces. The list of applications that use facial recognition will only increase as hardware becomes more cost-effective and more computationally powerful. As this technology becomes more prevalent in our lives, it is important to understand and protect the data provided to these companies. Any data transmitted should be encrypted …


Developing Detection And Mapping Of Roads Within Various Forms Of Media Using Opencv, Jordan C. Lyle 2023 University of Arkansas, Fayetteville

Developing Detection And Mapping Of Roads Within Various Forms Of Media Using Opencv, Jordan C. Lyle

Computer Science and Computer Engineering Undergraduate Honors Theses

OpenCV, and Computer Vision in general, has been a Computer Science topic that has interested me for a long time while completing my Bachelor’s degree at the University of Arkansas. As a result of this, I ended up choosing to utilize OpenCV in order to complete the task of detecting road-lines and mapping roads when given a wide variety of images. The purpose of my Honors research and this thesis is to detail the process of creating an algorithm to detect the road-lines such that the results are effective and instantaneous, as well as detail how Computer Vision can be …


Clueless: Revolutionizing Sustainable Fashion And Combating Overconsumption, Tanya Ravichandran 2023 California Polytechnic State University, San Luis Obispo

Clueless: Revolutionizing Sustainable Fashion And Combating Overconsumption, Tanya Ravichandran

Graphic Communication

“Clueless” revolutionizes sustainable fashion by combating wardrobe overconsumption and the industry’s carbon footprint, using AI to suggest personalized outfits from existing wardrobes tailored to weather and wear history. It enhances user engagement through features like outfit ‘shuffle’ and provides insights into wardrobe utilization and carbon impact.

It’s more than an app; it’s a step towards a greener wardrobe and a healthier planet.


Triple Helix: Ai-Artist-Audience Collaboration In A Performative Art Experience, Xuedan Zou 2023 Department of Computer Science, Dartmouth College

Triple Helix: Ai-Artist-Audience Collaboration In A Performative Art Experience, Xuedan Zou

Dartmouth College Master’s Theses

Imagine an art exhibition that morphs its content according to the audience’s experience like a chameleon, reflecting the audience’s mind and culture and turning the artist’s exhibition into the viewer’s. But when the viewers leave, the work fades back to the creator’s original work and waits for the next audience. In this project, my team introduced an interactive exhibition called "Triple Helix," where audience members were provided the opportunity to alter the artworks created by the artist, thus imbuing them with their own perspectives. This interactive exhibition was held at three physical-locations and online, and a comprehensive user study was …


Digital Commons powered by bepress