Open Access. Powered by Scholars. Published by Universities.®
Graphics and Human Computer Interfaces Commons™
Open Access. Powered by Scholars. Published by Universities.®
- Institution
-
- Singapore Management University (313)
- University of Dayton (48)
- University of Arkansas, Fayetteville (10)
- University of Malaya (8)
- San Jose State University (7)
-
- Technological University Dublin (7)
- City University of New York (CUNY) (6)
- California Polytechnic State University, San Luis Obispo (5)
- Southern Adventist University (5)
- St. Mary's University (5)
- Institute of Business Administration (4)
- University of Nevada, Las Vegas (4)
- California State University, San Bernardino (3)
- Montclair State University (3)
- University of Nebraska at Omaha (3)
- Dakota State University (2)
- Governors State University (2)
- Minnesota State University Moorhead (2)
- Nova Southeastern University (2)
- Old Dominion University (2)
- Rochester Institute of Technology (2)
- The University of Akron (2)
- University of Nebraska - Lincoln (2)
- University of South Carolina (2)
- Arkansas Tech University (1)
- Ateneo de Manila University (1)
- Beirut Arab University (1)
- Bridgewater State University (1)
- Brigham Young University (1)
- Coastal Carolina University (1)
- Keyword
-
- Gamification (8)
- Deep learning (7)
- Visualization (7)
- Machine Learning (6)
- Collaboration (5)
-
- Deep Learning (5)
- Face recognition (5)
- Multimodal (5)
- Usability (5)
- Data visualization (4)
- Database (4)
- Education (4)
- Eye tracking (4)
- Few-shot learning (4)
- Food recognition (4)
- Graph neural networks (4)
- Human-computer interaction (4)
- Knowledge Graph (4)
- Recipe retrieval (4)
- Recommendation (4)
- Trust (4)
- Virtual worlds (4)
- Algorithms (3)
- Avatars (3)
- Click-through data (3)
- Clustering (3)
- Computer science (3)
- Databases (3)
- Design (3)
- E-commerce (3)
- Publication Year
- Publication
-
- Research Collection School Of Computing and Information Systems (310)
- Computer Science Faculty Publications (32)
- MAICS: The Modern Artificial Intelligence and Cognitive Science Conference (12)
- Student Works (2000-2009) (8)
- Graduate Theses and Dissertations (5)
-
- Conference papers (4)
- Articles (3)
- Campus Research Month (3)
- College of Engineering: Graduate Celebration Programs (3)
- Computer Science and Computer Engineering Undergraduate Honors Theses (3)
- Department of Computer Science Faculty Scholarship and Creative Works (3)
- Dissertations and Theses Collection (Open Access) (3)
- International Conference on Information and Communication Technologies (3)
- Master's Projects (3)
- Presentations - 2026 (3)
- All Capstone Projects (2)
- CCAC Theses and Dissertations (2)
- Computer Engineering (2)
- Computer Science Working Papers (2)
- Computer Science and Software Engineering (2)
- Dissertations, Theses, and Capstone Projects (2)
- Electronic Theses, Projects, and Dissertations (2)
- Posters - 2026 (2)
- Publications (2)
- Publications and Research (2)
- Research & Publications (2)
- SWITCH (2)
- Student Academic Conference (2)
- Theses and Dissertations (2)
- Theses/Capstones/Creative Projects (2)
- Publication Type
- File Type
Articles 1 - 30 of 473
Full-Text Articles in Graphics and Human Computer Interfaces
From Data To Decision-Making: The Role Of Local Digital Twins In Cross-Domain Management Within Municipalities – A Research-In-Progress Study In Veenendaal, Diana M.E. Boekman, Koen Smit, Guido Ongena, Rob Peters
From Data To Decision-Making: The Role Of Local Digital Twins In Cross-Domain Management Within Municipalities – A Research-In-Progress Study In Veenendaal, Diana M.E. Boekman, Koen Smit, Guido Ongena, Rob Peters
Communications of the IIMA
Municipalities are facing increasingly complex, interconnected challenges in areas like housing, climate adaptation, mobility, and social policy. Local Digital Twins (LDTs) are seen as a promising tool to make this complexity more understandable and support decision-making. At the same time, both literature and practice show that few initiatives get past the pilot phase, even though getting through that phase is essential for successful long-term adoption.
This paper presents a research-in-progress study on the development and application of an implementation method for LDT technology within the municipality of Veenendaal, based on human values rather than driven by technological possibilities. Based on …
A Pruning-Based Question-Answering For Interactive Video Search: A Simple Baseline, Yu Tong Cheng, Phuong Anh Nguyen, Chong-Wah Ngo
A Pruning-Based Question-Answering For Interactive Video Search: A Simple Baseline, Yu Tong Cheng, Phuong Anh Nguyen, Chong-Wah Ngo
Research Collection School Of Computing and Information Systems
There are various factors affecting the performance of video search. An imprecise query will enlarge search space and reduce the discriminative power of ranking functions. This problem is further exacerbated by the presence of numerous visually or semantically similar videos in large datasets. Consequently, users need to painstakingly browse through many highly similar candidates to locate the search target, leading to increased cognitive load and inefficient searching. Ideally, engaging users through interactive questioning to resolve uncertainties in the search process is an effective strategy for progressively narrowing down the search space. However, despite rapid advances in deep learning, generating informative …
Happycal: Designing Text And Image-Based Supports For Savouring Positive Work Experiences, Molly Stewart, Minghao Cai, Anthony Tang, Sam Liu, Chris Mosunic, Sowmya Somanath
Happycal: Designing Text And Image-Based Supports For Savouring Positive Work Experiences, Molly Stewart, Minghao Cai, Anthony Tang, Sam Liu, Chris Mosunic, Sowmya Somanath
Research Collection School Of Computing and Information Systems
Savouring positive work experiences can promote positive affect and well-being at work, yet there is limited guidance on how digital applications can support workers to engage in savouring. We developed HappyCal, a work-focused savouring application offering two forms of savouring support: text-based, a common modality in workplace reflection tools, and images, a largely unexplored approach in work-related savouring. We conducted an exploratory qualitative study where participants (N=36) used HappyCal over five days and engaged in savouring through either a text-only modality (n=17) or text input paired with image output (n=19). We found that (1) participants in both groups reported heightened …
Harvest Scanner, Alexander Murphy
Harvest Scanner, Alexander Murphy
Posters - 2026
In current times, people can find themselves at the whims of markets and may be spending more than they realize or want to on regular, everyday goods. New tools can help users keep track of the goods they are paying for. Harvest Scanner was developed to scan and track local grocery prices from stores using their publicly available website information. It was developed in Python using PyQt5 for GUI. The database is stored as an SQL file with Python using SQLite engine. Users will be able to view local grocery prices in a database interface (GUI). There are many features …
Topshelf, Ayden Jay Soliz
Topshelf, Ayden Jay Soliz
Posters - 2026
With so many great video games releasing each year, it becomes challenging to keep up with the latest. Players find it difficult to maintain an updated list of future games to play, and many existing online trackers have become too complicated to use. TopShelf is designed to be a simple video game backlogging website that will track games for the player. By connecting to an online video game database API, users can add/drop games from their personal list and enable tracking and receive emails for platform releasing. Gamers can leave all the tracking and updates responsibilities to TopShelf
Match-A-Fit, Adan Diaz De Leon, Juan Marco Saca Dada, Brianna Mendoza, Arsalan Kataneh, Theophile Nsabimana, Pedro Jacobo
Match-A-Fit, Adan Diaz De Leon, Juan Marco Saca Dada, Brianna Mendoza, Arsalan Kataneh, Theophile Nsabimana, Pedro Jacobo
Presentations - 2026
Welcome to Match-a-Fit! Match-a-Fit is an iOS application that allows the user to create a digital closet by uploading images of their clothing items. With AI, the program can generate outfits based on the digital closet, the time, and the occasion. Match-a-Fit’s purpose is designed to help users who struggle to get ready, run out of time, or can’t decide on an outfit, by easily generating outfit options based on the occasion.
Foxbuddy, Luis Eduardo Garza Jr.
Foxbuddy, Luis Eduardo Garza Jr.
Presentations - 2026
Problem:
•Many people are still unprepared incase of an emergency. (42%-46% are prepared for an emergency)
•Supplies can be scattered, expired, or forgotten. •Reliable guidance is often not easy to access.
Chopchop: The Digital Cookbook, Dominc Mcdevitt, Shane Misley, Katie Cerda, Kobie Henson, Adolfo Duran
Chopchop: The Digital Cookbook, Dominc Mcdevitt, Shane Misley, Katie Cerda, Kobie Henson, Adolfo Duran
Presentations - 2026
Problem With traditional recipe organization methods,
● Recipes are scattered across paper, PDFs, Word docs, and notes
● Paper recipes can be lost, damaged, or left at home
● Digital recipes are difficult to edit, store, and organize
● Manually typing or updating recipes is time-consuming
● Sharing recipes is inconvenient and often confusing
● Limited or inconsistent cloud access reduces accessibility
● Formatting is messy and inconsistent across platforms
● No simple, centralized system for managing recipes
Cellscout: Visual Analytics For Mining Biomarkers In Cell State Discovery, Rui Sheng, Zelin Zang, Jiachen Wang, Yan Luo, Zixin Chen, Yan Zhou, Shaolun Ruan, Huamin Qu
Cellscout: Visual Analytics For Mining Biomarkers In Cell State Discovery, Rui Sheng, Zelin Zang, Jiachen Wang, Yan Luo, Zixin Chen, Yan Zhou, Shaolun Ruan, Huamin Qu
Research Collection School Of Computing and Information Systems
Cell state discovery is crucial for understanding biological systems and enhancing medical outcomes. A key aspect of this process is identifying distinct biomarkers that define specific cell states. However, difficulties arise from the co-discovery process of cell states and biomarkers: biologists often use dimensionality reduction to visualize cells in a two-dimensional space. Then they usually interpret visually clustered cells as distinct states, from which they seek to identify unique biomarkers. However, this assumption is often this assumption often fails to hold due to internal inconsistencies in a cluster, making the process trial-and-error and highly uncertain. Therefore, biologists urgently need effective …
Swimming In Uncertainty: Filling Data Gaps And Providing An Educational Platform For Beach Water Quality At Tybee Island, Georgia, Lukas Roberson
Swimming In Uncertainty: Filling Data Gaps And Providing An Educational Platform For Beach Water Quality At Tybee Island, Georgia, Lukas Roberson
College of Graduate Studies: Theses & Dissertations
@font-face {font-family:"Cambria Math"; panose-1:2 4 5 3 5 4 6 3 2 4; mso-font-charset:0; mso-generic-font-family:roman; mso-font-pitch:variable; mso-font-signature:-536870145 1107305727 0 0 415 0;}p.MsoNormal, li.MsoNormal, div.MsoNormal {mso-style-unhide:no; mso-style-qformat:yes; mso-style-parent:""; margin:0in; mso-pagination:widow-orphan; font-size:12.0pt; font-family:"Times New Roman",serif; mso-fareast-font-family:"Times New Roman";}.MsoChpDefault {mso-style-type:export-only; mso-default-props:yes; mso-font-kerning:0pt; mso-ligatures:none;}div.WordSection1 {page:WordSection1;}
Swimming in beaches water contaminated with high levels of bacteria can make you sick. Current monitoring at the public beaches on Tybee Island consists of weekly monitoring and enumeration of fecal indicator bacteria that takes 24 hours for results. If the number of bacteria exceed regulatory limits, a public health advisory is issued, and affected waters are retested until …
Addressing Sparsity For Knowledge Graph Completion: Data And Model Perspectives, Ran Liu
Addressing Sparsity For Knowledge Graph Completion: Data And Model Perspectives, Ran Liu
Dissertations and Theses Collection (Open Access)
Knowledge graphs (KGs) are powerful tools for structuring factual knowledge into relational triples, yet their practical utility is often adversely affected by data sparsity. Many entities and relations are associated with only a few observations, which limits the quality of learned embeddings and weakens generalization in downstream tasks. The problem of sparsity led to two interrelated challenges. Firstly, it restricts the informativeness of training samples: positive examples are scarce, and conventional negative sampling often produces trivial or redundant negatives that resulting in limited guidance. Secondly, in few-shot relation learning scenarios, sparsity worsens distribution shifts between training and test relations, as …
Deep Graph Anomaly Detection: A Survey And New Perspectives, Hezhe Qiao, Hanghang Tong, Nanyang Technological University, Irwin King, Charu Aggarwal, Guansong Pang
Deep Graph Anomaly Detection: A Survey And New Perspectives, Hezhe Qiao, Hanghang Tong, Nanyang Technological University, Irwin King, Charu Aggarwal, Guansong Pang
Research Collection School Of Computing and Information Systems
Graph anomaly detection (GAD), which aims to identify unusual graph instances (e.g., nodes, edges, subgraphs, or graphs), has attracted increasing attention in recent years due to its significance in a wide range of applications. Deep learning approaches, graph neural networks (GNNs) in particular, have been emerging as a promising paradigm for GAD, owing to its strong capability in capturing complex structure and/or node attributes in graph data. Considering the large number of methods proposed for GNN-based GAD, it is of paramount importance to summarize the methodologies and findings in the existing GAD studies, so that we can pinpoint effective model …
Affinitytune: A Prompt-Tuning Framework For Few-Shot Anomaly Detection On Graphs, Jingyan Chen, Guanghui Zhu, Guansong Pang, Chunfeng Yuan, Yihua Huang
Affinitytune: A Prompt-Tuning Framework For Few-Shot Anomaly Detection On Graphs, Jingyan Chen, Guanghui Zhu, Guansong Pang, Chunfeng Yuan, Yihua Huang
Research Collection School Of Computing and Information Systems
Graph anomaly detection (GAD) is a critical task with applications in domains such as networking, finance, and bioinformatics. % However, the scarcity of labeled anomalies and the limitations of unsupervised methods hinder effective detection. % While semi-supervised and few-shot learning approaches offer improvements, they struggle with knowledge transfer and rely heavily on labeled data. % Recent advancements in prompt tuning on graphs provide a promising direction, but their application to heterophilous graphs in anomaly detection remains underexplored. % In this work, we propose AffinityTune, a novel framework for few-shot graph anomaly detection based on prompt tuning. % Our approach introduces …
Hd-Epic: A Highly-Detailed Egocentric Video Dataset, Toby Perrett, Ahmad Darkhalil, Saptarshi Sinha, Omar Emara, Sam Pollard, Kranti Kumar Parida, Kaiting Liu, Prajwal Gatti, Siddhant Bansal, Kevin Flanagan, Jacob Chalk, Zhifan Zhu, Rhodri Guerrier, Fahd Abdelazim, Bin Zhu, Davide Moltisanti, Michael Wray, Hazel Doughty, Dima Damen
Hd-Epic: A Highly-Detailed Egocentric Video Dataset, Toby Perrett, Ahmad Darkhalil, Saptarshi Sinha, Omar Emara, Sam Pollard, Kranti Kumar Parida, Kaiting Liu, Prajwal Gatti, Siddhant Bansal, Kevin Flanagan, Jacob Chalk, Zhifan Zhu, Rhodri Guerrier, Fahd Abdelazim, Bin Zhu, Davide Moltisanti, Michael Wray, Hazel Doughty, Dima Damen
Research Collection School Of Computing and Information Systems
We present a validation dataset of newly-collected kitchenbased egocentric videos, manually annotated with highly detailed and interconnected ground-truth labels covering: recipe steps, fine-grained actions, ingredients with nutritional values, moving objects, and audio annotations. Importantly, all annotations are grounded in 3D through digital twinning of the scene, fixtures, object locations, and primed with gaze. Footage is collected from unscripted recordings in diverse home environments, making HDEPIC the first dataset collected in-the-wild but with detailed annotations matching those in controlled lab environments. We show the potential of our highly-detailed annotations through a challenging VQA benchmark of 26K questions assessing the capability to …
Guest Editorial: When Multimedia Meets Food: Multimedia Computing For Food Data Analysis And Applications, Weiqing Min, Shuqiang Jiang, Petia Radeva, Vladimir Pavlovic, Chong-Wah Ngo, Kiyoharu Aizawa, Wanqing Li
Guest Editorial: When Multimedia Meets Food: Multimedia Computing For Food Data Analysis And Applications, Weiqing Min, Shuqiang Jiang, Petia Radeva, Vladimir Pavlovic, Chong-Wah Ngo, Kiyoharu Aizawa, Wanqing Li
Research Collection School Of Computing and Information Systems
Food is central in our life for its fundamental role in our survival, health, mood and culture. The deployment of various networks (e.g., IoT and mobile networks), devices (e.g., hyperspectral imaging devices, electronic nose/tongue), databases (e.g., nutrition tables and food compositional databases), recipe-sharing websites (e.g., Yummly and Meishijie) and social media (e.g., Twitter and Weibo) has generated unprecedented volumes of multi-modal food data. Such multi-source multi-modal food data provides new perspectives to analyze and understand food consumption via multimedia computing. Riding on the wave of AI, food-oriented multimedia computing integrates AI, multimedia technology and food science to enable a wide …
Full-Stack Web Applications: Industry Standard Frameworks, Libraries & Technologies, Yassine Chahid, Patrick Slattery
Full-Stack Web Applications: Industry Standard Frameworks, Libraries & Technologies, Yassine Chahid, Patrick Slattery
Publications and Research
This research explores emerging full-stack web development technologies across front-end, back-end, and DevSecOps domains. It evaluates modern tools including Django, React, and TypeScript—focusing on their key features such as compile-time error checking—through to the development of a web application. By examining documentation for the frameworks Node.js, Next.js, Tailwind CSS, and others, along with the deployment tools Docker and Git for version/release control, the study analyzes how these innovations speed up development, improve existing practices, and have often replaced older technologies. Cloud solutions for tasks such as authentication and deployment will also be evaluated, along with various web-application technology stacks and …
Application Of Graph Neural Networks With Phase Space Graphs, Parker H. Cole, Ryan Benton, Ralf Riedel, David Bourrie
Application Of Graph Neural Networks With Phase Space Graphs, Parker H. Cole, Ryan Benton, Ralf Riedel, David Bourrie
Shelby Hall Graduate Research Forum Posters
Non-linear phase-space analysis models data represented as a graph transitioning between states in the time domain. By studying data transitions, we can predict the time a particular behavior occurs and classify the events (states) in a system. For example, we could classify neurological sensor data to determine if a person is asleep (state), or predict the direction in which a stock will move (transitions) based on micro trade patterns.
Previous research has demonstrated success in phase-space graphs in classifying malware, detecting network intrusions, and predicting seizures. However, the solutions either require calculating global graph features as inputs to a classifier, …
Flowing Together Or Alone: Impact Of Collaboration In The Metaverse, Fiona Fui-Hoon Nah, Brenda Eschenbrenner, Langtao Chen
Flowing Together Or Alone: Impact Of Collaboration In The Metaverse, Fiona Fui-Hoon Nah, Brenda Eschenbrenner, Langtao Chen
Research Collection School Of Computing and Information Systems
The metaverse is the next-generation Internet (Web3) that facilitates social connections and collaborations in a virtual world environment. Given the potential of the metaverse to provide more satisfying and effective means of remote collaborations, exploring the possibility of leveraging the metaverse for these endeavors is warranted. Therefore, an important question to address is whether greater engagement occurs when tasks are completed collaboratively versus individually in the metaverse. We address this question by drawing on flow and transportation theories to hypothesize the effect of carrying out a creative task in the metaverse collaboratively versus alone on one's cognitive absorption, a contextually …
La Creatividad En Peligro: Como La Inteligencia Artificial Es Un Reto Para Los Artistas., Nathaly Cisneros
La Creatividad En Peligro: Como La Inteligencia Artificial Es Un Reto Para Los Artistas., Nathaly Cisneros
Capstones
Los artistas digitales han creado obras maestras que nos han dejado sin aliento con sus pinceles digitales, lápices y pinturas. Desde retratos que parecen saltar de la pantalla hasta paisajes que nos transportan a mundos desconocidos, su arte ha sido una fuente constante de inspiración.
Pero en los últimos años, una nueva fuerza ha comenzado a cambiar el juego. La inteligencia artificial ha estado avanzando a pasos agigantados y ahora se perfila como una amenaza para el futuro de los artistas digitales. ¿Qué significa esto para el arte y la creatividad?
Link: https://docs.google.com/document/d/1xe8UxDMekX_SwiIppyt_JppK8M-lB-YWNWGyeyShlJM/edit?usp=sharing
Visualization Of Paleocurrents On A Web Application Using Gplates, Anjan Sapkota
Visualization Of Paleocurrents On A Web Application Using Gplates, Anjan Sapkota
MS in Computer Science Theses
Paleocurrents are flow directions derived from features of sedimentary rocks that reveal the direction of the current of wind or water that deposited the sediment. In 2015, Brand et al. created a global database of paleocurrents, which contains over 1,000,000 measurements worldwide: North America, South America, Australia, Great Britain, parts of Western Europe, China, Africa are fairly well represented; Antarctica, Eastern Europe, and Asia are modestly represented and Russia is poorly represented. The contribution of this thesis is a web application that uses the GPlates’ Application Programming Interface (API) to visualize global paleocurrents through time in an interactive way based …
Using Llms To Establish Implicit User Sentiment Of Software Desirability, Sherri Weitl-Harms, John D. Hastings, Jonah Lum
Using Llms To Establish Implicit User Sentiment Of Software Desirability, Sherri Weitl-Harms, John D. Hastings, Jonah Lum
Research & Publications
This study explores the use of LLMs for providing quantitative zero-shot sentiment analysis of implicit software desirability, addressing a critical challenge in product evaluation where traditional review scores, though convenient, fail to capture the richness of qualitative user feedback. Innovations include establishing a method that 1) works with qualitative user experience data without the need for explicit review scores, 2) focuses on implicit user satisfaction, and 3) provides scaled numerical sentiment analysis, offering a more nuanced understanding of user sentiment, instead of simply classifying sentiment as positive, neutral, or negative.
Data is collected using the Microsoft Product Desirability Toolkit (PDT), …
Replay-And-Forget-Free Graph Class-Incremental Learning: A Task Profiling And Prompting Approach, Chaoxi Niu, Guansong Pang, Ling Chen, Bing Liu
Replay-And-Forget-Free Graph Class-Incremental Learning: A Task Profiling And Prompting Approach, Chaoxi Niu, Guansong Pang, Ling Chen, Bing Liu
Research Collection School Of Computing and Information Systems
Class-incremental learning (CIL) aims to continually learn a sequence of tasks, with each task consisting of a set of unique classes. Graph CIL (GCIL) follows the same setting but needs to deal with graph tasks (e.g., node classification in a graph). The key characteristic of CIL lies in the absence of task identifiers (IDs) during inference, which causes a significant challenge in separating classes from different tasks (i.e., inter-task class separation). Being able to accurately predict the task IDs can help address this issue, but it is a challenging problem. In this paper, we show theoretically that accurate task ID …
Eyetraes : Fine-Grained, Low-Latency Eye Tracking Via Adaptive Event Slicing, Argha Sen, Panahetipola Mudiyanselage Nuwan Bandara, Ila Gokarn, Thivya Kandappu, Archan Misra
Eyetraes : Fine-Grained, Low-Latency Eye Tracking Via Adaptive Event Slicing, Argha Sen, Panahetipola Mudiyanselage Nuwan Bandara, Ila Gokarn, Thivya Kandappu, Archan Misra
Research Collection School Of Computing and Information Systems
Eye-tracking technology has gained significant attention in recent years due to its wide range of applications in humancomputer interaction, virtual and augmented reality, and wearable health. Traditional RGB camera-based eye-tracking systems often struggle with poor temporal resolution and computational constraints, limiting their effectiveness in capturing rapid eye movements. To address these limitations, we propose EyeTrAES, a novel approach using neuromorphic event cameras for high-fidelity tracking of natural pupillary movement that shows significant kinematic variance. One of EyeTrAES’s highlights is the use of a novel adaptive windowing/slicing algorithm that ensures just the right amount of descriptive asynchronous event data accumulation within …
Eyegraph : Modularity-Aware Spatio Temporal Graph Clustering For Continuous Event-Based Eye Tracking, Panahetipola Mudiyanselage Nuwan Bandara, Thivya Kandappu, Archan Misra, Ila Gokarn, Archan Misra
Eyegraph : Modularity-Aware Spatio Temporal Graph Clustering For Continuous Event-Based Eye Tracking, Panahetipola Mudiyanselage Nuwan Bandara, Thivya Kandappu, Archan Misra, Ila Gokarn, Archan Misra
Research Collection School Of Computing and Information Systems
Continuous tracking of eye movement dynamics plays a significant role in developing a broad spectrum of human-centered applications, such as cognitive skills (visual attention and working memory) modeling, human-machine interaction, biometric user authentication, and foveated rendering. Recently neuromorphic cameras have garnered significant interest in the eye-tracking research community, owing to their sub-microsecond latency in capturing intensity changes resulting from eye movements. Nevertheless, the existing approaches for event-based eye tracking suffer from several limitations: dependence on RGB frames, label sparsity, and training on datasets collected in controlled lab environments that do not adequately reflect real-world scenarios. To address these limitations, in …
Transitioning Our Website To Libguides Cms, Samantha Duncan, Eric Resnis
Transitioning Our Website To Libguides Cms, Samantha Duncan, Eric Resnis
Library Faculty Presentations
In this presentation, we describe how we used data from rapid and in-depth student usability testing to assist with the redesign of the library’s website as we finally transitioned to LibGuides CMS. Using the Springy tools; LibGuides, LibGuides CMS, LibCal, and LibWizard we outlined how we were able to create and carry out this highly effective testing, resulting in a better understanding of how our students navigate our site and how to improve it. During this journey, attendees were provided with the detailed and some might say lengthy process that was undertaken to achieve our goals. We did this by …
Enhancing Recipe Retrieval With Foundation Models: A Data Augmentation Perspective, Fangzhou Song, Bin Zhu, Yanbin Hao, Shuo Wang
Enhancing Recipe Retrieval With Foundation Models: A Data Augmentation Perspective, Fangzhou Song, Bin Zhu, Yanbin Hao, Shuo Wang
Research Collection School Of Computing and Information Systems
Learning recipe and food image representation in common embedding space is non-trivial but crucial for cross-modal recipe retrieval. In this paper, we propose a new perspective for this problem by utilizing foundation models for data augmentation. Leveraging on the remarkable capabilities of foundation models (i.e., Llama2 and SAM), we propose to augment recipe and food image by extracting alignable information related to the counterpart. Specifically, Llama2 is employed to generate a textual description from the recipe, aiming to capture the visual cues of a food image, and SAM is used to produce image segments that correspond to key ingredients in …
Onerestore : A Universal Restoration Framework For Composite Degradation, Yu Guo, Yuan Gao, Yuxu Lu, Huilin Zhu, Ryan Wen Liu, Shengfeng He
Onerestore : A Universal Restoration Framework For Composite Degradation, Yu Guo, Yuan Gao, Yuxu Lu, Huilin Zhu, Ryan Wen Liu, Shengfeng He
Research Collection School Of Computing and Information Systems
In real-world scenarios, image impairments often manifest as composite degradations, presenting a complex interplay of elements such as low light, haze, rain, and snow. Despite this reality, existing restoration methods typically target isolated degradation types, thereby falling short in environments where multiple degrading factors coexist. To bridge this gap, our study proposes a versatile imaging model that consolidates four physical corruption paradigms to accurately represent complex, composite degradation scenarios. In this context, we propose OneRestore, a novel transformer-based framework designed for adaptive, controllable scene restoration. The proposed framework leverages a unique cross-attention mechanism, merging degraded scene descriptors with image features, …
Video Editing For Video Retrieval, Bin Zhu, Kevin Flanagan, Adriano Fragomeni, Michael Wray, Dima Damen
Video Editing For Video Retrieval, Bin Zhu, Kevin Flanagan, Adriano Fragomeni, Michael Wray, Dima Damen
Research Collection School Of Computing and Information Systems
Though pre-training vision-language models have demonstrated significant benefits in boosting video-text retrieval performance from large-scale web videos, fine-tuning still plays a critical role with manually annotated clips with start and end times, which requires considerable human effort. To address this issue, we explore an alternative cheaper source of annotations, single timestamps, for video-text retrieval. We initialise clips from timestamps in a heuristic way to warm up a retrieval model. Then a video clip editing method is proposed to refine the initial rough boundaries to improve retrieval performance. A student-teacher network is introduced for video clip editing: the teacher model is …
Beat-It : Beat-Synchronized Multi-Condition 3d Dance Generation, Zikai Huang, Xuemiao Xu, Cheng Xu, Huaidong Zhang, Chenxi Zheng, Jing Qin, Shengfeng He
Beat-It : Beat-Synchronized Multi-Condition 3d Dance Generation, Zikai Huang, Xuemiao Xu, Cheng Xu, Huaidong Zhang, Chenxi Zheng, Jing Qin, Shengfeng He
Research Collection School Of Computing and Information Systems
Dance, as an art form, fundamentally hinges on the precise synchronization with musical beats. However, achieving aesthetically pleasing dance sequences from music is challenging, with existing methods often falling short in controllability and beat alignment. To address these shortcomings, this paper introduces Beat-It, a novel framework for beat-specific, key pose-guided dance generation. Unlike prior approaches, Beat-It uniquely integrates explicit beat awareness and key pose guidance, effectively resolving two main issues: the misalignment of generated dance motions with musical beats, and the inability to map key poses to specific beats, critical for practical choreography. Our approach disentangles beat conditions from music …
Robust Image Classification System Via Cloud Computing, Aligned Multimodal Embeddings, Centroids And Neighbours, Wei Lun Koh, Boon Yong Koh, Bing Tian Dai
Robust Image Classification System Via Cloud Computing, Aligned Multimodal Embeddings, Centroids And Neighbours, Wei Lun Koh, Boon Yong Koh, Bing Tian Dai
Research Collection School Of Computing and Information Systems
We propose a framework for a cloud-based application of an image classification system that is highly accessible, maintains data confidentiality, and robust to incorrect training labels. The end-to-end system is implemented using Amazon Web Services (AWS), with a detailed guide provided for replication, enhancing the ways which researchers can collaborate with a community of users for mutual benefits. A front-end web application allows users across the world to securely log in, contribute labelled training images conveniently via a drag-and-drop approach, and use that same application to query an up-to-date model that has knowledge of images from the community of users. …