Open Access. Powered by Scholars. Published by Universities.®
Graphics and Human Computer Interfaces Commons™
Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Databases and Information Systems (38)
- Artificial Intelligence and Robotics (29)
- Social and Behavioral Sciences (26)
- Engineering (22)
- Software Engineering (21)
-
- Other Computer Sciences (19)
- Arts and Humanities (16)
- Data Science (14)
- Computer Engineering (12)
- Theory and Algorithms (11)
- Art and Design (10)
- Psychology (10)
- Business (9)
- Education (9)
- Electrical and Computer Engineering (8)
- Human Factors Psychology (8)
- Information Security (8)
- Numerical Analysis and Scientific Computing (8)
- Communication (7)
- Interactive Arts (7)
- Medicine and Health Sciences (7)
- OS and Networks (7)
- Sociology (6)
- Science and Technology Studies (5)
- Architecture (4)
- Biomedical Engineering and Bioengineering (4)
- Communication Technology and New Media (4)
- Institution
-
- Singapore Management University (82)
- Old Dominion University (8)
- San Jose State University (8)
- Clemson University (7)
- Dartmouth College (6)
-
- Chapman University (5)
- Southern Adventist University (4)
- University of Arkansas, Fayetteville (4)
- University of Dayton (4)
- University of Denver (4)
- City University of New York (CUNY) (3)
- Louisiana State University (3)
- Mississippi State University (3)
- Universitas Negeri Yogyakarta (3)
- Air Force Institute of Technology (2)
- Arkansas Tech University (2)
- California Polytechnic State University, San Luis Obispo (2)
- Michigan Technological University (2)
- Southern Methodist University (2)
- Technological University Dublin (2)
- The University of Akron (2)
- University of Central Florida (2)
- University of Minnesota Morris Digital Well (2)
- University of Nebraska - Lincoln (2)
- American University in Cairo (1)
- Beirut Arab University (1)
- Bentley University (1)
- Bridgewater College (1)
- Brigham Young University (1)
- Bucknell University (1)
- Keyword
-
- Computer Science (8)
- Machine learning (6)
- Virtual reality (6)
- Visualization (6)
- Augmented reality (5)
-
- Computer vision (5)
- Augmented Reality (4)
- Blind (4)
- Daniel Felix Ritchie School of Engineering and Computer Science (4)
- Deep learning (4)
- Eye tracking (4)
- Human computer interaction (4)
- Human-computer interaction (4)
- Older adults (4)
- Reinforcement learning (4)
- Accessibility (3)
- Artificial intelligence (3)
- Deep Learning (3)
- Generative adversarial networks (3)
- Graph neural networks (3)
- Human-centered computing (3)
- Mixed Reality (3)
- Privacy (3)
- Usability (3)
- VR (3)
- Virtual Reality (3)
- 3D Localization (2)
- Acoustic sensing (2)
- Algorithms (2)
- Artificial Intelligence (2)
- Publication
-
- Research Collection School Of Computing and Information Systems (80)
- Computer Science Faculty Publications (7)
- Master's Projects (7)
- Theses and Dissertations (7)
- Dartmouth College Master’s Theses (6)
-
- All Dissertations (5)
- Electronic Theses and Dissertations (5)
- Campus Research Month (3)
- Computer Science and Computer Engineering Undergraduate Honors Theses (3)
- Elinvo (Electronics, Informatics, and Vocational Education) (3)
- Publications and Research (3)
- ATU Scholars Symposium (2)
- All Theses (2)
- Dissertations and Theses Collection (Open Access) (2)
- Dissertations, Master's Theses and Master's Reports (2)
- Electronic Theses and Dissertations, 2020-2023 (2)
- Engineering Faculty Articles and Research (2)
- Honors Theses (2)
- LSU Doctoral Dissertations (2)
- Scholarly Horizons: University of Minnesota, Morris Undergraduate Journal (2)
- Student Scholar Symposium Abstracts and Posters (2)
- Williams Honors College, Honors Research Projects (2)
- 2023 (1)
- Academic Posters Collection (1)
- BAU Journal - Science and Technology (1)
- CISLA Senior Integrative Projects (1)
- College of Engineering Summer Undergraduate Research Program (1)
- College of Sciences Posters (1)
- Communication Sciences and Disorders Faculty Articles and Research (1)
- Computer Science ETDs (1)
- Publication Type
- File Type
Articles 61 - 90 of 191
Full-Text Articles in Graphics and Human Computer Interfaces
Studying Memes During Covid Lockdown As A Lens Through Which To Understand Video-Mediated Communication Interactions, Tatyana Claytor
Studying Memes During Covid Lockdown As A Lens Through Which To Understand Video-Mediated Communication Interactions, Tatyana Claytor
Electronic Theses and Dissertations, 2020-2023
The purpose of this study is to analyze image macros about video-mediated communication (VMC) created during the time frame of 2020-2021 when people all over the world started using Zoom and VMC for work and school. It is a unique opportunity to study how users' interactions with themselves and with others were affected at a time when a lot of people started using the technology at the same time. Because the focus is on interactions, I narrowed it down to three topics to analyze the memes: presence, self, and space and place to analyze the memes. I chose memes relating …
Signings Of Graphs And Sign-Symmetric Signed Graphs, Ahmad Asiri
Signings Of Graphs And Sign-Symmetric Signed Graphs, Ahmad Asiri
Theses and Dissertations
In this dissertation, we investigate various aspects of signed graphs, with a particular focus on signings and sign-symmetric signed graphs. We begin by examining the complete graph on six vertices with one edge deleted ($K_6$\textbackslash e) and explore the different ways of signing this graph up to switching isomorphism. We determine the frustration index (number) of these signings and investigate the existence of sign-symmetric signed graphs. We then extend our study to the $K_6$\textbackslash 2e graph and the McGee graph with exactly two negative edges. We investigate the distinct ways of signing these graphs up to switching isomorphism and demonstrate …
Visual And Spatial Audio Mismatching In Virtual Environments, Zachary Lawrence Garris
Visual And Spatial Audio Mismatching In Virtual Environments, Zachary Lawrence Garris
Theses and Dissertations
This paper explores how vision affects spatial audio perception in virtual reality. We created four virtual environments with different reverb and room sizes, and recorded binaural clicks in each one. We conducted two experiments: one where participants judged the audio-visual match, and another where they pointed to the click direction. We found that vision influences spatial audio perception and that congruent audio-visual cues improve accuracy. We suggest some implications for virtual reality design and evaluation.
Understanding The Role Of Interactivity And Explanation In Adaptive Experiences, Lijie Guo
Understanding The Role Of Interactivity And Explanation In Adaptive Experiences, Lijie Guo
All Dissertations
Adaptive experiences have been an active area of research in the past few decades, accompanied by advances in technology such as machine learning and artificial intelligence. Whether the currently ongoing research on adaptive experiences has focused on personalization algorithms, explainability, user engagement, or privacy and security, there is growing interest and resources in developing and improving these research focuses. Even though the research on adaptive experiences has been dynamic and rapidly evolving, achieving a high level of user engagement in adaptive experiences remains a challenge. %????? This dissertation aims to uncover ways to engage users in adaptive experiences by incorporating …
All Hands On Deck: Choosing Virtual End Effector Representations To Improve Near Field Object Manipulation Interactions In Extended Reality, Roshan Venkatakrishnan
All Hands On Deck: Choosing Virtual End Effector Representations To Improve Near Field Object Manipulation Interactions In Extended Reality, Roshan Venkatakrishnan
All Dissertations
Extended reality, or "XR", is the adopted umbrella term that is heavily gaining traction to collectively describe Virtual reality (VR), Augmented reality (AR), and Mixed reality (MR) technologies. Together, these technologies extend the reality that we experience either by creating a fully immersive experience like in VR or by blending in the virtual and "real" worlds like in AR and MR.
The sustained success of XR in the workplace largely hinges on its ability to facilitate efficient user interactions. Similar to interacting with objects in the real world, users in XR typically interact with virtual integrants like objects, menus, windows, …
Anatoview: Using Interactive 3d Visualizations With Augmented Reality Support For Laypersons’ Medical Education In Informed Consent Processes, Michelle Chen
Dartmouth College Master’s Theses
AnatoView is an interactive multimedia educational application that visualizes medical procedures in three-dimensional (3D), augmented reality (AR) space. By providing visual and spatial information of medical procedures, AnatoView acts as a learning supplement for laypersons/patients in informed consent (IC) processes — wherein instructional content is traditionally limited to purely spoken explanations that lead to poor patient comprehension. We design a mixed study and conduct a randomized, controlled trial with 15 laypersons as participants: administering a traditional IC process to a control group, and an IC process supplemented by the use of AnatoView to experimental groups. As a primary outcome, medical …
The Effects Of Primary And Secondary Task Workloads On Cybersickness In Immersive Virtual Active Exploration Experiences, Rohith Venkatakrishnan
The Effects Of Primary And Secondary Task Workloads On Cybersickness In Immersive Virtual Active Exploration Experiences, Rohith Venkatakrishnan
All Dissertations
Virtual reality (VR) technology promises to transform humanity. The technology enables users to explore and interact with computer-generated environments that can be simulated to approximate or deviate from reality. This creates an endless number of ways to propitiously apply the technology in our lives. It follows that large technological conglomerates are pushing for the widespread adoption of VR, financing the creation of the Metaverse - a hypothetical representation of the next iteration of the internet.
Even with VR technology's continuous growth, its widespread adoption remains long overdue. This can largely be attributed to an affliction called cybersickness, an analog to …
Hyperbolic Graph Topic Modeling Network With Continuously Updated Topic Tree, Ce Zhang, Rex Ying, Hady Wirawan Lauw
Hyperbolic Graph Topic Modeling Network With Continuously Updated Topic Tree, Ce Zhang, Rex Ying, Hady Wirawan Lauw
Research Collection School Of Computing and Information Systems
Connectivity across documents often exhibits a hierarchical network structure. Hyperbolic Graph Neural Networks (HGNNs) have shown promise in preserving network hierarchy. However, they do not model the notion of topics, thus document representations lack semantic interpretability. On the other hand, a corpus of documents usually has high variability in degrees of topic specificity. For example, some documents contain general content (e.g., sports), while others focus on specific themes (e.g., basketball and swimming). Topic models indeed model latent topics for semantic interpretability, but most assume a flat topic structure and ignore such semantic hierarchy. Given these two challenges, we propose a …
Flood: A Flexible Invariant Learning Framework For Out-Of-Distribution Generalization On Graphs, Yang Liu, Xiang Ao, Fuli Feng, Yunshan Ma, Kuan Li, Tat‑Seng Chua, Qing He
Flood: A Flexible Invariant Learning Framework For Out-Of-Distribution Generalization On Graphs, Yang Liu, Xiang Ao, Fuli Feng, Yunshan Ma, Kuan Li, Tat‑Seng Chua, Qing He
Research Collection School Of Computing and Information Systems
Graph Neural Networks (GNNs) have achieved remarkable success in various domains but most of them are developed under the in-distribution assumption. Under out-of-distribution (OOD) settings, they suffer from the distribution shift between the training set and the test set and may not generalize well to the test distribution. Several methods have tried the invariance principle to improve the generalization of GNNs in OOD settings. However, in previous solutions, the graph encoder is immutable after the invariant learning and cannot be adapted to the target distribution flexibly. Confronting the distribution shift, a flexible encoder with refinement to the target distribution can …
Socialz: Multi-Feature Social Fuzz Testing, Francisco Zanartu, Christoph Treude, Markus Wagner
Socialz: Multi-Feature Social Fuzz Testing, Francisco Zanartu, Christoph Treude, Markus Wagner
Research Collection School Of Computing and Information Systems
Online social networks have become an integral aspect of our daily lives and play a crucial role in shaping our relationships with others. However, bugs and glitches, even minor ones, can cause anything from frustrating problems to serious data leaks that can have farreaching impacts on millions of users. To mitigate these risks, fuzz testing, a method of testing with randomised inputs, can provide increased confidence in the correct functioning of a social network. However, implementing traditional fuzz testing methods can be prohibitively difficult or impractical for programmers outside of the network’s development team. To tackle this challenge, we present …
Contrastive Video Question Answering Via Video Graph Transformer, Junbin Xiao Xiao, Pan Zhou, Angela Yao, Yicong Li, Richang Hong, Shuicheng Yan, Tat-Seng Chua
Contrastive Video Question Answering Via Video Graph Transformer, Junbin Xiao Xiao, Pan Zhou, Angela Yao, Yicong Li, Richang Hong, Shuicheng Yan, Tat-Seng Chua
Research Collection School Of Computing and Information Systems
We propose to perform video question answering (VideoQA) in a Contrastive manner via a Video Graph Transformer model (CoVGT). CoVGT’s uniqueness and superiority are three-fold: 1) It proposes a dynamic graph transformer module which encodes video by explicitly capturing the visual objects, their relations and dynamics, for complex spatio-temporal reasoning. 2) It designs separate video and text transformers for contrastive learning between the video and text to perform QA, instead of multi-modal transformer for answer classification. Fine-grained video-text communication is done by additional cross-modal interaction modules. 3) It is optimized by the joint fully- and self-supervised contrastive objectives between the …
Balanced Blended Space: Foundational Human–Ai Dialogues In A Symmetry-Based Mediation Framework, David Smith
Balanced Blended Space: Foundational Human–Ai Dialogues In A Symmetry-Based Mediation Framework, David Smith
Publications and Research
This working paper documents the early development of the Balanced Blended Space (BBS) framework through a series of iterative interactions between a cognitive agent (human researcher) and a computational agent (AI system) conducted in 2023. The work is motivated by the need for a universal theoretical model capable of describing the integration of physical, virtual, and conceptual spaces, particularly in response to increasing fragmentation across contemporary communication systems.
BBS is proposed as a symmetry-based mediation framework in which relationships between domains—such as physical and virtual space, cognition and computation, and multiple sensory modalities—are treated as structurally equivalent and mappable. Central …
Fine-Grained Domain Adaptive Crowd Counting Via Point-Derived Segmentation, Yongtuo Liu, Dan Xu, Sucheng Ren, Hanjie Wu, Hongmin Cai, Shengfeng He
Fine-Grained Domain Adaptive Crowd Counting Via Point-Derived Segmentation, Yongtuo Liu, Dan Xu, Sucheng Ren, Hanjie Wu, Hongmin Cai, Shengfeng He
Research Collection School Of Computing and Information Systems
Due to domain shift, a large performance drop is usually observed when a trained crowd counting model is deployed in the wild. While existing domain-adaptive crowd counting methods achieve promising results, they typically regard each crowd image as a whole and reduce domain discrepancies in a holistic manner, thus limiting further improvement of domain adaptation performance. To this end, we propose to untangle domain-invariant crowd and domain-specific background from crowd images and design a fine-grained domain adaption method for crowd counting. Specifically, to disentangle crowd from background, we propose to learn crowd segmentation from point-level crowd counting annotations in a …
A Review On The Effects Of Chanting And Solfeggio Frequencies On Well-Being, Xuyu Yang, Fiona Fui-Hoon Nah, Fen Lin
A Review On The Effects Of Chanting And Solfeggio Frequencies On Well-Being, Xuyu Yang, Fiona Fui-Hoon Nah, Fen Lin
Research Collection School Of Computing and Information Systems
This paper presents our literature review on how chanting and solfeggio frequencies affect brain activity, enhance well-being, and inform the design of sound therapy. We call for more scientific research to investigate and better understand the effects of chanting and solfeggio frequencies on well-being.
Robust And Parallel Segmentation Model (Rpsm) For Early Detection Of Skin Cancer Disease Using Heterogeneous Distributions, Nancy Zreika, Ali El-Zaart, Abdallah El Chakik
Robust And Parallel Segmentation Model (Rpsm) For Early Detection Of Skin Cancer Disease Using Heterogeneous Distributions, Nancy Zreika, Ali El-Zaart, Abdallah El Chakik
BAU Journal - Science and Technology
Melanoma is the most common dangerous type of skin cancer; however, it is preventable if it is diagnosed early. Diagnosis of Melanoma would be improved if an accurate skin image segmentation model is available. Many computer vision methods have been investigated, yet the problem of finding a consistent and robust model that extracts the best threshold value, persists. This paper suggests a novel image segmentation approach using a multilevel cross entropy thresholding algorithm based on heterogeneous distributions. The proposed strategy searches the problem space by segmenting the image into several levels, and applying for each level one of the three …
Deep-Learning Realtime Upsampling Techniques In Video Games, Biruk Mengistu
Deep-Learning Realtime Upsampling Techniques In Video Games, Biruk Mengistu
Scholarly Horizons: University of Minnesota, Morris Undergraduate Journal
This paper addresses the challenge of keeping up with the ever-increasing graphical complexity of video games and introduces a deep-learning approach to mitigating it. As games get more and more demanding in terms of their graphics, it becomes increasingly difficult to maintain high-quality images while also ensuring good performance. This is where deep learning super sampling (DLSS) comes in. The paper explains how DLSS works, including the use of convolutional autoencoder neural networks and various other techniques and technologies. It also covers how the network is trained and optimized, as well as how it incorporates temporal antialiasing and frame generation …
How Photorealistic Images Are Generated, Nahom Ketema
How Photorealistic Images Are Generated, Nahom Ketema
University Honors Theses
The field of computer graphics looks into how computers can be used to generate images. From using some trigonometry to plot 3D objects to using rays to calculate the lighting of an object, there are a variety of ways that we can use to draw objects onto a screen. For this thesis, we will be looking at a few of those methods to determine how photorealistic images are generated.
Interstice, Shravan Rao
Interstice, Shravan Rao
Masters Theses
When I was about three years old, I distinctly remember being too small to see what was on top of the table. A couple of years later, when I could see those objects, I thought the world around me had grown smaller. In a way, it did, as I experienced, lived, captured, remembered, and shared the space repeatedly. This sense of the world shrinking was exaggerated during the Covid-19 pandemic, allowing new behaviours and modes of interaction to emerge. Continually shaping our modern lives, virtual technologies redefine how we access and share information and stories or even explore new places. …
An Investigation Into Machine Learning Techniques For Designing Dynamic Difficulty Agents In Real-Time Games, Ryan Adare Dunagan
An Investigation Into Machine Learning Techniques For Designing Dynamic Difficulty Agents In Real-Time Games, Ryan Adare Dunagan
Electronic Theses and Dissertations
Video games are an incredibly popular pastime enjoyed by people of all ages world wide. Many different kinds of games exist, but most games feature some elements of the player overcoming some challenge, usually through gameplay. These challenges are insurmountable for some people and may turn them off to video games as a pastime. Games can be made more accessible to players of little skill and/or experience through the use of Dynamic Difficulty Adjustment (DDA) systems that adjust the difficulty of the game in response to the player’s performance. This research seeks to establish the effectiveness of machine learning techniques …
Generalizing Graph Neural Networks Across Graphs, Time, And Tasks, Zhihao Wen
Generalizing Graph Neural Networks Across Graphs, Time, And Tasks, Zhihao Wen
Dissertations and Theses Collection (Open Access)
Graph-structured data are ubiquitous across numerous real-world contexts, encompassing social networks, commercial graphs, bibliographic networks, and biological systems. Delving into the analysis of these graphs can yield significant understanding pertaining to their corresponding application fields.Graph representation learning offers a potent solution to graph analytics challenges by transforming a graph into a low-dimensional space while preserving its information to the greatest extent possible. This conversion into low-dimensional vectors enables the efficient computation of subsequent graph algorithms. The majority of prior research has concentrated on deriving node representations from a single, static graph. However, numerous real-world situations demand rapid generation of representations …
Mosaic: Spatially-Multiplexed Edge Ai Optimization Over Multiple Concurrent Video Sensing Streams, Ila Gokarn, Hemanth Sabbella, Yigong Hu, Tarek Abdelzaher, Archan Misra
Mosaic: Spatially-Multiplexed Edge Ai Optimization Over Multiple Concurrent Video Sensing Streams, Ila Gokarn, Hemanth Sabbella, Yigong Hu, Tarek Abdelzaher, Archan Misra
Research Collection School Of Computing and Information Systems
Sustaining high fidelity and high throughput of perception tasks over vision sensor streams on edge devices remains a formidable challenge, especially given the continuing increase in image sizes (e.g., generated by 4K cameras) and complexity of DNN models. One promising approach involves criticality-aware processing, where the computation is directed selectively to "critical" portions of individual image frames. We introduce MOSAIC, a novel system for such criticality-aware concurrent processing of multiple vision sensing streams that provides a multiplicative increase in the achievable throughput with negligible loss in perception fidelity. MOSAIC determines critical regions from images received from multiple vision …
Scanet: Self-Paced Semi-Curricular Attention Network For Non-Homogeneous Image Dehazing, Yu Guo, Yuan Gao, Ryan Wen Liu, Yuxu Lu, Jingxiang Qu, Shengfeng He, Ren Wenqi
Scanet: Self-Paced Semi-Curricular Attention Network For Non-Homogeneous Image Dehazing, Yu Guo, Yuan Gao, Ryan Wen Liu, Yuxu Lu, Jingxiang Qu, Shengfeng He, Ren Wenqi
Research Collection School Of Computing and Information Systems
The presence of non-homogeneous haze can cause scene blurring, color distortion, low contrast, and other degradations that obscure texture details. Existing homogeneous dehazing methods struggle to handle the non-uniform distribution of haze in a robust manner. The crucial challenge of non-homogeneous dehazing is to effectively extract the non-uniform distribution features and reconstruct the details of hazy areas with high quality. In this paper, we propose a novel self-paced semi-curricular attention network, called SCANet, for non-homogeneous image dehazing that focuses on enhancing haze-occluded regions. Our approach consists of an attention generator network and a scene re-construction network. We use the luminance …
Towards A Smaller Student: Capacity Dynamic Distillation For Efficient Image Retrieval, Yi Xie, Huaidong Zhang, Xuemiao Xu, Jianqing Zhu, Shengfeng He
Towards A Smaller Student: Capacity Dynamic Distillation For Efficient Image Retrieval, Yi Xie, Huaidong Zhang, Xuemiao Xu, Jianqing Zhu, Shengfeng He
Research Collection School Of Computing and Information Systems
Previous Knowledge Distillation based efficient image retrieval methods employ a lightweight network as the student model for fast inference. However, the lightweight student model lacks adequate representation capacity for effective knowledge imitation during the most critical early training period, causing final performance degeneration. To tackle this issue, we propose a Capacity Dynamic Distillation framework, which constructs a student model with editable representation capacity. Specifically, the employed student model is initially a heavy model to fruitfully learn distilled knowledge in the early training epochs, and the student model is gradually compressed during the training. To dynamically adjust the model capacity, our …
Position-Guided Text Prompt For Vision-Language Pre-Training, Alex Jinpeng Wang, Pan Zhou, Mike Zheng Shou, Yan Shuicheng
Position-Guided Text Prompt For Vision-Language Pre-Training, Alex Jinpeng Wang, Pan Zhou, Mike Zheng Shou, Yan Shuicheng
Research Collection School Of Computing and Information Systems
Vision-Language Pre-Training (VLP) has shown promising capabilities to align image and text pairs, facilitating a broad variety of cross-modal learning tasks. However, we observe that VLP models often lack the visual grounding/localization capability which is critical for many downstream tasks such as visual reasoning. In this work, we propose a novel Position-guided Text Prompt (PTP) paradigm to enhance the visual grounding ability of cross-modal models trained with VLP. Specifically, in the VLP phase, PTP divides the image into N x N blocks, and identifies the objects in each block through the widely used object detector in VLP. It then reformulates …
Class-Incremental Exemplar Compression For Class-Incremental Learning, Zilin Luo, Yaoyao Liu, Bernt Schiele, Qianru Sun
Class-Incremental Exemplar Compression For Class-Incremental Learning, Zilin Luo, Yaoyao Liu, Bernt Schiele, Qianru Sun
Research Collection School Of Computing and Information Systems
Exemplar-based class-incremental learning (CIL) finetunes the model with all samples of new classes but few-shot exemplars of old classes in each incremental phase, where the "few-shot" abides by the limited memory budget. In this paper, we break this "few-shot" limit based on a simple yet surprisingly effective idea: compressing exemplars by downsampling non-discriminative pixels and saving "many-shot" compressed exemplars in the memory. Without needing any manual annotation, we achieve this compression by generating 0-1 masks on discriminative pixels from class activation maps (CAM). We propose an adaptive mask generation model called class-incremental masking (CIM) to explicitly resolve two difficulties of …
Freestyle Layout-To-Image Synthesis, Han Xue, Zhiwu Huang, Qianru Sun, Li Song, Wenjun Zhang
Freestyle Layout-To-Image Synthesis, Han Xue, Zhiwu Huang, Qianru Sun, Li Song, Wenjun Zhang
Research Collection School Of Computing and Information Systems
Typical layout-to-image synthesis (LIS) models generate images for a close set of semantic classes, e.g., 182 common objects in COCO-Stuff. In this work, we explore the freestyle capability of the model, i.e., how far can it generate unseen semantics (e.g., classes, attributes, and styles) onto a given layout, and call the task Freestyle LIS (FLIS). Thanks to the development of large-scale pre-trained language-image models, a number of discriminative models (e.g., image classification and object detection) trained on limited base classes are empowered with the ability of unseen class prediction. Inspired by this, we opt to leverage large-scale pre-trained text-to-image diffusion …
Curricular Contrastive Regularization For Physics-Aware Single Image Dehazing, Yu Zheng, Jiahui Zhan, Shengfeng He, Yong Du
Curricular Contrastive Regularization For Physics-Aware Single Image Dehazing, Yu Zheng, Jiahui Zhan, Shengfeng He, Yong Du
Research Collection School Of Computing and Information Systems
Considering the ill-posed nature, contrastive regularization has been developed for single image dehazing, introducing the information from negative images as a lower bound. However, the contrastive samples are non-consensual, as the negatives are usually represented distantly from the clear (i.e., positive) image, leaving the solution space still under-constricted. Moreover, the interpretability of deep dehazing models is underexplored towards the physics of the hazing process. In this paper, we propose a novel curricular contrastive regularization targeted at a consensual contrastive space as opposed to a non-consensual one. Our negatives, which provide better lower-bound constraints, can be assembled from 1) the hazy …
Venus: A Geometrical Representation For Quantum State Visualization, Shaolun Ruan, Ribo Yuan, Qiang Guan, Yanna Lin, Ying Mao, Weiwen Jiang, Zhepeng Wang, Wei Xu, Yong Wang
Venus: A Geometrical Representation For Quantum State Visualization, Shaolun Ruan, Ribo Yuan, Qiang Guan, Yanna Lin, Ying Mao, Weiwen Jiang, Zhepeng Wang, Wei Xu, Yong Wang
Research Collection School Of Computing and Information Systems
Visualizations have played a crucial role in helping quantum computing users explore quantum states in various quantum computing applications. Among them, Bloch Sphere is the widely-used visualization for showing quantum states, which leverages angles to represent quantum amplitudes. However, it cannot support the visualization of quantum entanglement and superposition, the two essential properties of quantum computing. To address this issue, we propose VENUS, a novel visualization for quantum state representation. By explicitly correlating 2D geometric shapes based on the math foundation of quantum computing characteristics, VENUS effectively represents quantum amplitudes of both the single qubit and two qubits for quantum …
Ifundit: Visual Profiling Of Fund Investment Styles, Rong Zhang, Bon Kyung Ku, Yong Wang, Xuanwu Yue, Siyuan Liu, Ke Li, Huamin Qu
Ifundit: Visual Profiling Of Fund Investment Styles, Rong Zhang, Bon Kyung Ku, Yong Wang, Xuanwu Yue, Siyuan Liu, Ke Li, Huamin Qu
Research Collection School Of Computing and Information Systems
Mutual funds are becoming increasingly popular with the emergence of Internet finance. Clear profiling of a fund's investment style is crucial for fund managers to evaluate their investment strategies, and for investors to understand their investment. However, it is challenging to profile a fund's investment style as it requires a comprehensive analysis of complex multi-dimensional temporal data. In addition, different fund managers and investors have different focuses when analysing a fund's investment style. To address the issue, we propose iFUNDit, an interactive visual analytic system for fund investment style analysis. The system decomposes a fund's critical features into performance attributes …
Gnnlens: A Visual Analytics Approach For Prediction Error Diagnosis Of Graph Neural Networks., Zhihua Jin, Yong Wang, Qianwen Wang, Yao Ming, Tengfei Ma, Huamin Qu
Gnnlens: A Visual Analytics Approach For Prediction Error Diagnosis Of Graph Neural Networks., Zhihua Jin, Yong Wang, Qianwen Wang, Yao Ming, Tengfei Ma, Huamin Qu
Research Collection School Of Computing and Information Systems
Graph Neural Networks (GNNs) aim to extend deep learning techniques to graph data and have achieved significant progress in graph analysis tasks (e.g., node classification) in recent years. However, similar to other deep neural networks like Convolutional Neural Networks (CNNs) and Recurrent Neural Networks (RNNs), GNNs behave like a black box with their details hidden from model developers and users. It is therefore difficult to diagnose possible errors of GNNs. Despite many visual analytics studies being done on CNNs and RNNs, little research has addressed the challenges for GNNs. This paper fills the research gap with an interactive visual analysis …