Open Access. Powered by Scholars. Published by Universities.®

Discipline
Institution
Keyword
Publication Year
Publication
Publication Type
File Type

Articles 541 - 570 of 2362

Full-Text Articles in Graphics and Human Computer Interfaces

Experiences Of Autistic Twitch Livestreamers: “I Have Made Easily The Most Meaningful And Impactful Relationships”, Terrance Mok, Anthony Tang, Adam Mccrimmon, Lora Oehlberg Oct 2023

Experiences Of Autistic Twitch Livestreamers: “I Have Made Easily The Most Meaningful And Impactful Relationships”, Terrance Mok, Anthony Tang, Adam Mccrimmon, Lora Oehlberg

Research Collection School Of Computing and Information Systems

We present perspectives from 10 autistic Twitch streamers regarding their experiences as livestreamers and how autism uniquely colors their experiences. Livestreaming offers a social online experience distinct from in-person, face-to-face communication, where autistic people tend to encounter challenges. Our reflexive thematic analysis of interviews with 10 participants showcases autistic livestreamers’ perspectives in their own words. Our findings center on the importance of having streamers establishing connections with other, sharing autistic identities, controlling a space for social interaction, personal growth, and accessibility challenges. In our discussion, we highlight the crucial value of having a medium for autistic representation, as well as …


Unsupervised Anomaly Detection In Medical Images With A Memory-Augmented Multi-Level Cross-Attentional Masked Autoencoder, Yu Tian, Guansong Pang, Yuyuan Liu, Chong Wang, Yuanhong Chen, Fengbei Liu, Rajvinder Singh, Johan W. Verjans, Mengyu Wang, Gustavo Carneiro Oct 2023

Unsupervised Anomaly Detection In Medical Images With A Memory-Augmented Multi-Level Cross-Attentional Masked Autoencoder, Yu Tian, Guansong Pang, Yuyuan Liu, Chong Wang, Yuanhong Chen, Fengbei Liu, Rajvinder Singh, Johan W. Verjans, Mengyu Wang, Gustavo Carneiro

Research Collection School Of Computing and Information Systems

Unsupervised anomaly detection (UAD) aims to find anomalous images by optimising a detector using a training set that contains only normal images. UAD approaches can be based on reconstruction methods, self-supervised approaches, and Imagenet pre-trained models. Reconstruction methods, which detect anomalies from image reconstruction errors, are advantageous because they do not rely on the design of problem-specific pretext tasks needed by self-supervised approaches, and on the unreliable translation of models pre-trained from non-medical datasets. However, reconstruction methods may fail because they can have low reconstruction errors even for anomalous images. In this paper, we introduce a new reconstruction-based UAD approach …


Ubisurface: A Robotic Touch Surface For Supporting Mid-Air Planar Interactions In Room-Scale Vr, Ryota Gomi, Kazuki Takashima, Yuki Onishi, Kazuyuki Fujita, Yoshifumi Kitamura Oct 2023

Ubisurface: A Robotic Touch Surface For Supporting Mid-Air Planar Interactions In Room-Scale Vr, Ryota Gomi, Kazuki Takashima, Yuki Onishi, Kazuyuki Fujita, Yoshifumi Kitamura

Research Collection School Of Computing and Information Systems

Room-scale VR has been considered an alternative to physical office workspaces. For office activities, users frequently require planar input methods, such as typing or handwriting, to quickly record annotations to virtual content. However, current off-The-shelf VR HMD setups rely on mid-Air interactions, which can cause arm fatigue and decrease input accuracy. To address this issue, we propose UbiSurface, a robotic touch surface that can automatically reposition itself to physically present a virtual planar input surface (VR whiteboard, VR canvas, etc.) to users and to permit them to achieve accurate and fatigue-less input while walking around a virtual room. We design …


Balanced Blended Space: Proposing A Universal Theoretical Framework For Combinative Reality, David Smith, Frederick Bianchi Oct 2023

Balanced Blended Space: Proposing A Universal Theoretical Framework For Combinative Reality, David Smith, Frederick Bianchi

Publications and Research

In today's fragmented societies, a unified framework for communication and collaboration across different realities is crucial. We introduce Balanced Blended Space (BBS) as a framework for describing combinative reality, encompassing virtual, physical, and conceptual realms, all intrinsically connected. Interactions within these environments shape our perceptual space. This paper outlines key axiomatic assumptions, criteria for a universal framework, and fundamental terminology. We identify deep symmetries enabling the BBS framework, including Cognitive and Computational Symmetry, Physical and Virtual Symmetry, Mediation Pathway Symmetry, Space-Time Symmetry, and Sensory Symmetry. We propose tests to determine its viability, emphasizing virtual intelligence as a collaborative partner. We …


Ai Vs. Ai: Can Ai Detect Ai-Generated Images?, Samah S. Baraheem, Tam Van Nguyen Sep 2023

Ai Vs. Ai: Can Ai Detect Ai-Generated Images?, Samah S. Baraheem, Tam Van Nguyen

Computer Science Faculty Publications

The proliferation of Artificial Intelligence (AI) models such as Generative Adversarial Net- works (GANs) has shown impressive success in image synthesis. Artificial GAN-based synthesized images have been widely spread over the Internet with the advancement in generating naturalistic and photo-realistic images. This might have the ability to improve content and media; however, it also constitutes a threat with regard to legitimacy, authenticity, and security. Moreover, implementing an automated system that is able to detect and recognize GAN-generated images is significant for image synthesis models as an evaluation tool, regardless of the input modality. To this end, we propose a framework …


Edge Distraction-Aware Salient Object Detection, Sucheng Ren, Wenxi Liu, Jianbo Jiao, Guoqiang Han, Shengfeng He Sep 2023

Edge Distraction-Aware Salient Object Detection, Sucheng Ren, Wenxi Liu, Jianbo Jiao, Guoqiang Han, Shengfeng He

Research Collection School Of Computing and Information Systems

Integrating low-level edge features has been proven to be effective in preserving clear boundaries of salient objects. However, the locality of edge features makes it difficult to capture globally salient edges, leading to distraction in the final predictions. To address this problem, we propose to produce distraction-free edge features by incorporating cross-scale holistic interdependencies between high-level features. In particular, we first formulate our edge features extraction process as a boundary-filling problem. In this way, we enforce edge features to focus on closed boundaries instead of those disconnected background edges. Second, we propose to explore cross-scale holistic contextual connections between every …


Graph-Level Anomaly Detection Via Hierarchical Memory Networks, Chaoxi Niu, Guansong Pang, Ling Chen Sep 2023

Graph-Level Anomaly Detection Via Hierarchical Memory Networks, Chaoxi Niu, Guansong Pang, Ling Chen

Research Collection School Of Computing and Information Systems

Graph-level anomaly detection aims to identify abnormal graphs that exhibit deviant structures and node attributes compared to the majority in a graph set. One primary challenge is to learn normal patterns manifested in both fine-grained and holistic views of graphs for identifying graphs that are abnormal in part or in whole. To tackle this challenge, we propose a novel approach called Hierarchical Memory Networks (HimNet), which learns hierarchical memory modules---node and graph memory modules---via a graph autoencoder network architecture. The node-level memory module is trained to model fine-grained, internal graph interactions among nodes for detecting locally abnormal graphs, while the …


One Font Doesn’T Fit All: The Influence Of Digital Text Personalization On Comprehension In Child And Adolescent Readers, Shannon M. Sheppard, Susanne L. Nobles, Anton Palma, Sophie Kajfez, Marjorie Jordan, Kathy Crowley, Sofie Beier Aug 2023

One Font Doesn’T Fit All: The Influence Of Digital Text Personalization On Comprehension In Child And Adolescent Readers, Shannon M. Sheppard, Susanne L. Nobles, Anton Palma, Sophie Kajfez, Marjorie Jordan, Kathy Crowley, Sofie Beier

Communication Sciences and Disorders Faculty Articles and Research

Reading comprehension is an essential skill. It is unclear whether and to what degree typography and font personalization may impact reading comprehension in younger readers. With advancements in technology, it is now feasible to personalize digital reading formats in general technology tools, but this feature is not yet available for many educational tools. The current study aimed to investigate the effect of character width and inter-letter spacing on reading speed and comprehension. We enrolled 94 children (kindergarten–8th grade) and compared performance with six font variations on a word-level semantic decision task (Experiment 1) and a passage-level comprehension task (Experiment 2). …


Human Recognition Theory And Facial Recognition Technology: A Topic Modeling Approach To Understanding The Ethical Implication Of A Developing Algorithmic Technologies Landscape On How We View Ourselves And Are Viewed By Others, Hajer Albalawi Aug 2023

Human Recognition Theory And Facial Recognition Technology: A Topic Modeling Approach To Understanding The Ethical Implication Of A Developing Algorithmic Technologies Landscape On How We View Ourselves And Are Viewed By Others, Hajer Albalawi

Electronic Theses and Dissertations, 2020-2023

The emergence of algorithmic-driven technology has significantly impacted human life in the current century. Algorithms, as versatile constructs, hold different meanings across various disciplines, including computer science, mathematics, social science, and human-artificial intelligence studies. This study defines algorithms from an ethical perspective as the foundation of an information society and focuses on their implications in the context of human recognition. Facial recognition technology, driven by algorithms, has gained widespread use, raising important ethical questions regarding privacy, bias, and accuracy. This dissertation aims to explore the impact of algorithms on machine perception of human individuals and how humans perceive one another …


Studying Memes During Covid Lockdown As A Lens Through Which To Understand Video-Mediated Communication Interactions, Tatyana Claytor Aug 2023

Studying Memes During Covid Lockdown As A Lens Through Which To Understand Video-Mediated Communication Interactions, Tatyana Claytor

Electronic Theses and Dissertations, 2020-2023

The purpose of this study is to analyze image macros about video-mediated communication (VMC) created during the time frame of 2020-2021 when people all over the world started using Zoom and VMC for work and school. It is a unique opportunity to study how users' interactions with themselves and with others were affected at a time when a lot of people started using the technology at the same time. Because the focus is on interactions, I narrowed it down to three topics to analyze the memes: presence, self, and space and place to analyze the memes. I chose memes relating …


Signings Of Graphs And Sign-Symmetric Signed Graphs, Ahmad Asiri Aug 2023

Signings Of Graphs And Sign-Symmetric Signed Graphs, Ahmad Asiri

Theses and Dissertations

In this dissertation, we investigate various aspects of signed graphs, with a particular focus on signings and sign-symmetric signed graphs. We begin by examining the complete graph on six vertices with one edge deleted ($K_6$\textbackslash e) and explore the different ways of signing this graph up to switching isomorphism. We determine the frustration index (number) of these signings and investigate the existence of sign-symmetric signed graphs. We then extend our study to the $K_6$\textbackslash 2e graph and the McGee graph with exactly two negative edges. We investigate the distinct ways of signing these graphs up to switching isomorphism and demonstrate …


Visual And Spatial Audio Mismatching In Virtual Environments, Zachary Lawrence Garris Aug 2023

Visual And Spatial Audio Mismatching In Virtual Environments, Zachary Lawrence Garris

Theses and Dissertations

This paper explores how vision affects spatial audio perception in virtual reality. We created four virtual environments with different reverb and room sizes, and recorded binaural clicks in each one. We conducted two experiments: one where participants judged the audio-visual match, and another where they pointed to the click direction. We found that vision influences spatial audio perception and that congruent audio-visual cues improve accuracy. We suggest some implications for virtual reality design and evaluation.


Understanding The Role Of Interactivity And Explanation In Adaptive Experiences, Lijie Guo Aug 2023

Understanding The Role Of Interactivity And Explanation In Adaptive Experiences, Lijie Guo

All Dissertations

Adaptive experiences have been an active area of research in the past few decades, accompanied by advances in technology such as machine learning and artificial intelligence. Whether the currently ongoing research on adaptive experiences has focused on personalization algorithms, explainability, user engagement, or privacy and security, there is growing interest and resources in developing and improving these research focuses. Even though the research on adaptive experiences has been dynamic and rapidly evolving, achieving a high level of user engagement in adaptive experiences remains a challenge. %????? This dissertation aims to uncover ways to engage users in adaptive experiences by incorporating …


All Hands On Deck: Choosing Virtual End Effector Representations To Improve Near Field Object Manipulation Interactions In Extended Reality, Roshan Venkatakrishnan Aug 2023

All Hands On Deck: Choosing Virtual End Effector Representations To Improve Near Field Object Manipulation Interactions In Extended Reality, Roshan Venkatakrishnan

All Dissertations

Extended reality, or "XR", is the adopted umbrella term that is heavily gaining traction to collectively describe Virtual reality (VR), Augmented reality (AR), and Mixed reality (MR) technologies. Together, these technologies extend the reality that we experience either by creating a fully immersive experience like in VR or by blending in the virtual and "real" worlds like in AR and MR.

The sustained success of XR in the workplace largely hinges on its ability to facilitate efficient user interactions. Similar to interacting with objects in the real world, users in XR typically interact with virtual integrants like objects, menus, windows, …


Anatoview: Using Interactive 3d Visualizations With Augmented Reality Support For Laypersons’ Medical Education In Informed Consent Processes, Michelle Chen Aug 2023

Anatoview: Using Interactive 3d Visualizations With Augmented Reality Support For Laypersons’ Medical Education In Informed Consent Processes, Michelle Chen

Dartmouth College Master’s Theses

AnatoView is an interactive multimedia educational application that visualizes medical procedures in three-dimensional (3D), augmented reality (AR) space. By providing visual and spatial information of medical procedures, AnatoView acts as a learning supplement for laypersons/patients in informed consent (IC) processes — wherein instructional content is traditionally limited to purely spoken explanations that lead to poor patient comprehension. We design a mixed study and conduct a randomized, controlled trial with 15 laypersons as participants: administering a traditional IC process to a control group, and an IC process supplemented by the use of AnatoView to experimental groups. As a primary outcome, medical …


The Effects Of Primary And Secondary Task Workloads On Cybersickness In Immersive Virtual Active Exploration Experiences, Rohith Venkatakrishnan Aug 2023

The Effects Of Primary And Secondary Task Workloads On Cybersickness In Immersive Virtual Active Exploration Experiences, Rohith Venkatakrishnan

All Dissertations

Virtual reality (VR) technology promises to transform humanity. The technology enables users to explore and interact with computer-generated environments that can be simulated to approximate or deviate from reality. This creates an endless number of ways to propitiously apply the technology in our lives. It follows that large technological conglomerates are pushing for the widespread adoption of VR, financing the creation of the Metaverse - a hypothetical representation of the next iteration of the internet.

Even with VR technology's continuous growth, its widespread adoption remains long overdue. This can largely be attributed to an affliction called cybersickness, an analog to …


Hyperbolic Graph Topic Modeling Network With Continuously Updated Topic Tree, Ce Zhang, Rex Ying, Hady Wirawan Lauw Aug 2023

Hyperbolic Graph Topic Modeling Network With Continuously Updated Topic Tree, Ce Zhang, Rex Ying, Hady Wirawan Lauw

Research Collection School Of Computing and Information Systems

Connectivity across documents often exhibits a hierarchical network structure. Hyperbolic Graph Neural Networks (HGNNs) have shown promise in preserving network hierarchy. However, they do not model the notion of topics, thus document representations lack semantic interpretability. On the other hand, a corpus of documents usually has high variability in degrees of topic specificity. For example, some documents contain general content (e.g., sports), while others focus on specific themes (e.g., basketball and swimming). Topic models indeed model latent topics for semantic interpretability, but most assume a flat topic structure and ignore such semantic hierarchy. Given these two challenges, we propose a …


Flood: A Flexible Invariant Learning Framework For Out-Of-Distribution Generalization On Graphs, Yang Liu, Xiang Ao, Fuli Feng, Yunshan Ma, Kuan Li, Tat‑Seng Chua, Qing He Aug 2023

Flood: A Flexible Invariant Learning Framework For Out-Of-Distribution Generalization On Graphs, Yang Liu, Xiang Ao, Fuli Feng, Yunshan Ma, Kuan Li, Tat‑Seng Chua, Qing He

Research Collection School Of Computing and Information Systems

Graph Neural Networks (GNNs) have achieved remarkable success in various domains but most of them are developed under the in-distribution assumption. Under out-of-distribution (OOD) settings, they suffer from the distribution shift between the training set and the test set and may not generalize well to the test distribution. Several methods have tried the invariance principle to improve the generalization of GNNs in OOD settings. However, in previous solutions, the graph encoder is immutable after the invariant learning and cannot be adapted to the target distribution flexibly. Confronting the distribution shift, a flexible encoder with refinement to the target distribution can …


Socialz: Multi-Feature Social Fuzz Testing, Francisco Zanartu, Christoph Treude, Markus Wagner Jul 2023

Socialz: Multi-Feature Social Fuzz Testing, Francisco Zanartu, Christoph Treude, Markus Wagner

Research Collection School Of Computing and Information Systems

Online social networks have become an integral aspect of our daily lives and play a crucial role in shaping our relationships with others. However, bugs and glitches, even minor ones, can cause anything from frustrating problems to serious data leaks that can have farreaching impacts on millions of users. To mitigate these risks, fuzz testing, a method of testing with randomised inputs, can provide increased confidence in the correct functioning of a social network. However, implementing traditional fuzz testing methods can be prohibitively difficult or impractical for programmers outside of the network’s development team. To tackle this challenge, we present …


Contrastive Video Question Answering Via Video Graph Transformer, Junbin Xiao Xiao, Pan Zhou, Angela Yao, Yicong Li, Richang Hong, Shuicheng Yan, Tat-Seng Chua Jul 2023

Contrastive Video Question Answering Via Video Graph Transformer, Junbin Xiao Xiao, Pan Zhou, Angela Yao, Yicong Li, Richang Hong, Shuicheng Yan, Tat-Seng Chua

Research Collection School Of Computing and Information Systems

We propose to perform video question answering (VideoQA) in a Contrastive manner via a Video Graph Transformer model (CoVGT). CoVGT’s uniqueness and superiority are three-fold: 1) It proposes a dynamic graph transformer module which encodes video by explicitly capturing the visual objects, their relations and dynamics, for complex spatio-temporal reasoning. 2) It designs separate video and text transformers for contrastive learning between the video and text to perform QA, instead of multi-modal transformer for answer classification. Fine-grained video-text communication is done by additional cross-modal interaction modules. 3) It is optimized by the joint fully- and self-supervised contrastive objectives between the …


Balanced Blended Space: Foundational Human–Ai Dialogues In A Symmetry-Based Mediation Framework, David Smith Jul 2023

Balanced Blended Space: Foundational Human–Ai Dialogues In A Symmetry-Based Mediation Framework, David Smith

Publications and Research

This working paper documents the early development of the Balanced Blended Space (BBS) framework through a series of iterative interactions between a cognitive agent (human researcher) and a computational agent (AI system) conducted in 2023. The work is motivated by the need for a universal theoretical model capable of describing the integration of physical, virtual, and conceptual spaces, particularly in response to increasing fragmentation across contemporary communication systems.

BBS is proposed as a symmetry-based mediation framework in which relationships between domains—such as physical and virtual space, cognition and computation, and multiple sensory modalities—are treated as structurally equivalent and mappable. Central …


Fine-Grained Domain Adaptive Crowd Counting Via Point-Derived Segmentation, Yongtuo Liu, Dan Xu, Sucheng Ren, Hanjie Wu, Hongmin Cai, Shengfeng He Jul 2023

Fine-Grained Domain Adaptive Crowd Counting Via Point-Derived Segmentation, Yongtuo Liu, Dan Xu, Sucheng Ren, Hanjie Wu, Hongmin Cai, Shengfeng He

Research Collection School Of Computing and Information Systems

Due to domain shift, a large performance drop is usually observed when a trained crowd counting model is deployed in the wild. While existing domain-adaptive crowd counting methods achieve promising results, they typically regard each crowd image as a whole and reduce domain discrepancies in a holistic manner, thus limiting further improvement of domain adaptation performance. To this end, we propose to untangle domain-invariant crowd and domain-specific background from crowd images and design a fine-grained domain adaption method for crowd counting. Specifically, to disentangle crowd from background, we propose to learn crowd segmentation from point-level crowd counting annotations in a …


A Review On The Effects Of Chanting And Solfeggio Frequencies On Well-Being, Xuyu Yang, Fiona Fui-Hoon Nah, Fen Lin Jul 2023

A Review On The Effects Of Chanting And Solfeggio Frequencies On Well-Being, Xuyu Yang, Fiona Fui-Hoon Nah, Fen Lin

Research Collection School Of Computing and Information Systems

This paper presents our literature review on how chanting and solfeggio frequencies affect brain activity, enhance well-being, and inform the design of sound therapy. We call for more scientific research to investigate and better understand the effects of chanting and solfeggio frequencies on well-being.


Robust And Parallel Segmentation Model (Rpsm) For Early Detection Of Skin Cancer Disease Using Heterogeneous Distributions, Nancy Zreika, Ali El-Zaart, Abdallah El Chakik Jun 2023

Robust And Parallel Segmentation Model (Rpsm) For Early Detection Of Skin Cancer Disease Using Heterogeneous Distributions, Nancy Zreika, Ali El-Zaart, Abdallah El Chakik

BAU Journal - Science and Technology

Melanoma is the most common dangerous type of skin cancer; however, it is preventable if it is diagnosed early. Diagnosis of Melanoma would be improved if an accurate skin image segmentation model is available. Many computer vision methods have been investigated, yet the problem of finding a consistent and robust model that extracts the best threshold value, persists. This paper suggests a novel image segmentation approach using a multilevel cross entropy thresholding algorithm based on heterogeneous distributions. The proposed strategy searches the problem space by segmenting the image into several levels, and applying for each level one of the three …


Deep-Learning Realtime Upsampling Techniques In Video Games, Biruk Mengistu Jun 2023

Deep-Learning Realtime Upsampling Techniques In Video Games, Biruk Mengistu

Scholarly Horizons: University of Minnesota, Morris Undergraduate Journal

This paper addresses the challenge of keeping up with the ever-increasing graphical complexity of video games and introduces a deep-learning approach to mitigating it. As games get more and more demanding in terms of their graphics, it becomes increasingly difficult to maintain high-quality images while also ensuring good performance. This is where deep learning super sampling (DLSS) comes in. The paper explains how DLSS works, including the use of convolutional autoencoder neural networks and various other techniques and technologies. It also covers how the network is trained and optimized, as well as how it incorporates temporal antialiasing and frame generation …


How Photorealistic Images Are Generated, Nahom Ketema Jun 2023

How Photorealistic Images Are Generated, Nahom Ketema

University Honors Theses

The field of computer graphics looks into how computers can be used to generate images. From using some trigonometry to plot 3D objects to using rays to calculate the lighting of an object, there are a variety of ways that we can use to draw objects onto a screen. For this thesis, we will be looking at a few of those methods to determine how photorealistic images are generated.


Interstice, Shravan Rao Jun 2023

Interstice, Shravan Rao

Masters Theses

When I was about three years old, I distinctly remember being too small to see what was on top of the table. A couple of years later, when I could see those objects, I thought the world around me had grown smaller. In a way, it did, as I experienced, lived, captured, remembered, and shared the space repeatedly. This sense of the world shrinking was exaggerated during the Covid-19 pandemic, allowing new behaviours and modes of interaction to emerge. Continually shaping our modern lives, virtual technologies redefine how we access and share information and stories or even explore new places. …


An Investigation Into Machine Learning Techniques For Designing Dynamic Difficulty Agents In Real-Time Games, Ryan Adare Dunagan Jun 2023

An Investigation Into Machine Learning Techniques For Designing Dynamic Difficulty Agents In Real-Time Games, Ryan Adare Dunagan

Electronic Theses and Dissertations

Video games are an incredibly popular pastime enjoyed by people of all ages world wide. Many different kinds of games exist, but most games feature some elements of the player overcoming some challenge, usually through gameplay. These challenges are insurmountable for some people and may turn them off to video games as a pastime. Games can be made more accessible to players of little skill and/or experience through the use of Dynamic Difficulty Adjustment (DDA) systems that adjust the difficulty of the game in response to the player’s performance. This research seeks to establish the effectiveness of machine learning techniques …


Generalizing Graph Neural Networks Across Graphs, Time, And Tasks, Zhihao Wen Jun 2023

Generalizing Graph Neural Networks Across Graphs, Time, And Tasks, Zhihao Wen

Dissertations and Theses Collection (Open Access)

Graph-structured data are ubiquitous across numerous real-world contexts, encompassing social networks, commercial graphs, bibliographic networks, and biological systems. Delving into the analysis of these graphs can yield significant understanding pertaining to their corresponding application fields.Graph representation learning offers a potent solution to graph analytics challenges by transforming a graph into a low-dimensional space while preserving its information to the greatest extent possible. This conversion into low-dimensional vectors enables the efficient computation of subsequent graph algorithms. The majority of prior research has concentrated on deriving node representations from a single, static graph. However, numerous real-world situations demand rapid generation of representations …


Mosaic: Spatially-Multiplexed Edge Ai Optimization Over Multiple Concurrent Video Sensing Streams, Ila Gokarn, Hemanth Sabbella, Yigong Hu, Tarek Abdelzaher, Archan Misra Jun 2023

Mosaic: Spatially-Multiplexed Edge Ai Optimization Over Multiple Concurrent Video Sensing Streams, Ila Gokarn, Hemanth Sabbella, Yigong Hu, Tarek Abdelzaher, Archan Misra

Research Collection School Of Computing and Information Systems

Sustaining high fidelity and high throughput of perception tasks over vision sensor streams on edge devices remains a formidable challenge, especially given the continuing increase in image sizes (e.g., generated by 4K cameras) and complexity of DNN models. One promising approach involves criticality-aware processing, where the computation is directed selectively to "critical" portions of individual image frames. We introduce MOSAIC, a novel system for such criticality-aware concurrent processing of multiple vision sensing streams that provides a multiplicative increase in the achievable throughput with negligible loss in perception fidelity. MOSAIC determines critical regions from images received from multiple vision …