Open Access. Powered by Scholars. Published by Universities.®
Graphics and Human Computer Interfaces Commons™
Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Databases and Information Systems (473)
- Artificial Intelligence and Robotics (400)
- Engineering (335)
- Software Engineering (307)
- Social and Behavioral Sciences (264)
-
- Other Computer Sciences (241)
- Computer Engineering (178)
- Arts and Humanities (142)
- Theory and Algorithms (133)
- Education (126)
- OS and Networks (97)
- Medicine and Health Sciences (93)
- Numerical Analysis and Scientific Computing (93)
- Systems Architecture (81)
- Art and Design (79)
- Business (78)
- Programming Languages and Compilers (78)
- Psychology (77)
- Communication (74)
- Electrical and Computer Engineering (73)
- Data Storage Systems (71)
- Information Security (69)
- Life Sciences (58)
- Data Science (51)
- Educational Technology (48)
- Library and Information Science (40)
- Communication Technology and New Media (39)
- Institution
-
- Singapore Management University (938)
- University of Dayton (114)
- Air Force Institute of Technology (98)
- Old Dominion University (97)
- California Polytechnic State University, San Luis Obispo (96)
-
- University of Arkansas, Fayetteville (89)
- University of Nebraska - Lincoln (51)
- City University of New York (CUNY) (48)
- Technological University Dublin (48)
- University of Malaya (42)
- San Jose State University (37)
- Dartmouth College (34)
- Embry-Riddle Aeronautical University (24)
- Clemson University (23)
- Purdue University (23)
- Rochester Institute of Technology (23)
- The University of Akron (22)
- Chapman University (20)
- Edith Cowan University (20)
- University of Kentucky (18)
- Michigan Technological University (16)
- University of Central Florida (15)
- Southern Adventist University (13)
- California State University, San Bernardino (12)
- Kennesaw State University (12)
- St. Mary's University (12)
- Nova Southeastern University (11)
- University of Minnesota Morris Digital Well (11)
- Louisiana State University (10)
- University of Nevada, Las Vegas (10)
- Keyword
-
- Virtual reality (62)
- Visualization (46)
- Computer graphics (38)
- Computer vision (37)
- Accessibility (36)
-
- Human-computer interaction (35)
- Augmented reality (33)
- Usability (31)
- Machine learning (29)
- Computer Science (25)
- Data visualization (25)
- Machine Learning (25)
- Artificial intelligence (24)
- Deep learning (24)
- Virtual Reality (23)
- HCI (22)
- Computer science (20)
- Eye tracking (20)
- Human computer interaction (20)
- User experience (20)
- Design (19)
- Deep Learning (16)
- Education (16)
- Feature extraction (15)
- Graph Neural Networks (15)
- Graphics (15)
- VR (15)
- Gamification (14)
- Image processing (14)
- Applied sciences (13)
- Publication Year
- Publication
-
- Research Collection School Of Computing and Information Systems (912)
- Computer Science Faculty Publications (135)
- Theses and Dissertations (98)
- Master's Theses (50)
- Graduate Theses and Dissertations (43)
-
- Student Works (2000-2009) (33)
- Computer Science and Computer Engineering Undergraduate Honors Theses (31)
- 3-D Printed Model Structural Files (29)
- Publications and Research (28)
- Dartmouth College Master’s Theses (24)
- Williams Honors College, Honors Research Projects (22)
- Master's Projects (20)
- Conference papers (19)
- Frameless (19)
- Computer Science and Software Engineering (18)
- All Dissertations (17)
- Dissertations and Theses Collection (Open Access) (16)
- Dissertations, Master's Theses and Master's Reports (16)
- H-Workload 2017: Models and Applications (Works in Progress) (15)
- Computer Engineering (14)
- Electronic Theses and Dissertations (14)
- Theses : Honours (14)
- Honors Theses (13)
- MAICS: The Modern Artificial Intelligence and Cognitive Science Conference (12)
- AFIT Patents (11)
- CCAC Theses and Dissertations (11)
- Engineering Faculty Articles and Research (11)
- Scholarly Horizons: University of Minnesota, Morris Undergraduate Journal (10)
- Inquiry: The University of Arkansas Undergraduate Research Journal (9)
- Publications (9)
- Publication Type
- File Type
Articles 481 - 510 of 2362
Full-Text Articles in Graphics and Human Computer Interfaces
Dynamic Meta-Path Guided Temporal Heterogeneous Graph Neural Networks, Yugang Ji, Chuan Shi, Yuan Fang
Dynamic Meta-Path Guided Temporal Heterogeneous Graph Neural Networks, Yugang Ji, Chuan Shi, Yuan Fang
Research Collection School Of Computing and Information Systems
Graph Neural Networks (GNNs) have become the de facto standard for representation learning on topological graphs, which usually derive effective node representations via message passing from neighborhoods. Although GNNs have achieved great success, previous models are mostly confined to static and homogeneous graphs. However, there are multiple dynamic interactions between different-typed nodes in real-world scenarios like academic networks and e-commerce platforms, forming temporal heterogeneous graphs (THGs). Limited work has been done for representation learning on THGs and the challenges are in two aspects. First, there are abundant dynamic semantics between nodes while traditional techniques like meta-paths can only capture static …
Effects Of Mindfulness And Emotion Regulation On Aesthetics: A Theoretical Model From Hedonic Perspective Of Processing Fluency, Geng-Bao Lin, Fiona Fui-Hoon Nah, Choon Ling Sia
Effects Of Mindfulness And Emotion Regulation On Aesthetics: A Theoretical Model From Hedonic Perspective Of Processing Fluency, Geng-Bao Lin, Fiona Fui-Hoon Nah, Choon Ling Sia
Research Collection School Of Computing and Information Systems
Research has shown that processing fluency positively impacts perceived aesthetics, with pleasure mediating the relationship. Considering the important role of pleasure, we propose studying the role of emotion regulation in moderating the mediated relationship from processing fluency to perceived aesthetics. Based on our hypotheses, individuals’ emotion regulation strategies are expected to have moderating effects on the relationship between processing fluency and perceived aesthetics such that cognitive reappraisal positively moderates the relationship from processing fluency to pleasure, and expressive suppression negatively moderates the relationship from pleasure to perceived aesthetics. Trait mindfulness is also expected to influence perceived aesthetics through emotion regulation …
Random Walk Methods For Geometry Representation Agnostic Transport, Dario R. Seyb
Random Walk Methods For Geometry Representation Agnostic Transport, Dario R. Seyb
Dartmouth College Ph.D Dissertations
In computer graphics, we use geometry representations to model a wide range of virtual scenes—from the fantastical worlds shown in animated movies to intricate mechanical parts.
These representations provide the context for transport problems—light transport is used to produce images of virtual scenes and diffusive transport to simulate distributions of quantities like heat.
There are many types of representations each with their own advantages.
For example, explicit ones make it easy to directly manipulate surfaces, while implicit representations allow for intuitive modeling by non-technical users and straightforward integration into machine learning systems.
Unfortunately, many algorithms that work on these digital …
Trust: The Feature That Vending Machines And Atms Share, But Simplygo Lacks, Sun Sun Lim
Trust: The Feature That Vending Machines And Atms Share, But Simplygo Lacks, Sun Sun Lim
Research Collection College of Integrative Studies
The article discussed the intricacies of trust in the SimplyGo debacle and highlighted how the design of physical interfaces like vending machines and ATMs and digital interfaces from apps like Grab, Parking.sg and ShopBack have critical features to instil trust. People need to be reassured that their transactions have proceeded as they should, and thay have not been short-changed.
Enhancing Research Productivity: Seamless Integration Of Personal Devices And Hpc Resources With The Cybershuttle Notebook Gateway, Yasith Jayawardana, Dimuthu Wannipurage, Eroma Abeysinghe, Suresh Marru
Enhancing Research Productivity: Seamless Integration Of Personal Devices And Hpc Resources With The Cybershuttle Notebook Gateway, Yasith Jayawardana, Dimuthu Wannipurage, Eroma Abeysinghe, Suresh Marru
Computer Science Faculty Publications
Scientists often utilize personal laptops and workstations for initial research stages and turn to high-performance computing (HPC) supercomputers for compute-intensive tasks. However, seamless transitions between these environments are vital for enhancing productivity and accelerating research progress. Our paper presents the Cybershuttle Notebook Gateway, an open-source framework crafted to streamline this transition, optimize resource utilization, and reduce time-to-science for researchers. Leveraging JupyterLab, the framework extends kernel mechanics for seamless provisioning and connection to remote HPC cluster kernels. We delve into its architecture, which separates user authentication, kernel provisioning, and remote file system access. Additionally, we highlight practical capabilities like analyzing network …
An Analysis Of Precision: Occlusion And Perspective Geometry’S Role In 6d Pose Estimation, Jeffrey Choate, Derek Worth, Scott Nykl, Clark N. Taylor, Brett J. Borghetti, Christine M. Schubert Kabban
An Analysis Of Precision: Occlusion And Perspective Geometry’S Role In 6d Pose Estimation, Jeffrey Choate, Derek Worth, Scott Nykl, Clark N. Taylor, Brett J. Borghetti, Christine M. Schubert Kabban
Faculty Publications
Achieving precise 6 degrees of freedom (6D) pose estimation of rigid objects from color images is a critical challenge with wide-ranging applications in robotics and close-contact aircraft operations. This study investigates key techniques in the application of YOLOv5 object detection convolutional neural network (CNN) for 6D pose localization of aircraft using only color imagery. Traditional object detection labeling methods suffer from inaccuracies due to perspective geometry and being limited to visible key points. This research demonstrates that with precise labeling, a CNN can predict object features with near-pixel accuracy, effectively learning the distinct appearance of the object due to perspective …
A-Disetrac Advanced Analytic Dashboard For Distributed Eye Tracking, Yasasi Abeysinghe, Bhanuka Mahanama, Gavindya Jayawardena, Yasith Jayawardena, Mohan Sunkara, Andrew T. Duchowski, Vikas Ashok, Sampath Jayarathna
A-Disetrac Advanced Analytic Dashboard For Distributed Eye Tracking, Yasasi Abeysinghe, Bhanuka Mahanama, Gavindya Jayawardena, Yasith Jayawardena, Mohan Sunkara, Andrew T. Duchowski, Vikas Ashok, Sampath Jayarathna
Computer Science Faculty Publications
Understanding how individuals focus and perform visual searches during collaborative tasks can help improve user engagement. Eye tracking measures provide informative cues for such understanding. This article presents A-DisETrac, an advanced analytic dashboard for distributed eye tracking. It uses off-the-shelf eye trackers to monitor multiple users in parallel, compute both traditional and advanced gaze measures in real-time, and display them on an interactive dashboard. Using two pilot studies, the system was evaluated in terms of user experience and utility, and compared with existing work. Moreover, the system was used to study how advanced gaze measures such as ambient-focal coefficient K …
All In One Place: Ensuring Usable Access To Online Shopping Items For Blind Users, Yash Prakash, Akshay Kolgar Nayak, Mohan Sunkara, Sampath Jayarathna, Hae-Na Lee, Vikas Ashok
All In One Place: Ensuring Usable Access To Online Shopping Items For Blind Users, Yash Prakash, Akshay Kolgar Nayak, Mohan Sunkara, Sampath Jayarathna, Hae-Na Lee, Vikas Ashok
Computer Science Faculty Publications
Perusing web data items such as shopping products is a core online user activity. To prevent information overload, the content associated with data items is typically dispersed across multiple webpage sections over multiple web pages. However, such content distribution manifests an unintended side effect of significantly increasing the interaction burden for blind users, since navigating to-and-fro between different sections in different pages is tedious and cumbersome with their screen readers. While existing works have proposed methods for the context of a single webpage, solutions enabling usable access to content distributed across multiple webpages are few and far between. In this …
Improving Usability Of Data Charts In Multimodal Documents For Low Vision Users, Yash Prakash, Akshay Kolgar Nayak, Shoaib Mohammed Alyaan, Pathan Aseef Khan, Hae-Na Lee, Vikas Ashok
Improving Usability Of Data Charts In Multimodal Documents For Low Vision Users, Yash Prakash, Akshay Kolgar Nayak, Shoaib Mohammed Alyaan, Pathan Aseef Khan, Hae-Na Lee, Vikas Ashok
Computer Science Faculty Publications
Data chart visualizations and text are often paired in news articles, online blogs, and academic publications to present complex data. While chart visualizations offer graphical summaries of the data, the accompanying text provides essential context and explanation. Associating information from text and charts is straightforward for sighted users but presents significant challenges for individuals with low vision, especially on small-screen devices such as smartphones. The visual nature of charts coupled with the layout of the text inherently makes it difficult for low vision users to mentally associate chart data with text and comprehend the content due to their dependence on …
Breaking Down Computer Networking Instructional Videos: Automatic Summarization With Video Attributes And Language Models, Totok Sukardiyono, Muhammad Irfan Luthfi, Nisa Dwi Septiyanti
Breaking Down Computer Networking Instructional Videos: Automatic Summarization With Video Attributes And Language Models, Totok Sukardiyono, Muhammad Irfan Luthfi, Nisa Dwi Septiyanti
Elinvo (Electronics, Informatics, and Vocational Education)
Instructional videos have become a popular tool for teaching complex topics in computer networking. However, these videos can often be lengthy and time-consuming, making it difficult for learners to obtain the key information they need. In this study, we propose an approach that leverages automatic summarization and language models to generate concise and informative summaries of instructional videos. To enhance the performance of the summarization algorithm, we also incorporate video attributes that provide contextual information about the video content. Using a dataset of computer networking tutorials, we evaluate the effectiveness of the proposed method and show that it significantly improves …
Differences In Software Usability Level Based On User Background, Abdur Rohman Sholeh, Agung Fatwanto
Differences In Software Usability Level Based On User Background, Abdur Rohman Sholeh, Agung Fatwanto
Elinvo (Electronics, Informatics, and Vocational Education)
The development of software must consider usability as one of its key success indicators. Relatively few studies discuss the factors influencing usability, including users' backgrounds. The purpose of this research was to investigate the impact of user background, specifically gender, class (year of college admission), and frequency of use on the rating of usability. This research utilized a descriptive quantitative method with instruments: the usability matrix of the Computer System Usability Questionnaire (CSUQ), the System Usability Scale (SUS), the Usability Metric for User Experience (UMUX), and the Net Promoter Score (NPS). The research object was the Learning Management System (LMS) …
Usability Of Mobile Application For Implementing Genetic Counselling Intervention Among Thalassemia Patients And Caregivers: A Case Study Of Cyber Gen, Henri Setiawan, Nur Hidayat, Atun Farihatun, Marlina Indriastuti, Rudi Kurniawan, Andan Firmansyah, Esti Andarini, Yudisa Diaz Lutfi Sandi
Usability Of Mobile Application For Implementing Genetic Counselling Intervention Among Thalassemia Patients And Caregivers: A Case Study Of Cyber Gen, Henri Setiawan, Nur Hidayat, Atun Farihatun, Marlina Indriastuti, Rudi Kurniawan, Andan Firmansyah, Esti Andarini, Yudisa Diaz Lutfi Sandi
Elinvo (Electronics, Informatics, and Vocational Education)
Utilization of communication and information technology has been widely used in the health sector, especially nursing. As one of the nursing interventions for thalassemia patients and caregivers, genetic counseling is not only done face to face but can use android-based telenursing facilities through the complete features available in the Cyber Gen application. This study aims to measure the usability level of Cyber Gen application as an indirect genetic counseling medium for thalassemia patients. This application was developed with four main services: basic information about diseases, consultation rooms, social support, and direct surveys. This application is built using the Flutter Framework, …
Reducing Food Scarcity: The Benefits Of Urban Farming, S.A. Claudell, Emilio Mejia
Reducing Food Scarcity: The Benefits Of Urban Farming, S.A. Claudell, Emilio Mejia
Journal of Nonprofit Innovation
Urban farming can enhance the lives of communities and help reduce food scarcity. This paper presents a conceptual prototype of an efficient urban farming community that can be scaled for a single apartment building or an entire community across all global geoeconomics regions, including densely populated cities and rural, developing towns and communities. When deployed in coordination with smart crop choices, local farm support, and efficient transportation then the result isn’t just sustainability, but also increasing fresh produce accessibility, optimizing nutritional value, eliminating the use of ‘forever chemicals’, reducing transportation costs, and fostering global environmental benefits.
Imagine Doris, who is …
Deep Learning Image Analysis To Isolate And Characterize Different Stages Of S-Phase In Human Cells, Kevin A. Boyd, Rudranil Mitra, John Santerre, Christopher L. Sansam
Deep Learning Image Analysis To Isolate And Characterize Different Stages Of S-Phase In Human Cells, Kevin A. Boyd, Rudranil Mitra, John Santerre, Christopher L. Sansam
SMU Data Science Review
Abstract. This research used deep learning for image analysis by isolating and characterizing distinct DNA replication patterns in human cells. By leveraging high-resolution microscopy images of multiple cells stained with 5-Ethynyl-2′-deoxyuridine (EdU), a replication marker, this analysis utilized Convolutional Neural Networks (CNNs) to perform image segmentation and to provide robust and reliable classification results. First multiple cells in a field of focus were identified using a pretrained CNN called Cellpose. After identifying the location of each cell in the image a python script was created to crop out each cell into individual .tif files. After careful annotation, a CNN was …
Adaptable Object And Animation System For Game Development, Isaiah Turner
Adaptable Object And Animation System For Game Development, Isaiah Turner
Masters Theses & Specialist Projects
In contemporary times, video games have swiftly evolved into a prominent medium, excelling in both entertainment and narrative delivery, positioning themselves as significant rivals to traditional forms such as film and theater. The burgeoning popularity of gaming has led to a surge in aspiring game developers seeking to craft their own creations, driven by both commercial aspirations and personal passion. However, a common challenge faced by these individuals involves the considerable time investment required to acquire essential skills and establish a foundational framework for their projects. Accessible game development engines that offer a diverse range of fundamental features play a …
Triple Helix: Ai-Artist-Audience Collaboration In A Performative Art Experience, Xuedan Zou
Triple Helix: Ai-Artist-Audience Collaboration In A Performative Art Experience, Xuedan Zou
Dartmouth College Master’s Theses
Imagine an art exhibition that morphs its content according to the audience’s experience like a chameleon, reflecting the audience’s mind and culture and turning the artist’s exhibition into the viewer’s. But when the viewers leave, the work fades back to the creator’s original work and waits for the next audience. In this project, my team introduced an interactive exhibition called "Triple Helix," where audience members were provided the opportunity to alter the artworks created by the artist, thus imbuing them with their own perspectives. This interactive exhibition was held at three physical-locations and online, and a comprehensive user study was …
Data-Centric Image Super-Resolution In Magnetic Resonance Imaging: Challenges And Opportunities, Mamata Shrestha
Data-Centric Image Super-Resolution In Magnetic Resonance Imaging: Challenges And Opportunities, Mamata Shrestha
Graduate Theses and Dissertations
Super-resolution has emerged as a crucial research topic in the field of Magnetic Resonance Imaging (MRI) where it plays an important role in understanding and analysis of complex, qualitative, and quantitative characteristics of tissues at high resolutions. Deep learning techniques have been successful in achieving state-of-the-art results for super-resolution. These deep learning-based methods heavily rely on a substantial amount of data. Additionally, they require a pair of low-resolution and high-resolution images for supervised training which is often unavailable. Particularly in MRI super-resolution, it is often impossible to have low-resolution and high-resolution training image pairs. To overcome this, existing methods for …
Enhanced Privacy-Enabled Face Recognition Using Κ-Identity Optimization, Ryan Karl
Enhanced Privacy-Enabled Face Recognition Using Κ-Identity Optimization, Ryan Karl
Department of Electrical and Computer Engineering: Dissertations, Theses, and Student Research
Facial recognition is becoming more and more prevalent in the daily lives of the common person. Law enforcement utilizes facial recognition to find and track suspects. The newest smartphones have the ability to unlock using the user's face. Some door locks utilize facial recognition to allow correct users to enter restricted spaces. The list of applications that use facial recognition will only increase as hardware becomes more cost-effective and more computationally powerful. As this technology becomes more prevalent in our lives, it is important to understand and protect the data provided to these companies. Any data transmitted should be encrypted …
Mrim: Lightweight Saliency-Based Mixed-Resolution Imaging For Low-Power Pervasive Vision, Jiyan Wu, Vithurson Subasharan, Minh Anh Tuan Tran, Kasun Pramuditha Gamlath, Archan Misra
Mrim: Lightweight Saliency-Based Mixed-Resolution Imaging For Low-Power Pervasive Vision, Jiyan Wu, Vithurson Subasharan, Minh Anh Tuan Tran, Kasun Pramuditha Gamlath, Archan Misra
Research Collection School Of Computing and Information Systems
While many pervasive computing applications increasingly utilize real-time context extracted from a vision sensing infrastructure, the high energy overhead of DNN-based vision sensing pipelines remains a challenge for sustainable in-the-wild deployment. One common approach to reducing such energy overheads is the capture and transmission of lower-resolution images to an edge node (where the DNN inferencing task is executed), but this results in an accuracy-vs-energy tradeoff, as the DNN inference accuracy typically degrades with a drop in resolution. In this work, we introduce MRIM, a simple but effective framework to tackle this tradeoff. Under MRIM, the vision sensor platform first executes …
Self-Supervised Pseudo Multi-Class Pre-Training For Unsupervised Anomaly Detection And Segmentation In Medical Images, Yu Tian, Fengbei Liu, Guansong Pang, Yuanhong Chen, Yuyuan Liu, Johan W. Verjans, Rajvinder Singh, Gustavo Carneiro
Self-Supervised Pseudo Multi-Class Pre-Training For Unsupervised Anomaly Detection And Segmentation In Medical Images, Yu Tian, Fengbei Liu, Guansong Pang, Yuanhong Chen, Yuyuan Liu, Johan W. Verjans, Rajvinder Singh, Gustavo Carneiro
Research Collection School Of Computing and Information Systems
Unsupervised anomaly detection (UAD) methods are trained with normal (or healthy) images only, but during testing, they are able to classify normal and abnormal (or disease) images. UAD is an important medical image analysis (MIA) method to be applied in disease screening problems because the training sets available for those problems usually contain only normal images. However, the exclusive reliance on normal images may result in the learning of ineffective low-dimensional image representations that are not sensitive enough to detect and segment unseen abnormal lesions of varying size, appearance, and shape. Pre-training UAD methods with self-supervised learning, based on computer …
Graph Contrastive Learning With Stable And Scalable Spectral Encoding, Deyu Bo, Yuan Fang, Yang Liu, Chuan Shi
Graph Contrastive Learning With Stable And Scalable Spectral Encoding, Deyu Bo, Yuan Fang, Yang Liu, Chuan Shi
Research Collection School Of Computing and Information Systems
Graph contrastive learning (GCL) aims to learn representations by capturing the agreements between different graph views. Traditional GCL methods generate views in the spatial domain, but it has been recently discovered that the spectral domain also plays a vital role in complementing spatial views. However, existing spectral-based graph views either ignore the eigenvectors that encode valuable positional information, or suffer from high complexity when trying to address the instability of spectral features. To tackle these challenges, we first design an informative, stable, and scalable spectral encoder, termed EigenMLP, to learn effective representations from the spectral features. Theoretically, EigenMLP is invariant …
Clueless: Revolutionizing Sustainable Fashion And Combating Overconsumption, Tanya Ravichandran
Clueless: Revolutionizing Sustainable Fashion And Combating Overconsumption, Tanya Ravichandran
Graphic Communication
“Clueless” revolutionizes sustainable fashion by combating wardrobe overconsumption and the industry’s carbon footprint, using AI to suggest personalized outfits from existing wardrobes tailored to weather and wear history. It enhances user engagement through features like outfit ‘shuffle’ and provides insights into wardrobe utilization and carbon impact.
It’s more than an app; it’s a step towards a greener wardrobe and a healthier planet.
Leveraging Artificial Intelligence For Team Cognition In Human-Ai Teams, Beau Schelble
Leveraging Artificial Intelligence For Team Cognition In Human-Ai Teams, Beau Schelble
All Dissertations
Advances in artificial intelligence (AI) technologies have enabled AI to be applied across a wide variety of new fields like cryptography, art, and data analysis. Several of these fields are social in nature, including decision-making and teaming, which introduces a new set of challenges for AI research. While each of these fields has its unique challenges, the area of human-AI teaming is beset with many that center around the expectations and abilities of AI teammates. One such challenge is understanding team cognition in these human-AI teams and AI teammates' ability to contribute towards, support, and encourage it. Team cognition is …
Mermaid: A Dataset And Framework For Multimodal Meme Semantic Understanding, Shaun Toh, Adriel Kuek, Wen Haw Chong, Roy Ka Wei Lee
Mermaid: A Dataset And Framework For Multimodal Meme Semantic Understanding, Shaun Toh, Adriel Kuek, Wen Haw Chong, Roy Ka Wei Lee
Research Collection School Of Computing and Information Systems
Memes are widely used to convey cultural and societal issues and have a significant impact on public opinion. However, little work has been done on understanding and explaining the semantics expressed in multimodal memes. To fill this research gap, we introduce MERMAID, a dataset consisting of 3,633 memes annotated with their entities and relations, and propose a novel MERF pipeline that extracts entities and their relationships in memes. Our framework combines state-of-the-art techniques from natural language processing and computer vision to extract text and image features and infer relationships between entities in memes. We evaluate the proposed framework on a …
Video Sentiment Analysis For Child Safety, Yee Sen Tan, Nicole Anne Huiying Teo, Ezekiel En Zhe Ghe, Jolie Zhi Yi Fong, Zhaoxia Wang
Video Sentiment Analysis For Child Safety, Yee Sen Tan, Nicole Anne Huiying Teo, Ezekiel En Zhe Ghe, Jolie Zhi Yi Fong, Zhaoxia Wang
Research Collection School Of Computing and Information Systems
The proliferation of online video content underscores the critical need for effective sentiment analysis, particularly in safeguarding children from potentially harmful material. This research addresses this concern by presenting a multimodal analysis method for assessing video sentiment, categorizing it as either positive (child-friendly) or negative (potentially harmful). This method leverages three key components: text analysis, facial expression analysis, and audio analysis, including music mood analysis, resulting in a comprehensive sentiment assessment. Our evaluation results validate the effectiveness of this approach, making significant contributions to the field of video sentiment analysis and bolstering child safety measures. This research serves as a …
Depwignn: A Depth-Wise Graph Neural Network For Multi-Hop Spatial Reasoning In Text, Shuaiyi Li, Yang Deng, Wai Lam
Depwignn: A Depth-Wise Graph Neural Network For Multi-Hop Spatial Reasoning In Text, Shuaiyi Li, Yang Deng, Wai Lam
Research Collection School Of Computing and Information Systems
Spatial reasoning in text plays a crucial role in various real-world applications. Existing approaches for spatial reasoning typically infer spatial relations from pure text, which overlook the gap between natural language and symbolic structures. Graph neural networks (GNNs) have showcased exceptional proficiency in inducing and aggregating symbolic structures. However, classical GNNs face challenges in handling multi-hop spatial reasoning due to the over-smoothing issue, i.e., the performance decreases substantially as the number of graph layers increases. To cope with these challenges, we propose a novel Depth-Wise Graph Neural Network (DepWiGNN). Specifically, we design a novel node memory scheme and aggregate the …
Unifying Text, Tables, And Images For Multimodal Question Answering, Haohao Luo, Ying Shen, Yang Deng
Unifying Text, Tables, And Images For Multimodal Question Answering, Haohao Luo, Ying Shen, Yang Deng
Research Collection School Of Computing and Information Systems
Multimodal question answering (MMQA), which aims to derive the answer from multiple knowledge modalities (e.g., text, tables, and images), has received increasing attention due to its board applications. Current approaches to MMQA often rely on single-modal or bi-modal QA models, which limits their ability to effectively integrate information across all modalities and leverage the power of pre-trained language models. To address these limitations, we propose a novel framework called UniMMQA, which unifies three different input modalities into a text-to-text format by employing position-enhanced table linearization and diversified image captioning techniques. Additionally, we enhance cross-modal reasoning by incorporating a multimodal rationale …
Developing Detection And Mapping Of Roads Within Various Forms Of Media Using Opencv, Jordan C. Lyle
Developing Detection And Mapping Of Roads Within Various Forms Of Media Using Opencv, Jordan C. Lyle
Computer Science and Computer Engineering Undergraduate Honors Theses
OpenCV, and Computer Vision in general, has been a Computer Science topic that has interested me for a long time while completing my Bachelor’s degree at the University of Arkansas. As a result of this, I ended up choosing to utilize OpenCV in order to complete the task of detecting road-lines and mapping roads when given a wide variety of images. The purpose of my Honors research and this thesis is to detail the process of creating an algorithm to detect the road-lines such that the results are effective and instantaneous, as well as detail how Computer Vision can be …
The Propagation And Execution Of Malware In Images, Piper Hall
The Propagation And Execution Of Malware In Images, Piper Hall
Cybersecurity Undergraduate Research Showcase
Malware has become increasingly prolific and severe in its consequences as information systems mature and users become more reliant on computing in their daily lives. As cybercrime becomes more complex in its strategies, an often-overlooked manner of propagation is through images. In recent years, several high-profile vulnerabilities in image libraries have opened the door for threat actors to steal money and information from unsuspecting users. This paper will explore the mechanisms by which these exploits function and how they can be avoided.
Performative Mixing For Immersive Audio, Brian A. Elizondo
Performative Mixing For Immersive Audio, Brian A. Elizondo
LSU Doctoral Dissertations
Immersive multichannel audio can be produced with specialized setups of loudspeakers, often surrounding the audience. These setups can feature as few as four loudspeakers or more than 300. Performative mixing in these environments requires a bespoke solution offering intuitive gestural control. Beyond the usual faders for gain control, advancements in multichannel sound demand interfaces capable of quickly positioning sounds between channels. The Quad Cartesian Positioner is such a solution in the form of a Eurorack module for surround mixing for use in live or studio performances.
Diffusion/mixing methods for live multichannel immersive music often rely on the repurposing of hardware …