Open Access. Powered by Scholars. Published by Universities.®
Artificial Intelligence and Robotics Commons™
Open Access. Powered by Scholars. Published by Universities.®
- Institution
-
- Singapore Management University (176)
- Technological University Dublin (28)
- Old Dominion University (20)
- University of Arkansas, Fayetteville (20)
- University of Dayton (16)
-
- City University of New York (CUNY) (14)
- California Polytechnic State University, San Luis Obispo (11)
- San Jose State University (11)
- Dartmouth College (8)
- University of Malaya (8)
- Clemson University (7)
- Embry-Riddle Aeronautical University (7)
- University of Texas at Arlington (6)
- University of Nebraska - Lincoln (5)
- Rochester Institute of Technology (4)
- University of Kentucky (4)
- Central Washington University (3)
- Michigan Technological University (3)
- Montclair State University (3)
- University of Denver (3)
- University of New Mexico (3)
- California State University, San Bernardino (2)
- Dakota State University (2)
- Fort Hays State University (2)
- Georgia Southern University (2)
- Illinois Math and Science Academy (2)
- LSU New Orleans (2)
- Missouri State University (2)
- New Jersey Institute of Technology (2)
- Southern Adventist University (2)
- Keyword
-
- Artificial intelligence (18)
- Computer vision (17)
- Machine Learning (14)
- Machine learning (13)
- Deep learning (12)
-
- Artificial Intelligence (11)
- AI (8)
- Robotics (7)
- Virtual reality (7)
- Visualization (7)
- Augmented reality (6)
- Eye tracking (6)
- HCI (6)
- Image classification (6)
- Automation (5)
- Codes (5)
- Deep Learning (5)
- Feature extraction (5)
- Mental workload (5)
- Personality (5)
- Reinforcement learning (5)
- Accessibility (4)
- Artificial Intelligence (AI) (4)
- Classification (4)
- Computer Science (4)
- Computer Vision (4)
- Human Computer Interaction (4)
- Large Language Models (4)
- Pattern recognition (4)
- Semantics (4)
- Publication Year
- Publication
-
- Research Collection School Of Computing and Information Systems (171)
- H-Workload 2017: Models and Applications (Works in Progress) (14)
- Conference papers (12)
- MAICS: The Modern Artificial Intelligence and Cognitive Science Conference (12)
- Graduate Theses and Dissertations (10)
-
- Publications and Research (10)
- Computer Science Faculty Publications (9)
- Computer Science and Computer Engineering Undergraduate Honors Theses (9)
- Master's Theses (7)
- Dartmouth College Master’s Theses (6)
- Master's Projects (6)
- All Dissertations (5)
- Student Works (2020-2029) (5)
- College of Engineering Summer Undergraduate Research Program (4)
- Dissertations and Theses Collection (Open Access) (4)
- Frameless (4)
- SWITCH (4)
- Theses and Dissertations--Computer Science (4)
- All Master's Theses (3)
- Department of Computer Science Faculty Scholarship and Creative Works (3)
- Dissertations, Master's Theses and Master's Reports (3)
- International Journal of Aviation, Aeronautics, and Aerospace (3)
- Student Works (2000-2009) (3)
- Theses and Dissertations (3)
- All Theses (2)
- College of Graduate Studies: Theses & Dissertations (2)
- Computer Science ETDs (2)
- Computer Science Working Papers (2)
- Computer Science and Engineering Dissertations - Archive (2)
- Dartmouth College Ph.D Dissertations (2)
- Publication Type
- File Type
Articles 271 - 300 of 414
Full-Text Articles in Artificial Intelligence and Robotics
Brave New World Reboot: Technology’S Role In Consumer Manipulation And Implications For Privacy And Transparency, Allie Mertensotto
Brave New World Reboot: Technology’S Role In Consumer Manipulation And Implications For Privacy And Transparency, Allie Mertensotto
Marketing Undergraduate Honors Theses
Most consumers are aware that our data is being obtained and collected through the use of our devices we keep in our homes or even on our person throughout the day. But, it is understated how much data is being collected. Conversations you have with your peers – in a close proximity of a device – are being used to tailor advertising. The advertisements you receive on your devices are uniquely catered to your individual person, due to the fact it consistently uses our data to produce efficient and personal ads. On the flip side, our government is also tapping …
Analog Spiking Neural Network Implementing Spike Timing-Dependent Plasticity On 65 Nm Cmos, Luke Vincent
Analog Spiking Neural Network Implementing Spike Timing-Dependent Plasticity On 65 Nm Cmos, Luke Vincent
Graduate Theses and Dissertations
Machine learning is a rapidly accelerating tool and technology used for countless applications in the modern world. There are many digital algorithms to deploy a machine learning program, but the most advanced and well-known algorithm is the artificial neural network (ANN). While ANNs demonstrate impressive reinforcement learning behaviors, they require large power consumption to operate. Therefore, an analog spiking neural network (SNN) implementing spike timing-dependent plasticity is proposed, developed, and tested to demonstrate equivalent learning abilities with fractional power consumption compared to its digital adversary.
Exploring Ai And Multiplayer In Java, Ronni Kurtzhals
Exploring Ai And Multiplayer In Java, Ronni Kurtzhals
Student Academic Conference
I conducted research into three topics: artificial intelligence, package deployment, and multiplayer servers in Java. This research came together to form my project presentation on the implementation of these topics, which I felt accurately demonstrated the various things I have learned from my courses at Moorhead State University. Several resources were consulted throughout the project, including the work of W3Schools and StackOverflow as well as relevant assignments and textbooks from previous classes. I found this project relevant to computer science and information systems for several reasons, such as the AI component and use of SQL data tables; but it was …
Learning And Simulation Algorithms For Constraint Physical Systems, Shuqi Yang
Learning And Simulation Algorithms For Constraint Physical Systems, Shuqi Yang
Dartmouth College Master’s Theses
This thesis explores two computational approaches to learn and simulate complex physical systems exhibiting constraint characteristics. The target applications encompass both solids and fluids. On the solid side, we proposed a new family of data-driven simulators to predict the behaviors of an unknown physical system by learning its underpinning constraints. We devised a neural projection operator facilitated by an embedded recursive neural network to interactively enforce the learned underpinning constraints and to predict its various physical behaviors. Our method can automatically uncover a broad range of constraints from observation point data, such as length, angle, bending, collision, boundary effects, and …
Adversarial Meta Sampling For Multilingual Low-Resource Speech Recognition, Yubei Xiao, Ke Gong, Pan Zhou, Guolin Zheng, Xiaodan Liang, Liang Lin
Adversarial Meta Sampling For Multilingual Low-Resource Speech Recognition, Yubei Xiao, Ke Gong, Pan Zhou, Guolin Zheng, Xiaodan Liang, Liang Lin
Research Collection School Of Computing and Information Systems
Human doctors with well-structured medical knowledge can diagnose a disease merely via a few conversations with patients about symptoms. In contrast, existing knowledgegrounded dialogue systems often require a large number of dialogue instances to learn as they fail to capture the correlations between different diseases and neglect the diagnostic experience shared among them. To address this issue, we propose a more natural and practical paradigm, i.e., low-resource medical dialogue generation, which can transfer the diagnostic experience from source diseases to target ones with a handful of data for adaptation. It is capitalized on a commonsense knowledge graph to characterize the …
3d Dental Biometrics: Automatic Pose-Invariant Dental Arch Extraction And Matching, Xin Zhong, Zhiyuan Zhang
3d Dental Biometrics: Automatic Pose-Invariant Dental Arch Extraction And Matching, Xin Zhong, Zhiyuan Zhang
Research Collection School Of Computing and Information Systems
A novel automatic pose-invariant dental arch extraction and matching framework is developed for 3D dental identification using laser-scanned dental plasters. In our previous attempt [1-5], 3D point-based algorithms have been developed and they have shown a few advantages over existing 2D dental identifications. This study is a continuous effort in developing arch-based algorithms to extract and match dental arch feature in an automatic and pose-invariant way. As best as we know, this is the first attempt at automatic dental arch extraction and matching for 3D dental identification. A Radial Ray Algorithm (RRA) is proposed by projecting dental arch shape from …
Light Field Compression And Manipulation Via Residual Convolutional Neural Network, Eisa Hedayati
Light Field Compression And Manipulation Via Residual Convolutional Neural Network, Eisa Hedayati
Dissertations, Master's Theses and Master's Reports
Light field (LF) imaging has gained significant attention due to its recent success in microscopy, 3-dimensional (3D) displaying and rendering, augmented and virtual reality usage. Postprocessing of LF enables us to extract more information from a scene compared to traditional cameras. However, the use of LF is still a research novelty because of the current limitations in capturing high-resolution LF in all of its four dimensions. While researchers are actively improving methods of capturing high-resolution LF's, using simulation, it is possible to explore a high-quality captured LF's properties. The immediate concerns following the LF capture are its storage and processing …
"Who Can Help Me?'': Knowledge Infused Matching Of Support Seekers And Support Providers During Covid-19 On Reddit, Manas Gaur, Kaushik Roy, Aditya Sharma, Biplav Srivastava, Amit Sheth
"Who Can Help Me?'': Knowledge Infused Matching Of Support Seekers And Support Providers During Covid-19 On Reddit, Manas Gaur, Kaushik Roy, Aditya Sharma, Biplav Srivastava, Amit Sheth
Publications
During the ongoing COVID-19 crisis, subreddits on Reddit, such as r/Coronavirus saw a rapid growth in user's requests for help (support seekers - SSs) including individuals with varying professions and experiences with diverse perspectives on care (support providers - SPs). Currently, knowledgeable human moderators match an SS with a user with relevant experience, i.e, an SP on these subreddits. This unscalable process defers timely care. We present a medical knowledge-infused approach to efficient matching of SS and SPs validated by experts for the users affected by anxiety and depression, in the context of with COVID-19. After matching, each SP to …
Converting Optical Videos To Infrared Videos Using Attention Gan And Its Impact On Target Detection And Classification Performance, Mohammad Shahab Uddin, Reshad Hoque, Kazi Aminul Islam, Chiman Kwan, David Gribben, Jiang Li
Converting Optical Videos To Infrared Videos Using Attention Gan And Its Impact On Target Detection And Classification Performance, Mohammad Shahab Uddin, Reshad Hoque, Kazi Aminul Islam, Chiman Kwan, David Gribben, Jiang Li
Electrical & Computer Engineering Faculty Publications
To apply powerful deep-learning-based algorithms for object detection and classification in infrared videos, it is necessary to have more training data in order to build high-performance models. However, in many surveillance applications, one can have a lot more optical videos than infrared videos. This lack of IR video datasets can be mitigated if optical-to-infrared video conversion is possible. In this paper, we present a new approach for converting optical videos to infrared videos using deep learning. The basic idea is to focus on target areas using attention generative adversarial network (attention GAN), which will preserve the fidelity of target areas. …
Detecting Surface Interactions Via A Wearable Microphone To Improve Augmented Reality Text Entry, R. Habibi
Detecting Surface Interactions Via A Wearable Microphone To Improve Augmented Reality Text Entry, R. Habibi
Dissertations, Master's Theses and Master's Reports
This thesis investigates whether we can detect and distinguish between surface interaction events such as tapping or swiping using a wearable mic from a surface. Also, what are the advantages of new text entry methods such as tapping with two fingers simultaneously to enter capital letters and punctuation? For this purpose, we conducted a remote study to collect audio and video of three different ways people might interact with a surface. We also built a CNN classifier to detect taps. Our results show that we can detect and distinguish between surface interaction events such as tap or swipe via a …
Virtual Tutor Personality In Computer Assisted Language Learning, Johanna Dobbriner, Cathy Ennis, Robert J. Ross
Virtual Tutor Personality In Computer Assisted Language Learning, Johanna Dobbriner, Cathy Ennis, Robert J. Ross
Conference papers
The use of intelligent virtual agents in language learning has increased in recent years. Studies into several aspects of personalisation aiming to increase user engagement are an ongoing research topic with avatar personality being one such aspect. As a step towards our development of intelligent virtual avatars, we present two of our initial experiments to explore differences in user interaction with two contrasting avatar personalities -- P1: open-minded, friendly and sociable and P2: closed-off, curt and distant. Each user interacted with a single personality in a video-call setting and gave feedback on the interaction. Our expectations, that P1 would be …
Fractal And Edge-Based Techniques For Kidney Enhancement And Segmentation On Magnetic Resonance Images (Mri), Alaá Rateb Mahmoud Al-Shamasneh
Fractal And Edge-Based Techniques For Kidney Enhancement And Segmentation On Magnetic Resonance Images (Mri), Alaá Rateb Mahmoud Al-Shamasneh
Student Works (2020-2029)
Recently, many rapid developments in digital medical imaging have made further contributions to healthcare systems. However, the segmentation of regions of interest in medical images plays a vital role in assisting doctors in their medical diagnoses and for the early detection of disease. Since health issues related to the kidneys are increasing exponentially, this thesis focused on developing methods for the segmentation of MRI images of the kidney. Kidney images frequently suffer from low contrast, low resolution and noise, and are blur. Hence, it is necessary to enhance the images in order to improve the segmentation. Therefore, the current thesis …
A Study Of Multi-Task And Region-Wise Deep Learning For Food Ingredient Recognition, Jingjing Chen, Bin Zhu, Chong-Wah Ngo, Tat-Seng Chua, Yu-Gang Jiang
A Study Of Multi-Task And Region-Wise Deep Learning For Food Ingredient Recognition, Jingjing Chen, Bin Zhu, Chong-Wah Ngo, Tat-Seng Chua, Yu-Gang Jiang
Research Collection School Of Computing and Information Systems
Food recognition has captured numerous research attention for its importance for health-related applications. The existing approaches mostly focus on the categorization of food according to dish names, while ignoring the underlying ingredient composition. In reality, two dishes with the same name do not necessarily share the exact list of ingredients. Therefore, the dishes under the same food category are not mandatorily equal in nutrition content. Nevertheless, due to limited datasets available with ingredient labels, the problem of ingredient recognition is often overlooked. Furthermore, as the number of ingredients is expected to be much less than the number of food categories, …
Global Context Aware Convolutions For 3d Point Cloud Understanding, Zhiyuan Zhang, Binh-Son Hua, Wei Chen, Yibin Tian, Sai-Kit Yeung
Global Context Aware Convolutions For 3d Point Cloud Understanding, Zhiyuan Zhang, Binh-Son Hua, Wei Chen, Yibin Tian, Sai-Kit Yeung
Research Collection School Of Computing and Information Systems
Recent advances in deep learning for 3D point clouds have shown great promises in scene understanding tasks thanks to the introduction of convolution operators to consume 3D point clouds directly in a neural network. Point cloud data, however, could have arbitrary rotations, especially those acquired from 3D scanning. Recent works show that it is possible to design point cloud convolutions with rotation invariance property, but such methods generally do not perform as well as translation-invariant only convolution. We found that a key reason is that compared to point coordinates, rotation-invariant features consumed by point cloud convolution are not as distinctive. …
Espade: An Efficient And Semantically Secure Shortest Path Discovery For Outsourced Location-Based Services, Bharath K. Samanthula, Divyadharshini Karthikeyan, Boxiang Dong, K. Anitha Kumari
Espade: An Efficient And Semantically Secure Shortest Path Discovery For Outsourced Location-Based Services, Bharath K. Samanthula, Divyadharshini Karthikeyan, Boxiang Dong, K. Anitha Kumari
Department of Computer Science Faculty Scholarship and Creative Works
With the rapid growth of smart devices and technological advancements in tracking geospatial data, the demand for Location-Based Services (LBS) is facing a constant rise in several domains, including military, healthcare and transportation. It is a natural step to migrate LBS to a cloud environment to achieve on-demand scalability and increased resiliency. Nonetheless, outsourcing sensitive location data to a third-party cloud provider raises a host of privacy concerns as the data owners have reduced visibility and control over the outsourced data. In this paper, we consider outsourced LBS where users want to retrieve map directions without disclosing their location information. …
Gesture Enhanced Comprehension Of Ambiguous Human-To-Robot Instructions, Weerakoon Mudiyanselage Dulanga Kaveesha Weerakoon, Vigneshwaran Subbaraju, Nipuni Karumpulli, Minh Anh Tuan Tran, Qianli Xu, U-Xuan Tan, Joo Hwee Lim, Archan Misra
Gesture Enhanced Comprehension Of Ambiguous Human-To-Robot Instructions, Weerakoon Mudiyanselage Dulanga Kaveesha Weerakoon, Vigneshwaran Subbaraju, Nipuni Karumpulli, Minh Anh Tuan Tran, Qianli Xu, U-Xuan Tan, Joo Hwee Lim, Archan Misra
Research Collection School Of Computing and Information Systems
This work demonstrates the feasibility and benefits of using pointing gestures, a naturally-generated additional input modality, to improve the multi-modal comprehension accuracy of human instructions to robotic agents for collaborative tasks.We present M2Gestic, a system that combines neural-based text parsing with a novel knowledge-graph traversal mechanism, over a multi-modal input of vision, natural language text and pointing. Via multiple studies related to a benchmark table top manipulation task, we show that (a) M2Gestic can achieve close-to-human performance in reasoning over unambiguous verbal instructions, and (b) incorporating pointing input (even with its inherent location uncertainty) in M2Gestic results in a significant …
Knowledge Enhanced Neural Fashion Trend Forecasting, Yunshan Ma, Yujuan Ding, Xun Yang, Lizi Liao, Wai Keung Wong, Tat-Seng Chua
Knowledge Enhanced Neural Fashion Trend Forecasting, Yunshan Ma, Yujuan Ding, Xun Yang, Lizi Liao, Wai Keung Wong, Tat-Seng Chua
Research Collection School Of Computing and Information Systems
Fashion trend forecasting is a crucial task for both academia and industry. Although some efforts have been devoted to tackling this challenging task, they only studied limited fashion elements with highly seasonal or simple patterns, which could hardly reveal the real fashion trends. Towards insightful fashion trend forecasting, this work focuses on investigating fine-grained fashion element trends for specific user groups. We first contribute a large-scale fashion trend dataset (FIT) collected from Instagram with extracted time series fashion element records and user information. Furthermore, to effectively model the time series data of fashion elements with rather complex patterns, we propose …
Cognition And Context-Aware Computing: Towards A Situation-Aware System With A Case Study In Aviation, Justin C. Wilson
Cognition And Context-Aware Computing: Towards A Situation-Aware System With A Case Study In Aviation, Justin C. Wilson
Computer Science and Engineering Theses and Dissertations
In aviation, flight instructors seek to comprehend the intent and awareness of their students. With this awareness, derived from in-flight observation and post-flight examination, a human instructor can infer the internal contexts of their student aviators as they perform. It is this understanding that is fundamental for evaluating student development. Further, a well-understood construct for describing the state of knowledge about a dynamic environment is known as situational awareness (SA). Often pilot error is associated with SA [80], and it is fundamental to flight safety and mission execution. If these contexts can be automatically inferred, instructors and students can more …
A 3d Image-Guided System To Improve Myocardial Revascularization Decision-Making For Patients With Coronary Artery Disease, Haipeng Tang
A 3d Image-Guided System To Improve Myocardial Revascularization Decision-Making For Patients With Coronary Artery Disease, Haipeng Tang
Dissertations
OBJECTIVES. Coronary artery disease (CAD) is the most common type of heart disease and kills over 360,000 people a year in the United States. Myocardial revascularization (MR) is a standard interventional treatment for patients with stable CAD. Fluoroscopy angiography is real-time anatomical imaging and routinely used to guide MR by visually estimating the percent stenosis of coronary arteries. However, a lot of patients do not benefit from the anatomical information-guided MR without functional testing. Single-photon emission computed tomography (SPECT) myocardial perfusion imaging (MPI) is a widely used functional testing for CAD evaluation but limits to the absence of anatomical information. …
Unsupervised Monocular Depth Estimation With Multi-Scale Structural Similarity Powered Loss Function, Kohan Ali
Unsupervised Monocular Depth Estimation With Multi-Scale Structural Similarity Powered Loss Function, Kohan Ali
Student Works (2020-2029)
Depth Estimation refers to a set of techniques and algorithms that aim to obtain a representation of spatial information of a scene. Nowadays specific hardware such as sensors, radars and multiple-view-recording cameras are being used in order to acquire depth data of a scene. Modern approaches use deep learning to address this task by trying to learn depth information in a supervised manner. However, this approach requires a large amount ground-truth data for a particular scene so that a model can be trained successfully. Also preparing ground-truth data for a range of environments is a challenging and expensive task to …
Self-Trained Deep Ordinal Regression For End-To-End Video Anomaly Detection, Guansong Pang, Cheng Yan, Chunhua Shen, Anton Van Den Hengel, Xiao Bai
Self-Trained Deep Ordinal Regression For End-To-End Video Anomaly Detection, Guansong Pang, Cheng Yan, Chunhua Shen, Anton Van Den Hengel, Xiao Bai
Research Collection School Of Computing and Information Systems
Depression is among the most prevalent mental disorders, affecting millions of people of all ages globally. Machine learning techniques have shown effective in enabling automated detection and prediction of depression for early intervention and treatment. However, they are challenged by the relative scarcity of instances of depression in the data. In this work we introduce a novel deep multi-task recurrent neural network to tackle this challenge, in which depression classification is jointly optimized with two auxiliary tasks, namely one-class metric learning and anomaly ranking. The auxiliary tasks introduce an inductive bias that improves the classification model’s generalizability on small depression …
Visual Commonsense R-Cnn, Tan Wang, Jianqiang Huang, Hanwang Zhang, Qianru Sun
Visual Commonsense R-Cnn, Tan Wang, Jianqiang Huang, Hanwang Zhang, Qianru Sun
Research Collection School Of Computing and Information Systems
We present a novel unsupervised feature representation learning method, Visual Commonsense Region-based Convolutional Neural Network (VC R-CNN), to serve as an improved visual region encoder for high-level tasks such as captioning and VQA. Given a set of detected object regions in an image (e.g., using Faster R-CNN), like any other unsupervised feature learning methods (e.g., word2vec), the proxy training objective of VC R-CNN is to predict the contextual objects of a region. However, they are fundamentally different: the prediction of VC R-CNN is by using causal intervention: P(Y|do(X)), while others are by using the conventional likelihood: P(Y|X). This is also …
Vision And Sensor-Based Signer-Independent Framework For Arabic Sign Language Recognition, Al-Shamayleh Ahmad Sami Abd Alkareem
Vision And Sensor-Based Signer-Independent Framework For Arabic Sign Language Recognition, Al-Shamayleh Ahmad Sami Abd Alkareem
Student Works (2020-2029)
Hearing and speech-impairment disability is widespread throughout the world. At present, 15 million people have this disability in the Arab world, and about 86% of them come from low- and middle-income countries. Meanwhile, sign language (SL) can be classified into standard Arabic sign language (ArSL) and local Arabic sign language (LArSL). ArSL is the formal standard and is the more acceptable SL in the Arab world; it is also considered as the medium of instructions for schools and universities as well as television news, shows and programmes. With the absence of usable ArSL recognition (ArSLR) platforms, hearing- and speech-impaired people …
Video Synthesis From The Stylegan Latent Space, Lei Zhang
Video Synthesis From The Stylegan Latent Space, Lei Zhang
Master's Projects
Generative models have shown impressive results in generating synthetic images. However, video synthesis is still difficult to achieve, even for these generative models. The best videos that generative models can currently create are a few seconds long, distorted, and low resolution. For this project, I propose and implement a model to synthesize videos at 1024x1024x32 resolution that include human facial expressions by using static images generated from a Generative Adversarial Network trained on the human facial images. To the best of my knowledge, this is the first work that generates realistic videos that are larger than 256x256 resolution from single …
Speech Processing In Computer Vision Applications, Nicholas Waterworth
Speech Processing In Computer Vision Applications, Nicholas Waterworth
Computer Science and Computer Engineering Undergraduate Honors Theses
Deep learning has been recently proven to be a viable asset in determining features in the field of Speech Analysis. Deep learning methods like Convolutional Neural Networks facilitate the expansion of specific feature information in waveforms, allowing networks to create more feature dense representations of data. Our work attempts to address the problem of re-creating a face given a speaker's voice and speaker identification using deep learning methods. In this work, we first review the fundamental background in speech processing and its related applications. Then we introduce novel deep learning-based methods to speech feature analysis. Finally, we will present our …
Connecting The Dots For People With Autism: A Data-Driven Approach To Designing And Evaluating A Global Filter, Viseth Sean
Connecting The Dots For People With Autism: A Data-Driven Approach To Designing And Evaluating A Global Filter, Viseth Sean
Computational and Data Sciences (PhD) Dissertations
"Social communication is the use of language in social contexts. It encompasses social interaction, social cognition, pragmatics, and language processing” [3]. One presumed prerequisite of social communication is visual attention–the focus of this work. “Visual attention is a process that directs a tiny fraction of the information arriving at primary visual cortex to high-level centers involved in visual working memory and pattern recognition” [7]. This process involves the integration of two streams: the global and local streams; the global stream rapidly processes the scene, and the local stream processes details. This integration is important to social communication in that attending …
Storage Management Strategy In Mobile Phones For Photo Crowdsensing, En Wang, Zhengdao Qu, Xinyao Liang, Xiangyu Meng, Yongjian Yang, Dawei Li, Weibin Meng
Storage Management Strategy In Mobile Phones For Photo Crowdsensing, En Wang, Zhengdao Qu, Xinyao Liang, Xiangyu Meng, Yongjian Yang, Dawei Li, Weibin Meng
Department of Computer Science Faculty Scholarship and Creative Works
In mobile crowdsensing, some users jointly finish a sensing task through the sensors equipped in their intelligent terminals. In particular, the photo crowdsensing based on Mobile Edge Computing (MEC) collects pictures for some specific targets or events and uploads them to nearby edge servers, which leads to richer data content and more efficient data storage compared with the common mobile crowdsensing; hence, it has attracted an important amount of attention recently. However, the mobile users prefer uploading the photos through Wifi APs (PoIs) rather than cellular networks. Therefore, photos stored in mobile phones are exchanged among users, in order to …
Zero-Shot Ingredient Recognition By Multi-Relational Graph Convolutional Network, Jingjing Chen, Liangming Pan, Zhipeng Wei, Xiang Wang, Chong-Wah Ngo, Tat-Seng Chua
Zero-Shot Ingredient Recognition By Multi-Relational Graph Convolutional Network, Jingjing Chen, Liangming Pan, Zhipeng Wei, Xiang Wang, Chong-Wah Ngo, Tat-Seng Chua
Research Collection School Of Computing and Information Systems
Recognizing ingredients for a given dish image is at the core of automatic dietary assessment, attracting increasing attention from both industry and academia. Nevertheless, the task is challenging due to the difficulty of collecting and labeling sufficient training data. On one hand, there are hundred thousands of food ingredients in the world, ranging from the common to rare. Collecting training samples for all of the ingredient categories is difficult. On the other hand, as the ingredient appearances exhibit huge visual variance during the food preparation, it requires to collect the training samples under different cooking and cutting methods for robust …
Gdface: Gated Deformation For Multi-View Face Image Synthesis, Xuemiao Xu, Keke Li, Cheng Xu, Shengfeng He
Gdface: Gated Deformation For Multi-View Face Image Synthesis, Xuemiao Xu, Keke Li, Cheng Xu, Shengfeng He
Research Collection School Of Computing and Information Systems
Photorealistic multi-view face synthesis from a single image is an important but challenging problem. Existing methods mainly learn a texture mapping model from the source face to the target face. However, they fail to consider the internal deformation caused by the change of poses, leading to the unsatisfactory synthesized results for large pose variations. In this paper, we propose a Gated Deformable Face Synthesis Network to model the deformation of faces that aids the synthesis of the target face image. Specifically, we propose a dual network that consists of two modules. The first module estimates the deformation of two views …
Accessibility Of Deepfakes, Andrew L. Collings
Accessibility Of Deepfakes, Andrew L. Collings
Cybersecurity Undergraduate Research Showcase
The danger posed by falsified media, commonly referred to as deepfakes, has been well researched and documented. The software Faceswap to was used to swap the faces of two politician (Joe Biden and Donald Trump). The testing was performed using an affordable consumer GPU (an AMD Radeon RX 570) over 100,000 iterations. The process and results for the two attempts with the best results (and largest differences) were recorded. The result was ultimately unconvincing, while the software was able to recreate the facial structure the lighting and skin tone did not blend at all.