Open Access. Powered by Scholars. Published by Universities.®
Artificial Intelligence and Robotics Commons™
Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Engineering (52)
- Data Science (34)
- Computer Engineering (24)
- Electrical and Computer Engineering (24)
- Medicine and Health Sciences (22)
-
- Databases and Information Systems (15)
- Software Engineering (15)
- Other Computer Sciences (14)
- Statistics and Probability (13)
- Theory and Algorithms (13)
- Numerical Analysis and Scientific Computing (12)
- Medical Specialties (10)
- Statistical Models (9)
- Applied Mathematics (8)
- Life Sciences (8)
- Social and Behavioral Sciences (8)
- Analytical, Diagnostic and Therapeutic Techniques and Equipment (7)
- Information Security (7)
- Signal Processing (6)
- Applied Statistics (5)
- Graphics and Human Computer Interfaces (5)
- Mechanical Engineering (5)
- OS and Networks (5)
- Programming Languages and Compilers (5)
- Bioinformatics (4)
- Biomedical Engineering and Bioengineering (4)
- Computational Engineering (4)
- Institution
-
- Singapore Management University (15)
- San Jose State University (14)
- California Polytechnic State University, San Luis Obispo (13)
- City University of New York (CUNY) (12)
- University of Kentucky (10)
-
- Southern Methodist University (8)
- University of South Florida (6)
- Clemson University (5)
- Technological University Dublin (5)
- Thomas Jefferson University (5)
- Old Dominion University (4)
- Purdue University (4)
- The University of Southern Mississippi (4)
- University of Texas at El Paso (4)
- West Virginia University (4)
- Western University (4)
- California State University, San Bernardino (3)
- Kennesaw State University (3)
- Loyola University Chicago (3)
- University of Arkansas, Fayetteville (3)
- University of Nevada, Las Vegas (3)
- Virginia Commonwealth University (3)
- Wayne State University (3)
- Bucknell University (2)
- Dartmouth College (2)
- East Tennessee State University (2)
- Florida Institute of Technology (2)
- Grand Valley State University (2)
- LSU New Orleans (2)
- Louisiana State University (2)
- Publication Year
- Publication
-
- Research Collection School Of Computing and Information Systems (14)
- Master's Projects (13)
- Master's Theses (12)
- Dissertations, Theses, and Capstone Projects (8)
- Theses and Dissertations (8)
-
- SMU Data Science Review (7)
- Theses and Dissertations--Computer Science (6)
- USF Tampa Graduate Theses and Dissertations (6)
- Electrical and Computer Engineering Publications (4)
- Graduate Theses, Dissertations, and Problem Reports (ETD) (4)
- Open Access Theses & Dissertations (4)
- All Dissertations (3)
- Computer Science: Faculty Publications and Other Works (3)
- Conference papers (3)
- Dissertations (3)
- Electrical & Computer Engineering Theses & Dissertations (3)
- Electronic Theses, Projects, and Dissertations (3)
- Honors Theses (3)
- The Summer Undergraduate Research Fellowship (SURF) Symposium (3)
- UNLV Theses, Dissertations, Professional Papers, and Capstones (3)
- Wayne State University Dissertations (3)
- All Theses (2)
- College of Engineering Summer Undergraduate Research Program (2)
- Dissertations and Theses (2)
- Doctoral Dissertations (2)
- Electronic Theses and Dissertations (2)
- Faculty Conference Papers and Presentations (2)
- Graduate Studies Theses and Dissertations 2026 (2)
- Graduate Theses and Dissertations (2)
- Honors Scholar Theses (2)
- Publication Type
Articles 91 - 120 of 201
Full-Text Articles in Artificial Intelligence and Robotics
Federated Learning For Protecting Medical Data Privacy, Abhishek Reddy Punreddy
Federated Learning For Protecting Medical Data Privacy, Abhishek Reddy Punreddy
Master's Projects
Deep learning is one of the most advanced machine learning techniques, and its prominence has increased in recent years. Language processing, predictions in medical research and pattern recognition are few of the numerous fields in which it is widely utilized. Numerous modern medical applications benefit greatly from the implementation of machine learning (ML) models and the disruptive innovations in the entire modern health care system. It is extensively used for constructing accurate and robust statistical models from large volumes of medical data collected from a variety of sources in contemporary healthcare systems [1]. Due to privacy concerns that restrict access …
Liquid Tab, Nathan Hulet
Liquid Tab, Nathan Hulet
Williams Honors College, Honors Research Projects
Guitar transcription is a complex task requiring significant time, skill, and musical knowledge to achieve accurate results. Since most music is recorded and processed digitally, it would seem like many tools to digitally analyze and transcribe the audio would be available. However, the problem of automatic transcription presents many more difficulties than are initially evident. There are multiple ways to play a guitar, many diverse styles of playing, and every guitar sounds different. These problems become even more difficult considering the varying qualities of recordings and levels of background noise.
Machine learning has proven itself to be a flexible tool …
Breast Density Classification Using Deep Learning, Conrad Thomas Testagrose
Breast Density Classification Using Deep Learning, Conrad Thomas Testagrose
UNF Graduate Theses and Dissertations
Breast density screenings are an accepted means to determine a patient's predisposed risk of breast cancer development. Although the direct correlation is not fully understood, breast cancer risk increases with higher levels of mammographic breast density. Radiologists visually assess a patient's breast density using mammogram images and assign a density score based on four breast density categories outlined by the Breast Imaging and Reporting Data Systems (BI-RADS). There have been efforts to develop automated tools that assist radiologists with increasing workloads and to help reduce the intra- and inter-rater variability between radiologists. In this thesis, I explored two deep-learning-based approaches …
Wildfire Spread Prediction Using Attention Mechanisms In U-Net, Kamen Haresh Shah, Kamen Haresh Shah
Wildfire Spread Prediction Using Attention Mechanisms In U-Net, Kamen Haresh Shah, Kamen Haresh Shah
Master's Theses
An investigation into using attention mechanisms for better feature extraction in wildfire spread prediction models. This research examines the U-net architecture to achieve image segmentation, a process that partitions images by classifying pixels into one of two classes. The deep learning models explored in this research integrate modern deep learning architectures, and techniques used to optimize them. The models are trained on 12 distinct observational variables derived from the Google Earth Engine catalog. Evaluation is conducted with accuracy, Dice coefficient score, ROC-AUC, and F1-score. This research concludes that when augmenting U-net with attention mechanisms, the attention component improves feature suppression …
Overview Of The Clpsych 2022 Shared Task: Capturing Moments Of Change In Longitudinal User Posts, Adam Tsakalidis, Jenny Chim, Iman Munire Bilal, Ayah Zirikly, Dana Atzil-Slonim, Federico Nanni, Philip Resnik, Manas Gaur, Kaushik Roy, Becky Inkster, Jeff Leintz, Maria Liakata
Overview Of The Clpsych 2022 Shared Task: Capturing Moments Of Change In Longitudinal User Posts, Adam Tsakalidis, Jenny Chim, Iman Munire Bilal, Ayah Zirikly, Dana Atzil-Slonim, Federico Nanni, Philip Resnik, Manas Gaur, Kaushik Roy, Becky Inkster, Jeff Leintz, Maria Liakata
Publications
We provide an overview of the CLPsych 2022 Shared Task, which focusses on the automatic identification of Moments of Change in longitudinal posts by individuals on social media and its connection with information regarding mental health . This year's task introduced the notion of longitudinal modelling of the text generated by an individual online over time, along with appropriate temporally sensitive evaluation metrics. The Shared Task consisted of two subtasks: (a) the main task of capturing changes in an individual's mood (drastic changes-`Switches'- and gradual changes -`Escalations'- on the basis of textual content shared online; and subsequently (b) the sub-task …
Towards Understanding The Faults Of Javascript-Based Deep Learning Systems, Lili Quan, Qianyu Guo, Xiaofei Xie, Sen Chen, Xiaohong Li, Yang Liu
Towards Understanding The Faults Of Javascript-Based Deep Learning Systems, Lili Quan, Qianyu Guo, Xiaofei Xie, Sen Chen, Xiaohong Li, Yang Liu
Research Collection School Of Computing and Information Systems
Quality assurance is of great importance for deep learning (DL) systems, especially when they are applied in safety-critical applications. While quality issues of native DL applications have been extensively analyzed, the issues of JavaScript-based DL applications have never been systematically studied. Compared with native DL applications, JavaScript-based DL applications can run on major browsers, making the platform- and device-independent. Specifically, the quality of JavaScript-based DL applications depends on the 3 parts: the application, the third-party DL library used and the underlying DL framework (e.g., TensorFlow.js), called JavaScript-based DL system. In this paper, we conduct the first empirical study on the …
Deep Learning For Coverage-Guided Fuzzing: How Far Are We?, Siqi Li, Xiaofei Xie, Yun Lin, Yuekang Li, Ruitao Feng, Xiaohong Li, Weimin Ge, Jin Song Dong
Deep Learning For Coverage-Guided Fuzzing: How Far Are We?, Siqi Li, Xiaofei Xie, Yun Lin, Yuekang Li, Ruitao Feng, Xiaohong Li, Weimin Ge, Jin Song Dong
Research Collection School Of Computing and Information Systems
Fuzzing is a widely-used software vulnerability discovery technology, many of which are optimized using coverage-feedback. Recently, some techniques propose to train deep learning (DL) models to predict the branch coverage of an arbitrary input owing to its always-available gradients etc. as a guide. Those techniques have proved their success in improving coverage and discovering bugs under different experimental settings. However, DL models, usually as a magic black-box, are notoriously lack of explanation. Moreover, their performance can be sensitive to the collected runtime coverage information for training, indicating potentially unstable performance. In this work, we conduct a systematic empirical study on …
Machine Learning And Scalable Informatics Methods To Predict Disease Status From Multimodal Biomedical Data, Hossein Mohammadian Foroushani
Machine Learning And Scalable Informatics Methods To Predict Disease Status From Multimodal Biomedical Data, Hossein Mohammadian Foroushani
McKelvey School of Engineering Graduate Student Theses & Dissertations
Biological understanding of complex diseases such as stroke and obesity is critical for the advancement of medicine. Further knowledge discovery can provide effective biomarkers to improve disease diagnosis and prognosis, identify driver mutations, predict individual genetic susceptibility for early prevention and effective disease management, and facilitate development of personalized drugs. Stroke is the second leading cause of death and long-term disability in the world. Thus, stroke management is a time-sensitive emergency. The initial hours after stroke onset map the trajectory of subsequent neurologic complications. Cerebral edema develops hours to days after acute ischemic stroke and may result in midline shift …
Deep Active Genetic Learning With Evidential Uncertainty For Agriculture Crops And Lake Water Quality Assessment, Oguz M. Aranay
Deep Active Genetic Learning With Evidential Uncertainty For Agriculture Crops And Lake Water Quality Assessment, Oguz M. Aranay
Legacy Theses & Dissertations (2009 - 2024)
Despite significant advancements in the field of machine learning, there are two issues that still require further exploration. First, how to learn from a small dataset; and second, how to select appropriate features from the data. Although there exist many techniques to address these issues, choosing a combination of the techniques from these two groups is challenging, and worth investigating. To address these concerns, this thesis presents a learning framework that is based on a deep learning model utilizing active learning (with evidential uncertainty as a basis for acquisition function) for the first issue and a genetic algorithm for the …
Applied Deep Learning: Case Studies In Computer Vision And Natural Language Processing, Md Reshad Ul Hoque
Applied Deep Learning: Case Studies In Computer Vision And Natural Language Processing, Md Reshad Ul Hoque
Electrical & Computer Engineering Theses & Dissertations
Deep learning has proved to be successful for many computer vision and natural language processing applications. In this dissertation, three studies have been conducted to show the efficacy of deep learning models for computer vision and natural language processing. In the first study, an efficient deep learning model was proposed for seagrass scar detection in multispectral images which produced robust, accurate scars mappings. In the second study, an arithmetic deep learning model was developed to fuse multi-spectral images collected at different times with different resolutions to generate high-resolution images for downstream tasks including change detection, object detection, and land cover …
Textual Emotion Detection Approaches: A Survey, Mahinda Mahmoud Samy Zidan, Ibrahim Elhenawy, Ahmed R. Abas, Mahmoud Othman
Textual Emotion Detection Approaches: A Survey, Mahinda Mahmoud Samy Zidan, Ibrahim Elhenawy, Ahmed R. Abas, Mahmoud Othman
Future Computing and Informatics Journal
Over the past decades, social media attracted individuals to express their feelings on any topic or item, resulting in an incremental growth in the size of created data. These feelings and unstructured data paved the path for business organizations to gather information and build statistical analysis. Various machine learning and natural language processing-based approaches are used for sentiment and emotion analysis. Moreover, deep learning-based approaches recently gained popularity due to their remarkable performance in text analysis. This paper provides a comprehensive overview of the prominent machine learning models applied in emotion analysis. It explores various emotion analysis taxonomies, in addition …
Machine Learning With Kay, Lasith Niroshan, James Carswell
Machine Learning With Kay, Lasith Niroshan, James Carswell
Conference Papers
Computational power is very important when training Deep Learning (DL) models with large amounts of data (Wooldridge, 2021). Hence, High-Performance Computing (HPC) can be leveraged to reduce computational cost, and the Irish Centre for High-End Computing (ICHEC) provides significant infrastructure and services for research and development to both academia and industry. A portion of ICHEC's HPC system has been allocated for institutional access, and this paper presents a case study of how to use Kay (Ireland's national supercomputer) in the remote sensing domain. Specifically, this study uses clusters of Kay Graphics Processing Units (GPUs) for training DL models to extract …
An Empirical Study On Sampling Approaches For 3d Image Classification Using Deep Learning, Nicholas Michelette
An Empirical Study On Sampling Approaches For 3d Image Classification Using Deep Learning, Nicholas Michelette
Theses and Dissertations
A 3D classification method requires more training data than a 2D image classification method to achieve good performance. These training data usually come in the form of multiple 2D images (e.g., slices in a CT scan) or point clouds (e.g., 3D CAD modeling) for volumetric object representation. The amount of data required to complete this higher dimension problem comes with the cost of requiring more processing time and space. This problem can be mitigated with data size reduction (i.e., sampling). In this thesis, we empirically study and compare the classification performance and deep learning training time of PointNet utilizing uniform …
Training Thinner And Deeper Neural Networks: Jumpstart Regularization, Carles Riera, Camilo Rey, Thiago Serra, Eloi Puertas, Oriol Pujol
Training Thinner And Deeper Neural Networks: Jumpstart Regularization, Carles Riera, Camilo Rey, Thiago Serra, Eloi Puertas, Oriol Pujol
Faculty Conference Papers and Presentations
Neural networks are more expressive when they have multiple layers. In turn, conventional training methods are only successful if the depth does not lead to numerical issues such as exploding or vanishing gradients, which occur less frequently when the layers are sufficiently wide. However, increasing width to attain greater depth entails the use of heavier computational resources and leads to overparameterized models. These subsequent issues have been partially addressed by model compression methods such as quantization and pruning, some of which relying on normalization-based regularization of the loss function to make the effect of most parameters negligible. In this work, …
Weakly-Supervised Tumor Purity Prediction From Frozen H&E Stained Slides, Matthew Brendel, Vanesa Getseva, Majd Al Assaad, Michael Sigouros, Alexandros Sigaras, Troy Kane, Pegah Khosravi, Juan Miguel Mosquera, Olivier Elemento, Iman Hajirasouliha
Weakly-Supervised Tumor Purity Prediction From Frozen H&E Stained Slides, Matthew Brendel, Vanesa Getseva, Majd Al Assaad, Michael Sigouros, Alexandros Sigaras, Troy Kane, Pegah Khosravi, Juan Miguel Mosquera, Olivier Elemento, Iman Hajirasouliha
Publications and Research
Background
Estimating tumor purity is especially important in the age of precision medicine. Purity estimates have been shown to be critical for correction of tumor sequencing results, and higher purity samples allow for more accurate interpretations from next-generation sequencing results. Molecular-based purity estimates using computational approaches require sequencing of tumors, which is both time-consuming and expensive.
Methods
Here we propose an approach, weakly-supervised purity (wsPurity), which can accurately quantify tumor purity within a digitally captured hematoxylin and eosin (H&E) stained histological slide, using several types of cancer from The Cancer Genome Atlas (TCGA) as a proof-of-concept.
Findings
Our model predicts …
Simultaneous Energy Harvesting And Gait Recognition Using Piezoelectric Energy Harvester, Dong Ma, Guohao Lan, Weitao Xu, Mahbub Hassan, Wen Hu
Simultaneous Energy Harvesting And Gait Recognition Using Piezoelectric Energy Harvester, Dong Ma, Guohao Lan, Weitao Xu, Mahbub Hassan, Wen Hu
Research Collection School Of Computing and Information Systems
Piezoelectric energy harvester, which generates electricity from stress or vibrations, is gaining increasing attention as a viable solution to extend battery life in wearables. Recent research further reveals that, besides generating energy, PEH can also serve as a passive sensor to detect human gait power-efficiently because its stress or vibration patterns are significantly influenced by the gait. However, as PEHs are not designed for precise measurement of motion, achievable gait recognition accuracy remains low with conventional classification algorithms. The accuracy deteriorates further when the generated electricity is stored simultaneously. To classify gait reliably while simultaneously storing generated energy, we make …
A Machine Learning And Deep Learning Framework For Binary, Ternary, And Multiclass Emotion Classification Of Covid-19 Vaccine-Related Tweets, Aditya Dubey
Honors Scholar Theses
My research mines public emotion toward the Covid-19 vaccine based on Twitter data collected over the past 6-12 months. This project is centered around building and developing machine learning and deep learning models to perform natural language processing of short-form text, which in our case tweets. These tweets are all vaccine-related tweets and the goal of the classification task is for our models to accurately classify a tweet into one of four emotion groups: Apprehension/Anticipation, Sadness/Anger/Frustration, Joy/Humor/Sarcasm, and Gratitude/Relief. Given this data and the goal of the paper, we aim to answer the following questions: (1) Can a framework be …
Improved Sensor-Based Human Activity Recognition Via Hybrid Convolutional And Recurrent Neural Networks, Sonia Perez-Gamboa
Improved Sensor-Based Human Activity Recognition Via Hybrid Convolutional And Recurrent Neural Networks, Sonia Perez-Gamboa
Electronic Theses, Projects, and Dissertations
Non-intrusive sensor-based human activity recognition is utilized in a spectrum of applications including fitness tracking devices, gaming, health care monitoring, and smartphone applications. Deep learning models such as convolutional neural networks (CNNs) and long short-term memory (LSTMs) recurrent neural networks provide a way to achieve human activity recognition accurately and effectively. This project designed and explored a variety of multi-layer hybrid deep learning architectures which aimed to improve human activity recognition performance by integrating local features and was scale invariant with dependencies of activities. We achieved a 94.7% activity recognition rate on the University of California, Irvine public domain dataset …
Data-Driven Design And Analysis Of Next Generation Mobile Networks For Anomaly Detection And Signal Classification With Fast, Robust And Light Machine Learning, Muhammed Furkan Küçük
Data-Driven Design And Analysis Of Next Generation Mobile Networks For Anomaly Detection And Signal Classification With Fast, Robust And Light Machine Learning, Muhammed Furkan Küçük
USF Tampa Graduate Theses and Dissertations
This research focuses on machine (and deep) learning applications (including clustering,anomaly detection and signal classification) for self-organizing and next generation mobile networks in wireless communications. Specifically, this dissertation document will address the three different topics.
First, in the study titled “Performance analysis of neural network topologies and hyperparameters for deep clustering”, we explore the relationship between the clustering performance and network complexity. Deep learning found its initial footing in supervised applications such as image and voice recognition successes of which were followed by deep generative models across similar domains. In recent years, researchers have proposed creative learning representations to utilize …
Entity Based Sentiment Analysis For Textual Health Advice, Dae Lim Chung
Entity Based Sentiment Analysis For Textual Health Advice, Dae Lim Chung
Computer Science Senior Theses
This work explores entity based sentiment analysis for textual health advice through deep learning. We fine tuned a pretrained BERT model to analyze sentiments across five different predetermined categories which consist of food, medicine, disease, exercise, and vitality for three different sentiments: positive, negative, and neutral. Original set of annotated medical dataset from Dartmouth College’s Persist Lab was used to conduct the experiments. For the aim of tailoring the data for the purpose of entity based sentiment analysis, we explored data transformation techniques to generate optimum training examples. During the experiments, we were able to discover that the wide variety …
Riconv++: Effective Rotation Invariant Convolutions For 3d Point Clouds Deep Learning, Zhiyuan Zhang, Binh-Son Hua, Sai-Kit Yeung
Riconv++: Effective Rotation Invariant Convolutions For 3d Point Clouds Deep Learning, Zhiyuan Zhang, Binh-Son Hua, Sai-Kit Yeung
Research Collection School Of Computing and Information Systems
3D point clouds deep learning is a promising field of research that allows a neural network to learn features of point clouds directly, making it a robust tool for solving 3D scene understanding tasks. While recent works show that point cloud convolutions can be invariant to translation and point permutation, investigations of the rotation invariance property for point cloud convolution has been so far scarce. Some existing methods perform point cloud convolutions with rotation-invariant features, existing methods generally do not perform as well as translation-invariant only counterpart. In this work, we argue that a key reason is that compared to …
Segmentation Of Intracranial Structures From Noncontrast Ct Images With Deep Learning, Evan Porter
Segmentation Of Intracranial Structures From Noncontrast Ct Images With Deep Learning, Evan Porter
Wayne State University Dissertations
Presented in this work is an investigation of the application of artificially intelligent algorithms, namely deep learning, to generate segmentations for the application in functional avoidance radiotherapy treatment planning. Specific applications of deep learning for functional avoidance include generating hippocampus segmentations from computed tomography (CT) images and generating synthetic pulmonary perfusion images from four-dimensional CT (4DCT).A single institution dataset of 390 patients treated with Gamma Knife stereotactic radiosurgery was created. From these patients, the hippocampus was manually segmented on the high-resolution MR image and used for the development of the data processing methodology and model testing. It was determined that …
An Analysis On Adversarial Machine Learning: Methods And Applications, Ali Dabouei
An Analysis On Adversarial Machine Learning: Methods And Applications, Ali Dabouei
Graduate Theses, Dissertations, and Problem Reports (ETD)
Deep learning has witnessed astonishing advancement in the last decade and revolutionized many fields ranging from computer vision to natural language processing. A prominent field of research that enabled such achievements is adversarial learning, investigating the behavior and functionality of a learning model in presence of an adversary. Adversarial learning consists of two major trends. The first trend analyzes the susceptibility of machine learning models to manipulation in the decision-making process and aims to improve the robustness to such manipulations. The second trend exploits adversarial games between components of the model to enhance the learning process. This dissertation aims to …
License Plate Image Quality Enhancement Utilizing Super Resolution Generative Adversarial Networks, Mark Moelter
License Plate Image Quality Enhancement Utilizing Super Resolution Generative Adversarial Networks, Mark Moelter
College of Graduate Studies: Theses & Dissertations
This thesis focuses primarily on enhancing the image quality of blurred license plates through the use of Super-Resolution Generative Adversarial Networks (SRGANs) [1]. We propose a synthetic dataset with SRGAN model to promote blurred image quality enhancement, and allow for model evaluation on a multitude of image input and output size combinations. SRGAN is mainly used for low-resolution image enhancement, but by heavily blurring the input images, the model is tested on its ability to blindly deblur and upsample images to the desired super-resolution (SR) size. The model enhances the image quality to nearly that of the reference images. The …
A Study On Human Face Expressions Using Convolutional Neural Networks And Generative Adversarial Networks, Sriramm Muthyala Sudhakar
A Study On Human Face Expressions Using Convolutional Neural Networks And Generative Adversarial Networks, Sriramm Muthyala Sudhakar
Master's Projects
Human beings express themselves via words, signs, gestures, and facial emotions. Previous research using pre-trained convolutional models had been done by freezing the entire network and running the models without the use of any image processing techniques. In this research, we attempt to enhance the accuracy of many deep CNN architectures like ResNet and Senet, using a variety of different image processing techniques like Image Data Generator, Histogram Equalization, and UnSharpMask. We used FER 2013, which is a dataset containing multiple classes of images. While working on these models, we decided to take things to the next level, and we …
Smart City Management Using Machine Learning Techniques, Mostafa Zaman
Smart City Management Using Machine Learning Techniques, Mostafa Zaman
Theses and Dissertations
In response to the growing urban population, "smart cities" are designed to improve people's quality of life by implementing cutting-edge technologies. The concept of a "smart city" refers to an effort to enhance a city's residents' economic and environmental well-being via implementing a centralized management system. With the use of sensors and actuators, smart cities can collect massive amounts of data, which can improve people's quality of life and design cities' services. Although smart cities contain vast amounts of data, only a percentage is used due to the noise and variety of the data sources. Information and communication technology (ICT) …
Enhancing Zero-Shot And Few-Shot Text Classification Using Pre-Trained Language Models, Yanan Chen
Enhancing Zero-Shot And Few-Shot Text Classification Using Pre-Trained Language Models, Yanan Chen
Theses and Dissertations (Comprehensive)
In recent years, the community of natural language processing (NLP) has seen amazing progress in the development of pre-trained language models (PLMs). The novel paradigm of PLMs does not require labeled data, allowing us to experiment with increased training scale through employing freely available colossal online self-training corpus to push the limits. Language models (LMs), such as GPT, BERT and T5, have achieved high performance on a wide range of NLP tasks. Meanwhile, research on zero-shot and few-shot text classification has received increasing attention. As labelling can be costly and time-consuming, how to perform data augmentation (DA) and enhance the …
Cross-Modal Food Retrieval: Learning A Joint Embedding Of Food Images And Recipes With Semantic Consistency And Attention Mechanism, Hao Wang, Doyen Sahoo, Chenghao Liu, Ke Shu, Palakorn Achananuparp, Ee-Peng Lim, Steven C. H. Hoi
Cross-Modal Food Retrieval: Learning A Joint Embedding Of Food Images And Recipes With Semantic Consistency And Attention Mechanism, Hao Wang, Doyen Sahoo, Chenghao Liu, Ke Shu, Palakorn Achananuparp, Ee-Peng Lim, Steven C. H. Hoi
Research Collection School Of Computing and Information Systems
Food retrieval is an important task to perform analysis of food-related information, where we are interested in retrieving relevant information about the queried food item such as ingredients, cooking instructions, etc. In this paper, we investigate cross-modal retrieval between food images and cooking recipes. The goal is to learn an embedding of images and recipes in a common feature space, such that the corresponding image-recipe embeddings lie close to one another. Two major challenges in addressing this problem are 1) large intra-variance and small inter-variance across cross-modal food data; and 2) difficulties in obtaining discriminative recipe representations. To address these …
Task Classification During Visual Search Using Classic Machine Learning And Deep Learning, Devangi Vilas Chinchankar
Task Classification During Visual Search Using Classic Machine Learning And Deep Learning, Devangi Vilas Chinchankar
Master's Projects
In an average human life, the eyes not only passively scan visual scenes, but most times end up actively performing tasks including, but not limited to, searching, comparing, and counting. As a result of the advances in technology, we are observing a boost in the average screen time. Humans are now looking at an increasing number of screens and in turn images and videos. Understanding what scene a user is looking at and what type of visual task is being performed can be useful in developing intelligent user interfaces, and in virtual reality and augmented reality devices. In this research, …
Analyzing And Detecting Android Malware And Deepfake, Md Shohel Rana
Analyzing And Detecting Android Malware And Deepfake, Md Shohel Rana
Dissertations
Rapid advances in artificial intelligence (AI), machine learning (ML), and deep learning (DL) over the past several decades have produced a variety of technologies and tools that, among numerous cybersecurity issues, have enticed cybercriminals and hackers to design malware for the Android operating systems and/or manipulate multimedia. For example, high-quality and realistic fake videos, images, or audios have been created to spread misinformation and propaganda, foment political discord and hate, or even harass and blackmail people; these manipulated, high-quality and realistic videos became known recently as Deepfake. There has been much work done in recent years on malware analysis and …