Open Access. Powered by Scholars. Published by Universities.®

Computer Engineering Commons

Open Access. Powered by Scholars. Published by Universities.®

Deep learning

Discipline
Institution
Publication Year
Publication
Publication Type

Articles 91 - 120 of 347

Full-Text Articles in Computer Engineering

Text-To-Sql: A Methodical Review Of Challenges And Models, Ali Buğra Kanburoğlu, Faik Boray Tek May 2024

Text-To-Sql: A Methodical Review Of Challenges And Models, Ali Buğra Kanburoğlu, Faik Boray Tek

Turkish Journal of Electrical Engineering and Computer Sciences

This survey focuses on Text-to-SQL, automated translation of natural language queries into SQL queries. Initially, we describe the problem and its main challenges. Then, by following the PRISMA systematic review methodology, we survey the existing Text-to-SQL review papers in the literature. We apply the same method to extract proposed Text-to-SQL models and classify them with respect to used evaluation metrics and benchmarks. We highlight the accuracies achieved by various models on Text-to-SQL datasets and discuss execution-guided evaluation strategies. We present insights into model training times and implementations of different models. We also explore the availability of Text-to-SQL datasets in non-English …


Dpafy-Gcaps: Denoising Patch-And-Amplify Gabor Capsule Network For The Recognition Of Gastrointestinal Diseases, Henrietta Adjei Pokuaa, Adeboya Felix Adekoya, Benjamin Asubam Weyori, Owusu Nyarko-Boateng May 2024

Dpafy-Gcaps: Denoising Patch-And-Amplify Gabor Capsule Network For The Recognition Of Gastrointestinal Diseases, Henrietta Adjei Pokuaa, Adeboya Felix Adekoya, Benjamin Asubam Weyori, Owusu Nyarko-Boateng

Turkish Journal of Electrical Engineering and Computer Sciences

Deep learning (DL) models have performed tremendously well in image classification. This good performance can be attributed to the availability of massive data in most domains. However, some domains are known to have few datasets, especially the health sector. This makes it difficult to develop domain-specific high-performing DL algorithms for these fields. The field of health is critical and requires accurate detection of diseases. In the United States Gastrointestinal diseases are prevalent and affect 60 to 70 million people. Ulcerative colitis, polyps, and esophagitis are some gastrointestinal diseases. Colorectal polyps is the third most diagnosed malignancy in the world. This …


Deep Learning Based Local Path Planning Method For Moving Robots, Zesen Liu, Sheng Bi, Chuanhong Guo, Yankui Wang, Min Dong May 2024

Deep Learning Based Local Path Planning Method For Moving Robots, Zesen Liu, Sheng Bi, Chuanhong Guo, Yankui Wang, Min Dong

Journal of System Simulation

Abstract: In order to integrate visual information into the robot navigation process, improve the robot's recognition rate of various types of obstacles, and reduce the occurrence of dangerous events, a local path planning network based on two-dimensional CNN and LSTM is designed, and a local path planning approach based on deep learning is proposed. The network uses the image from camera and the global path to generate the current steering angle required for obstacle avoidance and navigation. A simulated indoor scene is built for training and validating the network. A path evaluation method that uses the total length and the …


Toward Intuitive 3d Interactions In Virtual Reality: A Deep Learning- Based Dual-Hand Gesture Recognition Approach, Trudi Di Qi, Franceli L. Cibrian, Meghna Raswan, Tyler Kay, Hector M. Camarillo-Abad, Yuxin Wen May 2024

Toward Intuitive 3d Interactions In Virtual Reality: A Deep Learning- Based Dual-Hand Gesture Recognition Approach, Trudi Di Qi, Franceli L. Cibrian, Meghna Raswan, Tyler Kay, Hector M. Camarillo-Abad, Yuxin Wen

Engineering Faculty Articles and Research

Dual-hand gesture recognition is crucial for intuitive 3D interactions in virtual reality (VR), allowing the user to interact with virtual objects naturally through gestures using both handheld controllers. While deep learning and sensor-based technology have proven effective in recognizing single-hand gestures for 3D interactions, research on dual-hand gesture recognition for VR interactions is still underexplored. In this work, we introduce CWT-CNN-TCN, a novel deep learning model that combines a 2D Convolution Neural Network (CNN) with Continuous Wavelet Transformation (CWT) and a Temporal Convolution Network (TCN). This model can simultaneously extract features from the time-frequency domain and capture long-term dependencies using …


Machine Learning Security For Tactical Operations, Dr. Denaria Fields, Shakiya A. Friend, Andrew Hermansen, Dr. Tugba Erpek, Dr. Yalin E. Sagduyu May 2024

Machine Learning Security For Tactical Operations, Dr. Denaria Fields, Shakiya A. Friend, Andrew Hermansen, Dr. Tugba Erpek, Dr. Yalin E. Sagduyu

Military Cyber Affairs

Deep learning finds rich applications in the tactical domain by learning from diverse data sources and performing difficult tasks to support mission-critical applications. However, deep learning models are susceptible to various attacks and exploits. In this paper, we first discuss application areas of deep learning in the tactical domain. Next, we present adversarial machine learning as an emerging attack vector and discuss the impact of adversarial attacks on the deep learning performance. Finally, we discuss potential defense methods that can be applied against these attacks.


Incremental Image Dehazing Algorithm Based On Multiple Transfer Attention, Jinyang Wei, Keping Wang, Yi Yang, Shumin Fei Apr 2024

Incremental Image Dehazing Algorithm Based On Multiple Transfer Attention, Jinyang Wei, Keping Wang, Yi Yang, Shumin Fei

Journal of System Simulation

Abstract: In order to improve the processing ability of the depth-neural network dehazing algorithm to the supplementary data set, and to make the network differently process the image features of different importance to improve the dehazing ability of the network, an incremental dehazing algorithm based on multiple migration of attention is proposed. The teacher's attention generation network in the form of Encoder-Decoder extracts the multiple attention of labels and haze, which is used it as the label of the characteristic migration media network to constrain the network training to form the migration media attention as close as possible to the …


Preserving Location Authenticity: Multi-Sensor System To Thwart Gps Spoofing In Self-Driving Vehicles, Peng Jiang Apr 2024

Preserving Location Authenticity: Multi-Sensor System To Thwart Gps Spoofing In Self-Driving Vehicles, Peng Jiang

Electrical & Computer Engineering Theses & Dissertations

The ubiquity of the Global Positioning System (GPS) has cemented its role as the cornerstone for an array of location-based services and navigation systems, spanning applications from autonomous vehicles and drones to maritime vessels and wearable technology. Nonetheless, ensuring the integrity of reported geographical coordinates poses a formidable challenge, owing to the proliferation of diverse GPS spoofing tools. This predicament is compounded by the pervasive availability of tools like Fake GPS, Lockito, and software-defined radios, enabling even unsophisticated users to commandeer and disseminate counterfeit GPS coordinates. This dissertation undertakes the task of devising an encompassing and resilient framework, integrating a …


Computational Modeling And Analysis Of Facial Expressions And Gaze For Discovery Of Candidate Behavioral Biomarkers For Children And Young Adults With Autism Spectrum Disorder, Megan Anita Witherow Apr 2024

Computational Modeling And Analysis Of Facial Expressions And Gaze For Discovery Of Candidate Behavioral Biomarkers For Children And Young Adults With Autism Spectrum Disorder, Megan Anita Witherow

Electrical & Computer Engineering Theses & Dissertations

Facial expression production and perception in autism spectrum disorder (ASD) suggest the potential presence of behavioral biomarkers that may stratify individuals on the spectrum into prognostic or treatment subgroups. High-speed internet and the ease of technology have enabled remote, scalable, affordable, and timely access to medical care, such as measurements of ASDrelated behaviors in familiar environments to complement clinical observation. Machine and deep learning (DL)-based analysis of video tracking (VT) of expression production and eye tracking (ET) of expression perception may aid stratification biomarker discovery for children and young adults with ASD. However, there are open challenges in 1) facial …


Convolutional Neural Network In Motion Detection For Physiotherapy Exercise Movement, Dika Fikri Laistulloh, Anik Nur Handayani, Rosa Andrie Asmara, Phillip Taw Apr 2024

Convolutional Neural Network In Motion Detection For Physiotherapy Exercise Movement, Dika Fikri Laistulloh, Anik Nur Handayani, Rosa Andrie Asmara, Phillip Taw

Knowledge Engineering and Data Science

Physiotherapy focuses on movement and optimal utilization of the patient's potential. Exercise Therapy is a physiotherapy procedure that specifically focuses exercises on active and passive movements. Cerebral Palsy (CP) patients are one of the sufferers of motor disorders of the upper extremities. Cerebral Palsy (CP) patients suffer from disorders in motor functions of the upper extremities. Physiotherapy Exercise Movement has 4 categories of movement exercises for the therapy of people with upper extremity body disorders: Elbow flexor strengthening in sitting using free weights, lifting an object up, reaching diagonally in sitting, and reaching from a low surface to a high …


A Deep Learning Convolutional Neural Network For Antenna Near-Field Prediction And Surrogate Modeling, Md Rayhan Khan, Constantinos L. Zekios, Shubhendu Bhardwaj, Stavros V. Georgakopoulos Mar 2024

A Deep Learning Convolutional Neural Network For Antenna Near-Field Prediction And Surrogate Modeling, Md Rayhan Khan, Constantinos L. Zekios, Shubhendu Bhardwaj, Stavros V. Georgakopoulos

Department of Electrical and Computer Engineering: Faculty Publications

This study investigates the use of deep learning techniques for building a generalized surrogate model that can accurately and very efficiently predict antenna performance parameters. Notably, we focus on applications where a substantial amount of simulation time is required and prior data is available for deep learning use. Specifically, for these applications, we introduce deep learning models that efficiently and reliably model the near-field of the antenna. These models, in turn, accurately predict far-field properties and essential antenna metrics, such as the reflection coefficient. To demonstrate the efficiency of our method, the widely used rectangular patch antenna is considered, encompassing …


Motion Magnification-Inspired Feature Manipulation For Deepfake Detection, Aydamir Mirzayev, Hamdi Di̇bekli̇oğlu Feb 2024

Motion Magnification-Inspired Feature Manipulation For Deepfake Detection, Aydamir Mirzayev, Hamdi Di̇bekli̇oğlu

Turkish Journal of Electrical Engineering and Computer Sciences

Recent advances in deep learning, increased availability of large-scale datasets, and improvement of accelerated graphics processing units facilitated creation of an unprecedented amount of synthetically generated media content with impressive visual quality. Although such technology is used predominantly for entertainment, there is widespread practice of using deepfake technology for malevolent ends. This potential for malicious use necessitates the creation of detection methods capable of reliably distinguishing manipulated video content. In this work we aim to create a learning-based detection method for synthetically generated videos. To this end, we attempt to detect spatiotemporal inconsistencies by leveraging a learning-based magnification-inspired feature manipulation …


Automated Identification Of Vehicles In Very High-Resolution Uav Orthomosaics Using Yolov7 Deep Learning Model, Esra Yildirim, Umut Güneş Seferci̇k, Taşkın Kavzoğlu Feb 2024

Automated Identification Of Vehicles In Very High-Resolution Uav Orthomosaics Using Yolov7 Deep Learning Model, Esra Yildirim, Umut Güneş Seferci̇k, Taşkın Kavzoğlu

Turkish Journal of Electrical Engineering and Computer Sciences

The utilization of remote sensing products for vehicle detection through deep learning has gained immense popularity, especially due to the advancement of unmanned aerial vehicles (UAVs). UAVs offer millimeter-level spatial resolution at low flight altitudes, which surpasses traditional airborne platforms. Detecting vehicles from very high-resolution UAV data is crucial in numerous applications, including parking lot and highway management, traffic monitoring, search and rescue missions, and military operations. Obtaining UAV data at desired periods allows the detection and tracking of target objects even several times during a day. Despite challenges such as diverse vehicle characteristics, traffic congestion, and hardware limitations, the …


Action Recognition Model Of Directed Attention Based On Cosine Similarity, Chen Li, Ming He, Chen Dong, Wei Li Jan 2024

Action Recognition Model Of Directed Attention Based On Cosine Similarity, Chen Li, Ming He, Chen Dong, Wei Li

Journal of System Simulation

Abstract: Aiming at the lack of directionality of traditional dot product attention, this paper proposes a directed attention model (DAM) based on cosine similarity. To effectively represent the direction relationship between the spatial and temporal features of video frames, the paper defines the relationship function in the attention mechanism using the cosine similarity theory, which can remove the absolute value of the relationship between features. To reduce the computational burden of the attention mechanism, the operation is decomposed from two dimensions of time and space. The computational complexity is further optimized by combining linear attention operation. The experiment is divided …


Deep Transfer Learning For Detection Of Upper And Lower Body Movements: Transformer With Convolutional Neural Network, Kyle Lacroix, Davoud Gholamiangonabadi, Ana Luisa Trejos, Katarina Grolinger Jan 2024

Deep Transfer Learning For Detection Of Upper And Lower Body Movements: Transformer With Convolutional Neural Network, Kyle Lacroix, Davoud Gholamiangonabadi, Ana Luisa Trejos, Katarina Grolinger

Electrical and Computer Engineering Publications

When humans repeat the same motion, the tendons, muscles, and nerves can be damaged, causing Repetitive Stress Injuries (RSI). If the repetitive motions that lead to RSI are recognized early, actions can be taken to prevent these injuries. As Human Activity Recognition (HAR) aims to identify activities employing wearable or environment sensors, HAR is the first step toward identifying repetitive motions. Deep learning models, such as Convolutional Neural Networks (CNNs), have seen great success in recognizing activities for participants whose data are used in the model training; however, their accuracy drops for new participants as people move in different ways. …


Benchmarking And Enhancing Generalization In Multilingual Speech Emotion Recognition, Mohamed Osman Ismael Jan 2024

Benchmarking And Enhancing Generalization In Multilingual Speech Emotion Recognition, Mohamed Osman Ismael

Theses and Dissertations

Speech Emotion Recognition (SER) is pivotal in advancing human-computer interaction by enabling machines to understand and respond to human emotions. Despite significant progress with self-supervised learning models, SER systems often struggle with generalization across diverse languages and unseen data distributions, limiting their real-world applicability. This thesis addresses these challenges by first introducing a large-scale benchmark to evaluate the robustness and adaptability of state-of-the-art SER models in both in-domain and out-of-domain settings. The benchmark includes a diverse set of multilingual datasets, emphasizing cross-lingual and out-of-domain evaluations to assess model generalization. Surprisingly, we find that the Whisper model, originally designed for automatic …


Novel Approach To Music Analysis Using Apache Spark, Nidhi Zare Jan 2024

Novel Approach To Music Analysis Using Apache Spark, Nidhi Zare

Master's Projects

Music is one of the most common source of entertainment. Every user has their own taste of music and prefer to listen music that adheres to their taste and mood. There are various categories, called as music genres in which music can be classified. This research project addresses the challenge in music genre classification by using various deep learning models such as Convolutional Neural Networks (CNNs), Recurrent Neural Networks (RNNs), Very Deep Convolutional Networks (VGGNet), ResNet and others. The primary objective of this research is to enhance the accuracy of music genre classification using a distributed computing framework Apache Spark. …


Skin Cancer Detection Using Reinforcement Learning, Vikas Chercadu Jan 2024

Skin Cancer Detection Using Reinforcement Learning, Vikas Chercadu

Master's Projects

Advances in medical diagnostics have increasingly harnessed the power of artificial

intelligence, offering substantial improvements in early and accurate disease identifi- cation. This paper elaborates on a novel integration of Multi-Agent Reinforcement

Learning (MARL) with Deep Learning for the early detection of skin cancer, one of the most prevalent and lethal forms of cancer when left unchecked. Our project capitalizes on the sophisticated VGG16 Network for the extraction of detailed features from the widely-utilized HAM10000 dermatoscopic dataset, enhancing these features with additional color, texture, and shape analysis. Utilizing a custom-designed MARL environment, we facilitate a collaboration among multiple intelligent agents, …


Comparitive Analysis Of Time Series Forecasting Using Frequency Informed Dense Neural Networks And Lstm, Avinash Mangalore Suresh Jan 2024

Comparitive Analysis Of Time Series Forecasting Using Frequency Informed Dense Neural Networks And Lstm, Avinash Mangalore Suresh

Master's Projects

Time series forecasting influences our lives on a daily basis, being a versatile tool in various application areas like environmental studies, finance, medicine and much more. While there are many established statistical and deep learning approaches to model time series data, each implementation comes with their own set of drawbacks or areas of improvements. Most of the existing deep learning architectures and research have focused on modeling time series data in the time-domain exclusively. However training deep learning models in the time-domain has some drawbacks, mainly due to the inherent temporal dependence of each time-step on the time-steps before it, …


Temporal Dynamics In Diabetes Prediction: A Sensor-Driven Time-Series Exploration, Monica Meduri Jan 2024

Temporal Dynamics In Diabetes Prediction: A Sensor-Driven Time-Series Exploration, Monica Meduri

Master's Projects

Diabetes is a lifelong illness that, if not detected or managed appropriately, turns into serious complications. Correct glucose forecasting is critical to ensuring timely interventions, thereby minimizing risks of hyperglycemia and hypoglycemia, and optimizing the management strategies of the disease. Classical machine learning models have been applied in the blood glucose forecasting problem for a long time, however, usage of transformer-based architectures is still scarce within the literature. Due to the self-attention mechanism, transformers can capture temporal relationships very effectively, which makes them suitable for time-series data. TFT is a novel framework proposed here to utilize time-series data from CGM …


Classification Of Sow Postures Using Convolutional Neural Network And Depth Images, Md Towfiqur Rahman, Tami M. Brown-Brandl, Gary A. Rohrer, Sudhendu R. Sharma, Yeyin Shi Jan 2024

Classification Of Sow Postures Using Convolutional Neural Network And Depth Images, Md Towfiqur Rahman, Tami M. Brown-Brandl, Gary A. Rohrer, Sudhendu R. Sharma, Yeyin Shi

Department of Agricultural and Biological Systems Engineering: Faculty Publications

The United States swine industry reports an average preweaning mortality of approximately 16% where approximately 6% of them are attributed to piglets overlayed by sows. Detecting postural transitions and estimating sows’ time budgets for different postures are valuable information for breeders and engineering design of farrowing facilities to eventually reduce piglet death. Computer vision tools can help monitor changes in animal posture accurately and efficiently. To create a more robust system and eliminate varying lighting issues within a day including daytime/ nighttime differences, there is an advantage to using depth cameras over digital cameras. In this study, a computer vision …


Improvements In Biomedical Image Analysis With Computational Intelligence And Data Fusion Techniques, Akanksha Maurya Jan 2024

Improvements In Biomedical Image Analysis With Computational Intelligence And Data Fusion Techniques, Akanksha Maurya

Doctoral Dissertations

"An estimated 2 million new cases of basal cell carcinoma (BCC) are diagnosed each year in the United States, making it one of the most common skin cancers. Earlier detection of these cancers enables less invasive biopsies. Clinical detection consists of a preliminary visual observation of these skin lesions by an experienced dermatologist making it a specialized task highly dependent on their time, availability, and resources. Hence, there is a need for automating this process that can assist healthcare staff. In recent years, deep learning (DL) has been used extensively and successfully to diagnose different cancers in dermoscopic images. Telangiectasia …


A Memory Efficient Deep Recurrent Q-Learning Approach For Autonomous Wildfire Surveillance, Jeremy A. Cantor Jan 2024

A Memory Efficient Deep Recurrent Q-Learning Approach For Autonomous Wildfire Surveillance, Jeremy A. Cantor

UNF Graduate Theses and Dissertations

Previous literature demonstrates that autonomous UAVs (unmanned aerial vehicles) have the po- tential to be utilized for wildfire surveillance. This advanced technology empowers firefighters by providing them with critical information, thereby facilitating more informed decision-making processes. This thesis applies deep Q-learning techniques to the problem of control policy design under the objective that the UAVs collectively identify the maximum number of locations that are under fire, assuming the UAVs can share their observations. The prohibitively large state space underlying the control policy motivates a neural network approximation, but prior work used only convolutional layers to extract spatial fire information from …


Deep-Learning Approaches To Predict Remaining Useful Life Of Hard Disks, Rohan Mohapatra Jan 2024

Deep-Learning Approaches To Predict Remaining Useful Life Of Hard Disks, Rohan Mohapatra

Master's Projects

On a daily basis, data centers process huge volumes of data using inexpensive hard disks. Data stored in these disks serve a range of critical functional needs from financial, and healthcare to aerospace. As such, premature disk failure and consequent loss of data can be catastrophic. To mitigate the risk of failures, cloud storage providers perform condition-based monitoring and replace hard disks before they fail. By estimating the remaining useful life (RUL) of hard disk drives, one can predict the time-to-failure of a particular device and replace it at the right time, ensuring maximum utilization whilst reducing operational costs. We …


Multimodal Fusion For Audio-Image And Video Action Recognition, Muhammad B. Shaikh, Douglas Chai, Syed M. S. Islam, Naveed Akhtar Jan 2024

Multimodal Fusion For Audio-Image And Video Action Recognition, Muhammad B. Shaikh, Douglas Chai, Syed M. S. Islam, Naveed Akhtar

Research outputs 2022 to 2026

Multimodal Human Action Recognition (MHAR) is an important research topic in computer vision and event recognition fields. In this work, we address the problem of MHAR by developing a novel audio-image and video fusion-based deep learning framework that we call Multimodal Audio-Image and Video Action Recognizer (MAiVAR). We extract temporal information using image representations of audio signals and spatial information from video modality with the help of Convolutional Neutral Networks (CNN)-based feature extractors and fuse these features to recognize respective action classes. We apply a high-level weights assignment algorithm for improving audio-visual interaction and convergence. This proposed fusion-based framework utilizes …


Using Artificial Intelligence To Diagnose Demented In The Elderly, Ohood Fadil Alwan Dec 2023

Using Artificial Intelligence To Diagnose Demented In The Elderly, Ohood Fadil Alwan

Al-Esraa University College Journal for Engineering Sciences

In addition to possible other symptoms, memory loss and impairment are the hallmarks of Demented or Alzheimer’s disease (AD). Despite the fact that dementia is incurable and has a significant negative impact on patients' lives, an early diagnosis can help start the right treatment and prevent additional brain damage. Over the years, machine learning techniques have been used to classify AD; nevertheless, the efficacy of the results depends on the use of multi-step classifiers and manually created features. Thanks to recent advances in deep learning, patterns may now be classified using neural networks' final stage. In order to diagnose dementia …


Sc-Fuse: A Feature Fusion Approach For Unpaved Road Detection From Remotely Sensed Images, Aniruddh Saxena Dec 2023

Sc-Fuse: A Feature Fusion Approach For Unpaved Road Detection From Remotely Sensed Images, Aniruddh Saxena

School of Computing: Dissertations, Theses, and Student Research

Road network extraction from remote sensing imagery is crucial for numerous applications, ranging from autonomous navigation to urban and rural planning. A particularly challenging aspect is the detection of unpaved roads, often underrepresented in research and data. These roads display variability in texture, width, shape, and surroundings, making their detection quite complex. This thesis addresses these challenges by creating a specialized dataset and introducing the SC-Fuse model.

Our custom dataset comprises high resolution remote sensing imagery which primarily targets unpaved roads of the American Midwest. To capture the diverse seasonal variation and their impact, the dataset includes images from different …


Melanoma Detection Based On Deep Learning Networks, Sanjay Devaraneni Dec 2023

Melanoma Detection Based On Deep Learning Networks, Sanjay Devaraneni

Electronic Theses, Projects, and Dissertations

Our main objective is to develop a method for identifying melanoma enabling accurate assessments of patient’s health. Skin cancer, such as melanoma can be extremely dangerous if not detected and treated early. Detecting skin cancer accurately and promptly can greatly increase the chances of survival. To achieve this, it is important to develop a computer-aided diagnostic support system. In this study a research team introduces a sophisticated transfer learning model that utilizes Resnet50 to classify melanoma. Transfer learning is a machine learning technique that takes advantage of trained models, for similar tasks resulting in time saving and enhanced accuracy by …


A Comparative Study Of Yolo Models And A Transformer-Based Yolov5 Model For Mass Detection In Mammograms, Damla Coşkun, Dervi̇ş Karaboğa, Alper Baştürk, Bahri̇ye Akay, Özkan Ufuk Nalbantoğlu, Serap Doğan, İshak Paçal, Meryem Altin Karagöz Nov 2023

A Comparative Study Of Yolo Models And A Transformer-Based Yolov5 Model For Mass Detection In Mammograms, Damla Coşkun, Dervi̇ş Karaboğa, Alper Baştürk, Bahri̇ye Akay, Özkan Ufuk Nalbantoğlu, Serap Doğan, İshak Paçal, Meryem Altin Karagöz

Turkish Journal of Electrical Engineering and Computer Sciences

Breast cancer is a prevalent form of cancer across the globe, and if it is not diagnosed at an early stage it can be life-threatening. In order to aid in its diagnosis, detection, and classification, computer-aided detection (CAD) systems are employed. You Only Look Once (YOLO)-based CAD algorithms have become very popular owing to their highly accurate results for object detection tasks in recent years. Therefore, the most popular YOLO models are implemented to compare the performance in mass detection with various experiments on the INbreast dataset. In addition, a YOLO model with an integrated Swin Transformer in its backbone …


An In-Depth Analysis Of Domain Adaptation In Computer And Robotic Vision, Muhammad Hassan Tanveer, Zainab Fatima, Shehnila Zardari, David A. Guerra-Zubiaga Nov 2023

An In-Depth Analysis Of Domain Adaptation In Computer And Robotic Vision, Muhammad Hassan Tanveer, Zainab Fatima, Shehnila Zardari, David A. Guerra-Zubiaga

Faculty Articles

This review article comprehensively delves into the rapidly evolving field of domain adaptation in computer and robotic vision. It offers a detailed technical analysis of the opportunities and challenges associated with this topic. Domain adaptation methods play a pivotal role in facilitating seamless knowledge transfer and enhancing the generalization capabilities of computer and robotic vision systems. Our methodology involves systematic data collection and preparation, followed by the application of diverse assessment metrics to evaluate the efficacy of domain adaptation strategies. This study assesses the effectiveness and versatility of conventional, deep learning-based, and hybrid domain adaptation techniques within the domains of …


Infrared Imaging Segmentation Employing An Explainable Deep Neural Network, Xinfei Liao, Dan Wang, Zairan Li, Nilanjan Dey, Rs Simon, Fuqian Shi Oct 2023

Infrared Imaging Segmentation Employing An Explainable Deep Neural Network, Xinfei Liao, Dan Wang, Zairan Li, Nilanjan Dey, Rs Simon, Fuqian Shi

Turkish Journal of Electrical Engineering and Computer Sciences

Explainable AI (XAI) improved by a deep neural network (DNN) of a residual neural network (ResNet) and long short-term memory networks (LSTMs), termed XAIRL, is proposed for segmenting foot infrared imaging datasets. First, an infrared sensor imaging dataset is acquired by a foot infrared sensor imaging device and preprocessed. The infrared sensor image features are then defined and extracted with XAIRL being applied to segment the dataset. This paper compares and discusses our results with XAIRL. Evaluation indices are applied to perform various measurements for foot infrared image segmentation including accuracy, precision, recall, F1 score, intersection over union (IoU), Dice …