Open Access. Powered by Scholars. Published by Universities.®

Computer Sciences Commons

Open Access. Powered by Scholars. Published by Universities.®

Deep Learning

Discipline
Institution
Publication Year
Publication
Publication Type
File Type

Articles 91 - 120 of 432

Full-Text Articles in Computer Sciences

"Deep Learning For Microscope Image Denoising", Nasreen Buhn, Sriya Adunur, Guy Hagen, Jonathan Ventura Oct 2024

"Deep Learning For Microscope Image Denoising", Nasreen Buhn, Sriya Adunur, Guy Hagen, Jonathan Ventura

College of Engineering Summer Undergraduate Research Program

In order to avoid damaging live cells, optical microscope imaging must be conducted under low-excitation light intensity and/or short exposure times, resulting in low signal-to-noise ratios (SNR). Deep learning methods offer an effective solution for removing microscope noise, utilizing algorithms that are able to reconstruct finer features in low SNR images. This research explores the denoising capability of several deep learning methods based on PSNR and SSIM. Tested methods include traditional approaches (BMED), supervised learning (CARE and Restormer), and unsupervised methods (Noise2Fast, N2V, SSD-Unsupervised, and SASSID). The Restormer model, which employs an encoder-decoder transformer architecture and progressive learning, stood out …


Retrofitting A Legacy Cutlery Washing Machine Using Computer Vision, Hua Leong Fwa Oct 2024

Retrofitting A Legacy Cutlery Washing Machine Using Computer Vision, Hua Leong Fwa

Research Collection School Of Computing and Information Systems

Industry 4.0, the digitalization of manufacturing promises to lead to lowered cost, efficient processes and even discovery of new business models. However, many of the enterprises have huge investments in legacy machines which are not 'smart'. In this study, we thus designed a cost-efficient solution to retrofit a legacy conveyor belt-based cutlery washing machine with a commodity web camera. We then applied computer vision (using both traditional image processing and deep learning techniques) to infer the speed and utilization of the machine. We detailed the algorithms that we designed for computing both speed andutilization. With the existing operational constraints of …


Developing Empathetic Ai: Exploring The Potential Of Artificial Intelligence To Understand And Simulate Family Dynamics And Cultural Identity, Emily Barnes, James Hutson Sep 2024

Developing Empathetic Ai: Exploring The Potential Of Artificial Intelligence To Understand And Simulate Family Dynamics And Cultural Identity, Emily Barnes, James Hutson

Faculty Scholarship

The rapid advancement of Artificial Intelligence (AI) has significantly impacted various domains. Yet, the exploration of AI's potential to develop a deep understanding of family culture and identity remains underexplored. This study introduces the concept of "a love of grandma and apple pie" to symbolize the potential of various AI to internalize and appreciate familial relationships, cultural traditions, and personal identity. The proposed study would investigate how an advanced deep learning model, trained on diverse unstructured datasets—including multimedia data from 100 families-could learn and reflect human-like emotions, values, and cultural understanding. Utilizing Convolutional Neural Networks (CNNs) for visual data processing …


The Impact Of Model Variations On The Robustness Of Deep Learning Models In Adversarial Settings, Firuz Juraev, Mohammed Abuhamad, Simon S. Woo, George K. Thiruvathukal, Tamer Abuhmed Aug 2024

The Impact Of Model Variations On The Robustness Of Deep Learning Models In Adversarial Settings, Firuz Juraev, Mohammed Abuhamad, Simon S. Woo, George K. Thiruvathukal, Tamer Abuhmed

Computer Science: Faculty Publications and Other Works

Rapid advancements of deep learning are accelerating adoption in a wide variety of applications, including safety-critical applications such as self-driving vehicles, drones, robots, and surveillance systems. These advancements include applying variations of sophisticated techniques that improve the performance of models. However, such models are not immune to adversarial manipulations, which can cause the system to misbehave and remain unnoticed by experts. The frequency of modifications to existing deep learning models necessitates thorough analysis to determine the impact on models’ robustness. In this work, we present an experimental evaluation of the effects of model modifications on deep learning model robustness using …


Development And Optimization Of A 1-Dimensional Convolutional Neural Network-Based Keyword Spotting Model For Fpga Acceleration, Trysten E. Dembeck Aug 2024

Development And Optimization Of A 1-Dimensional Convolutional Neural Network-Based Keyword Spotting Model For Fpga Acceleration, Trysten E. Dembeck

Masters Theses

Spoken Keyword Spotting (KWS) has steadily remained one of the most studied and implemented technologies in human-facing artificially intelligent systems and has enabled them to detect specific keywords in utterances. Modern machine learning models, such as the variants of deep neural networks, have significantly improved the performance and accuracy of these systems over other rudimentary techniques. However, they often demand substantial computational resources, use large parameter spaces, and introduce latencies that limit their real-time applicability and offline use. These speed and memory requirements have become a tremendous problem where faster and more efficient KWS methods dominate and better meet industry …


Transformer-Based Deep Learning Prediction Of 10-Degree Humphrey Visual Field Tests From 24-Degree Data, Min Shi, Anagha Lokhande, Yu Tian, Yan Luo, Mohammad Eslami, Saber Kazeminasab, Tobias Elze, Lucy Shen, Louis Pasquale, Sarah Wellik, Carlos Gustavo De Moraes, Jonathan Myers, Nazlee Zebardast, David Friedman, Michael Boland, Mengyu Wang Aug 2024

Transformer-Based Deep Learning Prediction Of 10-Degree Humphrey Visual Field Tests From 24-Degree Data, Min Shi, Anagha Lokhande, Yu Tian, Yan Luo, Mohammad Eslami, Saber Kazeminasab, Tobias Elze, Lucy Shen, Louis Pasquale, Sarah Wellik, Carlos Gustavo De Moraes, Jonathan Myers, Nazlee Zebardast, David Friedman, Michael Boland, Mengyu Wang

Wills Eye Hospital Papers

PURPOSE: To predict 10-2 Humphrey visual fields (VFs) from 24-2 VFs and associated non-total deviation features using deep learning.

METHODS: We included 5189 reliable 24-2 and 10-2 VF pairs from 2236 patients, and 28,409 reliable pairs of macular OCT scans and 24-2 VF from 19,527 eyes of 11,560 patients. We developed a transformer-based deep learning model using 52 total deviation values and nine VF test features to predict 68 10-2 total deviation values. The mean absolute error, root mean square error, and the R2 were evaluation metrics. We further evaluated whether the predicted 10-2 VFs can improve the structure-function relationship …


High Prevalence Of Artifacts In Optical Coherence Tomography With Adequate Signal Strength, Wei-Chun Lin, Aaron Coyner, Charles Amankwa, Abigail Lucero, Gadi Wollstein, Joel Schuman, Hiroshi Ishikawa Aug 2024

High Prevalence Of Artifacts In Optical Coherence Tomography With Adequate Signal Strength, Wei-Chun Lin, Aaron Coyner, Charles Amankwa, Abigail Lucero, Gadi Wollstein, Joel Schuman, Hiroshi Ishikawa

Wills Eye Hospital Papers

PURPOSE: This study aims to investigate the prevalence of artifacts in optical coherence tomography (OCT) images with acceptable signal strength and evaluate the performance of supervised deep learning models in improving OCT image quality assessment.

METHODS: We conducted a retrospective study on 4555 OCT images from 546 patients, with each image having an acceptable signal strength (≥6). A comprehensive analysis of prevalent OCT artifacts was performed, and five pretrained convolutional neural network models were trained and tested to infer images based on quality.

RESULTS: Our results showed a high prevalence of artifacts in OCT images with acceptable signal strength. Approximately …


Checklist For Reproducibility Of Deep Learning In Medical Imaging, Mana Moassefi, Yashbir Singh, Gian Marco Conte, Bardia Khosravi, Pouria Rouzrokh, Sanaz Vahdati, Nabile Safdar, Linda Moy, Felipe Kitamura, Amilcare Gentili, Paras Lakhani, Nina Kottler, Safwan Halabi, Joseph Yacoub, Yuankai Hou, Khaled Younis, Bradley Erickson, Elizabeth Krupinski, Shahriar Faghani Aug 2024

Checklist For Reproducibility Of Deep Learning In Medical Imaging, Mana Moassefi, Yashbir Singh, Gian Marco Conte, Bardia Khosravi, Pouria Rouzrokh, Sanaz Vahdati, Nabile Safdar, Linda Moy, Felipe Kitamura, Amilcare Gentili, Paras Lakhani, Nina Kottler, Safwan Halabi, Joseph Yacoub, Yuankai Hou, Khaled Younis, Bradley Erickson, Elizabeth Krupinski, Shahriar Faghani

Department of Radiology Faculty Papers

The application of deep learning (DL) in medicine introduces transformative tools with the potential to enhance prognosis, diagnosis, and treatment planning. However, ensuring transparent documentation is essential for researchers to enhance reproducibility and refine techniques. Our study addresses the unique challenges presented by DL in medical imaging by developing a comprehensive checklist using the Delphi method to enhance reproducibility and reliability in this dynamic field. We compiled a preliminary checklist based on a comprehensive review of existing checklists and relevant literature. A panel of 11 experts in medical imaging and DL assessed these items using Likert scales, with two survey …


Interpretable And Evidential Deep Learning For Medical Image Analysis, Sai Chandra Kosaraju Aug 2024

Interpretable And Evidential Deep Learning For Medical Image Analysis, Sai Chandra Kosaraju

UNLV Theses, Dissertations, Professional Papers, and Capstones

Automatic histopathological Whole Slide Image (WSI) analysis has been highlighted along with the advancements in microscopic imaging techniques, but manual examination and diagnosis of WSIs are time-consuming and tiresome. Recently, deep convolutional neural networks have succeeded in histopathological image analysis. Especially, Convolutional Neural Networks CNN models such as Inception and DenseNet have achieved effective performance. However, automatic histopathological WSI analysis still has significant drawbacks such as considering deep learning as black-box models, predicting disease independently on a small part of images (patch images) extracted from WSIs, limitations in predicting a single slide-based score for a patient, and capturing disease-specific morphology …


Ensemble Learning For Accurate Prediction Of Heart Sounds Using Gammatonegram Images, Sinam Ashinikumar Singh, Sinam Ajitkumar Singh, Aheibam Dinamani Singh Jul 2024

Ensemble Learning For Accurate Prediction Of Heart Sounds Using Gammatonegram Images, Sinam Ashinikumar Singh, Sinam Ajitkumar Singh, Aheibam Dinamani Singh

Turkish Journal of Electrical Engineering and Computer Sciences

The analysis of heart sound signals constitutes a pivotal domain in healthcare, with the prediction of imbalanced heart sounds offering critical diagnostic insights. However, the inherent diversity in cardiac sound patterns presents a substantial challenge in predicting imbalanced signals. Many scientific disciplines have focused a great deal of emphasis on the problem of class inequality. We introduce an ensemble learning approach employing a convolutional neural network model-based deep learning algorithm to effectively tackle the challenges associated with predicting imbalanced heart sound signals. We use a Gammatone filter bank to extract relevant features from the heard sound signal. Our approach leverages …


From Graph Theory For Robust Deep Networks To Graph Learning For Multimodal Cancer Analysis, Asim Waqas Jun 2024

From Graph Theory For Robust Deep Networks To Graph Learning For Multimodal Cancer Analysis, Asim Waqas

USF Tampa Graduate Theses and Dissertations

This dissertation explores the intersection of graph theory and deep learning, focusing on enhancing the robustness of deep neural networks (DNNs) and applying these advancements to complex problems like cancer diagnosis and treatment. We investigate the structural properties of graphs and their influence on neural network performance, particularly in multimodal learning. The work delves into the design space of DNN architectures using graph-theoretic measures, transforming graphs into DNN architectures for various tasks, and examining their robustness against noise and adversarial attacks. The study extends to medical imaging, highlighting advanced DNN architectures like U-Net for brain tumor segmentation. It addresses the …


Machine Learning Multimodal Framework For Fake News Detection And Mitigation, Nada A. Gaballah Jun 2024

Machine Learning Multimodal Framework For Fake News Detection And Mitigation, Nada A. Gaballah

Theses and Dissertations

Social media has become our new reality, people wake up every morning and the first thing they do before getting out of bed, is check their social media. Nowadays, people rarely read newspapers, they even rarely watch TV news or listen to radio broadcasts. In recent years, we have witnessed lots of fake news roaming social media every second, with people simply believing it and spreading it even more without checking the credibility of this news. This fake news affected several domains like what happened in the US election in 2016 and again in 2020, the false information about Covid-19 …


Context In Computer Vision: A Taxonomy, Multi-Stage Integration, And A General Framework, Xuan Wang Jun 2024

Context In Computer Vision: A Taxonomy, Multi-Stage Integration, And A General Framework, Xuan Wang

Dissertations, Theses, and Capstone Projects

Contextual information has been widely used in many computer vision tasks, such as object detection, video action detection, image classification, etc. Recognizing a single object or action out of context could be sometimes very challenging, and context information may help improve the understanding of a scene or an event greatly. However, existing approaches design specific contextual information mechanisms for different detection tasks.

In this research, we first present a comprehensive survey of context understanding in computer vision, with a taxonomy to describe context in different types and levels. Then we proposed MultiCLU, a new multi-stage context learning and utilization framework, …


Morp: Monocular Orientation Regression Pipeline, Jacob Gunderson Jun 2024

Morp: Monocular Orientation Regression Pipeline, Jacob Gunderson

Master's Theses

Orientation estimation of objects plays a pivotal role in robotics, self-driving cars, and augmented reality. Beyond mere position, accurately determining the orientation of objects is essential for constructing precise models of the physical world. While 2D object detection has made significant strides, the field of orientation estimation still faces several challenges. Our research addresses these hurdles by proposing an efficient pipeline which facilitates rapid creation of labeled training data and enables direct regression of object orientation from a single image. We start by creating a digital twin of a physical object using an iPhone, followed by generating synthetic images using …


Singleadv: Single-Class Target-Specific Attack Against Interpretable Deep Learning Systems, Eldor Abdukhamidov, Mohammed Abuhamad, George K. Thiruvathukal, Hyoungshick Kim, Tamer Abuhmed May 2024

Singleadv: Single-Class Target-Specific Attack Against Interpretable Deep Learning Systems, Eldor Abdukhamidov, Mohammed Abuhamad, George K. Thiruvathukal, Hyoungshick Kim, Tamer Abuhmed

Computer Science: Faculty Publications and Other Works

In this paper, we present a novel Single-class target-specific Adversarial attack called SingleADV. The goal of SingleADV is to generate a universal perturbation that deceives the target model into confusing a specific category of objects with a target category while ensuring highly relevant and accurate interpretations. The universal perturbation is stochastically and iteratively optimized by minimizing the adversarial loss that is designed to consider both the classifier and interpreter costs in targeted and non-targeted categories. In this optimization framework, ruled by the first- and second-moment estimations, the desired loss surface promotes high confidence and interpretation score of adversarial samples. By …


Context Aware Music Recommendation And Playlist Generation, Elias Mann May 2024

Context Aware Music Recommendation And Playlist Generation, Elias Mann

SMU Journal of Undergraduate Research

There are many reasons people listen to music, and the type of music is largely determined by what the listener may be doing while they listen. For example, one may listen to one type of music while commuting, another while exercising, and yet another while relaxing. Without access to the physiological state of the user, current music recommendation methods rely on collaborative filtering - recommending music based on what other similar users listen to - and content based filtering - recommending songs based on their similarities to songs the user already prefers. With the rise in popularity of smart devices …


Missing Wedge Completion Via Unsupervised Learning With Coordinate Networks, Dave Van Veen, Jesús G Galaz-Montoya, Liyue Shen, Philip Baldwin, Akshay S Chaudhari, Dmitry Lyumkis, Michael F Schmid, Wah Chiu, John Pauly May 2024

Missing Wedge Completion Via Unsupervised Learning With Coordinate Networks, Dave Van Veen, Jesús G Galaz-Montoya, Liyue Shen, Philip Baldwin, Akshay S Chaudhari, Dmitry Lyumkis, Michael F Schmid, Wah Chiu, John Pauly

Faculty, Staff and Students Publications

Cryogenic electron tomography (cryoET) is a powerful tool in structural biology, enabling detailed 3D imaging of biological specimens at a resolution of nanometers. Despite its potential, cryoET faces challenges such as the missing wedge problem, which limits reconstruction quality due to incomplete data collection angles. Recently, supervised deep learning methods leveraging convolutional neural networks (CNNs) have considerably addressed this issue; however, their pretraining requirements render them susceptible to inaccuracies and artifacts, particularly when representative training data is scarce. To overcome these limitations, we introduce a proof-of-concept unsupervised learning approach using coordinate networks (CNs) that optimizes network weights directly against input …


Ai: Is The Future Robots?, Ruthanne Fischer May 2024

Ai: Is The Future Robots?, Ruthanne Fischer

University Honors College

The goal of this thesis is to analyze how Artificial Intelligence is actively changing. The newly developed technology that is known as Artificial Intelligence can potentially have implications that are not yet known. Artificial Intelligence impacts the business world in many ways. Things such as ChatGPT and Deepfakes are some of the biggest forms of Artificial Intelligence today. It is important to know how Artificial Intelligence is changing every day because no one knows the true implications of a technology like this. It is crucial for companies to understand how this technology can potentially be harmful to their employees. Artificial …


Satellite Image Analysis And Sidewalk Classification Using Deep Learning Models, Subeksha Khanal May 2024

Satellite Image Analysis And Sidewalk Classification Using Deep Learning Models, Subeksha Khanal

Master's Theses

Lack of sidewalk pavement can contribute to pedestrian fatalities and injuries in the USA. Although many researchers have conducted research on sidewalk detection, there are not many publicly available datasets that we would work on for sidewalk classification. In this study, I conducted classification tasks using pretrained CNN models including VGG16 and ResNet50. I extended these models by adding custom layers at the top of pretrained layers, employing various techniques to improve the classification accuracy.

The dataset comprises 4,731 images of sidewalk based on occlusion levels, including images where the sidewalk is visible from overhead and instances where sidewalk is …


Robust And Trustworthy Deep Learning: Attacks, Defenses And Designs, Bingyin Zhao May 2024

Robust And Trustworthy Deep Learning: Attacks, Defenses And Designs, Bingyin Zhao

All Dissertations

Deep neural networks (DNNs) have achieved unprecedented success in many fields. However, robustness and trustworthiness have become emerging concerns since DNNs are vulnerable to various attacks and susceptible to data distributional shifts. Attacks such as data poisoning and out-of-distribution scenarios such as natural corruption significantly undermine the performance and robustness of DNNs in model training and inference and impose uncertainty and insecurity on the deployment in real-world applications. Thus, it is crucial to investigate threats and challenges against deep neural networks, develop corresponding countermeasures, and dig into design tactics to secure their safety and reliability. The works investigated in this …


Lecture-Style Tutorial: Towards Graph Foundation Models, Chuan Shi, Cheng Yang, Yuan Fang, Lichao Sun, Philip Yu May 2024

Lecture-Style Tutorial: Towards Graph Foundation Models, Chuan Shi, Cheng Yang, Yuan Fang, Lichao Sun, Philip Yu

Research Collection School Of Computing and Information Systems

Emerging as fundamental building blocks for diverse artificial intelligence applications, foundation models have achieved notable success across natural language processing and many other domains. Concurrently, graph machine learning has gradually evolved from shallow methods to deep models to leverage the abundant graph-structured data that constitute an important pillar in the data ecosystem for artificial intelligence. Naturally, the emergence and homogenization capabilities of foundation models have piqued the interest of graph machine learning researchers. This has sparked discussions about developing a next-generation graph learning paradigm, one that is pre-trained on broad graph data and can be adapted to a wide range …


Learning Scene Semantics For 3d Scene Retrieval, Natalie Gleason May 2024

Learning Scene Semantics For 3d Scene Retrieval, Natalie Gleason

Honors Theses

This project presents a comprehensive exploration into semantics-driven 3D scene retrieval, aiming to bridge the gap between 2D sketches/images and 3D models. Through four distinct research objectives, this project endeavors to construct a foundational infrastructure, develop methodologies for quantifying semantic similarity, and advance a semantics-based retrieval framework for 2D scene sketch-based and image-based 3D scene retrieval. Leveraging WordNet as a foundational semantic ontology library, the research proposes the construction of an extensive hierarchical scene semantic tree, enriching 2D/3D scenes with encoded semantic information. The methodologies for semantic similarity computation utilize this semantic tree to bridge the semantic disparity between 2D …


Modeling The Spatiotemporal Variations Of The Magnetic Field In Active Regions On The Sun Using Deep Neural Networks, Godwill Asare Mensah Mensah May 2024

Modeling The Spatiotemporal Variations Of The Magnetic Field In Active Regions On The Sun Using Deep Neural Networks, Godwill Asare Mensah Mensah

Open Access Theses & Dissertations

Solar active regions are areas on the Sun's surface that have especially strong magnetic fields. Active regions are usually linked to a number of phenomena that can have serious detrimental consequences on technology and, in turn, human life. Examples of these phenomena include solar flares and coronal mass ejections, or CMEs. The precise predictionof solar flares and coronal mass ejections is still an open problem since the fundamental processes underpinning the formation and development of active regions are still not well understood. One key area of research at the intersection of solar physics and artificial intelligence is deriving insights from …


Detection And Classification Of Diabetic Retinopathy Using Deep Learning Models, Aishat Olatunji May 2024

Detection And Classification Of Diabetic Retinopathy Using Deep Learning Models, Aishat Olatunji

Electronic Theses and Dissertations

Healthcare analytics leverages extensive patient data for data-driven decision-making, enhancing patient care and results. Diabetic Retinopathy (DR), a complication of diabetes, stems from damage to the retina’s blood vessels. It can affect both type 1 and type 2 diabetes patients. Ophthalmologists employ retinal images for accurate DR diagnosis and severity assessment. Early detection is crucial for preserving vision and minimizing risks. In this context, we utilized a Kaggle dataset containing patient retinal images, employing Python’s versatile tools. Our research focuses on DR detection using deep learning techniques. We used a publicly available dataset to apply our proposed neural network and …


Crash Detecting System Using Deep Learning, Yogesh Reddy Muddam May 2024

Crash Detecting System Using Deep Learning, Yogesh Reddy Muddam

Electronic Theses, Projects, and Dissertations

Accidents pose a significant risk to both individual and property safety, requiring effective detection and response systems. This work introduces an accident detection system using a convolutional neural network (CNN), which provides an impressive accuracy of 86.40%. Trained on diverse data sets of images and videos from various online sources, the model exhibits complex accident detection and classification and is known for its prowess in image classification and visualization.

CNN ensures better accident detection in various scenarios and road conditions. This example shows its adaptability to a real-world accident scenario and enhances its effectiveness in detecting early events. A key …


Representation Learning For Generative Models With Applications To Healthcare, Astronautics, And Aviation, Van Minh Nguyen May 2024

Representation Learning For Generative Models With Applications To Healthcare, Astronautics, And Aviation, Van Minh Nguyen

Theses and Dissertations

This dissertation explores applications of representation learning and generative models to challenges in healthcare, astronautics, and aviation.

The first part investigates the use of Generative Adversarial Networks (GANs) to synthesize realistic electronic health record (EHR) data. An initial attempt at training a GAN on the MIMIC-IV dataset encountered stability and convergence issues, motivating a deeper study of 1-Lipschitz regularization techniques for Auxiliary Classifier GANs (AC-GANs). An extensive ablation study on the CIFAR-10 dataset found that Spectral Normalization is key for AC-GAN stability and performance, while Weight Clipping fails to converge without Spectral Normalization. Analysis of the training dynamics provided further …


Evaluation Of An End-To-End Radiotherapy Treatment Planning Pipeline For Prostate Cancer, Mohammad Daniel El Basha, Court Laurence, Carlos Eduardo Cardenas, Julianne Pollard-Larkin, Steven Frank, David T. Fuentes, Falk Poenisch, Zhiqian H. Yu May 2024

Evaluation Of An End-To-End Radiotherapy Treatment Planning Pipeline For Prostate Cancer, Mohammad Daniel El Basha, Court Laurence, Carlos Eduardo Cardenas, Julianne Pollard-Larkin, Steven Frank, David T. Fuentes, Falk Poenisch, Zhiqian H. Yu

Dissertations and Theses (Open Access)

Radiation treatment planning is a crucial and time-intensive process in radiation therapy. This planning involves carefully designing a treatment regimen tailored to a patient’s specific condition, including the type, location, and size of the tumor with reference to surrounding healthy tissues. For prostate cancer, this tumor may be either local, locally advanced with extracapsular involvement, or extend into the pelvic lymph node chain. Automating essential parts of this process would allow for the rapid development of effective treatment plans and better plan optimization to enhance tumor control for better outcomes.

The first objective of this work, to automate the treatment …


Pedestrian Pathing Prediction Using Complex Contextual Behavioral Data In High Foot Traffic Settings, Laurel Bingham May 2024

Pedestrian Pathing Prediction Using Complex Contextual Behavioral Data In High Foot Traffic Settings, Laurel Bingham

All Graduate Theses and Dissertations, Fall 2023 to Present

Ensuring the safe integration of autonomous vehicles into real-world environments requires a comprehensive understanding of pedestrian behavior. This study addresses the challenge of predicting the movement and crossing intentions of pedestrians, a crucial aspect in the development of fully autonomous vehicles.

The research focuses on leveraging Honda's TITAN dataset, comprising 700 unique clips captured by moving vehicles in high-foot-traffic areas of Tokyo, Japan. Each clip provides detailed contextual information, including human-labeled tags for individuals and vehicles, encompassing attributes such as age, motion status, and communicative actions. Long Short-Term Memory (LSTM) networks were employed and trained on various combinations of contextual …


Deep Learning In Indus Valley Script Digitization, Deva Munikanta Reddy Atturu May 2024

Deep Learning In Indus Valley Script Digitization, Deva Munikanta Reddy Atturu

Theses and Dissertations

This research introduces ASR-net(Ancient Script Recognition), a groundbreaking system that automatically digitizes ancient Indus seals by converting them into coded text, similar to Optical Character Recognition for modern languages. ASR-net, with an 95% success rate in identifying individual symbols, aims to address the crucial need for automated techniques in deciphering the enigmatic Indus script. Initially Yolov3 is utilized to create the bounding boxes around each graphemes present in the Indus Valley Seal. In addition to that we created M-net(Mahadevan) model to encode the graphemes. Beyond digitization, the paper proposes a new research challenge called the Motif Identification Problem (MIP) related …


An Intelligent Framework Towards Fully Autonomous Driving Fueled By Smart Roads, Muhammad Jalal Khan Apr 2024

An Intelligent Framework Towards Fully Autonomous Driving Fueled By Smart Roads, Muhammad Jalal Khan

Thesis/ Dissertation Defenses

Autonomous vehicles (AVs) are transforming next-generation autonomous mobility. Such vehicles promise to increase road safety, improve traffic efficiency, reduce vehicle emissions, and enhance mobility. The development of AVs involves the integration of various disciplines and technologies, i.e., sensors, communication, computation, and artificial intelligence (AI), to achieve higher levels of autonomous driving (AD). The main objective of this dissertation is to design and develop a novel approach for achieving higher levels of automation in AD through an end-to-end intelligent framework. This involves addressing the challenges of technological augmentation of road infrastructure to support intelligent transport system (ITS) services, service satisfaction in …