Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Artificial Intelligence and Robotics (31)
- Data Science (13)
- Medicine and Health Sciences (13)
- Engineering (10)
- Social and Behavioral Sciences (7)
-
- Statistics and Probability (7)
- Applied Statistics (6)
- Medical Specialties (6)
- Computer Engineering (5)
- Mathematics (5)
- Numerical Analysis and Scientific Computing (5)
- Other Computer Engineering (5)
- Theory and Algorithms (5)
- Databases and Information Systems (4)
- Diseases (4)
- Electrical and Computer Engineering (4)
- Life Sciences (4)
- Analysis (3)
- Analytical, Diagnostic and Therapeutic Techniques and Equipment (3)
- Arts and Humanities (3)
- Business (3)
- Categorical Data Analysis (3)
- Environmental Indicators and Impact Assessment (3)
- Environmental Monitoring (3)
- Environmental Sciences (3)
- Information Security (3)
- Oncology (3)
- Institution
-
- Chapman University (15)
- San Jose State University (6)
- California Polytechnic State University, San Luis Obispo (3)
- Southern Methodist University (3)
- University of Louisville (3)
-
- Boise State University (2)
- College of Saint Benedict and Saint John's University (2)
- DePaul University (2)
- Georgia Southern University (2)
- Louisiana State University (2)
- University of Denver (2)
- University of Kentucky (2)
- University of Nevada, Las Vegas (2)
- Virginia Commonwealth University (2)
- Western Kentucky University (2)
- Association of Arab Universities (1)
- Claremont Colleges (1)
- Clemson University (1)
- Embry-Riddle Aeronautical University (1)
- Emory University School of Law (1)
- Grand Valley State University (1)
- Kennesaw State University (1)
- Marshall University (1)
- Michigan Technological University (1)
- Murray State University (1)
- Portland State University (1)
- Rowan University (1)
- SAE Institute Australasia (1)
- Technological University Dublin (1)
- Thomas Jefferson University (1)
- Publication Year
- Publication
-
- Master's Projects (6)
- Mathematics, Physics, and Computer Science Faculty Articles and Research (6)
- Electronic Theses and Dissertations (4)
- Engineering Faculty Articles and Research (4)
- Master's Theses (3)
-
- SMU Data Science Review (3)
- Boise State University Theses and Dissertations (2)
- College of Computing and Digital Media Dissertations (2)
- College of Graduate Studies: Theses & Dissertations (2)
- Theses and Dissertations (2)
- All College Thesis Program, 2016-2019 (1)
- All Dissertations (1)
- CMC Senior Theses (1)
- CSB and SJU Distinguished Thesis (1)
- Computational and Data Sciences (MS) Theses (1)
- Computational and Data Sciences (PhD) Dissertations (1)
- Department of Radiology Faculty Papers (1)
- Dissertations and Theses (1)
- Dissertations, Master's Theses and Master's Reports (1)
- Economics Faculty Articles and Research (1)
- Electrical & Computer Engineering Faculty Research (1)
- Faculty & Staff Research and Creative Activity (1)
- Faculty Articles (1)
- Faculty Publications (1)
- Graduate Masters Theses (1)
- Graduate Student Theses, Dissertations, & Professional Papers (1)
- Graduate Theses, Dissertations, and Problem Reports (ETD) (1)
- Hadhramout University Journal of Natural & Applied Sciences (1)
- Institute for ECHO Articles and Research (1)
- Irish Communication Review (1)
- Publication Type
Articles 31 - 60 of 70
Full-Text Articles in Other Computer Sciences
Scite: The Next Generation Of Citations, Sean Rife, Domenic Rosati, Joshua M. Nicholson
Scite: The Next Generation Of Citations, Sean Rife, Domenic Rosati, Joshua M. Nicholson
Faculty & Staff Research and Creative Activity
Key points
- While the importance of citation context has long been recognized, simple citation counts remain as a crude measure of importance.
- Providing citation context should support the publication of careful science instead of headline‐grabbing and salami‐sliced non‐replicable studies.
- Machine learning has enabled the extraction of citation context for the first time, and made the classification of citation types at scale possible.
Ensemble Protein Inference Evaluation, Kyle Lee Lucke
Ensemble Protein Inference Evaluation, Kyle Lee Lucke
Graduate Student Theses, Dissertations, & Professional Papers
The Protein inference problem is becoming an increasingly important tool that aids in the characterization of complex proteomes and analysis of complex protein samples. In bottom-up shotgun proteomics experiments the metrics for evaluation (like AUC and calibration error) are based on an often imperfect target-decoy database. These metrics make the inherent assumption that all of the proteins in the target set are present in the sample being analyzed. In general, this is not the case, they are typically a mix of present and absent proteins. To objectively evaluate inference methods, protein standard datasets are used. These datasets are special in …
Unsupervised Structural Graph Node Representation Learning, Mikel Joaristi
Unsupervised Structural Graph Node Representation Learning, Mikel Joaristi
Boise State University Theses and Dissertations
Unsupervised Graph Representation Learning methods learn a numerical representation of the nodes in a graph. The generated representations encode meaningful information about the nodes' properties, making them a powerful tool for tasks in many areas of study, such as social sciences, biology or communication networks. These methods are particularly interesting because they facilitate the direct use of standard Machine Learning models on graphs. Graph representation learning methods can be divided into two main categories depending on the information they encode, methods preserving the nodes connectivity information, and methods preserving nodes' structural information. Connectivity-based methods focus on encoding relationships between nodes, …
Exploring The Eating Disorder Examination Questionnaire, Clinical Impairment Assessment, And Autism Quotient To Identify Eating Disorder Vulnerability: A Cluster Analysis, Natalia Stewart Rosenfield, Erik Linstead
Exploring The Eating Disorder Examination Questionnaire, Clinical Impairment Assessment, And Autism Quotient To Identify Eating Disorder Vulnerability: A Cluster Analysis, Natalia Stewart Rosenfield, Erik Linstead
Engineering Faculty Articles and Research
Eating disorders are very complicated and many factors play a role in their manifestation. Furthermore, due to the variability in diagnosis and symptoms, treatment for an eating disorder is unique to the individual. As a result, there are numerous assessment tools available, which range from brief survey questionnaires to in-depth interviews conducted by a professional. One of the many benefits to using machine learning is that it offers new insight into datasets that researchers may not previously have, particularly when compared to traditional statistical methods. The aim of this paper was to employ k-means clustering to explore the Eating Disorder …
Improving Spellchecking For Children: Correction And Design, Brody Downs
Improving Spellchecking For Children: Correction And Design, Brody Downs
Boise State University Theses and Dissertations
Children commonly use software applications such as search engines and word processors in the classroom environment. However, a major barrier to using these programs successfully is the ability of children to type and spell effectively. While many programs make use of spellcheckers to provide spelling corrections to their users, they are designed for more traditional users (i.e., adults) and have proven inadequate for children. The aims of this work is twofold: first, to address the types of spelling errors children make by researching, developing, and evaluating algorithms to generate and rank candidate spelling suggestions; and second, to evaluate the impact …
Critical Media, Information, And Digital Literacy: Increasing Understanding Of Machine Learning Through An Interdisciplinary Undergraduate Course, Barbara R. Burke, Elena Machkasova
Critical Media, Information, And Digital Literacy: Increasing Understanding Of Machine Learning Through An Interdisciplinary Undergraduate Course, Barbara R. Burke, Elena Machkasova
Irish Communication Review
Widespread use of Artificial Intelligence in all areas of today’s society creates a unique problem: algorithms used in decision-making are generally not understandable to those without a background in data science. Thus, those who use out-of-the-box Machine Learning (ML) approaches in their work and those affected by these approaches are often not in a position to analyze their outcomes and applicability.
Our paper describes and evaluates our undergraduate course at the University of Minnesota Morris, which fosters understanding of the main ideas behind ML. With Communication, Media & Rhetoric and Computer Science faculty expertise, students from a variety of majors, …
A Machine Learning Approach To Delineating Neighborhoods From Geocoded Appraisal Data, Rao Hamza Ali, Josh Graves, Stanley Wu, Jenny Lee, Erik Linstead
A Machine Learning Approach To Delineating Neighborhoods From Geocoded Appraisal Data, Rao Hamza Ali, Josh Graves, Stanley Wu, Jenny Lee, Erik Linstead
Engineering Faculty Articles and Research
Identification of neighborhoods is an important, financially-driven topic in real estate. It is known that the real estate industry uses ZIP (postal) codes and Census tracts as a source of land demarcation to categorize properties with respect to their price. These demarcated boundaries are static and are inflexible to the shift in the real estate market and fail to represent its dynamics, such as in the case of an up-and-coming residential project. Delineated neighborhoods are also used in socioeconomic and demographic analyses where statistics are computed at a neighborhood level. Current practices of delineating neighborhoods have mostly ignored the information …
Pathways To The Native Storyteller: A Method To Enable Computational Story Understanding, Aramide O. Kehinde
Pathways To The Native Storyteller: A Method To Enable Computational Story Understanding, Aramide O. Kehinde
College of Computing and Digital Media Dissertations
The primary objective of this thesis is to develop a method that uses machine learning algorithms to enable computational story understanding. This research is conducted with the aim of establishing a system called the Native Storyteller that plans and creates storytelling experiences for human users. The paper first establishes the desired capabilities of the system and then deep dives into how to enable story understanding, which is the core ability the system needs to function. As such, the research places emphasis on natural language processing and its application to solving key problems in this context. Namely, machine representation of story …
Using Color Thresholding And Contouring To Understand Coral Reef Biodiversity, Scott Vuong Tran
Using Color Thresholding And Contouring To Understand Coral Reef Biodiversity, Scott Vuong Tran
Master's Projects
This paper presents research outcomes of understanding coral reef biodiversity through the usage of various computer vision applications and techniques. It aims to help further analyze and understand the coral reef biodiversity through the usage of color thresholding and contouring onto images of the ARMS plates to extract groups of microorganisms based on color. The results are comparable to the manual markup tool developed to do the same tasks and shows that the manual process can be sped up using computer vision. The paper presents an automated way to extract groups of microorganisms based on color without the use of …
Understanding Impact Of Twitter Feed On Bitcoin Price And Trading Patterns, Ashrit Deebadi
Understanding Impact Of Twitter Feed On Bitcoin Price And Trading Patterns, Ashrit Deebadi
Master's Projects
‘‘Cryptocurrency trading was one of the most exciting jobs of 2017’’. ‘‘Bit- coin’’,‘‘Blockchain’’, ‘‘Bitcoin Trading’’ were the most searched words in Google during 2017. High return on investment has attracted many people towards this crypto market. Existing research has shown that the trading price is completely based on speculation, and its trading volume is highly impacted by news media. This paper discusses the existing work to evaluate the sentiment and price of the cryptocurrency, the issues with the current trading models. It builds possible solutions to understand better the semantic orientation of text by comparing different machine learning techniques and …
Ml-Medic: A Preliminary Study Of An Interactive Visual Analysis Tool Facilitating Clinical Applications Of Machine Learning For Precision Medicine, Laura Stevens, David Kao, Jennifer Hall, Carsten Görg, Kaitlyn Abdo, Erik Linstead
Ml-Medic: A Preliminary Study Of An Interactive Visual Analysis Tool Facilitating Clinical Applications Of Machine Learning For Precision Medicine, Laura Stevens, David Kao, Jennifer Hall, Carsten Görg, Kaitlyn Abdo, Erik Linstead
Engineering Faculty Articles and Research
Accessible interactive tools that integrate machine learning methods with clinical research and reduce the programming experience required are needed to move science forward. Here, we present Machine Learning for Medical Exploration and Data-Inspired Care (ML-MEDIC), a point-and-click, interactive tool with a visual interface for facilitating machine learning and statistical analyses in clinical research. We deployed ML-MEDIC in the American Heart Association (AHA) Precision Medicine Platform to provide secure internet access and facilitate collaboration. ML-MEDIC’s efficacy for facilitating the adoption of machine learning was evaluated through two case studies in collaboration with clinical domain experts. A domain expert review was also …
Novel Inference Methods For Generalized Linear Models Using Shrinkage Priors And Data Augmentation., Arinjita Bhattacharyya
Novel Inference Methods For Generalized Linear Models Using Shrinkage Priors And Data Augmentation., Arinjita Bhattacharyya
Electronic Theses and Dissertations
Generalized linear models have broad applications in biostatistics and sociology. In a regression setup, the main target is to find a relevant set of predictors out of a large collection of covariates. Sparsity is the assumption that only a few of these covariates in a regression setup have a meaningful correlation with an outcome variate of interest. Sparsity is incorporated by regularizing the irrelevant slopes towards zero without changing the relevant predictors and keeping the resulting inferences intact. Frequentist variable selection and sparsity are addressed by popular techniques like Lasso, Elastic Net. Bayesian penalized regression can tackle the curse of …
An Analysis Of The Success Of Farmers Markets In Kentucky Using Logistic Regression And Support Vector Machines, Jeron Russell
An Analysis Of The Success Of Farmers Markets In Kentucky Using Logistic Regression And Support Vector Machines, Jeron Russell
Mahurin Honors College Capstone Experience/Thesis Projects
The purpose of this research is to look at the relationship that market-specific, economic, and demographic variables have with the success of farmers markets in Kentucky. It additionally seeks to build a tool for predicting farmers market success that could be used by policy makers to aid in decision-making processes concerning farmers markets. Logistic regression and Support Vector Machines (SVMs) are used on data acquired from the Kentucky Department of Agriculture and the American Community Survey in order to analyze the data in a traditional statistical approach as well as a machine learning approach. The results included an SVM model …
A Probabilistic Machine Learning Framework For Cloud Resource Selection On The Cloud, Syeduzzaman Khan
A Probabilistic Machine Learning Framework For Cloud Resource Selection On The Cloud, Syeduzzaman Khan
University of the Pacific Theses and Dissertations
The execution of the scientific applications on the Cloud comes with great flexibility, scalability, cost-effectiveness, and substantial computing power. Market-leading Cloud service providers such as Amazon Web service (AWS), Azure, Google Cloud Platform (GCP) offer various general purposes, memory-intensive, and compute-intensive Cloud instances for the execution of scientific applications. The scientific community, especially small research institutions and undergraduate universities, face many hurdles while conducting high-performance computing research in the absence of large dedicated clusters. The Cloud provides a lucrative alternative to dedicated clusters, however a wide range of Cloud computing choices makes the instance selection for the end-users. This thesis …
Implementation Considerations For Mitigating Bias In Supervised Machine Learning, Bardia Bijani Aval
Implementation Considerations For Mitigating Bias In Supervised Machine Learning, Bardia Bijani Aval
CSB and SJU Distinguished Thesis
Machine Learning (ML) is an important component of computer science and a mainstream way of making sense of large amounts of data. Although the technology is establishing new possibilities in different fields, there are also problems to consider, one of which is bias. Due to the inductive reasoning of ML algorithms in creating mathematical models, the predictions and trends found by the models will never necessarily be true – just more or less probable. Knowing this, it is unreasonable for us to expect the applied deductive reasoning of these models to ever be fully unbiased. Therefore, it is important that …
Sparsity And Weak Supervision In Quantum Machine Learning, Seyran Saeedi
Sparsity And Weak Supervision In Quantum Machine Learning, Seyran Saeedi
Theses and Dissertations
Quantum computing is an interdisciplinary field at the intersection of computer science, mathematics, and physics that studies information processing tasks on a quantum computer. A quantum computer is a device whose operations are governed by the laws of quantum mechanics. As building quantum computers is nearing the era of commercialization and quantum supremacy, it is essential to think of potential applications that we might benefit from. Among many applications of quantum computation, one of the emerging fields is quantum machine learning. We focus on predictive models for binary classification and variants of Support Vector Machines that we expect to be …
Detecting Myocardial Infarctions Using Machine Learning Methods, Aniruddh Mathur
Detecting Myocardial Infarctions Using Machine Learning Methods, Aniruddh Mathur
Master's Projects
Myocardial Infarction (MI), commonly known as a heart attack, occurs when one of the three major blood vessels carrying blood to the heart get blocked, causing the death of myocardial (heart) cells. If not treated immediately, MI may cause cardiac arrest, which can ultimately cause death. Risk factors for MI include diabetes, family history, unhealthy diet and lifestyle. Medical treatments include various types of drugs and surgeries which can prove very expensive for patients due to high healthcare costs. Therefore, it is imperative that MI is diagnosed at the right time. Electrocardiography (ECG) is commonly used to detect MI. ECG …
Editorial: Machine Learning In Biomolecular Simulations, Gennady M. Verkhivker, Vojtech Spiwok, Francesco Luigi Gervasio
Editorial: Machine Learning In Biomolecular Simulations, Gennady M. Verkhivker, Vojtech Spiwok, Francesco Luigi Gervasio
Mathematics, Physics, and Computer Science Faculty Articles and Research
"Interest in machine learning is growing in all fields of science, industry, and business. This interest was not primarily initiated by new theoretical findings. Interestingly, the theoretical basis of the majority of machine learning techniques, such as artificial neural networks, decision trees, or kernel methods, have been known for a relatively long time. Instead, there are other effects that triggered the recent boom of machine learning."
Predicting Switch-Like Behavior In Proteins Using Logistic Regression On Sequence-Based Descriptors, Benjamin Strauss
Predicting Switch-Like Behavior In Proteins Using Logistic Regression On Sequence-Based Descriptors, Benjamin Strauss
Master's Projects
Ligands can bind at specific protein locations, inducing conformational changes such as those involving secondary structure. Identifying these possible switches from sequence, including homology, is an important ongoing area of research. We attempt to predict possible secondary structure switches from sequence in proteins using machine learning, specifically a logistic regression approach with 48 N-acetyltransferases as our learning set and 5 sirtuins as our test set. Validated residue binary assignments of 0 (no change in secondary structure) and 1 (change in secondary structure) were determined (DSSP) from 3D X-ray structures for sets of virtually identical chains crystallized under different conditions. Our …
Identifying Depression In The National Health And Nutrition Examination Survey Data Using A Deep Learning Algorithm, Jihoon Oh, Kyongsik Yun, Uri Maoz, Tae-Suk Kim, Jeong-Ho Chae
Identifying Depression In The National Health And Nutrition Examination Survey Data Using A Deep Learning Algorithm, Jihoon Oh, Kyongsik Yun, Uri Maoz, Tae-Suk Kim, Jeong-Ho Chae
Psychology Faculty Articles and Research
Background
As depression is the leading cause of disability worldwide, large-scale surveys have been conducted to establish the occurrence and risk factors of depression. However, accurately estimating epidemiological factors leading up to depression has remained challenging. Deep-learning algorithms can be applied to assess the factors leading up to prevalence and clinical manifestations of depression.
Methods
Customized deep-neural-network and machine-learning classifiers were assessed using survey data from 19,725 participants from the NHANES database (from 1999 through 2014) and 4949 from the South Korea NHANES (K-NHANES) database in 2014.
Results
A deep-learning algorithm showed area under the receiver operating characteristic curve (AUCs) …
Evaluating Projections And Developing Projection Models For Daily Fantasy Basketball, Eric C. Evangelista
Evaluating Projections And Developing Projection Models For Daily Fantasy Basketball, Eric C. Evangelista
Master's Theses
Daily fantasy sports (DFS) has grown in popularity with millions of participants throughout the world. However, studies have shown that most profits from DFS contests are won by only a small percentage of players. This thesis addresses the challenges faced by DFS participants by evaluating sources that provide player projections for NBA DFS contests and by developing machine learning models that produce competitive player projections.
External sources are evaluated by constructing daily lineups based on the projections offered and evaluating those lineups in the context of all potential lineups, as well as those submitted by participants in competitive FanDuel DFS …
Using Computer Vision To Quantify Coral Reef Biodiversity, Niket Bhodia
Using Computer Vision To Quantify Coral Reef Biodiversity, Niket Bhodia
Master's Projects
The preservation of the world’s oceans is crucial to human survival on this planet, yet we know too little to begin to understand anthropogenic impacts on marine life. This is especially true for coral reefs, which are the most diverse marine habitat per unit area (if not overall) as well as the most sensitive. To address this gap in knowledge, simple field devices called autonomous reef monitoring structures (ARMS) have been developed, which provide standardized samples of life from these complex ecosystems. ARMS have now become successful to the point that the amount of data collected through them has outstripped …
Stock Market Prediction Using Ensemble Of Graph Theory, Machine Learning And Deep Learning Models, Pratik Patil
Stock Market Prediction Using Ensemble Of Graph Theory, Machine Learning And Deep Learning Models, Pratik Patil
Master's Projects
Efficient Market Hypothesis (EMH) is the cornerstone of the modern financial theory and it states that it is impossible to predict the price of any stock using any trend, fundamental or technical analysis. Stock trading is one of the most important activities in the world of finance. Stock price prediction has been an age-old problem and many researchers from academia and business have tried to solve it using many techniques ranging from basic statistics to machine learning using relevant information such as news sentiment and historical prices. Even though some studies claim to get prediction accuracy higher than a random …
The New Legal Landscape For Text Mining And Machine Learning, Matthew Sag
The New Legal Landscape For Text Mining And Machine Learning, Matthew Sag
Faculty Articles
Now that the dust has settled on the Authors Guild cases, this Article takes stock of the legal context for TDM research in the United States. This reappraisal begins in Part I with an assessment of exactly what the Authors Guild cases did and did not establish with respect to the fair use status of text mining. Those cases held unambiguously that reproducing copyrighted works as one step in the process of knowledge discovery through text data mining was transformative, and thus ultimately a fair use of those works. Part I explains why those rulings followed inexorably from copyright's most …
Ifocus: A Framework For Non-Intrusive Assessment Of Student Attention Level In Classrooms, Narayanan Veliyath
Ifocus: A Framework For Non-Intrusive Assessment Of Student Attention Level In Classrooms, Narayanan Veliyath
College of Graduate Studies: Theses & Dissertations
The process of learning is not merely determined by what the instructor teaches, but also by how the student receives that information. An attentive student will naturally be more open to obtaining knowledge than a bored or frustrated student. In recent years, tools such as skin temperature measurements and body posture calculations have been developed for the purpose of determining a student's affect, or emotional state of mind. However, measuring eye-gaze data is particularly noteworthy in that it can collect measurements non-intrusively, while also being relatively simple to set up and use. This paper details how data obtained from such …
Predicting Post-Procedural Complications Using Neural Networks On Mimic-Iii Data, Namratha Mohan
Predicting Post-Procedural Complications Using Neural Networks On Mimic-Iii Data, Namratha Mohan
LSU Master's Theses
The primary focus of this paper is the creation of a Machine Learning based algorithm for the analysis of large health based data sets. Our input was extracted from MIMIC-III, a large Health Record database of more than 40,000 patients. The main question was to predict if a patient will have complications during certain specified procedures performed in the hospital. These events are denoted by the icd9 code 996 in the individuals' health record. The output of our predictive model is a binary variable which outputs the value 1 if the patient is diagnosed with the specific complication or 0 …
A Transfer Learning Approach For Sentiment Classification., Omar Abdelwahab
A Transfer Learning Approach For Sentiment Classification., Omar Abdelwahab
Electronic Theses and Dissertations
The idea of developing machine learning systems or Artificial Intelligence agents that would learn from different tasks and be able to accumulate that knowledge with time so that it functions successfully on a new task that it has not seen before is an idea and a research area that is still being explored. In this work, we will lay out an algorithm that allows a machine learning system or an AI agent to learn from k different domains then uses some or no data from the new task for the system to perform strongly on that new task. In order …
Cleaver: Classification Of Everyday Activities Via Ensemble Recognizers, Samantha Hsu
Cleaver: Classification Of Everyday Activities Via Ensemble Recognizers, Samantha Hsu
Master's Theses
Physical activity can have immediate and long-term benefits on health and reduce the risk for chronic diseases. Valid measures of physical activity are needed in order to improve our understanding of the exact relationship between physical activity and health. Activity monitors have become a standard for measuring physical activity; accelerometers in particular are widely used in research and consumer products because they are objective, inexpensive, and practical. Previous studies have experimented with different monitor placements and classification methods. However, the majority of these methods were developed using data collected in controlled, laboratory-based settings, which is not reliably representative of real …
Cryptovisor: A Cryptocurrency Advisor Tool, Matthew Baldree, Paul Widhalm, Brandon Hill, Matteo Ortisi
Cryptovisor: A Cryptocurrency Advisor Tool, Matthew Baldree, Paul Widhalm, Brandon Hill, Matteo Ortisi
SMU Data Science Review
In this paper, we present a tool that provides trading recommendations for cryptocurrency using a stochastic gradient boost classifier trained from a model labeled by technical indicators. The cryptocurrency market is volatile due to its infancy and limited size making it difficult for investors to know when to enter, exit, or stay in the market. Therefore, a tool is needed to provide investment recommendations for investors. We developed such a tool to support one cryptocurrency, Bitcoin, based on its historical price and volume data to recommend a trading decision for today or past days. This tool is 95.50% accurate with …
A Multiple Classifier System For Predicting Best-Selling Amazon Products, Michael Kranzlein
A Multiple Classifier System For Predicting Best-Selling Amazon Products, Michael Kranzlein
Master of Science in Computer Science Theses
In this work, I examine a dataset of Amazon product metadata and propose a heterogeneous multiple classifier system for the task of identifying best-selling products in multiple categories. This system of classifiers consumes the product description and the featured product image as input and feeds them through binary classifiers of the following types: Convolutional Neural Network, Na¨ıve Bayes, Random Forest, Ridge Regression, and Support Vector Machine. While each individual model is largely successful in identifying best-selling products from non best-selling products and from worst-selling products, the multiple classifier system is shown to be stronger than any individual model in the …