Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Computer Sciences (21)
- Data Science (15)
- Statistical Models (13)
- Artificial Intelligence and Robotics (10)
- Mathematics (8)
-
- Statistical Methodology (8)
- Biostatistics (7)
- Business (7)
- Applied Mathematics (6)
- Life Sciences (6)
- Multivariate Analysis (6)
- Other Computer Sciences (6)
- Social and Behavioral Sciences (6)
- Categorical Data Analysis (5)
- Engineering (5)
- Longitudinal Data Analysis and Time Series (5)
- Analysis (4)
- Animal Sciences (4)
- Medicine and Health Sciences (4)
- Economics (3)
- Environmental Sciences (3)
- Numerical Analysis and Computation (3)
- Probability (3)
- Animal Studies (2)
- Aquaculture and Fisheries (2)
- Bioinformatics (2)
- Clinical Trials (2)
- Institution
-
- Southern Methodist University (4)
- California Polytechnic State University, San Luis Obispo (3)
- Western Kentucky University (3)
- Air Force Institute of Technology (2)
- City University of New York (CUNY) (2)
-
- Dartmouth College (2)
- Georgia Southern University (2)
- Kennesaw State University (2)
- Louisiana State University (2)
- Marshall University (2)
- East Tennessee State University (1)
- Louisiana Tech University (1)
- Loyola Marymount University and Loyola Law School (1)
- Michigan Technological University (1)
- Minnesota State University, Mankato (1)
- Missouri State University (1)
- Missouri University of Science and Technology (1)
- Murray State University (1)
- New Jersey Institute of Technology (1)
- University of Arkansas, Fayetteville (1)
- University of Denver (1)
- University of Louisville (1)
- University of Montana (1)
- University of Nebraska - Lincoln (1)
- University of New Hampshire (1)
- University of New Mexico (1)
- University of Texas Rio Grande Valley (1)
- Virginia Commonwealth University (1)
- Publication Year
- Publication
-
- SMU Data Science Review (4)
- Master's Theses (3)
- Theses and Dissertations (3)
- College of Graduate Studies: Theses & Dissertations (2)
- Economics Faculty Publications (2)
-
- Electronic Theses and Dissertations (2)
- Theses, Dissertations and Capstones (2)
- All Graduate Theses, Dissertations, and Other Capstone Projects (1)
- Computer Science Senior Theses (1)
- Department of Agricultural and Biological Systems Engineering: Faculty Publications (1)
- Dissertations (1)
- Dissertations, Master's Theses and Master's Reports (1)
- Dissertations, Theses, and Capstone Projects (1)
- Doctoral Dissertations (1)
- Electrical and Computer Engineering ETDs (1)
- Faculty Articles (1)
- Faculty Publications (1)
- Graduate Student Theses, Dissertations, & Professional Papers (1)
- Graduate Theses/Dissertations (1)
- Honors Theses and Capstones (1)
- Honors Thesis (1)
- Industrial Engineering Undergraduate Honors Theses (1)
- LSU Doctoral Dissertations (1)
- LSU Master's Theses (1)
- Mahurin Honors College Capstone Experience/Thesis Projects (1)
- Mathematics and Statistics Faculty Research & Creative Works (1)
- Murray State Theses and Dissertations (1)
- Open Educational Resources (1)
- Quantitative Social Science Undergraduate Senior Theses (1)
- Senior Design Project For Engineers (1)
- Publication Type
- File Type
Articles 31 - 42 of 42
Full-Text Articles in Applied Statistics
Novel Inference Methods For Generalized Linear Models Using Shrinkage Priors And Data Augmentation., Arinjita Bhattacharyya
Novel Inference Methods For Generalized Linear Models Using Shrinkage Priors And Data Augmentation., Arinjita Bhattacharyya
Electronic Theses and Dissertations
Generalized linear models have broad applications in biostatistics and sociology. In a regression setup, the main target is to find a relevant set of predictors out of a large collection of covariates. Sparsity is the assumption that only a few of these covariates in a regression setup have a meaningful correlation with an outcome variate of interest. Sparsity is incorporated by regularizing the irrelevant slopes towards zero without changing the relevant predictors and keeping the resulting inferences intact. Frequentist variable selection and sparsity are addressed by popular techniques like Lasso, Elastic Net. Bayesian penalized regression can tackle the curse of …
An Analysis Of The Success Of Farmers Markets In Kentucky Using Logistic Regression And Support Vector Machines, Jeron Russell
An Analysis Of The Success Of Farmers Markets In Kentucky Using Logistic Regression And Support Vector Machines, Jeron Russell
Mahurin Honors College Capstone Experience/Thesis Projects
The purpose of this research is to look at the relationship that market-specific, economic, and demographic variables have with the success of farmers markets in Kentucky. It additionally seeks to build a tool for predicting farmers market success that could be used by policy makers to aid in decision-making processes concerning farmers markets. Logistic regression and Support Vector Machines (SVMs) are used on data acquired from the Kentucky Department of Agriculture and the American Community Survey in order to analyze the data in a traditional statistical approach as well as a machine learning approach. The results included an SVM model …
Evaluating An Ordinal Output Using Data Modeling, Algorithmic Modeling, And Numerical Analysis, Martin Keagan Wynne Brown
Evaluating An Ordinal Output Using Data Modeling, Algorithmic Modeling, And Numerical Analysis, Martin Keagan Wynne Brown
Murray State Theses and Dissertations
Data and algorithmic modeling are two different approaches used in predictive analytics. The models discussed from these two approaches include the proportional odds logit model (POLR), the vector generalized linear model (VGLM), the classification and regression tree model (CART), and the random forests model (RF). Patterns in the data were analyzed using trigonometric polynomial approximations and Fast Fourier Transforms. Predictive modeling is used frequently in statistics and data science to find the relationship between the explanatory (input) variables and a response (output) variable. Both approaches prove advantageous in different cases depending on the data set. In our case, the data …
Habitat Associations And Reproduction Of Fishes On The Northwestern Gulf Of Mexico Shelf Edge, Elizabeth Marie Keller
Habitat Associations And Reproduction Of Fishes On The Northwestern Gulf Of Mexico Shelf Edge, Elizabeth Marie Keller
LSU Doctoral Dissertations
Several of the northwestern Gulf of Mexico (GOM) shelf-edge banks provide critical hard bottom habitat for coral and fish communities, supporting a wide diversity of ecologically and economically important species. These sites may be fish aggregation and spawning sites and provide important habitat for fish growth and reproduction. Already designated as habitat areas of particular concern, many of these banks are also under consideration for inclusion in the expansion of the Flower Garden Banks National Marine Sanctuary. This project aimed to gain a more comprehensive understanding of the communities and fish species on shelf-edge banks by way of gonad histology, …
Field Drilling Data Cleaning And Preparation For Data Analytics Applications, Daniel Cardoso Braga
Field Drilling Data Cleaning And Preparation For Data Analytics Applications, Daniel Cardoso Braga
LSU Master's Theses
Throughout the history of oil well drilling, service providers have been continuously striving to improve performance and reduce total drilling costs to operating companies. Despite constant improvement in tools, products, and processes, data science has not played a large part in oil well drilling. With the implementation of data science in the energy sector, companies have come to see significant value in efficiently processing the massive amounts of data produced by the multitude of internet of thing (IOT) sensors at the rig. The scope of this project is to combine academia and industry experience to analyze data from 13 different …
Data Patterns Discovery Using Unsupervised Learning, Rachel A. Lewis
Data Patterns Discovery Using Unsupervised Learning, Rachel A. Lewis
College of Graduate Studies: Theses & Dissertations
Self-care activities classification poses significant challenges in identifying children’s unique functional abilities and needs within the exceptional children healthcare system. The accuracy of diagnosing a child's self-care problem, such as toileting or dressing, is highly influenced by an occupational therapists’ experience and time constraints. Thus, there is a need for objective means to detect and predict in advance the self-care problems of children with physical and motor disabilities. We use clustering to discover interesting information from self-care problems, perform automatic classification of binary data, and discover outliers. The advantages are twofold: the advancement of knowledge on identifying self-care problems in …
Predicting National Basketball Association Success: A Machine Learning Approach, Adarsh Kannan, Brian Kolovich, Brandon Lawrence, Sohail Rafiqi
Predicting National Basketball Association Success: A Machine Learning Approach, Adarsh Kannan, Brian Kolovich, Brandon Lawrence, Sohail Rafiqi
SMU Data Science Review
In this paper, we present a machine learning based approach to projecting the success of National Basketball Association (NBA) draft prospects. With the proliferation of data, analytics have increasingly be- come a critical component in the assessment of professional and collegiate basketball players. We leverage player biometric data, college statistics, draft selection order, and positional breakdown as modelling features in our prediction algorithms. We found that a player's draft pick and their college statistics are the best predictors of their longevity in the National Basketball Association.
Cognitive Virtual Admissions Counselor, Kumar Raja Guvindan Raju, Cory Adams, Raghuram Srinivas
Cognitive Virtual Admissions Counselor, Kumar Raja Guvindan Raju, Cory Adams, Raghuram Srinivas
SMU Data Science Review
Abstract. In this paper, we present a cognitive virtual admissions counselor for the Master of Science in Data Science program at Southern Methodist University. The virtual admissions counselor is a system capable of providing potential students accurate information at the time that they want to know it. After the evaluation of multiple technologies, Amazon’s LEX was selected to serve as the core technology for the virtual counselor chatbot. Student surveys were leveraged to collect and generate training data to deploy the natural language capability. The cognitive virtual admissions counselor platform is currently capable of providing an end-to-end conversational dialog to …
Old English Character Recognition Using Neural Networks, Sattajit Sutradhar
Old English Character Recognition Using Neural Networks, Sattajit Sutradhar
College of Graduate Studies: Theses & Dissertations
Character recognition has been capturing the interest of researchers since the beginning of the twentieth century. While the Optical Character Recognition for printed material is very robust and widespread nowadays, the recognition of handwritten materials lags behind. In our digital era more and more historical, handwritten documents are digitized and made available to the general public. However, these digital copies of handwritten materials lack the automatic content recognition feature of their printed materials counterparts. We are proposing a practical, accurate, and computationally efficient method for Old English character recognition from manuscript images. Our method relies on a modern machine learning …
Empirical Methods For Predicting Student Retention- A Summary From The Literature, Matt Bogard
Empirical Methods For Predicting Student Retention- A Summary From The Literature, Matt Bogard
Economics Faculty Publications
The vast majority of the literature related to the empirical estimation of retention models includes a discussion of the theoretical retention framework established by Bean, Braxton, Tinto, Pascarella, Terenzini and others (see Bean, 1980; Bean, 2000; Braxton, 2000; Braxton et al, 2004; Chapman and Pascarella, 1983; Pascarell and Ternzini, 1978; St. John and Cabrera, 2000; Tinto, 1975) This body of research provides a starting point for the consideration of which explanatory variables to include in any model specification, as well as identifying possible data sources. The literature separates itself into two major camps including research related to the hypothesis testing …
Empirical Methods-A Review: With An Introduction To Data Mining And Machine Learning, Matt Bogard
Empirical Methods-A Review: With An Introduction To Data Mining And Machine Learning, Matt Bogard
Economics Faculty Publications
This presentation was part of a staff workshop focused on empirical methods and applied research. This includes a basic overview of regression with matrix algebra, maximum likelihood, inference, and model assumptions. Distinctions are made between paradigms related to classical statistical methods and algorithmic approaches. The presentation concludes with a brief discussion of generalization error, data partitioning, decision trees, and neural networks.
Machine Learning Approaches For Determining Effective Seeds For K -Means Algorithm, Kaveephong Lertwachara
Machine Learning Approaches For Determining Effective Seeds For K -Means Algorithm, Kaveephong Lertwachara
Doctoral Dissertations
In this study, I investigate and conduct an experiment on two-stage clustering procedures, hybrid models in simulated environments where conditions such as collinearity problems and cluster structures are controlled, and in real-life problems where conditions are not controlled. The first hybrid model (NK) is an integration between a neural network (NN) and the k-means algorithm (KM) where NN screens seeds and passes them to KM. The second hybrid (GK) uses a genetic algorithm (GA) instead of the neural network. Both NN and GA used in this study are in their simplest-possible forms.
In the simulated data sets, I investigate two …