Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Applied Statistics (13)
- Statistical Models (8)
- Engineering (7)
- Applied Mathematics (6)
- Statistical Methodology (6)
-
- Other Statistics and Probability (5)
- Mathematics (4)
- Social and Behavioral Sciences (4)
- Civil and Environmental Engineering (3)
- Computer Sciences (3)
- Multivariate Analysis (3)
- Numerical Analysis and Computation (3)
- Social Statistics (3)
- Business (2)
- Categorical Data Analysis (2)
- Discrete Mathematics and Combinatorics (2)
- Earth Sciences (2)
- Economics (2)
- Electrical and Computer Engineering (2)
- Engineering Physics (2)
- Law (2)
- Legal Studies (2)
- Life Sciences (2)
- Other Applied Mathematics (2)
- Other Civil and Environmental Engineering (2)
- Other Legal Studies (2)
- Physics (2)
- Institution
-
- East Tennessee State University (2)
- LSU New Orleans (2)
- Michigan Technological University (2)
- The University of Akron (2)
- City University of New York (CUNY) (1)
-
- Claremont Colleges (1)
- Colby College (1)
- Georgia Southern University (1)
- Gettysburg College (1)
- Illinois State University (1)
- Louisiana State University (1)
- Marshall University (1)
- Old Dominion University (1)
- Rose-Hulman Institute of Technology (1)
- Southern Methodist University (1)
- Stephen F. Austin State University (1)
- University of Montana (1)
- University of Nebraska - Lincoln (1)
- University of Nevada, Las Vegas (1)
- University of New Mexico (1)
- University of North Florida (1)
- University of Northern Iowa (1)
- Vanderbilt University Law School (1)
- Western Kentucky University (1)
- Keyword
-
- Probability (4)
- Naive Bayes (2)
- Nonparametric (2)
- <p>Probabilities.</p> <p>Paired comparisons (Statistics)</p> <p>Logistic regression analysis.</p> (1)
- Academic -- UNF -- Engineering; Particle Filtering; State Estimation; Water Resources; Groundwater Hydrology (1)
-
- Academic -- UNF -- Master of Science in Electrical Engineering; Dissertations (1)
- Actuarial Risk Assessment (1)
- Actuarial Risk Assessment Tool (1)
- Agricultural drought risk (1)
- Algorithm (1)
- Bag of Words (1)
- Bail (1)
- Bayesian Network (1)
- Bayesian analysis (1)
- Bias (1)
- Big data (1)
- Bond (1)
- Boosted Decision Trees (1)
- Bootstrap (1)
- Bootstrap simulation (1)
- Burden of proof (1)
- Cargo Tank Truck (1)
- Causal state splitting reconstruction (1)
- Cholinergic modulation (1)
- Climate-information (1)
- Clue (1)
- Clustering (1)
- College Basketball (1)
- College room allotment (1)
- Conditional Probability (1)
- Publication
-
- Electronic Theses and Dissertations (3)
- Dissertations, Master's Theses and Master's Reports (2)
- LSU New Orleans Theses and Dissertations (2)
- Williams Honors College, Honors Research Projects (2)
- College of Graduate Studies: Theses & Dissertations (1)
-
- Computer Science Faculty Publications (1)
- Department of Civil and Environmental Engineering: Dissertations, Theses, and Student Research (1)
- Graduate Student Theses, Dissertations, & Professional Papers (1)
- Honors Program Theses (1)
- Honors Theses (1)
- Journal of Humanistic Mathematics (1)
- LSU Doctoral Dissertations (1)
- Masters Theses & Specialist Projects (1)
- Mathematics & Statistics Faculty Publications (1)
- Publications and Research (1)
- Rose-Hulman Undergraduate Mathematics Journal (1)
- SMU Data Science Review (1)
- Shared Knowledge Conference (1)
- Student Research – Stevenson Center (1)
- Theses, Dissertations and Capstones (1)
- UNF Graduate Theses and Dissertations (1)
- UNLV Gaming Research & Review Journal (1)
- Vanderbilt Law School Faculty Publications (1)
- Publication Type
Articles 1 - 28 of 28
Full-Text Articles in Probability
Analysis Of Ranked Gene Tree Probability Distributions Under The Coalescent Process For Detecting Anomaly Zones, Anastasiia Kim
Analysis Of Ranked Gene Tree Probability Distributions Under The Coalescent Process For Detecting Anomaly Zones, Anastasiia Kim
Shared Knowledge Conference
In phylogenetic studies, gene trees are used to reconstruct species tree. Under the multispecies coalescent model, gene trees topologies may differ from that of species trees. The incorrect gene tree topology (one that does not match the species tree) that is more probable than the correct one is termed anomalous gene tree (AGT). Species trees that can generate such AGTs are said to be in the anomaly zone (AZ). In this region, the method of choosing the most common gene tree as the estimate of the species tree will be inconsistent and will converge to an incorrect species tree when …
Statistical Investigation Of Road And Railway Hazardous Materials Transportation Safety, Amirfarrokh Iranitalab
Statistical Investigation Of Road And Railway Hazardous Materials Transportation Safety, Amirfarrokh Iranitalab
Department of Civil and Environmental Engineering: Dissertations, Theses, and Student Research
Transportation of hazardous materials (hazmat) in the United States (U.S.) constituted 22.8% of the total tonnage transported in 2012 with an estimated value of more than 2.3 billion dollars. As such, hazmat transportation is a significant economic activity in the U.S. However, hazmat transportation exposes people and environment to the infrequent but potentially severe consequences of incidents resulting in hazmat release. Trucks and trains carried 63.7% of the hazmat in the U.S. in 2012 and are the major foci of this dissertation. The main research objectives were 1) identification and quantification of the effects of different factors on occurrence and …
Probabilities Involving Standard Trirectangular Tetrahedral Dice Rolls, Rulon Olmstead, Doneliezer Baize
Probabilities Involving Standard Trirectangular Tetrahedral Dice Rolls, Rulon Olmstead, Doneliezer Baize
Rose-Hulman Undergraduate Mathematics Journal
The goal is to be able to calculate probabilities involving irregular shaped dice rolls. Here it is attempted to model the probabilities of rolling standard tri-rectangular tetrahedral dice on a hard surface, such as a table top. The vertices and edges of a tetrahedron were projected onto the surface of a sphere centered at the center of mass of the tetrahedron. By calculating the surface areas bounded by the resultant geodesics, baseline probabilities were achieved. Using a 3D printer, dice were constructed of uniform density and the results of rolling them were recorded. After calculating the corresponding confidence intervals, the …
Season-Ahead Forecasting Of Water Storage And Irrigation Requirements – An Application To The Southwest Monsoon In India, Arun Ravindranath, Naresh Devineni, Upmanu Lall, Paulina Concha Larrauri
Season-Ahead Forecasting Of Water Storage And Irrigation Requirements – An Application To The Southwest Monsoon In India, Arun Ravindranath, Naresh Devineni, Upmanu Lall, Paulina Concha Larrauri
Publications and Research
Water risk management is a ubiquitous challenge faced by stakeholders in the water or agricultural sector. We present a methodological framework for forecasting water storage requirements and present an application of this methodology to risk assessment in India. The application focused on forecasting crop water stress for potatoes grown during the monsoon season in the Satara district of Maharashtra. Pre-season large-scale climate predictors used to forecast water stress were selected based on an exhaustive search method that evaluates for highest ranked probability skill score and lowest root-mean-squared error in a leave-one-out cross-validation mode. Adaptive forecasts were made in the years …
Yelp’S Review Filtering Algorithm, Yao Yao, Ivelin Angelov, Jack Rasmus-Vorrath, Mooyoung Lee, Daniel W. Engels
Yelp’S Review Filtering Algorithm, Yao Yao, Ivelin Angelov, Jack Rasmus-Vorrath, Mooyoung Lee, Daniel W. Engels
SMU Data Science Review
In this paper, we present an analysis of features influencing Yelp's proprietary review filtering algorithm. Classifying or misclassifying reviews as recommended or non-recommended affects average ratings, consumer decisions, and ultimately, business revenue. Our analysis involves systematically sampling and scraping Yelp restaurant reviews. Features are extracted from review metadata and engineered from metrics and scores generated using text classifiers and sentiment analysis. The coefficients of a multivariate logistic regression model were interpreted as quantifications of the relative importance of features in classifying reviews as recommended or non-recommended. The model classified review recommendations with an accuracy of 78%. We found that reviews …
Generalizing Multistage Partition Procedures For Two-Parameter Exponential Populations, Rui Wang
Generalizing Multistage Partition Procedures For Two-Parameter Exponential Populations, Rui Wang
LSU New Orleans Theses and Dissertations
ANOVA analysis is a classic tool for multiple comparisons and has been widely used in numerous disciplines due to its simplicity and convenience. The ANOVA procedure is designed to test if a number of different populations are all different. This is followed by usual multiple comparison tests to rank the populations. However, the probability of selecting the best population via ANOVA procedure does not guarantee the probability to be larger than some desired prespecified level. This lack of desirability of the ANOVA procedure was overcome by researchers in early 1950's by designing experiments with the goal of selecting the best …
Surprise Vs. Probability As A Metric For Proof, Edward K. Cheng, Matthew Ginther
Surprise Vs. Probability As A Metric For Proof, Edward K. Cheng, Matthew Ginther
Vanderbilt Law School Faculty Publications
In this Symposium issue celebrating his career, Professor Michael Risinger in Leveraging Surprise proposes using "the fundamental emotion of surprise" as a way of measuring belief for purposes of legal proof. More specifically, Professor Risinger argues that we should not conceive of the burden of proof in terms of probabilities such as 51%, 95%, or even "beyond a reasonable doubt." Rather, the legal system should reference the threshold using "words of estimative surprise" -asking jurors how surprised they would be if the fact in question were not true. Toward this goal (and being averse to cardinality), he suggests categories such …
Distribution Of A Sum Of Random Variables When The Sample Size Is A Poisson Distribution, Mark Pfister
Distribution Of A Sum Of Random Variables When The Sample Size Is A Poisson Distribution, Mark Pfister
Electronic Theses and Dissertations
A probability distribution is a statistical function that describes the probability of possible outcomes in an experiment or occurrence. There are many different probability distributions that give the probability of an event happening, given some sample size n. An important question in statistics is to determine the distribution of the sum of independent random variables when the sample size n is fixed. For example, it is known that the sum of n independent Bernoulli random variables with success probability p is a Binomial distribution with parameters n and p: However, this is not true when the sample size …
The Expected Number Of Patterns In A Random Generated Permutation On [N] = {1,2,...,N}, Evelyn Fokuoh
The Expected Number Of Patterns In A Random Generated Permutation On [N] = {1,2,...,N}, Evelyn Fokuoh
Electronic Theses and Dissertations
Previous work by Flaxman (2004) and Biers-Ariel et al. (2018) focused on the number of distinct words embedded in a string of words of length n. In this thesis, we will extend this work to permutations, focusing on the maximum number of distinct permutations contained in a permutation on [n] = {1,2,...,n} and on the expected number of distinct permutations contained in a random permutation on [n]. We further considered the problem where repetition of subsequences are as a result of the occurrence of (Type A and/or Type B) replications. Our method of enumerating the Type A replications causes double …
Pretrial Release And Failure-To-Appear In Mclean County, Il, Jonathan Monsma
Pretrial Release And Failure-To-Appear In Mclean County, Il, Jonathan Monsma
Student Research – Stevenson Center
Actuarial risk assessment tools increasingly have been employed in jurisdictions across the U.S. to assist courts in the decision of whether someone charged with a crime should be detained or released prior to their trial. These tools should be continually monitored and researched by independent 3rd parties to ensure that these powerful tools are being administered properly and used in the most proficient way as to provide socially optimal results. McLean County, Illinois began using the Public Safety Assessment-CourtTM (PSA-Court or simply PSA) risk assessment tool beginning in 2016. This study culls data from the McLean County Jail …
Mixed Logical And Probabilistic Reasoning In The Game Of Clue, Todd W. Neller, Ziqian Luo
Mixed Logical And Probabilistic Reasoning In The Game Of Clue, Todd W. Neller, Ziqian Luo
Computer Science Faculty Publications
Neller and Ziqian Luo ’18 presented a means of mixed logical and probabilistic reasoning with knowledge in the popular deductive mystery game Clue. Using at-least constraints, we more efficiently represented and reasoned about cardinality constraints on Clue card deal knowledge, and then employed a WalkSAT-based solution sampling algorithm with a tabu search metaheuristic in order to estimate the probabilities of unknown card places.
Risk Assessment Of Dropped Cylindrical Objects In Offshore Operations, Adelina Steven
Risk Assessment Of Dropped Cylindrical Objects In Offshore Operations, Adelina Steven
LSU New Orleans Theses and Dissertations
Dropped object are defined as any object that fall under its own weight from a previously static position or fell due to an applied force from equipment or a moving object. It is among the top ten causes of injuries and fatality in oil and gas industry. To solve this problem, several in-house tools and guidelines is developed over time to assess the risk of dropped objects on the sub-sea structures. This thesis focuses on compiling and comparing those methods in hope to improve the recommended practices available in the market. A simple modification is done on the in-house tools …
Golden Arm: A Probabilistic Study Of Dice Control In Craps, Donald R. Smith, Robert Scott Iii
Golden Arm: A Probabilistic Study Of Dice Control In Craps, Donald R. Smith, Robert Scott Iii
UNLV Gaming Research & Review Journal
This paper calculates how much control a craps shooter must possess on dice outcomes to eliminate the house advantage. A golden arm is someone who has dice control (or a rhythm roller or dice influencer). There are various strategies for dice control in craps. We discuss several possibilities of dice control that would result in several different mathematical models of control. We do not assert whether dice control is possible or not (there is a lack of published evidence). However, after studying casino-legal methods described by dice-control advocates, we can see only one realistic mathematical model that describes the resulting …
Evaluation Of Using The Bootstrap Procedure To Estimate The Population Variance, Nghia Trong Nguyen
Evaluation Of Using The Bootstrap Procedure To Estimate The Population Variance, Nghia Trong Nguyen
Electronic Theses and Dissertations
The bootstrap procedure is widely used in nonparametric statistics to generate an empirical sampling distribution from a given sample data set for a statistic of interest. Generally, the results are good for location parameters such as population mean, median, and even for estimating a population correlation. However, the results for a population variance, which is a spread parameter, are not as good due to the resampling nature of the bootstrap method. Bootstrap samples are constructed using sampling with replacement; consequently, groups of observations with zero variance manifest in these samples. As a result, a bootstrap variance estimator will carry a …
Score Test And Likelihood Ratio Test For Zero-Inflated Binomial Distribution And Geometric Distribution, Xiaogang Dai
Score Test And Likelihood Ratio Test For Zero-Inflated Binomial Distribution And Geometric Distribution, Xiaogang Dai
Masters Theses & Specialist Projects
The main purpose of this thesis is to compare the performance of the score test and the likelihood ratio test by computing type I errors and type II errors when the tests are applied to the geometric distribution and inflated binomial distribution. We first derive test statistics of the score test and the likelihood ratio test for both distributions. We then use the software package R to perform a simulation to study the behavior of the two tests. We derive the R codes to calculate the two types of error for each distribution. We create lots of samples to approximate …
General Stochastic Integral And Itô Formula With Application To Stochastic Differential Equations And Mathematical Finance, Jiayu Zhai
LSU Doctoral Dissertations
A general stochastic integration theory for adapted and instantly independent stochastic processes arises when we consider anticipative stochastic differential equations. In Part I of this thesis, we conduct a deeper research on the general stochastic integral introduced by W. Ayed and H.-H. Kuo in 2008. We provide a rigorous mathematical framework for the integral in Chapter 2, and prove that the integral is well-defined. Then a general Itô formula is given. In Chapter 3, we present an intrinsic property, near-martingale property, of the general stochastic integral, and Doob-Meyer's decomposition for near-submartigales. We apply the new stochastic integration theory to several …
Density Estimation Of Spatio-Temporal Point Patterns Using Moran’S Statistics, Jennifer L. Lorio, Norou Diawara, Lance A. Waller
Density Estimation Of Spatio-Temporal Point Patterns Using Moran’S Statistics, Jennifer L. Lorio, Norou Diawara, Lance A. Waller
Mathematics & Statistics Faculty Publications
Moran’s Index is a statistic that measures spatial autocorrelation, quantifying the degree of dispersion (or spread) of objects in space. When investigating data in an area, a single Moran statistic may not give a sufficient summary of the autocorrelation spread. However, by partitioning the area and taking the Moran statistic of each subarea, we discover patterns of the local neighbors not otherwise apparent. In this paper, we consider the model of the spread of an infectious disease, incorporate time factor, and simulate a multilevel Poisson process where the dependence among the levels is captured by the rate of increase of …
Predicting The Next Us President By Simulating The Electoral College, Boyan Kostadinov
Predicting The Next Us President By Simulating The Electoral College, Boyan Kostadinov
Journal of Humanistic Mathematics
We develop a simulation model for predicting the outcome of the US Presidential election based on simulating the distribution of the Electoral College. The simulation model has two parts: (a) estimating the probabilities for a given candidate to win each state and DC, based on state polls, and (b) estimating the probability that a given candidate will win at least 270 electoral votes, and thus win the White House. All simulations are coded using the high-level, open-source programming language R. One of the goals of this paper is to promote computational thinking in any STEM field by illustrating how probabilistic …
Preference Probability Based On Ranks - A New Approach Using Logistic Regression With Zero Intercept, Oluwagbenga David Agboola
Preference Probability Based On Ranks - A New Approach Using Logistic Regression With Zero Intercept, Oluwagbenga David Agboola
Theses, Dissertations and Capstones
Many probability models have been proposed to describe rankings. One of these is the BradleyTerry model, which is based on observed pairwise preferences. For this study, we reverse the case and propose a new approach for estimating pairwise preference probabilities based on observed rankings. The new approach uses logistic regression with zero intercept as the statistical model that fits this situation. In order to implement the model, we first estimate the parameter using maximum likelihood estimation. Then we evaluate this estimation using numerical approximation procedures. We consider three such procedures: bisection method, Newton-Raphson method, and improved Newton’s method. Using simulated …
Mechanism Design, Matching Theory And The Stable Roommates Problem, Yashaswi Mohanty
Mechanism Design, Matching Theory And The Stable Roommates Problem, Yashaswi Mohanty
Honors Theses
This thesis consists of two independent albeit related chapters. The first chapter introduces concepts from mechanism design and matching theory, and discusses potential applications of this theory, particularly in relation to dorm allocations in colleges. The second chapter investigates a subset of the dorm allocation problem, namely that of matching roommates. In particular, the paper looks at the probability of solvability of random instances of the stable roommates game under the condition that preferences are not completely random and exogenous but endogenously determined through a dependence on room choice. These probabilities are estimated using Monte-Carlo simulations and then compared with …
Comparing Various Machine Learning Statistical Methods Using Variable Differentials To Predict College Basketball, Nicholas Bennett
Comparing Various Machine Learning Statistical Methods Using Variable Differentials To Predict College Basketball, Nicholas Bennett
Williams Honors College, Honors Research Projects
The purpose of this Senior Honors Project is to research, study, and demonstrate newfound knowledge of various machine learning statistical techniques that are not covered in the University of Akron’s statistics major curriculum. This report will be an overview of three machine-learning methods that were used to predict NCAA Basketball results, specifically, the March Madness tournament. The variables used for these methods, models, and tests will include numerous variables kept throughout the season for each team, along with a couple variables that are used by the selection committee when tournament teams are being picked. The end goal is to find …
A Review Of The Utility Of Bayesian Network Models, Luke Magyar
A Review Of The Utility Of Bayesian Network Models, Luke Magyar
Williams Honors College, Honors Research Projects
Bayesian Networks are probabilistic models built from conditional probability tables that relate two observable instances to one another in parent-child fashion. The networks’ strength lies in their ability to use inferential logic to make likelihood assessments about a parent node based on an observation of its child. Additionally, they make it very easy to combine quantitative data with qualitative knowledge from industry experts. These abilities make them very attractive for use as formulation tools in the paint and rubber industries. Paint and rubber formulation has long proven to be a challenging task because companies have a difficult time compiling the …
Offline And Online Density Estimation For Large High-Dimensional Data, Aref Majdara
Offline And Online Density Estimation For Large High-Dimensional Data, Aref Majdara
Dissertations, Master's Theses and Master's Reports
Density estimation has wide applications in machine learning and data analysis techniques including clustering, classification, multimodality analysis, bump hunting and anomaly detection. In high-dimensional space, sparsity of data in local neighborhood makes many of parametric and nonparametric density estimation methods mostly inefficient.
This work presents development of computationally efficient algorithms for high-dimensional density estimation, based on Bayesian sequential partitioning (BSP). Copula transform is used to separate the estimation of marginal and joint densities, with the purpose of reducing the computational complexity and estimation error. Using this separation, a parallel implementation of the density estimation algorithm on a 4-core CPU is …
Application Of Remote Sensing And Machine Learning Modeling To Post-Wildfire Debris Flow Risks, Priscilla Addison
Application Of Remote Sensing And Machine Learning Modeling To Post-Wildfire Debris Flow Risks, Priscilla Addison
Dissertations, Master's Theses and Master's Reports
Historically, post-fire debris flows (DFs) have been mostly more deadly than the fires that preceded them. Fires can transform a location that had no history of DFs to one that is primed for it. Studies have found that the higher the severity of the fire, the higher the probability of DF occurrence. Due to high fatalities associated with these events, several statistical models have been developed for use as emergency decision support tools. These previous models used linear modeling approaches that produced subpar results. Our study therefore investigated the application of nonlinear machine learning modeling as an alternative. Existing models …
An Analysis Of Equity-Linked Insurance Pricing, Clara C. Ortgies
An Analysis Of Equity-Linked Insurance Pricing, Clara C. Ortgies
Honors Program Theses
This comprehensive study of equity-linked insurance options will explore the pricing of certificates of deposit and life insurance options using a present value method. With this study, I will be able to construct and price various equity-linked insurance products, with a focus on life insurance, that insurance companies could then sell to prospective customers. I will use concepts and formulas based in actuarial math, probability theory, and financial engineering in order to construct, price, and analyze new equity-linked insurance products. The fundamental methodology I will use involves applying pricing theory based on the expected value of the insurance payoff present …
Effect Of Neuromodulation Of Short-Term Plasticity On Information Processing In Hippocampal Interneuron Synapses, Elham Bayat Mokhtari
Effect Of Neuromodulation Of Short-Term Plasticity On Information Processing In Hippocampal Interneuron Synapses, Elham Bayat Mokhtari
Graduate Student Theses, Dissertations, & Professional Papers
Neurons convey information about the complex dynamic environment in the form of signals. Computational neuroscience provides a theoretical foundation toward enhancing our understanding of nervous system. The aim of this dissertation is to present techniques to study the brain and how it processes information in particular neurons in hippocampus.
We begin with a brief review of the history of neuroscience and biological background of basic neurons. To appreciate the importance of information theory, familiarity with the information theoretic basics is required, these basics are presented in Chapter 2. In Chapter 3, we use information theory to estimate the amount of …
Some New And Generalized Distributions Via Exponentiation, Gamma And Marshall-Olkin Generators With Applications, Hameed Abiodun Jimoh
Some New And Generalized Distributions Via Exponentiation, Gamma And Marshall-Olkin Generators With Applications, Hameed Abiodun Jimoh
College of Graduate Studies: Theses & Dissertations
Three new generalized distributions developed via completing risk, gamma generator, Marshall-Olkin generator and exponentiation techniques are proposed and studied. Structural properties including quantile functions, hazard rate functions, moment, conditional moments, mean deviations, R\'enyi entropy, distribution of order statistics and maximum likelihood estimates are presented. Monte Carlo simulation is employed to examine the performance of the proposed distributions. Applications of the generalized distributions to real lifetime data are presented to illustrate the usefulness of the models.
Particle Filters For State Estimation Of Confined Aquifers, Graeme Field
Particle Filters For State Estimation Of Confined Aquifers, Graeme Field
UNF Graduate Theses and Dissertations
Mathematical models are used in engineering and the sciences to estimate properties of systems of interest, increasing our understanding of the surrounding world and driving technological innovation. Unfortunately, as the systems of interest grow in complexity, so to do the models necessary to accurately describe them. Analytic solutions for problems with such models are provably intractable, motivating the use of approximate yet still accurate estimation techniques. Particle filtering methods have emerged as a popular tool in the presence of such models, spreading from its origins in signal processing to a diverse set of fields throughout engineering and the sciences including …