Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Statistics and Probability (27)
- Computer Sciences (23)
- Social and Behavioral Sciences (17)
- Business (13)
- Medicine and Health Sciences (13)
-
- Statistical Models (13)
- Artificial Intelligence and Robotics (12)
- Applied Statistics (11)
- Engineering (10)
- Longitudinal Data Analysis and Time Series (9)
- Categorical Data Analysis (6)
- Applied Mathematics (5)
- Diseases (4)
- Electrical and Computer Engineering (4)
- Finance and Financial Management (4)
- Information Security (4)
- Law (4)
- Mathematics (4)
- Sports Studies (4)
- Statistical Methodology (4)
- Theory and Algorithms (4)
- Analytical, Diagnostic and Therapeutic Techniques and Equipment (3)
- Business Analytics (3)
- Business Intelligence (3)
- Computer Engineering (3)
- Diagnosis (3)
- Digital Communications and Networking (3)
- Environmental Sciences (3)
- Keyword
-
- Machine Learning (16)
- NLP (15)
- Data Science (14)
- Machine learning (11)
- CNN (10)
-
- Natural language processing (9)
- Deep Learning (8)
- Time series (8)
- Deep learning (7)
- Neural Networks (6)
- Random Forest (6)
- COVID-19 (5)
- Classification (5)
- LSTM (5)
- ARIMA (4)
- Clustering (4)
- Computer vision (4)
- GAN (4)
- LLM (4)
- Natural Language Processing (4)
- XGBoost (4)
- AI (3)
- BERT (3)
- Bias (3)
- ERCOT (3)
- Emotion (3)
- Forecasting (3)
- LLMs (3)
- Logistic regression (3)
- ML (3)
Articles 121 - 124 of 124
Full-Text Articles in Data Science
Forecasting Power Consumption In Pennsylvania During The Covid-19 Pandemic: A Sarimax Model With External Covid-19 And Unemployment Variables, Jackson Au, Javier Saldaña Jr., Ben Spanswick, John Santerre
Forecasting Power Consumption In Pennsylvania During The Covid-19 Pandemic: A Sarimax Model With External Covid-19 And Unemployment Variables, Jackson Au, Javier Saldaña Jr., Ben Spanswick, John Santerre
SMU Data Science Review
In this paper, we present how electrical consumption can reveal insight into the novel COVID-19 pandemic spread. We analyze electrical power consumption provided by PPL Electric Utilities, Department of Labor’s unemployment claims, and the COVID-19 cases/deaths for the State of Pennsylvania to study the impact of the pandemic on the infrastructure. Using a SARIMA model as our benchmark and we analyzed the use of a SARIMAX model to forecast the power consumption in Pennsylvania 14 days ahead. Our work quantifies and illuminates the effect that the strict legislation passed to minimize the spread of COVID19 had a on power consumption. …
Compressed Dna Representation For Efficient Amr Classification, John Partee, Robert Hazell, Anjli Solsi, John Santerre
Compressed Dna Representation For Efficient Amr Classification, John Partee, Robert Hazell, Anjli Solsi, John Santerre
SMU Data Science Review
In this paper, we explore a representation methodology for the compression of DNA isolates. Using lossless string compression via tokenization of frequently repeated segments of DNA, we reduce the length of the isolates to be counted as k-mers for classification. With this new representation, we apply a previously established feature sampling method to dramatically reduce the feature space. In understanding the genetic diversity, we also look at conserving biological function across these spaces. Using a random forest model we were able to predict the resistance or susceptibility of bacteria with 85-90\% accuracy, with a 30-50\% reduction in overall isolate length, …
Spoken Language Recognition On Open-Source Datasets, Brady Arendale, Samira Zarandioon, Ryan Goodwin, Douglas Reynolds
Spoken Language Recognition On Open-Source Datasets, Brady Arendale, Samira Zarandioon, Ryan Goodwin, Douglas Reynolds
SMU Data Science Review
The field of speaker and language recognition is constantly being researched and developed, but much of this research is done on private or expensive datasets, making the field more inaccessible than many other areas of machine learning. In addition, many papers make performance claims without comparing their models to other recent research. With the recent development of public multilingual speech corpora such as Mozilla's Common Voice as well as several single-language corpora, we now have the resources to attempt to address both of these problems. We construct an eight-language dataset from Common Voice and a Google Bengali corpus as well …
Predicting Attrition - A Driver For Creating Value, Realizing Strategy, And Refining Key Hr Processes, Kevin Mendonsa, Maureen Stolberg, Vivek Viswanathan, Scott Crum
Predicting Attrition - A Driver For Creating Value, Realizing Strategy, And Refining Key Hr Processes, Kevin Mendonsa, Maureen Stolberg, Vivek Viswanathan, Scott Crum
SMU Data Science Review
Talent is the most important asset for every organization's success. While attrition (or churn) and turnover can refer to both employees and customers, this paper will focus on employee attrition only. Many organizations accept attrition as an inevitable cost of doing business and do nothing to adopt or implement mitigating strategies to combat it. World class companies on the other hand take deliberate measures to understand, control and mitigate attrition (turnover) at every stage. Unmitigated attrition can have a devastating effect on an organization's bottom line and market value. In addition, the “invisible" costs of low employee morale, reduced employee …