Open Access. Powered by Scholars. Published by Universities.®

Computer Sciences Commons

Open Access. Powered by Scholars. Published by Universities.®

Classification

Discipline
Institution
Publication Year
Publication
Publication Type
File Type

Articles 121 - 150 of 375

Full-Text Articles in Computer Sciences

An Improved Version Of Multi-View K-Nearest Neighbors (Mvknn) For Multipleview Learning, Eli̇fe Öztürk Kiyak, Derya Bi̇rant, Kökten Ulaş Bi̇rant Jan 2021

An Improved Version Of Multi-View K-Nearest Neighbors (Mvknn) For Multipleview Learning, Eli̇fe Öztürk Kiyak, Derya Bi̇rant, Kökten Ulaş Bi̇rant

Turkish Journal of Electrical Engineering and Computer Sciences

Multi-view learning (MVL) is a special type of machine learning that utilizes more than one views, where views include various descriptions of a given sample. Traditionally, classification algorithms such as k-nearest neighbors (KNN) are designed for learning from single-view data. However, many real-world applications involve datasets with multiple views and each view may contain different and partly independent information, which makes the traditional single-view classification approaches ineffective. Therefore, this article proposes an improved MVL algorithm, called multi-view k-nearest neighbors (MVKNN), based on the existing KNN algorithm. The experimental results conducted in this research show that a significant improvement is achieved …


A Linear Programming Approach To Multiple Instance Learning, Emel Şeyma Küçükaşci, Mustafa Gökçe Baydoğan, Zeki̇ Caner Taşkin Jan 2021

A Linear Programming Approach To Multiple Instance Learning, Emel Şeyma Küçükaşci, Mustafa Gökçe Baydoğan, Zeki̇ Caner Taşkin

Turkish Journal of Electrical Engineering and Computer Sciences

Multiple instance learning (MIL) aims to classify objects with complex structures and covers a wide range of real-world data mining applications. In MIL, objects are represented by a bag of instances instead of a single instance, and class labels are provided only for the bags. Some of the earlier MIL methods focus on solving MIL problem under the standard MIL assumption, which requires at least one positive instance in positive bags and all remaining instances are negative. This study proposes a linear programming framework to learn instance level contributions to bag label without emposing the standart assumption. Each instance of …


Diagnosis Of Paroxysmal Atrial Fibrillation From Thirty-Minute Heart Ratevariability Data Using Convolutional Neural Networks, Murat Sürücü, Yalçin İşler, Resul Kara Jan 2021

Diagnosis Of Paroxysmal Atrial Fibrillation From Thirty-Minute Heart Ratevariability Data Using Convolutional Neural Networks, Murat Sürücü, Yalçin İşler, Resul Kara

Turkish Journal of Electrical Engineering and Computer Sciences

Paroxysmal atrial fibrillation (PAF) is the initial stage of atrial fibrillation, one of the most common arrhythmia types. PAF worsens with time and affects the patient?s life quality negatively. In this study, we aimed to diagnose PAF early, so patients can start taking precautions before this disease gets worse. We used the atrial fibrillation prediction database, an open data from Physionet and constructed our approach using convolutional neural networks. Heart rate variability (HRV) features are calculated from time-domain measures, frequency-domain measures using power spectral density estimations (fast Fourier transform, Lomb-Scargle, and Welch periodogram), time-frequencydomain measures using wavelet transform, and nonlinear …


Automatic Subtyping Of Individuals With Primary Progressive Aphasia, Charalambos Themistocleous, Bronte Ficek, Kimberly Webster, Dirk B. Den Ouden, Argye Hillis, Kyrana Tsapkini Jan 2021

Automatic Subtyping Of Individuals With Primary Progressive Aphasia, Charalambos Themistocleous, Bronte Ficek, Kimberly Webster, Dirk B. Den Ouden, Argye Hillis, Kyrana Tsapkini

Communication Sciences and Disorders Faculty Articles and Research

Background:

The classification of patients with primary progressive aphasia (PPA) into variants is time-consuming, costly, and requires combined expertise by clinical neurologists, neuropsychologists, speech pathologists, and radiologists.

Objective:

The aim of the present study is to determine whether acoustic and linguistic variables provide accurate classification of PPA patients into one of three variants: nonfluent PPA, semantic PPA, and logopenic PPA.

Methods:

In this paper, we present a machine learning model based on deep neural networks (DNN) for the subtyping of patients with PPA into three main variants, using combined acoustic and linguistic information elicited automatically via acoustic and linguistic analysis. …


Neural Network Supervised And Reinforcement Learning For Neurological, Diagnostic, And Modeling Problems, Donald Wunsch Iii Jan 2021

Neural Network Supervised And Reinforcement Learning For Neurological, Diagnostic, And Modeling Problems, Donald Wunsch Iii

Masters Theses

“As the medical world becomes increasingly intertwined with the tech sphere, machine learning on medical datasets and mathematical models becomes an attractive application. This research looks at the predictive capabilities of neural networks and other machine learning algorithms, and assesses the validity of several feature selection strategies to reduce the negative effects of high dataset dimensionality. Our results indicate that several feature selection methods can maintain high validation and test accuracy on classification tasks, with neural networks performing best, for both single class and multi-class classification applications. This research also evaluates a proof-of-concept application of a deep-Q-learning network (DQN) to …


A Comparison Of Instructional Efficiency Models In Third Level Education, Murali Rajendran Jan 2021

A Comparison Of Instructional Efficiency Models In Third Level Education, Murali Rajendran

Dissertations

This study investigates the validity and sensitivity of a novel model of instructional efficiency: the parabolic model. The novel model is compared against state-of-the-art models present in instructional design today; Likelihood model, Deviational model and Multidimensional model. This models is based on the assumption that optimal mental workload and high performance leads to high efficiency, while other models assume that low mental workload and high performance leads to high efficiency. The investigation makes use of two instructional design conditions: a direct instructions approach to learning and its extension with a collaborative activity. A control group received the former instructional design …


Plant Species Identification In The Wild Based On Images Of Organs, Meghana Kovur Jan 2021

Plant Species Identification In The Wild Based On Images Of Organs, Meghana Kovur

Graduate Theses, Dissertations, and Problem Reports (ETD)

Image-based plant species identification in the wild is a difficult problem for several reasons. First, the input data is subject to a very high degree of variability because it is captured under fully unconstrained conditions. The same plant species may look very different in different images, while different species can often appear very similar, challenging even the recognition skills of human experts in the field. The large intra-class and small inter-class image variability makes this a fine-grained visual classification problem. One way to cope with this variability and to reduce image background noise is to predict species based on the …


Identification And Classification Of Radio Pulsar Signals Using Machine Learning, Di Pang Jan 2021

Identification And Classification Of Radio Pulsar Signals Using Machine Learning, Di Pang

Graduate Theses, Dissertations, and Problem Reports (ETD)

Automated single-pulse search approaches are necessary as ever-increasing amount of observed data makes the manual inspection impractical. Detecting radio pulsars using single-pulse searches, however, is a challenging problem for machine learning because pul- sar signals often vary significantly in brightness, width, and shape and are only detected in a small fraction of observed data.

The research work presented in this dissertation is focused on development of ma- chine learning algorithms and approaches for single-pulse searches in the time domain. Specifically, (1) We developed a two-stage single-pulse search approach, named Single- Pulse Event Group IDentification (SPEGID), which automatically identifies and clas- …


Gene Selection For Cancer Classification: A New Hybrid Filter-C5.0 Approach For Breast Cancer Risk Prediction, Mohammed Hamim, Ismail El Moudden, Hicham Moutachaouik, Mustapha Hain Jan 2021

Gene Selection For Cancer Classification: A New Hybrid Filter-C5.0 Approach For Breast Cancer Risk Prediction, Mohammed Hamim, Ismail El Moudden, Hicham Moutachaouik, Mustapha Hain

Department of Medicine Faculty Publications

Despite the significant progress made in data mining technologies in recent years, breast cancer risk prediction and diagnosis at an early stage using DNA microarray technology still a real challenging task. This challenge comes especially from the high-dimensionality in gene expression data, i.e., an enormous number of genes versus a few tens of subjects (samples). To overcome this problem of data imbalance, a gene selection phase becomes a crucial step for gene expression data analysis. This study proposes a new Decision Tree model-based attributes (genes) selection strategy, which incorporates two stages: fisher-score-based filter technique and the gene selection ability of …


Impact Of Image Segmentation Techniques On Celiac Disease Classification Usingscale Invariant Texture Descriptors For Standard Flexible Endoscopic Systems, Manarbek Saken, Munkhtsetseg Banzragch Yağci, Nejat Yumuşak Jan 2021

Impact Of Image Segmentation Techniques On Celiac Disease Classification Usingscale Invariant Texture Descriptors For Standard Flexible Endoscopic Systems, Manarbek Saken, Munkhtsetseg Banzragch Yağci, Nejat Yumuşak

Turkish Journal of Electrical Engineering and Computer Sciences

Celiac disease (CD) is quite common and is a proximal small bowel disease that develops as a permanentintolerance to gluten and other cereal proteins in cereals. It is considered as one of the most di?icult diseases to diagnose.Histopathological evidence of small bowel biopsies taken during endoscopy remains the gold standard for diagnosis.Therefore, computer-aided detection (CAD) systems in endoscopy are a newly emerging technology to enhance thediagnostic accuracy of the disease and to save time and manpower. For this reason, a hybrid machine learning methodshave been applied for the CAD of celiac disease. Firstly, a context-based optimal multilevel thresholding technique wasemployed …


The Nearest Polyhedral Convex Conic Regions For High-Dimensional Classification, Hakan Çevi̇kalp, Emre Çi̇men, Gürkan Öztürk Jan 2021

The Nearest Polyhedral Convex Conic Regions For High-Dimensional Classification, Hakan Çevi̇kalp, Emre Çi̇men, Gürkan Öztürk

Turkish Journal of Electrical Engineering and Computer Sciences

In the nearest-convex-model type classifiers, each class in the training set is approximated with a convexclass model, and a test sample is assigned to a class based on the shortest distance from the test sample to these classmodels. In this paper, we propose new methods for approximating the distances from test samples to the convex regionsspanned by training samples of classes. To this end, we approximate each class region with a polyhedral convex conicregion by utilizing polyhedral conic functions (PCFs) and its extension, extended PCFs. Then, we derive the necessary formulations for computing the distances from test samples to these …


A New Approach: Semisupervised Ordinal Classification, Ferda Ünal, Derya Bi̇rant, Özlem Şeker Jan 2021

A New Approach: Semisupervised Ordinal Classification, Ferda Ünal, Derya Bi̇rant, Özlem Şeker

Turkish Journal of Electrical Engineering and Computer Sciences

Semisupervised learning is a type of machine learning technique that constructs a classifier by learning from a small collection of labeled samples and a large collection of unlabeled ones. Although some progress has been made in this research area, the existing semisupervised methods provide a nominal classification task. However, semisupervised learning for ordinal classification is yet to be explored. To bridge the gap, this study combines two concepts ?semisupervised learning? and "ordinal classification" for the categorical class labels for the first time and introduces a new concept of "semisupervised ordinal classification". This paper proposes a new algorithm for semisupervised learning …


Understanding And Predicting Retractions Of Published Work, Sai Ajay Modukuri, Sarah Rajtmajer, Anna Cinzia Squicciarini, Jian Wu, C. Lee Giles Jan 2021

Understanding And Predicting Retractions Of Published Work, Sai Ajay Modukuri, Sarah Rajtmajer, Anna Cinzia Squicciarini, Jian Wu, C. Lee Giles

Computer Science Faculty Publications

Recent increases in the number of retractions of published papers reflect heightened attention and increased scrutiny in the scientific process motivated, in part, by the replication crisis. These trends motivate computational tools for understanding and assessment of the scholarly record. Here, we sketch the landscape of retracted papers in the Retraction Watch database, a collection of 19k records of published scholarly articles that have been retracted for various reasons (e.g., plagiarism, data error). Using metadata as well as features derived from full-text for a subset of retracted papers in the social and behavioral sciences, we develop a random forest classifier …


Signature Identification And Verification Systems: A Comparative Study On The Online And Offline Techniques, Nehal Hamdy Al-Banhawy, Heba Mohsen, Neveen I. Ghali Prof. Dec 2020

Signature Identification And Verification Systems: A Comparative Study On The Online And Offline Techniques, Nehal Hamdy Al-Banhawy, Heba Mohsen, Neveen I. Ghali Prof.

Future Computing and Informatics Journal

Handwritten signature identification and verification has become an active area of research in recent years. Handwritten signature identification systems are used for identifying the user among all users enrolled in the system while handwritten signature verification systems are used for authenticating a user by comparing a specific signature with his signature that is stored in the system. This paper presents a review for commonly used methods for preprocessing, feature extraction and classification techniques in signature identification and verification systems, in addition to a comparison between the systems implemented in the literature for identification techniques and verification techniques in online and …


Identification Of Ecg Anomalies Through Deep Deterministic Learning, Uzair Iqbal Nov 2020

Identification Of Ecg Anomalies Through Deep Deterministic Learning, Uzair Iqbal

Student Works (2020-2029)

Electrocardiography (ECG) is a primary diagnostic tool for measuring the malfunctioning of the heart muscles in the context of morbidity of different cardiac diseases and arrhythmia. Different existing techniques and methods delivered accurate cardiac diseases myocardial infarction (heart stroke) and atrial fibrillation recognition. However, there are still some flaws in existing methods like recognition of special myocardial infarction situation flattened T wave in “Non-Specific ST-T Changes (nsst-t)” and reduction of computational cost in cardiac diseases recognition. Accurate recognition of cardiac diseases along with least computational complexity and feature analysis of flattened T wave in myocardial infarction remains an open job. …


A Systematic Mapping Study On The Risk Factors Leading To Type Ii Diabetes Mellitus, Karar N. J Musafer, Fahrul Zaman Huyop, Mufeed J Ewadh, Eko Supriyanto, Mohammad Rava Oct 2020

A Systematic Mapping Study On The Risk Factors Leading To Type Ii Diabetes Mellitus, Karar N. J Musafer, Fahrul Zaman Huyop, Mufeed J Ewadh, Eko Supriyanto, Mohammad Rava

Karbala International Journal of Modern Science

Diabetes is one of the most common diseases that has had devastating effects on the general population. It is also among the most popular research trends in modern medicine. Thus, due to the complexity and desirability of this particular affliction, there is a lot of demand towards understanding this disease better, so that it can pave the way towards better solutions in combating diabetes. The aim of this review is to provide a categorization of the risk factors leading to Type II Diabetes. In order to provide a justification for the type of diabetes, an explanation is provided which covers …


Wait For It: Identifying 'On-Hold' Self-Admitted Technical Debt, Rungroj Maipradit, Christoph Treude, Hideaki Hata, Kenichi Matsumoto Sep 2020

Wait For It: Identifying 'On-Hold' Self-Admitted Technical Debt, Rungroj Maipradit, Christoph Treude, Hideaki Hata, Kenichi Matsumoto

Research Collection School Of Computing and Information Systems

Self-admitted technical debt refers to situations where a software developer knows that their current implementation is not optimal and indicates this using a source code comment. In this work, we hypothesize that it is possible to develop automated techniques to understand a subset of these comments in more detail, and to propose tool support that can help developers manage self-admitted technical debt more effectively. Based on a qualitative study of 333 comments indicating self-admitted technical debt, we first identify one particular class of debt amenable to automated management: on-hold self-admitted technical debt (on-hold SATD), i.e., debt which contains a condition …


A Unified Framework For Sparse Online Learning, Peilin Zhao, Dayong Wong, Pengcheng Wu, Steven C. H. Hoi Aug 2020

A Unified Framework For Sparse Online Learning, Peilin Zhao, Dayong Wong, Pengcheng Wu, Steven C. H. Hoi

Research Collection School Of Computing and Information Systems

The amount of data in our society has been exploding in the era of big data. This article aims to address several open challenges in big data stream classification. Many existing studies in data mining literature follow the batch learning setting, which suffers from low efficiency and poor scalability. To tackle these challenges, we investigate a unified online learning framework for the big data stream classification task. Different from the existing online data stream classification techniques, we propose a unified Sparse Online Classification (SOC) framework. Based on SOC, we derive a second-order online learning algorithm and a cost-sensitive sparse online …


Development And Identification Of Metrics To Predict The Impact Of Dimension Reduction Techniques On Classical Machine Learning Algorithms For Still Highway Images, Wasim Akram Khan Aug 2020

Development And Identification Of Metrics To Predict The Impact Of Dimension Reduction Techniques On Classical Machine Learning Algorithms For Still Highway Images, Wasim Akram Khan

All Graduate Theses and Dissertations, Spring 1920 to Summer 2023

We are witnessing an influx of data - images, texts, video, etc. Their high dimensionality and large volume make it challenging to apply machine learning to obtain actionable insight. This thesis explores several aspects pertaining to dimensional reduction: dimension reduction methods, metrics to measure distortion, image preprocessing, etc. Faster training and inference time on reduced data and smaller models which can be deployed on commodity hardware are a critical advantage of dimension reduction. For this study, classical machine learning methods were explored owing to their solid mathematical foundation and interpretability.

The dataset used is a time series of images from …


Data Mining For Structural Damage Identification Using Hybrid Artificial Neural Network Based Algorithm For Beam And Slab Girder, Gordan Meisam Aug 2020

Data Mining For Structural Damage Identification Using Hybrid Artificial Neural Network Based Algorithm For Beam And Slab Girder, Gordan Meisam

Student Works (2020-2029)

One of the approaches for structural health monitoring (SHM) consists of two major components, i.e. a network of sensors to collect the response data and an extraction method to obtain information on the structural health condition. Data mining (DM) is a novel data extraction technology which can employ for development of inverse analysis. Implementation of DM techniques in different areas of civil engineering has recently given very good results. However, application of DM in SHM is not used as much as expected, thus, many challenges are still ahead. Therefore, it is necessary to develop the applicability of DM in SHM. …


Computational Astronomy: Classification Of Celestial Spectra Using Machine Learning Techniques, Gayatri Milind Hungund May 2020

Computational Astronomy: Classification Of Celestial Spectra Using Machine Learning Techniques, Gayatri Milind Hungund

Master's Projects

Lightyears beyond the Planet Earth there exist plenty of unknown and unexplored stars and Galaxies that need to be studied in order to support the Big Bang Theory and also make important astronomical discoveries in quest of knowing the unknown. Sophisticated devices and high-power computational resources are now deployed to make a positive effort towards data gathering and analysis. These devices produce massive amount of data from the astronomical surveys and the data is usually in terabytes or petabytes. It is exhaustive to process this data and determine the findings in short period of time. Many details can be missed …


An Exploration Of Methods For Classifying Air-Written Letters From The Spanish Alphabet, Manuel Serna-Aguilera May 2020

An Exploration Of Methods For Classifying Air-Written Letters From The Spanish Alphabet, Manuel Serna-Aguilera

Computer Science and Computer Engineering Undergraduate Honors Theses

The ability to recognize human activity, especially air-writing, is an interesting challenge as one could identify any letter from many languages. I intend to investigate this problem of air-writing, but with the added twist of including the following letters from the Spanish alphabet: Á, É, Í, Ó, Ú, Ü, and Ñ. With this new alphabet, I set out to see what kinds of classifiers work best and on what kinds of data, since letters can be represented in multiple ways.

My tracking system will consist of a regular camera and a subject who will draw with a brightly colored marker …


Randomized And Evolutionary Approaches To Dataset Characterization, Feature Weighting, And Sampling In K-Nearest Neighbors, Suryoday Basak May 2020

Randomized And Evolutionary Approaches To Dataset Characterization, Feature Weighting, And Sampling In K-Nearest Neighbors, Suryoday Basak

Computer Science and Engineering Theses - Archive

K-Nearest Neighbors (KNN) has remained one of the most popular methods for supervised machine learning tasks. However, its performance often depends on the characteristics of the dataset and on appropriate feature scaling. In this thesis, characteristics of a dataset that make it suitable for being used within KNN are explored. As part of this, two new measures for dataset dispersion, called mean neighborhood target variance (MNTV), and mean neighborhood target entropy (MNTE) are developed to help determine the performance we expect while using KNN regressors and classifiers, respectively. It is empirically demonstrated that these measures of dispersion can be indicative …


Novel Inference Methods For Generalized Linear Models Using Shrinkage Priors And Data Augmentation., Arinjita Bhattacharyya May 2020

Novel Inference Methods For Generalized Linear Models Using Shrinkage Priors And Data Augmentation., Arinjita Bhattacharyya

Electronic Theses and Dissertations

Generalized linear models have broad applications in biostatistics and sociology. In a regression setup, the main target is to find a relevant set of predictors out of a large collection of covariates. Sparsity is the assumption that only a few of these covariates in a regression setup have a meaningful correlation with an outcome variate of interest. Sparsity is incorporated by regularizing the irrelevant slopes towards zero without changing the relevant predictors and keeping the resulting inferences intact. Frequentist variable selection and sparsity are addressed by popular techniques like Lasso, Elastic Net. Bayesian penalized regression can tackle the curse of …


Towards Multi-Modal Data Classification, Henry Ng May 2020

Towards Multi-Modal Data Classification, Henry Ng

UNLV Theses, Dissertations, Professional Papers, and Capstones

A feature fusion multi-modal neural network (MMN) is a network that combines different modalities at the feature level to perform a specific task. In this paper, we study the problem of training the fusion procedure for MMN. A recent study has found that training a multi-modal network that incorporates late fusion produces a network that has not learned the proper parameters for feature extraction. These late fusion models perform very well during training but fall short to its single modality counterpart when testing. We hypothesize that jointly trained MMN have weight space that is too large for effective training. To …


A Survey Of Feature Extraction And Fusion Of Deep Learning For Detection Of Abnormalities In Video Endoscopy Of Gastrointestinal-Tract, Hussam Ali, Muhammad Sharif, Mussarat Yasmin, Mubashir Husain Rehmani, Farhan Riaz Apr 2020

A Survey Of Feature Extraction And Fusion Of Deep Learning For Detection Of Abnormalities In Video Endoscopy Of Gastrointestinal-Tract, Hussam Ali, Muhammad Sharif, Mussarat Yasmin, Mubashir Husain Rehmani, Farhan Riaz

Publications

A standard screening procedure involves video endoscopy of the Gastrointestinal tract. It is a less invasive method which is practiced for early diagnosis of gastric diseases. Manual inspection of a large number of gastric frames is an exhaustive, time-consuming task, and requires expertise. Conversely, several computer-aided diagnosis systems have been proposed by researchers to cope with the dilemma of manual inspection of the massive volume of frames. This article gives an overview of different available alternatives for automated inspection, detection, and classification of various GI abnormalities. Also, this work elaborates techniques associated with content-based image retrieval and automated systems for …


Brain Disease Detection From Eegs: Comparing Spiking And Recurrent Neural Networks For Non-Stationary Time Series Classification, Hristo Stoev Jan 2020

Brain Disease Detection From Eegs: Comparing Spiking And Recurrent Neural Networks For Non-Stationary Time Series Classification, Hristo Stoev

Dissertations

Modeling non-stationary time series data is a difficult problem area in AI, due to the fact that the statistical properties of the data change as the time series progresses. This complicates the classification of non-stationary time series, which is a method used in the detection of brain diseases from EEGs. Various techniques have been developed in the field of deep learning for tackling this problem, with recurrent neural networks (RNN) approaches utilising Long short-term memory (LSTM) architectures achieving a high degree of success. This study implements a new, spiking neural network-based approach to time series classification for the purpose of …


An Analysis Of The Success Of Farmers Markets In Kentucky Using Logistic Regression And Support Vector Machines, Jeron Russell Jan 2020

An Analysis Of The Success Of Farmers Markets In Kentucky Using Logistic Regression And Support Vector Machines, Jeron Russell

Mahurin Honors College Capstone Experience/Thesis Projects

The purpose of this research is to look at the relationship that market-specific, economic, and demographic variables have with the success of farmers markets in Kentucky. It additionally seeks to build a tool for predicting farmers market success that could be used by policy makers to aid in decision-making processes concerning farmers markets. Logistic regression and Support Vector Machines (SVMs) are used on data acquired from the Kentucky Department of Agriculture and the American Community Survey in order to analyze the data in a traditional statistical approach as well as a machine learning approach. The results included an SVM model …


A Description Of A Humans Knowledge Using Artificial Intelligence, Dj Price Jan 2020

A Description Of A Humans Knowledge Using Artificial Intelligence, Dj Price

Mahurin Honors College Capstone Experience/Thesis Projects

There currently does not exist a way to easily view the relationships between a collection of written items (e.g. sports articles, diary entries, research papers). In recent years, novel machine learning methods have been developed which are very good at extracting semantic relationships from large numbers of documents. One of them is the (unsupervised) machine learning model Doc2Vec which constructs vectors for documents. The research project detailed in this paper uses this and other already existing algorithms to analyze the relationship between pieces of text. We set forth a broader ambition for this project before discussing the use and need …


Revised Polyhedral Conic Functions Algorithm For Supervised Classification, Gürhan Ceylan, Gürkan Öztürk Jan 2020

Revised Polyhedral Conic Functions Algorithm For Supervised Classification, Gürhan Ceylan, Gürkan Öztürk

Turkish Journal of Electrical Engineering and Computer Sciences

In supervised classification, obtaining nonlinear separating functions from an algorithm is crucial for prediction accuracy. This paper analyzes the polyhedral conic functions (PCF) algorithm that generates nonlinear separating functions by only solving simple subproblems. Then, a revised version of the algorithm is developed that achieves better generalization and fast training while maintaining the simplicity and high prediction accuracy of the original PCF algorithm. This is accomplished by making the following modifications to the subproblem: extension of the objective function with a regularization term, relaxation of a hard constraint set and introduction of a new error term. Experimental results show that …