Open Access. Powered by Scholars. Published by Universities.®

Communication Sciences and Disorders Commons

Open Access. Powered by Scholars. Published by Universities.®

Engineering

Institution
Keyword
Publication Year
Publication
Publication Type

Articles 31 - 60 of 68

Full-Text Articles in Communication Sciences and Disorders

Effects Of Vocal Fold Nodules On Glottal Cycle Measurements Derived From High-Speed Videoendoscopy In Children, Rita R. Patel, Harikrishnan Unnikrishnan, Kevin D. Donohue Apr 2016

Effects Of Vocal Fold Nodules On Glottal Cycle Measurements Derived From High-Speed Videoendoscopy In Children, Rita R. Patel, Harikrishnan Unnikrishnan, Kevin D. Donohue

Electrical and Computer Engineering Faculty Publications

The goal of this study is to quantify the effects of vocal fold nodules on vibratory motion in children using high-speed videoendoscopy. Differences in vibratory motion were evaluated in 20 children with vocal fold nodules (5–11 years) and 20 age and gender matched typically developing children (5–11 years) during sustained phonation at typical pitch and loudness. Normalized kinematic features of vocal fold displacements from the mid-membranous vocal fold point were extracted from the steady-state high-speed video. A total of 12 kinematic features representing spatial and temporal characteristics of vibratory motion were calculated. Average values and standard deviations (cycle-to-cycle variability) of …


Exploring Oculomotor Trends In Collegiate Athletes, Brett Whorley, Julie A. Honaker Apr 2016

Exploring Oculomotor Trends In Collegiate Athletes, Brett Whorley, Julie A. Honaker

UCARE: Research Products

Collaborative efforts to improve athlete safety without significantly hindering the rules of the games aim to develop a novel system to better measure and diagnose concussions. Provided that common signs of concussions include blurred vision, distant gaze, and dizziness, the Dizziness and Balance Disorders Lab at UNL believes that the simple oculomotor exam studied in this project may be applied to this procedure. Within the broader goal to better understand the causes, signs, symptoms, and prognosis of concussions, researchers desired to further investigate the results of this oculomotor test. The aim was to identify and interpret correlations between collegiate athlete …


Least-Squares Mapping From Kinematic Data To Acoustic Synthesis Parameters For Rehabilitative Acoustic Learning, Xiangyu Zhou Apr 2016

Least-Squares Mapping From Kinematic Data To Acoustic Synthesis Parameters For Rehabilitative Acoustic Learning, Xiangyu Zhou

Master's Theses (2009 -)

Thousands of people suffer from dysarthria resulting from neurological injury of the motor component of the motor-speech system, and need to rely on alternative methods to communicate in daily life, such as body language or text-to-speech [1] . However, there are currently very few effective rehabilitative therapies for helping these patients improve their speech. Because of this, research is needed to develop better rehabilitative therapies. One such area of research is the use of involuntary acoustic learning. The Speech and Swallowing lab at Marquette University has an Electromagnetic Articulography (EMA) system to collect kinematic data and a software system called …


Spatio-Temporal Progression Of Cortical Activity Related To Continuous Overt And Covert Speech Production In A Reading Task, Jonathan S. Brumberg, Dean J. Krusienski, Shreya Chakrabarti, Aysegul Gunduz, Peter Brunner, Anthony L. Ritaccio, Gerwin Schalk Jan 2016

Spatio-Temporal Progression Of Cortical Activity Related To Continuous Overt And Covert Speech Production In A Reading Task, Jonathan S. Brumberg, Dean J. Krusienski, Shreya Chakrabarti, Aysegul Gunduz, Peter Brunner, Anthony L. Ritaccio, Gerwin Schalk

Electrical & Computer Engineering Faculty Publications

How the human brain plans, executes, and monitors continuous and fluent speech has remained largely elusive. For example, previous research has defined the cortical locations most important for different aspects of speech function, but has not yet yielded a definition of the temporal progression of involvement of those locations as speech progresses either overtly or covertly. In this paper, we uncovered the spatio-temporal evolution of neuronal population-level activity related to continuous overt speech, and identified those locations that shared activity characteristics across overt and covert speech. Specifically, we asked subjects to repeat continuous sentences aloud or silently while we recorded …


An Axisymmetric Finite Element Model To Study The Earplug Contribution To The Bone Conduction Occlusion Effect, Martin K. Brummund, Franck Sgard, Yvan Petit, Frédéric Laville, Hugues Nélisse Jan 2015

An Axisymmetric Finite Element Model To Study The Earplug Contribution To The Bone Conduction Occlusion Effect, Martin K. Brummund, Franck Sgard, Yvan Petit, Frédéric Laville, Hugues Nélisse

Études primaires

An axisymmetric linear elasto-acoustic finite element (FE) model of an occluded human external ear is proposed to simulate the bone conduction occlusion effect (OE). The model consists of a cylindrical ear canal cavity surrounded by layers of biological tissues (skin, cartilage and bone) in which an earplug is inserted. Geometrical and material properties are taken from the literature. OEs are predicted for foam and silicone earplug FE-models using COMSOL Multiphysics (COMSOL®, Sweden). The FE-model is shown to predict the experimental OE measured in two healthy human reference groups wearing foam or silicone earplugs, satisfactorily. Deviations between model and experiment are …


Current Steering And Electrode Spanning With Partial Tripolar Stimulation Mode In Cochlear Implants, Ching-Chih Wu Oct 2014

Current Steering And Electrode Spanning With Partial Tripolar Stimulation Mode In Cochlear Implants, Ching-Chih Wu

Open Access Dissertations

Cochlear implants (CIs) partially restore hearing sensation to profoundly deaf people by electrically stimulating the surviving auditory neurons. However, CI users perform poorly in challenging listening tasks such as speech recognition in noise and Cochlear implants (CIs) partially restore hearing sensation to profoundly deaf people by electrically stimulating the surviving auditory neurons. However, CI users perform poorly in challenging listening tasks such as speech recognition in noise and music perception, possibly due to the small number of implanted electrodes and the large current spread of electric stimulation. Although current spread may be reduced using partial tripolar (pTP) stimulation mode, the …


Preliminary Test Of A Real-Time, Interactive Silent Speech Interface Based On Electromagnetic Articulograph, Jun Wang, Ashok Samal, Jordan R. Green Jun 2014

Preliminary Test Of A Real-Time, Interactive Silent Speech Interface Based On Electromagnetic Articulograph, Jun Wang, Ashok Samal, Jordan R. Green

School of Computing: Conference and Workshop Papers

A silent speech interface (SSI) maps articulatory movement data to speech output. Although still in experimental stages, silent speech interfaces hold significant potential for facilitating oral communication in persons after laryngectomy or with other severe voice impairments. Despite the recent efforts on silent speech recognition algorithm development using offline data analysis, online test of SSIs have rarely been conducted. In this paper, we present a preliminary, online test of a real-time, interactive SSI based on electromagnetic motion tracking. The SSI played back synthesized speech sounds in response to the user’s tongue and lip movements. Three English talkers participated in this …


The Electromagnetic Articulography Mandarin Accented English (Ema-Mae) Corpus Of Acoustic And 3d Articulatory Kinematic Data, Jeffrey J. Berry, An Ji, Michael T. Johnson May 2014

The Electromagnetic Articulography Mandarin Accented English (Ema-Mae) Corpus Of Acoustic And 3d Articulatory Kinematic Data, Jeffrey J. Berry, An Ji, Michael T. Johnson

Speech Pathology and Audiology Faculty Research and Publications

There is a significant need for more comprehensive electromagnetic articulography (EMA) datasets that can provide matched acoustics and articulatory kinematic data with good spatial and temporal resolution. The Marquette University Electromagnetic Articulography Mandarin Accented English (EMA-MAE) corpus provides kinematic and acoustic data from 40 gender and dialect balanced speakers representing 20 Midwestern standard American English L1 speakers and 20 Mandarin Accented English (MAE) L2 speakers, half Beijing region dialect and half are Shanghai region dialect. Three dimensional EMA data were collected at a 400 Hz sampling rate using the NDI Wave system, with articulatory sensors on the midsagittal lips, lower …


Sensorimotor Adaptation Of Speech Using Real-Time Articulatory Resynthesis, Jeffrey J. Berry, Cassandra North, Michael T. Johnson May 2014

Sensorimotor Adaptation Of Speech Using Real-Time Articulatory Resynthesis, Jeffrey J. Berry, Cassandra North, Michael T. Johnson

Speech Pathology and Audiology Faculty Research and Publications

Sensorimotor adaptation is an important focus in the study of motor learning for non-disordered speech, but has yet to be studied substantially for speech rehabilitation. Speech adaptation is typically elicited experimentally using LPC resynthesis to modify the sounds that a speaker hears himself producing. This method requires that the participant be able to produce a robust speech-acoustic signal and is therefore not well-suited for talkers with dysarthria. We have developed a novel technique using electromagnetic articulography (EMA) to drive an articulatory synthesizer. The acoustic output of the articulatory synthesizer can be perturbed experimentally to study auditory feedback effects on sensorimotor …


Computational Multimedia For Video Self Modeling, Ju Shen Jan 2014

Computational Multimedia For Video Self Modeling, Ju Shen

Theses and Dissertations--Computer Science

Video self modeling (VSM) is a behavioral intervention technique in which a learner models a target behavior by watching a video of oneself. This is the idea behind the psychological theory of self-efficacy - you can learn or model to perform certain tasks because you see yourself doing it, which provides the most ideal form of behavior modeling. The effectiveness of VSM has been demonstrated for many different types of disabilities and behavioral problems ranging from stuttering, inappropriate social behaviors, autism, selective mutism to sports training. However, there is an inherent difficulty associated with the production of VSM material. Prolonged …


Developing A Drug Delivery System For Treatment Of Vocal Fold Scarring, Aaron Michael Kosinski Oct 2013

Developing A Drug Delivery System For Treatment Of Vocal Fold Scarring, Aaron Michael Kosinski

Open Access Dissertations

Vocal fold scarring is an affliction that results in the formation of a disorganized and stiff extracellular matrix (ECM) with abnormal ECM component densities & structures including a significant increase in collagen deposition. It is caused by improper healing post injury and results in profound changes in the biomechanical properties of the vocal folds impairing their ability to generate a normal mucosal wave during phonation.

Finding an effective treatment for vocal fold scarring has been elusive. Currently, treatments seek temporary solutions that correct glottal incompetence and reduce stiffness caused by the scar through the augmentation of the vocal folds using …


Effects Of Hearing Aid Amplification On Robust Neural Coding Of Speech, Jonathan Daniel Boley Oct 2013

Effects Of Hearing Aid Amplification On Robust Neural Coding Of Speech, Jonathan Daniel Boley

Open Access Dissertations

Hearing aids are able to restore some hearing abilities for people with auditory impairments, but background noise remains a significant problem. Unfortunately, we know very little about how speech is encoded in the auditory system, particularly in impaired systems with prosthetic amplifiers. There is growing evidence that relative timing in the neural signals (known as spatiotemporal coding) is important for speech perception, but there is little research that relates spatiotemporal coding and hearing aid amplification.

This research uses a combination of computational modeling and physiological experiments to characterize how hearing aids affect vowel coding in noise at the level of …


Individual Articulator's Contribution To Phoneme Production, Jun Wang, Jordan R. Green, Ashok Samal May 2013

Individual Articulator's Contribution To Phoneme Production, Jun Wang, Jordan R. Green, Ashok Samal

School of Computing: Conference and Workshop Papers

Speech sounds are the result of coordinated movements of individual articulators. Understanding each articulator’s role in speech is fundamental not only for understanding how speech is produced, but also for optimizing speech assessments and treatments. In this paper, we studied the individual contributions of six articulators, tongue tip, tongue blade, tongue body front, tongue body back, upper lip, and lower lip to phoneme classification. A total of 3,838 vowel and consonant production samples were collected from eleven native English speakers. The results of speech movement classification using a support vector machine indicated that the tongue encoded significantly more information than …


Modeling Hrtf For Sound Localization In Normal Listeners And Bilateral Cochlear Implant Users, Douglas A. Miller Jan 2013

Modeling Hrtf For Sound Localization In Normal Listeners And Bilateral Cochlear Implant Users, Douglas A. Miller

Electronic Theses and Dissertations

Mathematical models can be very useful for understanding complicated systems and for testing algorithms through simulation that would be difficult or expensive to implement. This dissertation presents a model that attempts to simulate the sound localization performance of persons using bilateral cochlear implants. The expectation is that this model could prove to be a useful tool in developing new signal processing algorithms for neural encoding strategies.

The head related transfer function (HRTF) is a critical component of this model, and in the ideal case, provides the base characteristics of head shadow, torso and pinna effects. This defines the temporal, intensity …


Age-Related Changes To The Production Of Linguistic Prosody, Daniel Richard Barnes Jan 2013

Age-Related Changes To The Production Of Linguistic Prosody, Daniel Richard Barnes

Open Access Theses

The production of speech prosody (the rhythm, pausing, and intonation associated with natural speech) is critical to effective communication. The current study investigated the impact of age-related changes to physiology and cognition in relation to the production of two types of linguistic prosody: lexical stress and the disambiguation of syntactically ambiguous utterances. Analyses of the acoustic correlates of stress: speech intensity (or sound-pressure level; SPL), fundamental frequency (F0), key word/phrase duration, and pause duration revealed that both young and older adults effectively use these acoustic features to signal linguistic prosody, although the relative weighting of cues differed by group. Differences …


Whole-Word Recognition From Articulatory Movements For Silent Speech Interfaces, Jun Wang, Ashok Samal, Jordan R. Green, Frank Rudzicz Sep 2012

Whole-Word Recognition From Articulatory Movements For Silent Speech Interfaces, Jun Wang, Ashok Samal, Jordan R. Green, Frank Rudzicz

Department of Special Education and Communication Disorders: Faculty Publications

Articulation-based silent speech interfaces convert silently produced speech movements into audible words. These systems are still in their experimental stages, but have significant potential for facilitating oral communication in persons with laryngectomy or speech impairments. In this paper, we report the result of a novel, real-time algorithm that recognizes whole-words based on articulatory movements. This approach differs from prior work that has focused primarily on phoneme-level recognition based on articulatory features. On average, our algorithm missed 1.93 words in a sequence of twenty-five words with an average latency of 0.79 seconds for each word prediction using a data set of …


Tracking Articulator Movements Using Orientation Measurements, An Ji, Michael T. Johnson, Jeffrey J. Berry Jan 2012

Tracking Articulator Movements Using Orientation Measurements, An Ji, Michael T. Johnson, Jeffrey J. Berry

Speech Pathology and Audiology Faculty Research and Publications

This paper introduces a new method to track articulator movements, specifically jaw position and angle, using 5 degree of freedom (5 DOF) orientation data. The approach uses a quaternion rotation method to accomplish this jaw tracking during speech using a single senor on the mandibular incisor. Data were collected using the NDI Wave Speech Research System for one pilot subject with various speech tasks. The degree of jaw rotation from the proposed approach is compared with traditional geometric calculation. Results show that the quaternion based method is able to describe jaw angle trajectory and gives more accurate and smooth estimation …


Intelligibility Of Electrolarynx Speech Using A Novel Hands-Free Actuator, Brian Madden, Mark Nolan, Ted Burke, James Condron, Eugene Coyle Jan 2011

Intelligibility Of Electrolarynx Speech Using A Novel Hands-Free Actuator, Brian Madden, Mark Nolan, Ted Burke, James Condron, Eugene Coyle

Conference Papers

During voiced speech, the larynx provides quasi-periodic acoustic excitation of the vocal tract. In most electrolarynxes, mechanical vibrations are produced by a linear electromechanical actuator, the armature of which percusses against a metal or plastic plate at a frequency within the range of glottal excitation. In this paper, the intelligibility of speech produced using a novel hands-free actuator is compared to speech produced using a conventional electrolarynx. Two able-bodied speakers (one male, one female) performed a closed response test containing 28 monosyllabic words, once using a conventional electrolarynx and a second time using the novel design. The resulting audio recordings …


Augmented Control Of A Hands-Free Electrolarynx, Brian Madden, James Condron, Eugene Coyle Jan 2011

Augmented Control Of A Hands-Free Electrolarynx, Brian Madden, James Condron, Eugene Coyle

Conference Papers

During voiced speech, the larynx acts as the sound source, providing a quasi-periodic excitation of the vocal tract. Following a total laryngectomy, some people speak using an electrolarynx which employs an electromechanical actuator to perform the excitatory function of the absent larynx. Drawbacks of conventional electrolarynx designs include the monotonic sound emitted, the need for a free-hand to operate the device, and the difficulty experienced by many laryngectomees in adapting to its use. One improvement to the electrolarynx, which clinicians and users frequently suggest, is the provision of a convenient hands-free control facility. This would allow more natural use of …


Intelligibility Of Electrolarynx Speech Using A Novel Actuator, Brian Madden, Mark Nolan, Ted Burke, James Condron, Eugene Coyle Jun 2010

Intelligibility Of Electrolarynx Speech Using A Novel Actuator, Brian Madden, Mark Nolan, Ted Burke, James Condron, Eugene Coyle

Conference Papers

During voiced speech, the larynx provides quasi-periodic acoustic excitation of the vocal tract. Following a laryngectomy, some people speak using an electrolarynx which replaces the excitatory function of the absent larynx. Drawbacks of conventional electrolarynx designs include the buzzing monotonic sound emitted, the need for a free hand to operate the device, and difficulty experienced by many laryngectomees in adapting to its use. Despite these shortcomings, it remains the preferred method of speech rehabilitation for a substantial minority of laryngectomees. In most electrolarynxes, mechanical vibrations are produced by a linear electromechanical actuator, the armature of which percusses against a metal …


A Model For Electrical Communication Between Cochlear Implants And The Brain, Douglas A. Miller Jan 2009

A Model For Electrical Communication Between Cochlear Implants And The Brain, Douglas A. Miller

Electronic Theses and Dissertations

In the last thirty years, cochlear implants have become an invaluable instrument in the treatment of severe-to-profound hearing impairment. An important aspect of research in the continued development of cochlear implants is the in vivo assessment of signal processing algorithms intended to improve perception of speech and other auditory signals. In trying to determine how closely cochlear implant recipients process sound relative to the processing done by a normal auditory system, various assessment techniques have been applied. The most common technique has been measurement of auditory evoked potentials (AEPs), which involves the recording of neural responses to auditory stimulation. Depending …


Health Prognisis Of Electronics Via Power Profiling, Jonathan Amilcar Cervantes Jan 2009

Health Prognisis Of Electronics Via Power Profiling, Jonathan Amilcar Cervantes

Open Access Theses & Dissertations

The objective of this research is to investigate a new approach for the early detection of latent defects in electronic devices in the field. Reliability is assessed through the non-traditional approach of recording and evaluating the power profile of electronic devices within a deterministic state of operation. Traditionally, measuring the quiescent current (Iddq) of a device has been employed in manufacturing tests to detect defective parts prior to deployment to the field. However, the monitoring of the deterministic power signature (i.e. boot up or during a self-test routine) has never been exploited to monitor the health of a device in …


Spectral Analysis Of Pathological Acoustic Speech Waveforms, Priyanka Medida Jan 2009

Spectral Analysis Of Pathological Acoustic Speech Waveforms, Priyanka Medida

UNLV Theses, Dissertations, Professional Papers, and Capstones

Biomedical engineering is the application of engineering principles and techniques to the medical field. The design and problem solving skills of engineering are combined with medical and biological science, which improves medical disorder diagnosis and treatment. The purpose of this study is to develop an automated procedure for detecting excessive jitter in speech signals, which is useful for differentiating normal from pathologic speech. The fundamental motivation for this research is that tools are needed by speech pathologists and laryngologists for use in the early detection and treatment of laryngeal disorders. Acoustical analysis of speech was performed to analyze various features …


Gaussian Mixture Models And Neural Networks For Automatic Speaker Identification, Usha Gayatri Chalkapally Jul 2006

Gaussian Mixture Models And Neural Networks For Automatic Speaker Identification, Usha Gayatri Chalkapally

Electrical & Computer Engineering Theses & Dissertations

Automatic Speaker Recognition is the process of automatically recognizing who is speaking on the basis of individual information contained in speech signals. This technique of Automatic Speaker Recognition makes it possible to use the speaker's voice to verify their identity and control access to services such as voice dialing, banking by telephone, telephone shopping, database access services, information services, voice mail, security control for confidential information areas, and remote access to computers.

In this thesis, the techniques of Gaussian Mixture Models and Neural Networks for Automatic Speaker Identification are presented. Algorithms for Speaker Identification using Gaussian Mixture Models were developed, …


A Computer-Based Articulation Training Aid For Short Words (Cata), Mukund Devarajan Oct 2003

A Computer-Based Articulation Training Aid For Short Words (Cata), Mukund Devarajan

Electrical & Computer Engineering Theses & Dissertations

Several improvements in the vowel articulation training aid (VATA) are described, as well as the efforts to extend the visual feedback system to operate with short words in the form of consonant, vowel and consonant (CVC). The extended version of the visual feedback system is referred to as CATA (Computer-based Articulation Training Aid); the vowel version of the aid (VATA) only operates with ten American English monopthong vowels. Improvements in VATA include the use of a neural network (NN) recognizer method to prune a large database of vowel recordings to eliminate noisy and/or mispronounced tokens. The spectral jitter problem, previously …


Automatic Speaker Identification Using Reusable And Retrainable Binary-Pair Partitioned Neural Networks, Ashutosh Mishra Apr 2003

Automatic Speaker Identification Using Reusable And Retrainable Binary-Pair Partitioned Neural Networks, Ashutosh Mishra

Electrical & Computer Engineering Theses & Dissertations

This thesis presents an extension of the work previously done on speaker identification using Binary Pair Partitioned (BPP) neural networks. In the previous work, a separate network was used for each pair of speakers in the speaker population. Although the basic BPP approach did perform well and had a simple underlying algorithm, it had the obvious disadvantage of requiring an extremely large number of networks for speaker identification with large speaker populations. It also requires training of networks proportional to the square of the number of speakers under consideration, leading to a very large number of networks to be trained …


Yet Another Algorithm For Pitch Tracking (Yaapt), Kavita Kasi Oct 2002

Yet Another Algorithm For Pitch Tracking (Yaapt), Kavita Kasi

Electrical & Computer Engineering Theses & Dissertations

This thesis presents a pitch detection algorithm that is extremely robust for both high quality and telephone speech. The kernel method for this algorithm is the Normalized Cross Correlation (NCCF) reported by David Talkin [16]. Major innovations include: processing of the original acoustic signal and a nonlinearly processed version of the signal to partially restore very weak F0 components; intelligent peak picking to select multiple F0 candidates and assign merit factors; and, incorporation of highly robust pitch contours obtained from smoothed versions of low frequency portions of spectrograms. Dynamic programming is used to find the ''best" pitch track among all …


Minimum Mean Square Error Spectral Peak Envelope Estimation For Automatic Vowel Classification, Jaishree Venugopal Jul 2001

Minimum Mean Square Error Spectral Peak Envelope Estimation For Automatic Vowel Classification, Jaishree Venugopal

Electrical & Computer Engineering Theses & Dissertations

Spectral feature computations continue to be a very difficult problem for accurate machine recognition of speech. In this work, which focuses on vowels, a new spectral peak envelope method for vowel classification is developed, based on a missing frequency components model of speech recognition. According to the missing frequency components model, vowel recognition depends only on the spectral (harmonic) peaks. Smoothing and interpolation of the spectra, performed in the standard cepstral analysis method commonly used in automatic speech recognition, actually loses valuable information and results in reduced recognition accuracy. The new method for feature extraction presented in this thesis is …


Signal Modeling With Non-Uniform Time Sampling Of Features For Automatic Speech Recognition, Montri Karnjanadecha Jul 2000

Signal Modeling With Non-Uniform Time Sampling Of Features For Automatic Speech Recognition, Montri Karnjanadecha

Electrical & Computer Engineering Theses & Dissertations

This dissertation presents an investigation of non-uniform time sampling methods for spectral/temporal feature extraction in speech. Frame-based features were computed based on an encoding of the global spectral shape using a Discrete Cosine Transform. In most current “standard” methods, trajectory (dynamic) features are determined from frame-based parameters using a fixed time sampling, i.e., fixed block length and fixed block spacing. In this research, new methods are proposed and investigated in which block length and/or block spacing are variable. The idea was initially tested with HMM-based isolated word recognition, and a significant performance improvement resulted when a variable block length and …


Variability Analysis Of Discrete Cosine Transform Coefficient (Dctc) Features For Speech Processing, Bingjun Dai Oct 1998

Variability Analysis Of Discrete Cosine Transform Coefficient (Dctc) Features For Speech Processing, Bingjun Dai

Electrical & Computer Engineering Theses & Dissertations

In this research, the variability of Discrete Cosine Transform Coefficient (DCTC) features was investigated. Additionally, a new pitch-synchronous processing method was explored to increase the stability of features and to reduce window effects when compared to the regular method. The noise sources that lead to feature variability were analyzed, and different smoothing methods were tested. It was found that longer frames, frequency warping, time smoothing of the log spectrum, and DCS level time smoothing, all help reduce DCTC variability and increase classification performance. The pitch­ synchronous method was implemented with Matlab. Important processing methods, including pitch period estimation, time­ domain …