Open Access. Powered by Scholars. Published by Universities.®

Engineering

Institution
Keyword
Publication Year
Publication
Publication Type

Articles 1 - 30 of 47

Full-Text Articles in Speech and Hearing Science

Guitar Amplifier Directivity, Rachel C. Edelman, Brian E. Anderson, Samuel D. Bellows, Timothy W. Leishman Jan 2026

Guitar Amplifier Directivity, Rachel C. Edelman, Brian E. Anderson, Samuel D. Bellows, Timothy W. Leishman

Directivity

No abstract provided.


Revealing Spatiotemporal Neural Activation Patterns In Electrocorticography Recordings Of Human Speech Production By Mutual Information, Julio Kovacs, Dean Krusienski, Minu Maninder, Willy Wriggers Jan 2025

Revealing Spatiotemporal Neural Activation Patterns In Electrocorticography Recordings Of Human Speech Production By Mutual Information, Julio Kovacs, Dean Krusienski, Minu Maninder, Willy Wriggers

Mechanical & Aerospace Engineering Faculty Publications

Background

Spatiotemporal mapping of neural activity during continuous speech production has been traditionally approached using correlation coefficient (CC) analysis between cortical signals and speech recordings. A prior study employed this approach using electrocorticography (ECoG) data from participants who underwent invasive intracranial monitoring for epilepsy. However, CC cannot detect nonlinear relationships and is dominated by the correspondence between periods of silence and of non-silence.

New Method

We introduce the mutual information (MI) measure, which can capture both linear and nonlinear dependencies. We validated CC and MI on the sub-second spatiotemporal brain activity recorded during continuous speech tasks. To refine the results, …


Reliability Of An Extended Version Of The 3m™ Eargage Tool To Assess Earcanal Size And Assist Earplug Selection, Bastien Poissenot-Arrigoni, Laurence Martin, Alessia Negrini, Djamal Berbiche, Olivier Doutres, Franck Sgard Jan 2025

Reliability Of An Extended Version Of The 3m™ Eargage Tool To Assess Earcanal Size And Assist Earplug Selection, Bastien Poissenot-Arrigoni, Laurence Martin, Alessia Negrini, Djamal Berbiche, Olivier Doutres, Franck Sgard

Études primaires

Objective: Evaluate the ability of an extended version of the 3 MTM Eargage to estimate the earcanal size and assess the likelihood that a particular earplug can fit an individual’s earcanal, ultimately serving as a tool for selecting earplugs in the field. Design: Earcanal morphology, assessed through earcanal earmolds scans, is compared to earcanal size assessed with the extended eargage (EE) via box plots and Pearson linear correlations coefficients. Relations between attenuation measured on participants (for 6 different earplugs) and their earcanal size assessed with the EE are established via comparison tests. Study sample: 121 participants exposed to occupational noise …


Spatial Hearing In Simulated Reverberant Classroom Environments, Gabriel Seth Evan Weeldreyer May 2024

Spatial Hearing In Simulated Reverberant Classroom Environments, Gabriel Seth Evan Weeldreyer

Durham School of Architectural Engineering and Construction: Dissertations, Theses, and Student Research

Spatial hearing provides access to auditory spatial cues that promote speech perception in noisy listening situations. However, reverberation degrades auditory spatial cues and limits listeners’ ability to utilize these cues for segregating target speech from competing babble. Hence, spatial unmasking—an intelligibility benefit from a spatial separation between a target and masker—is reduced in reverberant environments as compared to free field. This work tests the hypothesis that interaural decorrelation, the result of increasing reverberation, will broaden the perceived auditory source width with a cascading effect of reduced auditory spatial acuity and subsequently poorer spatial unmasking. To understand the perceptual consequences of …


An Impedance Tube Technique For Estimating The Insertion Loss Of Earplugs, K. Carillo, O. Doutres, F. Sgard Jan 2024

An Impedance Tube Technique For Estimating The Insertion Loss Of Earplugs, K. Carillo, O. Doutres, F. Sgard

Études primaires

This paper proposes a quick and straightforward technique for estimating the insertion loss (IL) of earplugs measured on an acoustical test fixture (ATF) using a commercial impedance tube. In this method, the earplug's acoustic properties (i.e., its transmission loss and the reflection coefficient of its medial surface) are determined from its transfer matrix measured using the three-microphones impedance tube method modified here for the current application. The IL is then estimated using a one-dimensional analytical model of open and occluded earcanals based on the wavefield decomposition theory. The method is evaluated numerically and experimentally from 50 Hz to 6.5 kHz. …


Measurement Of The Local Static Mechanical Pressure Of Earplugs, Luiz G.C. Melo, Ahmed S. Dalaq, Franck Sgard, Olivier Doutres, Laurianne Legroux, Eric Wagnac Jan 2024

Measurement Of The Local Static Mechanical Pressure Of Earplugs, Luiz G.C. Melo, Ahmed S. Dalaq, Franck Sgard, Olivier Doutres, Laurianne Legroux, Eric Wagnac

Études primaires

Earplugs are used in various critical industrial sectors, such as construction, aviation, military and defense, transportation, and healthcare. However, they present inherent drawbacks, notably discomfort that can lead to inconsistent and incorrect use, thereby leaving a significant proportion of workers vulnerable to irreversible hearing damage, ranging from tinnitus to deafness. This discomfort is closely linked to a physical parameter known as static mechanical pressure (SMP), representing the pressure exerted by earplugs on the earcanal walls. Determining the SMP is crucial for developing earplugs that prioritize comfort. However, experimental studies in this area are scarce and there is no experimental set-up …


Trumpet Directivity From A Rotating Semicircular Array, Samuel D. Bellows, Joseph E. Avila, Timothy W. Leishman Sep 2023

Trumpet Directivity From A Rotating Semicircular Array, Samuel D. Bellows, Joseph E. Avila, Timothy W. Leishman

Directivity

The directivity function of a played musical instrument describes the angular dependence of its acoustic radiation and diffraction about the instrument, musician, and musician’s chair. Directivity influences sound in rehearsal, performance, and recording environments and signals in audio systems. Because high-resolution, spherically comprehensive measurements of played musical instruments have been unavailable in the past, the authors have undertaken research to produce and share such data for studies of musical instruments, simulations of acoustical environments, optimizations of microphone placements, and other applications. The authors acquired the data from repeated chromatic scales produced by a trumpet played at mezzo-forte in an anechoic …


A Proof Of Concept For Detection Of Palato-Lingual Contact Without Intraoral Electrodes In A Human Participant, Amitava Biswas May 2023

A Proof Of Concept For Detection Of Palato-Lingual Contact Without Intraoral Electrodes In A Human Participant, Amitava Biswas

Master's Theses

The author of this thesis conceived and investigated an unexpected electrical signal that may not be noticed during regular electroglottography (EGG) when electrodes are conventionally placed on the anterior surface of the throat. It appears that there is a measurable electrical signal from the EGG equipment when electrodes are placed over the frontal cheek (belly of zygomatic muscle), and the tongue is elevated to contact the hard palate, or the tongue is lowered to break the contact. Therefore, this phenomenon is designated as unconventional electroglottography (UEGG, Dr. Steven Cloud, private communication, July 17, 2022).

Two distinctive patterns of waveforms were …


Gamelan Gong Directivity Dataset, Samuel D. Bellows, Dallin T. Harwood, Kent L. Gee, Micah R. Shepherd Jan 2023

Gamelan Gong Directivity Dataset, Samuel D. Bellows, Dallin T. Harwood, Kent L. Gee, Micah R. Shepherd

Directivity

No abstract provided.


Brain Activity Associated With Taste Stimulation: A Mechanism For Neuroplastic Change?, Angela M. Dietsch, Ross M. Westemeyer, Douglas H. Schultz Jan 2023

Brain Activity Associated With Taste Stimulation: A Mechanism For Neuroplastic Change?, Angela M. Dietsch, Ross M. Westemeyer, Douglas H. Schultz

Department of Special Education and Communication Disorders: Faculty Publications

Purpose: Neuroplasticity may be enhanced by increasing brain activation and bloodflow in neural regions relevant to the target behavior.We administered precisely formulated and dosed taste stimuli to determine whether the associated brain activity patterns included areas that underlie swallowing control.

Methods: Five taste stimuli (unflavored, sour, sweet-sour, lemon, and orange suspensions) were administered in timing-regulated and temperature-controlled 3 mL doses via a customized pump/tubing system to 21 healthy adults during functional magnetic resonance imaging (fMRI). Whole-brain analyses of fMRI data assessed main effects of taste stimulation as well as differential effects of taste profile.

Results: Differences in …


Passive Earplug Including Helmholtz Resonators Arranged In Series To Achieve Broadband Near Zero Occlusion Effect At Low Frequencies, Kevin Carillo, Franck Sgard, Olivier Dazel, Olivier Doutres Jan 2023

Passive Earplug Including Helmholtz Resonators Arranged In Series To Achieve Broadband Near Zero Occlusion Effect At Low Frequencies, Kevin Carillo, Franck Sgard, Olivier Dazel, Olivier Doutres

Études primaires

The use of passive earplugs is often associated with the occlusion effect: a phenomenon described as the increased auditory perception of one's own physiological noise at low frequencies. As a notable acoustic discomfort, the occlusion effect penalizes the use and the efficiency of earplugs. This phenomenon is objectively characterized by the increase in sound pressure level in the occluded ear canal compared to the open ear canal. Taking inspiration from acoustic metamaterials, a new design of a three-dimensional printed “meta-earplug,” made of four Helmholtz resonators arranged in series, is proposed for achieving near zero objective occlusion effect measured on artificial …


Pressure Induced By Roll-Down Foam-Earplugs On Earcanal, Ahmed Dalaq, Luiz Melo, Franck Sgard, Olivier Doutres, Éric Wagnac Jan 2023

Pressure Induced By Roll-Down Foam-Earplugs On Earcanal, Ahmed Dalaq, Luiz Melo, Franck Sgard, Olivier Doutres, Éric Wagnac

Études primaires

Physical discomfort of earplug is a common complaint, so much so it jeopardizes the health and safety of workers, resulting in frequent noncompliance to proper wearing of earplugs. A pressing need thus arises to understand the underlying mechanics for the interaction of earplugs with earcanal walls. An idealized cylindrical geometry of earcanal is first treated with computational and analytical modeling to predict the pressure induced by roll-down cylindrical PVC foam earplugs. In order to predict representative pressure values and distribution, we estimated the hyperelastic properties of foam-based earplugs and characterized their behavior using a stent testing machine (J-CrimpTM machine) …


Augmentative And Alternative Communication Use, Service Delivery Experiences, And Communicative Participation For People With Amyotrophic Lateral Sclerosis, Betts Peters Jun 2022

Augmentative And Alternative Communication Use, Service Delivery Experiences, And Communicative Participation For People With Amyotrophic Lateral Sclerosis, Betts Peters

Dissertations and Theses

People with amyotrophic lateral sclerosis (ALS) often experience changes to their speech, and may use augmentative and alternative communication (AAC) devices and techniques to maintain the ability to communicate. The use of AAC may facilitate the participation of people with ALS in various life situations involving communication. There is limited data in the literature about the AAC approaches currently used by people with ALS, the professional services they receive to support communication, or the effects of AAC on their communicative participation. This dissertation involved a nationwide online survey of people with ALS, and comprises three papers intended to add to …


Kemar Hats Head Orientation Directivity, Samuel D. Bellows, Timothy W. Leishman Mar 2022

Kemar Hats Head Orientation Directivity, Samuel D. Bellows, Timothy W. Leishman

Directivity

This directivity data set for a KEMAR head head-and-torso simulator (HATS) includes head orientations in 14 directions in 5° steps starting from 0° to 40° and then in 10° steps from 40° to 90°. The full spherical measurements followed at an a = 0.97 m radius with the mouth aperture at the spherical center. The sampling density and distribution followed the AES 5° dual-equiangular sampling standard, omitting the south pole (θ = 180°). Thus, each spherical directivity assessment included 36 polar-angle θ samples and 72 azimuthal-angle ϕ samples. The presented data include 22 1/3-octave bands, ranging from 80 Hz …


Speaker Encoding For Zero-Shot Speech Synthesis, Tristin W. Cory Jan 2022

Speaker Encoding For Zero-Shot Speech Synthesis, Tristin W. Cory

Graduate Theses/Dissertations

Spoken communication, for many, is an essential part of everyday life. Some individuals can lose or not be born with the ability to speak. To function on a day-to-day basis, these individuals find other ways of communication. Adaptive speech synthesis is one of those ways. It recreates a user’s previous voice or creates a voice that blends with their regional dialect. Current adaptive speech synthesis techniques that achieve human-like speech require thirty minutes, to a few hours of high-quality audio recordings of a target speaker. This amount of recorded audio is not commonly possessed by people in need of a …


Average Speech Directivity, Samuel D. Bellows, Claire M. Pincock, Jennifer K. Whiting, Timothy W. Leishman Nov 2019

Average Speech Directivity, Samuel D. Bellows, Claire M. Pincock, Jennifer K. Whiting, Timothy W. Leishman

Directivity

Speech directivity describes the angular dependence of acoustic radiation from a talker’s mouth and nostrils and diffraction about his or her body and chair (if seated). It is an essential physical aspect of communication affecting sounds and signals in acoustical environments, audio, and telecommunication systems. Because high-resolution, spherically comprehensive measurements of live, phonetically balanced speech have been unavailable in the past, the authors have undertaken research to produce and share such data for simulations of acoustical environments, optimizations of microphone placements, speech studies, and other applications. The measurements included three male and three female talkers who repeated phonetically balanced passages …


Communication Of Deaf People Based On Myoware Muscle Sensor, Mohamad Ataya Jan 2019

Communication Of Deaf People Based On Myoware Muscle Sensor, Mohamad Ataya

Research Opportunities for Engineering Undergraduates (ROEU) Program 2018-19

The current project aims to create a product that enhances communication of deaf people. The developed product has the following features:

• a muscle sensor that transcribes sign language into actual letters

• a voice detector that translates speech into actual words

• the ability to receive and send emergency alerts.


Development And Validation For A Mobile Speech-In-Noise Audiometric Task, Tommy Peng Aug 2017

Development And Validation For A Mobile Speech-In-Noise Audiometric Task, Tommy Peng

McKelvey School of Engineering Graduate Student Theses & Dissertations

Traditional speech-in-noise hearing tests are performed by clinicians with specialized equipment. Furthermore, these tasks often present contextually weak sentences in background babble, which are poor representations of real-world situations. This study proposes a mobile audiometric task, Semantic Auditory Search, which uses the Android platform to bypass the need for specialized equipment and presents multiple tasks of two competing real-world conversations to estimate the user’s speech-in-noise hearing ability. Through linear regression models built from data of seventy-nine subjects, three Semantic Auditory Search metrics have been shown to have statistically significant (p < 0.05) with medium effects sizes for predicting QuickSIN SNR50. The internal consistency of the task was also high, with a Cronbach’s alpha of 0.88 or more across multiple metrics. In conclusion, this preliminary study suggests that Semantic Auditory Search can accurately and reliably perform as an automated speech-in-noise hearing test. It also has tremendous potential for extension into automated tests of cognitive function, as well.


Development Of Kinematic Templates For Automatic Pronunciation Assessment Using Acoustic-To-Articulatory Inversion, Deriq K. Jones Jul 2017

Development Of Kinematic Templates For Automatic Pronunciation Assessment Using Acoustic-To-Articulatory Inversion, Deriq K. Jones

Master's Theses (2009 -)

Computer-aided pronunciation training (CAPT) is a subcategory of computer-aided language learning (CALL) that deals with the correction of mispronunciation during language learning. For a CAPT system to be effective, it must provide useful and informative feedback that is comprehensive, qualitative, quantitative, and corrective. While the majority of modern systems address the first 3 aspects of feedback, most of these systems do not provide corrective feedback. As part of the National Science Foundation (NSF) funded study “RI: Small: Speaker Independent Acoustic-Articulator Inversion for Pronunciation Assessment”, the Marquette Speech and Swallowing Lab and Marquette Speech and Signal Processing Lab are conducting a …


The Effect Of Frequency Resolution On Intelligibility Sentence And Its Relevance To Cochlear Implant Design, Seth H. Roy Dec 2016

The Effect Of Frequency Resolution On Intelligibility Sentence And Its Relevance To Cochlear Implant Design, Seth H. Roy

School of Computing: Dissertations, Theses, and Student Research

The purpose of this study is to understand how electrical stimulation (as opposed to acoustical stimulation) of the auditory nerve is used in cochlear implants. Speech is a complex signal that changes rapidly in time and frequency domains. Since phonemes (the smallest unit of speech that distinguishes words) depend on nuanced differences in frequency patterns, it would be expected that a signal with drastically reduced frequency information would be of limited value for conveying speech. Such a frequency-poor signal is the object to be investigated in the present work. It is also the basis of the way speech is represented …


Effects Of Vocal Fold Nodules On Glottal Cycle Measurements Derived From High-Speed Videoendoscopy In Children, Rita R. Patel, Harikrishnan Unnikrishnan, Kevin D. Donohue Apr 2016

Effects Of Vocal Fold Nodules On Glottal Cycle Measurements Derived From High-Speed Videoendoscopy In Children, Rita R. Patel, Harikrishnan Unnikrishnan, Kevin D. Donohue

Electrical and Computer Engineering Faculty Publications

The goal of this study is to quantify the effects of vocal fold nodules on vibratory motion in children using high-speed videoendoscopy. Differences in vibratory motion were evaluated in 20 children with vocal fold nodules (5–11 years) and 20 age and gender matched typically developing children (5–11 years) during sustained phonation at typical pitch and loudness. Normalized kinematic features of vocal fold displacements from the mid-membranous vocal fold point were extracted from the steady-state high-speed video. A total of 12 kinematic features representing spatial and temporal characteristics of vibratory motion were calculated. Average values and standard deviations (cycle-to-cycle variability) of …


Spatio-Temporal Progression Of Cortical Activity Related To Continuous Overt And Covert Speech Production In A Reading Task, Jonathan S. Brumberg, Dean J. Krusienski, Shreya Chakrabarti, Aysegul Gunduz, Peter Brunner, Anthony L. Ritaccio, Gerwin Schalk Jan 2016

Spatio-Temporal Progression Of Cortical Activity Related To Continuous Overt And Covert Speech Production In A Reading Task, Jonathan S. Brumberg, Dean J. Krusienski, Shreya Chakrabarti, Aysegul Gunduz, Peter Brunner, Anthony L. Ritaccio, Gerwin Schalk

Electrical & Computer Engineering Faculty Publications

How the human brain plans, executes, and monitors continuous and fluent speech has remained largely elusive. For example, previous research has defined the cortical locations most important for different aspects of speech function, but has not yet yielded a definition of the temporal progression of involvement of those locations as speech progresses either overtly or covertly. In this paper, we uncovered the spatio-temporal evolution of neuronal population-level activity related to continuous overt speech, and identified those locations that shared activity characteristics across overt and covert speech. Specifically, we asked subjects to repeat continuous sentences aloud or silently while we recorded …


An Axisymmetric Finite Element Model To Study The Earplug Contribution To The Bone Conduction Occlusion Effect, Martin K. Brummund, Franck Sgard, Yvan Petit, Frédéric Laville, Hugues Nélisse Jan 2015

An Axisymmetric Finite Element Model To Study The Earplug Contribution To The Bone Conduction Occlusion Effect, Martin K. Brummund, Franck Sgard, Yvan Petit, Frédéric Laville, Hugues Nélisse

Études primaires

An axisymmetric linear elasto-acoustic finite element (FE) model of an occluded human external ear is proposed to simulate the bone conduction occlusion effect (OE). The model consists of a cylindrical ear canal cavity surrounded by layers of biological tissues (skin, cartilage and bone) in which an earplug is inserted. Geometrical and material properties are taken from the literature. OEs are predicted for foam and silicone earplug FE-models using COMSOL Multiphysics (COMSOL®, Sweden). The FE-model is shown to predict the experimental OE measured in two healthy human reference groups wearing foam or silicone earplugs, satisfactorily. Deviations between model and experiment are …


The Electromagnetic Articulography Mandarin Accented English (Ema-Mae) Corpus Of Acoustic And 3d Articulatory Kinematic Data, Jeffrey J. Berry, An Ji, Michael T. Johnson May 2014

The Electromagnetic Articulography Mandarin Accented English (Ema-Mae) Corpus Of Acoustic And 3d Articulatory Kinematic Data, Jeffrey J. Berry, An Ji, Michael T. Johnson

Speech Pathology and Audiology Faculty Research and Publications

There is a significant need for more comprehensive electromagnetic articulography (EMA) datasets that can provide matched acoustics and articulatory kinematic data with good spatial and temporal resolution. The Marquette University Electromagnetic Articulography Mandarin Accented English (EMA-MAE) corpus provides kinematic and acoustic data from 40 gender and dialect balanced speakers representing 20 Midwestern standard American English L1 speakers and 20 Mandarin Accented English (MAE) L2 speakers, half Beijing region dialect and half are Shanghai region dialect. Three dimensional EMA data were collected at a 400 Hz sampling rate using the NDI Wave system, with articulatory sensors on the midsagittal lips, lower …


Sensorimotor Adaptation Of Speech Using Real-Time Articulatory Resynthesis, Jeffrey J. Berry, Cassandra North, Michael T. Johnson May 2014

Sensorimotor Adaptation Of Speech Using Real-Time Articulatory Resynthesis, Jeffrey J. Berry, Cassandra North, Michael T. Johnson

Speech Pathology and Audiology Faculty Research and Publications

Sensorimotor adaptation is an important focus in the study of motor learning for non-disordered speech, but has yet to be studied substantially for speech rehabilitation. Speech adaptation is typically elicited experimentally using LPC resynthesis to modify the sounds that a speaker hears himself producing. This method requires that the participant be able to produce a robust speech-acoustic signal and is therefore not well-suited for talkers with dysarthria. We have developed a novel technique using electromagnetic articulography (EMA) to drive an articulatory synthesizer. The acoustic output of the articulatory synthesizer can be perturbed experimentally to study auditory feedback effects on sensorimotor …


Developing A Drug Delivery System For Treatment Of Vocal Fold Scarring, Aaron Michael Kosinski Oct 2013

Developing A Drug Delivery System For Treatment Of Vocal Fold Scarring, Aaron Michael Kosinski

Open Access Dissertations

Vocal fold scarring is an affliction that results in the formation of a disorganized and stiff extracellular matrix (ECM) with abnormal ECM component densities & structures including a significant increase in collagen deposition. It is caused by improper healing post injury and results in profound changes in the biomechanical properties of the vocal folds impairing their ability to generate a normal mucosal wave during phonation.

Finding an effective treatment for vocal fold scarring has been elusive. Currently, treatments seek temporary solutions that correct glottal incompetence and reduce stiffness caused by the scar through the augmentation of the vocal folds using …


Individual Articulator's Contribution To Phoneme Production, Jun Wang, Jordan R. Green, Ashok Samal May 2013

Individual Articulator's Contribution To Phoneme Production, Jun Wang, Jordan R. Green, Ashok Samal

School of Computing: Conference and Workshop Papers

Speech sounds are the result of coordinated movements of individual articulators. Understanding each articulator’s role in speech is fundamental not only for understanding how speech is produced, but also for optimizing speech assessments and treatments. In this paper, we studied the individual contributions of six articulators, tongue tip, tongue blade, tongue body front, tongue body back, upper lip, and lower lip to phoneme classification. A total of 3,838 vowel and consonant production samples were collected from eleven native English speakers. The results of speech movement classification using a support vector machine indicated that the tongue encoded significantly more information than …


Modeling Hrtf For Sound Localization In Normal Listeners And Bilateral Cochlear Implant Users, Douglas A. Miller Jan 2013

Modeling Hrtf For Sound Localization In Normal Listeners And Bilateral Cochlear Implant Users, Douglas A. Miller

Electronic Theses and Dissertations

Mathematical models can be very useful for understanding complicated systems and for testing algorithms through simulation that would be difficult or expensive to implement. This dissertation presents a model that attempts to simulate the sound localization performance of persons using bilateral cochlear implants. The expectation is that this model could prove to be a useful tool in developing new signal processing algorithms for neural encoding strategies.

The head related transfer function (HRTF) is a critical component of this model, and in the ideal case, provides the base characteristics of head shadow, torso and pinna effects. This defines the temporal, intensity …


Age-Related Changes To The Production Of Linguistic Prosody, Daniel Richard Barnes Jan 2013

Age-Related Changes To The Production Of Linguistic Prosody, Daniel Richard Barnes

Open Access Theses

The production of speech prosody (the rhythm, pausing, and intonation associated with natural speech) is critical to effective communication. The current study investigated the impact of age-related changes to physiology and cognition in relation to the production of two types of linguistic prosody: lexical stress and the disambiguation of syntactically ambiguous utterances. Analyses of the acoustic correlates of stress: speech intensity (or sound-pressure level; SPL), fundamental frequency (F0), key word/phrase duration, and pause duration revealed that both young and older adults effectively use these acoustic features to signal linguistic prosody, although the relative weighting of cues differed by group. Differences …


Whole-Word Recognition From Articulatory Movements For Silent Speech Interfaces, Jun Wang, Ashok Samal, Jordan R. Green, Frank Rudzicz Sep 2012

Whole-Word Recognition From Articulatory Movements For Silent Speech Interfaces, Jun Wang, Ashok Samal, Jordan R. Green, Frank Rudzicz

Department of Special Education and Communication Disorders: Faculty Publications

Articulation-based silent speech interfaces convert silently produced speech movements into audible words. These systems are still in their experimental stages, but have significant potential for facilitating oral communication in persons with laryngectomy or speech impairments. In this paper, we report the result of a novel, real-time algorithm that recognizes whole-words based on articulatory movements. This approach differs from prior work that has focused primarily on phoneme-level recognition based on articulatory features. On average, our algorithm missed 1.93 words in a sequence of twenty-five words with an average latency of 0.79 seconds for each word prediction using a data set of …