Open Access. Powered by Scholars. Published by Universities.®
- Institution
-
- Old Dominion University (15)
- Institut de Recherche Robert-Sauvé en santé et en sécurité du travail (6)
- Brigham Young University (5)
- University of Nebraska - Lincoln (5)
- Marquette University (3)
-
- Purdue University (2)
- Technological University Dublin (2)
- University of Denver (2)
- Missouri State University (1)
- Portland State University (1)
- The University of Southern Mississippi (1)
- University of Kentucky (1)
- Washington University in St. Louis (1)
- Wayne State University (1)
- Western Kentucky University (1)
- Keyword
-
- Earplug (6)
- Protège-tympan (6)
- Automatic speech recognition (5)
- Directivity (5)
- Ear (5)
-
- Oreille (5)
- Speech processing systems (5)
- Hearing protection (4)
- Protection de l'ouïe (4)
- Comfort assessment (3)
- Laryngectomy (3)
- Radiation (3)
- Speech (3)
- Speech synthesis (3)
- Évaluation du confort (3)
- Articulation (2)
- Binary control systems (2)
- Cochlear implant (2)
- Comfort criteria (2)
- Critère de confort (2)
- Daniel Felix Ritchie School of Engineering and Computer Science (2)
- Deaf people (2)
- Electro-larynx (2)
- Equipment testing (2)
- Essai du matériel (2)
- Hearing impaired (2)
- Humans (2)
- Intelligibility (2)
- Neural networks (2)
- Phonetics (2)
- Publication Year
- Publication
-
- Electrical & Computer Engineering Theses & Dissertations (13)
- Études primaires (6)
- Directivity (5)
- Conference Papers (2)
- Department of Special Education and Communication Disorders: Faculty Publications (2)
-
- Electronic Theses and Dissertations (2)
- Speech Pathology and Audiology Faculty Research and Publications (2)
- Dissertations and Theses (1)
- Durham School of Architectural Engineering and Construction: Dissertations, Theses, and Student Research (1)
- Electrical & Computer Engineering Faculty Publications (1)
- Electrical and Computer Engineering Faculty Publications (1)
- Graduate Theses/Dissertations (1)
- Master's Theses (1)
- Master's Theses (2009 -) (1)
- McKelvey School of Engineering Graduate Student Theses & Dissertations (1)
- Mechanical & Aerospace Engineering Faculty Publications (1)
- Open Access Dissertations (1)
- Open Access Theses (1)
- Research Opportunities for Engineering Undergraduates (ROEU) Program 2018-19 (1)
- School of Computing: Conference and Workshop Papers (1)
- School of Computing: Dissertations, Theses, and Student Research (1)
- WKU Administration Documents (1)
- Publication Type
Articles 1 - 30 of 47
Full-Text Articles in Speech and Hearing Science
Guitar Amplifier Directivity, Rachel C. Edelman, Brian E. Anderson, Samuel D. Bellows, Timothy W. Leishman
Guitar Amplifier Directivity, Rachel C. Edelman, Brian E. Anderson, Samuel D. Bellows, Timothy W. Leishman
Directivity
No abstract provided.
Revealing Spatiotemporal Neural Activation Patterns In Electrocorticography Recordings Of Human Speech Production By Mutual Information, Julio Kovacs, Dean Krusienski, Minu Maninder, Willy Wriggers
Revealing Spatiotemporal Neural Activation Patterns In Electrocorticography Recordings Of Human Speech Production By Mutual Information, Julio Kovacs, Dean Krusienski, Minu Maninder, Willy Wriggers
Mechanical & Aerospace Engineering Faculty Publications
Background
Spatiotemporal mapping of neural activity during continuous speech production has been traditionally approached using correlation coefficient (CC) analysis between cortical signals and speech recordings. A prior study employed this approach using electrocorticography (ECoG) data from participants who underwent invasive intracranial monitoring for epilepsy. However, CC cannot detect nonlinear relationships and is dominated by the correspondence between periods of silence and of non-silence.
New Method
We introduce the mutual information (MI) measure, which can capture both linear and nonlinear dependencies. We validated CC and MI on the sub-second spatiotemporal brain activity recorded during continuous speech tasks. To refine the results, …
Reliability Of An Extended Version Of The 3m™ Eargage Tool To Assess Earcanal Size And Assist Earplug Selection, Bastien Poissenot-Arrigoni, Laurence Martin, Alessia Negrini, Djamal Berbiche, Olivier Doutres, Franck Sgard
Reliability Of An Extended Version Of The 3m™ Eargage Tool To Assess Earcanal Size And Assist Earplug Selection, Bastien Poissenot-Arrigoni, Laurence Martin, Alessia Negrini, Djamal Berbiche, Olivier Doutres, Franck Sgard
Études primaires
Objective: Evaluate the ability of an extended version of the 3 MTM Eargage to estimate the earcanal size and assess the likelihood that a particular earplug can fit an individual’s earcanal, ultimately serving as a tool for selecting earplugs in the field. Design: Earcanal morphology, assessed through earcanal earmolds scans, is compared to earcanal size assessed with the extended eargage (EE) via box plots and Pearson linear correlations coefficients. Relations between attenuation measured on participants (for 6 different earplugs) and their earcanal size assessed with the EE are established via comparison tests. Study sample: 121 participants exposed to occupational noise …
Spatial Hearing In Simulated Reverberant Classroom Environments, Gabriel Seth Evan Weeldreyer
Spatial Hearing In Simulated Reverberant Classroom Environments, Gabriel Seth Evan Weeldreyer
Durham School of Architectural Engineering and Construction: Dissertations, Theses, and Student Research
Spatial hearing provides access to auditory spatial cues that promote speech perception in noisy listening situations. However, reverberation degrades auditory spatial cues and limits listeners’ ability to utilize these cues for segregating target speech from competing babble. Hence, spatial unmasking—an intelligibility benefit from a spatial separation between a target and masker—is reduced in reverberant environments as compared to free field. This work tests the hypothesis that interaural decorrelation, the result of increasing reverberation, will broaden the perceived auditory source width with a cascading effect of reduced auditory spatial acuity and subsequently poorer spatial unmasking. To understand the perceptual consequences of …
An Impedance Tube Technique For Estimating The Insertion Loss Of Earplugs, K. Carillo, O. Doutres, F. Sgard
An Impedance Tube Technique For Estimating The Insertion Loss Of Earplugs, K. Carillo, O. Doutres, F. Sgard
Études primaires
This paper proposes a quick and straightforward technique for estimating the insertion loss (IL) of earplugs measured on an acoustical test fixture (ATF) using a commercial impedance tube. In this method, the earplug's acoustic properties (i.e., its transmission loss and the reflection coefficient of its medial surface) are determined from its transfer matrix measured using the three-microphones impedance tube method modified here for the current application. The IL is then estimated using a one-dimensional analytical model of open and occluded earcanals based on the wavefield decomposition theory. The method is evaluated numerically and experimentally from 50 Hz to 6.5 kHz. …
Measurement Of The Local Static Mechanical Pressure Of Earplugs, Luiz G.C. Melo, Ahmed S. Dalaq, Franck Sgard, Olivier Doutres, Laurianne Legroux, Eric Wagnac
Measurement Of The Local Static Mechanical Pressure Of Earplugs, Luiz G.C. Melo, Ahmed S. Dalaq, Franck Sgard, Olivier Doutres, Laurianne Legroux, Eric Wagnac
Études primaires
Earplugs are used in various critical industrial sectors, such as construction, aviation, military and defense, transportation, and healthcare. However, they present inherent drawbacks, notably discomfort that can lead to inconsistent and incorrect use, thereby leaving a significant proportion of workers vulnerable to irreversible hearing damage, ranging from tinnitus to deafness. This discomfort is closely linked to a physical parameter known as static mechanical pressure (SMP), representing the pressure exerted by earplugs on the earcanal walls. Determining the SMP is crucial for developing earplugs that prioritize comfort. However, experimental studies in this area are scarce and there is no experimental set-up …
Trumpet Directivity From A Rotating Semicircular Array, Samuel D. Bellows, Joseph E. Avila, Timothy W. Leishman
Trumpet Directivity From A Rotating Semicircular Array, Samuel D. Bellows, Joseph E. Avila, Timothy W. Leishman
Directivity
The directivity function of a played musical instrument describes the angular dependence of its acoustic radiation and diffraction about the instrument, musician, and musician’s chair. Directivity influences sound in rehearsal, performance, and recording environments and signals in audio systems. Because high-resolution, spherically comprehensive measurements of played musical instruments have been unavailable in the past, the authors have undertaken research to produce and share such data for studies of musical instruments, simulations of acoustical environments, optimizations of microphone placements, and other applications. The authors acquired the data from repeated chromatic scales produced by a trumpet played at mezzo-forte in an anechoic …
A Proof Of Concept For Detection Of Palato-Lingual Contact Without Intraoral Electrodes In A Human Participant, Amitava Biswas
A Proof Of Concept For Detection Of Palato-Lingual Contact Without Intraoral Electrodes In A Human Participant, Amitava Biswas
Master's Theses
The author of this thesis conceived and investigated an unexpected electrical signal that may not be noticed during regular electroglottography (EGG) when electrodes are conventionally placed on the anterior surface of the throat. It appears that there is a measurable electrical signal from the EGG equipment when electrodes are placed over the frontal cheek (belly of zygomatic muscle), and the tongue is elevated to contact the hard palate, or the tongue is lowered to break the contact. Therefore, this phenomenon is designated as unconventional electroglottography (UEGG, Dr. Steven Cloud, private communication, July 17, 2022).
Two distinctive patterns of waveforms were …
Gamelan Gong Directivity Dataset, Samuel D. Bellows, Dallin T. Harwood, Kent L. Gee, Micah R. Shepherd
Gamelan Gong Directivity Dataset, Samuel D. Bellows, Dallin T. Harwood, Kent L. Gee, Micah R. Shepherd
Directivity
No abstract provided.
Brain Activity Associated With Taste Stimulation: A Mechanism For Neuroplastic Change?, Angela M. Dietsch, Ross M. Westemeyer, Douglas H. Schultz
Brain Activity Associated With Taste Stimulation: A Mechanism For Neuroplastic Change?, Angela M. Dietsch, Ross M. Westemeyer, Douglas H. Schultz
Department of Special Education and Communication Disorders: Faculty Publications
Purpose: Neuroplasticity may be enhanced by increasing brain activation and bloodflow in neural regions relevant to the target behavior.We administered precisely formulated and dosed taste stimuli to determine whether the associated brain activity patterns included areas that underlie swallowing control.
Methods: Five taste stimuli (unflavored, sour, sweet-sour, lemon, and orange suspensions) were administered in timing-regulated and temperature-controlled 3 mL doses via a customized pump/tubing system to 21 healthy adults during functional magnetic resonance imaging (fMRI). Whole-brain analyses of fMRI data assessed main effects of taste stimulation as well as differential effects of taste profile.
Results: Differences in …
Passive Earplug Including Helmholtz Resonators Arranged In Series To Achieve Broadband Near Zero Occlusion Effect At Low Frequencies, Kevin Carillo, Franck Sgard, Olivier Dazel, Olivier Doutres
Passive Earplug Including Helmholtz Resonators Arranged In Series To Achieve Broadband Near Zero Occlusion Effect At Low Frequencies, Kevin Carillo, Franck Sgard, Olivier Dazel, Olivier Doutres
Études primaires
The use of passive earplugs is often associated with the occlusion effect: a phenomenon described as the increased auditory perception of one's own physiological noise at low frequencies. As a notable acoustic discomfort, the occlusion effect penalizes the use and the efficiency of earplugs. This phenomenon is objectively characterized by the increase in sound pressure level in the occluded ear canal compared to the open ear canal. Taking inspiration from acoustic metamaterials, a new design of a three-dimensional printed “meta-earplug,” made of four Helmholtz resonators arranged in series, is proposed for achieving near zero objective occlusion effect measured on artificial …
Pressure Induced By Roll-Down Foam-Earplugs On Earcanal, Ahmed Dalaq, Luiz Melo, Franck Sgard, Olivier Doutres, Éric Wagnac
Pressure Induced By Roll-Down Foam-Earplugs On Earcanal, Ahmed Dalaq, Luiz Melo, Franck Sgard, Olivier Doutres, Éric Wagnac
Études primaires
Physical discomfort of earplug is a common complaint, so much so it jeopardizes the health and safety of workers, resulting in frequent noncompliance to proper wearing of earplugs. A pressing need thus arises to understand the underlying mechanics for the interaction of earplugs with earcanal walls. An idealized cylindrical geometry of earcanal is first treated with computational and analytical modeling to predict the pressure induced by roll-down cylindrical PVC foam earplugs. In order to predict representative pressure values and distribution, we estimated the hyperelastic properties of foam-based earplugs and characterized their behavior using a stent testing machine (J-CrimpTM machine) …
Augmentative And Alternative Communication Use, Service Delivery Experiences, And Communicative Participation For People With Amyotrophic Lateral Sclerosis, Betts Peters
Dissertations and Theses
People with amyotrophic lateral sclerosis (ALS) often experience changes to their speech, and may use augmentative and alternative communication (AAC) devices and techniques to maintain the ability to communicate. The use of AAC may facilitate the participation of people with ALS in various life situations involving communication. There is limited data in the literature about the AAC approaches currently used by people with ALS, the professional services they receive to support communication, or the effects of AAC on their communicative participation. This dissertation involved a nationwide online survey of people with ALS, and comprises three papers intended to add to …
Kemar Hats Head Orientation Directivity, Samuel D. Bellows, Timothy W. Leishman
Kemar Hats Head Orientation Directivity, Samuel D. Bellows, Timothy W. Leishman
Directivity
This directivity data set for a KEMAR head head-and-torso simulator (HATS) includes head orientations in 14 directions in 5° steps starting from 0° to 40° and then in 10° steps from 40° to 90°. The full spherical measurements followed at an a = 0.97 m radius with the mouth aperture at the spherical center. The sampling density and distribution followed the AES 5° dual-equiangular sampling standard, omitting the south pole (θ = 180°). Thus, each spherical directivity assessment included 36 polar-angle θ samples and 72 azimuthal-angle ϕ samples. The presented data include 22 1/3-octave bands, ranging from 80 Hz …
Speaker Encoding For Zero-Shot Speech Synthesis, Tristin W. Cory
Speaker Encoding For Zero-Shot Speech Synthesis, Tristin W. Cory
Graduate Theses/Dissertations
Spoken communication, for many, is an essential part of everyday life. Some individuals can lose or not be born with the ability to speak. To function on a day-to-day basis, these individuals find other ways of communication. Adaptive speech synthesis is one of those ways. It recreates a user’s previous voice or creates a voice that blends with their regional dialect. Current adaptive speech synthesis techniques that achieve human-like speech require thirty minutes, to a few hours of high-quality audio recordings of a target speaker. This amount of recorded audio is not commonly possessed by people in need of a …
Average Speech Directivity, Samuel D. Bellows, Claire M. Pincock, Jennifer K. Whiting, Timothy W. Leishman
Average Speech Directivity, Samuel D. Bellows, Claire M. Pincock, Jennifer K. Whiting, Timothy W. Leishman
Directivity
Speech directivity describes the angular dependence of acoustic radiation from a talker’s mouth and nostrils and diffraction about his or her body and chair (if seated). It is an essential physical aspect of communication affecting sounds and signals in acoustical environments, audio, and telecommunication systems. Because high-resolution, spherically comprehensive measurements of live, phonetically balanced speech have been unavailable in the past, the authors have undertaken research to produce and share such data for simulations of acoustical environments, optimizations of microphone placements, speech studies, and other applications. The measurements included three male and three female talkers who repeated phonetically balanced passages …
Communication Of Deaf People Based On Myoware Muscle Sensor, Mohamad Ataya
Communication Of Deaf People Based On Myoware Muscle Sensor, Mohamad Ataya
Research Opportunities for Engineering Undergraduates (ROEU) Program 2018-19
The current project aims to create a product that enhances communication of deaf people. The developed product has the following features:
• a muscle sensor that transcribes sign language into actual letters
• a voice detector that translates speech into actual words
• the ability to receive and send emergency alerts.
Development And Validation For A Mobile Speech-In-Noise Audiometric Task, Tommy Peng
Development And Validation For A Mobile Speech-In-Noise Audiometric Task, Tommy Peng
McKelvey School of Engineering Graduate Student Theses & Dissertations
Traditional speech-in-noise hearing tests are performed by clinicians with specialized equipment. Furthermore, these tasks often present contextually weak sentences in background babble, which are poor representations of real-world situations. This study proposes a mobile audiometric task, Semantic Auditory Search, which uses the Android platform to bypass the need for specialized equipment and presents multiple tasks of two competing real-world conversations to estimate the user’s speech-in-noise hearing ability. Through linear regression models built from data of seventy-nine subjects, three Semantic Auditory Search metrics have been shown to have statistically significant (p < 0.05) with medium effects sizes for predicting QuickSIN SNR50. The internal consistency of the task was also high, with a Cronbach’s alpha of 0.88 or more across multiple metrics. In conclusion, this preliminary study suggests that Semantic Auditory Search can accurately and reliably perform as an automated speech-in-noise hearing test. It also has tremendous potential for extension into automated tests of cognitive function, as well.
Development Of Kinematic Templates For Automatic Pronunciation Assessment Using Acoustic-To-Articulatory Inversion, Deriq K. Jones
Development Of Kinematic Templates For Automatic Pronunciation Assessment Using Acoustic-To-Articulatory Inversion, Deriq K. Jones
Master's Theses (2009 -)
Computer-aided pronunciation training (CAPT) is a subcategory of computer-aided language learning (CALL) that deals with the correction of mispronunciation during language learning. For a CAPT system to be effective, it must provide useful and informative feedback that is comprehensive, qualitative, quantitative, and corrective. While the majority of modern systems address the first 3 aspects of feedback, most of these systems do not provide corrective feedback. As part of the National Science Foundation (NSF) funded study “RI: Small: Speaker Independent Acoustic-Articulator Inversion for Pronunciation Assessment”, the Marquette Speech and Swallowing Lab and Marquette Speech and Signal Processing Lab are conducting a …
The Effect Of Frequency Resolution On Intelligibility Sentence And Its Relevance To Cochlear Implant Design, Seth H. Roy
The Effect Of Frequency Resolution On Intelligibility Sentence And Its Relevance To Cochlear Implant Design, Seth H. Roy
School of Computing: Dissertations, Theses, and Student Research
The purpose of this study is to understand how electrical stimulation (as opposed to acoustical stimulation) of the auditory nerve is used in cochlear implants. Speech is a complex signal that changes rapidly in time and frequency domains. Since phonemes (the smallest unit of speech that distinguishes words) depend on nuanced differences in frequency patterns, it would be expected that a signal with drastically reduced frequency information would be of limited value for conveying speech. Such a frequency-poor signal is the object to be investigated in the present work. It is also the basis of the way speech is represented …
Effects Of Vocal Fold Nodules On Glottal Cycle Measurements Derived From High-Speed Videoendoscopy In Children, Rita R. Patel, Harikrishnan Unnikrishnan, Kevin D. Donohue
Effects Of Vocal Fold Nodules On Glottal Cycle Measurements Derived From High-Speed Videoendoscopy In Children, Rita R. Patel, Harikrishnan Unnikrishnan, Kevin D. Donohue
Electrical and Computer Engineering Faculty Publications
The goal of this study is to quantify the effects of vocal fold nodules on vibratory motion in children using high-speed videoendoscopy. Differences in vibratory motion were evaluated in 20 children with vocal fold nodules (5–11 years) and 20 age and gender matched typically developing children (5–11 years) during sustained phonation at typical pitch and loudness. Normalized kinematic features of vocal fold displacements from the mid-membranous vocal fold point were extracted from the steady-state high-speed video. A total of 12 kinematic features representing spatial and temporal characteristics of vibratory motion were calculated. Average values and standard deviations (cycle-to-cycle variability) of …
Spatio-Temporal Progression Of Cortical Activity Related To Continuous Overt And Covert Speech Production In A Reading Task, Jonathan S. Brumberg, Dean J. Krusienski, Shreya Chakrabarti, Aysegul Gunduz, Peter Brunner, Anthony L. Ritaccio, Gerwin Schalk
Spatio-Temporal Progression Of Cortical Activity Related To Continuous Overt And Covert Speech Production In A Reading Task, Jonathan S. Brumberg, Dean J. Krusienski, Shreya Chakrabarti, Aysegul Gunduz, Peter Brunner, Anthony L. Ritaccio, Gerwin Schalk
Electrical & Computer Engineering Faculty Publications
How the human brain plans, executes, and monitors continuous and fluent speech has remained largely elusive. For example, previous research has defined the cortical locations most important for different aspects of speech function, but has not yet yielded a definition of the temporal progression of involvement of those locations as speech progresses either overtly or covertly. In this paper, we uncovered the spatio-temporal evolution of neuronal population-level activity related to continuous overt speech, and identified those locations that shared activity characteristics across overt and covert speech. Specifically, we asked subjects to repeat continuous sentences aloud or silently while we recorded …
An Axisymmetric Finite Element Model To Study The Earplug Contribution To The Bone Conduction Occlusion Effect, Martin K. Brummund, Franck Sgard, Yvan Petit, Frédéric Laville, Hugues Nélisse
An Axisymmetric Finite Element Model To Study The Earplug Contribution To The Bone Conduction Occlusion Effect, Martin K. Brummund, Franck Sgard, Yvan Petit, Frédéric Laville, Hugues Nélisse
Études primaires
An axisymmetric linear elasto-acoustic finite element (FE) model of an occluded human external ear is proposed to simulate the bone conduction occlusion effect (OE). The model consists of a cylindrical ear canal cavity surrounded by layers of biological tissues (skin, cartilage and bone) in which an earplug is inserted. Geometrical and material properties are taken from the literature. OEs are predicted for foam and silicone earplug FE-models using COMSOL Multiphysics (COMSOL®, Sweden). The FE-model is shown to predict the experimental OE measured in two healthy human reference groups wearing foam or silicone earplugs, satisfactorily. Deviations between model and experiment are …
The Electromagnetic Articulography Mandarin Accented English (Ema-Mae) Corpus Of Acoustic And 3d Articulatory Kinematic Data, Jeffrey J. Berry, An Ji, Michael T. Johnson
The Electromagnetic Articulography Mandarin Accented English (Ema-Mae) Corpus Of Acoustic And 3d Articulatory Kinematic Data, Jeffrey J. Berry, An Ji, Michael T. Johnson
Speech Pathology and Audiology Faculty Research and Publications
There is a significant need for more comprehensive electromagnetic articulography (EMA) datasets that can provide matched acoustics and articulatory kinematic data with good spatial and temporal resolution. The Marquette University Electromagnetic Articulography Mandarin Accented English (EMA-MAE) corpus provides kinematic and acoustic data from 40 gender and dialect balanced speakers representing 20 Midwestern standard American English L1 speakers and 20 Mandarin Accented English (MAE) L2 speakers, half Beijing region dialect and half are Shanghai region dialect. Three dimensional EMA data were collected at a 400 Hz sampling rate using the NDI Wave system, with articulatory sensors on the midsagittal lips, lower …
Sensorimotor Adaptation Of Speech Using Real-Time Articulatory Resynthesis, Jeffrey J. Berry, Cassandra North, Michael T. Johnson
Sensorimotor Adaptation Of Speech Using Real-Time Articulatory Resynthesis, Jeffrey J. Berry, Cassandra North, Michael T. Johnson
Speech Pathology and Audiology Faculty Research and Publications
Sensorimotor adaptation is an important focus in the study of motor learning for non-disordered speech, but has yet to be studied substantially for speech rehabilitation. Speech adaptation is typically elicited experimentally using LPC resynthesis to modify the sounds that a speaker hears himself producing. This method requires that the participant be able to produce a robust speech-acoustic signal and is therefore not well-suited for talkers with dysarthria. We have developed a novel technique using electromagnetic articulography (EMA) to drive an articulatory synthesizer. The acoustic output of the articulatory synthesizer can be perturbed experimentally to study auditory feedback effects on sensorimotor …
Developing A Drug Delivery System For Treatment Of Vocal Fold Scarring, Aaron Michael Kosinski
Developing A Drug Delivery System For Treatment Of Vocal Fold Scarring, Aaron Michael Kosinski
Open Access Dissertations
Vocal fold scarring is an affliction that results in the formation of a disorganized and stiff extracellular matrix (ECM) with abnormal ECM component densities & structures including a significant increase in collagen deposition. It is caused by improper healing post injury and results in profound changes in the biomechanical properties of the vocal folds impairing their ability to generate a normal mucosal wave during phonation.
Finding an effective treatment for vocal fold scarring has been elusive. Currently, treatments seek temporary solutions that correct glottal incompetence and reduce stiffness caused by the scar through the augmentation of the vocal folds using …
Individual Articulator's Contribution To Phoneme Production, Jun Wang, Jordan R. Green, Ashok Samal
Individual Articulator's Contribution To Phoneme Production, Jun Wang, Jordan R. Green, Ashok Samal
School of Computing: Conference and Workshop Papers
Speech sounds are the result of coordinated movements of individual articulators. Understanding each articulator’s role in speech is fundamental not only for understanding how speech is produced, but also for optimizing speech assessments and treatments. In this paper, we studied the individual contributions of six articulators, tongue tip, tongue blade, tongue body front, tongue body back, upper lip, and lower lip to phoneme classification. A total of 3,838 vowel and consonant production samples were collected from eleven native English speakers. The results of speech movement classification using a support vector machine indicated that the tongue encoded significantly more information than …
Modeling Hrtf For Sound Localization In Normal Listeners And Bilateral Cochlear Implant Users, Douglas A. Miller
Modeling Hrtf For Sound Localization In Normal Listeners And Bilateral Cochlear Implant Users, Douglas A. Miller
Electronic Theses and Dissertations
Mathematical models can be very useful for understanding complicated systems and for testing algorithms through simulation that would be difficult or expensive to implement. This dissertation presents a model that attempts to simulate the sound localization performance of persons using bilateral cochlear implants. The expectation is that this model could prove to be a useful tool in developing new signal processing algorithms for neural encoding strategies.
The head related transfer function (HRTF) is a critical component of this model, and in the ideal case, provides the base characteristics of head shadow, torso and pinna effects. This defines the temporal, intensity …
Age-Related Changes To The Production Of Linguistic Prosody, Daniel Richard Barnes
Age-Related Changes To The Production Of Linguistic Prosody, Daniel Richard Barnes
Open Access Theses
The production of speech prosody (the rhythm, pausing, and intonation associated with natural speech) is critical to effective communication. The current study investigated the impact of age-related changes to physiology and cognition in relation to the production of two types of linguistic prosody: lexical stress and the disambiguation of syntactically ambiguous utterances. Analyses of the acoustic correlates of stress: speech intensity (or sound-pressure level; SPL), fundamental frequency (F0), key word/phrase duration, and pause duration revealed that both young and older adults effectively use these acoustic features to signal linguistic prosody, although the relative weighting of cues differed by group. Differences …
Whole-Word Recognition From Articulatory Movements For Silent Speech Interfaces, Jun Wang, Ashok Samal, Jordan R. Green, Frank Rudzicz
Whole-Word Recognition From Articulatory Movements For Silent Speech Interfaces, Jun Wang, Ashok Samal, Jordan R. Green, Frank Rudzicz
Department of Special Education and Communication Disorders: Faculty Publications
Articulation-based silent speech interfaces convert silently produced speech movements into audible words. These systems are still in their experimental stages, but have significant potential for facilitating oral communication in persons with laryngectomy or speech impairments. In this paper, we report the result of a novel, real-time algorithm that recognizes whole-words based on articulatory movements. This approach differs from prior work that has focused primarily on phoneme-level recognition based on articulatory features. On average, our algorithm missed 1.93 words in a sequence of twenty-five words with an average latency of 0.79 seconds for each word prediction using a data set of …