Open Access. Powered by Scholars. Published by Universities.®

Signal Processing Commons

Open Access. Powered by Scholars. Published by Universities.®

Technological University Dublin

Discipline
Keyword
Publication Year
Publication
File Type

Articles 31 - 60 of 106

Full-Text Articles in Signal Processing

Bitrate Classification Of Twice-Encoded Audio Using Objective Quality Features, Colm Sloan, Damien Kelly, Naomi Harte, Anil C. Kokaram, Andrew Hines Jun 2016

Bitrate Classification Of Twice-Encoded Audio Using Objective Quality Features, Colm Sloan, Damien Kelly, Naomi Harte, Anil C. Kokaram, Andrew Hines

Conference papers

When a user uploads audio files to a music stream- ing service, these files are subsequently re-encoded to lower bitrates to target different devices, e.g. low bitrate for mobile. To save time and bandwidth uploading files, some users encode their original files using a lossy codec. The metadata for these files cannot always be trusted as users might have encoded their files more than once. Determining the lowest bitrate of the files allows the streaming service to skip the process of encoding the files to bitrates higher than that of the uploaded files, saving on processing and storage space. This …


Source Separation Approach To Video Quality Prediction In Computer Networks, Ruairí De Fréin May 2016

Source Separation Approach To Video Quality Prediction In Computer Networks, Ruairí De Fréin

Articles

Time-varying loads introduce errors in the estimated model parameters of service-level predictors in Computer Networks. A load-adjusted modification of a traditional unadjusted service-level predictor is contributed, based on Source Separation (SS). It mitigates these errors and improves service-quality predictions for Video-on-Demand (VoD) by :6 to 2dB.


Monitoring Voip Speech Quality For Chopped And Clipped Speech, Andrew Hines, Jan Skoglund, Anil C. Kokaram, Naomi Harte Jan 2016

Monitoring Voip Speech Quality For Chopped And Clipped Speech, Andrew Hines, Jan Skoglund, Anil C. Kokaram, Naomi Harte

Articles

No abstract provided.


Hvdc Systems Fault Analysis Using Various Signal Processing Techniques, Benish Paily Dec 2015

Hvdc Systems Fault Analysis Using Various Signal Processing Techniques, Benish Paily

Doctoral

The detection and fast clearance of faults are important for the safe and optimal operation of HVDC systems. In HVDC systems, various types of AC faults (rectifier & inverter side) and DC faults can occur. It is therefore necessary to detect the faults and classify them for better protection and diagnostics purposes. Various techniques for fault detection and classification in HVDC systems using signal processing techniques are presented and investigated in this research work. In this research work, it is shown that the wavelet transformation can effectively detect abrupt changes in system signals which are indicative of a fault. This …


Digital Signal Processing Foundations, David Dorran Jan 2015

Digital Signal Processing Foundations, David Dorran

Other resources

Signals are all around us and come in a wide variety of shapes and forms. When we speak we create pressure variations in the air which generate audio signals; earthquakes produce large seismic signals; healthcare professionals monitor ECG signals which capture the electrical activity of the heart; radio, internet and telephone signals are being transmitted across the world; the list of signals is endless! (see 2 minute video at pzdsp.com/vid1 for some examples) Digital signal processing (DSP) is primarily about making use of computers to help us analyse and manipulate signals in order to help us with our everyday lives. …


Measuring And Monitoring Speech Quality For Voice Over Ip With Polqa, Visqol And P.563, Andrew Hines, Eoin Gillen, Naomi Harte Jan 2015

Measuring And Monitoring Speech Quality For Voice Over Ip With Polqa, Visqol And P.563, Andrew Hines, Eoin Gillen, Naomi Harte

Conference papers

There are many types of degradation which can occur in Voice over IP (VoIP) calls. Of interest in this work are degradations which occur independently of the codec, hardware or network in use. Specifically, their effect on the subjective and objec- tive quality of the speech is examined. Since no dataset suit- able for this purpose exists, a new dataset (TCD-VoIP) has been created and has been made publicly available. The dataset con- tains speech clips suffering from a range of common call qual- ity degradations, as well as a set of subjective opinion scores on the clips from 24 …


Quantized Nonnegative Matrix Factorization, Ruairí De Fréin Jan 2014

Quantized Nonnegative Matrix Factorization, Ruairí De Fréin

Conference papers

Even though Nonnegative Matrix Factorization (NMF) in its original form performs rank reduction and signal compaction implicitly, it does not explicitly consider storage or transmission constraints. We propose a Frobenius-norm Quantized Nonnegative Matrix Factorization algorithm that is 1) almost as precise as traditional NMF for decomposition ranks of interest (with in 1-4dB), 2) admits to practical encoding techniques by learning a factorization which is simpler than NMF’s (by a factor of 20-70) and 3) exhibits a complexity which is comparable with state-of-the-art NMF methods. These properties are achieved by considering the quantization residual via an outer quantization optimization step, in …


Sweet (Small Wind Energy Estimation Tool), Thomas Woolmington, Keith Sunderland Oct 2013

Sweet (Small Wind Energy Estimation Tool), Thomas Woolmington, Keith Sunderland

Other resources

Description:

Generic Wind Energy Estimation Model incorporating Turbulence Intensity.

Requirements:

Manufacturer specific wind turbine characteristic

Mean wind speed at a specific height

Site conditions in terms of TI and Surface Roughness

Disclaimer:

This model has been designed as an estimation tool for estimating the wind power yield of a custom turbine for an idealised set of site conditions for a given year.

While the assumptions that are made in this estimation tool are mathematically plausible they should not in any way be interpreted as an alternative to a real world site assessment.

A list of assumptions to which the model …


Adaptive Ofdm For Wireless Interconnect In Confined Enclosures, Vit Sipal, Javier Gelabert, Christopher J. Stevens, Ben Allen, David Edwards Jul 2013

Adaptive Ofdm For Wireless Interconnect In Confined Enclosures, Vit Sipal, Javier Gelabert, Christopher J. Stevens, Ben Allen, David Edwards

Articles

This letter considers and recommends OFDM with adaptive subcarrier modulation as a suitable candidate for wireless UWB communication in computer chassis. A rigorous measurement campaign studies the guaranteed spectral efficiency. It concludes that enhancement of the existing WiMedia OFDM systems with a bandwidth of 528 MHz in order to support adaptive OFDM would enable data-rates above 1 Gbps over short ranges, i.e. the spectral efficiency would be doubled. Moreover, the guaranteed spectral efficiency is shown to increase with bandwidth, i.e. the guaranteed data-rate increases better than linearly with bandwidth.


The Presence Of Regional Accents In Electrolarynx Speech And The Resultant Effect On Overall Intelligibility., Brian Madden, Eugene Coyle Apr 2012

The Presence Of Regional Accents In Electrolarynx Speech And The Resultant Effect On Overall Intelligibility., Brian Madden, Eugene Coyle

Conference Papers

During voiced speech, the larynx provides quasi-periodic acoustic excitation of the vocal tract. In most electrolarynxes, mechanical vibrations are produced by a linear electromechanical actuator, the armature of which percusses against a metal or plastic plate at a frequency within the range of glottal excitation. In this paper, a phonological analysis of a section of results from an online perceptual intelligibility test was performed which compared speech produced using a novel hands-free electrolarynx and a commercially available electrolarynx. A portion of the test consisted of a closed-set format containing a selection of four sets of four random CVC audio samples …


On Inpainting The Adress Algorithm, Derry Fitzgerald, Dan Barry Jan 2012

On Inpainting The Adress Algorithm, Derry Fitzgerald, Dan Barry

Conference papers

The Adress algorithm has been demonstrated to be capable of separating sound sources from instantaneous linear mixtures, provided that the sources have a unique pan position in the stereo field. However, a shortcoming of the Adress algorithm is that all time-frequency bins outside of the chosen azimuth range are set to zero, resulting in audible artifacts in the resynthesised sound. Here we show that an inpainting algorithm based on NMF is capable of estimating these missing values and improves on the results obtained using Adress only.


On The Use Of Masking Filters In Sound Source Separation, Derry Fitzgerald, Rajesh Jaiswal Jan 2012

On The Use Of Masking Filters In Sound Source Separation, Derry Fitzgerald, Rajesh Jaiswal

Conference papers

Many sound source separation algorithms, such as NMF and related approaches, disregard phase information and operate only on magnitude or power spectrograms. In this context, generalised Wiener filters have been widely used to generate masks which are applied to the original complex-valued spectrogram before inversion to the time domain, as these masks have been shown to give good results. However, these masks may not be optimal from a perceptual point of view. To this end, we propose new families of masks and compare their performance to generalised Wiener filter masks using three different factorisation-based separation algorithms. Further, to-date no analysis …


User Assisted Separation Using Tensor Factorisations, Derry Fitzgerald Jan 2012

User Assisted Separation Using Tensor Factorisations, Derry Fitzgerald

Conference papers

Recent research has demonstrated that user assisted techniques, where the user provides a ”guide” version of the source to be separated, are capable of giving good sound source separation. Here the user sings or plays along with the target source, and the user input is used to guide the separation towards the source of interest. This is typically done in a factorisation framework, such as non-negative matrix factorisation. Here we extend such approaches to a tensor factorisation framework to deal with multichannel signals. Further, we demonstrate how this framework can be used to improve the output from other user assisted …


Distributed Formal Concept Analysis Algorithms Based On An Iterative Mapreduce Framework, Ruairí De Fréin, Biao Xu, Eric Robson, Mícheál Ó Fóghlú Jan 2012

Distributed Formal Concept Analysis Algorithms Based On An Iterative Mapreduce Framework, Ruairí De Fréin, Biao Xu, Eric Robson, Mícheál Ó Fóghlú

Conference papers

While many existing formal concept analysis algorithms are efficient, they are typically unsuitable for distributed implementation. Taking the MapReduce (MR) framework as our inspiration we introduce a distributed approach for performing formal concept mining. Our method has its novelty in that we use a light-weight MapReduce runtime called Twister which is better suited to iterative algorithms than recent distributed approaches. First, we describe the theoretical foundations underpinning our distributed formal concept analysis approach. Second, we provide a representative exemplar of how a classic centralized algorithm can be implemented in a distributed fashion using our methodology: we modify Ganter’s classic algorithm …


Vocal Separation Using Nearest Neighbours And Median Filtering, Derry Fitzgerald Jan 2012

Vocal Separation Using Nearest Neighbours And Median Filtering, Derry Fitzgerald

Conference papers

Recently, single channel vocal separation algorithms have been proposed which exploit the fact that most popular music can be regarded as a repeating musical background over which a locally non-repeating vocal signal is superimposed. In this paper we describe a novel vocal separator inspired by these approaches which finds the k nearest neighbours to each frame of a spectrogram of the mixture signal. The median value of these frames is then used as the estimate of the background music at the current frame. This is then used to generate a mask on the original complex-valued spectrogram before inversion to the …


Upmixing From Mono : A Source Separation Approach, Derry Fitzgerald Jul 2011

Upmixing From Mono : A Source Separation Approach, Derry Fitzgerald

Conference papers

We present a system for upmixing mono recordings to stereo through the use of sound source separation techniques. The use of sound source separation has the advantage of allowing sources to be placed at distinct points in the stereo field, resulting in more natural sounding upmixes. The system separates an input signal into a number of sources, which can then be imported into a digital audio workstation for upmixing to stereo. Considerations to be taken into account when upmixing are discussed, and a brief overview of the various sound source separation techniques used in the system are given. The effectiveness …


Shifted Nmf Using An Efficient Constant-Q Transform For Monaural Sound Source Separation, Rajesh Jaiswal, Derry Fitzgerald, Eugene Coyle, Scott Rickard Jun 2011

Shifted Nmf Using An Efficient Constant-Q Transform For Monaural Sound Source Separation, Rajesh Jaiswal, Derry Fitzgerald, Eugene Coyle, Scott Rickard

Conference papers

Non-negative Matrix Factorisation (NMF) based algorithms have found application in monaural audio source separation due to their ability to factorize audio spectrogram into additive part-based basis functions, which typically corresponds to individual notes or chords in music. These separated basis functions are usually greater in number than the active sources, hence clustering is needed for individual source signal synthesis. Although, many attempts have been made to improve the clustering of the basis functions to sources, much research is still required in this area. Recently, Shifted NMF based methods have been proposed as a means to avoid clustering these pitched basis …


User Assisted Source Separation Using Non-Negative Matrix Factorisation, Derry Fitzgerald Jun 2011

User Assisted Source Separation Using Non-Negative Matrix Factorisation, Derry Fitzgerald

Conference papers

Much research has been carried out on the use of non-negative matrix factorisation for the purpose of musical sound source separation. However, a notable shortcoming of non-negative matrix factorisation is that the recovered basis functions have to be clustered to sound sources for separation to take place. This has proved to be a difficult problem to solve. As a means of overcoming this problem, we introduce an extension to non-negative matrix factorisation which allows a user to guide the sepa- ration by singing, or playing along with, the source they want to separate. This is done through the use of …


Intelligibility Of Electrolarynx Speech Using A Novel Hands-Free Actuator, Brian Madden, Mark Nolan, Ted Burke, James Condron, Eugene Coyle Jan 2011

Intelligibility Of Electrolarynx Speech Using A Novel Hands-Free Actuator, Brian Madden, Mark Nolan, Ted Burke, James Condron, Eugene Coyle

Conference Papers

During voiced speech, the larynx provides quasi-periodic acoustic excitation of the vocal tract. In most electrolarynxes, mechanical vibrations are produced by a linear electromechanical actuator, the armature of which percusses against a metal or plastic plate at a frequency within the range of glottal excitation. In this paper, the intelligibility of speech produced using a novel hands-free actuator is compared to speech produced using a conventional electrolarynx. Two able-bodied speakers (one male, one female) performed a closed response test containing 28 monosyllabic words, once using a conventional electrolarynx and a second time using the novel design. The resulting audio recordings …


Augmented Control Of A Hands-Free Electrolarynx, Brian Madden, James Condron, Eugene Coyle Jan 2011

Augmented Control Of A Hands-Free Electrolarynx, Brian Madden, James Condron, Eugene Coyle

Conference Papers

During voiced speech, the larynx acts as the sound source, providing a quasi-periodic excitation of the vocal tract. Following a total laryngectomy, some people speak using an electrolarynx which employs an electromechanical actuator to perform the excitatory function of the absent larynx. Drawbacks of conventional electrolarynx designs include the monotonic sound emitted, the need for a free-hand to operate the device, and the difficulty experienced by many laryngectomees in adapting to its use. One improvement to the electrolarynx, which clinicians and users frequently suggest, is the provision of a convenient hands-free control facility. This would allow more natural use of …


Tempo Detection Using A Hybrid Multi-Band Approach, Mikel Gainza, Eugene Coyle Jan 2011

Tempo Detection Using A Hybrid Multi-Band Approach, Mikel Gainza, Eugene Coyle

Articles

In this paper, a novel tempo detection system is presented, which suggests the use of a hybrid multiband decomposition. The model tracks the periodicities of different signal property changes that manifest within different frequency bands by using the most appropriate onset/transient detectors for each frequency band. In addition, the proposed system applies a novel method to weight tempo candidates. Each contribution is evaluated by comparing the presented system against existing approaches using three different databases that comprises 1638 songs. These databases include the two publicly available database of songs used in the tempo evaluation contest of ISMIR 2004. These songs …


The Synchronized Short-Time-Fourier-Transform: Properties And Definitions For Multichannel Source Separation., Ruairí De Fréin, Scott Rickard Jan 2011

The Synchronized Short-Time-Fourier-Transform: Properties And Definitions For Multichannel Source Separation., Ruairí De Fréin, Scott Rickard

Articles

This paper proposes the use of a synchronized linear transform, the synchronized short-time-Fourier-transform (sSTFT), for time-frequency analysis of anechoic mixtures. We address the short comings of the commonly used time-frequency linear transform in multichannel settings, namely the classical short-time-Fourier-transform (cSTFT). We propose a series of desirable properties for the linear transform used in a multichannel source separation scenario: stationary invertibility, relative delay, relative attenuation, and finally delay invariant relative windowed-disjoint orthogonality (DIRWDO). Multisensor source separation techniques which operate in the time-frequency domain, have an inherent error unless consideration is given to the multichannel properties proposed in this paper. The sSTFT …


On Improving Electrooculogram-Based Computer Mouse Systems: The Accelerometer Trigger, Johnalan Keegan, Edward Burke, James Condron, Eugene Coyle Jan 2011

On Improving Electrooculogram-Based Computer Mouse Systems: The Accelerometer Trigger, Johnalan Keegan, Edward Burke, James Condron, Eugene Coyle

Conference Papers

Eye tracking is a well-established method of computer control for profoundly paralysed people (Anson et al., 2002). Cameras are commonly used to track eye movements (Morimoto et al., 2005) but one alternative is the bioelectrical signal known as the electrooculogram (EOG). There are some EOG mouse control systems that facilitate the use of GUI applications, but certain actions, which are straightforward using a conventional mouse, remain impossible. Unless the eyes are tracking a target, they move in saccades (jumps), making it impossible to voluntarily trace out smooth trajectories with one's gaze, as would be required to draw a smooth curve. …


Application Of Stochastic Diffusion For Hiding High Fidelity Encrypted Images, Jonathan Blackledge, Abdulrahman Al-Rawi Jan 2011

Application Of Stochastic Diffusion For Hiding High Fidelity Encrypted Images, Jonathan Blackledge, Abdulrahman Al-Rawi

Articles

Cryptography coupled with information hiding has received increased attention in recent years and has become a major research theme because of the importance of protecting encrypted information in any Electronic Data Interchange system in a way that is both discrete and covert. One of the essential limitations in any cryptography system is that the encrypted data provides an indication on its importance which arouses suspicion and makes it vulnerable to attack. Information hiding of Steganography provides a potential solution to this issue by making the data imperceptible, the security of the hidden information being a threat only if its existence …


Wavelet Based Islanding Detection Of Dc-Ac Inverter Interfaced Dg Systems, Mohamed Moin Hanif, Malabika Basu, Kevin Gaughan Sep 2010

Wavelet Based Islanding Detection Of Dc-Ac Inverter Interfaced Dg Systems, Mohamed Moin Hanif, Malabika Basu, Kevin Gaughan

Conference papers

The increased penetration of distributed generation (DG) often connected to grid through dc-ac interface has made islanding detection an important and challenging issue to power engineers. Several methods based on passive and active detection scheme have been proposed in the literature. While passive schemes have a large non detection zone (NDZ), concern has been raised on active method due to its degrading power quality effect. This paper proposes a wavelet based passive islanding detection scheme with almost zero NDZ for dc-ac inverter interfaced grid connected DGs. The key idea is to utilize the spectral changes in the higher frequency components …


Intelligibility Of Electrolarynx Speech Using A Novel Actuator, Brian Madden, Mark Nolan, Ted Burke, James Condron, Eugene Coyle Jun 2010

Intelligibility Of Electrolarynx Speech Using A Novel Actuator, Brian Madden, Mark Nolan, Ted Burke, James Condron, Eugene Coyle

Conference Papers

During voiced speech, the larynx provides quasi-periodic acoustic excitation of the vocal tract. Following a laryngectomy, some people speak using an electrolarynx which replaces the excitatory function of the absent larynx. Drawbacks of conventional electrolarynx designs include the buzzing monotonic sound emitted, the need for a free hand to operate the device, and difficulty experienced by many laryngectomees in adapting to its use. Despite these shortcomings, it remains the preferred method of speech rehabilitation for a substantial minority of laryngectomees. In most electrolarynxes, mechanical vibrations are produced by a linear electromechanical actuator, the armature of which percusses against a metal …


Information Hiding Using Stochastic Diffusion For The Covert Transmission Of Encrypted Images, Jonathan Blackledge Jun 2010

Information Hiding Using Stochastic Diffusion For The Covert Transmission Of Encrypted Images, Jonathan Blackledge

Conference papers

A principal weakness of all encryption systems is that the output data can be `seen' to be encrypted. In other words, encrypted data provides a 'flag' on the potential value of the information that has been encrypted. In this paper, we provide a novel approach to `hiding' encrypted data in a digital image. We consider an approach in which a plaintext image is encrypted with a cipher using the processes of `stochastic diffusion' and the output quantized into a 1-bit array generating a binary image cipher-text. This output is then `embedded' in a host image which is undertaken either in …


Measuring Variations Of Mimicry By Means Of Prosodic Cues In Task-Based Scenarios And Conversational Speech, Brian Vaughan, Celine De Looze Mar 2010

Measuring Variations Of Mimicry By Means Of Prosodic Cues In Task-Based Scenarios And Conversational Speech, Brian Vaughan, Celine De Looze

Other resources

Here, we address the measurement of mimicry, that is when speakers’ speech variations look like parallel patterns.

As a definition of mimicry, we often read in the literature description such as mimicry is “The situation where the observed behaviours of two inter-actants although dissimilar at the start of the interaction are moving towards behavioral matching”. These types of descriptions imply that mimicry is a linear phenomenon and that speakers tend to imitate over time. However, it can be assumed, especially when studying spontaneous speech, that there are rather phases of mimicry and non-mimicry and that mimicry should be rather …


On The Use Of A Dynamic Hybrid Tempo Detection Model For Beat Tracking, Mikel Gainza Jan 2010

On The Use Of A Dynamic Hybrid Tempo Detection Model For Beat Tracking, Mikel Gainza

Conference papers

In this paper, an approach that estimates the times at which musical beats occur is presented. The system uses a hybrid multi-band decomposition in order to estimate the music tempo. Following this, beat events are tracked by using a dynamic programming approach, which is updated by using short time tempo estimates. The hybrid decomposition is used in order to calculate the tempo by using different onset detection functions in different frequency bands. In addition, a method that estimates which frequency bands provide reliable periodicities is also presented. The accuracy of the model is evaluated by comparing the presented system against …


Exploiting Glottal Formant Parameters For Glottal Inverse Filtering And Parameterization, Alan O'Cinneide, David Dorran, Mikel Gainza, Eugene Coyle Jan 2010

Exploiting Glottal Formant Parameters For Glottal Inverse Filtering And Parameterization, Alan O'Cinneide, David Dorran, Mikel Gainza, Eugene Coyle

Conference papers

It is crucial for many methods of inverse filtering that the time domain information of the glottal source waveform is known, e.g. the location of the instant of glottal closure. It is often the case that this information is unknown and/or cannot be determined due to e.g. recording conditions which can corrupt the phase spectrum. In these scenarios, alternative strategies are required. This paper describes a method which, given the parameters of the glottal formant of the signal frame, can accurately parameterize the glottal shape source and vocal filter for a broad range of voice quality types and which is …