The Experts below are selected from a list of 309 Experts worldwide ranked by ideXlab platform
Richard Heusdens - One of the best experts on this subject based on the ideXlab platform.
-
speech energy redistribution for intelligibility improvement in noise based on a Perceptual Distortion measure
Computer Speech & Language, 2014Co-Authors: Cees H Taal, Richard C Hendriks, Richard HeusdensAbstract:Abstract A speech pre-processing algorithm is presented that improves the speech intelligibility in noise for the near-end listener. The algorithm improves intelligibility by optimally redistributing the speech energy over time and frequency according to a Perceptual Distortion measure, which is based on a spectro-temporal auditory model. Since this auditory model takes into account short-time information, transients will receive more amplification than stationary vowels, which has been shown to be beneficial for intelligibility of speech in noise. The proposed method is compared to unprocessed speech and two reference methods using an intelligibility listening test. Results show that the proposed method leads to significant intelligibility gains while still preserving quality. Although one of the methods used as a reference obtained higher intelligibility gains, this happened at the cost of decreased quality. Matlab code is provided.
-
a speech preprocessing strategy for intelligibility improvement in noise based on a Perceptual Distortion measure
International Conference on Acoustics Speech and Signal Processing, 2012Co-Authors: Cees H Taal, Richard C Hendriks, Richard HeusdensAbstract:A speech pre-processing algorithm is presented to improve the speech intelligibility in noise for the near-end listener. The algorithm improves the intelligibility by optimally redistributing the speech energy over time and frequency for a Perceptual Distortion measure, which is based on a spectro-temporal auditory model. In contrast to spectral-only models, short-time information is taken into account. As a consequence, the algorithm is more sensitive to transient regions, which will therefore receive more amplification compared to stationary vowels. It is known from literature that changing the vowel-transient energy ratio is beneficial for improving speech-intelligibility in noise. Objective intelligibility prediction results show that the proposed method has higher speech intelligibility in noise compared to two other reference methods, without modifying the global speech energy.
-
ICASSP - A speech preprocessing strategy for intelligibility improvement in noise based on a Perceptual Distortion measure
2012 IEEE International Conference on Acoustics Speech and Signal Processing (ICASSP), 2012Co-Authors: Cees H Taal, Richard C Hendriks, Richard HeusdensAbstract:A speech pre-processing algorithm is presented to improve the speech intelligibility in noise for the near-end listener. The algorithm improves the intelligibility by optimally redistributing the speech energy over time and frequency for a Perceptual Distortion measure, which is based on a spectro-temporal auditory model. In contrast to spectral-only models, short-time information is taken into account. As a consequence, the algorithm is more sensitive to transient regions, which will therefore receive more amplification compared to stationary vowels. It is known from literature that changing the vowel-transient energy ratio is beneficial for improving speech-intelligibility in noise. Objective intelligibility prediction results show that the proposed method has higher speech intelligibility in noise compared to two other reference methods, without modifying the global speech energy.
-
high resolution spherical quantization of sinusoidal parameters using a Perceptual Distortion measure
International Conference on Acoustics Speech and Signal Processing, 2005Co-Authors: P Korten, Jesper Jensen, Richard HeusdensAbstract:Sinusoidal modelling is a key technology in low rate audio coding, and methods for efficient quantization of sinusoidal parameters are therefore of high importance. We derive analytical formulas for the optimal entropy constrained unrestricted spherical quantizers for amplitude, phase and frequency, using a Perceptual Distortion measure. This is done both for a single sinusoid, and for multiple sinusoids distributed over multiple segments. The quantizers minimize a high-resolution approximation of the expected Distortion, while the corresponding quantization indices satisfy an entropy constraint. The quantizers turn out to be flexible and of low complexity, in the sense that they can be determined easily for varying bit rate requirements, without any sort of retraining or iterative procedures. In objective and subjective comparison tests, the proposed method is shown to outperform an existing state-of-the-art sinusoidal quantization scheme, where quantization of frequency parameters is done independently.
-
ICASSP (3) - High resolution spherical quantization of sinusoidal parameters using a Perceptual Distortion measure
Proceedings. (ICASSP '05). IEEE International Conference on Acoustics Speech and Signal Processing 2005., 1Co-Authors: P Korten, Jesper Jensen, Richard HeusdensAbstract:Sinusoidal modelling is a key technology in low rate audio coding, and methods for efficient quantization of sinusoidal parameters are therefore of high importance. We derive analytical formulas for the optimal entropy constrained unrestricted spherical quantizers for amplitude, phase and frequency, using a Perceptual Distortion measure. This is done both for a single sinusoid, and for multiple sinusoids distributed over multiple segments. The quantizers minimize a high-resolution approximation of the expected Distortion, while the corresponding quantization indices satisfy an entropy constraint. The quantizers turn out to be flexible and of low complexity, in the sense that they can be determined easily for varying bit rate requirements, without any sort of retraining or iterative procedures. In objective and subjective comparison tests, the proposed method is shown to outperform an existing state-of-the-art sinusoidal quantization scheme, where quantization of frequency parameters is done independently.
Richard C Hendriks - One of the best experts on this subject based on the ideXlab platform.
-
speech reinforcement with a globally optimized Perceptual Distortion measure for noisy reverberant channels
International Workshop on Acoustic Signal Enhancement, 2014Co-Authors: Joao B Crespo, Richard C HendriksAbstract:In this paper, a time-frequency weighting is proposed for speech reinforcement (near-end listening enhancement) in a noisy and reverberant environment, which optimizes a Perceptual Distortion measure globally for a number of time-frequency bins. Simulations confirm the optimality of the algorithm and a comparison is made to three reference methods using two additional instrumental measures.
-
speech energy redistribution for intelligibility improvement in noise based on a Perceptual Distortion measure
Computer Speech & Language, 2014Co-Authors: Cees H Taal, Richard C Hendriks, Richard HeusdensAbstract:Abstract A speech pre-processing algorithm is presented that improves the speech intelligibility in noise for the near-end listener. The algorithm improves intelligibility by optimally redistributing the speech energy over time and frequency according to a Perceptual Distortion measure, which is based on a spectro-temporal auditory model. Since this auditory model takes into account short-time information, transients will receive more amplification than stationary vowels, which has been shown to be beneficial for intelligibility of speech in noise. The proposed method is compared to unprocessed speech and two reference methods using an intelligibility listening test. Results show that the proposed method leads to significant intelligibility gains while still preserving quality. Although one of the methods used as a reference obtained higher intelligibility gains, this happened at the cost of decreased quality. Matlab code is provided.
-
speech reinforcement in noisy reverberant environments using a Perceptual Distortion measure
International Conference on Acoustics Speech and Signal Processing, 2014Co-Authors: Joao B Crespo, Richard C HendriksAbstract:In this paper, a time-frequency weighting is proposed for speech reinforcement (near-end listening enhancement) in a noisy and reverberant environment, which optimizes a Perceptual Distortion measure locally for each time-frequency bin. The algorithm acts as a dynamic range compressor, smearing out the energy of the clean speech along time. Simulations predict an intelligibility increase with respect to the unprocessed condition and two reference methods, for moderate smoothing windows, as measured by the optimized Distortion measure and two objective intelligibility measures.
-
IWAENC - Speech reinforcement with a globally optimized Perceptual Distortion measure for noisy reverberant channels
2014 14th International Workshop on Acoustic Signal Enhancement (IWAENC), 2014Co-Authors: Joao B Crespo, Richard C HendriksAbstract:In this paper, a time-frequency weighting is proposed for speech reinforcement (near-end listening enhancement) in a noisy and reverberant environment, which optimizes a Perceptual Distortion measure globally for a number of time-frequency bins. Simulations confirm the optimality of the algorithm and a comparison is made to three reference methods using two additional instrumental measures.
-
ICASSP - Speech reinforcement in noisy reverberant environments using a Perceptual Distortion measure
2014 IEEE International Conference on Acoustics Speech and Signal Processing (ICASSP), 2014Co-Authors: Joao B Crespo, Richard C HendriksAbstract:In this paper, a time-frequency weighting is proposed for speech reinforcement (near-end listening enhancement) in a noisy and reverberant environment, which optimizes a Perceptual Distortion measure locally for each time-frequency bin. The algorithm acts as a dynamic range compressor, smearing out the energy of the clean speech along time. Simulations predict an intelligibility increase with respect to the unprocessed condition and two reference methods, for moderate smoothing windows, as measured by the optimized Distortion measure and two objective intelligibility measures.
J B Rault - One of the best experts on this subject based on the ideXlab platform.
-
subband audio coding with synthesis filters minimizing a Perceptual Distortion
International Conference on Acoustics Speech and Signal Processing, 1997Co-Authors: K Gosse, Moreau F De Saintmartin, X Durot, Pierre Duhamel, J B RaultAbstract:The design of filter banks for source coding purposes classically relies on the perfect reconstruction (PR) property. However, several studies have shown that taking the quantization noise into account in the design could yield a noticeable reduction of the mean square reconstruction error. The purpose of this study is to show that Perceptual improvement can also be obtained in the particular audio coding context by relaxing the PR constraint. In this context, the mean square error is not relevant any more, and we define a new Perceptual Distortion criterion, making use of a simplified ear model, the MPE (mean Perceptual error). Then, synthesis filters are optimized so as to minimize this MPE. Finally, this MMPE (minimum MPE) filter bank is included in an audio coding scheme. Compared to the corresponding PR filter bank-based scheme by the means of POM (Perceptual objective measure), they show an improved audio quality.
-
ICASSP - Subband audio coding with synthesis filters minimizing a Perceptual Distortion
1997 IEEE International Conference on Acoustics Speech and Signal Processing, 1Co-Authors: K Gosse, X Durot, Pierre Duhamel, F. Moreau De Saint-martin, J B RaultAbstract:The design of filter banks for source coding purposes classically relies on the perfect reconstruction (PR) property. However, several studies have shown that taking the quantization noise into account in the design could yield a noticeable reduction of the mean square reconstruction error. The purpose of this study is to show that Perceptual improvement can also be obtained in the particular audio coding context by relaxing the PR constraint. In this context, the mean square error is not relevant any more, and we define a new Perceptual Distortion criterion, making use of a simplified ear model, the MPE (mean Perceptual error). Then, synthesis filters are optimized so as to minimize this MPE. Finally, this MMPE (minimum MPE) filter bank is included in an audio coding scheme. Compared to the corresponding PR filter bank-based scheme by the means of POM (Perceptual objective measure), they show an improved audio quality.
Peter Svensson - One of the best experts on this subject based on the ideXlab platform.
-
Multisensory modulation of experimentally evoked Perceptual Distortion of the face
Journal of oral rehabilitation, 2017Co-Authors: Lilja Kristin Dagsdottir, Ina Skyt, Lene Vase, Eduardo Castrillon, Lene Baad-hansen, Valeria Bellan, Peter SvenssonAbstract:Background Chronic orofacial pain patients often perceive the painful face area as ‘swollen’ without clinical signs, i.e., a Perceptual Distortion (PD). Local anesthetic (LA) injections in healthy participants are also associated with PD Objective The aim was to explore whether PD evoked by LA into the infraorbital region of could be modulated by adding mechanical stimulation (MS) to the affected area Methods MS was given with a brush and a 128 mN von Frey filament. First, sixty healthy participants were randomly divided into three groups: 1) LA control, 2) LA with MS, 3) Isotonic solution (ISO) with MS as an additional control condition. To further examine the role of a multisensory modulation an additional experiment was conducted. Twenty participants received LA with MS (filament) in addition to visual feedback of their distorted face. The results of the two experiments are presented together Results All three LA groups experienced PD, per contra PD was not reported in the ISO group. MS alone did not change the magnitude of PD: brush (p = 0.089), filament (p = 0.203). However, when the filament stimulation was combined with additional visual information of a distorted face there was observable decrease in PD (p = 0.002) Conclusion The findings indicate the importance of multisensory integration for PD, and represent a significant step forward in the understanding of the factors that may influence this common condition. Future studies are encouraged to investigate further the cortical processing for possible implications for PD in pain management. This article is protected by copyright. All rights reserved.
-
Perceptual Distortion of the tongue by lingual nerve block and topical application of capsaicin in healthy women
Clinical Oral Investigations, 2017Co-Authors: Mika Honda, Lilja Kristin Dagsdottir, Takashi Iida, Osamu Komiyama, Misao Kawara, Lene Baad-hansen, Peter SvenssonAbstract:Objectives The aim of this study was to examine reports of Perceptual Distortion evoked by transient deafferentation and burning pain as models of aspects of burning mouth syndrome (BMS). Materials and methods Sixteen healthy women took part in three experimental sessions that included exposure to lingual nerve block, capsaicin, and control substance. In each session, reported Perceptual Distortion and mechanical detection threshold (MDT) were assessed at four areas (the tongue, lower front teeth, lower lip, and right thumb) before and at 5, 15, 30 min and 1 and 3 h after the injection or application. A numerical rating scale (NRS) and a template matching procedure were used to quantify the Perceptual Distortions. Results There was a significantly higher MDT on the tongue during the lingual nerve block session at 5 min up until 1 h, with the perceived tongue size significantly increased at 5, 15, and 30 min and at 1 h compared to baseline ( P
-
Perceptual Distortion of the tongue by lingual nerve block and topical application of capsaicin in healthy women
Clinical Oral Investigations, 2017Co-Authors: Lilja Kristin Dagsdottir, Lene Baadhansen, Peter Svensson, Mika Honda, Takashi Iida, Osamu Komiyama, Misao KawaraAbstract:Objectives The aim of this study was to examine reports of Perceptual Distortion evoked by transient deafferentation and burning pain as models of aspects of burning mouth syndrome (BMS).
-
Reports of Perceptual Distortion of the face are common in patients with different types of chronic oro-facial pain.
Journal of oral rehabilitation, 2016Co-Authors: Lilja Kristin Dagsdottir, Ina Skyt, Lene Vase, Eduardo Castrillon, Lene Baad-hansen, Peter SvenssonAbstract:Anecdotally, chronic oro-facial pain patients may perceive the painful face area as 'swollen'. Because there are no clinical signs, these self-reported 'illusions' may represent Perceptual Distortions and can be speculated to contribute to the maintenance of oro-facial pain. This descriptive study investigated whether chronic oro-facial pain patients experience Perceptual Distortions - a kind of body image disruption. Sixty patients were consecutively recruited to fill in questionnaires regarding i) pain experience, ii) self-reports of Perceptual Distortion and iii) psychological condition. Perceptual Distortions were examined in the total group and in three diagnostic subgroups: i) painful post-traumatic trigeminal neuropathy (PPTN), ii) painful temporomandibular disorder (TMD) or iii) persistent idiopathic facial pain (PIFP). A large proportion of oro-facial pain patients reported Perceptual Distortions of the face (55·0%). In the diagnostic subgroups, Perceptual Distortions were most pronounced in PPTN patients (81·5%) but with no significant group differences. In the total group of chronic oro-facial pain patients, the present pain intensity explained 16·9% of the variance in magnitude of the Perceptual Distortions (R(2) = 16·9, F(31) = 6·3, P = 0·017). This study demonstrates that many chronic oro-facial pain patients may experience Perceptual Distortions. Future studies may clarify the mechanisms underlying Perceptual Distortions, which may point towards new complementary strategies for the management of chronic oro-facial pain.
-
experimental orofacial pain and sensory deprivation lead to Perceptual Distortion of the face in healthy volunteers
Experimental Brain Research, 2015Co-Authors: Lilja Kristin Dagsdottir, Ina Skyt, Lene Vase, Lene Baadhansen, Eduardo Castrillon, Peter SvenssonAbstract:Patients suffering from persistent orofacial pain may sporadically report that the painful area feels “swollen” or “differently,” a phenomenon that may be conceptualized as a Perceptual Distortion because there are no clinical signs of swelling present. Our aim was to investigate whether standardized experimental pain and sensory deprivation of specific orofacial test sites would lead to changes in the size perception of these face areas. Twenty-four healthy participants received either 0.2 mL hypertonic saline (HS) or local anesthetics (LA) into six regions (buccal, mental, lingual, masseter muscle, infraorbital and auriculotemporal nerve regions). Participants estimated the perceived size changes in percentage (0 % = no change, −100 % = half the size or +100 % = double the size), and somatosensory function was checked with tactile stimuli. The pain intensity was rated on a 0–10 Verbal Numerical Rating Scale (VNRS), and sets of psychological questionnaires were completed. HS and LA were associated with significant self-reported Perceptual Distortions as indicated by consistent increases in perceived size of the adjacent face areas (P ≤ 0.050). Perceptual Distortion was most pronounced in the buccal region, and the smallest increase was observed in the auriculotemporal region. HS was associated with moderate levels of pain VNRS = 7.3 ± 0.6. Weak correlations were found between HS-evoked Perceptual Distortion and level of dissociation in two regions (P < 0.050). Experimental pain and transient sensory deprivation evoked Perceptual Distortions in all face regions and overall demonstrated the importance of afferent inputs for the perception of the face. We propose that Perceptual Distortion may be an important phenomenon to consider in persistent orofacial pain conditions.
K Gosse - One of the best experts on this subject based on the ideXlab platform.
-
subband audio coding with synthesis filters minimizing a Perceptual Distortion
International Conference on Acoustics Speech and Signal Processing, 1997Co-Authors: K Gosse, Moreau F De Saintmartin, X Durot, Pierre Duhamel, J B RaultAbstract:The design of filter banks for source coding purposes classically relies on the perfect reconstruction (PR) property. However, several studies have shown that taking the quantization noise into account in the design could yield a noticeable reduction of the mean square reconstruction error. The purpose of this study is to show that Perceptual improvement can also be obtained in the particular audio coding context by relaxing the PR constraint. In this context, the mean square error is not relevant any more, and we define a new Perceptual Distortion criterion, making use of a simplified ear model, the MPE (mean Perceptual error). Then, synthesis filters are optimized so as to minimize this MPE. Finally, this MMPE (minimum MPE) filter bank is included in an audio coding scheme. Compared to the corresponding PR filter bank-based scheme by the means of POM (Perceptual objective measure), they show an improved audio quality.
-
ICASSP - Subband audio coding with synthesis filters minimizing a Perceptual Distortion
1997 IEEE International Conference on Acoustics Speech and Signal Processing, 1Co-Authors: K Gosse, X Durot, Pierre Duhamel, F. Moreau De Saint-martin, J B RaultAbstract:The design of filter banks for source coding purposes classically relies on the perfect reconstruction (PR) property. However, several studies have shown that taking the quantization noise into account in the design could yield a noticeable reduction of the mean square reconstruction error. The purpose of this study is to show that Perceptual improvement can also be obtained in the particular audio coding context by relaxing the PR constraint. In this context, the mean square error is not relevant any more, and we define a new Perceptual Distortion criterion, making use of a simplified ear model, the MPE (mean Perceptual error). Then, synthesis filters are optimized so as to minimize this MPE. Finally, this MMPE (minimum MPE) filter bank is included in an audio coding scheme. Compared to the corresponding PR filter bank-based scheme by the means of POM (Perceptual objective measure), they show an improved audio quality.