2015 Kleijn Optimizing speech intelligibility in a noisy environment.pdf (542.64 kB)

Optimizing speech intelligibility in a noisy environment: A unified view

Download (542.64 kB)
journal contribution
posted on 30.03.2021, 01:44 by Willem Kleijn, JB Crespo, RC Hendriks, P Petkov, B Sauert, P Vary
Modern communication technology facilitates communication from anywhere to anywhere. As a result, low speech intelligibility has become a common problem, which is exacerbated by the lack of feedback to the talker about the rendering environment. In recent years, a range of algorithms has been developed to enhance the intelligibility of speech rendered in a noisy environment. We describe methods for intelligibility enhancement from a unified vantage point. Before one defines a measure of intelligibility, the level of abstraction of the representation must be selected. For example, intelligibility can be measured on the message, the sequence of words spoken, the sequence of sounds, or a sequence of states of the auditory system. Natural measures of intelligibility defined at the message level are mutual information and the hit-or-miss criterion. The direct evaluation of high-level measures requires quantitative knowledge of human cognitive processing. Lower-level measures can be derived from higher-level measures by making restrictive assumptions. We discuss the implementation and performance of some specific enhancement systems in detail, including speech intelligibility index (SII)-based systems and systems aimed at enhancing the sound-field where it is perceived by the listener. We conclude with a discussion of the current state of the field and open problems. © 2015 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works.

History

Preferred citation

Kleijn, W. B., Crespo, J. B., Hendriks, R. C., Petkov, P., Sauert, B. & Vary, P. (2015). Optimizing speech intelligibility in a noisy environment: A unified view. IEEE Signal Processing Magazine, 32(2), 43-54. https://doi.org/10.1109/MSP.2014.2365594

Journal title

IEEE Signal Processing Magazine

Volume

32

Issue

2

Publication date

01/03/2015

Pagination

43-54

Publisher

IEEE

Publication status

Published

Contribution type

Article

Online publication date

10/02/2015

ISSN

1053-5888

eISSN

1558-0792

Article number

2

Language

en

Exports