Deep Learning Based Speech Quality Prediction / by Gabriel Mittag
By: Mittag, Gabriel, autor
Material type:
E-bookSeries: (T-Labs Series in Telecommunication Services, 2192-2829).Publisher: Cham : Springer International Publishing, 2022Edition: First edition 2022.Description: 1 recurso en línea (XIV, 165 páginas) : 58 ilustraciones, 54 ilustraciones a color.ISBN: 9783030914790.Subject: Aprendizaje automático
| Item type | Current library | Collection | Call number | Status | Date due | Barcode | Item holds | |
|---|---|---|---|---|---|---|---|---|
LIBRO-E NO PRÉSTAMO
|
Madrid Digital Acceso Electrónico (UEM) | Ciencias e Ingeniería | TK5105.8865 2022 EB (Browse shelf(Opens below)) | Acceso electrónico | eBook.01042409 |
Browsing Madrid Digital shelves, Shelving location: Acceso Electrónico (UEM) Close shelf browser (Hides shelf browser)
| TK5105.88573 2020 EB Decentralised Internet of Things : A Blockchain Perspective | TK5105.8863 .D34 2011 EB DAFX : digital audio effects | TK5105.8865 2019 EB VoIP Technology : Applications and Challenges | TK5105.8865 2022 EB Deep Learning Based Speech Quality Prediction | TK5105.8865 .A53 2016 EB VoIP and PBX Security and Forensics : A Practical Approach | TK5105.8865 .M33 2011 EB Asterisk : the definitive guide | TK5105.888 2021 EB Guide to Web Development with Java . Understanding Website Creation |
1. Introduction -- 2. Quality Assessment of Transmitted Speech -- 3. Neural Network Architectures for Speech Quality Prediction -- 4. Double-Ended Speech Quality Prediction Using Siamese Networks -- 5. Prediction of Speech Quality Dimensions With Multi-Task Learning -- 6. Bias-Aware Loss for Training From Multiple Datasets -- 7. NISQA - A Single-Ended Speech Quality Model -- 8. Conclusions -- A. Dataset Condition Tables -- B. Train and Validation Dataset Dimension Histograms -- References.
This book presents how to apply recent machine learning (deep learning) methods for the task of speech quality prediction. The author shows how recent advancements in machine learning can be leveraged for the task of speech quality prediction and provides an in-depth analysis of the suitability of different deep learning architectures for this task. The author then shows how the resulting model outperforms traditional speech quality models and provides additional information about the cause of a quality impairment through the prediction of the speech quality dimensions of noisiness, coloration, discontinuity, and loudness.
There are no comments on this title.