Real-time Speech and Music Classification by Large Audio Feature Space Extraction / by Florian Eyben
By: Eyben, Florian
Contributor(s): SpringerLink (Online service)
Material type:
E-bookSeries: (Springer Theses, Recognizing Outstanding Ph.D. Research, 2190-5053).Publisher: Cham : Springer International Publishing, 2016Edition: 1st ed.Description: 1 recurso en línea (XXXVIII, 298 páginas) : 41 ilustraciones, 39 ilustraciones en color.ISBN: 9783319272993.Subject: Interfaces de usuario
| Item type | Current library | Collection | Call number | Copy number | Status | Date due | Barcode | Item holds | |
|---|---|---|---|---|---|---|---|---|---|
LIBRO-E NO PRÉSTAMO
|
Madrid Digital Acceso Electrónico (UEM) | Ciencias e Ingeniería | TK7895.S65 E934 2016 EB (Browse shelf(Opens below)) | .i11589681 | Acceso electrónico | eBOOK .i11589681 |
Browsing Madrid Digital shelves, Shelving location: Acceso Electrónico (UEM) Close shelf browser (Hides shelf browser)
| TK7895 .S65 2019 EB Audio Processing and Speech Recognition : Concepts, Techniques and Research overviews | TK7895.S65 2022 EB Multilingual Phone Recognition in Indian Languages | TK7895.S65 D663 2017 EB Language modeling for automatic speech recognition of inflective languages : an applications-oriented approach using lexical data | TK7895.S65 E934 2016 EB Real-time Speech and Music Classification by Large Audio Feature Space Extraction | TK7895.S65 J643 2016 EB Emotion, Affect and Personality in Speech : The Bias of Language and Paralanguage | TK7895 .S65 M38 2016 EB The Conversational Interface : Talking to Smart Devices | TK7895.S65 R365 2017 EB Speech recognition using articulatory and excitation source features |
Abstract -- Introduction -- Acoustic Features and Modelling -- Standard Baseline Feature Sets -- Real-time Incremental Processing -- Real-life Robustness -- Evaluation -- Discussion and Outlook -- Appendix -- Mel-frequency Filterbank Parameters.
This book reports on an outstanding thesis that has significantly advanced the state-of-the-art in the automated analysis and classification of speech and music. It defines several standard acoustic parameter sets and describes their implementation in a novel, open-source, audio analysis framework called openSMILE, which has been accepted and intensively used worldwide. The book offers extensive descriptions of key methods for the automatic classification of speech and music signals in real-life conditions and reports on the evaluation of the framework developed and the acoustic parameter sets that were selected. It is not only intended as a manual for openSMILE users, but also and primarily as a guide and source of inspiration for students and scientists involved in the design of speech and music analysis methods that can robustly handle real-life conditions.
There are no comments on this title.