Real-time Speech and Music Classification by Large Audio Feature Space Extraction
Book information
Description
This book reports on an outstanding thesis that has significantly advanced the state-of-the-art in the automated analysis and classification of speech and music. It defines several standard acoustic parameter sets and describes their implementation in a novel, open-source, audio analysis framework called openSMILE, which has been accepted and intensively used worldwide. The book offers extensive descriptions of key methods for the automatic classification of speech and music signals in real-life conditions and reports on the evaluation of the framework developed and the acoustic parameter sets that were selected. It is not only intended as a manual for openSMILE users, but also and primarily as a guide and source of inspiration for students and scientists involved in the design of speech and music analysis methods that can robustly handle real-life conditions.
Similar books
Introducing Spoken Dialogue Systems into Intelligent Environments
2013 · PDF
Hierarchical Neural Network Structures for Phoneme Recognition
2013 · PDF
The Conversational Interface: Talking to Smart Devices
2016 · PDF
Situated Dialog in Speech-Based Human-Computer Interaction
2016 · PDF
Design of Video Quality Metrics with Multi-Way Data Analysis: A data driven approach
2016 · PDF
Building Dialogue POMDPs from Expert Dialogues: An end-to-end approach
2016 · PDF
Next Generation Intelligent Environments: Ambient Adaptive Systems
2016 · PDF
Haptic Teleoperation Systems: Signal Processing Perspective
2015 · PDF