Audio ML · 2024
Speech Phonation Mode Classification
Learning to hear how we sing.
The project
Built a system using HuBERT, wav2vec2, and wavelet features to classify phonation modes in singing. Improved classification accuracy by 12-15% over traditional signal processing methods by leveraging self-supervised learning representations.
Built with
Audio Processing · Self-Supervised Learning · HuBERT