Abstract
Based on the well-known relationship between vocal tract length (VTL) variation and linear frequency warping, we present a method for generating vocal tract length invariant (VTLI) features. These features are computed as translation invariant, correlation-type features in a log-frequency domain. In phoneme classification and recognition experiments on the TIMIT database, their discrimination capabilities and robustness to mismatches between training and test conditions turned out to be considerably better than for Mel-frequency cepstral coefficients (MFCCs). The best results are obtained when VTLI features and MFCCs are combined
| Originalsprache | Englisch |
|---|---|
| Titel | 2006 IEEE International Conference on Acoustics Speech and Signal Processing Proceedings |
| Seitenumfang | 4 |
| Herausgeber (Verlag) | IEEE |
| Erscheinungsdatum | 01.12.2006 |
| Seiten | 1025-1028 |
| Aufsatznummer | 1661453 |
| ISBN (Print) | 978-142440469-8 |
| DOIs | |
| Publikationsstatus | Veröffentlicht - 01.12.2006 |
| Veranstaltung | 2006 IEEE International Conference on Acoustics, Speech and Signal Processing - Toulouse, Frankreich Dauer: 14.05.2006 → 19.05.2006 Konferenznummer: 69350 |
UN SDGs
Dieser Output leistet einen Beitrag zu folgendem(n) Ziel(en) für nachhaltige Entwicklung
-
SDG 9 – Industrie, Innovation und Infrastruktur
Fingerprint
Untersuchen Sie die Forschungsthemen von „Frequency-Warping Invariant Features for Automatic Speech Recognition“. Zusammen bilden sie einen einzigartigen Fingerprint.Zitieren
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver