Modification of the speech feature extraction module for the improvement of the system for automatic lectures transcription

This contribution is about experiments with different speech feature extraction methods and strategies where the goal has been to improve the result and the resulting recognition rate of the speech recognizer of an automatic audio speech signal transcription system. The extraction of speech features...

Ausführliche Beschreibung

Gespeichert in:

Bibliographische Detailangaben
Hauptverfasser:	Chaloupka, J., Cerva, P., Silovsky, J., Zd'ansky, J., Nouza, J.
Format:	Tagungsbericht
Sprache:	eng
Schlagworte:	Automatic Speech Transcription Feature extraction Hidden Markov models LVCSR Mel frequency cepstral coefficient Speech Speech Feature Extraction Speech recognition Vocabulary
Online-Zugang:	Volltext bestellen
Tags:	Tag hinzufügen Keine Tags, Fügen Sie den ersten Tag hinzu!

Beschreibung
Zusammenfassung:	This contribution is about experiments with different speech feature extraction methods and strategies where the goal has been to improve the result and the resulting recognition rate of the speech recognizer of an automatic audio speech signal transcription system. The extraction of speech features is based on MFCC (Mel Frequency Cepstral Coefficients) and PLP (Perceptual Linear Prediction), which are normally used in different transcription systems around the world. The speech recognizer with different speech features has been tested on our speech database where audio (or video) recordings from archives of university lectures are stored. The result from our experiments is that we get higher recognition rate if PLP based audio speech features are used.
ISSN:	1334-2630