PofoliaShared via Pofolia

· 2015

Audio augmentation for speech recognition

Tom Ko, Vijayaditya Peddinti, Daniel Povey, Sanjeev Khudanpur

Short summary

Changing audio speed by factors of 0.9, 1.0, and 1.1 improved large-vocabulary speech recognition (LVCSR) performance by an average of 4.3% across four tasks with 100-1000 hours of training data.

AI-generated from the title and abstract; the full text is not read.

TakeawaysIn the app
Key pointsIn the app
Ask the paperIn the app

The rest is in the Pofolia app

Takeaways, key points and questions to the paper; new summaries every day for your field. Free.

Sign in on the web to open

Field: Signal Processing

Signal ProcessingComputer Science