Papers › wav2letter++: The Fastest Open-source Speech Recognition System

wav2letter++: The Fastest Open-source Speech Recognition System

18 Dec 2018arXiv:1812.07625archive 2025-07-28

Vineel Pratap, Awni Hannun, Qiantong Xu, Jeff Cai, Jacob Kahn, Gabriel Synnaeve, Vitaliy Liptchinsky, Ronan Collobert

This paper introduces wav2letter++, the fastest open-source deep learning speech recognition framework. wav2letter++ is written entirely in C++, and uses the ArrayFire tensor library for maximum efficiency. Here we explain the architecture and design of the wav2letter++ system and compare it to other major open-source speech recognition systems. In some cases wav2letter++ is more than 2x faster than other optimized frameworks for training end-to-end neural networks for speech recognition. We also show that wav2letter++'s training times scale linearly to 64 GPUs, the highest we tested, for models with 100 million parameters. High-performance frameworks enable fast iteration, which is often a crucial factor in successful research and model tuning on new datasets and tasks.

PaperPDFCode

In Syntology Open this paper in Syntology's Atlas, the map of the papers in Syntology's graph and their citations.

Code

bsridatta/wav2letter-Swedish mentioned on GitHubNOASSERTION report
facebookresearch/wav2letter mentioned on GitHub report
flashlight/wav2letter mentioned on GitHubNOASSERTION report
gcambara/wav2letter mentioned on GitHubNOASSERTION report
jakeju/wav2letter mentioned on GitHubNOASSERTION report
krantirk/Wav2letterPlus mentioned on GitHubNOASSERTION report
lilei-John/wav2letter mentioned on GitHubNOASSERTION report
mailong25/wav2letter mentioned on GitHubNOASSERTION report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Speech Recognition

Results from the paper archive 2025-07-28

No leaderboard rows for this paper in the archive.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections