Wav2Vec2
Model ● Covered in 4 stories + Follow
Wav2Vec2 is a speech recognition model that uses Connectionist Temporal Classification architecture for automatic speech recognition tasks. Recent developments include optimization for Graphcore's IPU processors through the Optimum library, support for processing arbitrarily long audio files through chunk-based processing, integration with n-gram language models to improve transcription accuracy, and extension to multilingual variants like XLS-R for low-resource language ASR applications.
Updated 8 August 2026
Specifications
No specifications recorded yet.
Latest developments
2022
- Graphcore and Hugging Face Launch New Lineup of IPU-Ready Transformers
- Making automatic speech recognition work on large files with Wav2Vec2 in 🤗 Transformers
- Boosting Wav2Vec2 with n-grams in 🤗 Transformers
2021
Relationships
Products & technology
- Integrated with Transformers · 1 source
- Hugging Face develops this model · 1 source
- XLS-R derived from this model · 1 source
- Transformers integrated with this model · 1 source