Wav2Vec2
Model ● Covered in 4 stories + Follow
Wav2Vec2 is a speech recognition model that uses Connectionist Temporal Classification architecture for automatic speech recognition tasks. Recent developments include optimization for Graphcore's IPU processors through the Optimum library, support for processing arbitrarily long audio files through chunk-based processing, integration with n-gram language models to improve transcription accuracy, and extension to multilingual variants like XLS-R for low-resource language ASR applications.
Updated 8 August 2026
Specifications
No specifications recorded yet.
Latest developments
Q2 2022
Q1 2022
- Making automatic speech recognition work on large files with Wav2Vec2 in 🤗 Transformers
- Boosting Wav2Vec2 with n-grams in 🤗 Transformers
Q4 2021
Relationships
Products & technology
- Integrated with Transformers · 1 source
- Hugging Face develops this model · 1 source
- XLS-R derived from this model · 1 source
- Transformers integrated with this model · 1 source