# Hugging Face — Making automatic speech recognition work on large files with Wav2Vec2 in 🤗 Transformers

- Company: Hugging Face (huggingface.co)
- Announced: 2022-02-01
- Category: not stated
- Coverage: not counted
- Announcement: no
- Group: routine
- Source: https://huggingface.co/blog/asr-chunking
- Record: https://forck.live/items/2227-making-automatic-speech-recognition-work-on-large-files-with-wav2vec2-in
- Subject: Platform
- Models affected: Wav2Vec2, facebook/wav2vec2-base-960h

To use chunking with stride to perform automatic speech recognition on arbitrarily long files using Wav2Vec2 in the Transformers library, leveraging the CTC architecture to maintain quality.

## Evidence

Verbatim from https://huggingface.co/blog/asr-chunking:

> This post explains how to use the specificities of the Connectionist Temporal Classification (CTC) architecture in order to achieve very good quality automatic speech recognition (ASR) even on arbitrarily long files or during live inference.

---

Record: https://forck.live/items/2227-making-automatic-speech-recognition-work-on-large-files-with-wav2vec2-in
Catalogue: https://forck.live/llms.txt
Current issue: https://forck.live/feed.md
