Model reference · open weights

wav2vec2-xls-r-cs-250

wav2vec2-xls-r-cs-250 is an open-weight audio or speech model from comodoro, listed in the AxForge catalogue. AxForge can bring it up on EU-owned hardware for you on request — with the licence handled where one is required.

Audio comodoro 1 variants 1.6M downloads/mo
Request this model on EU hardware All served models Not on the shared API today — deployed on request.

About

What wav2vec2-xls-r-cs-250 is

Czech wav2vec2-xls-r-300m-cs-250 This model is a fine-tuned version of facebook/wav2vec2-xls-r-300m on the commonvoice 8.0 dataset as well as other datasets listed below. It achieves the following results on the evaluation set: - Loss: 0.1271 - Wer: 0.1475 - Cer: 0.0329 The eval.py script results using a LM are: - WER: 0.07274312090176113 - CER: 0.021207369275558875 Model description Fine-tuned facebook/wav2vec2-large-xlsr-53 on Czech using the Common Voice dataset. When using this model, make sure that your speech input is sampled at 16kHz. The model can be used directly (without a language model) as follows: Evaluation The model can be evaluated using the attached eval.py script: Training and evaluation data The Common Voice 8.0 train and validation datasets were used for training, as well as the following datasets: - Šmídl, Luboš and Pražák, Aleš, 2013, OVM – Otázky Václava Moravce, LINDAT/CLARIAH-CZ digital library at the Institute of Formal and Applied Linguistics (ÚFAL), Faculty of Mathematics and Physics, Charles University, http://hdl.handle.net/11858/00-097C-0000-000D-EC98-3. - Pražák, Aleš and Šmídl, Luboš, 2012, Czech Parliament Meetings, LINDAT/CLARIAH-CZ digital library at the Institute of Formal and Applied Linguistics (ÚFAL), Faculty of Mathematics and Physics, Charles University, http://hdl.handle.net/11858/00-097C-0000-0005-CF9C-4. - Plátek, Ondřej; Dušek, Ondřej and Jurčíček, Filip, 2016, Vystadial 2016 – Czech data, LINDAT/CLARIAH-CZ digital library at the Institute of Formal and Applied Linguistics (ÚFAL), Faculty of Mathematics and Physics, Charles University, http://hdl.handle.net/11234/1-1740. Training hyperparameters The following hyperparameters were used during training: - learningrate: 0.0001 - trainbatchsize: 32 - evalbatchsize: 8 - seed: 42 - optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08 - lrschedulertype: linear - lrschedulerwarmupsteps: 800 - numepochs: 5 - mixedprecisiontraining: Native AMP Training results Framework versions - Transformers 4.16.2 - Pytorch 1.10.1+cu102 - Datasets 1.18.3 - Tokenizers 0.11.0

Summarised from the published model card. Read the full card on the HuggingFace links below.

Specifications

What it is

Makercomodoro
TypeAudio & music
Parameters (lead)315M
Variants1
Runs withtransformers
Based onfacebook/wav2vec2-xls-r-300m
Released2022-03-02
Popularity1.6M downloads / month
Likes3
LicenceOpen weights

How it works

How audio & music work

Audio or textinputAudio modelrecognise / synthesiseText or audiooutputSpeech-to-text turns audio into text; text-to-speech and music models turn text into audio.

Variants

Sizes & precisions

Open weights ship in several sizes and precisions. One page, all the variants — pick the one that fits your GPU. VRAM figures are estimates from model size.

VariantParamsPrecisionVRAMFits 16 GBWeights
wav2vec2-xls-r-300m-cs-250315MBF16~0.7 GBWeights ↗

Benchmarks

Reported results

As published on the model card — the maker's own numbers, not measured by AxForge.

TaskDatasetMetricScore
Automatic Speech RecognitionCommon Voice 8Test WER7.3
Automatic Speech RecognitionCommon Voice 8Test CER2.1
Automatic Speech RecognitionRobust Speech Event - Dev DataTest WER43.44
Automatic Speech RecognitionRobust Speech Event - Test DataTest WER38.5

Using it via the API

Call it like any OpenAI endpoint

Once AxForge deploys wav2vec2-xls-r-cs-250 for you, it answers on the OpenAI-compatible API — the same base URL and keys as every other model. (wav2vec2-xls-r-cs-250 below is illustrative; you get the exact model name on deployment.)

$ curl -sS https://api.axforge.ai/v1/audio/transcriptions \
  -H "Authorization: Bearer $AXFORGE_API_KEY" \
  -F model="wav2vec2-xls-r-cs-250" -F file=@audio.mp3

Details

Languages, data & research

Languages

cs

Trained / evaluated on

mozilla-foundation/common_voice_8_0 ovm pscr vystadial2016

Tags

transformers pytorch safetensors wav2vec2 automatic-speech-recognition generated_from_trainer hf-asr-leaderboard mozilla-foundation/common_voice_8_0 robust-speech-event xlsr-fine-tuning-week cs dataset:mozilla-foundation/common_voice_8_0 dataset:ovm dataset:pscr

Licence

Open weights

Open weights under apache-2.0 — commercial use is permitted. Deploy it on AxForge EU hardware on request. Read the licence ↗

Sources

Weights & code

Want wav2vec2-xls-r-cs-250 on EU-owned hardware?

Request this model on EU hardware See what’s served now

Explore

More audio & music

© 2026 AxForge · EU-hosted AI infrastructure Pricing Docs Trust Privacy Terms