Model reference · open weights

w2v-xls-r-uk

w2v-xls-r-uk is an open-weight audio or speech model from Yehor, listed in the AxForge catalogue. AxForge can bring it up on EU-owned hardware for you on request — with the licence handled where one is required.

Audio Yehor 1 variants 803k downloads/mo
Request this model on EU hardware All served models Not on the shared API today — deployed on request.

About

What w2v-xls-r-uk is

🚨🚨🚨 ATTENTION! 🚨🚨🚨 Use an updated model: https://huggingface.co/Yehor/w2v-bert-uk-v2.1 Community - Discord: https://bit.ly/discord-uds - Speech Recognition: https://t.me/speechrecognitionuk - Speech Synthesis: https://t.me/speechsynthesisuk See other Ukrainian models: https://github.com/egorsmkv/speech-recognition-uk Evaluation results Metrics (float16) using evaluate library with batchsize=1: - WER: 0.2024 metric, 20.24% - CER: 0.0364 metric, 3.64% - Accuracy on words: 79.76% - Accuracy on chars: 96.36% - Inference time: 63.4848 seconds - Audio duration: 16665.5212 seconds - RTF: 0.0038 Cite this work

Summarised from the published model card. Read the full card on the HuggingFace links below.

Specifications

What it is

MakerYehor
TypeAudio & music
Parameters (lead)315M
Variants1
Runs withtransformers
Based onfacebook/wav2vec2-xls-r-300m
Released2022-06-08
Popularity803k downloads / month
Likes8
LicenceOpen weights

How it works

How audio & music work

Audio or textinputAudio modelrecognise / synthesiseText or audiooutputSpeech-to-text turns audio into text; text-to-speech and music models turn text into audio.

Variants

Sizes & precisions

Open weights ship in several sizes and precisions. One page, all the variants — pick the one that fits your GPU. VRAM figures are estimates from model size.

VariantParamsPrecisionVRAMFits 16 GBWeights
w2v-xls-r-uk315MBF16~0.7 GBWeights ↗

Benchmarks

Reported results

As published on the model card — the maker's own numbers, not measured by AxForge.

TaskDatasetMetricScore
Automatic Speech Recognitioncommon_voice_10_0WER20.24
Automatic Speech Recognitioncommon_voice_10_0CER3.64

Using it via the API

Call it like any OpenAI endpoint

Once AxForge deploys w2v-xls-r-uk for you, it answers on the OpenAI-compatible API — the same base URL and keys as every other model. (w2v-xls-r-uk below is illustrative; you get the exact model name on deployment.)

$ curl -sS https://api.axforge.ai/v1/audio/transcriptions \
  -H "Authorization: Bearer $AXFORGE_API_KEY" \
  -F model="w2v-xls-r-uk" -F file=@audio.mp3

Details

Languages, data & research

Languages

uk

Trained / evaluated on

mozilla-foundation/common_voice_10_0

Tags

transformers safetensors wav2vec2 automatic-speech-recognition uk dataset:mozilla-foundation/common_voice_10_0 doi:10.57967/hf/4556 model-index endpoints_compatible deploy:azure

Licence

Open weights

Open weights under apache-2.0 — commercial use is permitted. Deploy it on AxForge EU hardware on request. Read the licence ↗

Sources

Weights & code

Want w2v-xls-r-uk on EU-owned hardware?

Request this model on EU hardware See what’s served now

Explore

More audio & music

© 2026 AxForge · EU-hosted AI infrastructure Pricing Docs Trust Privacy Terms