Model reference · open weights

wav2vec2-vi-vlsp2020

wav2vec2-vi-vlsp2020 is an open-weight audio or speech model from nguyenvulebinh, listed in the AxForge catalogue. AxForge can bring it up on EU-owned hardware for you on request — with the licence handled where one is required.

Licence fee required Audio nguyenvulebinh 1 variants 1.1M downloads/mo
Request a licence + hosting quote All served models Not on the shared API today — deployed on request.

About

What wav2vec2-vi-vlsp2020 is

Model description Our models use wav2vec2 architecture, pre-trained on 13k hours of Vietnamese youtube audio (un-label data) and fine-tuned on 250 hours labeled of VLSP ASR dataset on 16kHz sampled speech audio. You can find more description here Benchmark WER result on VLSP T1 testset: Usage [](https://colab.research.google.com/drive/1z3FQUQ2t7nIPR-dBR4bkcee6oCDGmcd4?usp=sharing) Model Parameters License The ASR model parameters are made available for non-commercial use only, under the terms of the Creative Commons Attribution-NonCommercial 4.0 International (CC BY-NC 4.0) license. You can find details at: https://creativecommons.org/licenses/by-nc/4.0/legalcode Contact nguyenvulebinh@gmail.com [](https://twitter.com/intent/follow?screenname=nguyenvulebinh)

Summarised from the published model card. Read the full card on the HuggingFace links below.

Specifications

What it is

Makernguyenvulebinh
TypeAudio & music
Variants1
Runs withtransformers
Released2022-11-04
Popularity1.1M downloads / month
Likes2
LicenceCommercial licence needed

How it works

How audio & music work

Audio or textinputAudio modelrecognise / synthesiseText or audiooutputSpeech-to-text turns audio into text; text-to-speech and music models turn text into audio.

Variants

Sizes & precisions

Open weights ship in several sizes and precisions. One page, all the variants — pick the one that fits your GPU. VRAM figures are estimates from model size.

VariantParamsPrecisionVRAMFits 16 GBWeights
wav2vec2-base-vi-vlsp2020BF16Weights ↗

Using it via the API

Call it like any OpenAI endpoint

Once AxForge deploys wav2vec2-vi-vlsp2020 for you, it answers on the OpenAI-compatible API — the same base URL and keys as every other model. (wav2vec2-vi-vlsp2020 below is illustrative; you get the exact model name on deployment.)

$ curl -sS https://api.axforge.ai/v1/audio/transcriptions \
  -H "Authorization: Bearer $AXFORGE_API_KEY" \
  -F model="wav2vec2-vi-vlsp2020" -F file=@audio.mp3

Details

Languages, data & research

Languages

vi

Trained / evaluated on

vlsp-asr-2020

Tags

transformers pytorch wav2vec2 automatic-speech-recognition audio vi dataset:vlsp-asr-2020 endpoints_compatible

Licence

Commercial licence needed

The weights are open but cc-by-nc-4.0 needs a commercial agreement for business use. AxForge can arrange that licence and host the model for you — you pay AxForge, we settle with the model’s maker. Ask us for a quote. Read the licence ↗

Sources

Weights & code

Want wav2vec2-vi-vlsp2020 on EU-owned hardware?

Request a licence + hosting quote See what’s served now

Explore

More audio & music

© 2026 AxForge · EU-hosted AI infrastructure Pricing Docs Trust Privacy Terms