Model reference · open weights

wav2vec2-indonesian-javanese-sundanese

wav2vec2-indonesian-javanese-sundanese is an open-weight audio or speech model from indonesian-nlp, listed in the AxForge catalogue. AxForge can bring it up on EU-owned hardware for you on request — with the licence handled where one is required.

Audio indonesian-nlp 1 variants 2.5M downloads/mo
Request this model on EU hardware All served models Not on the shared API today — deployed on request.

About

What wav2vec2-indonesian-javanese-sundanese is

Multilingual Speech Recognition for Indonesian Languages This is the model built for the project Multilingual Speech Recognition for Indonesian Languages. It is a fine-tuned facebook/wav2vec2-large-xlsr-53 model on the Indonesian Common Voice dataset, High-quality TTS data for Javanese - SLR41, and High-quality TTS data for Sundanese - SLR44 datasets. We also provide a live demo to test the model. When using this model, make sure that your speech input is sampled at 16kHz. Usage The model can be used directly (without a language model) as follows: Evaluation The model can be evaluated as follows on the Indonesian test data of Common Voice. Test Result: 11.57 % Training The Common Voice train, validation, and ... datasets were used for training as well as ... and ... # TODO The script used for training can be found here (will be available soon)

Summarised from the published model card. Read the full card on the HuggingFace links below.

Specifications

What it is

Makerindonesian-nlp
TypeAudio & music
Variants1
Runs withtransformers
Released2022-03-02
Popularity2.5M downloads / month
Likes15
LicenceOpen weights

How it works

How audio & music work

Audio or textinputAudio modelrecognise / synthesiseText or audiooutputSpeech-to-text turns audio into text; text-to-speech and music models turn text into audio.

Variants

Sizes & precisions

Open weights ship in several sizes and precisions. One page, all the variants — pick the one that fits your GPU. VRAM figures are estimates from model size.

VariantParamsPrecisionVRAMFits 16 GBWeights
wav2vec2-indonesian-javanese-sundaneseBF16Weights ↗

Benchmarks

Reported results

As published on the model card — the maker's own numbers, not measured by AxForge.

TaskDatasetMetricScore
Automatic Speech RecognitionCommon Voice 6.1Test WER4.056
Automatic Speech RecognitionCommon Voice 6.1Test CER1.472
Automatic Speech RecognitionCommon Voice 7Test WER4.492
Automatic Speech RecognitionCommon Voice 7Test CER1.577
Automatic Speech RecognitionRobust Speech Event - Dev DataTest WER48.94
Automatic Speech RecognitionRobust Speech Event - Test DataTest WER68.95

Using it via the API

Call it like any OpenAI endpoint

Once AxForge deploys wav2vec2-indonesian-javanese-sundanese for you, it answers on the OpenAI-compatible API — the same base URL and keys as every other model. (wav2vec2-indonesian-javanese-sundanese below is illustrative; you get the exact model name on deployment.)

$ curl -sS https://api.axforge.ai/v1/audio/transcriptions \
  -H "Authorization: Bearer $AXFORGE_API_KEY" \
  -F model="wav2vec2-indonesian-javanese-sundanese" -F file=@audio.mp3

Details

Languages, data & research

Languages

id jv sun

Trained / evaluated on

mozilla-foundation/common_voice_7_0 openslr magic_data titml

Tags

transformers pytorch wav2vec2 automatic-speech-recognition audio hf-asr-leaderboard id jv robust-speech-event speech su sun dataset:mozilla-foundation/common_voice_7_0 dataset:openslr

Licence

Open weights

Open weights under apache-2.0 — commercial use is permitted. Deploy it on AxForge EU hardware on request. Read the licence ↗

Sources

Weights & code

Want wav2vec2-indonesian-javanese-sundanese on EU-owned hardware?

Request this model on EU hardware See what’s served now

Explore

More audio & music

© 2026 AxForge · EU-hosted AI infrastructure Pricing Docs Trust Privacy Terms