Model reference · open weights

qwen3-forced-aligner-q4-k-m

Available as managed deployment Audio OpenVoiceOS Speech→text 1 variants 3k dl/mo

qwen3-forced-aligner-q4-k-m is an open-weight audio or speech model from OpenVoiceOS. AxForge deploys and operates it for you on dedicated EU-owned hardware — with the licence handled where one is required.

Available as managed deployment — configured and operated for you on dedicated EU hardware, quoted per deployment.

What it is

Released byOpenVoiceOS
TypeAudio & music
TaskSpeech→text
Released2026-02-19
Popularity3k downloads / month
LicenceUnknown

About

What qwen3-forced-aligner-q4-k-m is

license: apache-2.0

OVOS - Qwen3 Forced Aligner 0.6B Q4_K_M (GGUF)

This model is a quantized gguf-format export of Qwen/Qwen3-ForcedAligner-0.6B for ease of use in edge devices and CPU-based inference environments. The original model is transformed into gguf with F16 tensors by the script convert_hf_to_gguf.py and then further quantized, if needed, using the tool quantize from the same repo.

Requirements

The requirements can be installed as

Read the full model card
$ pip install git+https://github.com/femelo/py-qwen3-asr-cpp

Usage

from py_qwen3_asr_cpp.model import Qwen3ASRModel

# Initialize the model (it handles downloading from this repo)
model = Qwen3ASRModel(
    asr_model="qwen3-asr-0.6b-q4-k-m",
    align_model="qwen3-forced-aligner-0.6b-q4-k-m",
    n_threads=4
)

# Transcribe from file
result, alignment = model.transcribe_and_align("audio.mp3")
print(f"Detected Language: {result.language}")
print(f"Transcription: {result.text}")

Refer to https://github.com/femelo/py-qwen3-asr-cpp for more details.

Licensing

The license is derived from the original model: Apache 2.0. For more details, please refer to Qwen/Qwen3-ForcedAligner-0.6B.

From the published model card. Full card on the HuggingFace links in the sidebar.

How it works

How audio & music work

Audio or textinputAudio modelrecognise / synthesiseText or audiooutputSpeech-to-text turns audio into text; text-to-speech and music models turn text into audio.

Using it via the API

Call it like any OpenAI endpoint

Once AxForge deploys qwen3-forced-aligner-q4-k-m for you, it answers on the OpenAI-compatible API — the same base URL and keys as every other model. (qwen3-forced-aligner-q4-k-m below is illustrative; you get the exact model name on deployment.)

$ curl -sS https://api.axforge.ai/v1/audio/transcriptions \
  -H "Authorization: Bearer $AXFORGE_API_KEY" \
  -F model="qwen3-forced-aligner-q4-k-m" -F file=@audio.mp3

Create an account — your API key is available in the console. 3M free tokens every 30 days with every new account.

© 2026 AxForge · EU-hosted AI infrastructure Pricing Docs Trust Privacy Terms