Puzzle Multilingual ASR Test

Model: qwen3asr_small_v22_9ep_7292 /model
Task
Languages (none = auto-detect)
Context / Prompt
off
off
On silence after speech, the server finalizes the segment (commits the partial) automatically.
Restores punctuation on the committed transcript at VAD endpoints / finalizes (CPU model; the trailing mark is held back until a true end).
Labels each finalized segment with its dominant speaker (S1–S4, streaming Sortformer on CPU). Works for transcribe and translate. Requires VAD — enabling this turns VAD on automatically.
off
Restrict output to the selected language's script (needs exactly one language). E.g. English then never emits Korean/Arabic/Japanese.
Audio Input Device (mic session; e.g. BlackHole)
Input Gain (applies to mic & file, client-side)
0 dB
Audio File (decoded → 16 kHz mono, gain applied, streamed through model)
Idle
-∞ dBFS Show debug Clear history
Detected Language
Live Transcription
Speaker turns
History