A monolingual automatic speech recognition model for
Japanese
, fine-tuned from
openai/whisper-large-v3
. Part of
BuzzASR
,
a suite of 102 language-specialized ASR models (Findings of EMNLP 2026).
This model uses
full fine-tuning (native per-language tokenizer replacement + text multitask fine-tuning)
.
🏆
State-of-the-art (open-source).
On the combined FLEURS + Common Voice test set, this model
achieves the lowest CER of every open system we compare against: Whisper-large-v3, Omnilingual 1B/7B, MMS, Qwen3-ASR, and Cohere Transcribe.
Results (normalized CER / WER, %)
Test set
CER
WER
Whisper-large-v3 (zero-shot) CER
FLEURS
14.49
125.06
9.82
Common Voice 25
26.58
92.45
29.5
Combined
21.08
108.72
23.34
~1.1x CER reduction over Whisper zero-shot on the combined test set.
japanese huggingface.co is an AI model on huggingface.co that provides japanese's model effect (), which can be used instantly with this BuzzASR japanese model. huggingface.co supports a free trial of the japanese model, and also provides paid use of the japanese. Support call japanese model through api, including Node.js, Python, http.
japanese huggingface.co is an online trial and call api platform, which integrates japanese's modeling effects, including api services, and provides a free online trial of japanese, you can try japanese online for free by clicking the link below.
BuzzASR japanese online free url in huggingface.co:
japanese is an open source model from GitHub that offers a free installation service, and any user can find japanese on GitHub to install. At the same time, huggingface.co provides the effect of japanese install, users can directly use japanese installed effect in huggingface.co for debugging and trial. It also supports api for free installation.