mcit/qatari-speech-to-text
Qatari Speech To Text
WER (MSA)
8.4%
WER (Qatari dialect)
13.1%
WER (code-switched)
16.5%
RTF (A10G)
0.18
About this model
Whisper large-v3 fine-tune covering Modern Standard Arabic and the Qatari dialect, optimised for call-centre audio, public consultations, and broadcast material. Handles code-switching into English tokens common in service conversations.
Intended use
Transcription of citizen-facing calls, interviews, and public hearings by government entities and licensed contact-centre operators.
Training data lineage
Fine-tuned on 1,900 hours from the khaliji-speech-corpus with a Qatari-dialect weighted sampling schedule.
Limitations & bias notes
Word error rate rises noticeably on heavy Bedouin-register speech and on overlapping speakers in majlis-style recordings. Numerals spoken in mixed Arabic-English sequences are occasionally normalised inconsistently.
Evaluation metrics
| WER (MSA) | 8.4% |
| WER (Qatari dialect) | 13.1% |
| WER (code-switched) | 16.5% |
| RTF (A10G) | 0.18 |
Try it
Live sandboxcall_centre_qa_0912.wav
fixtureFeedback
Owning entity
- Updated
- 2026-06-30
- Latest version
- 3.0.0
- License
- Open
- Access
- Open
Trained on
khaliji-speech-corpus →