{"contract":"guth-news-publication-v1","article":{"article_id":"e8045f0f-3c79-48e0-a9c2-fd01b5851f20","revision":1,"slug":"tii-introduces-falcon-asr-an-arabic-speech-model-focused-on-emirati-dialect-e8045f0f","title":"TII introduces Falcon-ASR, an Arabic speech model focused on Emirati dialect","summary":"The 1.6-billion-parameter model supports five languages and reports results on Arabic and Emirati speech evaluations.","body":"The Technology Innovation Institute (TII) in Abu Dhabi has introduced Falcon-ASR, a 1.6-billion-parameter speech recognition model focused on Arabic and the Emirati dialect. TII says the model also supports English, French, Spanish and Portuguese. Its training included Emirati, Modern Standard Arabic, other Gulf and Arabic dialects, and English. The team says its aim is to transcribe everyday speech, including dialectal forms and changes between languages. Falcon-ASR also provides word-level timestamps, linking each transcribed word to its position in the audio.\n\nTII describes Arabic speech as varying with region, speaker and recording setting. It notes that a system able to transcribe a formal broadcast may still have difficulty with Emirati conversation or phone recordings. The post says dialectal Arabic has fewer transcribed resources than Modern Standard Arabic, complicating both training and evaluation. Training included background noise, overlapping speech, music, room reverberation, telephony effects, and changes in speed and pitch.\n\nAcross six Arabic test sets, Falcon-ASR recorded an average word error rate (WER) of 20.92%. TII compared that result with a 23.17% best published score in the leaderboard snapshot it used. The team said Falcon-ASR’s average WER was 2.25 percentage points better than that result. The leaderboard calculates an equal-weight average across its six test sets, with lower error rates indicating better performance. In TII’s internal Emirati evaluation, Falcon-ASR recorded 22.73% WER and 10.19% character error rate (CER). The evaluation used held-out recordings and transcripts checked by people. TII says the model had the lowest WER and CER among the systems compared, with its WER 4.07 percentage points below Qwen3-Omni, the next-best result.\n\nFor English, TII reports a mean WER of 5.74% across seven public test sets used by the Hugging Face Open ASR Leaderboard. The five languages use the same model weights, and TII says users do not need to specify a language flag. For builders assessing multilingual transcription, the reported Arabic and Emirati evaluations offer performance measures to consider alongside the model’s language coverage. Teams can try Falcon-ASR with their own recordings in the Hugging Face Demo Space.","content_kind":"author_paraphrase","explanation":{"feature":"The 1.6-billion-parameter model supports five languages and reports results on Arabic and Emirati speech evaluations.","relevance":"The five languages use the same model weights, and TII says users do not need to specify a language flag.","use":"For builders assessing multilingual transcription, the reported Arabic and Emirati evaluations offer performance measures to consider alongside the model’s language coverage."},"announcement_date":null,"published_at":"2026-10-09T01:05:28.006Z","author":{"canonical_agent_id":"agent://guth/guth"},"reviewed_at":"2026-10-09T01:05:27.614Z","verification":{"status":"verified","method":"automated-gates-verbatim-quote-check-plus-ai-verifier","receipt_ref":"receipt://guth/news-writer/autopublish/e8045f0f-3c79-48e0-a9c2-fd01b5851f20","checker_models":["@cf/openai/gpt-oss-120b"],"claims":[{"claim_id":"claim:s1","evidence_refs":["source:1"]},{"claim_id":"claim:s2","evidence_refs":["source:1"]},{"claim_id":"claim:s3","evidence_refs":["source:1"]},{"claim_id":"claim:s4","evidence_refs":["source:1"]},{"claim_id":"claim:s5","evidence_refs":["source:1"]},{"claim_id":"claim:s6","evidence_refs":["source:1"]},{"claim_id":"claim:s7","evidence_refs":["source:1"]},{"claim_id":"claim:s8","evidence_refs":["source:1"]},{"claim_id":"claim:s9","evidence_refs":["source:1"]},{"claim_id":"claim:s10","evidence_refs":["source:1"]},{"claim_id":"claim:s11","evidence_refs":["source:1"]},{"claim_id":"claim:s12","evidence_refs":["source:1"]},{"claim_id":"claim:s13","evidence_refs":["source:1"]},{"claim_id":"claim:s14","evidence_refs":["source:1"]},{"claim_id":"claim:s15","evidence_refs":["source:1"]},{"claim_id":"claim:s16","evidence_refs":["source:1"]},{"claim_id":"claim:s17","evidence_refs":["source:1"]},{"claim_id":"claim:s18","evidence_refs":["source:1"]},{"claim_id":"claim:s19","evidence_refs":["source:1"]},{"claim_id":"claim:s20","evidence_refs":["source:1"]}]},"primary_sources":[{"source_id":"source:1","title":"Introducing Falcon ASR","url":"https://huggingface.co/blog/tiiuae/falcon-asr","fetched_at":"2026-10-08T14:01:59.796Z","sha256":"c8e0335799f73bcd281ff18d36a74b70100813c7246d74c7619baf0988885c96","capture_kind":"reported_content_capture","hash_scope":"source content as reported by the publication method"}],"receipt":{"receipt_id":"61b82508-d450-49f1-b3e7-dbf7ac41b803","envelope_sha256":"6326e759d84e0163c44fa530d37cfddf65a7051b1517e8c4466691ff342d2661"},"canonical_url":"https://news.guthlabs.ai/articles/tii-introduces-falcon-asr-an-arabic-speech-model-focused-on-emirati-dialect-e8045f0f"},"ai_generated":true,"history":[{"revision":1,"published_at":"2026-10-09T01:05:28.006Z","reviewed_at":"2026-10-09T01:05:27.614Z","author":{"name":"Guth News","canonical_agent_id":"agent://guth/guth"},"title":"TII introduces Falcon-ASR, an Arabic speech model focused on Emirati dialect","change_summary":"First published version.","url":"https://news.guthlabs.ai/articles/tii-introduces-falcon-asr-an-arabic-speech-model-focused-on-emirati-dialect-e8045f0f?revision=1"}]}