Skip to content

TII releases Falcon-ASR, a 1.6B Arabic speech model

Short answer

The Technology Innovation Institute released Falcon-ASR, a 1.6-billion-parameter speech recognition model built for Arabic with a focus on the Emirati dialect. It scored 20.92% word error rate across six Arabic benchmarks, beating the best published result of 23.17%, and led on an internal Emirati evaluation. It also transcribes English, French, Spanish and Portuguese.

What this means for operators

For a support or sales team handling calls in the UAE or wider Gulf, transcription accuracy on dialectal Arabic has been a persistent weak point in voice automation pipelines — standard ASR tuned for Modern Standard Arabic often mangles everyday Emirati speech, which breaks downstream QA scoring, CRM logging and agent-assist tools. Falcon-ASR's word-level timestamps and lower error rate on Emirati and Gulf dialects make it a candidate for teams building or rebuilding call-transcription, voice-ticketing or meeting-notes pipelines in that region; the model is available to test now via a Hugging Face demo, though API access and native applications are only planned, so production integration isn't ready yet.

The Technology Innovation Institute in Abu Dhabi has released Falcon-ASR, a 1.6-billion-parameter speech recognition model built primarily for Arabic, with particular attention to the Emirati dialect.

On the Open Universal Arabic ASR Leaderboard's six test sets, Falcon-ASR achieved an average word error rate of 20.92%, against a best published result of 23.17% in the snapshot TII used — a 2.25 percentage point improvement. On TII's internal Emirati evaluation, the model recorded 22.73% WER and 10.19% character error rate, the lowest among the systems compared, beating the next-best result (Qwen3-Omni) by 4.07 percentage points on WER.

The model was trained on Emirati, Modern Standard Arabic, other Gulf and Arabic dialects, and English, with training data that included background noise, overlapping speech, music, reverberation and telephony effects to reflect real-world recording conditions such as calls and meetings.

Beyond Arabic, Falcon-ASR transcribes English, French, Spanish and Portuguese using the same model weights, without needing a language flag specified in advance. On the Hugging Face Open ASR Leaderboard's seven public English test sets, it posted a mean WER of 5.74%. The model also supports word-level timestamps, linking each transcribed word to its position in the audio.

Falcon-ASR builds on TII's earlier Falcon3-Audio work. A demo is available on Hugging Face for testing with user-supplied recordings; API access and native applications are described as planned but not yet available.

Source: Hugging Face

Next step

Visibility Analyzer

An AI news item will not tell you how assistants see your own site. The Visibility Analyzer checks that on your live site: which answers cite you, which pages an engine cannot retrieve, and what to fix first. Free to run.

Run a free visibility audit

Free to run. No card.

Fee
Free
Length
One run, minutes

Free tier: two analyses a day, no card required.