Start Free Trial
← All posts
One name, many languages

Arabic-English work: which Arabic do you mean?

Modern Standard Arabic is nobody's dinner-table language, and the dialects differ enormously. What diglossia means for interpreters, for speech recognition, and for captions on a screen.

4+ major dialect groups 7 min read Published July 2026

"Arabic" on a job ticket is one word for something no other major interpreting language quite has: a formal standard that everyone learns and nobody speaks at home, layered over regional varieties different enough to fail each other's listening tests. The defining fact of Arabic-English work is diglossia — and most of the practical decisions, from interpreter matching to whether speech recognition will hold up, follow from taking it seriously. Here is the map, and what to do with it.

Diglossia: the two-Arabics problem

Modern Standard Arabic (MSA, *fuṣḥā*) is the language of news broadcasts, speeches, legal documents, and formal writing across the Arab world. It is learned at school, understood by educated speakers everywhere — and spoken natively by no one. Conversation happens in regional dialects (*ʿāmmiyya*), which diverge from MSA and from each other in vocabulary, pronunciation, and grammar.

Speakers do not flip a binary switch between the two; they slide along a continuum, mixing more MSA into formal moments and more dialect into personal ones — often within a single answer. For an interpreter this means a deposition can move from document-register MSA to pure dialect the moment the witness describes what actually happened. For a speech recognition system it means the training data question — *which* Arabic did it learn? — is the whole ballgame.

The dialect map, in one table

VarietyWhereWhat to know
Egyptian (Masri)EgyptThe most widely *understood* dialect, thanks to decades of Egyptian film and television exported across the region
LevantineSyria, Lebanon, Jordan, PalestineInternally close enough to work across; a large share of US refugee and diaspora interpreting demand
Gulf (Khaleeji)Saudi Arabia, UAE, Kuwait, Qatar, and neighborsBusiness and energy-sector work; heavy English mixing among professionals
Maghrebi (incl. Darija)Morocco, Algeria, TunisiaThe furthest from the rest — Middle Eastern speakers often struggle with it; strong French and Amazigh influence
MSAFormal settings everywhereSpeeches, documents, broadcast; understood by educated speakers, conversational for none

Intelligibility across dialects is real but asymmetric and partial — most Arabs understand Egyptian better than Egyptians understand Darija, because media exposure ran one way. The practical consequence: a certified interpreter fluent in Levantine Arabic can be genuinely lost with a Moroccan client, and "we have an Arabic interpreter" is not yet a match. Ask which Arabic — on both sides.

What diglossia means for speech recognition

Arabic ASR has improved substantially, but the gains are uneven, and the reasons are the ones above: historically, training text skewed toward MSA — the written register — while the speech that needs transcribing is dialect. The traps to check for before relying on any tool:

Dialect coverage varies by vendor and by dialect. Strong MSA and Egyptian performance says little about Darija. Test on the dialect you actually work with, using real recorded speech, not a demo.
Code-switching stresses detection. Maghrebi speech weaves in French; Gulf professional speech weaves in English. A system locked to one language mangles the embedded other.
Names and religious formulae are frequent — *in shāʾ Allāh*, honorific phrases, Quranic quotation in formal speech — and are rendered inconsistently by systems that have not seen them enough.
A clean-audio test overstates field performance. Phone lines and clinic rooms are the real conditions; test there.

The general craft of getting the most out of ASR — audio path, dictionary, testing — is covered in how to improve live caption accuracy; with Arabic, dialect selection sits on top of all of it.

Right-to-left captions are their own discipline

Arabic runs right to left, and live captions inherit every classic bidirectional-text problem at speed. Before an event, verify the display path end to end:

  • Embedded left-to-right runs. Latin names, email addresses, and numerals inside an Arabic line force the renderer to switch direction mid-line — the classic place punctuation and word order visually scramble.
  • Two numeral systems are in live use. Western digits (123) dominate in the Maghreb; Eastern Arabic numerals (١٢٣) are common in the Mashreq. Readers handle both, but consistency within one caption stream matters.
  • Alignment and truncation. A display built left-to-right may align Arabic text incorrectly or truncate the wrong end of a long line. Check with real sentences, not lorem ipsum.
Run one bilingual test sentence — an Arabic line containing a Latin name and a number — on the actual screen before the event. It exercises every failure mode above in five seconds.

Names, honorifics, transliteration

Arabic names carry structure English records flatten. A *kunya* (Abu or Umm plus a child's name — "father of," "mother of") can be the name a person actually goes by. Honorifics — *ustādh*, *duktōr*, *ḥājj*, *shaykh* — encode respect an interpreter must judge how to carry. And a single Arabic name maps to many Latin spellings: Muhammad, Mohammed, Mohamed, and Mohamad are one name, and in legal and immigration files the variants can scatter one person across several records. Where spelling has consequences, fix the transliteration once — in a shared document or a caption dictionary — and hold to it.

Frequently asked

Why is Arabic especially hard for interpreters and speech recognition?

Because of diglossia: Modern Standard Arabic is the formal written and broadcast register, while everyday speech happens in regional dialects — Egyptian, Levantine, Gulf, Maghrebi — that differ substantially from MSA and from each other. Interpreters must be matched to the speaker's dialect, not just to "Arabic," and speech recognition trained mostly on MSA-register data performs unevenly on dialect speech. The fix in both cases is the same: identify the actual variety first.

Does speech recognition work on Arabic dialects?

Unevenly, and it is improving. Performance on MSA and widely represented dialects like Egyptian is generally stronger than on varieties like Moroccan Darija, and code-switching with French or English adds difficulty. The only trustworthy answer for your use case is an empirical one: test the tool on real audio in the dialect you work with, under field conditions. Unicaption's free tier gives you 30 minutes every week with no credit card, which is enough for exactly that test.

How do live captions handle right-to-left Arabic text?

The captions themselves render right to left; the risks live at the seams — Latin names, emails, and numerals embedded in an Arabic line can visually scramble on displays that handle bidirectional text poorly, and alignment or truncation can break on screens built for left-to-right scripts. Before an event, test one Arabic sentence containing a Latin name and a number on the actual display; it exercises every common failure in seconds.

Can AI replace an Arabic interpreter?

Not in the settings that matter most. AI captioning is genuinely useful for formal, MSA-register content and as a verification layer for numbers and names, and Unicaption is built for that copilot role. But dialect-native interpretation — an asylum interview in Darija, a medical conversation in colloquial Levantine — requires a human who lives in the variety, reads register and culture, and can ask for clarification. The realistic future is interpreters working with a live transcript, not being replaced by one.

Bring a checkable record to Arabic-English work

Live captions and translation beside the meeting, with names fixed once and nothing stored. 30 free minutes every week.

Start Free Trial →