AI chatbot development · UAE

AI chatbots that answer in Arabic and English.

Customers here write in Gulf dialect, in Modern Standard Arabic, and in Arabic typed in Latin letters. Webzenia settles which variety the bot answers in before the build starts.

Tell us what it would have to answer

The three decisions

What AI chatbot development decides.

The model is the part that is bought. What it may read, which Arabic it writes back in, and what it says when it does not know are decided.

Support assistantOnline · trained on your dataWhatsAppWebInstagramTlabt ams, mata yousalel talab?Talab #4821 fee el tareeq,yousal bukra el asr.from your order dataShukran! Agdar aghayyir el unwan?Resolved without a human68%this month→ human when unsureTrained on your data. Arabic and English. PDPL-ready.

What it may read.

Every answer traceable to one document you own. An answer you disagree with is then corrected at the document rather than argued with, which is the difference between a bot you can run and one you have to supervise.

Which Arabic it replies in.

The varieties it reads and the one it replies in are two separate decisions, and only the second is visible to the customer. Getting the first right and the second wrong produces a correct answer nobody continues.

What it does when it cannot answer.

Most assistants are judged on this turn, not on the easy ones. A refusal that reaches a person is a good outcome; a fluent guess is the failure that ends the pilot and takes the budget with it.

The reply register

Formal Arabic, or the Arabic customers write.

Both are correct Arabic. Only one of them reads as a person answering, and the difference decides whether a second message ever arrives.

Modern Standard ArabicOne register, every customerThe register they wrote inRead the variety, answer in it
How it readsLike a published notice, delivered in a chat window.Like the person at the counter, answering the question.
Latin-script ArabicRead as noise, or answered in a script the customer did not use.Recognised as Arabic and answered in the register it arrived in.
Mixed in one sentenceRead as two, and one half is quietly dropped.Read as one message and answered once.
When it is rightPublished policy, contract wording, anything government-facing.Support, sales, bookings, anything a person would say aloud.
Where it is decidedBy the model default, after launch, by nobody.In the scope document, before the build.
Which one answers your customers.

For anything conversational, the register the customer wrote in. A correct reply nobody continues is a ticket closed on paper and still open for the customer. Modern Standard Arabic is the right choice for published policy, contract wording and anything government-facing, and we say so when that is what the assistant is for. Either way the decision goes in the scope document, next to where an unclear message goes instead.

The UAE context

How customers judge the assistant.

Webzenia has worked with Gulf clients since 2018. Three things show up in almost every chat log we are handed, and none of them is a model problem.

  1. 01of 03
    Three Arabics, one threadHow the message actually arrives

    Customers mix Gulf dialect, formal Arabic and Latin letters.

    A Dubai Marina clinic on a mainland licence reads all three in one inbox. Research at MBZUAI in Abu Dhabi finds that pretraining on Modern Standard Arabic alone raises the error rate on ordinary dialect, and that the multilingual training which handles code-switching gives some of that back. The variety is a build decision.

    Our methodRead and write varieties agreed in writing before the first prompt.
  2. 02of 03
    The only Arabic deskWhy the refusal bar goes up

    For many companies this is their only Arabic channel.

    An Amazon.ae seller on a Meydan free zone licence has no Arabic-speaking second line behind it, so a wrong answer is the last thing the customer reads. That raises the bar on what it refuses, not on what it attempts, which is the opposite of how most bots are tuned.

    Our methodRefusal behaviour tested before the resolution rate is tuned.
  3. 03of 03
    The transcript is the recordPDPL, Federal Decree-Law 45 of 2021

    Chat transcripts are personal data your product creates.

    For a Downtown hotel group on a mainland licence, consent is captured in the widget before the first message is stored, and the transcript becomes a cross-border transfer the moment it reaches a hosted model. No adequacy list has been published, so the answer is a named processing location, not a lookup.

    Our methodConsent, retention and processing location written into the scope.

The scope

Inside a chatbot build.

Six deliverables, each with a named output. The Arabic decision is one of them, written down in the same breath as the English one.

Trained on your datayour contentWhat is your refund window?Refunds are accepted within 7 daysof delivery, on unused items.refund-policy.pdfFAQGrounded on 240 documentsno made-up answersanswers from your content, with the source cited

Answers with the source named.

Every answer traced to the document it came from, so an answer you disagree with is corrected at the source rather than argued with in a meeting.

RetrievalYour documents
Gulf Arabic & Englishauto-detectEnglishArabicGulf dialectMata yousal talabi?detected: Gulf ArabicTalabak yousal bukra el asr.Tabi rabt el tatabbu?replies in the language your customer actually types

The Arabic scope, written down.

Which varieties it reads, which one it writes back in, and what happens to a sentence that switches to English halfway. Agreed before the build, not after launch.

Gulf dialectLatin-script Arabic
Omnichannel deployone botOne build, every channelWebsite widgetWhatsAppInstagram DMFacebook MessengerIn-app chatthe same trained bot, live wherever your customers are

One build, on the channels you run.

The same assistant on the site and on WhatsApp, answering from one set of documents rather than from two configurations that drift apart.

WebsiteWhatsApp
Human handoffseamlessThis needs a human. Connecting you.handed to a human agentNNoor · SupportjoinedHi, I can see your order and thefull chat. Let me sort this out.Full history carried overno repeating themselvesthe bot knows when to bring in a person

Handover to a person, mid-thread.

The conversation, the register it was held in and what the assistant already tried, handed over in writing so the customer never starts again.

EscalationContext
Analytics & tuning83% resolvedResolved without a human83%Satisfaction4.5Top intentsOrder status90%Refund & return62%Product query48%Pricing30%watch what it handles, tune what it misses

What it answered, and what it refused.

The refusals are the useful half. Each one names a question your documents do not answer yet, which is the backlog the next month works down.

RefusalsTuning
PDPL & securityconsent-firstCall me at1234redactedPII redactionphone & email maskedConsent capturebefore data is storedProcessing locationnamed in the contractEncryptionin transit & at restbuilt to PDPL rules, with the processing location named

Transcripts, consent and retention.

Consent captured before the first message is stored, personal details redacted, a retention window agreed, and the processing location named in the contract.

PDPLConsent

Our stack

Built with retrieval and your own content.

Four layers decide how a chatbot behaves: what it retrieves, where that sits, which model reasons over it, and where it runs. Select one to see why.

LangChain
Why LangChain

LangChain composes retrieval, tools and memory into one application, which is what makes a grounded answer reproducible rather than a lucky prompt.

How we excel

We wire retrieval so every reply carries the document it came from, and we keep the trace of what was retrieved, so a wrong answer is debuggable rather than deniable.

RetrievalTraceabilitySources
LangChainCited
LlamaIndexCited
Vercel AI SDKPartial
Raw model APINone

How the build runs

From your content to a live assistant.

A chatbot is proved on the questions it gets wrong. The middle phase is a variety test, not a general check, because that is where a build in this market fails quietly.

01Weeks 1 to 2Scoped

Deciding what it may read.

We collect the documents each answer has to be traceable to, and mark the questions your team fields that no document covers. Those gaps are an output of this phase rather than a blocker to it: half of them are answered by writing one page, and the other half are the reason the assistant will refuse. The Arabic scope is agreed here, in the same session, and not left to the launch checklist.

  • Documentsingested and versioned
  • Gapslisted, not papered over
  • Arabicread and write varieties agreed
02Weeks 3 to 4Built

Testing how customers write.

The assistant is tested on real Gulf dialect, on Modern Standard Arabic, on Arabic typed in Latin letters and on sentences that switch to English halfway through. The refusal path is tested as hard as the answer path, because a fluent wrong answer is the one that costs the account. Nothing goes live on a test set written by the people who built it.

  • Varietiestested one at a time
  • Refusalstested as hard as the answers
  • Handoffproved on a real thread
03Month 2 onwardRunning

Closing the gaps it refused.

Every refusal names a question the documents do not answer yet, and that list is worked down month by month. The resolution rate then rises because coverage grew, which is the only way it should rise. A rate that climbs because the guardrails were loosened is the same number describing a worse assistant, and we report the two apart so the difference is visible.

  • Refusalsreviewed monthly
  • Coverageextended, not loosened
  • Reportingthe two reported apart

Reported on what it refused and why, so the resolution rate can only rise by covering more, never by guessing more.

Reported on refusals
RefusalsCoverageHandoffs

Our commitment

Four promises about the answers.

A chatbot fails in four ways: it invents, it traps, it answers in the wrong register, and it keeps a transcript nobody agreed to. We commit against all four.

  • It answers from your content or declines.

    Every answer is traceable to a document you own. Where nothing covers the question, the assistant says so and reaches a person. It does not fill the gap with a fluent sentence.

  • Always a route to a person.

    Any conversation can reach a human, carrying the thread and the register it was held in. No customer is left repeating themselves to a machine that has already run out of answers.

  • We write the Arabic scope down.

    Which varieties it reads and which one it answers in go in the scope document. Where an engagement does not fund proper Arabic, we write that down instead of shipping a translated bot and calling it bilingual.

  • Consent first, and we name where it runs.

    Consent is captured before the first message is stored, a retention window is agreed, and the contract names where transcripts are processed. Conversations with your customers are never used to train a model of ours.

Common questions

AI chatbots, answered.

Next step

Find out what it would have to read.

Send us the questions your team answers most and the documents the answers live in. We will tell you what an assistant can resolve, what it cannot, and which Arabic it should reply in.

Tell us what you need.

+971
Chat on WhatsApp