Vendor comparison

How Elbi Compares to Other AI Assistants

There is no shortage of AI companies selling into the Kingdom. There is a shortage of clarity about what they actually do — because most of them do several different things, most published capability lists overlap, and almost nobody publishes how they measure any of it. This page is an attempt to describe the landscape honestly, including where Elbi is not the answer.

Capability coverageVendor categoriesEvaluation questionsWhere we do not fit

There is no shortage of AI companies selling into the Kingdom. There is a shortage of clarity about what they actually do — because most of them do several different things, most published capability lists overlap, and almost nobody publishes how they measure any of it. This page is an attempt to describe the landscape honestly, including where Elbi is not the answer.

The market is not one market

The first thing to understand is that "AI company" covers at least five distinct businesses, and comparing across them produces nonsense. A national compute and model programme, a financial-crime platform, an autonomous software engineer, a business-intelligence consultancy and a conversational assistant are not competitors. They appear on the same lists because the same word is used for all of them.

When you are evaluating, the useful question is not "who are the AI vendors" but "who does the specific thing I need". For conversational assistants — an interface that answers questions in natural language, in Arabic and English, grounded in your own material — the genuine field is considerably smaller than the market maps suggest.

The second thing to understand is that capability lists are cheap. Almost every vendor publishes a similar-looking set of features, because the features are table stakes and the differences are in execution, deployment model and evidence. That is why the sections below are organised around the questions that separate vendors rather than around feature checklists.

The categories, and what each is actually for

National programmes

Organisations positioned around sovereign compute, national models and strategic capability. They exist to build infrastructure and capability at national scale. If you need a customer-service assistant next quarter, this is not the category you are shopping in.

Conversational assistants

Products whose purpose is answering questions in natural language against an organisation's own material. This is the category Elbi is in, and the one this page is genuinely about.

Vertical AI platforms

Deep capability in one domain — financial crime, fraud, risk, knowledge intelligence. Where your problem is that domain, a specialist will beat a generalist and should.

Decision intelligence

Platforms oriented toward analysis and decision support rather than customer interaction. Different buyer, different problem, frequently confused with conversational AI because both involve asking questions.

Autonomous engineering

Systems that write and modify software. Genuinely impressive and entirely unrelated to answering a customer asking where their order is.

Systems integrators and BI consultancies

Firms that assemble solutions from other people's technology. Often the right choice for complex programmes, and worth recognising as a different commercial relationship from buying a product.

What actually separates vendors in this category

Once you narrow to conversational assistants, the published feature lists converge almost completely. Everyone supports Arabic. Everyone has a widget. Everyone claims integrations. The differences that matter are structural and are rarely on the comparison page.

The first is where processing happens. An assistant that sends conversation content to a third-party model provider is architecturally different from one that does not, and for government, banking and healthcare buyers this frequently decides the procurement before any other criterion is examined. It is worth asking directly and asking for it in writing.

The second is what happens to Arabic. Nearly every vendor claims Arabic support, and the claim is nearly always true in the narrow sense that Arabic text produces Arabic replies. The meaningful question is whether it handles dialect, mixed Arabic-English input, and a genuinely mirrored interface — and the way to find out is to test it with your own material rather than to read the specification.

The third is evidence. Most vendors in this market publish capability claims and very few publish measurement. Asking how accuracy was measured, against what test set, in which language, is a question that separates vendors quickly and is entirely reasonable to ask.

The eight kinds of assistant sold under one name

The word “assistant” covers at least eight different architectures in this market. They have different buyers, different budgets and different failure modes, and conflating them is the most common error in an evaluation. This is the structural map, independent of who sells what.

How it is builtHow it answersWhere it usually runsWhere it falls down
Authored dialog treeButtons and menus written by hand in advanceVendor cloud, sometimes on a customer subdomainCannot answer anything unscripted
Off-the-shelf live chatRoutes to human agents; little or no AIVendor cloudNeeds staffed humans; no after-hours cover
Enterprise CXM suiteOmnichannel platform with a bot and an agent deskVendor cloudCost and integration weight
Sovereign or national LLMA large Arabic-first generative modelNational or vendor cloudHallucination; heavy compute
Grounded retrieval assistantRetrieves from a private corpus, then generates an answer from itVaries — cloud, private tenancy or on-premiseRetrieval quality and language fit
Agentic task executionTakes multi-step actions across other systemsVendor cloudUnbounded blast radius when it is wrong
Custom in-house buildCommissioned and built bespokeYour own estateMaintenance burden falls on you
Accessibility overlaySign-language avatar or accessibility toolingVendor cloudNot an assistant at all — routinely miscounted as one

Elbi is the fifth row — a grounded retrieval assistant with generative answering — deployed the way the seventh row is deployed. That combination is the whole design: the answer quality of a generative system, with the data control of something you built yourself and run in your own building.

“On-premise” has five different meanings

Almost every vendor in this market will say yes to on-premise deployment. The word is doing very different work in each case, and the difference only becomes visible in a compliance review — usually after signature. These are the five shapes it takes, from least to most contained.

What is on offerWho operates itDoes conversation data leave your estate?Does answering depend on an outside service?
Multi-tenant SaaS
Shared platform, vendor domain
The vendorYesUsually yes
Vendor-hosted on your subdomain
Looks like your estate in the address bar
The vendorYes — the subdomain is presentation, not locationUsually yes
Private cloud tenancy
Dedicated instance in a cloud region
The vendor, or youLeaves your building; may leave the countryOften yes
On-premise, with external model calls
Your servers, someone else’s model
YouThe prompt leaves on every single answerYes — architecturally, not by configuration
On-premise and self-contained
This is what Elbi is
YouNoNo

The fourth row is the one worth being careful about, because it satisfies a procurement checkbox that says “on-premise” while still sending the content of every question to a third party to be answered. Whether a system calls an outside model is an architectural property. It is decided when the system is designed, and it cannot be switched off afterwards.

The question that separates the fifth row from the other four takes one sentence to ask: if the building’s internet connection is cut, does the assistant still answer?

What runs where, in our own system

The claim above is only meaningful if it is specific, so here is the whole answering path, component by component. Every part that touches the content of a customer’s question runs on hardware inside the Kingdom that the deployment controls. None of it calls an outside service to produce an answer.

ComponentWhat it doesWhere it runsOutside calls to answer a question
Language modelWrites the answerHardware you control None
Retrieval indexFinds the passages the answer must come fromHardware you control None
Speech recognitionTurns speech into textHardware you control None
Speech synthesisSpeaks the answer, Arabic includedHardware you control None
Document OCRReads uploaded files, including scansHardware you control None
Conversation storeHolds messages until they are destroyedHardware you control None
Backup replicaEncrypted off-site copyA second host inside the Kingdom None

Two honest notes on scope. This table describes the assistant, not this marketing website — the site you are reading loads analytics from a third party, which is disclosed in full in our privacy policy and sub-processor register. And “no outside calls” is a statement about answering, not about your own integrations: if you connect the assistant to your CRM or your ticketing system, it will talk to those, because you told it to.

Questions worth asking every vendor, including us

Where is conversation data processed?Ask for a specific answer, in writing. "In the cloud" and "in the region" are different from "inside the Kingdom", and for regulated buyers the distinction is usually decisive.
Does it call an external model provider?This determines whether your customer conversations leave your control. It is an architectural property, not a configuration setting, and it cannot be turned off later if the answer is yes.
How was accuracy measured?Against what test set, in which language, by whom. A percentage without a described method is marketing. This is the single most revealing question you can ask.
Show me it failing.Ask something adjacent to the demo material but not in it. A good system declines; a weak one improvises fluently. Vendors rarely volunteer this and it takes ten seconds.
Test it in dialect.Not formal Arabic. Type the way your customers type, including a sentence mixing Arabic and English. This is where bilingual claims are separated from bilingual products.
What is the escalation path?How a customer reaches a person, how fast, and what that person sees. Ask to see the agent view rather than a description of it.
Who owns the deployment relationship?A product, a platform plus an integrator, or a consultancy engagement. These are different commercial relationships with different costs and different ongoing dependencies.
What happens to our data if we leave?Export, deletion and the practical mechanics of it. Worth settling before signature rather than during an exit.

Where Elbi is genuinely a good fit

Elbi is a conversational assistant for organisations that need bilingual customer or employee interaction, that care where their data is processed, and that want answers grounded in their own material rather than generated from a general model's training.

That profile fits government bodies, banks, healthcare providers, telecoms, universities and large enterprises operating in Arabic and English — organisations for whom "where does the data go" is a procurement question rather than a technical curiosity, and for whom an assistant that confidently invents an answer is a genuine risk rather than an amusing anecdote.

It also fits organisations that have accumulated a large amount of contact volume around a small number of repeated questions, and that would rather remove that volume than staff it. That description covers a remarkable proportion of the organisations we speak to, in every sector.

Where it is not

Elbi is not a national AI programme, and organisations looking for sovereign compute capacity, foundation-model development or national-scale strategic capability are looking at a different category entirely.

It is not a specialist financial-crime, fraud or risk platform. Where the problem is detecting money laundering or scoring transaction risk, a vertical specialist with years of domain modelling will outperform a conversational assistant, and it should.

It is not a business-intelligence or decision-support platform. Elbi answers questions about specific records and published material; it is not the right instrument for analytical work across large datasets, and asking it to be produces a poor version of a different product.

It is not an autonomous software engineering system. And it is not the bottom of the market — an assistant that processes everything inside the Kingdom on infrastructure you control costs more than one that forwards conversations to a shared external service. Our prices are published rather than quoted on request, so you can decide that before a call rather than after one.

Where the other categories are ahead of us

A comparison page that finds its author superior on every axis is not a comparison, it is an advertisement. These are the places where a cloud platform or an established suite will genuinely serve you better than we will today. If one of them is central to your requirement, you should weight it heavily — and we would rather you learned it here than three weeks into a pilot.

WhatsApp and other channelsThe assistant runs on your website. It is not an omnichannel product, and if WhatsApp is the channel your customers actually use, a suite built around messaging will get you there faster.
Satisfaction surveys in the widgetNo built-in NPS or CSAT step. Transcripts can be exported and surveyed separately, which is more work than a checkbox.
Emailed transcriptsNot built. A customer cannot ask for a copy of the conversation by email at the end of it.
Many brands on one deploymentOne organisation per deployment today. A group running several brands from one contract is better served elsewhere for now.
Automatic failoverSingle-node. Maintenance is a planned window rather than an invisible event, which is a real operational difference from a large cloud platform.
Very high simultaneous volumeSized for departmental and pilot workloads. Voice in particular is answered one caller at a time. It is not a national call centre replacement, and we will tell you the current ceiling in writing rather than describe it as unlimited.
Certification badgesWe hold no ISO 27001, no SOC 2 and no independent penetration test. The engineering underneath is documented and checkable, but the badge is a procurement asset we do not yet have. This is stated in full on our security page.

How to read vendor claims in this market

Logo walls are not deployments

A logo on a website can mean a signed contract, a pilot, a partnership, an investor relationship or a conversation. Ask which, and ask for a reference you can speak to.

Certifications are scoped

ISO 27001 and SOC 2 cover a defined scope. Ask what the scope was and when the audit was performed — both are on the certificate.

"Compliant" is not a certification

PDPL alignment is a claim about design. Ask what it means concretely: where data is processed, what the retention is, how erasure works.

User counts are not adoption

"80 million end users" usually counts people reached, not organisations served. Both numbers are interesting; they answer different questions.

Case studies are selected

Anonymised percentage improvements are real but chosen. Ask what the baseline was and how it was measured.

Partnerships are not endorsements

Being an NVIDIA Inception member or a Meta Business Partner is a programme membership. It says something about the company; it says nothing about the product's fit for you.

Capability coverage, counted

Each bar counts how many of the eight capabilities above that category meets. It is a count of published capability — not a benchmark. We have not run anyone else's system, so no accuracy, latency or cost figure for another vendor appears anywhere on this page.

Elbi
8
of 8
Top cloud chatbot platforms
2
of 8
On-premise chatbot builds
5
of 8

Where the gaps actually fall

The same eight rows, read down instead of across. The pattern matters more than the total: the cloud platforms are strong on conversation and weak on where the data goes, and the on-premise builds are the reverse of that only in part.

ElbiTop cloud chatbot platformsOn-premise chatbot builds
Arabic and English, native both waysMeets itMeets itMeets it
Live-agent handover with a full portalMeets itMeets itMeets it
On-premise / air-gapped deploymentMeets itDoes notMeets it
Zero external AI callsMeets itDoes notMeets it
Speech in and out on your own hardwareMeets itDoes notDoes not
Document understanding with Arabic OCRMeets itDoes notDoes not
Data never leaves the KingdomMeets itDoes notMeets it
Message content destroyed within 24 hoursMeets itDoes notDoes not

What each category is optimised for

No category is bad at everything. They are built to different briefs, and the brief is what you are really choosing between.

Speed to first deployment
Elbi
Top cloud chatbot platforms
On-premise chatbot builds
Control over where data sits
Elbi
Top cloud chatbot platforms
On-premise chatbot builds
Arabic handled as a first language
Elbi
Top cloud chatbot platforms
On-premise chatbot builds
Breadth without a build team
Elbi
Top cloud chatbot platforms
On-premise chatbot builds
Answerable in a compliance review
Elbi
Top cloud chatbot platforms
On-premise chatbot builds

Positions on each axis reflect what each category is designed to do, not measured scores.

A practical evaluation, in order

  1. Decide which category you are buying inConversational assistant, vertical platform, decision intelligence, integrator engagement. Most wasted evaluation time is spent comparing across categories that do not compete.
  2. Settle the data question firstWhere conversation data is processed, and whether an external model provider is involved. For regulated buyers this eliminates most of the field before anything else is examined, and doing it first saves months.
  3. Test with your own materialNot the vendor's demo content. Your policies, your product names, your customers' phrasing. Demos are built to succeed; your material is what the system will actually face.
  4. Test in the language your customers useDialect, mixed Arabic-English, informal phrasing, phone typing with autocorrect. This is where the field separates fastest.
  5. Ask it something it should not knowAdjacent to your material but not in it. Watch whether it declines or improvises. Ten seconds, and more informative than an hour of slides.
  6. Look at the escalation pathAsk to see what the human agent sees when a conversation is handed over. If it is a notification rather than the conversation, handover is decorative.
  7. Ask for the measurement methodHow accuracy was assessed, against what, in which language. Vendors who measure will tell you; vendors who do not will change the subject.
The most useful thing you can do in an evaluation costs ten seconds: ask the assistant something adjacent to its material but not in it. A system built to decline will decline. A system built to impress will improvise fluently — and it will improvise most convincingly on exactly the questions where being wrong costs you most.

A note on how this page is written

Everything above describes categories rather than ranking individual companies, and that is deliberate. We hold research on the vendors operating in this market — fifteen profiled firms across the Kingdom, the Gulf and internationally — and it is genuinely useful for understanding the landscape. It is not a sound basis for public claims about what another company can or cannot do.

The reason is simple: capability changes, published material lags reality in both directions, and a comparison table asserting that a named competitor lacks a feature is a claim we would have to substantiate on the day someone challenges it. Vendors who publish those tables are usually describing a competitor as they were eighteen months ago.

What we will do is answer specific questions about how we compare on a specific requirement, with the reasoning, in a conversation. That is more useful to a buyer than a matrix of ticks, and considerably more honest.

Common questions

It depends entirely on what you are buying. For bilingual conversational assistants in the Kingdom there is a genuine field of regional and international vendors. For national AI capability or vertical risk platforms we are not in the same category, and saying otherwise would waste your time.

Because we would have to defend every cell of it on the day someone challenges it, and capability changes faster than tables do. We will answer specific comparison questions directly, which is more useful and more honest than a matrix.

Structurally: conversation processing inside the Kingdom with no external model provider involved, answers grounded in your own material with a check before sending, and Arabic handled natively rather than through translation. Those are architectural properties rather than features, which is why they are hard to add later.

It depends entirely on what you are comparing us to, and the figures are published rather than held back — see the pricing table in our article on what an AI chatbot costs in Saudi Arabia. Against a cloud chatbot subscription we are several times the price, and for organisations whose requirements a cloud platform meets, it is the right purchase. Against a bespoke build from a systems integrator we are frequently below it, because we are not rebuilding the assistant each time. What we are never competing on is the bottom of the market: processing inside the Kingdom on infrastructure you control costs more than forwarding conversations to a shared external service, and if price is the only criterion that difference will not favour us.

Then you have a baseline, which is useful. The questions worth asking are what share of contact it actually resolves, what it does when it does not know, and where its conversation data is processed. Those three answers usually determine whether replacing it is worthwhile.

Frequently, yes. Large organisations often run a general assistant alongside a specialist one for a particular service line, and that is a reasonable architecture rather than a compromise.

Same material, same questions, same languages, for every vendor — including questions designed to fail. Most evaluations are decided by whichever vendor demoed most impressively, which measures sales skill rather than product fit.

See it on your own content

A working assistant on your own material, in both languages.