Multilingual evaluation
Testing in Arabic, its dialects and French, with code-switching and institutional terminology.
MX4 AI is an AI startup focused on the MENA region. We build and openly study AI that understands Arabic, its dialects and French, shows the evidence behind its answers, and is designed to stay under the control of the organization that deploys it.
Open research · Question 01
Our first question
The question is what changes when only the language of the question changes. Moroccan Darija is the first dialect we study.
Illustrative values, not results
Each dot is a passage, placed by meaning in a 2D projection: closer means more similar.
When we publish, we publish the data, code, results and limits together.
This is a research question, not a published result.
Engineering perspectives, not experimental studies.
Technology
Each dot is a passage, placed by meaning in a 2D projection: closer means more similar. The dashed ring marks the 0.70 threshold.
Testing in Arabic, its dialects and French, with code-switching and institutional terminology.
Answers tied to their supporting passages, or an explicit “not in the document”.
Open-weight models designed to run in environments the institution controls.
Language data with documented provenance and usage rights.
Use cases · Illustration
The AI prepares a sourced answer; your teams choose the sources and verify it before use.
The AI flags discrepancies with references; a reviewer decides if the file is complete.
The AI drafts; an authorized person checks and approves anything sent.
Illustrative scenarios, fictional content, not customer case studies.
Explore the use casesWork with MX4
Consulting and engineering, in four areas.
An AI technology startup focused on the MENA region, studying and building sovereign AI for Arabic, its dialects and French, grounded in evidence and designed to stay under the deploying organization’s control.
When we publish, we publish the data, code, results and limits together. Our notes are engineering perspectives, not experimental studies; questions we study are not presented as results.
No. They are illustrations: a fictional corpus and document, prewritten questions and answers. No AI model runs on this website.
Yes, selectively, where our research applies. Describe your context through the contact form.
Systems can be designed for on-premises, private-cloud, hybrid or isolated environments, depending on the institution’s security, connectivity, data-residency and operational requirements.
As a whole-system question, not an interface setting: Modern Standard Arabic, its dialects, French and code-switching affect retrieval, terminology, citations and when to refuse or defer to a person. Each needs its own evaluation, with people who speak it.
Contact
Write to us about our research questions, or a project where they matter.