EveryLanguageMatters

AcoliACH AfarAAR AfrikaansAFR AkanAKA BambaraBAM BasaaBAS BembaBEM DinkaDIN DioulaDYU EfikEFI EnglishENG EsanESN EweEWE FonFON FulfideFUL GaGAA GagnoabeteGNB GandaLUG HausaHAU IgboIBO IsokoISO JulaJOD KabuverdianuKEA KalangaKLQ KambaKAM KaondeKQN KikuyuKIK KimbunduKMB KinyarwandaKIN KongoKON KrioKRI LambaLAM LingalaLIN LoziLOZ LubakatangaLUB LubaluluaLUA LundaLUN LuoLUO LuvaleLUE AcoliACH AfarAAR AfrikaansAFR AkanAKA BambaraBAM BasaaBAS BembaBEM DinkaDIN DioulaDYU EfikEFI EnglishENG EsanESN EweEWE FonFON FulfideFUL GaGAA GagnoabeteGNB GandaLUG HausaHAU IgboIBO IsokoISO JulaJOD KabuverdianuKEA KalangaKLQ KambaKAM KaondeKQN KikuyuKIK KimbunduKMB KinyarwandaKIN KongoKON KrioKRI LambaLAM LingalaLIN LoziLOZ LubakatangaLUB LubaluluaLUA LundaLUN LuoLUO LuvaleLUE
MakhumaMAK MalagasyMLG MendeMEN NamaNAQ NandeNNE NdebeleNDE NdongaNDO NorthernsothoNSO NuerNUS NyanjaNYA NyankoleNYN PidginPCM PularFUF RundiRUN SangoSAG ShonaSNA SidamoSID SomaliSOM SouthernsothoSOT SouthndebeleNBL SukumaSUK SwahiliSWA SwatiSSW TigrinyaTIR TivTIV TongaTOI TsongaTSO TswanaTSN TumbukaTUM TwiTWI UmbunduUMB UrhoboURH VendaVEN WolayttaWAL WolofWOL XhosaXHO YorubaYOR ZuluZUL MakhumaMAK MalagasyMLG MendeMEN NamaNAQ NandeNNE NdebeleNDE NdongaNDO NorthernsothoNSO NuerNUS NyanjaNYA NyankoleNYN PidginPCM PularFUF RundiRUN SangoSAG ShonaSNA SidamoSID SomaliSOM SouthernsothoSOT SouthndebeleNBL SukumaSUK SwahiliSWA SwatiSSW TigrinyaTIR TivTIV TongaTOI TsongaTSO TswanaTSN TumbukaTUM TwiTWI UmbunduUMB UrhoboURH VendaVEN WolayttaWAL WolofWOL XhosaXHO YorubaYOR ZuluZUL

Our mission is to build foundational multilingual AI infrastructure serving every community around the world.

Creating a scalable multilingual data foundation to power AI across all languages. This ensures every community, in every region and language, can benefit from accessible and inclusive technology.

Multilingual Annotation Studio

Where the missing data gets made.

A model drafts, a speaker corrects, a second speaker clears it. Communities hold their own projects and decide what is fit to release.

Supporting underserved languages to become AI-ready

EveryLanguageMatters builds the data foundation that lets communities across Africa and the Global South own, correct and govern the language data their models are trained on.

How the studio works

Nothing reaches a dataset on one person's say-so.

Most annotation tools hand one item to one person and lock everyone else out. Here a prompt has a lane per language, and people work them at the same time.

Two-stage community review

one speaker corrects, another clears it or returns it with a reason

Parallel language lanes

a prompt is unfinished until every target language has something in it

Phrase-level judgement, kept

highlight a span, record what it should be, and the reasoning travels with the data

Releases that cannot overstate

review state is derived from the rows, never chosen at export

What a project collects

One workspace, whatever the task is.

A project declares its objective when it opens. The interface stays the same either way — a source, one or more outputs, and the judgements people make about them.

Translation

Parallel text in any direction, including pairs that never pass through English.

Sequence to sequence

Rewriting, simplification, structured extraction, instruction following.

Sentiment & classification

Labelled by speakers, not translated from an English-labelled set — which is how idiom gets lost.

Summarisation

Reduced in the same language, with reviewers checking for the overstatement it invites.

Question answering & NER

Annotated against conventions the community agrees for its own names and places.

Response & RAG evaluation

Preference between two answers, and whether a retrieved passage actually supports one.

Text, speech and image projects share one review workflow — learn it once.

Data & AI APIs

Models trained on reviewed data, behind a key.

One integration rather than a research project, so a clinic, a ministry or a two-person team can serve people in their own language without rebuilding any of this.

POST /api/v1/translate
curl https://www.everylanguagematters.com/api/v1/generate \
  -H "Authorization: Bearer $ELM_KEY" \
  -H "Content-Type: application/json" \
  -d '{"task": "translate",
       "text": "Take one tablet after meals.",
       "source": "eng",
       "targets": ["bem", "swa", "hau"]}'

Translation, any direction

with the measured quality of every language published openly

Sentiment & classification

trained on judgements from speakers rather than a translated label set

Sequence to sequence

one prompt into one or many languages in a single call

Scoped keys, free tier

usage returned with every response, not from a second endpoint

77 languages and growing

Every language on this list has somebody behind it.

A speaker who decided it was worth the trouble, a community that took it on, or a dataset that was built and released. More are added as evaluation sets are built for them — a language is listed when it can be measured, not before.

UG Acoli ACH ET Afar AAR ZA Afrikaans AFR GH Akan AKA ML Bambara BAM CM Basaa BAS ZM Bemba BEM SS Dinka DIN CI Dioula DYU NG Efik EFI -- English ENG NG Esan ESN GH Ewe EWE BJ Fon FON WA Fulfide FUL GH Ga GAA CI Gagnoabete GNB UG Ganda LUG WA Hausa HAU NG Igbo IBO NG Isoko ISO CI Jula JOD CV Kabuverdianu KEA ZW Kalanga KLQ KE Kamba KAM ZM Kaonde KQN KE Kikuyu KIK AO Kimbundu KMB RW Kinyarwanda KIN CD Kongo KON SL Krio KRI ZM Lamba LAM CD Lingala LIN ZM Lozi LOZ CD Lubakatanga LUB CD Lubalulua LUA ZM Lunda LUN KE Luo LUO ZM Luvale LUE ZM Makhuma MAK MG Malagasy MLG SL Mende MEN NA Nama NAQ CD Nande NNE ZW Ndebele NDE NA Ndonga NDO ZA Northernsotho NSO SS Nuer NUS ZM Nyanja NYA UG Nyankole NYN NG Pidgin PCM GN Pular FUF BI Rundi RUN CF Sango SAG ZW Shona SNA ET Sidamo SID SO Somali SOM ZA Southernsotho SOT ZA Southndebele NBL TZ Sukuma SUK EA Swahili SWA ZA Swati SSW ET Tigrinya TIR NG Tiv TIV ZM Tonga TOI ZA Tsonga TSO ZA Tswana TSN ZM Tumbuka TUM GH Twi TWI AO Umbundu UMB NG Urhobo URH ZA Venda VEN ET Wolaytta WAL SN Wolof WOL ZA Xhosa XHO NG Yoruba YOR ZA Zulu ZUL

Not on the list yet — tell us what to add next.

Showing of 77 languages

Get started with the Multilingual Annotation Studio.

Create a free account, open a project in your language, and start building the training data that frontier models are missing. No setup, no minimum commitment — a single corrected sentence is already a contribution.

EveryLanguageMatters — foundational multilingual AI infrastructure for every community, in every region and language.