BANTUNOMICSPlans
Access & pricing

Choose how you engage.

Four ways in — from a free public test to a full program subscription. Each is a real step, and you're never locked into climbing them in order.

Prove it free → measure it on your data → license the platform. The paid pilot is 100% creditable toward a subscription, so nothing you spend to evaluate is lost.

What every plan is measured in. BantuNomics is a suite of domain products — each its own body of work, built on published Bantu scholarship (dictionaries, grammars, and native fieldwork) and consented native audio. FSI is the foundational domain — everything decomposes back into it — but it is one product among several. Plans differ by how many languages and which domains you license.
FSI — the operating alphabet: every legal syllable of a language, verified. The foundation every other domain builds on.
Tone — the meaning carried by pitch and vowel length, in native audio.
Nouns — the noun-class system, aligned one-to-one across the family.
Verbs — the generative verb: one root unfolds into its whole paradigm.
Numbers — the numeral system, spoken in clean Bantu and code-switch.
Health — a standardized set of clinical concepts, consented and voiced.
Grammar — the concord system: how one noun class governs a whole sentence, certified.
Stories — connected speech, in three aligned modes.
Each is a full product in its own right — see the catalogue for its scope. Most carry consented native audio in two modes: clean Bantu, and the Bantu-and-English code-switch a speaker actually uses.
Public
Anyone — no account needed
FreeOpen access, forever

See the whole thesis and verify the gap yourself before you talk to anyone.

  • Includes
  • The public Alphabet Test — give any AI our task
  • The L26 benchmark leaderboard
  • Coverage counts + the essays
Evaluation
AI labs · self-serve
FreeSelf-serve · renewable

Confirm the foundation is real on your own models, hands-on, before any commitment.

  • Everything in Public, plus
  • The scored Alphabet Test on Bemba + Chewa, your assigned evaluation languages
  • L26 Lite — the six-track suite: English, Pinyin base + toned, Bemba, Kinyarwanda and Luvale, the last two free on top of your assigned pair
  • Every track run closed-book, then tool-assisted — so you see your model's tool-lift
  • Where it broke, by class of syllable — per-stratum diagnostics and a named failure mode, not just a number
  • Deterministic scoring — the same answer scores the same, every time
  • Graded against the live inventory, so your score tracks the real thing, not a stale snapshot
  • Test any model — web, API, or MCP. Fully headless: one key, and your agent runs the whole battery itself
  • Every run saved and scored — retest after training and show the improvement
  • Tell us what's missing — leave feedback on any run and help shape the benchmark
  • Share results with your team
  • Scores only — the inventories and audio begin at Pilot
Most teams start here
Validation Pilot
A scoped, paid proof on your data
Talk to us75 days · 100% creditable

The whole foundation across the family, plus full depth on the languages you choose — measured on your own held-out data, for one clean go/no-go.

  • Everything in Evaluation, plus
  • All 459 released FSIs — the complete operating alphabet for every released language, not only the ones you pick
  • Then the rest of the domains on your three chosen languages, as available:
    • Tone
    • Nouns
    • Verbs
    • Numbers
    • Health
    • Grammar
    • Stories
  • Plus the L26 Lite suite across ten Bantu languages
  • Consented native audio in both modes — clean Bantu and Bantu–English code-switch
  • Bantu-accented English — the same stories read by speakers of different Bantu first languages
  • The curation record — provenance, FSI revision history and the reproducibility kit
  • A pre-registered go/no-go — criteria locked at kickoff that depend only on what we control, never your stack
  • First scored result in days, not weeks
  • Private workspace, an embedded engineer and a native-linguist audit
  • Fully credited toward a subscription
Full Annual Subscription
A full commercial license to the living platform
Talk to usFull annual license

One subscription to the entire living ecosystem — not a dataset licensed once. Everything today, and everything added while you're subscribed.

  • Everything in Pilot, plus
  • All 459 released languages — and every domain we hold for each, not a chosen three
  • The full consented audio corpus, uncapped — both modes, clean Bantu and Bantu–English code-switch
  • Bulk export — HuggingFace and Parquet
  • The full machine surface — API and MCP across every product
  • Full curation provenance — the complete record behind every row
  • Everything added while you're subscribed — new languages and new products ship into the subscription as they release
  • Private data room + a say in the roadmap
Compare every plan · the same options, in detail

What each plan unlocks

The four options above, laid out product by product — so you can see exactly what opens up at each tier.

  PublicFree EvaluationFree · self-serve Validation Pilot75 days · creditable Full Annual SubscriptionTalk to us
Best forVerifying the gapScoring your models on the Alphabet TestProving value on your own dataBuilding on the whole platform
LanguagesCounts onlyBemba + Chewa (Alphabet Test · assigned)
5 languages (L26 Lite · scores only)
3 in full depth · all 459 FSIs · 10 on L26 LiteAll 459 released, growing
The domains — what each plan unlocks, product by product
FSI · operating alphabet · foundationalPublic test Alphabet Test · 2 langs + L26 Lite · 5 langs all 459 released all released
Tone · tonal homographsTeaser your langs
Nouns · aligned matrixTeaser your langs
Verbs · generative engineTeaser your langs
Numbers · numerals & voiceCalculator your langs
Health · clinical languageTeaser your langs
Grammar · concord matrixSample your langs
Stories · connected speechTeaser your langs
Consented native audioSample (per language) full, uncapped
Bulk export HuggingFace / Parquet
Private data roomPilot workspace engagement room
TermRenewable75 daysAnnual
PriceFreeFreeTalk to us*Talk to us

*The Validation Pilot fee is 100% creditable toward the Full Annual Subscription. Enterprise-only; access is provisioned by BantuNomics.

Grant & funder access · invite-only

For mission-aligned funders

A separate, non-commercial track for grant, philanthropic and impact capital. It funds the public-interest layer — paid native-speaker recording, non-overridable consent and provenance, open standards, and coverage chosen for health and education need rather than commercial demand. Commercial rights stay separately licensed.

Enquire about the Showroom
Good to know

Questions, answered

Do I have to evaluate before a pilot?

No. Each door is independent — start a pilot, or go straight to a subscription, without an evaluation first. An evaluation is welcome, not required.

Is the pilot fee refundable?

Better than refundable — it's 100% creditable. Convert to the annual subscription and the entire pilot fee comes straight off the first year.

What does "Evaluation is the foundation only" mean?

Evaluation opens the operating alphabet (FSI) for two languages so you can confirm the core is real. The other domains — Tone, Nouns, Verbs, and the rest — open at the Pilot step, on your chosen languages.

Which languages does an evaluation actually cover?

Two different things run on one account. The Alphabet Test is scored on your assigned evaluation languages, currently Bemba + Chewa. L26 Lite is a six-track suite covering five languages — English, Mandarin Pinyin (base and toned), Bemba, Kinyarwanda and Luvale. Kinyarwanda and Luvale come free, over and above your assigned pair. The two grade different target sets, so read their scores separately.

Is L26 Lite the same as the L26 benchmark on the board?

Same tracks, lighter run. L26 Lite is the instrument you point at your own model: a single pass per track, closed-book then tool-assisted, free and headless. The full L26 behind l26.ai is the measurement we run across the field — each language evaluated three to five times to assess repeatability, blind and scaffolded. Lite is single-pass and does not measure run-to-run variance; for a repeatability figure, read the board or talk to us. Cite results as L26 v1.0 — the benchmark is young and still iterating.

Does the subscription mean "every language, complete"?

No — and we're candid about it. It's a subscription to a living platform that keeps growing wider (more languages) and deeper (more per language). You get everything that exists today and everything added while you're subscribed.

Who is the Showroom for?

Mission-aligned funders — grant, philanthropic, and impact capital — not AI labs. It's an invite-only, non-commercial view of the work.

How do I know it's real?

Check it yourself in a minute: l26.ai shows today's frontier models acing English and failing the same alphabet test on a Bantu language.

One thing worth understanding about the pricing: BantuNomics is a living platform, not a finished dataset — it grows in two directions at once: wider, as we add more of the 500-plus Bantu languages, and deeper, as each language gains more words, recordings, pronunciation, and grammar. That's why the top tier is a subscription: you're licensing something that never stops improving, not buying a file that goes stale.