Run the scored Alphabet Test and the L26 Lite suite against the foundational layer of Bantu — closed-book, then tool-assisted. Deterministic scoring, calibration, and a saved history you can re-test against. You get the score, never the inventory.
Confirm your email and you're in — instant for AI-lab domains. No procurement, no data handling, nothing for legal to review.
Every test runs twice — first with no tools at all, then with everything your model has. The gap is the tool-lift, and it separates a model that knows a language from one that merely found a document. Scoring is deterministic, the English control proves the harness is sound, and because the answer key is never returned the benchmark cannot be contaminated. Read the full case, with results, on FSI →
You get the scored tests on your assigned languages — not the inventories, the audio, or the answer key. Those begin at the Validation Pilot, where you take languages you choose and everything held for them, measured on your own held-out data. The evaluation is how you decide whether that is worth your time.
| PublicNo account | EvaluationFree · self-serve | Validation Pilot75 days | Full Annual SubscriptionTalk to us | |
|---|---|---|---|---|
| Alphabet Test | Unscored demo | ✓ Scored + saved | ✓ | ✓ |
| L26 Lite suite | Scored, not saved | ✓ Saved + board | ✓ | ✓ |
| Tool-lift + calibration | — | ✓ | ✓ | ✓ |
| Retest + model comparison | — | ✓ | ✓ | ✓ |
| Team seats + sharing | — | ✓ | ✓ | ✓ |
| Syllable inventories | — | — | ✓ Your languages | ✓ All released |
| Consented audio | — | — | ✓ | ✓ Full program |
| Curation + provenance | — | — | ✓ | ✓ |
| Bulk export | — | — | — | ✓ |
| Domains | FSI demo | FSI | Every domain covered for your languages | Every product line |
Yes. No card, no procurement, no trial clock. Evaluation exists so you can get a real number before anyone asks you for a decision.
Confirm your email and you're in — instantly for recognised AI-lab domains. Other organisations are reviewed briefly before activation.
The assigned evaluation languages, currently Bemba + Chewa. They're set by BantuNomics and always shown live in your workspace.
No — and that's deliberate. You get rates, grades and failure classes, never the syllables. It keeps the benchmark uncontaminated and keeps your pipeline clean.
A suite, not one test: English and Pinyin as controls plus the Bantu tracks, each run closed-book then tool-assisted. You complete every track to earn a result.
Nothing automatic. If the gap matters to you, the Validation Pilot is where the data begins — and the pilot fee is fully creditable toward a subscription.
Bantu is the largest language family in Africa — 500+ languages, 400M+ speakers. BantuNomics has released complete syllable inventories for 459 of them. Right now, nobody has measured a frontier model against that. The first lab to do it is the first that can say anything credible about it.
Already have a key? Log in — we'll take you to your workspace.