Evidence

Check the work yourself

Our claims are tied to a specific benchmark and a published method. Where a result is benchmark-specific, we say so.

Nigerian register benchmark

Gemini Flash78.6
GPT-567.8
Gemini Pro65.4
Calibrated human analysts100.0

MIS™ scores on our Nigerian register/discourse benchmark using a contamination-aware holdout. Human performance demonstrated on this benchmark only; these figures are not a general model ranking.