← Back to site
Research
What our speech data changes, measured.
Results from training open models on our datasets, with public test sets and reproducible protocols.
September 2026
Whisper tiny, 27% fewer errors on Beninese French: how we got there
Eight hours of recordings made in our studio in Benin were enough to remove a quarter of the errors of this 75 MB model.