CommonLID Leaderboard

Results for the CommonLID and CommonLID-nano benchmarks. Headline metric: macro F1. Models are ranked by macro F1 within each tab; click a row to see per-language metrics.

🌐 Website • 📝 Blog post • 📄 Paper • 🆕 Add a model

CommonLID

Common Crawl's language identification benchmark, sampled from real-world web text and human-validated across hundreds of language varieties.

Reference • License: common-crawl-tou • Main score: macro_f1

Scoring scope

commonlid — sorted by Macro F1

Click a row to load per-language metrics.