CommonLID Leaderboard
Results for the CommonLID and CommonLID-nano benchmarks. Headline metric: macro F1. Models are ranked by macro F1 within each tab; click a row to see per-language metrics.
🌐 Website • 📝 Blog post • 📄 Paper • 🆕 Add a model
CommonLID
Common Crawl's language identification benchmark, sampled from real-world web text and human-validated across hundreds of language varieties.
Reference • License: common-crawl-tou • Main score: macro_f1
commonlid — sorted by Macro F1
Click a row to load per-language metrics.
Language | F1 | Precision | Recall | FPR (%) | GT | Predictions | Correct |
|---|
Source: commoncrawl/commonlid-results @ HEAD.