AfroLID
VerifiedA neural language identification toolkit that detects which of 517 African languages and varieties a text belongs to across 14 language families, reaching 97.41 macro-F1 after fine-tuning on SERENGETI. Developed by the UBC Deep Learning and NLP Lab and published at EMNLP 2022.
- Category
- AI Resources
- Pricing
- Free for research use
- Country
- 🌍 Pan-African
- Last verified
- 25 Aug 2026
{ "name": "AfroLID", "slug": "afrolid", "category": "AI", "country": "Pan-African", "docs_status": "LIVE", "licensing_required": "INSTITUTIONAL", "verified": true, "last_verified": "2026-08-25", "website": "https://huggingface.co/UBC-NLP/afrolid_1.5", "documentation_url": "https://huggingface.co/UBC-NLP/afrolid_1.5"}Verification history
- 25 Aug 2026 · live
- 22 Aug 2026 · live
- 19 Aug 2026 · live
- 16 Aug 2026 · live
- 13 Aug 2026 · live
- 10 Aug 2026 · live
- 7 Aug 2026 · live
- 4 Aug 2026 · live
Automated checks run every few days. See all recent status changes
Tags
Compare AfroLID
Side-by-side, verified specs against its closest nlp model alternatives.
Related in AI Resources
Cheetah
A massively multilingual natural language generation model supporting 517 African languages, outperforming baselines on five of seven AfroNLG tasks like summarization and translation. Developed by the UBC Deep Learning and NLP Lab and published at ACL 2024.
AfriBERTa
AfriBERTa is a multilingual masked language model (XLM-RoBERTa architecture, ~126M params) pretrained from scratch on 11 African languages including Amharic, Hausa, Igbo, Swahili, and Yoruba. Built by the Castorini lab (University of Waterloo) for text classification and Named Entity Recognition on low-resource African languages.
GhanaNLP ABENA
ABENA (A BERT Now in Akan) is a family of BERT, DistilBERT and RoBERTa language models for the Twi/Akan language covering both Asante and Akuapem dialects, released by the open-source GhanaNLP initiative. Distinct from GhanaNLP's Khaya translation product.