Arabic Leaderboards: Introducing Arabic Instruction Following, Updating AraGen, and More
Related stories
The Open Arabic LLM Leaderboard 2
QIMMA قِمّة ⛰: A Quality-First Arabic LLM Leaderboard
Introducing Falcon-H1-Arabic: Pushing the Boundaries of Arabic Language AI with Hybrid Architecture
Introducing the Open Leaderboard for Hebrew LLMs!
Falcon-Arabic: A Breakthrough in Arabic Language Models
📚 3LM: A Benchmark for Arabic LLMs in STEM and Code
Open ASR Leaderboard: Trends and Insights with New Multilingual & Long-Form Tracks
The Open ASR Leaderboard Adds Its First Global South Language
Position: AI Leaderboards Are Underserving the Global South: A Case Study from India
The paper argues that AI leaderboards are ill-suited for the Global South because they lack independent governance, conflict‑of‑interest policies, and mechanisms for metric evolution. It highlights that high‑quality regional benchmarks already exist—IndicSUPERB, MILU, and LAHAJA for India; IrokoBench for Africa; AlGhafa for Arabic—but are excluded from global leaderboards due to institutional design. A consultation with 58 Indian AI practitioners shows a clear preference for formal governance and disclosure‑based conflict management, leading the authors to propose regional leaderboards with independent governance as the solution.
Mizan: A National Benchmark for Evaluating Large Language Models on Iraqi Arabic and the Iraqi Civic Context
arXiv:2609.13980v1 Announce Type: cross Abstract: Arabic large-language-model (LLM) evaluation has matured around Modern Standard Arabic (MSA): aggregated leaderboards such as the Open Arabic LLM Lea...
AI tools in Arab University English classrooms: Looking back and forward
arXiv:2607. 05403v1 Announce Type: cross Abstract: This paper aims to synthesize empirical research on AI tools used to support English as a second/foreign language (EL2) learners in Arab University classrooms (AUCs) between Jan 1st 2023 and Aug 31st 2025.