Nahw: A Comprehensive Benchmark of Arabic Grammar Understanding, Error Detection, Correction, and Explanation
Nahw is a comprehensive benchmark for evaluating Arabic grammar understanding in large language models. It includes multiple-choice grammar questions (Nahw-MCQ), annotated passages for grammatical error detection, correction, and explanation (Nahw-Passage), and synthetic training data.
The benchmark covers theoretical and practical grammar knowledge in Modern Standard Arabic, supporting research on grammatical reasoning, error analysis, and educational NLP applications.
For a detailed description of the dataset construction, task hierarchy, evaluation setup, and experimental results, please refer to our paper:
https://aclanthology.org/2026.eacl-long.296.pdf
Nahw: A Comprehensive Benchmark of Arabic Grammar Understanding, Error Detection, Correction, and Explanation
Hamdy Mubarak, Majd Hawasly, Abubakr Mohamed
Qatar Computing Research Institute (QCRI), HBKU, Qatar
@inproceedings{mubarak2026nahw,
title={Nahw: A Comprehensive Benchmark of Arabic Grammar Understanding, Error Detection, Correction, and Explanation},
author={Mubarak, Hamdy and Hawasly, Majd and Mohamed, Abubakr},
booktitle={Proceedings of the 19th Conference of the European Chapter of the Association for Computational Linguistics (Volume 1: Long Papers)},
pages={6310--6328},
year={2026}
}If you use this dataset, benchmark, or any part of this repository in your research, please cite the paper accordingly. Thank you!