NoorGhateh: A Benchmark Dataset for Training and Evaluating Arabic Morphological Analysis Systems
This dataset provides a linguistically and morphologically annotated sample of 313 Arabic words drawn from a larger corpus of 223,690 words compiled from Sharaye al-Islam, a classical Arabic jurisprudential text. Each token includes segmentation, lemma, part-of-speech, and affix-level annotations th...
محفوظ في:
| المؤلفون الرئيسيون: | , |
|---|---|
| التنسيق: | Artigo |
| اللغة: | Inglês |
| منشور في: |
Ubiquity Press
2026-02-01
|
| سلاسل: | Journal of Open Humanities Data |
| الموضوعات: | |
| الوصول للمادة أونلاين: | https://account.openhumanitiesdata.metajnl.com/index.php/up-j-johd/article/view/409 |
| الوسوم: |
لا توجد وسوم, كن أول من يضع وسما على هذه التسجيلة!
|
