These projects are all structured linguistic datasets and tools for the computational processing and analysis of various Armenian language varieties.
The PROIEL Treebank is a linguistic dataset containing dependency treebank annotations for texts in ancient Indo-Euro...
A Universal Dependencies treebank for Eastern Armenian, providing manually annotated morphological and syntactic data...
A Universal Dependencies (UD) treebank for Western Armenian, containing manually annotated morphological and syntacti...
A curated speech corpus of Armenian question-answer dialogues designed for intonation and prosody studies. It contain...
ArmTDP-NER is a manually annotated gold-standard named entity recognition (NER) corpus for Modern Eastern Armenian, c...
A Universal Dependencies treebank for Classical Armenian, containing annotated texts from the Gospels and Movses Khor...
This repository is part of the AI2001 project, specifically for Armenian language linguistic datasets. The README sta...
A repository containing an annotated digital version of the Kouyoumdjian 1970 "A Comprehensive Dictionary: Armenian-E...
A Universal Dependencies treebank for Eastern Armenian, manually annotated from the ArmTDP v2.0 corpus. It includes e...
A fieldwork data archive for the Iranian Armenian dialect, containing audio recordings, transcriptions, and linguisti...
A Universal Dependencies treebank for Middle Armenian, manually annotated with morphological and syntactic data, deri...