These projects are all open-source computational linguistics and natural language processing resources specifically designed for processing and analyzing the Armenian language, including its various dialects and historical forms.
A Hunspell dictionary for the Armenian language, providing spellchecking support for the hy_AM locale. Includes files...
A rule-based morphological analyzer for Modern Eastern Armenian built with uniparser-morph. It performs lemmatization...
ArmSpeech is an offline Armenian speech recognition library and CLI tool built on Coqui STT, trained on a 15.7-hour A...
A Western Armenian NLP augmentation layer designed to improve LLM output quality by providing structured linguistic d...
A repository for phonetic analysis of the Armenian language, focusing on calculating letter and digraph frequencies f...
This repository contains NLP parsing models (likely dependency parsers and/or part-of-speech taggers) specifically tr...
A rule-based morphological analyzer for Classical Armenian (Grabar) built on the uniparser-morph framework. It provid...
This is a partially implemented Classical Armenian (Grabar) text digitization project. The README describes an ambiti...
A web application for finding Armenian rhymes using IPA phoneme similarity analysis with feature-aware algorithms. It...
A Java-based tool for transliterating Western Armenian text to English. It reads .txt files from a resources director...
A small NLP project that trains classifiers to distinguish between Eastern and Western Armenian dialects using Wikipe...
A testing framework for evaluating Armenian spellcheckers on precision, accuracy, and speed. It includes tools to run...
A JavaScript library for transliterating Classical Armenian (Grabar) text into English using Western Armenian liturgi...
Apertium monolingual language package for Armenian (Western/Eastern dialects) providing morphological analysis, gener...
This repository contains the code and data pipelines for a bachelor thesis project on Armenian NLP, specifically focu...
A Python package for collecting, processing, and normalizing a Western Armenian language corpus. It provides a full E...
A web application built with Next.js that translates Modern Literary Armenian (Grakan) to Classical Armenian (Grabar)...
A Python package for clustering and deduplicating Armenian news articles using NLP and machine learning techniques. I...
A modular Retrieval-Augmented Generation (RAG) system designed for Question Answering on Armenian Labor Law documents...
A research project exploring machine learning approaches for detecting loanwords in Armenian and predicting their lan...
A Python tool for Armenian speech-to-text conversion and task extraction from transcribed text. It records Armenian s...
A rule-based morphological analyzer for Modern Eastern Armenian built with uniparser-morph. It performs full morpholo...
A Python project that extracts and visualizes character co-occurrence networks from Armenian literary texts. It uses ...
A pipeline for Armenian Named Entity Recognition (NER) and network analysis. Downloads Armenian text data, preprocess...
A complete pipeline for training Word2Vec embeddings on Armenian text, including data preprocessing, model training w...
A demonstration repository for a pre-trained Armenian text embedding model, showcasing applications in text classific...
A Python tool for sentiment analysis on Armenian literary texts from the Eastern Armenian National Corpus (EANC). It ...
A small Python script for transcribing Eastern Armenian text into a custom French-inspired phonetic transcription sys...
A student project implementing multiple tokenization methods (BPE, WordPiece, SentencePiece, tiktoken) for Armenian l...
A capstone project evaluating GPT-3.5-Turbo's performance on Armenian language tasks, including extractive QA, multip...
A Python library for transliterating Armenian text written in Latin script (Romanized Armenian) back to the Armenian ...
A Java-based toolkit for processing Armenian text, featuring multiple components for conversion, data collection, and...
A Go-based toolkit for editing and processing Armenian text, consisting of three components: an API, a converter, and...
A professional-grade English-to-Eastern Armenian literary translation tool that uses LLMs to emulate a human translat...
This repository hosts an interactive web-based etymological dictionary for Western Armenian, featuring over 18,938 en...
Armenian Contexto is a semantic word guessing game implementation for Armenian language using FastText embeddings and...
A FastAPI service for transliterating Latin-script Armenian ("armlish") into the Eastern Armenian alphabet. It combin...
An implementation of Hidden Markov Models for Part-of-Speech tagging, specifically designed for the Armenian language...
A PHP library for phonetic normalization of Armenian surnames transcribed in Latin script, designed to handle inconsi...
A research project focused on improving NLP for the low-resource Armenian language by analyzing tokenizer inefficienc...
A specialized linguistic resource providing morphological parsing data for the Nor-Nakhichevan Armenian dialect, buil...
vankatum is a specialized Armenian hyphenation library that implements syllabification-based rules for Eastern (refor...
A Python web scraping tool that fetches synonyms and definitions for Armenian words from the bararan.ru online dictio...
A voice assistant web application for Armenian language input, supporting both speech (via browser Web Speech API or ...
A production-ready SaaS translation platform for Western Armenian built with Next.js, TypeScript, Supabase, and OpenA...
A comprehensive evaluation harness for an Armenian language learning pipeline that converts TikTok videos into Russia...
An experimental autocorrect engine specifically for the Armenian language, aiming to improve accuracy and typing expe...