Skip to results
MLSift
← Feed
routineNLP & Language ModelsBERT2607.04895

Ossetic-COT: Designing a morphologically annotated corpus and morphological analyzer for Ossetic

Anna Shatskikh, Alexey Sorokin

cs.CL

Abstract

In this work we present the first morphologically annotated corpus for Iron Ossetic that conforms to the Universal Dependencies schema. The corpus includes 5454 manually annotated sentences from the Iron Ossetic Corpus of Oral Texts, containing 74032 tokens. We use this corpus to train a BERT-based morphological analyzer. The analyzer achieves tag accuracy of 95.60%.

Topics

Classified with taxonomy v2 on Wed, 2 Sept 2026.

The PDF is 1–3 MB. Open it in your browser's viewer, or load it here.

Open PDF