-
Updated
Jan 28, 2021 - Python
#
annotated-corpus
Here are 4 public repositories matching this topic...
An annotated corpus of Turkish–English intra-word code-switching collected from Reddit and annotated using the TREN application. The corpus includes token-level language labels, Leipzig-style morphological glossing and structured source metadata.
morphology turkish corpus english corpus-linguistics code-switching annotated-corpus intra-word-code-switching
-
Updated
Feb 14, 2026
A Semantic-Annotated Corpus for Asturian Based on PropBank
-
Updated
Jun 23, 2026
Unified, canonical, open corpus of Biblical Hebrew, Greek, and English texts with morpheme-level linguistic annotation and cross-language alignment for research and computational analysis.
multilingual natural-language-processing morphology text-analysis linguistics open-data alignment computational-linguistics digital-humanities structured-data corpus-linguistics parallel-corpus ancient-texts philology biblical-studies parallel-text linguistic-annotation annotated-corpus biblical-text linguistic-corpus
-
Updated
Aug 6, 2026 - Python
Add this topic to your repo
To associate your repository with the annotated-corpus topic, visit your repo's landing page and select "manage topics."