Constarium
← Search

Table · dataset · 2026

yaturk-7lang

Listed in Hugging Face Datasets

YaTURK-7lang This dataset was used in the research paper No One-Size-Fits-All: Building Systems For Translation to Bashkir, Kazakh, Kyrgyz, Tatar and Chuvash Using Synthetic And Original Data.

Description

Dataset used for the online competition series "Machine Translation for Low-Resource Turkic Languages". This dataset was generated via Yandex.Translate. 📢 News – March 5, 2026 The dataset has been updated!

Additional 186729 sentence pairs were added to the dataset. Obtained from… See the full description on the dataset page: huggingface.co/datasets/dimakarp1996/YaTURK-7lang.

Links

Where it is published

Documentation and papers

Catalogue records · 1

Topics

Stated by source
tabular · text · text generation · translation
Provenance · 1 source records, 11 field assertions
SourceKeyLast seenRaw
Hugging Face Datasetsdimakarp1996/YaTURK-7lang8 d agoJSON v1
FieldAssertionExtractorEvidence
access_levelsource · Hugging Faceconnector:huggingface@1.0.0/gated
concepts[field].local:field:computer-science-aimapping · Hugging Faceconnector:huggingface@1.0.0
concepts[modality].hf_modality:tabularsource · Hugging Faceconnector:huggingface@1.0.0
concepts[modality].hf_modality:textsource · Hugging Faceconnector:huggingface@1.0.0
concepts[task].hf_task:text-generationsource · Hugging Faceconnector:huggingface@1.0.0/tags[task_categories:*]
concepts[task].hf_task:translationsource · Hugging Faceconnector:huggingface@1.0.0/tags[task_categories:*]
created_datesource · Hugging Faceconnector:huggingface@1.0.0
descriptionsource · Hugging Faceconnector:huggingface@1.0.0/description
publication_datesource · Hugging Faceconnector:huggingface@1.0.0
titlesource · Hugging Faceconnector:huggingface@1.0.0/id
updated_datesource · Hugging Faceconnector:huggingface@1.0.0