Data · dataset · 2016
A Central Asian Language Survey
Listed in DataCite
TablesWe have documented language varieties (either Turkic or Indo-European) spoken in 23 test sites by 88 informants belonging to the major ethnic groups of Kyrgyzstan, Tajikistan and Uzbekistan (Karakalpaks, Kazakhs, Kyrgyz, Tajiks, Uzbeks, Yaghnobis).
Description
The recorded linguistic material concerns 176 words of the extended Swadesh list and will be made publically available with the publication of this paper. Phonological diversity is measured by the Levenshtein distance and displayed as a consensus bootstrap tree and as multidimensional scaling plots.
Linguistic contact is measured as the number of borrowings, from one linguistic family into the other, according to a precision/recall analysis further validated by expert judgment. Concerning Turkic languages, the results of our sample do not support Kazakh and Karakalpak as distinct languages and indicate the existence of several separate Karakalpak varieties. Kyrgyz and Uzbek, on the other hand, appear quite homogeneous.
Read the rest (1 more)
Among the Indo-Iranian languages, the distinction between Tajik and Yaghnobi varieties is very clear-cut. More generally, the degree of borrowing is higher than average where language families are in contact in one of the many sorts of situations characterizing Central Asia: frequent bilingualism, shifting political boundaries, ethnic groups living outside the “mother” country.
Links
Where it is published
- Repository landing page brill.figshare.com/articles/dataset/A_Central_Asian_Language_Survey/3443090 ↗
landing page · from DataCite
- DOI doi.org/10.6084/m9.figshare.3443090 ↗
DOI / persistent id · from DataCite
Documentation and papers
- Creative Commons Attribution 4.0 International creativecommons.org/licenses/by/4.0/legalcode ↗
license · from DataCite
- IsSupplementTo 10.1163/22105832-00601015 doi.org/10.1163/22105832-00601015 ↗
publication · from DataCite
Catalogue records · 2
- DataCite API api.datacite.org/dois/10.6084/m9.figshare.3443090 ↗
metadata API · from DataCite
- DataCite Commons commons.datacite.org/doi.org/10.6084/m9.figshare.3443090 ↗
catalogue entry · from DataCite
Topics
- Stated by source
- Languages and literature
Related
- Inverse of same asA Central Asian Language Survey
Provenance · 1 source records, 9 field assertions
| Source | Key | Last seen | Raw |
|---|---|---|---|
| DataCite | 10.6084/m9.figshare.3443090 | 12 d ago | JSON v1 |
| Field | Assertion | Extractor | Evidence |
|---|---|---|---|
| access_level | source · DataCite | connector:datacite@1.0.0 | /data/attributes/rightsList |
| byte_size | source · DataCite | connector:datacite@1.0.0 | |
| concepts[field].fos:languages-and-literature | source · DataCite | connector:datacite@1.0.0 | |
| created_date | source · DataCite | connector:datacite@1.0.0 | |
| description | source · DataCite | connector:datacite@1.0.0 | /data/attributes/descriptions |
| license | source · DataCite | connector:datacite@1.0.0 | /data/attributes/rightsList |
| publication_date | source · DataCite | connector:datacite@1.0.0 | /data/attributes/dates |
| title | source · DataCite | connector:datacite@1.0.0 | /data/attributes/titles/0/title |
| updated_date | source · DataCite | connector:datacite@1.0.0 |