Turkic basic vocabularies

doi:10.5281/zenodo.3555174

Published November 28, 2019 | Version v1

Dataset Open

Turkic basic vocabularies

1. Max Planck Institute for the Science of Human History

The dataset represents basic vocabulary data across 32 Turkic languages. The basic vocabulary list merges the Leipzig-Jakarta 200 list (Haspelmath and Tadmor 2009) with the Jena 200 list (Anderson and Heggarty n.d.) and contains 254 different concepts. For each word in the dataset we provide an etymological analysis to establish cognacy classes on the basis of regular sound correspondences. Borrowings that can be identified using clearcut historical comparative criteria are excluded to provide a clearer phylogenetic signal. We deal with cases of synonymy in that we allow more than one word with a certain basic meaning in our dataset unless there is evidence that it is less basic than one of its synonyms. Singletons are removed from the dataset in case they have a non-singleton synonym that fits the criteria for basic status. The dataset also contains the tsv file edited in the EDICTOR tool (List 2017) and the file fed to BEAST in the original nexus format.

Files

Files (1.2 MB)

Name	Size	Download all
Savelyev&Robbeets_Turkic basic vocabularies.xls md5:f3b49972ba4ccb7f60cd80d943e19c54	599.6 kB	Download
turkic.nex md5:29db1c61c285b537b3a982e6e66dd558	39.0 kB	Download
turkic_alignment.tsv md5:14b84ebebc551778fcbc6a3ea2976624	593.9 kB	Download

Additional details

Eurasia3angle – Millet and beans, language and genes. The origin and dispersal of the Transeurasian family. 646612: European Commission

	All versions	This version
Views	2,647	447
Downloads	473	40
Data volume	6.4 GB	19.7 MB

Turkic basic vocabularies

Creators

Description

Files

Files (1.2 MB)

Additional details

Funding