Dataset Open Access

Turkic basic vocabularies

Savelyev, Alexander; Robbeets, Martine

The dataset represents basic vocabulary data across 32 Turkic languages. The basic vocabulary list merges the Leipzig-Jakarta 200 list (Haspelmath and Tadmor 2009) with the Jena 200 list (Anderson and Heggarty n.d.) and contains 254 different concepts. For each word in the dataset we provide an etymological analysis to establish cognacy classes on the basis of regular sound correspondences. Borrowings that can be identified using clearcut historical comparative criteria are excluded to provide a clearer phylogenetic signal. We deal with cases of synonymy in that we allow more than one word with a certain basic meaning in our dataset unless there is evidence that it is less basic than one of its synonyms. Singletons are removed from the dataset in case they have a non-singleton synonym that fits the criteria for basic status. The dataset also contains the tsv file edited in the EDICTOR tool (List 2017) as well as the file fed to BEAST in the original nexus format and in the XML format required by BEAST.

Files (87.1 MB)
Name Size
Savelyev&Robbeets_Turkic basic vocabularies.xls
md5:f3b49972ba4ccb7f60cd80d943e19c54
599.6 kB Download
turkic.nex
md5:64cacf7717943bcc0238861850f103e2
39.0 kB Download
turkic.xml
md5:0b6f07b382ee5246e3d466bc6c5875d8
44.3 kB Download
turkic_alignment.tsv
md5:14b84ebebc551778fcbc6a3ea2976624
593.9 kB Download
turkic_AnnotatedTree.trees
md5:e29bdb2620edd88ea373f0eefe942624
26.3 kB Download
turkic_DensiTree.trees
md5:ff748303f771fcbea87e239d2f69e52a
69.6 MB Download
turkic_logfile.log
md5:b6c171ba3294a53f5f7c922973f1bd7e
16.2 MB Download
245
79
views
downloads
All versions This version
Views 245163
Downloads 7966
Data volume 1.1 GB1.0 GB
Unique views 182152
Unique downloads 2721

Share

Cite as