Published March 21, 2021 | Version v1

Acquiring word order in Slovak as a foreign language: comparison of Slavic and Non-Slavic learners utilizing corpus data

  • 1. University of Presov, Slovakia
  • 2. Matej Bel University, Slovakia

Description

The dataset was compiled to conduct the investigation of word order errors in texts written by foreigners learning Slovak. The data come from a pre-pilot version of a Corpus of Texts of Students Learning Slovak as a Foreign Language (errkorp-0.1) which is under development and they were completed with texts obtained from lecturers of Slovak as a foreign language and collected from non-native speakers of Slovak attending university language courses abroad. All sentences with enclitic components were transcribed into Excel and were assigned annotation tags reflecting the investigated variables. The texts were divided by proficiency level into two categories: 54 texts at the lower proficiency levels A1 – B1 and 27 texts at the higher proficiency levels B2 – C1 (according to CEFR). The aim was to obtain approximately 50 errors of enclitic placement in both language categories of texts at both investigated levels (A1-B1 and B2-C1).

Files

Files (147.5 kB)

Name Size Download all
md5:ef3613c31ac1708d81ad5f72a9e76125
147.5 kB Download

Additional details

References

  • Slovenský národný korpus – errkorp-0.1. Bratislava: Jazykovedný ústav Ľ. Štúra SAV 2020. Available from: WWW: https://korpus.juls.savba.sk.