Published November 21, 2019 | Version v2.1.3

cldf/segments: Unicode Standard tokenization

  • 1. Max Planck Institute for the Science of Human History
  • 2. University of Zurich
  • 3. @shh-dlce
  • 4. CUNY Grad Center; @Google AI
  • 5. GIScience, Universität Zürich

Description

Unicode Standard tokenization routines and orthography profile segmentation

Files

cldf/segments-v2.1.3.zip

Files (30.5 kB)

Name Size Download all
md5:84d52ca1c23ad6f03ce3755a05281038
30.5 kB Preview Download

Additional details

Related works