Published May 31, 2019
| Version v1
Dataset
Open
A New Annotation Scheme for the Sejong Part-of-speech Tagged Corpus
Description
We produce Sejong-style morphological analysis and part-of-speech tagging results which have been the de facto standard for Korean language processing by using UDPipe (http://ufal.mff.cuni.cz/udpipe)
udpipe --tokenize --tag sjmorph.model input > output
see https://github.com/jungyeul/sjmorph
Files
Files
(77.5 MB)
| Name | Size | Download all |
|---|---|---|
|
md5:1476ac0d666bdb29b028e3d5fb91c114
|
77.5 MB | Download |