Published August 25, 2021
| Version v1
Dataset
Open
TREC
Authors/Creators
Description
TREC with 5,952 documents (i.e. questions), is a question classification dataset in which the task is classify a question into 6 main subject categories: such as human, location, entity, abbreviation, description and numeric value.
The files:
texts.txt: Document set (text). One per line.
score.txt: Document class whose index is associated with texts.txt
split_<k>.pkl: pandas DataFrame with k-cross validation partition
Files
score.txt
Files
(3.2 MB)
| Name | Size | Download all |
|---|---|---|
|
md5:d98b1b794ec93f8434b477964b931da4
|
11.9 kB | Preview Download |
|
md5:133465c9e8de42fb6d04ad7009904593
|
177.1 kB | Download |
|
md5:6b72780ffdbac68d51fb5731a5227951
|
89.0 kB | Download |
|
md5:6adf246a41613f08931dab373f311ba5
|
300.0 kB | Preview Download |
|
md5:24e989a314abfbfa4a34d561ae5c5e73
|
2.6 MB | Preview Download |