Published May 5, 2024 | Version v1

data_tutoria_ntic

Authors/Creators

Description

Context

This is the sentiment140 dataset. It contains 1,600,000 tweets extracted using the twitter api . The tweets have been annotated (0 = negative, 4 = positive) and they can be used to detect sentiment .

Content

It contains the following 6 fields:

  1. target: the polarity of the tweet (0 = negative, 2 = neutral, 4 = positive)

  2. ids: The id of the tweet ( 2087)

  3. date: the date of the tweet (Sat May 16 23:58:44 UTC 2009)

  4. flag: The query (lyx). If there is no query, then this value is NO_QUERY.

  5. user: the user that tweeted (robotickilldozr)

  6. text: the text of the tweet (Lyx is cool)

The creator of the dataset is Stanford University.

https://www-cs.stanford.edu/people/alecmgo/papers/TwitterDistantSupervision09.pdf

Files

trainingandtestdata.zip

Files (81.4 MB)

Name Size Download all
md5:1647eb110dd2492512e27b9a70d5d1bc
81.4 MB Preview Download