Published June 1, 2011 | Version 1.0

Large Movie Review Dataset

Description

IMDB dataset having 50K movie reviews for natural language processing or Text analytics.
This is a dataset for binary sentiment classification containing substantially more data than previous benchmark datasets. We provide a set of 25,000 highly polar movie reviews for training and 25,000 for testing. So, predict the number of positive and negative reviews using either classification or deep learning algorithms.
For more dataset information, please go through the following link,
http://ai.stanford.edu/~amaas/data/sentiment/.

Files

IMDB Dataset.csv

Files (109.4 MB)

Name Size Download all
md5:308443a50e5c993e7b8a1cdb95750026
66.2 MB Preview Download
md5:f57a333397a32d27a88c7632886f4479
43.2 MB Preview Download

Additional details

References

  • Andrew L. Maas, Raymond E. Daly, Peter T. Pham, Dan Huang, Andrew Y. Ng, and Christopher Potts. (2011). Learning Word Vectors for Sentiment Analysis. The 49th Annual Meeting of the Association for Computational Linguistics (ACL 2011).