Published June 1, 2011
| Version 1.0
Dataset
Open
Large Movie Review Dataset
Authors/Creators
- 1. Stanford
Description
IMDB dataset having 50K movie reviews for natural language processing or Text analytics.
This is a dataset for binary sentiment classification containing substantially more data than previous benchmark datasets. We provide a set of 25,000 highly polar movie reviews for training and 25,000 for testing. So, predict the number of positive and negative reviews using either classification or deep learning algorithms.
For more dataset information, please go through the following link,
http://ai.stanford.edu/~amaas/data/sentiment/.
Files
IMDB Dataset.csv
Files
(109.4 MB)
| Name | Size | Download all |
|---|---|---|
|
md5:308443a50e5c993e7b8a1cdb95750026
|
66.2 MB | Preview Download |
|
md5:f57a333397a32d27a88c7632886f4479
|
43.2 MB | Preview Download |
Additional details
References
- Andrew L. Maas, Raymond E. Daly, Peter T. Pham, Dan Huang, Andrew Y. Ng, and Christopher Potts. (2011). Learning Word Vectors for Sentiment Analysis. The 49th Annual Meeting of the Association for Computational Linguistics (ACL 2011).