HomeDatasetsstanfordnlp/imdb
I

stanfordnlp/imdb

Text Classification · stanfordnlp· 176.1K
["other"] 127 MB

Dataset Card for "imdb" Dataset Summary Large Movie Review Dataset. This is a dataset for binary sentiment classification containing substantially more data than previous benchmark datasets. We provide a set of 25,000 highly polar movie reviews for training, and 25,000 for testing. There is additional unlabeled data for use as well. Supported Tasks and Leaderboards More Information Needed Languages More Information Needed Dataset Structure… See the full description on the dataset page: https://huggingface.co/datasets/stanfordnlp/imdb.

Open in MLForge Sign up free Desktop app
# download instantly
mlforge datasets pull stanfordnlp/imdb

Dataset details

Task
Text Classification
Language
en
License
["other"]
Size
127 MB
Rows / images
100.0K
Classes
2
Creator
stanfordnlp
Downloads
176.1K
Source
huggingface_datasets
Updated
2024-01-04

About stanfordnlp/imdb

Table of Contents - Dataset Description - Dataset Summary - Supported Tasks and Leaderboards - Languages - Dataset Structure - Data Instances - Data Fields - Data Splits - Dataset Creation - Curation Rationale - Source Data - Annotations - Personal and Sensitive Information - Considerations for Using the Data - Social Impact of Dataset - Discussion of Biases - Other Known Limitations - Additional Information - Dataset Curators - Licensing Information - Citation Information - Contributions