HomeDatasetsstanfordnlp/sst2
S

stanfordnlp/sst2

Text Classification · stanfordnlp· 41.5K
["unknown"] 4.8 MB

Dataset Card for [Dataset Name] Dataset Summary The Stanford Sentiment Treebank is a corpus with fully labeled parse trees that allows for a complete analysis of the compositional effects of sentiment in language. The corpus is based on the dataset introduced by Pang and Lee (2005) and consists of 11,855 single sentences extracted from movie reviews. It was parsed with the Stanford parser and includes a total of 215,154 unique phrases from those parse trees, each… See the full description on the dataset page: https://huggingface.co/datasets/stanfordnlp/sst2.

Open in MLForge Sign up free Desktop app
# download instantly
mlforge datasets pull stanfordnlp/sst2

Dataset details

Task
Text Classification
Language
en
License
["unknown"]
Size
4.8 MB
Rows / images
70.0K
Classes
3
Creator
stanfordnlp
Downloads
41.5K
Source
huggingface_datasets
Updated
2024-01-04

About stanfordnlp/sst2

Table of Contents - Table of Contents - Dataset Description - Dataset Summary - Supported Tasks and Leaderboards - Languages - Dataset Structure - Data Instances - Data Fields - Data Splits - Dataset Creation - Curation Rationale - Source Data - Annotations - Personal and Sensitive Information - Considerations for Using the Data - Social Impact of Dataset - Discussion of Biases - Other Known Limitations - Additional Information - Dataset Curators - Licensing Information - Citation Information - Contributions