HomeDatasetsnyu-mll/blimp
B

nyu-mll/blimp

Text Classification · nyu-mll· 37.7K
["cc-by-4.0"] 48 MB

Dataset Card for "blimp" Dataset Summary BLiMP is a challenge set for evaluating what language models (LMs) know about major grammatical phenomena in English. BLiMP consists of 67 sub-datasets, each containing 1000 minimal pairs isolating specific contrasts in syntax, morphology, or semantics. The data is automatically generated according to expert-crafted grammars. Supported Tasks and Leaderboards More Information Needed Languages More Information… See the full description on the dataset page: https://huggingface.co/datasets/nyu-mll/blimp.

Open in MLForge Sign up free Desktop app
# download instantly
mlforge datasets pull nyu-mll/blimp

Dataset details

Task
Text Classification
Language
en
License
["cc-by-4.0"]
Size
48 MB
Creator
nyu-mll
Downloads
37.7K
Source
huggingface_datasets
Updated
2024-01-23

About nyu-mll/blimp

Table of Contents - Dataset Description - Dataset Summary - Supported Tasks and Leaderboards - Languages - Dataset Structure - Data Instances - Data Fields - Data Splits - Dataset Creation - Curation Rationale - Source Data - Annotations - Personal and Sensitive Information - Considerations for Using the Data - Social Impact of Dataset - Discussion of Biases - Other Known Limitations - Additional Information - Dataset Curators - Licensing Information - Citation Information - Contributions