HomeDatasetsrajpurkar/squad
S

rajpurkar/squad

Question Answering · rajpurkar· 217.5K
cc-by-sa-4.0 86 MB

Dataset Card for SQuAD Dataset Summary Stanford Question Answering Dataset (SQuAD) is a reading comprehension dataset, consisting of questions posed by crowdworkers on a set of Wikipedia articles, where the answer to every question is a segment of text, or span, from the corresponding reading passage, or the question might be unanswerable. SQuAD 1.1 contains 100,000+ question-answer pairs on 500+ articles. Supported Tasks and Leaderboards Question Answering.… See the full description on the dataset page: https://huggingface.co/datasets/rajpurkar/squad.

Open in MLForge Sign up free Desktop app
# download instantly
mlforge datasets pull rajpurkar/squad

Dataset details

Task
Question Answering
Language
en
License
cc-by-sa-4.0
Size
86 MB
Rows / images
98.2K
Classes
5
Creator
rajpurkar
Downloads
217.5K
Source
huggingface_datasets
Updated
2024-03-04

About rajpurkar/squad

Table of Contents - Dataset Card for "squad" - Table of Contents - Dataset Description - Dataset Summary - Supported Tasks and Leaderboards - Languages - Dataset Structure - Data Instances - plaintext - Data Fields - plaintext - Data Splits - Dataset Creation - Curation Rationale - Source Data - Initial Data Collection and Normalization - Who are the source language producers? - Annotations - Annotation process - Who are the annotators? - Personal and Sensitive Information - Considerations for Using the Data - Social Impact of Dataset - Discussion of Biases - Other Known Limitations - Additional Information - Dataset Curators - Licensing Information - Citation Information - Contributions