HomeDatasetszai-org/LongBench
L

zai-org/LongBench

Question Answering · zai-org· 55.4K
Unknown 4.8 MB

LongBench is a comprehensive benchmark for multilingual and multi-task purposes, with the goal to fully measure and evaluate the ability of pre-trained language models to understand long text. This dataset consists of twenty different tasks, covering key long-text application scenarios such as multi-document QA, single-document QA, summarization, few-shot learning, synthetic tasks, and code completion.

Open in MLForge Sign up free Desktop app
# download instantly
mlforge datasets pull zai-org/LongBench

Dataset details

Task
Question Answering
Language
en
License
Unknown
Size
4.8 MB
Creator
zai-org
Downloads
55.4K
Source
huggingface_datasets
Updated
2024-12-18

About zai-org/LongBench

LongBench is the first benchmark for bilingual, multitask, and comprehensive assessment of long context understanding capabilities of large language models. LongBench includes different languages (Chinese and English) to provide a more comprehensive evaluation of the large models' multilingual capabilities on long contexts. In addition, LongBench is composed of six major categories and twenty one different tasks, covering key long-text application scenarios such as single-document QA, multi-document QA, summarization, few-shot learning, synthetic tasks and code completion.