HomeDatasetsxlangai/BRIGHT
B

xlangai/BRIGHT

Text Retrieval · xlangai· 18.6K
cc-by-4.0 4.8 MB

BRIGHT benchmark BRIGHT is the first text retrieval benchmark that requires intensive reasoning to retrieve relevant documents. The queries are collected from diverse domains (StackExchange, LeetCode, and math competitions), all sourced from realistic human data. Experiments show that existing retrieval models perform poorly on BRIGHT, where the highest score is only 22.1 measured by nDCG@10. BRIGHT provides a good testbed for future retrieval research in more realistic and… See the full description on the dataset page: https://huggingface.co/datasets/xlangai/BRIGHT.

Open in MLForge Sign up free Desktop app
# download instantly
mlforge datasets pull xlangai/BRIGHT

Dataset details

Task
Text Retrieval
Language
en
License
cc-by-4.0
Size
4.8 MB
Creator
xlangai
Downloads
18.6K
Source
huggingface_datasets
Updated
2025-03-01

About xlangai/BRIGHT

BRIGHT benchmark BRIGHT is the first text retrieval benchmark that requires intensive reasoning to retrieve relevant documents. The queries are collected from diverse domains (StackExchange, LeetCode, and math competitions), all sourced from realistic human data. Experiments show that existing retrieval models perform poorly on BRIGHT, where the highest score is only 22.1 measured by nDCG@10. BRIGHT provides a good testbed for future retrieval research in more realistic and challenging settings. More details are in the paper.