HomeDatasetsturing-motors/Cauldron-JA
C

turing-motors/Cauldron-JA

Visual Question Answering · turing-motors· 15.0K
cc-by-4.0 134 GB

The Cauldron-JA is a Vision Language Model dataset that translates 'The Cauldron' into Japanese using the DeepL API. The Cauldron is a massive collection of 50 vision-language datasets (training sets only) that were used for the fine-tuning of the vision-language model Idefics2.

Open in MLForge Sign up free Desktop app
# download instantly
mlforge datasets pull turing-motors/Cauldron-JA

Dataset details

Task
Visual Question Answering
Language
ja
License
cc-by-4.0
Size
134 GB
Rows / images
1.5M
Creator
turing-motors
Downloads
15.0K
Source
huggingface_datasets
Updated
2024-10-24

About turing-motors/Cauldron-JA

The Cauldron-JA is a Vision Language Model dataset that translates 'The Cauldron' into Japanese using the DeepL API. The Cauldron is a massive collection of 50 vision-language datasets (training sets only) that were used for the fine-tuning of the vision-language model Idefics2.