Dataset Card for The Cauldron JA Dataset description The Cauldron JA is a Vision Language Model dataset that translates 'The Cauldron' into Japanese using the DeepL API. The Cauldron is a massive collection of 50 vision language datasets (training sets only) that were used for the fine tuning of the vision language model Idefics2. To create a Japanese Vision Language Dataset, datasets related to OCR, coding, and graphs were excluded because translating them into Japanese would result in a loss of data consistency. iam ocrvqa rendered text datikz websight plotqa Ultimately, The Cauldron JA consists of 44 sub datasets . Load the dataset To load the dataset, install the library datasets with pip install datasets . Then, to download and load the config ai2d for example. License The Cauldron JA follows the same license as The Cauldron. Each of the publicly available sub datasets present in the Cauldron are governed by specific licensing conditions. Therefore, when making use of them you must take into consideration each of the licenses governing each dataset. To the extent we have any rights in the prompts, these are licensed under CC BY 4.0. Citation References to the original datasets
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy