HuggingFaceM4/the_cauldron
Dataset Card for The Cauldron Dataset description The Cauldron is part of the Idefics2 release. It is a massive collection of 50 vision-language datasets (training sets only) that were used for the fine-tuning of the vision-language model Idefics2. Load the dataset To load the dataset, install the library datasets with pip install datasets. Then, from datasets import load_dataset ds = load_dataset("HuggingFaceM4/the_cauldron", "ai2d") to download… See the full description on the dataset page: https://huggingface.co/datasets/HuggingFaceM4/the_cauldron.
561393k
../
train-00000-of-00016-fafacecb273170e4.parquetdownload
train-00001-of-00016-6ae168979846d4a4.parquetdownload
train-00002-of-00016-ab4912d25145209f.parquetdownload
train-00003-of-00016-c66704a214afef50.parquetdownload
train-00004-of-00016-c6cace90761c5a76.parquetdownload
train-00005-of-00016-948e185c5d91409f.parquetdownload
train-00006-of-00016-ab2900f85239648c.parquetdownload
train-00007-of-00016-de332dcb241624bd.parquetdownload
train-00008-of-00016-f429c3cf2c436b91.parquetdownload
train-00009-of-00016-04c10deb03901c4c.parquetdownload
train-00010-of-00016-aeb20cfa7cb16748.parquetdownload
train-00011-of-00016-24c09239c38f0ef1.parquetdownload
train-00012-of-00016-a9801228ea02a682.parquetdownload
train-00013-of-00016-b2c786ccb8ca2c23.parquetdownload
train-00014-of-00016-490b4e96632c3112.parquetdownload
train-00015-of-00016-3bb91b2e801dff48.parquetdownload
