Team Ai
Datasetpublic

cloudcatcher2/VCBench

Overview VCBench provides a standardized framework for evaluating vision-language models. This document outlines the procedures for both standard evaluation and GPT-assisted evaluation of your model's outputs. 1. Standard Evaluation 1.1 Output Format Requirements Models must produce outputs in JSONL format with the following structure: {"id": <int>, "pred_answer": "<answer_letter>"} {"id": <int>, "pred_answer": "<answer_letter>"} ... Example File… See the full description on the dataset page: https://huggingface.co/datasets/cloudcatcher2/VCBench.

sourceHugging Faceapache-2.0updated 3mo agoView on Hugging Face
3likes1kdownloads
../
file2.png14 KBdownload
file231.png107 KBdownload
file233.png235 KBdownload
file261.png420 KBdownload
file281.png157 KBdownload
file3.png21 KBdownload
file303.png115 KBdownload
file311.png211 KBdownload
file342.png192 KBdownload

cloudcatcher2/VCBench · main · files are served by the source, never re-hosted here