Team Ai
Datasetpublic

Lin-Chen/MMStar

MMStar (Are We on the Right Way for Evaluating Large Vision-Language Models?) 🌐 Homepage | 🤗 Dataset | 🤗 Paper | 📖 arXiv | GitHub Dataset Details As shown in the figure below, existing benchmarks lack consideration of the vision dependency of evaluation samples and potential data leakage from LLMs' and LVLMs' training data. Therefore, we introduce MMStar: an elite vision-indispensible multi-modal benchmark, aiming to ensure each curated sample… See the full description on the dataset page: https://huggingface.co/datasets/Lin-Chen/MMStar.

sourceHugging Faceupdated 3y agoView on Hugging Face
54likes18kdownloads
7 commits on main
bc98d663y ago

Merge branch 'main' of https://huggingface.co/datasets/Lin-Chen/MMStar into main

chenlin
1aa29ae3y ago

update tsv file

chenlin
2b43bbb3y ago

Update README.md

Lin-Chen
92ee3653y ago

add tsv

chenlin
f59e5323y ago

Update README.md

Lin-Chen
a9beb093y ago

init

chenlin
ddafb6f3y ago

initial commit

Lin-Chen