Team Ai
Datasetpublic

AgentVidBench123/agentvidbench

AgentVidBench: A Multi-Hop Video Question Answering Benchmark for Evaluating MLLM Agents Agentic Video Understanding Benchmark — 100 multiple-choice video QA questions, 26 options each (A-Z; ~3.8% random baseline) Anonymous authors — under review. Layout . ├── README.md ├── questions.jsonl # 100 rows — one per question ├── videos.jsonl # 71 rows — one per unique video ├── videos/ │ └── video*.mp4 # 71 video files └── transcripts/… See the full description on the dataset page: https://huggingface.co/datasets/AgentVidBench123/agentvidbench.

sourceHugging Faceccupdated 17d agoView on Hugging Face
0likes68downloads

AgentVidBench123/agentvidbench · main · files are served by the source, never re-hosted here