Team Ai
Datasetpublic

cjerzak/MultimodalMathBenchmarks

MultimodalMathBenchmarks This repository contains the datasets for the paper Multiplication in Multimodal LLMs: Computation with Text, Image, and Audio Inputs (ACL Findings 2026). It covers the public benchmark datasets and their modality assets (text, images, and audio) used to evaluate the arithmetic capabilities of multimodal LLMs. Canonical Upload Manifest HF path Local source Count Purpose SharedMultimodalGrid.csv SavedData/SharedMultimodalGrid.csv… See the full description on the dataset page: https://huggingface.co/datasets/cjerzak/MultimodalMathBenchmarks.

sourceHugging Facegpl-2.0updated 2mo agoView on Hugging Face
0likes218downloads
mm_00523.mp34 linesDownload Raw Back to AudioFiles
1version https://git-lfs.github.com/spec/v12oid sha256:7ba2e14d94f0fa439e0e71e31e7b8f8407243cd096f4e14732f782001f5749603size 775684 
cjerzak/MultimodalMathBenchmarks · Team Ai