holvan/LongAudioSpan
LongAudioSpan: Spanning the Duration and Depth of Audio Comprehension Introduction LongAudioSpan is a benchmark for long-form audio comprehension, spanning diverse durations and cognitive depths. Questions come from two complementary paths: Native QA: questions drawn from the audio's natural content. Anchor QA: questions built around acoustic anchors planted into the audio. Each path is scored in its own mode: Accuracy: multiple choice… See the full description on the dataset page: https://huggingface.co/datasets/holvan/LongAudioSpan.
9904
