Team Ai
Datasetpublic

vxltxrllc/audio-function-calling-v1

Audio Function Calling Dataset (vxltxrllc/audio-function-calling-v1) Multimodal dataset containing conversational function-calling turns with pre-extracted Voxtral Mini (vxltxrllc/voxtral-mini-audio-extractor) latent audio features in bfloat16. Overview Audio Turns: 703 user turns processed across 137 multi-turn conversations (~85.7 minutes of effective speech audio). Feature Format: Unpadded sequence-trimmed bfloat16 tensors with shape [T_frames, 3072] stored in… See the full description on the dataset page: https://huggingface.co/datasets/vxltxrllc/audio-function-calling-v1.

sourceHugging Faceapache-2.0updated 4d agoView on Hugging Face
0likes44downloads
3 commits on main
e5fd05c4d ago

Create README.md

vxltxrsmxth
5f403f64d ago

Upload folder using huggingface_hub

vxltxrsmxth
e2102ba4d ago

initial commit

vxltxrsmxth