vxltxrllc/audio-function-calling-v1
Audio Function Calling Dataset (vxltxrllc/audio-function-calling-v1) Multimodal dataset containing conversational function-calling turns with pre-extracted Voxtral Mini (vxltxrllc/voxtral-mini-audio-extractor) latent audio features in bfloat16. Overview Audio Turns: 703 user turns processed across 137 multi-turn conversations (~85.7 minutes of effective speech audio). Feature Format: Unpadded sequence-trimmed bfloat16 tensors with shape [T_frames, 3072] stored in… See the full description on the dataset page: https://huggingface.co/datasets/vxltxrllc/audio-function-calling-v1.
044
Create README.md
Upload folder using huggingface_hub
initial commit
