datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
StreamingOmniDatasets
StreamingOmniDatasets v0.5.0 Streaming CoT Mixed
Four training-ready configs with slim viewer schemas. Structured CoT stores concise, auditable causal state updates rather than private model thinking. Full duplex, barge-in, and simultaneous listen/speak are intentionally deferred.
conversational-streaming-asr-benchmark
SquadStack Conversational Streaming ASR Benchmark (8 kHz)
Version 1.0.0 · maintained by SquadStack
Schema · Leaderboard · Latency · Submit a system · Licence · Terms of use
Key takeaways
What this is. 863 real Hindi–English telesales calls (5.53 hours of customer speech, 8 kHz phone audio), human-transcribed turn by turn, and 11 speech recognisers scored on them. The question it answers: which recogniser should run inside an Indian voice agent, judged on… See the full description on the dataset page: https://huggingface.co/datasets/Squadstack/conversational-streaming-asr-benchmark.tarteel-ai-EA-DI-combined-stream-streamingw2v_streaming_2
Dataset Card for "w2v_streaming_2"
More Information needed
streaming_options_xtts_v2baby_cry_streaming
