Team Ai
28 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01eturok-weizmann /laser-vibrations Laser Vibrations Dataset of laser speckle vibration recordings used to locate objects hidden inside a cardboard box. A 10×10 grid of lasers shines on the side of a box; as loudspeakers excite the box, each laser's speckle pattern shifts in proportion to the local surface vibration. The goal is to reconstruct the shape and location of an object inside the box from the vibration signals alone. Per-sample viewer metadata lives in data/metadata.jsonl. Full signal data and media files… See the full description on the dataset page: https://huggingface.co/datasets/eturok-weizmann/laser-vibrations.audion<1K0 likes998 downloads5mo agoHugging Face02ittailup /lascosas Dataset Card for "lascosas" More Information needed audio100K<n<1M0 likes138 downloads2y agoHugging Face03Praxel /codeswitch-pairs-lase-heldout Codeswitch Pairs LASE — Western held-out corpus 1043 held-out cross-script utterance pairs from 8 ElevenLabs Western Multilingual voices. Used to evaluate generalisation of speaker encoders trained on Praxel/codeswitch-pairs-lase. Each row is one synthesized utterance with metadata; pairs are reconstructed at evaluation time by joining on voice_id (same voice, different script = cross-script pair). Schema (manifest.jsonl) { "voice_id": "21m00Tcm4TlvDq8ikWAM"… See the full description on the dataset page: https://huggingface.co/datasets/Praxel/codeswitch-pairs-lase-heldout.audioaudio-classification1K<n<10K0 likes85 downloads5mo agoHugging Face04Praxel /codeswitch-pairs-lase Codeswitch Pairs LASE — training corpus 1118 same-voice cross-script utterance pairs (8 ElevenLabs Multilingual voices × en/hi/te/ta) used to train the LASE r1 speaker encoder. Each row is one synthesized utterance with metadata; pairs are reconstructed at evaluation time by joining on voice_id (same voice, different script = cross-script pair). Schema (manifest.jsonl) { "voice_id": "21m00Tcm4TlvDq8ikWAM", "lang": "en | hi | te | ta", "text": "the prompt text"… See the full description on the dataset page: https://huggingface.co/datasets/Praxel/codeswitch-pairs-lase.audioaudio-classificationn<1K0 likes78 downloads5mo agoHugging Face05Praxel /codeswitch-pairs-lase-indian Codeswitch Pairs LASE — Indian-accent held-out corpus 1369 held-out cross-script utterance pairs from 8 ElevenLabs Indian-English Multilingual voices. Surfaces the accent-conditional finding: off-the-shelf encoders cluster Indian-accent voices closely regardless of script, while Western voices show large script-conditional gaps. Each row is one synthesized utterance with metadata; pairs are reconstructed at evaluation time by joining on voice_id (same voice, different script =… See the full description on the dataset page: https://huggingface.co/datasets/Praxel/codeswitch-pairs-lase-indian.audioaudio-classification1K<n<10K2 likes64 downloads5mo agoHugging Face06AymanMansour /Lightning_Pseudo_labeled_lastaudio1K<n<10K0 likes41 downloads1y agoHugging Face07eturok-weizmann /laser-vibrations-no-duplicatesaudion<1K0 likes30 downloads6mo agoHugging Face08diffunity /lass-synth-retrieval-miniaudio1K<n<10K0 likes26 downloads9mo agoHugging Face09Codyfederer /last last This is a merged speech dataset containing 345 audio segments from 2 source datasets. Dataset Information Total Segments: 345 Speakers: 7 Languages: en Emotions: neutral, angry, happy, sad Original Datasets: 2 Dataset Structure Each example contains: audio: Audio file (WAV format, 16kHz sampling rate) text: Transcription of the audio speaker_id: Unique speaker identifier (made unique across all merged datasets) emotion: Detected emotion… See the full description on the dataset page: https://huggingface.co/datasets/Codyfederer/last.audioautomatic-speech-recognitionn<1K0 likes23 downloads1y agoHugging Face10mteb /lass-synthaudio1K<n<10K0 likes22 downloads9mo agoHugging Face11eturok-weizmann /laser-vibrations-test Laser Vibrations Dataset of laser speckle vibration recordings used to locate objects hidden inside a cardboard box. A 10×10 grid of lasers shines on the side of a box containing an object; as loudspeakers excite the box, the speckle patterns shift in proportion to the local surface vibration. Per-sample metadata is in data/metadata.jsonl; full signal data and media files live in per-sample subdirectories. Dataset Viewer Columns Column Type Description… See the full description on the dataset page: https://huggingface.co/datasets/eturok-weizmann/laser-vibrations-test.audion<1K1 likes22 downloads6mo agoHugging Face12andjelajo /librispeech_train_last_5pctaudio10K<n<100K0 likes21 downloads2y agoHugging Face13diffunity /lass-synthaudio1K<n<10K0 likes19 downloads9mo agoHugging Face14mteb /lass-synth-t2a LASST2ARetrieval An MTEB dataset Massive Text Embedding Benchmark Language-Queried Audio Source Separation (LASS) dataset for text-to-audio retrieval. Retrieve audio clips corresponding to natural language text descriptions/captions.The original dataset is based on the AudioCaps dataset.The source audio has been synthesized by mixing two audio with their labelled snr ratio as indicated in the dataset. Task category t2a Domains AudioScene Reference… See the full description on the dataset page: https://huggingface.co/datasets/mteb/lass-synth-t2a.audioother1K<n<10K0 likes19 downloads9mo agoHugging Face15mteb /lass-synth-a2t LASSA2TRetrieval An MTEB dataset Massive Text Embedding Benchmark Language-Queried Audio Source Separation (LASS) dataset for audio-to-text retrieval. Retrieve text descriptions/captions for audio clips using natural language queries.The original dataset is based on the AudioCaps dataset.The source audio has been synthesized by mixing two audio with their labelled snr ratio as indicated in the dataset. Task category a2t Domains AudioScene Reference… See the full description on the dataset page: https://huggingface.co/datasets/mteb/lass-synth-a2t.audioother1K<n<10K0 likes18 downloads9mo agoHugging Face16plnguyen2908 /LASER-benchaudio1K<n<10K1 likes17 downloads11mo agoHugging Face17Shabarigirish /codeswitch-pairs-lase-indian Codeswitch Pairs LASE — Indian-accent held-out corpus 1369 held-out cross-script utterance pairs from 8 ElevenLabs Indian-English Multilingual voices. Surfaces the accent-conditional finding: off-the-shelf encoders cluster Indian-accent voices closely regardless of script, while Western voices show large script-conditional gaps. Each row is one synthesized utterance with metadata; pairs are reconstructed at evaluation time by joining on voice_id (same voice, different script =… See the full description on the dataset page: https://huggingface.co/datasets/Shabarigirish/codeswitch-pairs-lase-indian.audioaudio-classification1K<n<10K0 likes12 downloads5mo agoHugging Face18Vu165 /lastdance-asraudion<1K0 likes12 downloads1mo agoHugging Face19mo27harakani /egyptian_myvoice-last_testaudion<1K0 likes10 downloads5mo agoHugging Face20chiyuanhsiao /TTS_finetune_last3_self_attn_v1audion<1K0 likes7 downloads2y agoHugging Face21LasseRogers2111 /stt_lowrank_finetuningaudion<1K0 likes7 downloads1y agoHugging Face22chiyuanhsiao /xlsr2_1b_v2_kmeans_10k_Llama-3.2-11B-Vision-Instruct_interleaf_last_5_selfattn_ls960_TTS_v12audion<1K0 likes6 downloads2y agoHugging Face23AymanMansour /kaggle_Pseudo_labeled_lastaudio1K<n<10K0 likes6 downloads1y agoHugging Face24archivartaunik /siamen-laskin-sania-dyrachkin Саня Дырачкін Metadata Author: Сямён Ласкін Title: Саня Дырачкін Narrator: Source Group: Дзіцячыя Source: http://staroeradio.ru Notes The original audio files are preserved as-is: no conversion; no re-encoding; no filename changes inside each split folder, except removing one common top-level archive folder when present. To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders. Target maximum split size:… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/siamen-laskin-sania-dyrachkin.audion<1K0 likes6 downloads4mo agoHugging Face25archivartaunik /ivan-laskou-andrei-andyrei Андрэй-Андырэй Metadata Author: Іван Ласкоў Title: Андрэй-Андырэй Narrator: Source Group: Дзіцячыя Source: http://staroeradio.ru Notes The original audio files are preserved as-is: no conversion; no re-encoding; no filename changes inside each split folder, except removing one common top-level archive folder when present. To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders. Target maximum split size:… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/ivan-laskou-andrei-andyrei.audion<1K0 likes5 downloads4mo agoHugging Face26archivartaunik /uladzimir-karatkevich-grubae-i-laskavae Грубае і ласкавае Metadata Author: Уладзімір Караткевіч Title: Грубае і ласкавае Narrator: Source Group: Аўдыёкнігі Source: Notes The original audio files are preserved as-is: no conversion; no re-encoding; no filename changes inside each split folder, except removing one common top-level archive folder when present. To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders. Target maximum split size:… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/uladzimir-karatkevich-grubae-i-laskavae.audion<1K0 likes5 downloads4mo agoHugging Face27speed1 /lasmaraudion<1K0 likes4 downloads3y agoHugging Face28wilsonslz /LASCAAIaudion<1K0 likes4 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.