datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
laser-vibrations
Laser Vibrations
Dataset of laser speckle vibration recordings used to locate objects hidden inside a cardboard box.
A 10×10 grid of lasers shines on the side of a box; as loudspeakers excite the box, each laser's
speckle pattern shifts in proportion to the local surface vibration. The goal is to reconstruct the
shape and location of an object inside the box from the vibration signals alone.
Per-sample viewer metadata lives in data/metadata.jsonl.
Full signal data and media files… See the full description on the dataset page: https://huggingface.co/datasets/eturok-weizmann/laser-vibrations.lascosas
Dataset Card for "lascosas"
More Information needed
codeswitch-pairs-lase-heldout
Codeswitch Pairs LASE — Western held-out corpus
1043 held-out cross-script utterance pairs from 8 ElevenLabs Western Multilingual voices. Used to evaluate generalisation of speaker encoders trained on Praxel/codeswitch-pairs-lase.
Each row is one synthesized utterance with metadata; pairs are reconstructed at evaluation time by joining on voice_id (same voice, different script = cross-script pair).
Schema (manifest.jsonl)
{
"voice_id": "21m00Tcm4TlvDq8ikWAM"… See the full description on the dataset page: https://huggingface.co/datasets/Praxel/codeswitch-pairs-lase-heldout.codeswitch-pairs-lase
Codeswitch Pairs LASE — training corpus
1118 same-voice cross-script utterance pairs (8 ElevenLabs Multilingual voices × en/hi/te/ta) used to train the LASE r1 speaker encoder.
Each row is one synthesized utterance with metadata; pairs are reconstructed at evaluation time by joining on voice_id (same voice, different script = cross-script pair).
Schema (manifest.jsonl)
{
"voice_id": "21m00Tcm4TlvDq8ikWAM",
"lang": "en | hi | te | ta",
"text": "the prompt text"… See the full description on the dataset page: https://huggingface.co/datasets/Praxel/codeswitch-pairs-lase.codeswitch-pairs-lase-indian
Codeswitch Pairs LASE — Indian-accent held-out corpus
1369 held-out cross-script utterance pairs from 8 ElevenLabs Indian-English Multilingual voices. Surfaces the accent-conditional finding: off-the-shelf encoders cluster Indian-accent voices closely regardless of script, while Western voices show large script-conditional gaps.
Each row is one synthesized utterance with metadata; pairs are reconstructed at evaluation time by joining on voice_id (same voice, different script =… See the full description on the dataset page: https://huggingface.co/datasets/Praxel/codeswitch-pairs-lase-indian.Lightning_Pseudo_labeled_lastlaser-vibrations-no-duplicateslass-synth-retrieval-minilast
last
This is a merged speech dataset containing 345 audio segments from 2 source datasets.
Dataset Information
Total Segments: 345
Speakers: 7
Languages: en
Emotions: neutral, angry, happy, sad
Original Datasets: 2
Dataset Structure
Each example contains:
audio: Audio file (WAV format, 16kHz sampling rate)
text: Transcription of the audio
speaker_id: Unique speaker identifier (made unique across all merged datasets)
emotion: Detected emotion… See the full description on the dataset page: https://huggingface.co/datasets/Codyfederer/last.lass-synthlaser-vibrations-test
Laser Vibrations
Dataset of laser speckle vibration recordings used to locate objects hidden inside a cardboard box.
A 10×10 grid of lasers shines on the side of a box containing an object; as loudspeakers excite the box,
the speckle patterns shift in proportion to the local surface vibration. Per-sample metadata is in
data/metadata.jsonl; full signal data and media files live in per-sample subdirectories.
Dataset Viewer Columns
Column
Type
Description… See the full description on the dataset page: https://huggingface.co/datasets/eturok-weizmann/laser-vibrations-test.librispeech_train_last_5pctlass-synthlass-synth-t2a
LASST2ARetrieval
An MTEB dataset
Massive Text Embedding Benchmark
Language-Queried Audio Source Separation (LASS) dataset for text-to-audio retrieval. Retrieve audio clips corresponding to natural language text descriptions/captions.The original dataset is based on the AudioCaps dataset.The source audio has been synthesized by mixing two audio with their labelled snr ratio as indicated in the dataset.
Task category
t2a
Domains
AudioScene
Reference… See the full description on the dataset page: https://huggingface.co/datasets/mteb/lass-synth-t2a.lass-synth-a2t
LASSA2TRetrieval
An MTEB dataset
Massive Text Embedding Benchmark
Language-Queried Audio Source Separation (LASS) dataset for audio-to-text retrieval. Retrieve text descriptions/captions for audio clips using natural language queries.The original dataset is based on the AudioCaps dataset.The source audio has been synthesized by mixing two audio with their labelled snr ratio as indicated in the dataset.
Task category
a2t
Domains
AudioScene
Reference… See the full description on the dataset page: https://huggingface.co/datasets/mteb/lass-synth-a2t.LASER-benchcodeswitch-pairs-lase-indian
Codeswitch Pairs LASE — Indian-accent held-out corpus
1369 held-out cross-script utterance pairs from 8 ElevenLabs Indian-English Multilingual voices. Surfaces the accent-conditional finding: off-the-shelf encoders cluster Indian-accent voices closely regardless of script, while Western voices show large script-conditional gaps.
Each row is one synthesized utterance with metadata; pairs are reconstructed at evaluation time by joining on voice_id (same voice, different script =… See the full description on the dataset page: https://huggingface.co/datasets/Shabarigirish/codeswitch-pairs-lase-indian.lastdance-asregyptian_myvoice-last_testTTS_finetune_last3_self_attn_v1stt_lowrank_finetuningxlsr2_1b_v2_kmeans_10k_Llama-3.2-11B-Vision-Instruct_interleaf_last_5_selfattn_ls960_TTS_v12kaggle_Pseudo_labeled_lastsiamen-laskin-sania-dyrachkin
Саня Дырачкін
Metadata
Author: Сямён Ласкін
Title: Саня Дырачкін
Narrator:
Source Group: Дзіцячыя
Source: http://staroeradio.ru
Notes
The original audio files are preserved as-is:
no conversion;
no re-encoding;
no filename changes inside each split folder, except removing one common top-level archive folder when present.
To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders.
Target maximum split size:… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/siamen-laskin-sania-dyrachkin.ivan-laskou-andrei-andyrei
Андрэй-Андырэй
Metadata
Author: Іван Ласкоў
Title: Андрэй-Андырэй
Narrator:
Source Group: Дзіцячыя
Source: http://staroeradio.ru
Notes
The original audio files are preserved as-is:
no conversion;
no re-encoding;
no filename changes inside each split folder, except removing one common top-level archive folder when present.
To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders.
Target maximum split size:… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/ivan-laskou-andrei-andyrei.uladzimir-karatkevich-grubae-i-laskavae
Грубае і ласкавае
Metadata
Author: Уладзімір Караткевіч
Title: Грубае і ласкавае
Narrator:
Source Group: Аўдыёкнігі
Source:
Notes
The original audio files are preserved as-is:
no conversion;
no re-encoding;
no filename changes inside each split folder, except removing one common top-level archive folder when present.
To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders.
Target maximum split size:… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/uladzimir-karatkevich-grubae-i-laskavae.lasmarLASCAAI
