process
BanglaASRQwen3-4B-Instruct-2507-heretic-v3-quantized-processing-i1-GGUFBLAST_PROCESSING-3.2-1B-i1-GGUFQwen3-4B-Instruct-2507-heretic-v3-quantized-processing-GGUFeagle2hg-processor-groot-n1p5BLAST_PROCESSING-3.2-1B-GGUFGitBag_-_ultrainteract_multiturn_1_iter_processed_lr_3e-7_eta_1e2_555134_1726805998-ggufbangla_tts_female
Datasets
All datasets matching “process”RekaDaily-10k-processed
RekaDaily-10k (processed)
Short first-person clips cut from the RekaDaily-10k
recordings —
unscripted daily-life video collected through Claru, Reka's
data collection marketplace, recorded by paid collectors in their own homes and
workplaces on head-mounted and handheld phones, across multiple regions.
Every clip carries one dense caption and a multi-question Q&A exchange
written in the second person ("What am I doing in this video?"), so the corpus
drops straight into… See the full description on the dataset page: https://huggingface.co/datasets/RekaAI/RekaDaily-10k-processed.ptb-xl-processedProcessed_Interiorverseparliament_hearings_processed
Preprocessed parliament hearings ASR dataset to truecased form.
Original dataset: https://lindat.mff.cuni.cz/repository/xmlui/handle/11234/1-3126
dataset_info:
features:
- name: id
dtype: string
- name: audio
dtype:
audio:
sampling_rate: 16000
- name: transcription
sequence: string
splits:
- name: train
num_bytes: 53645064353.18
num_examples: 191455
- name: test
num_bytes: 740331298.0
num_examples: 2726… See the full description on the dataset page: https://huggingface.co/datasets/jkot/parliament_hearings_processed.processed_datasetsdrh-System-Prompt-processed
