Team Ai
Datasetpublic

C1Tech/Persian-ASR-Benchmark

This dataset consists of 3 hours of 16kHz audio collected from diverse environments to better represent real-world scenarios. The recordings were sourced from audiobooks, YouTube, and other public sources, ensuring a wide variety of speech styles and acoustic conditions. One key advantage of this dataset is that it was collected from recent sources within the last few months, ensuring no overlap with training data and fairness for evaluating other STT models. To enable a robust and fair… See the full description on the dataset page: https://huggingface.co/datasets/C1Tech/Persian-ASR-Benchmark.

sourceHugging Faceupdated 3mo agoView on Hugging Face
4likes128downloads
12 commits on main
eedc3d63mo ago

Update README.md

hosseini-ait
b4ae86e3mo ago

Update README.md

ArashAzma
236be193mo ago

Update README.md

ArashAzma
e8a31b311mo ago

Update README.md

ArashAzma
2d680de11mo ago

Upload assets

ArashAzma
cb30c5511mo ago

Update README.md

ArashAzma
7dcc62311mo ago

Update README.md

ArashAzma
1fba2e911mo ago

Update README.md

sinichiparsa
3958e4211mo ago

Update README.md

sinichiparsa
0d0be3e11mo ago

Update README.md

sinichiparsa
0d3f2721y ago

Upload dataset

ArashAzma
dfdaf4a1y ago

initial commit

ArashAzma