Team Ai
Datasetpublicgated

ivrit-ai/audio-v2

This dataset contains >20k hours of Hebrew audio, all licensed under the ivrit.ai v1 license. It wa released on April 20th, 2025. You can find the full list of sources in this dataset under the dataset's sources.txt. Paper: https://arxiv.org/abs/2307.08720 If you use our datasets, the following quote is preferable: @misc{marmor2023ivritai, title={ivrit.ai: A Comprehensive Dataset of Hebrew Speech for AI Research and Development}, author={Yanir Marmor and Kinneret Misgav and Yair… See the full description on the dataset page: https://huggingface.co/datasets/ivrit-ai/audio-v2.

sourceHugging Faceotherupdated 8mo agoView on Hugging Face
3likes2.8kdownloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.

ivrit-ai/audio-v2 · Team Ai