Team Ai
Datasetpublic

Aniket-Tathe-08/Custom_common_voice_dataset_using_RVC

Custom Data Augmentation for low resource ASR using Bark and Retrieval-Based Voice Conversion Custom common_voice_v11 corpus with a custom voice was was created using RVC(Retrieval-Based Voice Conversion) The model underwent 200 epochs of training, utilizing a total of 1 hour of audio clips. The data was scraped from Youtube. The audio in the custom generated dataset is of a YouTuber named Ajay Pandey Description license: cc0-1.0 language: - hi… See the full description on the dataset page: https://huggingface.co/datasets/Aniket-Tathe-08/Custom_common_voice_dataset_using_RVC.

sourceHugging Facecc0-1.0updated 3y agoView on Hugging Face
0likes604downloads
Dataset Card

Custom Data Augmentation for low resource ASR using Bark and Retrieval-Based Voice Conversion

Custom commonvoicev11 corpus with a custom voice was was created using RVC(Retrieval-Based Voice Conversion)

The model underwent 200 epochs of training, utilizing a total of 1 hour of audio clips. The data was scraped from Youtube. The audio in the custom generated dataset is of a YouTuber named Ajay Pandey

Description


license: cc0-1.0 language:

  • —hi prettyname: Custom Common Voice sizecategories:
  • —10K<n<100K viewer: true source_datasets:
  • —commonvoicev11 ---
  • —Paper: https://arxiv.org/abs/2311.14836

Licensing Information

Public Domain, CC-0

Citation Information

@inproceedings{commonvoice:2020,
  author = {Ardila, R. and Branson, M. and Davis, K. and Henretty, M. and Kohler, M. and Meyer, J. and Morais, R. and Saunders, L. and Tyers, F. M. and Weber, G.},
  title = {Common Voice: A Massively-Multilingual Speech Corpus},
  booktitle = {Proceedings of the 12th Conference on Language Resources and Evaluation (LREC 2020)},
  pages = {4211--4215},
  year = 2020
}

Contact Information

Anand Kamble Aniket Tathe