psd
RyanYr_-_reflect_llm8B_llm-mstlrg-om2-80k_SFTt2_psdp-T2-ggufRyanYr_-_reflect_llm8B_llm-mstlrg-om2-80k_SFTt2_psdp-T2_b.5-ggufRyanYr_-_reflect_llm8B_llm-mstlrg-om2-80k_SFTt2_psdp-t02_b1.0-ggufRyanYr_-_reflect_llm8B_llm-mstlrg-om2-80k_SFTt2_psdp-t02-ggufRyanYr_-_reflect_llama8B_om2-mistral460k_sft-t1_psdp-t1-ggufRyanYr_-_reflect_llm8B_llm-mstlrg-om2-80k_SFTt2_psdp-t02_b.5-ggufRyanYr_-_reflect_llama8B_llama-mstlrg-om2-80k_sft-t2_psdp-t02_b.5-ggufRyanYr_-_reflect_llama8B_llama-mstlrg-om2-80k_sft-t2_psdp-t02-gguf
Datasets
All datasets matching “psd”CROWS-multilingual
CROWS Multilingual
CROWS Multilingual contains 821,361 WAV recordings across 16 languages:
812,464 in main data, 2,260 in the breadth risk partition, and 6,637 in the predicted proportion risk partition. The original
release was selected by the passing_files.csv inclusion list. The publisher is
psdn-ai.
Upload status: Complete. All 821,361 recordings have been packaged,
decode-checked, and uploaded. See release_manifest.json for artifact checksums.
Main and synthetic… See the full description on the dataset page: https://huggingface.co/datasets/psdn-ai/CROWS-multilingual.bangla-10k
Bangla-10K: A Challenging, Metadata-Rich Corpus of Read and Conversational Bengali Speech from India and Bangladesh
Bangla-10K is a 10,816-hour Bengali speech corpus with
624,951 recordings from India and Bangladesh: a 10,070.8-hour core corpus
(567,323 recordings) and a separately collected 745.1-hour evaluation set
(57,628 recordings). It combines scripted single-speaker read speech with
natural multi-speaker conversations for Bengali automatic speech recognition
(ASR).
The… See the full description on the dataset page: https://huggingface.co/datasets/psdn-ai/bangla-10k.kemono_psdnumo-indic-speech
Numo Indic Speech Dataset - Open source
Numo Indic Speech is a quality-filtered, crowdsourced speech corpus collected
by Poseidon AI, Inc. It contains 353,349 recordings totaling 3,845.5 hours
across Bengali, Hindi, Tamil, and Telugu. The corpus consists of scripted,
single-speaker read speech for automatic speech recognition (ASR), with
reference transcripts, speaker metadata, and automated validation metrics for
every recording.
Corpus
3,845.5 hours · 353… See the full description on the dataset page: https://huggingface.co/datasets/psdn-ai/numo-indic-speech.train_psdPSD_image_datset_samples
InfoBay.AI SVG Dataset Catalogue
Overview
The InfoBay.AI SVG Dataset Catalogue is a professionally curated collection of 873,754 scalable vector graphics (SVG) designed for Artificial Intelligence, Computer Vision, UI/UX Design, Search Systems, Frontend Development, Design Automation, and Multimodal Foundation Models.
The dataset contains high-quality vector assets spanning user interface components, branding materials, logos, technology icons, healthcare graphics… See the full description on the dataset page: https://huggingface.co/datasets/InfoBayAI/PSD_image_datset_samples.
