1uckyan/code-switch_chunks
Dataset Summary This dataset is a curated compilation of SECoMiCSC, DevCECoMiCSC, and BAAI/CS-Dialogue, specifically processed for Code-Switching ASR research. root/ ├── audio/ │ ├── SECoMiCSC/ # Chunked segments from SECoMiCSC │ ├── DevCECoMiCSC/ # Chunked segments from DevCECoMiCSC │ └── CS_Dialogue/ # Extracted <MIX> segments from BAAI/CS-Dialogue ├── metadata.jsonl # Universal index containing paths, transcripts, and metadata └──… See the full description on the dataset page: https://huggingface.co/datasets/1uckyan/code-switch_chunks.
Update README.md
Update README.md
Update README.md
Update README.md
Upload data_preperation.py
Upload folder using huggingface_hub
Delete folder metadata.jsonl with huggingface_hub
Delete folder audio with huggingface_hub
Update README.md
Upload folder using huggingface_hub
Upload folder using huggingface_hub
Upload folder using huggingface_hub
Upload folder using huggingface_hub
Upload folder using huggingface_hub
Delete folder metadata.jsonl with huggingface_hub
Delete folder audio with huggingface_hub
Upload folder using huggingface_hub
Upload folder using huggingface_hub
Delete folder metadata.jsonl with huggingface_hub
Delete folder audio with huggingface_hub
Upload folder using huggingface_hub
Delete folder DevCECoMiCSC with huggingface_hub
Delete folder SECoMiCSC with huggingface_hub
Upload folder using huggingface_hub
Upload folder using huggingface_hub
Delete folder DevCECoMiCSC with huggingface_hub
Delete folder SECoMiCSC with huggingface_hub
Upload folder using huggingface_hub
Create README.md
Upload folder using huggingface_hub
initial commit
