Team Ai
Datasetpublic

BDRC/tibetan-script-classification-benchmark

Tibetan Script Classification Benchmark Holdout benchmark for 6-class Tibetan script classification. Test split only — not used during training. All images are BDRC manuscript page scans, balanced by subclass. Class Images Subclasses Danyig 60 DraDring: 25, DraRing: 9, Drathung: 17, Gongshabma: 3, Tsegdrig: 6 Druma 60 Dhumri: 22, DruDring: 20, DruRing: 10, Druchen: 2, Druthung: 6 Gyuyig 60 Khyuyig: 31, Tsumachug: 15, Yigchung: 14 Pedri 60 Peri: 44, Petsuk: 16… See the full description on the dataset page: https://huggingface.co/datasets/BDRC/tibetan-script-classification-benchmark.

sourceHugging Facemitupdated 3mo agoView on Hugging Face
0likes31downloads
split_stats.md23 linesDownload Raw Back to root
1# Split statistics2 3- **Source:** `parquet`4- **Total images:** 7205 6 7## Images per split8 9| Split | Total |10|-------|------:|11| test | 720 |12 13## Images per class (per split)14 15| Class | train | val | test | **All** |16|-------|------:|------:|------:|------:|17| Danyig | 0 | 0 | 120 | 120 |18| Druma | 0 | 0 | 120 | 120 |19| Gyuyig | 0 | 0 | 120 | 120 |20| Pedri | 0 | 0 | 120 | 120 |21| Tsugdri | 0 | 0 | 120 | 120 |22| Uchen | 0 | 0 | 120 | 120 |23