datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
plantvillage-tiny
PlantVillage (tiny)
This is a debug-grade subset, not a faithful subsample for analysis.
50 images per class drawn from the full PlantVillage dataset is too few
to represent class-level visual diversity. Use it for iterating on
training-loop code, smoke-testing pipelines, or any situation where you
want the data structure but not the data scale. For actual classifier
training or evaluation, use
geraldmc/plantvillage-full.
What's in this dataset
A stratified subsample… See the full description on the dataset page: https://huggingface.co/datasets/geraldmc/plantvillage-tiny.radgenome-ct-reshaped-tiny
RadGenome ChestCT Reshaped Tiny Dataset
This dataset contains resized chest CT scans from the RadGenome-ChestCT dataset.
Dataset Details
Original Resolution: 900x900xN
Resized Resolution: 300x300xN
Format: NIfTI (.nii.gz)
Number of Volumes: 253
Space Reduction: ~89% (resized to 1/9th of original spatial dimensions)
Dataset Structure
Each entry contains:
volumename: Name of the CT volume file (string)
anatomy: Anatomical region information (string)
sentence:… See the full description on the dataset page: https://huggingface.co/datasets/nahidhasan/radgenome-ct-reshaped-tiny.plantvillage-tiny
PlantVillage (tiny)
This is a debug-grade subset, not a faithful subsample for analysis.
50 images per class drawn from the full PlantVillage dataset is too few
to represent class-level visual diversity. Use it for iterating on
training-loop code, smoke-testing pipelines, or any situation where you
want the data structure but not the data scale. For actual classifier
training or evaluation, use
geraldmc/plantvillage-full.
What's in this dataset
A stratified… See the full description on the dataset page: https://huggingface.co/datasets/sriyaayayaay/plantvillage-tiny.plantdoc-tiny
PlantDoc — tiny variant
Warning: this is a debug-grade subset, not a faithful subsample for analysis. Use it for test suites, smoke tests, and notebook iteration. For substantive work, use geraldmc/plantdoc-full.
A 164-image stratified subsample of geraldmc/plantdoc-full, built to support fast iteration. Loading is roughly a 50 MB download instead of ~950 MB, and a pass over the dataset takes seconds rather than minutes.
Quick start
from datasets import… See the full description on the dataset page: https://huggingface.co/datasets/yashigupta1dev/plantdoc-tiny.plantdoc-tiny
PlantDoc — tiny variant
Warning: this is a debug-grade subset, not a faithful subsample for analysis. Use it for test suites, smoke tests, and notebook iteration. For substantive work, use geraldmc/plantdoc-full.
A 164-image stratified subsample of geraldmc/plantdoc-full, built to support fast iteration. Loading is roughly a 50 MB download instead of ~950 MB, and a pass over the dataset takes seconds rather than minutes.
Quick start
from datasets import load_dataset… See the full description on the dataset page: https://huggingface.co/datasets/geraldmc/plantdoc-tiny.Self-distill-logits-convnext-tiny
ConvNeXt-Tiny experiments — results and prediction archive
Access is configured. The full artifact migration is not complete. This initial publication contains the verified retained-checkpoint/prediction audit and the immutable report-source index, not the bulk checkpoints or full-logit arrays.
Published evidence
Contents
Verified before/after audit
Exact pretrained and epoch-5 random-KL checkpoint identities; 50,000 aligned prediction rows; 78.824% → 79.868%… See the full description on the dataset page: https://huggingface.co/datasets/dlsmarta/Self-distill-logits-convnext-tiny.
