datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
herislab-ca-training-data
CA_Training_Data -- Convolutional Autoencoder (Track A)
Curated dataset for training and evaluating the Convolutional Autoencoder anomaly detection model.
Approach
The autoencoder is trained only on normal (no-fault) images. At inference, high reconstruction error indicates an anomaly/fault.
Structure
train/normal/ -- Normal images for autoencoder training
electric_motor/ -- 168 PNG (Electric Motor Thermal Fault Diagnosis, no_fault class)… See the full description on the dataset page: https://huggingface.co/datasets/Ryanflash/herislab-ca-training-data.matt-training-imgMy dataset for training SDXL & SD 1.5
lanternfly_swatter_training
Spotted Lanternfly Classification Dataset
Dataset Description
This dataset contains images for binary classification of spotted lanternflies (Lycorma delicatula), an invasive species causing significant damage to agriculture and ecosystems in the United States. The dataset is designed to train machine learning models to identify dead or squashed lanternflies from photographs, supporting community-driven environmental monitoring and pest management efforts.
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/rlogh/lanternfly_swatter_training.football-dataset_training
Dataset Description
This dataset contains football (soccer) field images captured from a tactical camera perspective. The dataset contains two images per frame;
One image highlights only the green color, while all other colors are grayed out.
The other image grays out the green color, while all other colors remain in full color.
The dataset is designed for computer vision research in sports analytics, including player tracking, field understanding, and tactical analysis.
Field… See the full description on the dataset page: https://huggingface.co/datasets/chimp-ll/football-dataset_training.Roof_Training_Images_2training
87cc2s/training
Numerosity training data, organized by task then by source:
counting-training/
synthetic/ <- abstract shapes generator
real-world/ <- real photographs, ~19 public counting/detection datasets
ans-training/
synthetic/ <- abstract shapes generator
real-world/ <- same-category pairs of real photographs
Each *-training/<domain>/ folder has its own dataset_card.md (schema +
generation/curation details), manifest.json (summary stats), and… See the full description on the dataset page: https://huggingface.co/datasets/87cc2s/training.autotrain-data-hannah-training-demo
AutoTrain Dataset for project: hannah-training-demo
Dataset Description
This dataset has been automatically processed by AutoTrain for project hannah-training-demo.
Languages
The BCP-47 code for the dataset's language is unk.
Dataset Structure
Data Instances
A sample from this dataset looks as follows:
[
{
"image": "<842x1392 RGBA PIL image>",
"target": 0
},
{
"image": "<1004x1516 RGBA PIL image>",
"target": 0
}
]… See the full description on the dataset page: https://huggingface.co/datasets/slushily/autotrain-data-hannah-training-demo.Training
