datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
pascal-voc
Pascal VOC
Dataset Summary
The Pascal Visual Object Classes (VOC) dataset is a widely used benchmark in the field of computer vision. It is designed for object detection, image classification, semantic segmentation, and action classification tasks. The dataset provides a comprehensive set of annotated images covering 20 object classes, allowing researchers to evaluate and compare the performance of various algorithms.
Note: This dataset repository contains all editions of… See the full description on the dataset page: https://huggingface.co/datasets/merve/pascal-voc.Pascal_VOCpascal_digits_10class_remapPASCAL_VOC_backup_from_JimmyUnleashedpascal-voc
Pascal VOC
Dataset Summary
The Pascal Visual Object Classes (VOC) dataset is a widely used benchmark in the field of computer vision. It is designed for object detection, image classification, semantic segmentation, and action classification tasks. The dataset provides a comprehensive set of annotated images covering 20 object classes, allowing researchers to evaluate and compare the performance of various algorithms.
Note: This dataset repository contains all editions of… See the full description on the dataset page: https://huggingface.co/datasets/yizhangdev/pascal-voc.pascal-parts-sam3pascal_rawsc_Pascal
Dataset Card for "sc_Pascal"
More Information needed
pascal-code-generation-2mbmy-single-image-dataset
Dataset Card for "my-single-image-dataset"
More Information needed
pascal-context-fixnepali-slrpascal-stahl-numarkdown
Document OCR using NuMarkdown-8B-Thinking
This dataset contains markdown-formatted OCR results from images in ShaitanRa/PascalStahl using NuMarkdown-8B-Thinking.
Processing Details
Source Dataset: ShaitanRa/PascalStahl
Model: numind/NuMarkdown-8B-Thinking
Number of Samples: 5
Processing Time: 4.1 minutes
Processing Date: 2026-04-29 11:54 UTC
Configuration
Image Column: image
Output Column: markdown
Dataset Split: train
Batch Size: 16
Max Model Length: 16… See the full description on the dataset page: https://huggingface.co/datasets/ShaitanRa/pascal-stahl-numarkdown.PASCAL_Segmentation
Dataset Card for "PASCAL_Segmentation"
More Information needed
PascalQnA100100 Pascal Q and A
60% with an input string of some kind
open-orca-slimorca-deduped-cleaned-corrected-for-pascal-txtThis is a modified version of the slimorca-deduped-cleaned-corrected dataset.
It contains English only characters.
Open Orca Slim for Pascal Developers is a subset of the original Open Orca dataset .
Open Orca Slim for Pascal Developers dataset was created with:
from datasets import load_dataset
# Coded by Gemini
def biggest_char_code(input_string):
"""
Returns the largest character code in a string.
"""
if not input_string:
return None # Handle empty string case
largest_code… See the full description on the dataset page: https://huggingface.co/datasets/schuler/open-orca-slimorca-deduped-cleaned-corrected-for-pascal-txt.pedro_pascal_lora
Dataset Card for "pedro_pascal_lora"
More Information needed
PascalRelabeled-PASCAL-VOC-2012pascal-code-generation-18mbbuzz_sources_292_pascalfeuilleton_datasetPascal-raw_trainpascal_synthetic_data_valpascal_synthetic_data_test_trainPascal-raw_testcv_pascal_vocPASCAL_overpascal-code-generation-2mbPascal-raw_val
