Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ce-amtic /ProcVQA-20M-annotationsgated ProcVQA-20M Annotations Project Page | arXiv | Code | Model | Media This repository contains the text annotations for the ProcVQA-20M dataset. The full image files are hosted separately on ProcVQA-20M-media. Overview This dataset is constructed from over 26 embodied datasets, comprising: 20M QA pairs for training 330K original trajectories 50M annotated frames from ~5,000 hours of manipulation data 200+ different tasks Dataset Structure The… See the full description on the dataset page: https://huggingface.co/datasets/ce-amtic/ProcVQA-20M-annotations.image10K<n<100K1 likes764 downloads5mo agoHugging Face02titoruizh /Drone-Orthomosaic-Vehicles-Yolo-annotation Dataset Tailings Mining Vehicles & Instruments (High-Res Drone Imagery) Dataset Summary This dataset contains high-resolution aerial imagery focused on vehicle detection and geotechnical monitoring instruments within active mining environments (tailings dams). The data was acquired using a DJI Zenmuse P1 sensor at 120m altitude. Photogrammetric Context The images originate from large-scale georeferenced orthomosaics generated from bi-daily… See the full description on the dataset page: https://huggingface.co/datasets/titoruizh/Drone-Orthomosaic-Vehicles-Yolo-annotation.imageobject-detection1K<n<10K3 likes492 downloads9mo agoHugging Face03GlowBond /TT100K_Annotationedimage10K<n<100K0 likes302 downloads7mo agoHugging Face04marijanic /map-annotation-tool-data Map Annotation Tool Media This dataset contains Docling-extracted figure crops grouped by source PDF. It is the media input for map-annotation-tool; human annotations are maintained separately in the application repository and submitted through pull requests. Contents media/figures/brgm-v1/: 1,512 figures from the earlier BRGM parsing batch, of which 156 had legacy annotations at migration time. media/figures/brgm-v2/: 6,046 figures from the newer BRGM parsing… See the full description on the dataset page: https://huggingface.co/datasets/marijanic/map-annotation-tool-data.imageimage-classification1K<n<10K1 likes293 downloads1mo agoHugging Face05DessertDrops /White_Blood_Cells_with_annotationimagen<1K0 likes244 downloads10mo agoHugging Face06WPRM /human_annotation_web_rm_version_1{ "total_stats": { "total_action_data": 19016, "total_website": 50, "action_type": { "bid": 852, "coord": 848 }, "level": { "easy": 486, "medium": 856, "hard": 358 }, "viewport_type": { "full": 707, "laptop": 698, "mobile": 295 }, "judge_type": { "string_match": 858, "url_match": 836… See the full description on the dataset page: https://huggingface.co/datasets/WPRM/human_annotation_web_rm_version_1.image10K<n<100K0 likes213 downloads2y agoHugging Face07antokun /The-Oxford-IIIT-Pet-Dataset-With-Annotationsimage0 likes208 downloads2y agoHugging Face08CLIPAMharic /AmharicCLIP-annotation AmharicCLIP Annotation Dataset 69,629 images organized by category for Amharic caption annotation. Structure images/ animals10/ 18,644 images — 10 animal classes cat/ dog/ horse/ ... intel/ 11,998 images — 6 scene classes forest/ mountain/ ... fruits360/ 38,987 images — 131 fruit classes apple/ banana/ ... Image URL Format… See the full description on the dataset page: https://huggingface.co/datasets/CLIPAMharic/AmharicCLIP-annotation.imageimage-to-text0 likes205 downloads4mo agoHugging Face09touati-kamel /forest-fire-annotations Forest Fire Detection Dataset — Auto-Annotated Bounding-box annotated version of touati-kamel/forest-fire-dataset, built for training forest-fire / smoke / fog object detection models. Overview This dataset contains video frames auto-labeled with bounding boxes for fire and smoke-related visual phenomena, using a zero-shot open-vocabulary object detector (Grounding DINO). It is derived from the original touati-kamel/forest-fire-dataset image classification dataset… See the full description on the dataset page: https://huggingface.co/datasets/touati-kamel/forest-fire-annotations.image10K<n<100K0 likes191 downloads2mo agoHugging Face10jeffrey423 /ToothXpert.MM-OPG-Annotationsimage0 likes173 downloads8mo agoHugging Face11LianeMarilin /4k-video-annotations 4K Video Annotations — Shot Segmentation and Camera Motion This dataset contains 12 frame-accurate shot clips segmented from five short cinematic video sequences. Every clip is paired with a detailed, manually reviewed annotation covering visible content, subject actions, shot scale, camera angle, camera movement, movement direction, stabilization, composition, lighting, color, pacing, transitions, timecodes, and technical properties. The footage depicts a tense nighttime… See the full description on the dataset page: https://huggingface.co/datasets/LianeMarilin/4k-video-annotations.imagen<1K0 likes171 downloads26d agoHugging Face12alakxender /od-syn-page-annotations 📦 Dhivehi Synthetic Document Layout + Textline Dataset This dataset contains synthetically generated image-document pairs with detailed layout annotations and ground-truth Dhivehi text extractions.It’s designed for document layout analysis , visual document understanding , OCR fine-tuning, and related tasks specifically for Dhivehi script. 📋 Dataset Summary Total Examples: ~58,738 Image Content: Synthetic Dhivehi documents generated to simulate real-world layouts… See the full description on the dataset page: https://huggingface.co/datasets/alakxender/od-syn-page-annotations.imagetext-classification10K<n<100K0 likes166 downloads1y agoHugging Face13alakxender /od-syn-page-annotations-com 📦 Dhivehi Synthetic Document Layout + Textline Dataset This dataset contains synthetically generated image-document pairs with detailed layout annotations and ground-truth Dhivehi text extractions.It’s designed for document layout analysis, visual document understanding, OCR fine-tuning, and related tasks specifically for Dhivehi script. Note: this version image are compressed. Raw version 📁 Repository: Hugging Face Datasets 📋 Dataset Summary Total Examples: ~58… See the full description on the dataset page: https://huggingface.co/datasets/alakxender/od-syn-page-annotations-com.imageimage-classification10K<n<100K0 likes135 downloads1y agoHugging Face14larmkaixian /malaysia-trash-annotation Notice: The foundational manuscript detailing the methodology, taxonomy (Build 20260418), and benchmark performance of the MTA dataset is currently Under Review. If you are utilizing this dataset for research, please bookmark this repository. The official BibTeX citation will be published here upon the paper's acceptance. malaysian-trash-annotation > Build-20260418-x1 -x1 refer to dataset without any augmentation applied to. Check out the Roboflow Universe… See the full description on the dataset page: https://huggingface.co/datasets/larmkaixian/malaysia-trash-annotation.imageimage-segmentation1K<n<10K0 likes133 downloads4mo agoHugging Face15ebowwa /usd-side-coco-annotations USD Side Detection Dataset (Front/Back) A refined COCO-format dataset for detecting US Dollar currency and classifying whether the front or back side is visible. Dataset Summary Total Images: 3,618 Total Annotations: 3,746 Format: COCO + HuggingFace JSONL Classes: 24 (denominations × front/back × authentic/counterfeit) Classification Accuracy: 100% (all Front/Back classified) Split Images Annotations Train 2,671 2,738 Valid 597 627 Test 350 381… See the full description on the dataset page: https://huggingface.co/datasets/ebowwa/usd-side-coco-annotations.imageobject-detection1K<n<10K0 likes129 downloads10mo agoHugging Face16Machine-Learning-Oncology /orena-frame-annotations ORena FOCUS 2026 — FRAME supplementary annotations (public half) Supplementary VQA annotations produced by MLO-Lab for the ORena FOCUS 2026 FRAME track. This is the openly releasable half; the LapChole-FOCUS half is withheld under that dataset's usage agreement until the organisers publish it. rows vqa/heico_derived.jsonl — HeiCo-FOCUS 7,608 vqa/hernia_mesh.jsonl — hernia videos 1,350 total 8,958 Also included: raw/hernia_mesh_annotations/ (12 frame-level… See the full description on the dataset page: https://huggingface.co/datasets/Machine-Learning-Oncology/orena-frame-annotations.imagevisual-question-answeringn<1K0 likes127 downloads1mo agoHugging Face17KGB0 /pick_and_place_annotation_imageThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "Unitree_G1_Inspire", "total_episodes": 1035, "total_frames": 322070, "total_tasks": 14, "total_videos": 2070, "total_chunks": 2, "chunks_size": 1000, "fps": 30, "splits": { "train": "0:1035" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/KGB0/pick_and_place_annotation_image.imagerobotics100K<n<1M0 likes123 downloads7mo agoHugging Face18BowerApp /bower-waste-annotations Dataset Card for waste annotations made by the recycling solution Bower The data offered by Bower (Sugi Group AB) in collaboration with Google.org Dataset Summary The bower-waste-annotations dataset consists of 1440 images of waste and various consumer items taken by consumer phone cameras. The images are annotated with Material type and Object type classes, listed below. The images and annotations has been manually reviewed to ensure correctness. It is assumed… See the full description on the dataset page: https://huggingface.co/datasets/BowerApp/bower-waste-annotations.image1K<n<10K4 likes115 downloads2y agoHugging Face19Machine-Learning-Oncology /orena-segment-annotations ORena FOCUS 2026 — SEGMENT supplementary annotations (public half) Supplementary VQA annotations produced by MLO-Lab for the ORena FOCUS 2026 SEGMENT track. This is the openly releasable half; the LapChole-FOCUS half is withheld under that dataset's usage agreement until the organisers publish it. rows vqa/heico_derived.jsonl — HeiCo-FOCUS 1,300 vqa/hernia_mesh.jsonl — hernia videos 140 total 1,440 Also included: raw/hernia_mesh_annotations/ (12 frame-level… See the full description on the dataset page: https://huggingface.co/datasets/Machine-Learning-Oncology/orena-segment-annotations.imagevisual-question-answering1K<n<10K0 likes114 downloads1mo agoHugging Face20netprtony /pokemon-cards-image-and-annotationsimage1K<n<10K0 likes112 downloads1y agoHugging Face21tubasid /toy-car-annotation-YOLOHey everyone, In my final year project, I created Smart Traffic Management System.The project was to manage traffic lights' delays based on the number of vehicles on road.I made everything worked using Raspberry Pi and pre-recorded videos but it was a "final year project", it was needed to be tested by changing videos frequently which was a kind of hustle. Collecting tons of videos and loading them in Pi was not too hard but it would have cost time, by every time changing names of videos in… See the full description on the dataset page: https://huggingface.co/datasets/tubasid/toy-car-annotation-YOLO.imageimage-classificationn<1K0 likes108 downloads3y agoHugging Face22capitaletech /real-resumes-section-detection-annotationsimage1K<n<10K0 likes108 downloads9mo agoHugging Face23rpzhou /HyperNeRF-Annotation This is the language query annotations for the HyperNeRF dataset, which are used in 4DLangSplat For original dataset, please visit https://github.com/google/hypernerf. For the usage of annotations, please visit https://github.com/zrporz/4DLangSplat license: cc-by-nc-4.0 imagen<1K0 likes99 downloads2y agoHugging Face24sirunchained /fruit-dataset-annotation FruitDet – Object Detection Dataset A community-driven fruit image dataset annotated for object detection, original repo.The images are real-world fruit photos collected from various environments. All images have been resized to 920×1080 pixels, and each fruit instance is labeled with a bounding box in YOLO format. Features Real-world images from diverse sources 19 fruit/vegetable categories, each containing 30 images All images resized to 920×1080 pixels… See the full description on the dataset page: https://huggingface.co/datasets/sirunchained/fruit-dataset-annotation.imageobject-detection1K<n<10K0 likes93 downloads3mo agoHugging Face25nilsho01 /vhr-buildings-automatic-annotation vhr-buildings-automatic-annotation 18,557 image–mask pairs for training building segmentation models: 4,301 chips whose labels a human inspected and kept, plus 14,256 chips a label-quality classifier judged usable without human review. A model trained on this pool reaches PQ 47.79 ± 0.23 on a held-out, human-reviewed benchmark, against 37.65 ± 1.29 for the same architecture trained on the human-reviewed chips alone — so the automatically filtered majority carries real signal… See the full description on the dataset page: https://huggingface.co/datasets/nilsho01/vhr-buildings-automatic-annotation.imageimage-segmentation10K<n<100K0 likes70 downloads9d agoHugging Face26owlgebra-ai /amz-image-annotationsimage1M<n<10M0 likes68 downloads8mo agoHugging Face27mateoguaman /annotations_only_10pct_gpt5_miniimage100K<n<1M0 likes60 downloads1y agoHugging Face28ImageIN /ImageIn_annotations_resized_images Dataset Card for ImageIn_annotations_resized_images More Information needed imageimage-classification1K<n<10K0 likes58 downloads3y agoHugging Face29openpecha /OCR-Tibetan_line_segmentation_coordinate_annotation Tibetan OCR Line Segementation Coordinates Annotation This dataset contains coordinated annotation data for lines. It includes features related to text lines, image details, and processing methods used for data annotation. Features line_id: Text Line Information line_coordinates: Coordinates of the text lines source_image: Image filename or identifier image_size: Size of the images in pixel format: Image Format bdrc_work_id: Identifier for BDRC work image_url:… See the full description on the dataset page: https://huggingface.co/datasets/openpecha/OCR-Tibetan_line_segmentation_coordinate_annotation.image10K<n<100K1 likes54 downloads1y agoHugging Face30ImageIN /ImageIn_annotationsInitial annotated dataset derived from ImageIN/IA_unlabelled imageimage-classification1K<n<10K1 likes53 downloads10mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.