document-ai
RMC-AIDA-L_organise_the_document_bag
RMC-AIDA-L_organise_the_document_bag
📋 Overview
This dataset uses an extended format based on LeRobot and is fully compatible with LeRobot.
Robot Type: realman_rmc_aidal
| Codebase Version: v2.1
End-Effector Type: two_finger_gripper
🏠 Scene Types
This dataset covers the following scene types:
office
🤖 Atomic Actions
This dataset includes the following atomic actions:
grasp
place
pick
pull
📊 Dataset Statistics… See the full description on the dataset page: https://huggingface.co/datasets/RoboCOIN/RMC-AIDA-L_organise_the_document_bag.Unified_Document_Understanding_Dataset
Dataset Card for Read-Parsing-Describe: Unified Scientific Document Understanding
Read-Parsing-Describe (Unified Scientific Document Understanding) is a pioneering multimodal benchmark designed to train and evaluate models on the complex structures of scientific documents. Unlike traditional document datasets that treat visual elements merely as isolated layout blocks, RPD transforms document parsing into an accessibility-driven, cross-modal grounding task.
🌟 Key… See the full description on the dataset page: https://huggingface.co/datasets/chen-doc-ai/Unified_Document_Understanding_Dataset.AI2_Alphabot_2_stamp_document
AI2_Alphabot_2_stamp_document
Dataset Description
This dataset uses an extended format based on LeRobot and is fully compatible with LeRobot.
Task Preview
View Video Directly
Overview
Total Episodes: 987
Total Frames: 369702
FPS: 30
Dataset Size: 7.18 GB
Robot Name: AI2_Alphabot_2
End-Effector Type: two_finger_end_effector
Teleoperation Type: vr_controller
Sensors: cam_front_chest_rgb,
cam_front_head_rgb,
cam_left_wrist_rgb… See the full description on the dataset page: https://huggingface.co/datasets/RoboCOIN/AI2_Alphabot_2_stamp_document.ocr-document-processing-eval
ocr_document_processing_eval
Document-processing OCR proxy for digitization, KYC-like numeric fields, and RAG extraction checks.
Repo: himalaya-ai/ocr-document-processing-eval
Task: document_processing_ocr
Main raw file: *.ocr.jsonl with image, ocr, source_repo, and language/provenance columns.
Optional fine-tuning/eval file: *.sharegpt.json with messages and images.
Core Columns
id: unique sample identifier
image: relative path to the image file
ocr:… See the full description on the dataset page: https://huggingface.co/datasets/himalaya-ai/ocr-document-processing-eval.clearocr-invoice-document-ai
clearOCR Invoice Document AI Dataset
This dataset shows a complete invoice document AI workflow built around clearOCR.
It contains 423 high-confidence invoice examples with:
original invoice images,
OCR text generated by clearOCR,
Markdown reconstruction of the document,
structured invoice JSON generated by a local fine-tuned extraction model,
visual verification metadata.
The dataset demonstrates how clearOCR can serve as the OCR layer in an invoice automation pipeline where… See the full description on the dataset page: https://huggingface.co/datasets/Lukaszl/clearocr-invoice-document-ai.document-ai-paper-collection
Document AI Paper Collection
Curated repository of Document AI research papers with machine-readable catalog metadata for exploration, filtering, and downstream tooling.
Dataset Summary
Purpose: support literature navigation and benchmark tracking for Document AI topics.
Unit of data: one PDF paper file plus one catalog row in CSV/JSON.
Primary use cases:
build paper browsers and search interfaces
benchmark-oriented reading lists (DLA, parsing, OCR, structure)… See the full description on the dataset page: https://huggingface.co/datasets/tuandunghcmut/document-ai-paper-collection.
