Team Ai
Datasetpublic

llamaindex/ExtractBench

ExtractBench Quick links: [🌐 Website] [📜 Paper] [💻 Code] Given a document and a schema, a system returns structured data with evidence. The input is a full document, born-digital or scanned, and a schema written by the user. The output is a schema-valid JSON object, with the source page and a bounding box for each value as evidence. It must return correct, exhaustive values (including repeated records), correctly use null for absent information, and ground each extracted… See the full description on the dataset page: https://huggingface.co/datasets/llamaindex/ExtractBench.

sourceHugging Faceapache-2.0updated 7d agoView on Hugging Face
34likes16kdownloads

llamaindex/ExtractBench · main · files are served by the source, never re-hosted here