Team Ai
Datasetpublic

pdfqa/pdfQA-Benchmark

pdfQA: Diverse, Challenging, and Realistic Question Answering over PDFs pdfQA is a structured benchmark collection for document-level question answering and PDF understanding research. The dataset is organized to support: Raw document processing research Structured extraction pipelines Retrieval-augmented QA End-to-end document reasoning systems It preserves original documents alongside structured derivatives to enable reproducible evaluation across preprocessing strategies.… See the full description on the dataset page: https://huggingface.co/datasets/pdfqa/pdfQA-Benchmark.

sourceHugging Facemitupdated 7mo agoView on Hugging Face
5likes7kdownloads
settings

This repository belongs to pdfqa on Hugging Face.

Team Ai never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namepdfQA-Benchmark
visibilitypublic
licencemit
gatedno
ownerpdfqa
Account settings
pdfqa/pdfQA-Benchmark · Team Ai