Team Ai
Datasetpublic

RKB109/rag-evaluation-lab-20260918-dataset

RAG Evaluation Lab Synthetic Dataset Summary This dataset contains 14 training examples and 4 held-out examples for RAG systems often ship without a stable regression set or failure taxonomy. Every record is synthetic and includes: input: query, event, or feature description label: expected class, route, relation, or evidence category context: synthetic supporting context source: fictional source identifier variant: generation pattern synthetic: always true… See the full description on the dataset page: https://huggingface.co/datasets/RKB109/rag-evaluation-lab-20260918-dataset.

sourceHugging Facecc-by-4.0updated 22d agoView on Hugging Face
0likes65downloads
Dataset Card

RAG Evaluation Lab Synthetic Dataset

Summary

This dataset contains 14 training examples and 4 held-out examples for RAG systems often ship without a stable regression set or failure taxonomy.

Every record is synthetic and includes:

  • —input: query, event, or feature description
  • —label: expected class, route, relation, or evidence category
  • —context: synthetic supporting context
  • —source: fictional source identifier
  • —variant: generation pattern
  • —synthetic: always true

Uses

  • —Reproducible unit and integration tests
  • —Baseline model training
  • —Evaluation harness development
  • —Schema and architecture demonstrations

Limitations

Synthetic cases validate the harness, not a production RAG system. Teams must add representative domain examples.

This dataset does not represent real users, patients, customers, production traffic, or licensed media. It must not be presented as real-world evidence.

Related Model

RKB109/rag-evaluation-lab-20260918-model