Team Ai
Datasetpublic

Chulinz/Text2SQL-Decisions-Benchmark

Text2SQL-Decisions Benchmark Seed v0.1 A small, frozen English evaluation seed, separate from Text2SQL-Decisions. It contains 24 SQL-choice questions on three newly authored synthetic schemas and 20 end-to-end questions for the Olist application. It is not a large or human-reviewed benchmark. Questions and gold were authored by the same Codex assistant before running Clef, so execution checks do not establish independent semantic review. SQL-choice track: 24 examples… See the full description on the dataset page: https://huggingface.co/datasets/Chulinz/Text2SQL-Decisions-Benchmark.

sourceHugging Facecc0-1.0updated 17h agoView on Hugging Face
0likes
3 commits on main
8c6318717h ago

Record completed 25000-row Clef audit with split-separated metrics

Chulinz
3990e6918h ago

Publish frozen SQL-choice and ecommerce end-to-end benchmark seed

Chulinz
2abae3f18h ago

initial commit

Chulinz