datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
dbbench-mysql-synth
DBBench-Style MySQL Synthetic SFT Dataset (English)
This directory contains a DBBench-style (MySQL / English) synthetic dataset for SFT.
It is designed for:
SFT training in a DBBench-like interaction format
Auditing data quality (SQL / answers / difficulty / distribution)
Reproducing and debugging queries on MySQL
Future DAgger / hard example mining
This dataset is synthetic and not the official DBBench data. It is meant to teach DBBench-style skills (SQL generation, result… See the full description on the dataset page: https://huggingface.co/datasets/acomagu/dbbench-mysql-synth.openalex_dbbench_synth_v5
OpenAlex-Inspired Synthetic SQL Agent Dataset
(MySQL/MariaDB, Schema-Aware, Teacher-Guided)
This dataset contains fully synthetic multi-turn SQL agent trajectories
generated over an OpenAlex-inspired relational schema.
It is designed to improve SQL-agent performance in structured,
tool-driven environments such as SQL-agent benchmarks
(e.g., AgentBench-style database tasks).
✅ No real OpenAlex data is included.
All schema definitions and rows are programmatically generated synthetic… See the full description on the dataset page: https://huggingface.co/datasets/tussiiiii/openalex_dbbench_synth_v5.dbbench_sft_dataset_react_v4_plus20dbbench_plus_alfworld_v2dbbench_v4_plus_alfworld_v5_mixeddbbench_sft_dataset_react_v4_plus20_fix16_v4openalex_dbbench_synth_v6
OpenAlex-Inspired Synthetic SQL Agent Dataset
(MySQL/MariaDB, Schema-Aware, Teacher-Guided)
This dataset contains fully synthetic multi-turn SQL agent trajectories
generated over an OpenAlex-inspired relational schema.
It is designed to improve SQL-agent performance in structured,
tool-driven environments such as SQL-agent benchmarks
(e.g., AgentBench-style database tasks).
✅ No real OpenAlex data is included.
All schema definitions and rows are programmatically generated synthetic… See the full description on the dataset page: https://huggingface.co/datasets/tussiiiii/openalex_dbbench_synth_v6.dbbench_x1_plus_alfworld_x1dbbench_sft_dataset_react
DBBench SFT Dataset (ReAct Format — AgentBench Compatible)
Overview
Synthetic SFT dataset for DBBench (AgentBench, ICLR 2024).
All tables, data, and queries are independently generated to avoid test data leakage.
Format
ReAct text format matching the AgentBench DBBench evaluation protocol:
[user] System prompt (Action: Operation / Action: Answer instructions)
[agent] Ok.
[user] Question + table name + column headers
[agent] Thinking + Action: Operation +… See the full description on the dataset page: https://huggingface.co/datasets/u-10bei/dbbench_sft_dataset_react.dbbench_and_alfworld_sft_dataset_v4
DBBench + ALFWorld SFT Dataset (Merged)
Overview
This dataset is a simple concatenation (merge) of the following two synthetic SFT datasets:
ALFWorld Trajectory Dataset: moroqq/sft_alfworld_trajectory_dataset_v2
https://huggingface.co/datasets/moroqq/sft_alfworld_trajectory_dataset_v2
DBBench SFT Dataset (ReAct Format): u-10bei/dbbench_sft_dataset_react_v4
https://huggingface.co/datasets/u-10bei/dbbench_sft_dataset_react_v4
The goal is to provide a single dataset… See the full description on the dataset page: https://huggingface.co/datasets/moroqq/dbbench_and_alfworld_sft_dataset_v4.dbbench_sft_dataset_react_v2
DBBench SFT Dataset (ReAct Format — AgentBench Compatible)
Overview
Synthetic SFT dataset for DBBench (AgentBench, ICLR 2024).
All tables, data, and queries are independently generated to avoid test data leakage.
Format
ReAct text format matching the AgentBench DBBench evaluation protocol:
[user] System prompt (Action: Operation / Action: Answer instructions)
[agent] Ok.
[user] Question + table name + column headers
[agent] Thinking + Action: Operation +… See the full description on the dataset page: https://huggingface.co/datasets/u-10bei/dbbench_sft_dataset_react_v2.dbbench_and_alfworld_sft_dataset
DBBench + ALFWorld SFT Dataset (Merged)
Overview
This dataset is a simple concatenation (merge) of the following two synthetic SFT datasets:
ALFWorld Trajectory Dataset: u-10bei/sft_alfworld_trajectory_dataset_v5
https://huggingface.co/datasets/u-10bei/sft_alfworld_trajectory_dataset_v5
DBBench SFT Dataset (ReAct Format): u-10bei/dbbench_sft_dataset_react_v4
https://huggingface.co/datasets/u-10bei/dbbench_sft_dataset_react_v4
The goal is to provide a single dataset… See the full description on the dataset page: https://huggingface.co/datasets/moroqq/dbbench_and_alfworld_sft_dataset.dbbench_cleaned_for_agentbench
DBBench Cleaned for AgentBench
u-10bei/dbbench_sft_dataset_react_v4(1,200 件)に対してクレンジング処理を施したデータセット。
AgentBench DBBench 評価用の SFT 訓練データとしてそのまま使用可能。
混合利用を想定: 本データセットは mark-22/dbbench-spider-3500(1,697 件)と混合し、合計 2,897 件 の SFT データとして使用することを想定しています。
Dataset Summary
Metric
Value
Total rows
1,200
Source
u-10bei/dbbench_sft_dataset_react_v4
Avg messages per item
6.7
Items with Final Answer1,200 / 1,200 (100%)
Columns
id, messages, metadata… See the full description on the dataset page: https://huggingface.co/datasets/mark-22/dbbench_cleaned_for_agentbench.dbbench_sft_dataset_react_v4_plus20_fix16_v1openalex_dbbench_synth_v3
OpenAlex-Inspired Synthetic SQL Agent Dataset
(MySQL/MariaDB, Teacher-Guided)
This dataset contains fully synthetic multi-turn SQL agent trajectories
generated over an OpenAlex-inspired relational schema.
It is intended to help models learn tool-driven SQL reasoning skills
that are commonly evaluated in SQL-agent benchmarks (e.g., AgentBench DB tasks),
such as multi-step querying, aggregation, and database modifications.
✅ No real OpenAlex data is included.
All schema and rows are… See the full description on the dataset page: https://huggingface.co/datasets/tussiiiii/openalex_dbbench_synth_v3.dbbench_sft_dataset_react_v4
DBBench SFT Dataset (ReAct Format — AgentBench Compatible)
Overview
Synthetic SFT dataset for DBBench (AgentBench, ICLR 2024).
All tables, data, and queries are independently generated to avoid test data leakage.
Format
ReAct text format matching the AgentBench DBBench evaluation protocol:
[user] System prompt (Action: Operation / Action: Answer instructions)
[agent] Ok.
[user] Question + table name + column headers
[agent] Thinking + Action: Operation +… See the full description on the dataset page: https://huggingface.co/datasets/u-10bei/dbbench_sft_dataset_react_v4.dbbench_and_alfworld_sft_dataset_v3
DBBench + ALFWorld SFT Dataset (Merged)
Overview
This dataset is a simple concatenation (merge) of the following two synthetic SFT datasets:
ALFWorld Trajectory Dataset: moroqq/sft_alfworld_trajectory_dataset_v2
https://huggingface.co/datasets/moroqq/sft_alfworld_trajectory_dataset_v2
DBBench SFT Dataset (ReAct Format): moroqq/dbbench_sft_dataset200
https://huggingface.co/datasets/moroqq/dbbench_sft_dataset200
The goal is to provide a single dataset repo that… See the full description on the dataset page: https://huggingface.co/datasets/moroqq/dbbench_and_alfworld_sft_dataset_v3.dbbench_sft_dataset_react_v4_plus20_fix16_v2openalex_dbbench_synth_v4
OpenAlex-Inspired Synthetic SQL Agent Dataset
(MySQL/MariaDB, Schema-Aware, Teacher-Guided)
This dataset contains fully synthetic multi-turn SQL agent trajectories
generated over an OpenAlex-inspired relational schema.
It is designed to improve SQL-agent performance in structured,
tool-driven environments such as SQL-agent benchmarks
(e.g., AgentBench-style database tasks).
✅ No real OpenAlex data is included.
All schema definitions and rows are programmatically generated synthetic… See the full description on the dataset page: https://huggingface.co/datasets/tussiiiii/openalex_dbbench_synth_v4.dbbench-spider-3500
DBBench-Spider-3500
AgentBench DBBench 評価ハーネスと完全互換のフォーマットで生成した SFT 訓練データセット。
Spider データセット (Yale NLP) の 3,500 問を GPT-OSS-120B (Groq) に解かせ、正解したトラジェクトリ 1,697 件 を収録。
混合利用を想定: 本データセットは mark-22/dbbench_cleaned_for_agentbench(1,200 件)と混合し、合計 2,897 件 の SFT データとして使用することを想定しています。
Dataset Summary
Metric
Value
Total trajectories
1,697
Difficulty: Medium
1,406
Difficulty: Hard
291
Avg messages per item
13.2
Unique databases (db_id)
159
Source questions3… See the full description on the dataset page: https://huggingface.co/datasets/mark-22/dbbench-spider-3500.agentbench_sft_mix_alfworld_dbbench_v1
AgentBench SFT Mix (ALFWorld + DBBench)
This dataset is a mixed SFT dataset created by concatenating and shuffling:
u-10bei/sft_alfworld_trajectory_dataset_v5
u-10bei/dbbench_sft_dataset_react_v4
Fields
messages: multi-turn chat messages (role/content)
tools: optional tool schemas (if present)
Credits
This dataset is a mixed and reformatted version of the original datasets listed above.
Please refer to each source dataset for their respective licenses and… See the full description on the dataset page: https://huggingface.co/datasets/tussiiiii/agentbench_sft_mix_alfworld_dbbench_v1.openalex_dbbench_synth_v2
OpenAlex-inspired Synthetic SQL Agent Dataset (SQLite, Teacher-Guided)
This dataset contains synthetic multi-turn agent trajectories
generated over an OpenAlex-inspired SQLite schema.
It is intended to improve model performance on
database reasoning benchmark tasks (e.g., DBBench),
especially aggregation-heavy and tool-driven SQL.
This version uses a teacher LLM to generate natural language
questions and reasoning thoughts, while SQL execution is fully verified.
Key… See the full description on the dataset page: https://huggingface.co/datasets/tussiiiii/openalex_dbbench_synth_v2.distilled_dbbench_dataset_2_cleanedDeepthinking-alfworld_and_dbbench_spider_v1dbbench_sft_dataset_react_v4_plus20_fix16_v4adbbench_sft_dataset_org_v3
dbbench_sft_dataset_org_v3
A DBBench-style SFT dataset combining 295 stratified examples from dbbench_u-10bei_sft_dataset_modified_v1 (100 from INSERT/INSERT_error_recovery + 195 from other types) and 5 handcrafted examples from dbbench_sft_dataset_org_v1.
Overview
Total: 300 conversations
100 randomly sampled from INSERT and INSERT_error_recovery types in modified_v1
195 randomly sampled from other types (UPDATE, aggregation, comparison, etc.) in modified_v1
5… See the full description on the dataset page: https://huggingface.co/datasets/ShogoMu/dbbench_sft_dataset_org_v3.dbbench_cleaned_v1dbbench_and_alfworld_sft_dataset_v2
DBBench + ALFWorld SFT Dataset (Merged)
Overview
This dataset is a simple concatenation (merge) of the following two synthetic SFT datasets:
ALFWorld Trajectory Dataset: moroqq/sft_alfworld_trajectory_dataset_v5_cleaned
https://huggingface.co/datasets/moroqq/sft_alfworld_trajectory_dataset_v5_cleaned
DBBench SFT Dataset (ReAct Format): moroqq/dbbench_sft_dataset200
https://huggingface.co/datasets/moroqq/dbbench_sft_dataset200
The goal is to provide a single dataset… See the full description on the dataset page: https://huggingface.co/datasets/moroqq/dbbench_and_alfworld_sft_dataset_v2.dbbench_v3_rlvmr_taggeddbbench_sft_dataset_react_v3
DBBench SFT Dataset (ReAct Format — AgentBench Compatible)
Overview
Synthetic SFT dataset for DBBench (AgentBench, ICLR 2024).
All tables, data, and queries are independently generated to avoid test data leakage.
Format
ReAct text format matching the AgentBench DBBench evaluation protocol:
[user] System prompt (Action: Operation / Action: Answer instructions)
[agent] Ok.
[user] Question + table name + column headers
[agent] Thinking + Action: Operation +… See the full description on the dataset page: https://huggingface.co/datasets/u-10bei/dbbench_sft_dataset_react_v3.
