datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
browse_compaguvis-stage-2aguvis-stage-1android-controlGAIA-annotatedtoolcallingpost-train-bench-traces
PostTrainBench Sessions by Benchmark
Derived from akseljoonas/posttrainbench-sessions on 2026-04-20.
This dataset exports each source row as one viewer-compatible JSONL trace and groups traces by benchmark.
Layout
benchmarks.json: benchmark catalog and counts
benchmarks/<benchmark>/index.json: metadata index for one benchmark
benchmarks/<benchmark>/<job_id>.jsonl: one converted session trace per source row
Benchmarks
Benchmark
Sessions
aime2025
19… See the full description on the dataset page: https://huggingface.co/datasets/smolagents/post-train-bench-traces.gaia-tracesaguvis-stage-testguiact-web-singlecodeagent-tracesbenchmark-v1smolagents-merged-filteredhermes-function-calling-v1-formatted-code-agenthermes-codeagentsmolagents-toolcalling-mergedsmolagents-mergedanswersresultstraining-tracesglaive-function-calling-with-reasoningsmoltalk2_smolagents_toolcalling_french
Description
This is the SFT/smolagents_toolcalling_traces_think subset of HuggingFaceTB/smoltalk2, a tool calling dataset.We've translated the prompt and final anwser into French, while the rest (the tool call trace) remains in English.
smolagentstool-scrapingsynthetic-tracestrace-generation-taskssmolagents_benchmark_200smol_agents_benchmark_300CodeARC-Problems-TestCases-FilteredSecretAgenda_Game_Smolagents_Gemmascope_SAE_Data
