Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01snorkelai /Tau2-Bench-Airline-With-Code-Agents Dataset Card for a Code Agent Version of Tau Bench 2 Airline Dataset Summary This dataset includes sample traces and associated metadata from multi-turn interactions between an code agent and AI assistant. The dataset is based on the Airline environment from Tau^2 Bench and contains traces from both the original version and a version made at Snorkel AI using code agents to solve the same tasks (indicator in the version field; details below). Curated by: Snorkel AI… See the full description on the dataset page: https://huggingface.co/datasets/snorkelai/Tau2-Bench-Airline-With-Code-Agents.tabulartext-generationn<1K9 likes188 downloads10mo agoHugging Face02snorkelai /Tau2-Bench-Verified-Airline-With-Code-Agents Dataset Card for a Code Agent Version of Tau Bench 2 Airline Dataset Summary This dataset includes sample traces and associated metadata from multi-turn interactions between an code agent and AI assistant, along with the original verion of the tasks with more bespoke tools. The dataset is based on a verified version of the Airline environment from Sierra.ai's Tau^2 Bench with the verified version from Amazon AGI group here. You can find an earlier version of the dataset… See the full description on the dataset page: https://huggingface.co/datasets/snorkelai/Tau2-Bench-Verified-Airline-With-Code-Agents.tabularn<1K3 likes165 downloads7mo agoHugging Face03smolagents /codeagent-tracestext10K<n<100K3 likes94 downloads1y agoHugging Face04beatsprom /ai-code-generation-swe-agents-2026 💻 AI Code Generation, SWE Agents & Program Synthesis Dataset (2026 Edition) A structured research dataset featuring 3,181 domain-verified research papers and 771 official code repositories focused on Autonomous Software Engineering Agents (SWE-bench), Program Synthesis, DeepSeek-Coder-V2, Qwen2.5-Coder, Test-Driven Code Repair, Self-Healing Software, AST Semantic Modeling, and Formal Logic Verification (2023–2026). Built with Universal Scientific Engine V17.1 Gold, providing 47… See the full description on the dataset page: https://huggingface.co/datasets/beatsprom/ai-code-generation-swe-agents-2026.tabularfeature-extractionn<1K0 likes75 downloads2mo agoHugging Face05juliensimon /agent-traces-code-review-pipeline Agent Traces: code-review-pipeline Synthetic multi-agent workflow traces with LLM-enriched content for the code-review-pipeline domain. Part of the juliensimon/open-agent-traces collection — 10 datasets covering diverse domains and workflow patterns. What is this dataset? This dataset contains 2,035 events across 50 workflow runs, each representing a complete multi-agent execution trace. Every trace includes: Agent reasoning — chain-of-thought for each agent step LLM… See the full description on the dataset page: https://huggingface.co/datasets/juliensimon/agent-traces-code-review-pipeline.tabular1K<n<10K1 likes59 downloads7mo agoHugging Face06agent-data /misc-merged-claude-code-traces-v1 MISC Unification of Public Claude Code Traces A unified dataset of 32,133 deduplicated Claude API conversation traces focused on software engineering and code generation tasks. This dataset merges and normalizes traces from 10 different source datasets into a single, consistent format. Dataset Description This dataset contains real Claude API interaction traces capturing software engineering workflows including: Code generation and modification Bug fixing and… See the full description on the dataset page: https://huggingface.co/datasets/agent-data/misc-merged-claude-code-traces-v1.text10K<n<100K0 likes58 downloads8mo agoHugging Face07smolagents /hermes-function-calling-v1-formatted-code-agenttext1K<n<10K3 likes55 downloads1y agoHugging Face08smolagents /hermes-codeagenttext1K<n<10K0 likes50 downloads1y agoHugging Face09PersonalAILab /AFM-CodeAgent-RL-Dataset Data Introduction This dataset serves as the core training data for Agent Foundation Models (AFMs), specifically designed to elicit end-to-end multi-agent reasoning capabilities in large language models. Built on the novel "Chain-of-Agents (CoA)" paradigm, the dataset leverages a multi-agent distillation framework to transform collaboration processes from state-of-the-art multi-agent systems into trajectory data suitable for supervised fine-tuning (SFT), simulating dynamic… See the full description on the dataset page: https://huggingface.co/datasets/PersonalAILab/AFM-CodeAgent-RL-Dataset.text10K<n<100K1 likes49 downloads1y agoHugging Face10adityasoni17 /agentic-code-search-rollouts-sample10textn<1K0 likes45 downloads7mo agoHugging Face11akseljoonas /codeagent-tracestext1K<n<10K0 likes38 downloads1y agoHugging Face12SultanR /arxiv-to-code-agentic-tool-calling arxiv-to-code-agentic-tool-calling Multi-turn tool-calling dataset where an assistant implements ML papers in PyTorch through file-creation and command-execution tool calls. Built from lucidrains' (Phil Wang) open-source paper implementations. There are ~217 repositories on Codeberg, each implementing a different ML paper. This dataset reverse-engineers those into synthetic coding conversations. What's in it 199 conversations, each covering one repository. Every… See the full description on the dataset page: https://huggingface.co/datasets/SultanR/arxiv-to-code-agentic-tool-calling.tabulartext-generationn<1K0 likes38 downloads8mo agoHugging Face13akseljoonas /codeagent-traces-answerstext1K<n<10K1 likes37 downloads1y agoHugging Face14Hennara /code_agent_datatabular10K<n<100K0 likes32 downloads4d agoHugging Face15agent-data /code-contests-sandboxes-traces-terminus-2text10K<n<100K0 likes15 downloads8mo agoHugging Face16math-extraction-comp /EpistemeAI__Fireball-Meta-Llama-3.2-8B-Instruct-agent-003-128k-code-DPOtabular1K<n<10K0 likes13 downloads2y agoHugging Face17akseljoonas /codeagent-traces-tool-roletext1K<n<10K0 likes12 downloads1y agoHugging Face18akseljoonas /codeagent-traces-user-roletext1K<n<10K1 likes11 downloads1y agoHugging Face19maanas-writer /mem_agent-model_based-memagent-1-5b-separate-step720-infbench-code-debug-test-c27000-t4096-10s-atabularn<1K0 likes11 downloads10mo agoHugging Face20donghuna /generated_code-Agent0 likes10 downloads2y agoHugging Face21maanas-writer /mem_agent-model_based-rl-memoryagent-7b-infbench-code-debug-test-c8192-t4096-1000s-a-fullcontexttabular1K<n<10K0 likes10 downloads11mo agoHugging Face22maanas-writer /mem_agent-model_based-rl-memoryagent-7b-infbench-code-debug-test-c8192-t4096-1000s-a-nocontexttabular1K<n<10K0 likes10 downloads11mo agoHugging Face23maanas-writer /mem_agent-model_based-rl-memoryagent-7b-infbench-code-debug-test-c31000-t4096-1000s-agnostictabular1K<n<10K0 likes9 downloads11mo agoHugging Face24maanas-writer /mem_agent-model_based-rl-memoryagent-14b-infbench-code-debug-test-c31000-t4096-1000s-agnostictabular1K<n<10K0 likes9 downloads11mo agoHugging Face25maanas-writer /mem_agent-model_based-rl-memoryagent-7b-infbench-code-debug-test-c27000-t4096-1000s-agnostictabular1K<n<10K0 likes9 downloads11mo agoHugging Face26maanas-writer /mem_agent-model_based-rl-memoryagent-14b-infbench-code-debug-test-c27000-t4096-1000s-agnostictabular1K<n<10K0 likes9 downloads11mo agoHugging Face27maanas-writer /mem_agent-model_based-rl-memoryagent-7b-infbench-code-debug-test-c8192-t4096-1000s-agnostictabular1K<n<10K0 likes9 downloads11mo agoHugging Face28maanas-writer /mem_agent-model_based-memagent-1-5b-step1024-infbench-code-debug-test-c8192-t4096-10s-agnostictabularn<1K0 likes9 downloads10mo agoHugging Face29maanas-writer /mem_agent-model_based-qwen3-1-5b-oldgrpo-2086-infbench-code-debug-test-c8192-t4096-1000s-agnostitabular1K<n<10K0 likes9 downloads10mo agoHugging Face30maanas-writer /mem_agent-model_based-memagent-1-5b-separate-step720-infbench-code-debug-test-c27000-t4096-1000stabular1K<n<10K0 likes9 downloads10mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.