Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01sumitaidev /agent-memory-resilience-benchmark Agent Memory Resilience & Poisoning Benchmark Dataset Summary This benchmark dataset evaluates resilience, negative transfer, and memory poisoning mitigation in autonomous LLM agent architectures (such as LangGraph, AutoGen, and CrewAI). When autonomous agents record distilled self-reflections after attempting tasks, external stochastic failures or subtle API deprecations often cause agents to commit defective strategies into episodic memory. Under standard… See the full description on the dataset page: https://huggingface.co/datasets/sumitaidev/agent-memory-resilience-benchmark.tabularreinforcement-learning1K<n<10K1 likes137 downloads17d agoHugging Face02emgena /emgena_agent_vector_memory_leak_guard_mcp_teaser 🛡️ Vector DB & Memory - HNSW Drift & Embedding Leak Guard (Evaluation Teaser) ⚡ Official Free Evaluation Teaser (50 Verified Scenarios + Executable MCP Server)🏆 Get the Full Production Package & Commercial EULA on Gumroad:👉 Purchase Full Package on Gumroad🏷️ Use coupon code LAUNCH20 for €20 off at checkout! 🌟 Domain Overview & Safety Features Monitors vector database index fragmentation, embedding distance drift, and stale semantic cache corruption for… See the full description on the dataset page: https://huggingface.co/datasets/emgena/emgena_agent_vector_memory_leak_guard_mcp_teaser.textn<1K0 likes116 downloads10d agoHugging Face03AmanPriyanshu /tool-reasoning-sft-MEMORY-mem_agent-sft-data-cleaned-rectified-408k mem_agent-sft-data-cleaned-rectified Multi-turn long-context memory-agent SFT dataset with explicit reasoning traces, structured tool calls, and sequential chunk-processing sub-chains. Schema Column Type Description messages string (JSON) JSON-serialized list of {role, content} dicts. Roles: system, user, reasoning, tool_call, tool_output, answer core_chain_OR_subcall string "core_chain" (full orchestration trace) or "subcall" (single chunk-processing step)… See the full description on the dataset page: https://huggingface.co/datasets/AmanPriyanshu/tool-reasoning-sft-MEMORY-mem_agent-sft-data-cleaned-rectified-408k.texttext-generation100K<n<1M0 likes87 downloads8mo agoHugging Face04Gde05 /agent-memory-bench-corpus agent-memory-bench: the experience corpus The neutral feed for a preregistered, execution-graded benchmark of memory layers for coding agents. Every memory product under test ingests these same bytes through its own write path, then an agent is given real coding work in a real repository where success depends on something established in an earlier session, and the artifact is graded by execution: the task's tests pass or they do not. There is no LLM judge anywhere in the primary… See the full description on the dataset page: https://huggingface.co/datasets/Gde05/agent-memory-bench-corpus.texttext-generation1K<n<10K0 likes85 downloads1mo agoHugging Face05emgena /emgena_agent_episodic_memory_pruner_mcp_teaser 🛡️ Agent Architecture - Episodic Memory Pruner & Context Optimizer (Evaluation Teaser) ⚡ Official Free Evaluation Teaser (50 Verified Scenarios + Executable MCP Server)🏆 Get the Full Production Package & Commercial EULA on Gumroad:👉 Purchase Full Package on Gumroad🏷️ Use coupon code LAUNCH20 for €20 off at checkout! 🌟 Domain Overview & Safety Features Automated attention entropy pruner condensing agent observation histories, preventing context window… See the full description on the dataset page: https://huggingface.co/datasets/emgena/emgena_agent_episodic_memory_pruner_mcp_teaser.textn<1K0 likes84 downloads10d agoHugging Face06kushalicious /agent-memory-benchmark Agent Memory Compression & Evaluation Benchmark This dataset is a controlled evaluation testbed designed to benchmark long-term memory architectures for conversational AI agents. It stress-tests how agents handle long conversations with complex fact dynamics. Dataset Structure 1. conversation.json A 100-turn synthetic conversation (50 user, 50 assistant turns) containing embedded facts categorized under: Simple Facts: Baseline retrieval details.… See the full description on the dataset page: https://huggingface.co/datasets/kushalicious/agent-memory-benchmark.textquestion-answeringn<1K0 likes81 downloads4mo agoHugging Face07ICML-2026-agent-repro /repro-learning-to-share-selective-memory-for-efficient-parallel-agentic-systems-traces Agent traces Agent sessions published from a Trackio Logbook. tabularn<1K0 likes72 downloads2mo agoHugging Face08HieuNguyenDang /long-horizon-agent-memory Long-Horizon Agent-Memory Benchmark A benchmark for evaluating agent memory and long-horizon consistency. Each case is an event stream (a multi-step conversation/trajectory) overlaid with stages that probe individual memory capabilities, each carrying a failure-mode label. Structure (40 cases, 288 stages) events[] — the full incremental event stream (one record per line in cases.jsonl). stages[] — a sparse scoring overlay: each stage has a capability, a probe… See the full description on the dataset page: https://huggingface.co/datasets/HieuNguyenDang/long-horizon-agent-memory.textquestion-answeringn<1K0 likes65 downloads4mo agoHugging Face09huyxdang /adaption-agent-memory-augmented This dataset is a remastered version prepared using Adaption's Adaptive Data platform. adaption-agent-memory (augmented) This dataset contains samples of conversations between a user and an assistant, paired with the corresponding structured memory updates extracted for long-term storage. Each sample includes the full dialogue context, existing memory state, and the resulting JSON output containing new narrative summaries and atomic facts with specific keys and values.… See the full description on the dataset page: https://huggingface.co/datasets/huyxdang/adaption-agent-memory-augmented.text10K<n<100K0 likes64 downloads29d agoHugging Face10huyxdang /adaption-agent-memory-augmented-v1 This dataset is a remastered version prepared using Adaption's Adaptive Data platform. adaption-agent-memory (augmented) This dataset contains samples of conversations between a user and an assistant, paired with the corresponding structured memory updates extracted for long-term storage. Each sample includes the full dialogue context, existing memory state, and the resulting JSON output containing new narrative summaries and atomic facts with specific keys and values.… See the full description on the dataset page: https://huggingface.co/datasets/huyxdang/adaption-agent-memory-augmented-v1.text10K<n<100K0 likes52 downloads28d agoHugging Face11DeepNLP /ai-agent-memory AI Agent Memory Agent Meta and Traffic Dataset in AI Agent Marketplace | AI Agent Directory | AI Agent Index from DeepNLP This dataset is collected from AI Agent Marketplace Index and Directory at http://www.deepnlp.org, which contains AI Agents's meta information such as agent's name, website, description, as well as the monthly updated Web performance metrics, including Google,Bing average search ranking positions, Github Stars, Arxiv References, etc. The dataset is helpful for AI… See the full description on the dataset page: https://huggingface.co/datasets/DeepNLP/ai-agent-memory.textn<1K2 likes29 downloads2y agoHugging Face12maanas-writer /mem_agent-model_based-rl-memoryagent-14b-docfinqa-train-c4096-t4096-1000s-agnostictabular1K<n<10K0 likes23 downloads11mo agoHugging Face13trentdoney /agent-memory-research-corpus Agent Memory Research Corpus (AMRC) A public, citable dataset for agent memory research and systems. This dataset catalogues papers, systems, benchmarks, and design patterns related to long-term memory in autonomous agents. It is intended to serve as a canonical reference corpus for researchers and practitioners building memory-augmented agents. Dataset Summary Field Value Repository https://huggingface.co/datasets/trentdoney/agent-memory-research-corpus… See the full description on the dataset page: https://huggingface.co/datasets/trentdoney/agent-memory-research-corpus.textn<1K0 likes18 downloads5mo agoHugging Face14maanas-writer /mem_agent-model_based-rl-memoryagent-7b-docfinqa-train-c4096-t4096-1000s-agnostic-nocontexttabular1K<n<10K0 likes15 downloads11mo agoHugging Face15Ev3lynx727 /agent-memory-graph Agent Memory Graph Unified knowledge graph combining structured entity data from mcp-agents-ark and semantic content from mempalace (ChromaDB vector store). Sanitized for public release. Usage from datasets import load_dataset # Unified (27 cols) — use this for most training tasks ds = load_dataset("Ev3lynx727/agent-memory-graph", "unified", split="train") # Ark entities only (12 cols) — structured entity graph ds_ark =… See the full description on the dataset page: https://huggingface.co/datasets/Ev3lynx727/agent-memory-graph.text1K<n<10K0 likes15 downloads3mo agoHugging Face16maanas-writer /mem_agent-model_based-rl-memoryagent-7b-docfinqa-train-c8192-t4096-1000s-agnostictabular1K<n<10K0 likes14 downloads11mo agoHugging Face17maanas-writer /mem_agent-model_based-rl-memoryagent-7b-triviaqa-llama-memorization-val-c4096-t2048-fullcontexttabularn<1K0 likes13 downloads11mo agoHugging Face18maanas-writer /mem_agent-model_based-rl-memoryagent-7b-docfinqa-train-c27000-t4096-1000s-agnostictabular1K<n<10K0 likes13 downloads11mo agoHugging Face19maanas-writer /mem_agent-model_based-rl-memoryagent-14b-docfinqa-train-c8192-t4096-1000s-agnostictabular1K<n<10K0 likes13 downloads11mo agoHugging Face20maanas-writer /mem_agent-bertscore-rl-memoryagent-14b-docfinqa-train-c4096-t4096-1000s-agnostictabular1K<n<10K0 likes12 downloads11mo agoHugging Face21maanas-writer /mem_agent-model_based-rl-memoryagent-7b-docfinqa-train-c4096-t4096-1000s-agnostic-fullcontexttabular1K<n<10K0 likes12 downloads11mo agoHugging Face22maanas-writer /mem_agent-model_based-rl-memoryagent-14b-triviaqa-val-c4096-t2048-1000s-agnostictabular1K<n<10K0 likes12 downloads11mo agoHugging Face23maanas-writer /mem_agent-bertscore-rl-memoryagent-14b-housingqa-test-c512-t256-1000s-agnostictabular1K<n<10K0 likes11 downloads11mo agoHugging Face24maanas-writer /mem_agent-model_based-rl-memoryagent-7b-housingqa-test-c32000-t512-1000s-agnostictabular1K<n<10K0 likes11 downloads11mo agoHugging Face25maanas-writer /mem_agent-model_based-rl-memoryagent-7b-docfinqa-train-c4096-t4096-1000s-agnostictabular1K<n<10K0 likes10 downloads11mo agoHugging Face26maanas-writer /mem_agent-model_based-rl-memoryagent-7b-barexamqa-train-c32000-t128-1000s-agnostictabular1K<n<10K0 likes10 downloads11mo agoHugging Face27maanas-writer /mem_agent-model_based-rl-memoryagent-14b-ruler-qa-test-c32000-t1024-1000s-agnostictabularn<1K0 likes10 downloads11mo agoHugging Face28maanas-writer /mem_agent-model_based-rl-memoryagent-14b-barexamqa-train-c32000-t128-1000s-agnostictabular1K<n<10K0 likes10 downloads11mo agoHugging Face29maanas-writer /mem_agent-model_based-rl-memoryagent-7b-gsminf-ops_12-c32000-t2048-1000s-agnostictabular1K<n<10K0 likes10 downloads11mo agoHugging Face30maanas-writer /mem_agent-model_based-rl-memoryagent-14b-housingqa-test-c32000-t512-1000s-agnostictabular1K<n<10K0 likes10 downloads11mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.