memory
Datasets
All datasets matching “memory”llm-memoryThis repository contains the results of all experiments (inlcuding every single hyperparameter run) reported in the following paper:
Orhan AE (2023) Recognition, recall, and retention of few-shot memories in large language models. arXiv:2303.17557.
A brief description of the directories included in this repository:
evals: contains the results of all recognition experiments
recalls: contains the results of all recall experiments
re-evals: contains the results of all recognition experiments… See the full description on the dataset page: https://huggingface.co/datasets/eminorhan/llm-memory.transformers-gh-memory
huggingface/transformers issues and pull requests, as a funes memory
Every issue and pull request of huggingface/transformers
with activity since 2018-01-01 — opening bodies, comments, reviews, inline review comments and PR
diffs — chunked, embedded and written to a Lance table by
funes, so the tracker can be searched by meaning and read
back thread by thread. Kept fresh every few minutes by the
funes-github Space.
Use it
Set funes up for your agent the usual way… See the full description on the dataset page: https://huggingface.co/datasets/dacorvo/transformers-gh-memory.MemoryAgentBench
🚧 Update
(Sep 29th, 2025) We updated our paper, where we removed some in-efficient and high-cost samples. We also added a sub-sample of DetectiveQA.
(July 7th, 2025) We released the initial version of our datasets.
(July 22nd, 2025) We modify the datasets slightly, adding the keypoints in LRU and change the uuid into qa_pair_ids. The question_ids is only used in Longmemeval task.
(July 26th, 2025) We fixed bug on qa_pair_ids.
(Aug.5th, 2025) We removed the… See the full description on the dataset page: https://huggingface.co/datasets/ai-hyz/MemoryAgentBench.memoryarena
MemoryArena Dataset
Overview
This dataset contains structured multi-session agentic tasks with question [list], answer [list] with necessary background context. Each row in the jsonl represents a agentic task [dict] with multiple subtasks, their corresponding answers, and background information.
Dataset Structure
Each line in the JSONL file is a dictionary with the following fields:
id (int): Unique identifier for each agentic task entry
questions… See the full description on the dataset page: https://huggingface.co/datasets/ZexueHe/memoryarena.RPent-memory
RPent Memory
Memory dataset used by RPent.
Memory layers
Memory layer
Stored content
Reuse scope
Global Memory
Cross-task general rules and failure patterns
All tasks
Task-family Memory
Strategies and precautions validated within a specific task family
Similar tasks and their variants
Task-specific Memory
Execution records and procedures from a single task run
Reference for the current task only
Apply memory only when its stated prerequisites… See the full description on the dataset page: https://huggingface.co/datasets/RLinf/RPent-memory.Memoryvla
Memoryvla(robokit 采集)
一个任务一个目录,任务下面一批一个目录:
<任务>/<批次>/hdf5/ 原始 HDF5、robokit_dataset/ RLDS、source_meta/ 采集配置与清洗报告。
Cover_the_building_block_with_a_cup,_then_lift_up_the_cup_covering_the_block.(347 段)
批次
episode
段数
RLDS
b0
0..156
157
✓
b2
0..24
25
—
b3
25..45
21
—
b1
157..300
144
✓
Open_the_drawer,_put_the_fruit_and_the_cup_from_the_table_inside,_and_close_the_drawer.(255 段)
批次
episode
段数
RLDS
b0_32
0..32
33
✓
b33_64… See the full description on the dataset page: https://huggingface.co/datasets/shaohuan1/Memoryvla.
