MiniCPM
Datasets
All datasets matching “MiniCPM”Transmem_ecsd_minicpm5_1b_hotpotqa_n4_n8MiniCPM5-1B-atlas
juiceb0xc0de/MiniCPM5-1B-atlas
A brain atlas for openbmb/MiniCPM5-1B, a 1B on-device model with a 130k bilingual vocabulary. This is not a chat dataset or a benchmark. It is an internal-mechanics map, built by running activations through a corpus of prompts and scoring what each layer, component, head, and feature direction is doing.
If you want to know which parts of this model are safe to edit, where its output-vocabulary directions live, or which layers are carrying the most… See the full description on the dataset page: https://huggingface.co/datasets/juiceb0xc0de/MiniCPM5-1B-atlas.minicpm5-swe-native-eval-archive
MiniCPM5 原生 SWE 评测归档:32 run / 5842条任务记录
历史100-turn协议为11run/2200题,新600-turn协议为15run/3000题。各组独立目录;按同题配对,并保留调度和review差异。新协议表格见本文后半部分。
历史100-turn协议:11 run
11个run、2200题;同一固定100 Verified +100 Pro,各run终态与scratch/serving清理已核验。每题一个原始CC轨迹JSON,不重复保存每轮完整请求历史。
run
Verified 正确/已评分
V review
Pro 正确/已评分
P review
midtrain
51/93 (54.84%)
7
44/89 (49.44%)
11
step500
31/69 (44.93%)
31
21/60 (35.00%)
40
step1000
39/93 (41.94%)
7
19/83 (22.89%)
17
step1500
18/52… See the full description on the dataset page: https://huggingface.co/datasets/eigentom/minicpm5-swe-native-eval-archive.MiniCPM-RobotManip-LIBERO
MiniCPM-RobotManip LIBERO
This dataset contains the four LIBERO suites converted to LeRobot v3 format
for the MiniCPM-RobotManip LIBERO full-parameter fine-tuning example in
starVLA.
Dataset summary
Suite
Episodes
Frames
Videos
LIBERO-10
358
95,740
716
LIBERO-Goal
405
48,131
810
LIBERO-Object
450
66,294
900
LIBERO-Spatial
423
51,707
846
Total
1,636
261,872
3,272
Format: LeRobot v3
Frequency: 20 Hz
Cameras: agent view and wrist view
Video… See the full description on the dataset page: https://huggingface.co/datasets/openbmb/MiniCPM-RobotManip-LIBERO.latent-state-tracking-minicpm5
Tracking and Intervening on Latent State Dynamics in a Small Language Agent (MiniCPM5-2B)
Date: 2026-09-17
Model studied: openbmb/MiniCPM5-2B (2.52B params, 42 layers, hidden dim 2048)
Hardware: single RTX 3070 Ti (8GB) — all experiments run on consumer-grade hardware
Summary
We ask whether a small (2.5B-parameter) language model's hidden-state trajectory during generation contains a stable, low-dimensional structure that (a) is linearly decodable into task type… See the full description on the dataset page: https://huggingface.co/datasets/B2J/latent-state-tracking-minicpm5.minicpm5-sft3-3turn
minicpm5-sft3-3turn
Training / validation data for stage 1 (SFT) of the writing ladder behind
baiango/minicpm5-ul4b-story.
English fiction: each record is a templated writing instruction plus a
three-segment story distilled from teacher poolside/laguna-s-2.1
(3-turn chunked generation, automated quality gates).
Generation and training code for the full ladder: baiango/minicpm5-story-ladder.
Dataset summary
1,637 train / 40 valid records, one uniform schema across… See the full description on the dataset page: https://huggingface.co/datasets/baiango/minicpm5-sft3-3turn.
