datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
piperx-old-ab-rgbd-5903
PiperX Old AB RGB-D — 5,903 accepted points
This public dataset export contains the exact 5,903 accepted training Episodes recorded by the old AB SE(3) perturbation collector. The export is bound to the aggregate state.json truth and contains 5,903 unique schedule_index values across 124 closed shards (shard-00000 through shard-00123). Audit rejections remain audit records and are not training Episodes.
Data
Schema: piperx_lerobot_se3_rgbd_v2
Accepted Episodes /… See the full description on the dataset page: https://huggingface.co/datasets/Travor278/piperx-old-ab-rgbd-5903.eli5-human-vs-ai
ELI5 Human vs AI (long-form)
This dataset is for training and evaluating AI-writing detectors. It was built
as a clean way to compare known AI text against known human text: every human
answer predates ChatGPT by more than three years, so it is genuinely human by
construction, and every AI answer was written by a named 2026 model, so its origin
is certain too. Most detection datasets have to guess at their labels; this one
does not.
ELI5 answers were chosen because they are… See the full description on the dataset page: https://huggingface.co/datasets/mild-rgb/eli5-human-vs-ai.3d-dlp-repro-genericshapes-rgb
GenericShapes-RGB — synthetic RGB-voxel tabletop scenes
Training/evaluation corpus built for an independent reproduction of ICML 2026 paper #10351,
3D-DLP: Self-supervised 3D Object-centric Scene Representation Learning
(OpenReview vIotI25gJz, code
github.com/Eubooks3003/3d-dlp).
The paper's GenericShapes corpus (Appendix B.2) is described but not released, and the authors'
released generator scripts/generate_ply.py
writes colourless point clouds — the "RGB-coloured variant used… See the full description on the dataset page: https://huggingface.co/datasets/rvt832/3d-dlp-repro-genericshapes-rgb.aita-human-vs-ai
AITA Human-vs-AI corpus (2026 generators)
A second human-vs-AI corpus, a companion to
mild-rgb/eli5-human-vs-ai,
in a deliberately different register: first-person judgment narratives from
r/AmItheAsshole, versus same-title posts written by seven 2026 models. 2,900
questions, one human post and one AI post each; ~414 documents per generator.
The human side is redacted — reconstruct it from Scruples
The human posts are verbatim r/AmItheAsshole text, obtained via… See the full description on the dataset page: https://huggingface.co/datasets/mild-rgb/aita-human-vs-ai.fr3-peg-rlteacher-10k-rgb-native320-20261010
FR3 Peg-in-hole: RL-Teacher 10K (native 320x180, 10 Hz)
10,000 successful continuous simulated episodes in LeRobot Dataset v3.
This is state-teacher-generated RGB data, not a distilled RGB policy and not
MimicGen or real human demonstrations. Published as public with manual access
approval. Source generation finished on 2026-10-10 KST (Cube on October 9).
Data contract
637,745 synchronized rows at 10 Hz; three H.264 camera streams.
Views: side_a, side_b, gripper.… See the full description on the dataset page: https://huggingface.co/datasets/jaeikkim/fr3-peg-rlteacher-10k-rgb-native320-20261010.fr3-cube-rlteacher-10k-rgb-native320-20261010
FR3 Cube stacking: RL-Teacher 10K (native 320x180, 10 Hz)
10,000 successful continuous simulated episodes in LeRobot Dataset v3.
This is state-teacher-generated RGB data, not a distilled RGB policy and not
MimicGen or real human demonstrations. Published as public with manual access
approval. Source generation finished on 2026-10-10 KST (Cube on October 9).
Data contract
896,714 synchronized rows at 10 Hz; three H.264 camera streams.
Views: third_person_0… See the full description on the dataset page: https://huggingface.co/datasets/jaeikkim/fr3-cube-rlteacher-10k-rgb-native320-20261010.fr3-cube-rlteacher-10k-rgb-native320-sharedvis-v3-20261010
FR3 cube — rlteacher, shared visual v3, 10K
10,000 successful full episodes; 856,119 frames; LeRobot v3, actual 10 Hz, three natively rendered 320×180 views.
Default action is pde/eef-delta@v1: desired metric bare-fr3_hand delta in robot_base, xyz meters, axis-angle radians, grip 0=open/1=closed. Original controller actions remain in action.native_osc. Observation at t is paired with the command for t→t+1. The 22D state is joint9 + hand xyz/axis-angle6 + previous metric action7… See the full description on the dataset page: https://huggingface.co/datasets/jaeikkim/fr3-cube-rlteacher-10k-rgb-native320-sharedvis-v3-20261010.fr3-cube-mimicgen-10k-rgb-native320-sharedvis-v3-20261010
FR3 cube — mimicgen, shared visual v3, 10K
10,000 successful full episodes; 1,415,205 frames; LeRobot v3, actual 10 Hz, three natively rendered 320×180 views.
Default action is pde/eef-delta@v1: desired metric bare-fr3_hand delta in robot_base, xyz meters, axis-angle radians, grip 0=open/1=closed. Original controller actions remain in action.native_diik. Observation at t is paired with the command for t→t+1. The 22D state is joint9 + hand xyz/axis-angle6 + previous metric… See the full description on the dataset page: https://huggingface.co/datasets/jaeikkim/fr3-cube-mimicgen-10k-rgb-native320-sharedvis-v3-20261010.fr3-peg-mimicgen-10k-rgb-native320-sharedvis-v3-20261010
FR3 peg — mimicgen, shared visual v3, 10K
10,000 successful full episodes; 1,040,000 frames; LeRobot v3, actual 10 Hz, three natively rendered 320×180 views.
Default action is pde/eef-delta@v1: desired metric bare-fr3_hand delta in robot_base, xyz meters, axis-angle radians, grip 0=open/1=closed. Original controller actions remain in action.native_diik. Observation at t is paired with the command for t→t+1. The 22D state is joint9 + hand xyz/axis-angle6 + previous metric action7… See the full description on the dataset page: https://huggingface.co/datasets/jaeikkim/fr3-peg-mimicgen-10k-rgb-native320-sharedvis-v3-20261010.
