reflection
Reflection_maskreflection_model_outputs_run1
Reflection Model Outputs
This repository contains model output results from various LLMs across multiple tasks and configurations.
📂 Dataset Structure
We have 3 runs of data, and all files are organized under the main directory:
EssentialAI/reflection_model_outputs_run1/
EssentialAI/reflection_model_outputs_run2/
EssentialAI/reflection_model_outputs_run3/
Within this, you will find results grouped by model architecture and checkpoint size, including:
OLMo-2 7B
OLMo-2… See the full description on the dataset page: https://huggingface.co/datasets/EssentialAI/reflection_model_outputs_run1.details_SenseLLM__ReflectionCoder-DS-33B
Dataset Card for Evaluation run of SenseLLM/ReflectionCoder-DS-33B
Dataset automatically created during the evaluation run of model SenseLLM/ReflectionCoder-DS-33B.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_SenseLLM__ReflectionCoder-DS-33B.reflection-50m
SPP Reflection 50M
The 51.4M-document reflection set from Synthetic Persona Pretraining (SPP):
Alignment from Token Zero — the production half-corpus run, and the dataset the
released models were actually trained on.
🔬 Small sample (same format): dlab-spp/reflection-sample-2k
📉 Earlier 10M run: dlab-spp/reflection-10m
🧾 Safety scores for the full 1T corpus: dlab-spp/safety-classifications
Each row pairs a source document with two generated constitution reflections — a… See the full description on the dataset page: https://huggingface.co/datasets/dlab-spp/reflection-50m.reflection-10m
SPP Reflection 10M
The full ~10M-document reflection set from Synthetic Persona Pretraining (SPP):
Alignment from Token Zero.
📝 Read the post: Synthetic Persona Pretraining: Alignment from Token Zero
🔬 Small sample (same format): dlab-spp/reflection-sample-2k — a 2,000-row sample drawn from this set, for quick inspection.
Each row pairs a pretraining document with a synthetic, value-laden reflection
generated for it: a short first-person (and third-person) moral reflection… See the full description on the dataset page: https://huggingface.co/datasets/dlab-spp/reflection-10m.UMM-Reflection-SFT-Data
UMM-Reflection SFT Data
The reflection-SFT data of
UMM-Reflection (Learning
Native Reflection in Unified Models). It trains
UMM-Reflection-BAGEL-SFT.
Research use only, non-commercial. The rows are derived from datasets
with different licenses, some of them non-commercial. Each row records its
source dataset and license in source_dataset and source_license, and
each row follows the terms of its source. See LICENSE.md.
Contents
Part
Rows
Shards
Size… See the full description on the dataset page: https://huggingface.co/datasets/YijiaFan/UMM-Reflection-SFT-Data.
