datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
system-one-270m-data
system-one-270m-data
25,002 synthetic typed decisions: a piece of state, a question, a
caller-supplied option set, and a soft target distribution over those options.
Built to train kaivoss/system-one-270m,
an open take on the System One model class (TypeSafe
Jev,
Laya).
Schema
Field
Type
Meaning
prompt
string
the full rendered prompt, state + question + lettered options
letters
list[string]
the option letters in play, ["A", "B", ...]
target… See the full description on the dataset page: https://huggingface.co/datasets/kaivoss/system-one-270m-data.system-one-mini-data
System One Mini Synthetic Diagnostics
Deterministic synthetic controlled-intervention summaries used by DavidHatley/system-one-mini. The records contain no real coding-agent traces, private repositories, personal data, external model outputs, or teacher labels.
This dataset and model are independent research artifacts, not reproductions of Jev or RLCD.
Splits
Split
Rows
train
20,000
validation
2,000
calibration
2,000
development_renderer
2,000… See the full description on the dataset page: https://huggingface.co/datasets/DavidHatley/system-one-mini-data.system-one
Najd System One — public research draft
Decision tasks in English, MSA and Saudi Arabic. This draft is for research and debugging; independent linguistic and label review is pending.
Pack
Cases
License
controlled-development
600
CC-BY-4.0
controlled-validation
600
CC-BY-4.0
controlled-reserved_evaluation
1200
CC-BY-4.0
natural-development
216
CC-BY-4.0
references/arbanking77
1155
CC-BY-SA-4.0
references/massive
576
CC-BY-4.0
references/paired-tool-use
150… See the full description on the dataset page: https://huggingface.co/datasets/najdresearch/system-one.SystemOne
SystemOne-4B-Agentic-AGI Dataset
High-quality training & evaluation dataset for a hybrid System One + Agentic model.
This dataset is designed to teach and evaluate:
Fast, calibrated, typed decisions (Choice / Score / Noul)
Parallel multi-question evaluation
Confidence-aware behaviour
Agentic tool use, planning and reflection
Production-style scenarios (support, risk, routing, moderation, etc.)
Design Principles
Atomic & compositional – Prefer narrow… See the full description on the dataset page: https://huggingface.co/datasets/roskosmos19/SystemOne.system-one-eval
System One Eval
Structured State Reasoning Probe, v0.1.
This is a small, openly keyed diagnostic set, not a population benchmark or a blinded leaderboard. It contains 60 originally authored synthetic text tasks and 70 named questions covering policy precedence, cross-row aggregation, joins, boundaries, scheduling, access control, evidence limits, and multi-step arithmetic. There are no images or borrowed public-benchmark items in this package.
What is in the data… See the full description on the dataset page: https://huggingface.co/datasets/blazeofchi/system-one-eval.system-one-datasets
System One Datasets
Typed-decision datasets for System One models, normalized to the /v1/systemone wire format.
Every row is one typed decision (noul, choice, or score) whose state and question, once decoded, are
the body of a POST /v1/systemone request, the API served by TypeSafe's Jev and by open reimplementations such
as openjev. Use the rows for evaluation, calibration, regression tests, or training data selection.
This dataset is not affiliated with or endorsed by TypeSafe… See the full description on the dataset page: https://huggingface.co/datasets/zchee/system-one-datasets.
