datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
system-one-decisions
Featuring Labeled Customer Emails and Support Responses
🔧 Synthetic IT Ticket Generator — Custom Dataset
Create a dataset tailored to your own queues & priorities (no PII).
👉 Generate custom data
Define your queues, priorities, language
Need an on-prem AI to auto-classify tickets?→ Open Ticket AI
There are 2 Versions of the dataset, the new version has more tickets, but only languages english and german. So please look at both files, to find what best fits… See the full description on the dataset page: https://huggingface.co/datasets/pngwn/system-one-decisions.systemone-lite-general
systemone-lite-general
Synthetic typed-decision rows for
systemone-lite
(letter-alias choice labels for causal LM SFT).
Not affiliated with TypeSafe AI / Jev. Labels are rule-based, not human prefs.
Splits
Split
Rows
Notes
train
32 400
Stratified mix of 3 gyms
test
3 600
iid held-out by task
test_hard
5 400
layout / paraphrase / option-subset shift
full
36 000
train + iid test
Gyms
TicketDungeon: ticket.route… See the full description on the dataset page: https://huggingface.co/datasets/dwidlee/systemone-lite-general.system-one-270m-data
system-one-270m-data
25,002 synthetic typed decisions: a piece of state, a question, a
caller-supplied option set, and a soft target distribution over those options.
Built to train kaivoss/system-one-270m,
an open take on the System One model class (TypeSafe
Jev,
Laya).
Schema
Field
Type
Meaning
prompt
string
the full rendered prompt, state + question + lettered options
letters
list[string]
the option letters in play, ["A", "B", ...]
target… See the full description on the dataset page: https://huggingface.co/datasets/kaivoss/system-one-270m-data.system-one-mini-data
System One Mini Synthetic Diagnostics
Deterministic synthetic controlled-intervention summaries used by DavidHatley/system-one-mini. The records contain no real coding-agent traces, private repositories, personal data, external model outputs, or teacher labels.
This dataset and model are independent research artifacts, not reproductions of Jev or RLCD.
Splits
Split
Rows
train
20,000
validation
2,000
calibration
2,000
development_renderer
2,000… See the full description on the dataset page: https://huggingface.co/datasets/DavidHatley/system-one-mini-data.systemone-lite-phase2
systemone-lite-phase2
Typed System One distill rows (task / state / instructions / criteria /
label_alias) for systemone-lite.
Revision (2026-09-26)
Paired with Hub model revision action-v2-qwen
(dwidlee/systemone-lite-0.5b).
Change
Detail
Chess
staged_v1 — piece + destination, option caps ≤8
Spatial
action_v2 — legal-only actions; Connect4 drop≤3 + win_now
Sokoban test
Deadlock alerts balanced 75/75 yes/no; remap_alert_prob≈0.35
Postmortem:… See the full description on the dataset page: https://huggingface.co/datasets/dwidlee/systemone-lite-phase2.system-one
Najd System One — public research draft
Decision tasks in English, MSA and Saudi Arabic. This draft is for research and debugging; independent linguistic and label review is pending.
Pack
Cases
License
controlled-development
600
CC-BY-4.0
controlled-validation
600
CC-BY-4.0
controlled-reserved_evaluation
1200
CC-BY-4.0
natural-development
216
CC-BY-4.0
references/arbanking77
1155
CC-BY-SA-4.0
references/massive
576
CC-BY-4.0
references/paired-tool-use
150… See the full description on the dataset page: https://huggingface.co/datasets/najdresearch/system-one.SystemOne
SystemOne-4B-Agentic-AGI Dataset
High-quality training & evaluation dataset for a hybrid System One + Agentic model.
This dataset is designed to teach and evaluate:
Fast, calibrated, typed decisions (Choice / Score / Noul)
Parallel multi-question evaluation
Confidence-aware behaviour
Agentic tool use, planning and reflection
Production-style scenarios (support, risk, routing, moderation, etc.)
Design Principles
Atomic & compositional – Prefer narrow… See the full description on the dataset page: https://huggingface.co/datasets/roskosmos19/SystemOne.system-one-eval
System One Eval
Structured State Reasoning Probe, v0.1.
This is a small, openly keyed diagnostic set, not a population benchmark or a blinded leaderboard. It contains 60 originally authored synthetic text tasks and 70 named questions covering policy precedence, cross-row aggregation, joins, boundaries, scheduling, access control, evidence limits, and multi-step arithmetic. There are no images or borrowed public-benchmark items in this package.
What is in the data… See the full description on the dataset page: https://huggingface.co/datasets/blazeofchi/system-one-eval.system-one-datasets
System One Datasets
Typed-decision datasets for System One models, normalized to the /v1/systemone wire format.
Every row is one typed decision (noul, choice, or score) whose state and question, once decoded, are
the body of a POST /v1/systemone request, the API served by TypeSafe's Jev and by open reimplementations such
as openjev. Use the rows for evaluation, calibration, regression tests, or training data selection.
This dataset is not affiliated with or endorsed by TypeSafe… See the full description on the dataset page: https://huggingface.co/datasets/zchee/system-one-datasets.
