openjev
Datasets
All datasets matching “openjev”Open-Jev
Open-Jev: typed decision datasets
Open-Jev turns a state and a question into a typed decision: a yes/no probability, a distribution over choices, independent label probabilities, or a discrete numeric/ordinal decision. This repository publishes twelve separate, frozen data configs from the Open-Jev project, together with original manifests, exact raw records, source code and reconstruction instructions.
These are controlled, mostly synthetic tasks and reference labels. They are… See the full description on the dataset page: https://huggingface.co/datasets/ZefanCai/Open-Jev.Open-Jev-v1.1
Open-Jev v1.1 data
This release publishes 326,619 typed decision records, including 147,139 training records, from the frozen community-hard-mix-v2-final mixture used for the new Open-Jev 27B training stage. It is a redistributable projection, not the exact complete training or evaluation dataset: 2,053 Wikispeedia records are omitted because a separate redistribution license for the archived graph/path data has not been verified.
Project website · Code · Previous data release ·… See the full description on the dataset page: https://huggingface.co/datasets/ZefanCai/Open-Jev-v1.1.openjev-mixture
OpenJev training mixture
279 classification and multiple-choice tasks, 323,466 rows, normalised into
one typed-decision format so a single model can be trained across all of
them. Assembled from tasksource plus
two curated ordinal datasets.
Built for OpenJev. The format is
plain JSONL, so nothing here requires that code.
Composition
primitive
tasks
what it is
choice
234
pick one of K options
noul
34
yes/no
score
11
one level on an ordered scale… See the full description on the dataset page: https://huggingface.co/datasets/s1lv3rj1nx/openjev-mixture.open-jev-laya-benchOpenJevData-140k
Dataset Card for OpenJevData-140k
Dataset Summary
OpenJevData-140k is a curated release of the data collection used to train OpenJev-4B. It contains 146,738 decision-making examples across 19 task categories, organized into SFT and RL splits.
Each example presents a state, a question, and a request-specific set of natural-language options. The data include hard answers and soft probability distributions, with candidate sets ranging from 2 to 77 options. Languages… See the full description on the dataset page: https://huggingface.co/datasets/shenjunhao/OpenJevData-140k.APUS-OpenJev-Eval-Frozen80
APUS-OpenJev Eval Frozen80
Private, fixed development evaluation panel: 80 decisions from five task families, 79 parent groups. This is the same panel previously used for official Jev, xDAN-openJet low/high, Laya, and checkpoint comparisons. This publication packages existing data without resampling, changing labels, candidate order, or input content.
This is not an independent test set. It has repeatedly informed debugging and checkpoint selection. Repeated runs measure… See the full description on the dataset page: https://huggingface.co/datasets/apus-ailab/APUS-OpenJev-Eval-Frozen80.
