datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
l0-qwen3-1.7b-compression-bs32-n16-32k-t1-no-eos-seqmean-verl091-146102-rollouts
RL training rollouts
l0_Qwen3-1.7B_compression_bs32_n16_32k_t1_no_eos_verl091_seqmean
One verified gzip JSONL shard per training step; 512 responses per shard.
Historical compression MathVerify after thinking, without an EOS gate.
f-cov-l4096-qwen3-1.7b-compression-bs32-n16-32k-t1-no-eos-seqmean-verl091-146102-rollouts
RL training rollouts
f_cov_l0_4096_no_eos_Qwen3-1.7B_compression_bs32_n16_32k_t1_seqmean_verl091
One verified gzip JSONL shard per training step; 512 responses per shard.
Historical compression MathVerify after thinking, without an EOS gate.
dl_alchemy_seq9p6m_context1024less-is-moe-s1-calibration-128-seq8192
Less-is-MoE S1K calibration data — 128 samples, seq_length 8192
This is the fixed calibration artifact used to prune GPT-OSS-120B,
Qwen3.5-122B-A10B, and the Gemma-4-26B-A4B causal language tower. It uses the same 128 source rows as the full-length variant:
yentinglin/s1K-1.1-trl-format revision
58a01564d278477da20ead1bcf1cde8e31f36251, train, followed by
Dataset.shuffle(seed=1234) and the first 128 nonempty messages rows.
For pruning, concatenate messages[].content with one… See the full description on the dataset page: https://huggingface.co/datasets/jayzou3773/less-is-moe-s1-calibration-128-seq8192.nca-paper-share20-seq_len_2048-657M
nca-paper-share20-seq_len_2048-657M
Procedurally generated Neural Cellular Automata trajectories (Lee et al. 2026), as flat uint16 token-id .bin files. Random NCA rules are rolled out on a 12×12 grid of 10 cell states and tokenized by 2×2 patches (base-10); only high-complexity rules survive a gzip-ratio filter (kept iff in (0.5, 1.0)). Token ids: 10,000 patch ids plus two grid delimiters (start=10000, end=10001); vocab = 10,002.
Configuration
param
value
grid
12×12… See the full description on the dataset page: https://huggingface.co/datasets/alexkstern/nca-paper-share20-seq_len_2048-657M.dyck-k128-seq_len_2048-1B
dyck-k128-seq_len_2048-1B
Procedurally generated k-shuffle Dyck bracket sequences (Hu et al. 2025, arXiv:2502.19249), as flat uint16 token-id .bin files. Token ids are 0-based: opening bracket type i is id i and its matching close is i + k, so ids span [0, 2k) and the vocabulary is 2k = 256.
Grammar parameters
param
value
k (bracket types)
128
max_depth
16
p_open
0.5
seq_length
2048
file
split
tokens
train.bin
train
999,999,488
val.bin
val
10,000… See the full description on the dataset page: https://huggingface.co/datasets/alexkstern/dyck-k128-seq_len_2048-1B.CultriX__SeQwence-14B-details
Dataset Card for Evaluation run of CultriX/SeQwence-14B
Dataset automatically created during the evaluation run of model CultriX/SeQwence-14B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/CultriX__SeQwence-14B-details.flowzap-sequence-workflows
sequence-workflows
A synchronized FlowZap template corpus with 242 canonical templates sourced from https://flowzap.xyz/sitemap-templates.xml and organized by primary Use Case.
Organization Model
Top-level folders are primary Use Cases from the FlowZap Templates dropdown.
Second-level folders preserve the original source domain from the FlowZap app index.
Each template keeps all matched Use Cases in metadata.json and the generated JSON/CSV indexes.
Templates that do not… See the full description on the dataset page: https://huggingface.co/datasets/Jules-OC/flowzap-sequence-workflows.grpo-qwen3-1.7b-taco-easy-3200-bs32-n8-seqs16-32k-146102-rollouts
grpo_Qwen3-1.7B_TACO-easy-3200_bs32_n8_seqs16_32k_1epoch rollouts
This dataset contains one compressed JSONL shard for every completed training
step. The step and rollout_index columns uniquely locate a rollout within
this training run. Run metadata and per-step row counts are recorded in
rollout_manifest.json.
sequelbox__Llama3.1-8B-PlumCode-details
Dataset Card for Evaluation run of sequelbox/Llama3.1-8B-PlumCode
Dataset automatically created during the evaluation run of model sequelbox/Llama3.1-8B-PlumCode
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sequelbox__Llama3.1-8B-PlumCode-details.modernbert_encoder_sp_seq_512_csedm_fold1CultriX__SeQwence-14B-v5-details
Dataset Card for Evaluation run of CultriX/SeQwence-14B-v5
Dataset automatically created during the evaluation run of model CultriX/SeQwence-14B-v5
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/CultriX__SeQwence-14B-v5-details.openvid-frame-sequences-1M
OpenVid Frame Sequences — 1M adjacent frame pairs
Short, single-shot frame sequences cut from OpenVid-1M,
built to train and evaluate models on what changes between two frames half a second apart.
One sample = 10 consecutive frames, 0.5 s apart (a 4.5 s span) → 9 adjacent frame pairs.
[f00] --0.5s--> [f01] --0.5s--> [f02] ... [f09]
^ the thing you describe / predict
Sequences
116,596
Frames per sequence
10 (0.5 s apart, t = 0.0 … 4.5 s)
Adjacent frame… See the full description on the dataset page: https://huggingface.co/datasets/junha1125/openvid-frame-sequences-1M.CultriX__SeQwence-14Bv1-details
Dataset Card for Evaluation run of CultriX/SeQwence-14Bv1
Dataset automatically created during the evaluation run of model CultriX/SeQwence-14Bv1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/CultriX__SeQwence-14Bv1-details.nca-paper-seq_len_1024-164M
nca-paper-seq_len_1024-164M
Procedurally generated Neural Cellular Automata trajectories (Lee et al. 2026), as flat uint16 token-id .bin files. Random NCA rules are rolled out on a 12×12 grid of 10 cell states and tokenized by 2×2 patches (base-10); only high-complexity rules survive a gzip-ratio filter (kept iff in (0.5, 1.0)). Token ids: 10,000 patch ids plus two grid delimiters (start=10000, end=10001); vocab = 10,002.
Configuration
param
value
grid
12×12
colors… See the full description on the dataset page: https://huggingface.co/datasets/alexkstern/nca-paper-seq_len_1024-164M.modernbert_encoder_sp_seq_512_dbe22kt_fold1qwen35-08b-seqlen-ablation-0919sequelbox__gemma-2-9B-MOTH-details
Dataset Card for Evaluation run of sequelbox/gemma-2-9B-MOTH
Dataset automatically created during the evaluation run of model sequelbox/gemma-2-9B-MOTH
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sequelbox__gemma-2-9B-MOTH-details.sequelbox__Llama3.1-8B-MOTH-details
Dataset Card for Evaluation run of sequelbox/Llama3.1-8B-MOTH
Dataset automatically created during the evaluation run of model sequelbox/Llama3.1-8B-MOTH
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sequelbox__Llama3.1-8B-MOTH-details.nca-paper-share10-seq_len_1024-164M
nca-paper-share10-seq_len_1024-164M
Procedurally generated Neural Cellular Automata trajectories (Lee et al. 2026), as flat uint16 token-id .bin files. Random NCA rules are rolled out on a 12×12 grid of 10 cell states and tokenized by 2×2 patches (base-10); only high-complexity rules survive a gzip-ratio filter (kept iff in (0.5, 1.0)). Token ids: 10,000 patch ids plus two grid delimiters (start=10000, end=10001); vocab = 10,002.
Configuration
param
value
grid
12×12… See the full description on the dataset page: https://huggingface.co/datasets/alexkstern/nca-paper-share10-seq_len_1024-164M.nca-paper-share10-seq_len_1024-164M-seed2
nca-paper-share10-seq_len_1024-164M-seed2
Procedurally generated Neural Cellular Automata trajectories (Lee et al. 2026), as flat uint16 token-id .bin files. Random NCA rules are rolled out on a 12×12 grid of 10 cell states and tokenized by 2×2 patches (base-10); only high-complexity rules survive a gzip-ratio filter (kept iff in (0.5, 1.0)). Token ids: 10,000 patch ids plus two grid delimiters (start=10000, end=10001); vocab = 10,002.
Configuration
param
value
grid… See the full description on the dataset page: https://huggingface.co/datasets/alexkstern/nca-paper-share10-seq_len_1024-164M-seed2.cleand_sequelbox_Celestia3-DeepSeek-R1-0528元データ: https://huggingface.co/datasets/sequelbox/Celestia3-DeepSeek-R1-0528
データ件数: 88,443
平均トークン数: 2143
最大トークン数: 31,680
合計トークン数: 189,577,005
ファイル形式: JSONL
ファイルサイズ: 812.4 MB
adaption-nist-biosafety-seq-bench
This dataset is a remastered version prepared using Adaption's Adaptive Data platform.
adaption-nist_biosafety_seq_bench
This dataset comprises structured biological sequence records designed as a strategic benchmark for training AI systems in biosafety, biosecurity, and synthetic biology governance. Each sample includes genomic data, organism identifiers, risk classification labels, and review status metadata to support pathogen detection and function prediction tasks. The… See the full description on the dataset page: https://huggingface.co/datasets/joduor/adaption-nist-biosafety-seq-bench.sequelbox__Llama3.1-8B-PlumChat-details
Dataset Card for Evaluation run of sequelbox/Llama3.1-8B-PlumChat
Dataset automatically created during the evaluation run of model sequelbox/Llama3.1-8B-PlumChat
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sequelbox__Llama3.1-8B-PlumChat-details.CultriX__SeQwence-14B-EvolMerge-details
Dataset Card for Evaluation run of CultriX/SeQwence-14B-EvolMerge
Dataset automatically created during the evaluation run of model CultriX/SeQwence-14B-EvolMerge
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/CultriX__SeQwence-14B-EvolMerge-details.nca-paper-share200-seq_len_2048-6.5Bsequelbox__Llama3.1-8B-PlumMath-details
Dataset Card for Evaluation run of sequelbox/Llama3.1-8B-PlumMath
Dataset automatically created during the evaluation run of model sequelbox/Llama3.1-8B-PlumMath
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sequelbox__Llama3.1-8B-PlumMath-details.adaption-ebolavirus-protein-sequences
This dataset is a remastered version prepared using Adaption's Adaptive Data platform.
adaption-ebolavirus_protein_sequences
This dataset contains amino acid sequences for seven key proteins from various Ebola and Marburg virus genomes, including strains like Zaire, Sudan, and Tai Forest. Each entry provides the protein identifier, name, strain information, and the full sequence intended for generating embeddings using models like ESM-2 or ProtT5. The collection includes major… See the full description on the dataset page: https://huggingface.co/datasets/joduor/adaption-ebolavirus-protein-sequences.ebolavirus_protein_sequences_INITIAL
This dataset is a remastered version prepared using Adaption's Adaptive Data platform.
adaption-ebolavirus_protein_sequences
This dataset contains amino acid sequences for seven key proteins from various Ebola and Marburg virus genomes, including strains like Zaire, Sudan, and Tai Forest. Each entry provides the protein identifier, name, strain information, and the full sequence intended for generating embeddings using models like ESM-2 or ProtT5. The collection includes major… See the full description on the dataset page: https://huggingface.co/datasets/joduor/ebolavirus_protein_sequences_INITIAL.sequelbox__Llama3.1-70B-PlumChat-details
Dataset Card for Evaluation run of sequelbox/Llama3.1-70B-PlumChat
Dataset automatically created during the evaluation run of model sequelbox/Llama3.1-70B-PlumChat
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sequelbox__Llama3.1-70B-PlumChat-details.
