datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
piper-checkpointsCheckpoints for Piper text to speech system.
checkpoint
Dataset Card for LLaVA-Video-178K
Uses
This dataset is used for the training of the LLaVA-Video model. We only allow the use of this dataset for academic research and education purpose. For OpenAI GPT-4 generated data, we recommend the users to check the OpenAI Usage Policy.
Data Sources
For the training of LLaVA-Video, we utilized video-language data from five primary sources:
LLaVA-Video-178K: This dataset includes 178,510 caption entries, 960,792 open-ended… See the full description on the dataset page: https://huggingface.co/datasets/YYF111/checkpoint.checkpointsofficeqa-checkpoint-eval-data
Checkpoint evaluation plot data
Snapshot: 2026-09-14T16:26:45.684890+00:00. Aggregate inputs to notes/Sept-2-2026.md performance figures.
No model execution, grading, publication, or source-result changes were performed to make this export.
Contents
checkpoint_evaluations: 454 checkpoint rows, one evaluation per run/iteration/protocol; score, mean output tokens, mean steps, and the existing two-sided 95% confidence bounds.
pareto_points: current mean-token/USD… See the full description on the dataset page: https://huggingface.co/datasets/YWZBrandon/officeqa-checkpoint-eval-data.k1rl-checkpointsunimaia-checkpoints
UniMaia historical checkpoint archive
This dataset repository is the reproducibility archive for intermediate
UniMaia experiment checkpoints. Model artifacts are released under GPL-3.0;
this does not relicense the separately distributed UniMaia, chessnets, or
chesseval software.
checkpoint_catalog.json inventories the retained 201 checkpoints from 70
experiment directories, totaling approximately 1.23 TiB before conversion.
The reproducibility retention policy keeps every… See the full description on the dataset page: https://huggingface.co/datasets/shermansiu/unimaia-checkpoints.ce-checkpointsssl-checkpoints
ssl-checkpoints
The code to load the checkpoints to follow...
The repository is organised as follows:
Each folder corresponds to the data set used for our experiment.
Each subfolder represents the corresponding SSL technique used.
These subfolders contain the checkpoints for each transformation/pretext task considered. The five checkpoint files correspond to
the transformation Baseline, SimClr, Orthogonality, LoRot and DCL, respectively, described in the blog.
rhan-checkpoints-rollingcheckpointsbbflow-checkpoints-transferICH-17-model-checkpointsrlcc-exp3-k1-k3-checkpoint-trajectoryfinetuning-checkpointsmlebench-lite-baseline-checkpointsourmethod_128batchsize_continue_continue_checkpointspost_train_ablate_removegan_checkpoint_20-80Krea-2-Turbo-Checkpoint-Format-Benchmark
Krea 2 Turbo ComfyUI Format Fidelity Benchmark
This release is a paired, deterministic comparison of eight Krea 2 Turbo checkpoint formats in ComfyUI: BF16, FP8 Scaled, INT8 ConvRot, MXFP8, NVFP4, INT4 ConvRot W4A4, GGUF Q8_0, and GGUF Q4_K_M. It contains 240 scored 1024×1024 images, saved float32 decoded tensors and final latents, every denoising trajectory, raw metric tables, telemetry, statistical comparisons, and reproduction code.
Main result
BF16 is the… See the full description on the dataset page: https://huggingface.co/datasets/Merserk/Krea-2-Turbo-Checkpoint-Format-Benchmark.Molab-checkpointsDnD-checkpoints-and-logsThis is the pretrained checkpoints and logs for All DnD experiments! 🎉🎉🎉
This repo includes common sense reasoning, math, coding and multimodal tasks, along with ablation studies and exploration experiments.
You can also refer to the online file organizing each experiment's results here.
We highly appreciate your enthusiasm for DnD and we are constantly improving it for the better. 🤗🤗
marl_checkpointfile.checkpointsopenvla-checkpointstest
routing_analysis-checkpoints
routing_analysis checkpoint archive
This public dataset repository stores checkpoint files from the
routing_analysis filesystem snapshot while preserving their original paths
under routing_analysis/.
The tree routing_analysis/finetuning/phase2_full/checkpoints/ is explicitly
excluded. All other regular files classified under checkpoint directories are
included, including small code and configuration files needed to keep those
checkpoint directories complete.
Files are uploaded… See the full description on the dataset page: https://huggingface.co/datasets/lylybig/routing_analysis-checkpoints.openevo-q03-native-checkpoint-recovery-20261005
旧 Q03 原生训练存档 / Historical Q03 native training checkpoints
这是旧 Q03 SEED-style 实验 Stream A 的六套原生存档39/40/79/80/119/120。包含原始模型权重、优化器数值和模型配置,126文件,131,428,022,496B;原始字节未重写。它们用于研究与恢复历史训练状态。Q03具有本地Stage1和OpenEVO实验条件;此包不宣称原论文精确复现、A160完成或PR677新实验结果。
本仓只公开模型/优化器数值。原始随机数、加载器和环境状态60文件936,704B保留在获授权的私有恢复档案;公开包单独不能恢复原始完整随机/数据状态。原始任务、轨迹、teacher分析、Memory/Skill/Agent材料、评估与final panel、日志、控制面和凭据都不在本包。
恢复方法 / Restore
以manifest.json逐文件SHA256和size为准,从不可变HF… See the full description on the dataset page: https://huggingface.co/datasets/openevo-recovery/openevo-q03-native-checkpoint-recovery-20261005.weact-native-science-checkpoint
Native WeAct science checkpoint
The fixed eight-question, seven-arm pilot has reached its human-review checkpoint. All 56 planned task records are in pilot_v4/. The frozen 3,000-question test has not been run. No accuracy score or retraining conclusion is claimed.
The independent Serper, Jina and E2B checks passed, and each backend passed native webpage extraction with the live auxiliary model. The corrected runtime uses original questions, native hard/soft routing, the… See the full description on the dataset page: https://huggingface.co/datasets/Corning/weact-native-science-checkpoint.roberta-pt-checkpointsrhan-nxa-checkpoints-rollingruleforge-checkpointsrhan-checkpoints
