optimization
speculators-ci-datasets
speculator-tutorial
Raw vs. on-policy regenerated conversation data for training speculative-decoding
drafters (EAGLE-3 / DFlash / DSpark style), with the original source data kept alongside
so you can see exactly what regeneration changes and why it matters.
Prompts come from UltraChat-200k. The verifier / teacher model is Qwen/Qwen3-8B.
Why regenerate at all?
A speculative-decoding drafter is trained to predict what the verifier would say next.
If you train it… See the full description on the dataset page: https://huggingface.co/datasets/inference-optimization/speculators-ci-datasets.Alexandria_geometry_optimization_paths_PBE_2D
Cite this dataset Schmidt, J., Hoffmann, N., Wang, H., Borlido, P., Carriço, P. J. M. A., Cerqueira, T. F. T., Botti, S., and Marques, M. A. L. Alexandria geometry optimization paths PBE 2D. ColabFit, 2025. https://doi.org/10.60732/8781419f
This dataset has been curated and formatted for the ColabFit Exchange
This dataset is also available on the ColabFit Exchange:
https://materials.colabfit.org/id/DS_6pieq95jrqpn_0
Visit the ColabFit… See the full description on the dataset page: https://huggingface.co/datasets/colabfit/Alexandria_geometry_optimization_paths_PBE_2D.Alexandria_geometry_optimization_paths_PBE_3D
Cite this dataset Schmidt, J., Hoffmann, N., Wang, H., Borlido, P., Carriço, P. J. M. A., Cerqueira, T. F. T., Botti, S., and Marques, M. A. L. Alexandria geometry optimization paths PBE 3D. ColabFit, 2024. https://doi.org/10.60732/c88da7df
This dataset has been curated and formatted for the ColabFit Exchange
This dataset is also available on the ColabFit Exchange:
https://materials.colabfit.org/id/DS_s6gf4z2hcjqy_0
Visit the ColabFit… See the full description on the dataset page: https://huggingface.co/datasets/colabfit/Alexandria_geometry_optimization_paths_PBE_3D.QP-Benchmark
QP-Benchmark
Benchmark instances for Warm-IP,
an open-source solver for large-scale convex quadratic programs (QPs),
from the paper Warm-IP: A Path-Following ADMM Warm Start for
Interior-Point Quadratic Programming (Aslani, Tefagh, Jhanwar,
Zarepisheh; preprint link to be added). Every instance is a convex QP
minimize 0.5 x'Qx + q'x + c
subject to constraints (one-sided or two-sided; see below)
stored as an HDF5 (.h5) file, one folder per family:
MPC_data/ 64… See the full description on the dataset page: https://huggingface.co/datasets/Radiotherapy-Optimization/QP-Benchmark.tasktrove-hparam-optimization
TaskTrove hyperparameter optimization artifacts
This dataset is the campaign-level artifact archive for the TaskTrove hyperparameter and backend experiment. It
contains experiment configurations, reward and timing curves, evaluation records, checkpoint-cleanup records, the
A0 same-source base evaluation workspace, and the generated analysis website.
The experiment report, policy, and tracker remain in the parent experiment directory. See
PUBLISHED_ARTIFACTS.md for the separately… See the full description on the dataset page: https://huggingface.co/datasets/penfever/tasktrove-hparam-optimization.experimental-optimization
