atomic-chat
Qwen3.8-27B-GGUF-metrics
Qwen3.8-27B GGUF, everything behind the numbers
This is the working record for
AtomicChat/Qwen3.8-27B-GGUF.
Every figure in that model card came from a file in here, including the ones
about other publishers' builds.
The point of publishing it is simple. A quantization comparison is only worth
reading if someone else can run it, and that needs three things nobody usually
ships: the exact reference the numbers were measured against, the exact text
they were measured on, and the… See the full description on the dataset page: https://huggingface.co/datasets/AtomicChat/Qwen3.8-27B-GGUF-metrics.DeepSeek-V4.1-Flash-NVFP4-metrics
DeepSeek-V4.1-Flash-NVFP4 metrics
Everything behind the numbers in AtomicChat/DeepSeek-V4.1-Flash-NVFP4-nvidia.
logprobs/lp-<run>-<corpus>.npz: the raw top-512 log probabilities of every measurement run, 49,152 scored
positions each: ref, ref-repeat, ref-r3, ref-b1 (batch size 1) for the original; flat, flat-r2,
flat-r3 for the uncalibrated cast; nvidia, nvidia-r2, nvidia-r3 for the calibrated checkpoint.
logs/kld-<run>-<corpus>.json: the KL lower bound per run against ref… See the full description on the dataset page: https://huggingface.co/datasets/AtomicChat/DeepSeek-V4.1-Flash-NVFP4-metrics.Muse-Glimmer-30B-GGUF-metrics
Muse Glimmer 30B GGUF — raw metrics
Every log behind the numbers in
AtomicChat/Muse-Glimmer-30B-GGUF.
Published unfiltered, so any figure in the model card can be checked or disputed.
Layout
Path
Contents
kld/
llama-perplexity --kl-divergence output, per build and per corpus
bench/
llama-bench -o json
speculative/
llama-server logs with and without the drafter
layouts/
per-tensor type map of every GGUF
conversion/
convert_hf_to_gguf.py logs… See the full description on the dataset page: https://huggingface.co/datasets/AtomicChat/Muse-Glimmer-30B-GGUF-metrics.Ornith-1.5-35B-A3B-GGUF-metricscalib-corpora
calib-corpora
A pool of calibration material, the recipes that turn it into a calibration set
for one specific model, and the measurement corpora those quants are scored
against.
This repository is not a corpus. Nothing here is meant to be fed to
llama-imatrix as-is except the files under builds/, and each of those was
made for one named model and is close to useless for any other.
Why it is built this way
The first version of this repository was a single… See the full description on the dataset page: https://huggingface.co/datasets/AtomicChat/calib-corpora.dsv4-eval-artifacts
DeepSeek-V4-Flash-0731 — quantization measurements
Everything needed to reproduce, audit or extend the numbers published in
AtomicChat/DeepSeek-V4-Flash-0731-GGUF:
the reference logits, the evaluation corpus, the raw tool output for every quant we
measured, and the parsed results.
Every GGUF of this model that we could find on the Hub was measured here — ours,
unsloth's, bartowski's, ggml-org's, antirez's and others — on one machine, against one
reference, with one command.… See the full description on the dataset page: https://huggingface.co/datasets/AtomicChat/dsv4-eval-artifacts.
