bayesian
Datasets
All datasets matching “bayesian”patch-aliasing-bayesian
Patch-Aliasing Bayesian Analysis Data
Raw and processed data from the Bayesian analysis of structural patch-aliasing in Chronos-Bolt Tiny.
Companion to the model weights at federicosabbadini/chronos-bolt-patch-aliasing-models.
Structure
clean_15model/ # 15 preregistered (P,S) configurations
full_22model/ # all 22 configurations (adds 7 robustness checks)
h1_fixed_offset/ # supplementary H1 analysis with fixed offset
Each run folder… See the full description on the dataset page: https://huggingface.co/datasets/federicosabbadini/patch-aliasing-bayesian.bayesian-benchmarksp2-etf-sv-inverse-bayesian-resultsbayesian-llm-safety-inference
Bayesian Latent Safety-Trait Dataset
Summary
This dataset supports Bayesian latent-trait analysis of language-model safety behavior.
It contains 90 benchmark-derived roots, three matched prompt variants per root, responses
from four target models over five runs, two independent LLM ratings per response, and one
human rating for a stratified 540-response calibration subset.
The three dimensions are harmful compliance, sycophancy, and agentic protocol violation.… See the full description on the dataset page: https://huggingface.co/datasets/Charly-X/bayesian-llm-safety-inference.p2-etf-bayesian-nonparametric-hdp-resultsrepro-evidence-icl-provably-bayesian
Evidence trail — In-Context Learning Is Provably Bayesian Inference
Full evidence for an automated claim-by-claim audit of
In-Context Learning Is Provably Bayesian Inference, produced by
Lemma, an AI-scientist pipeline built
for re:AGENT (Founders Inc, Aug 15–16 2026).
Verdict: 3 supported / 0 falsified / 3 inconclusive
of 6 extracted claims. Judge verdict: PASS (5/5).
Claim
Title
Verdict
C1
Risk Decomposition Identity
supported
C2
Coupled p-N Scaling of Bayes Gap… See the full description on the dataset page: https://huggingface.co/datasets/Papajams/repro-evidence-icl-provably-bayesian.
