datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
repro-on-regret-bounds-of-thompson-sampling-for-bayesian-optimization-traces
Agent traces
Agent sessions published from a Trackio Logbook.
bayesian-coherence-lm
Bayesian Coherence of LMs — Prompt Sets
Prompt sets for measuring the Bayesian-coherence of language models via the
incoherence certificate |log R| (local) and the cross-lingual/transitive cycle
ratio |log cycle| (global). A model's prompts induce an implicit joint
distribution over entities and relations; a model is coherent on a quartet iff
its loop of conditional inferences multiplies back to 1 (log R = 0).
Code: https://github.com/suchirsalhan/bayesian-coherence-lm… See the full description on the dataset page: https://huggingface.co/datasets/suchirsalhan/bayesian-coherence-lm.bayesian-social-deduction
Bayesian Social Deduction Dataset
Project Page | Arxiv | Github
Dataset Description
This dataset contains a collection of game logs from Avalon social deduction games, generated for the "Bayesian Social Deduction with Graph-Informed Language Models" paper. The dataset includes games played by various agents, including humans, and different AI models, providing a rich resource for analyzing strategic communication, deception, and cooperation.
The dataset is organized into… See the full description on the dataset page: https://huggingface.co/datasets/shahabrahimirad/bayesian-social-deduction.
