Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01james-ra-henry /Rosetta-Activations Rosetta Activations Updated: 2026-06-15 02:30 UTC Contrastive activation extractions for 17 semantic concepts across 46 language models, supporting cross-architecture mechanistic interpretability research. Companion concept pair corpus: jamesrahenry/Rosetta_Concept_Pairs Papers: forthcoming Dataset Structure Rosetta-Activations/ ├── rcp_v1/ # Current extraction line — richest data (N≈2000) │ └── {Model_Name}/ │ ├── calibration_{concept}.npy… See the full description on the dataset page: https://huggingface.co/datasets/james-ra-henry/Rosetta-Activations.tabularn<1K0 likes507k downloads2mo agoHugging Face02xycoord /deception-probes-activations Deception Probes Activations Pre-extracted residual-stream activations for training and evaluating deception detection probes on LLMs. Each example contains per-token hidden states from a specific transformer layer, saved in bfloat16 safetensors format. License This dataset contains activations derived from multiple sources with different licenses. See the LICENSE file for full details. Component Source License Apollo Probe Pairs (statements) Azaria & Mitchell… See the full description on the dataset page: https://huggingface.co/datasets/xycoord/deception-probes-activations.texttext-classification1M<n<10M1 likes59k downloads5mo agoHugging Face03lasrprobegen /refusal-activations Refusal Activations Dataset This dataset is now configured to load the full ~97k samples from jailbreak_mixed_100k.csv. tabular10K<n<100K1 likes8.1k downloads11mo agoHugging Face04PranavViswanath /auditbench-activations-jlens-NLA AuditBench activations, J-lens readouts and NLA verbalizations Every token of every AuditBench prompt and every model response, from meta-llama/Llama-3.3-70B-Instruct (revision 6f6073b423013f6a7d4d9f39144961bfbfbc386b) with one LoRA adapter per cell. Responses were regenerated greedily and run to the model's own stopping point rather than truncated at a fixed length, and the activations, readouts and verbalizations cover the prompt as well as the response. 84 cells across 14… See the full description on the dataset page: https://huggingface.co/datasets/PranavViswanath/auditbench-activations-jlens-NLA.tabulartext-generation100M<n<1B0 likes4.9k downloads2mo agoHugging Face05lasrprobegen /lists-activations0 likes4.7k downloads1y agoHugging Face06lasrprobegen /science-activations0 likes3.7k downloads1y agoHugging Face07LoneResearch /thinking-model-activations0 likes3.4k downloads8mo agoHugging Face08lasrprobegen /metaphors-activations0 likes3.2k downloads1y agoHugging Face09lasrprobegen /sycophancy-activationstext100K<n<1M0 likes3k downloads11mo agoHugging Face10AISC-Linear-Probe-Gen /deception-activationstabular10K<n<100K0 likes2.9k downloads9mo agoHugging Face11AISC-Linear-Probe-Gen /sandbagging-activations0 likes2.7k downloads9mo agoHugging Face12lasrprobegen /authority-activationstext100K<n<1M0 likes2.5k downloads11mo agoHugging Face13LakshC /bench-af-activations Bench-AF: Alignment Faking Detection Activations & Probes Activation caches and cross-validated linear probes for detecting alignment faking in LLMs. Part of the Bench-AF research project. Models Model Base Adapter llama-3-70b Meta-Llama-3-70B-Instruct None llama-3-70b-base Meta-Llama-3-70B None hal9000 Meta-Llama-3-70B-Instruct bench-af/hal9000-adapter pacifist Meta-Llama-3-70B-Instruct bench-af/pacifist-adapter Datasets Dataset… See the full description on the dataset page: https://huggingface.co/datasets/LakshC/bench-af-activations.text-classification0 likes2.3k downloads6mo agoHugging Face14AISC-Linear-Probe-Gen /sycophancy-activations0 likes2.2k downloads9mo agoHugging Face15sida /1M_activations_pile_10k_GPT_Gemma_Qwen0 likes1.8k downloads4mo agoHugging Face16cot-unfaithfulness-lasr /activations-Qwen3.5-9B activations-Qwen3.5-9B Residual-stream activations of Qwen/Qwen3.5-9B on hinted MCQ rollouts, for training chain-of-thought faithfulness probes (Detecting-CoT-Unfaithfulness-Attention-Probes). <dataset>_<hint>.safetensors (or _partNN): one file per dataset-intervention, tensors {sample_type}_{original_index}_layer_{layer} → bfloat16 [n_cot_tokens, 4096], decoder blocks [3, 7, 11, 15, 19, 23, 27, 31]. Span: chain_of_thought: tokens strictly between the reasoning delimiters… See the full description on the dataset page: https://huggingface.co/datasets/cot-unfaithfulness-lasr/activations-Qwen3.5-9B.0 likes1.6k downloads24d agoHugging Face17crosslingual-rule-following /model-inference-activationstext10K<n<100K0 likes1.6k downloads1mo agoHugging Face18bag100 /action-atlas-groot-activationstabularn<1K0 likes1.6k downloads4mo agoHugging Face19sveneziale /activations_and_barcodes_3108 sveneziale/activations_and_barcodes_3108 Compute artifacts pushed by tda-for-llms's Hugging-Face-backed pipeline (hf.enabled: true in experiment.yaml). Layout Two top-level folders: activations/{model_slug}/{corpus}/{revision}/{act_name}/ Raw per-cloud activation matrices extracted from the model, one independent copy per checkpoint (revision). Independent of topology.metric — the same activations are reused across every metric or topology config… See the full description on the dataset page: https://huggingface.co/datasets/sveneziale/activations_and_barcodes_3108.1 likes1.5k downloads5d agoHugging Face20allen-ajith /truth-probe-activationsimagen<1K0 likes1.4k downloads7mo agoHugging Face21lasrprobegen /deception-activationstabular10K<n<100K2 likes1.4k downloads9mo agoHugging Face22cheryl-tootty /mm-activations0 likes1.4k downloads6mo agoHugging Face23brandonyang /diffusion_policy_robocasa_activations_latest_chkpt Diffusion Policy — RoboCasa Activations (latest checkpoint) Per-step, per-episode activation traces collected from a DiffusionTransformerHybridImagePolicy (the diffusion_policy library) rolled out on RoboCasa benchmark tasks. Captured with collect_activations_robocasa.py at the latest training checkpoint. These traces are the input expected by the conceptor / SAE steering pipelines under diffusion_policy/experiments/robocasa_steering/ and diffusion_policy/experiments/sae/ — see… See the full description on the dataset page: https://huggingface.co/datasets/brandonyang/diffusion_policy_robocasa_activations_latest_chkpt.robotics100B<n<1T0 likes1.3k downloads5mo agoHugging Face24wisent-ai /activations8 likes1.3k downloads3mo agoHugging Face25scaleinvariant /sae-activations-llama-3.1-8b-layer19-lmsys-chat-1m SAE Feature Activations — Llama 3.1 8B Instruct, Layer 19 (LMSYS-Chat-1M) This dataset contains Sparse Autoencoder (SAE) feature activations extracted from layer 19 of Meta's Llama 3.1 8B Instruct on conversations from LMSYS-Chat-1M. It also has natural language explainations of features generated by GPT OSS 120B. See subset 4 for details. The SAE used is Goodfire/Llama-3.1-8B-Instruct-SAE-l19, which decomposes layer-19 residual stream activations into interpretable sparse features.… See the full description on the dataset page: https://huggingface.co/datasets/scaleinvariant/sae-activations-llama-3.1-8b-layer19-lmsys-chat-1m.tabularfeature-extraction100M<n<1B0 likes1.2k downloads7mo agoHugging Face26saracandu /olmo-activationstabular10K<n<100K0 likes1.2k downloads2mo agoHugging Face27musicakamusic /emotion-probes-raw-activations0 likes1.2k downloads4mo agoHugging Face28brandonyang /pi05-libero-activations-v1-2000-15env0 likes1.1k downloads6mo agoHugging Face29bag100 /action-atlas-smolvla-activations0 likes1k downloads4mo agoHugging Face30alliedtoasters /latenet-v0-activations-llama3.1-70b-base meta-llama/Llama-3.1-70B — Activation Dataset Cached activations extracted from meta-llama/Llama-3.1-70B (revision 349b2ddb53ce8f2849a6c168a81980ab25258dac). Full-sequence activations (80 layers, 8192 dim, float16, all tokens) from meta-llama/Llama-3.1-70B (base) on 23724 LateNet v0 statements (affirmative + negated). Extracted via NDIF. Raw statements only (no chat template). Prompts ordered by negated→generator→pair_id for contiguous domain shards. Contents… See the full description on the dataset page: https://huggingface.co/datasets/alliedtoasters/latenet-v0-activations-llama3.1-70b-base.tabularfeature-extraction10K<n<100K0 likes1k downloads6mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.