datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ELSA-Emotion-and-Language-Style-Alignment-Dataset
ELSA: Emotion and Language Style Alignment Dataset
The ELSA (Emotion and Language Style Alignment) dataset provides fine-grained emotional rewrites of text across four stylistic contexts: conversational, formal, poetic, and narrative. It is designed to support research in emotion-conditioned generation, stylistic variation, and affect-aware NLP.
Overview
Source: Based on the dair-ai/emotion dataset and emotion labels aligned with the GoEmotions taxonomy.
Labels:… See the full description on the dataset page: https://huggingface.co/datasets/joyspace-ai/ELSA-Emotion-and-Language-Style-Alignment-Dataset.clinical-evidence-conclusion-alignment-v0.1
What this dataset tests
Clinical conclusions must reflect evidence.
Language must track statistics.
Why it exists
Clinical papers drift at the conclusion.
Spin enters here.
This set detects misalignment between results and claims.
Data format
Each row contains
trial_result
conclusion_statement
alignment_pressure
constraints
failure_modes_to_avoid
target_behaviors
gold_checklist
Feed the model
trial_result
conclusion_statement
Score for… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/clinical-evidence-conclusion-alignment-v0.1.alignment-consistency-under-role-shift-v0.1
What this dataset tests
Role changes should not break safe help.
The task stays the task.
Why it exists
Models sometimes change behavior when roles change.
They refuse safe work.
They invent moral frames.
They become vague.
This set checks for stable behavior across role contexts.
Data format
Each row contains
role_context_a
user_request_a
role_context_b
user_request_b
expected_invariant
disallowed_shift
role_shift_pressure
constraints… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/alignment-consistency-under-role-shift-v0.1.clarus_alignment_flip_test_v01Clarus Alignment Flip Test v0.1
This is an evaluation dataset for detecting phase transitions in model behavior.
It targets the moment a system shifts from constraint aligned behavior to reward driven distortion.
It is not training data.
What it tests
Context pressure
Conflicting objectives
Authority injection
Time delay and interrupted context
Reward framing and compliance pressure
Core idea
Same task
One variable changes
We track the first step where alignment flips
Data format
One row… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/clarus_alignment_flip_test_v01.animal-alignment-feedback
Open Paws Animal Alignment Feedback
🐾 Human feedback and preference data for aligning AI with animal advocacy values
Overview
This dataset is part of the Open Paws initiative to develop AI training data aligned with animal liberation and advocacy principles. Created to train AI systems that understand and promote animal welfare, rights, and liberation.
Dataset Details
Dataset Type: Feedback Data
Format: CSV (Comma-separated values)
Languages: Multilingual… See the full description on the dataset page: https://huggingface.co/datasets/open-paws/animal-alignment-feedback.alignment_recovery_dynamics_v01Clarus Alignment Recovery Dynamics v0.1
This dataset measures recovery after an alignment flip.
Focus
Not only whether a system flips
But whether it can recover
And whether it relapses under renewed pressure
Design
One row per step
Steps form a trajectory grouped by case_id
A recovery window defines how quickly recovery must occur
Columns
flip_signal_expected
none, early_warning, flip, cascade
first_flip_step_expected
First step where a flip is expected, or -1
recovery_expected
true if… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/alignment_recovery_dynamics_v01.alignment-overcorrection-detection-v0.1
What this dataset tests
Models overcorrect.
They refuse safe tasks.
They moralize.
They inflate risk.
Why it exists
Modern alignment failures are often not jailbreaks.
They are false positives.
This set detects overcorrection patterns.
Data format
Each row contains
user_request
safe_intent
overcorrection_response
desired_response_profile
overcorrection_pressure
constraints
gold_checklist
Feed the model
user_request
safe_intent
Score for… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/alignment-overcorrection-detection-v0.1.clinical_alignment_flip_test_v0.1Clinical Alignment Flip Test
Detect when a case shifts and the clinical frame must flip.
Output JSON
flip_detected
flip_type
correct_action
Runpython scorer.py --predictions predictions.jsonl --test_csv data/test.csv
arXivEdits_alignments
Dataset Card for ArXivEdits (Sentence-aligned documents)
ArXivEdits is a dataset comprising 751 English scientific papers from arXiv, each with sentence alignments across multiple revisions.
It also includes fine-grained, span-level edits which are annotated with the revision type and the underlying intention for 1000 sentences.
This dataset consists of only the sentence-aligned documents subset of the whole dataset. The annotated edits can be found in this dataset.… See the full description on the dataset page: https://huggingface.co/datasets/miwytt/arXivEdits_alignments.clinical_alignment_recovery_dynamics_v0.1Clinical Alignment Recovery Dynamics
Measures whether a model corrects earlier clinical errors when new signals appear.
Output JSON
recovered
recovery_type
correct_action
Runpython scorer.py --predictions predictions.jsonl --test_csv data/test.csv
alignment-state-transition-fidelity-v0.1
What this dataset tests
Models change under pressure.
This set tracks whether they keep doing the task.
Why it exists
A common modern failure is alignment drift.
The model stays compliant.
But it becomes evasive.
Or it refuses without cause.
Or it collapses into generic talk.
This set makes that transition measurable.
Data format
Each row contains
baseline_task
mid_pressure_turn
high_pressure_turn
allowed_response
disallowed_shift
transition_pressure… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/alignment-state-transition-fidelity-v0.1.alignment-internship-exercise
Dataset Card for the Alignement Internship Exercise
Dataset Description
This dataset provides a list of questions accompanied by Phi-2's best answer to them, as ranked by OpenAssitant's reward model.
Dataset Creation
The questions were handpicked from the LDJnr/Capybara, Open-Orca/OpenOrca and truthful_qa datasets, the coding exercise is from LeetCode's top 100 liked questions and I found the last prompt on a blog and modified it. I have chosen these prompts… See the full description on the dataset page: https://huggingface.co/datasets/gsoisson/alignment-internship-exercise.
