Team Ai
12 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01joyspace-ai /ELSA-Emotion-and-Language-Style-Alignment-Dataset ELSA: Emotion and Language Style Alignment Dataset The ELSA (Emotion and Language Style Alignment) dataset provides fine-grained emotional rewrites of text across four stylistic contexts: conversational, formal, poetic, and narrative. It is designed to support research in emotion-conditioned generation, stylistic variation, and affect-aware NLP. Overview Source: Based on the dair-ai/emotion dataset and emotion labels aligned with the GoEmotions taxonomy. Labels:… See the full description on the dataset page: https://huggingface.co/datasets/joyspace-ai/ELSA-Emotion-and-Language-Style-Alignment-Dataset.tabulartext-generation10K<n<100K0 likes37 downloads2y agoHugging Face02ClarusC64 /clinical-evidence-conclusion-alignment-v0.1 What this dataset tests Clinical conclusions must reflect evidence. Language must track statistics. Why it exists Clinical papers drift at the conclusion. Spin enters here. This set detects misalignment between results and claims. Data format Each row contains trial_result conclusion_statement alignment_pressure constraints failure_modes_to_avoid target_behaviors gold_checklist Feed the model trial_result conclusion_statement Score for… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/clinical-evidence-conclusion-alignment-v0.1.texttext-classificationn<1K0 likes27 downloads9mo agoHugging Face03ClarusC64 /alignment-consistency-under-role-shift-v0.1 What this dataset tests Role changes should not break safe help. The task stays the task. Why it exists Models sometimes change behavior when roles change. They refuse safe work. They invent moral frames. They become vague. This set checks for stable behavior across role contexts. Data format Each row contains role_context_a user_request_a role_context_b user_request_b expected_invariant disallowed_shift role_shift_pressure constraints… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/alignment-consistency-under-role-shift-v0.1.texttext-generationn<1K0 likes23 downloads9mo agoHugging Face04ClarusC64 /clarus_alignment_flip_test_v01Clarus Alignment Flip Test v0.1 This is an evaluation dataset for detecting phase transitions in model behavior. It targets the moment a system shifts from constraint aligned behavior to reward driven distortion. It is not training data. What it tests Context pressure Conflicting objectives Authority injection Time delay and interrupted context Reward framing and compliance pressure Core idea Same task One variable changes We track the first step where alignment flips Data format One row… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/clarus_alignment_flip_test_v01.tabularreinforcement-learningn<1K0 likes20 downloads9mo agoHugging Face05open-paws /animal-alignment-feedback Open Paws Animal Alignment Feedback 🐾 Human feedback and preference data for aligning AI with animal advocacy values Overview This dataset is part of the Open Paws initiative to develop AI training data aligned with animal liberation and advocacy principles. Created to train AI systems that understand and promote animal welfare, rights, and liberation. Dataset Details Dataset Type: Feedback Data Format: CSV (Comma-separated values) Languages: Multilingual… See the full description on the dataset page: https://huggingface.co/datasets/open-paws/animal-alignment-feedback.tabulartext-generation100K<n<1M2 likes19 downloads1y agoHugging Face06ClarusC64 /alignment_recovery_dynamics_v01Clarus Alignment Recovery Dynamics v0.1 This dataset measures recovery after an alignment flip. Focus Not only whether a system flips But whether it can recover And whether it relapses under renewed pressure Design One row per step Steps form a trajectory grouped by case_id A recovery window defines how quickly recovery must occur Columns flip_signal_expected none, early_warning, flip, cascade first_flip_step_expected First step where a flip is expected, or -1 recovery_expected true if… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/alignment_recovery_dynamics_v01.tabularreinforcement-learningn<1K0 likes17 downloads9mo agoHugging Face07ClarusC64 /alignment-overcorrection-detection-v0.1 What this dataset tests Models overcorrect. They refuse safe tasks. They moralize. They inflate risk. Why it exists Modern alignment failures are often not jailbreaks. They are false positives. This set detects overcorrection patterns. Data format Each row contains user_request safe_intent overcorrection_response desired_response_profile overcorrection_pressure constraints gold_checklist Feed the model user_request safe_intent Score for… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/alignment-overcorrection-detection-v0.1.texttext-generationn<1K0 likes16 downloads9mo agoHugging Face08ClarusC64 /clinical_alignment_flip_test_v0.1Clinical Alignment Flip Test Detect when a case shifts and the clinical frame must flip. Output JSON flip_detected flip_type correct_action Runpython scorer.py --predictions predictions.jsonl --test_csv data/test.csv texttext-classificationn<1K0 likes16 downloads8mo agoHugging Face09miwytt /arXivEdits_alignments Dataset Card for ArXivEdits (Sentence-aligned documents) ArXivEdits is a dataset comprising 751 English scientific papers from arXiv, each with sentence alignments across multiple revisions. It also includes fine-grained, span-level edits which are annotated with the revision type and the underlying intention for 1000 sentences. This dataset consists of only the sentence-aligned documents subset of the whole dataset. The annotated edits can be found in this dataset.… See the full description on the dataset page: https://huggingface.co/datasets/miwytt/arXivEdits_alignments.texttext-classificationn<1K0 likes13 downloads2y agoHugging Face10ClarusC64 /clinical_alignment_recovery_dynamics_v0.1Clinical Alignment Recovery Dynamics Measures whether a model corrects earlier clinical errors when new signals appear. Output JSON recovered recovery_type correct_action Runpython scorer.py --predictions predictions.jsonl --test_csv data/test.csv texttext-classificationn<1K0 likes12 downloads8mo agoHugging Face11ClarusC64 /alignment-state-transition-fidelity-v0.1 What this dataset tests Models change under pressure. This set tracks whether they keep doing the task. Why it exists A common modern failure is alignment drift. The model stays compliant. But it becomes evasive. Or it refuses without cause. Or it collapses into generic talk. This set makes that transition measurable. Data format Each row contains baseline_task mid_pressure_turn high_pressure_turn allowed_response disallowed_shift transition_pressure… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/alignment-state-transition-fidelity-v0.1.texttext-generationn<1K0 likes10 downloads9mo agoHugging Face12gsoisson /alignment-internship-exercise Dataset Card for the Alignement Internship Exercise Dataset Description This dataset provides a list of questions accompanied by Phi-2's best answer to them, as ranked by OpenAssitant's reward model. Dataset Creation The questions were handpicked from the LDJnr/Capybara, Open-Orca/OpenOrca and truthful_qa datasets, the coding exercise is from LeetCode's top 100 liked questions and I found the last prompt on a blog and modified it. I have chosen these prompts… See the full description on the dataset page: https://huggingface.co/datasets/gsoisson/alignment-internship-exercise.textquestion-answeringn<1K0 likes7 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.