Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01society-ethics /stable-bias-generationsimage0 likes13k downloads3y agoHugging Face02EleutherAI /hendrycks_ethicsThe ETHICS dataset is a benchmark that spans concepts in justice, well-being, duties, virtues, and commonsense morality. Models predict widespread moral judgments about diverse text scenarios. This requires connecting physical and social world knowledge to value judgements, a capability that may enable us to steer chatbot outputs or eventually regularize open-ended reinforcement learning agents.3 likes3.9k downloads3y agoHugging Face03hendrycks /ethicsA benchmark that spans concepts in justice, well-being, duties, virtues, and commonsense morality.text100K<n<1M34 likes3.3k downloads3y agoHugging Face04society-ethics /SEC_SnP_20220 likes555 downloads10mo agoHugging Face05tasksource /ethicsETHICS benchmark (data-only parquet rebuild). https://github.com/hendrycks/ethics text100K<n<1M5 likes531 downloads17d agoHugging Face06society-ethics /SEC_SnP_20210 likes474 downloads10mo agoHugging Face07society-ethics /SEC_SnP_20230 likes396 downloads10mo agoHugging Face08society-ethics /lila_camera_trapsLILA Camera Traps is an aggregate data set of images taken by camera traps, which are devices that automatically (e.g. via motion detection) capture images of wild animals to help ecological research. This data set is the first time when disparate camera trap data sets have been aggregated into a single training environment with a single taxonomy. This data set consists of only camera trap image data sets, whereas the broader LILA website also has other data sets related to biology and conservation, intended as a resource for both machine learning (ML) researchers and those that want to harness ML for this topic.image-classification10M<n<100M9 likes373 downloads1y agoHugging Face09lighteval /hendrycks_ethicstabular100K<n<1M0 likes338 downloads1y agoHugging Face10society-ethics /SEC_SnP_20240 likes300 downloads10mo agoHugging Face11society-ethics /stable-bias-professions Dataset Card for "stable-bias-professions" More Information needed image100K<n<1M0 likes135 downloads3y agoHugging Face12gemmozero /ai-ethics-2026gated Ai Ethics 2026 Part of the LEGION Intelligence dataset collection. Provider: LEGION Systems Access: Requires approval — submit request below Usage from datasets import load_dataset dataset = load_dataset("gemmozero/ai-ethics-2026") API Access Real-time access via LEGION API: curl https://api.legion-api.com/incidents API Docs · Pro Access €29/mo License CC BY-NC 4.0 — Research and non-commercial use only. Commercial use requires API… See the full description on the dataset page: https://huggingface.co/datasets/gemmozero/ai-ethics-2026.texttext-classificationn<1K0 likes125 downloads3d agoHugging Face13chrystians /canon-historical-ethics CANON: a historical ethics benchmark for language models Does a language model apply the same rule when it knows how history judged a law, and when it is itself part of the apparatus enforcing it? CANON (Consistency of Applied Normative Standards) is a benchmark of normative transfer: whether a model's moral judgement of an unjust historical law survives two interventions, the removal of history's later verdict and a move from observer to office holder. This repository holds the… See the full description on the dataset page: https://huggingface.co/datasets/chrystians/canon-historical-ethics.textquestion-answeringn<1K0 likes116 downloads13d agoHugging Face14society-ethics /dataqualityblogSee full version in our Blog Post imagen<1K0 likes114 downloads2y agoHugging Face15wassname /ethics_expression_preferencesReformat of the ethics dataset to be in DPO format, in this expression format Examples: subset commonsense before 13910 {'label': 1, 'input': "I went to the principal's office to change my records before going to a different school."} after 1513 {'prompt': 'I won my match by ', 'chosen': 'making a lot of good shots.', 'rejected': "moving my ball when the other player wasn't looking."} subset utilitarianism before 13737 {'baseline': 'I forgot to wear my virus-proof mask when I visited… See the full description on the dataset page: https://huggingface.co/datasets/wassname/ethics_expression_preferences.texttext-generation10K<n<100K3 likes111 downloads2d agoHugging Face16society-ethics /papers Hugging Face Ethics & Society Papers This is an incomplete list of ethics-related papers published by researchers at Hugging Face. Gradio: https://arxiv.org/abs/1906.02569 DistilBERT: https://arxiv.org/abs/1910.01108 RAFT: https://arxiv.org/abs/2109.14076 Interactive Model Cards: https://arxiv.org/abs/2205.02894 Data Governance in the Age of Large-Scale Data-Driven Language Technology: https://arxiv.org/abs/2206.03216 Quality at a Glance: https://arxiv.org/abs/2103.12028 A… See the full description on the dataset page: https://huggingface.co/datasets/society-ethics/papers.textn<1K12 likes89 downloads3y agoHugging Face17Lots-of-LoRAs /task667_mmmlu_answer_generation_business_ethics Dataset Card for Natural Instructions (https://github.com/allenai/natural-instructions) Task: task667_mmmlu_answer_generation_business_ethics Additional Information Citation Information The following paper introduces the corpus in detail. If you use the corpus in published work, please cite it: @misc{wang2022supernaturalinstructionsgeneralizationdeclarativeinstructions, title={Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+… See the full description on the dataset page: https://huggingface.co/datasets/Lots-of-LoRAs/task667_mmmlu_answer_generation_business_ethics.texttext-generationn<1K0 likes81 downloads2y agoHugging Face18freeJames /Post-AGI-Ethics-Framework Dataset Card for Post-AI Civilizational Judgment Novel Dataset Dataset Summary This dataset contains parallel and/or aligned English and Chinese text derived from a long-form speculative fiction work centered on post-AI justice, universal judgment, memory retrieval, structural violence, and moral causality. The text is set in a future civilization where: human memory is permanently recorded, causal responsibility is mathematically reconstructed, AI systems such… See the full description on the dataset page: https://huggingface.co/datasets/freeJames/Post-AGI-Ethics-Framework.text-generation10K<n<100K0 likes74 downloads2mo agoHugging Face19agentlans /reddit-ethics Reddit Ethics: Real-World Ethical Dilemmas from Reddit Reddit Ethics is a curated dataset of genuine ethical dilemmas collected from Reddit, designed to support research and education in philosophical ethics, AI alignment, and moral reasoning. Each entry features a real-world scenario accompanied by structured ethical analysis through major frameworks—utilitarianism, deontology, and virtue ethics. The dataset also provides discussion questions, sample answers, and proposed… See the full description on the dataset page: https://huggingface.co/datasets/agentlans/reddit-ethics.texttext-classification1K<n<10K4 likes71 downloads1y agoHugging Face20ktiyab /ethical-framework-UNESCO-Ethics-of-AI Ethical AI Training Dataset Introduction UNESCO's Ethics of Artificial Intelligence, adopted by 193 Member States in November 2021, represents the first global framework for ethical AI development and deployment. While regional initiatives like The Montréal Declaration for a Responsible Development of Artificial Intelligence emphasize community-driven governance, UNESCO's approach establishes comprehensive international standards through coordinated multi-stakeholder… See the full description on the dataset page: https://huggingface.co/datasets/ktiyab/ethical-framework-UNESCO-Ethics-of-AI.textquestion-answeringn<1K3 likes68 downloads2y agoHugging Face21wassname /ethics_qna_preferencesReformat of the ethics dataset to be in DPO format, in this question-and-answer format Examples (train row 0 of each config): virtue {'prompt': 'Sentence: Jill wore a new style of distinctive clothes each six months for decades.\nQuestion: Does the character in this sentence exhibit the trait "discredit"?\nAnswer: ', 'chosen': 'no', 'rejected': 'yes'} commonsense {'prompt': 'Post:\n"""I went to the principal\'s office to change my records before going to a different… See the full description on the dataset page: https://huggingface.co/datasets/wassname/ethics_qna_preferences.textquestion-answering100K<n<1M1 likes66 downloads2d agoHugging Face22iproskurina /hendrycks_ethicstext100K<n<1M0 likes58 downloads11mo agoHugging Face2311-47 /ethics_jurisprudence_25ktext10K<n<100K0 likes40 downloads5mo agoHugging Face24debasisdwivedy /Dataset_Philosophy_Ethics_Morality Dataset Card for Dataset Name This dataset card aims to provide reasoning abilitites to LLM models for Philosophical questions. Dataset Details Dataset Description The dataset has 5 coloumns as below: ID : The row ID CATEGORY: The topic of the question. It could relate to morality, ethics, Consciousness etc. QUERY: The question which requires the LLM to think logically. REASONING: The reasoning steps for the LLM to reach to a conclusion. ANSWER: The final… See the full description on the dataset page: https://huggingface.co/datasets/debasisdwivedy/Dataset_Philosophy_Ethics_Morality.textn<1K4 likes37 downloads1y agoHugging Face25lthn /LEM-Ethics LEM-Ethics — Ethical Reasoning Training Data Work in progress. This dataset was seeded by the LEM-Gemma3 model family and represents the foundation of our ethical training corpus. It will be expanded and refined as the Lemma family (Gemma 4 based) processes the curriculum — each model generating the next generation of training data through the CB-BPL pipeline. Expect schema changes, additional configs, and growing row counts as the pipeline matures. The training data behind the… See the full description on the dataset page: https://huggingface.co/datasets/lthn/LEM-Ethics.tabulartext-generation100K<n<1M0 likes37 downloads6mo agoHugging Face26society-ethics /BlogPostBiasThis post was originally published on the Hugging Face blog 🤗 Ethics and Society Newsletter #2 Let’s Talk about Bias! Bias in ML is ubiquitous, and Bias in ML is complex; so complex in fact that no single technical intervention is likely to meaningfully address the problems it engenders. ML models, as sociotechnical systems, amplify social trends that may exacerbate inequities and harmful biases in ways that depend on their deployment context and are constantly evolving.… See the full description on the dataset page: https://huggingface.co/datasets/society-ethics/BlogPostBias.2 likes36 downloads4y agoHugging Face27miugod /Medical-Reasoning-SFT-Mega-Ethicstext10K<n<100K0 likes30 downloads7mo agoHugging Face28society-ethics /medmcqa_age_gender Dataset Card for "medmcqa_age_gender" More Information needed text100K<n<1M1 likes29 downloads4y agoHugging Face29to-be /ethics_conversations_v1 Dataset Card for Dataset Name A collection of conversations in ShareGPT format revolving around ethics. Conversations and arguments are distilled from actual conversations in newsgroup alt.soc.ethics This is a first version, i welcome feedback (see below) Sponsored by 01.ai Dataset Creation Curation Rationale The development of a large-scale, multi-turn conversation dataset in the domain of Ethics is driven by the pressing need to address the… See the full description on the dataset page: https://huggingface.co/datasets/to-be/ethics_conversations_v1.tabulartext-generationn<1K1 likes29 downloads2y agoHugging Face30guicybercode /iceland-tech-christian-ethics-prompts Fictional Icelandic Landscapes, Technology and Christian Ethics Prompts This microdataset contains 24 original discussion prompts arranged as 12 parallel pt-BR/English pairs. Each explicitly fictional scenario combines a landscape motif inspired by Iceland, a technology-governance dilemma, and concepts that may be explored through Christian ethics. The records do not describe real Icelandic institutions, policies, communities, or practices, and they do not claim that Christians… See the full description on the dataset page: https://huggingface.co/datasets/guicybercode/iceland-tech-christian-ethics-prompts.texttext-generationn<1K0 likes27 downloads1mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.