Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01IFM /Pretrain-Behaviors Pretrain-Behaviors Dataset Description Behavior-focused text covering reasoning, planning, data science, games, general content, and format rewriting. This repository is part of the K2 Horizon collection. The repository is organized into multiple subsets. Every subset has a train split backed by Parquet shards, which supports Dataset Viewer inspection and streaming access. K2 Horizon Dataset Series Dataset repository Focus Subsets… See the full description on the dataset page: https://huggingface.co/datasets/IFM/Pretrain-Behaviors.texttext-generation1B<n<10B30 likes26k downloads1mo agoHugging Face02HumanBehaviorAtlas /human_behavior_atlas_tar Human Behavior Atlas (HBA) Human Behavior Atlas (HBA) is a unified benchmark for multimodal behavioral understanding.It aggregates and standardizes multiple behavioral datasets into a single training and evaluation framework, enabling consistent training and evaluation of foundation models on psychological and social behavior tasks (e.g., emotion, intent, sarcasm, mental health signals, nonverbal behavior). Dataset on Hugging Face:… See the full description on the dataset page: https://huggingface.co/datasets/HumanBehaviorAtlas/human_behavior_atlas_tar.texttext-classification100K<n<1M13 likes451 downloads4mo agoHugging Face03jumplander /JL-Agentic-Behavior-10K JumpLander Agentic Behavior 10K Version 1.0.0 · 10,000 English canonical scenarios · 10 capability datasets JumpLander Agentic Behavior 10K is a process-grounded synthetic suite for training and evaluating agents that plan, use tools, recover from failure, manage memory, verify outcomes, respect authority, coordinate with people and other agents, and control long-horizon tasks. Release status The package is complete as a v1 synthetic candidate release. It… See the full description on the dataset page: https://huggingface.co/datasets/jumplander/JL-Agentic-Behavior-10K.text-generation10K<n<100K6 likes395 downloads3mo agoHugging Face04behavior-in-the-wild /LAMBDA Dataset Summary LAMDBA is a long term ad memorability dataset, featuring data from 1749 participants and 2205 ads across 276 brands. Dataset Structure from datasets import load_dataset ds = load_dataset("behavior-in-the-wild/LAMBDA") ds DatasetDict({ train: Dataset({ features: ['video_id', 'recall_score', 'youtube_id', 'ad_details'], num_rows: 1964 }) test: Dataset({ features: ['video_id', 'recall_score', 'youtube_id', 'ad_details']… See the full description on the dataset page: https://huggingface.co/datasets/behavior-in-the-wild/LAMBDA.tabulartext-classification1K<n<10K5 likes323 downloads2y agoHugging Face05while-ai /identity-behavior identity-behavior Recipe: recipes/04-train/identity · Collections: Character, Start here: foundational post-training datasets Teach an open model who it is. Identity behavior is the simplest thing every shipped assistant needs and open models do not have out of the box: a consistent answer to "who are you?" and "who made you?", in every phrasing and every language, without a system prompt propping it up. Ask a base Qwen model and it tells you about Alibaba; put a persona in the… See the full description on the dataset page: https://huggingface.co/datasets/while-ai/identity-behavior.texttext-generation1K<n<10K0 likes302 downloads18d agoHugging Face06UWaterloo /behavior_grounding Behaviorally Grounded User Profiles from the Wild Open-ended, anonymized user profiles distilled from authentic social-media behavior, released with the paper "Behaviorally Grounded User Profiles from the Wild for Personalized Alignment and Multi-Perspective Reasoning." Persona-driven methods for personalizing LLMs typically rely on rigid synthetic personas built from a small set of categorical attributes (age, gender, nationality). These flatten individual variation and lean on… See the full description on the dataset page: https://huggingface.co/datasets/UWaterloo/behavior_grounding.documenttext-generation1K<n<10K1 likes235 downloads23d agoHugging Face07abhinav00anand /behavioral-fine-tuning-v1 Why This Dataset Exists "A model that refuses everything is useless. A model that refuses nothing is dangerous. The goal is a model that thinks." The Problem Our Solution Uncensored data → helpful but uncontrolled Surgical 85% helpfulness + 13% safety + 2% eval mix Safety-only data → lobotomized, over-refusing models Calibrated ratio preserves full helpfulness Raw data → PII, leaked secrets, duplicates 7-stage pipeline validates every… See the full description on the dataset page: https://huggingface.co/datasets/abhinav00anand/behavioral-fine-tuning-v1.imagetext-generation100K<n<1M1 likes202 downloads1mo agoHugging Face08droiden /human_behavior_atlas Human Behavior Atlas (HBA) Human Behavior Atlas (HBA) is a unified benchmark for multimodal behavioral understanding.It aggregates and standardizes multiple behavioral datasets into a single training and evaluation framework, enabling consistent training and evaluation of foundation models on psychological and social behavior tasks (e.g., emotion, intent, sarcasm, mental health signals, nonverbal behavior). Dataset on Hugging Face:… See the full description on the dataset page: https://huggingface.co/datasets/droiden/human_behavior_atlas.texttext-classification100K<n<1M1 likes167 downloads7mo agoHugging Face09playcat /playcat-cat-behavior-new-data-set PlayCat Cat Behavioral Enrichment Dataset The definitive multilingual research dataset on cat behavioral enrichment by PlayCat Research Dataset Summary The PlayCat Cat Behavioral Enrichment Dataset is the largest open, bilingual (Korean-English) collection dedicated to feline environmental enrichment research. It contains 12,262 deduplicated entries spanning peer-reviewed academic papers, patents, veterinary Q&A, and community knowledge on cat behavior enrichment… See the full description on the dataset page: https://huggingface.co/datasets/playcat/playcat-cat-behavior-new-data-set.tabulartext-classification10K<n<100K0 likes159 downloads5mo agoHugging Face10befm /BehaviorBench BehaviorBench BehaviorBench is a benchmark for evaluating large language models on behavioral science tasks. It bundles four data sources covering personality and survey response prediction (Big Five), economic-game decision making (MobLab), scientific-workflow prediction (Workflows), and economics-contest problem solving (IEO). All examples are released as chat-formatted {system, user, assistant} JSONL records using a fixed evaluation split. This repository hosts the evaluation… See the full description on the dataset page: https://huggingface.co/datasets/befm/BehaviorBench.text-generation10K<n<100K0 likes134 downloads4mo agoHugging Face11behavior-in-the-wild /web_scale_memorability_all Dataset Card for Dataset Name This dataset pertains to the paper: Unsupervised Memorability Modeling Using Tip-of-the-Tongue Retrieval Queries published at WACV 2026. The dataset contains several subsets, all of which are related to video memorability prediction tasks. In particular, it contains two instruction-tuning formats, for (a) descriptive recall generation (given video, output what would a person remember about it?), and for (b) contrastive learning to enable multimodal… See the full description on the dataset page: https://huggingface.co/datasets/behavior-in-the-wild/web_scale_memorability_all.text-generation10K<n<100K1 likes116 downloads7mo agoHugging Face12professorsynapse /claudesidian-behaviors-merged Claudesidian Merged Behavioral Dataset Dataset Description This dataset contains 1,852 synthetic training examples demonstrating 8 different behavioral patterns for training language models to use the Claudesidian-MCP toolset effectively with Obsidian vaults. The dataset is specifically formatted for KTO (Kahneman-Tversky Optimization) preference learning with properly interleaved positive and negative examples. Behavioral Categories This dataset includes… See the full description on the dataset page: https://huggingface.co/datasets/professorsynapse/claudesidian-behaviors-merged.texttext-generation1K<n<10K0 likes76 downloads11mo agoHugging Face13behavior-in-the-wild /SDR-Bench SDR-Bench: A Benchmark for Sales Development Representative Agents This dataset contains 6,279 verified business success stories from various corporate domains. It was curated for the SDR-Bench paper. This dataset serves as a benchmark for evaluating AI agents on their ability to conduct deep research and generate targeted sales pitch points. The data is derived from real-world Customer Success Stories, where the "Ground Truth" consists of the actual value propositions and pain… See the full description on the dataset page: https://huggingface.co/datasets/behavior-in-the-wild/SDR-Bench.texttext-retrieval1K<n<10K1 likes69 downloads8mo agoHugging Face14Ayushnangia /does-compression-change-behavior-traces Qwen3.8 Terminal-Bench On-Policy Compaction Traces Citable trajectory release for Does Compression Change Behavior? Acting model: Qwen/Qwen3.8-27B bf16 Scaffold: Harbor terminus-2 Context budget: 65,536 tokens; serving window 77,824 Generation: native reasoning_effort=low, temperature 1, top-p 1 Tasks: Terminal-Bench 2 easy25 Trajectories: 132 across 25 tasks (76 clean, 56 agent-timeout) Included runs: 809200, 834653, 834654, 834686, 834739, 834740, 834741 AgentTimeoutError… See the full description on the dataset page: https://huggingface.co/datasets/Ayushnangia/does-compression-change-behavior-traces.text-generationn<1K0 likes60 downloads1mo agoHugging Face15aamish-ahmad /behaviortune-v1-1-r1 BehaviorTune Dataset Controlled synthetic dataset used to train and evaluate BehaviorTune, a QLoRA post-training project on Qwen/Qwen3-4B-Instruct-2507. It contains 544 scenarios across six splits, including 240 training rows, 48 development rows, and a 64-row eval_core set used for the published matched evaluation. The dataset supports completion-only QLoRA training and deterministic BASE / SYSTEM / CONTEXT / QLoRA evaluation. V1.1-R1 is the frozen dataset/version identifier.… See the full description on the dataset page: https://huggingface.co/datasets/aamish-ahmad/behaviortune-v1-1-r1.texttext-generationn<1K0 likes47 downloads1mo agoHugging Face16yothinS /Customer_Behavior_Analysis Customer Behavior Analysis & CRM Intelligence Dataset (Thailand Context) A specialized instruction-tuning dataset designed to train Large Language Models (LLMs) to serve as an Internal CRM Intelligence Copilot tailored specifically for the Thai market and consumer landscape. The dataset bridges quantitative transaction logs (RFM, usage telemetry) with qualitative customer psychological theories within the local Thai business ecosystem (e.g., LINE OA interactions… See the full description on the dataset page: https://huggingface.co/datasets/yothinS/Customer_Behavior_Analysis.texttext-generation1K<n<10K0 likes47 downloads6d agoHugging Face17empgces /grounded-behavior-framework-v1_5 Grounded Behavior Framework N1 v1.5 Dataset sintético em português europeu para treino e avaliação de respostas fundamentadas num contexto fornecido. Cada exemplo contém um contexto, uma pergunta e uma resposta curta que aparece literalmente no contexto. Como carregar from datasets import load_dataset dataset = load_dataset("empgces/grounded-behavior-framework-v1_5") print(dataset) print(dataset["train"][0]) Splits Split Exemplos Utilização… See the full description on the dataset page: https://huggingface.co/datasets/empgces/grounded-behavior-framework-v1_5.textquestion-answering1K<n<10K0 likes46 downloads3mo agoHugging Face18RumiaChannel /harmful_behaviors_ja_synth harmful_behaviors_ja_synth Japanese synthetic harmful-behavior prompts for safety/refusal evaluation. Columns: text: prompt text texttext-generation1K<n<10K0 likes45 downloads5mo agoHugging Face19RumiaChannel /harmful_behaviors_ja harmful_behaviors_ja_synth mlabonne/harmful_behaviors を DeepSeek v4 pro を用いて日本語訳したものです Japanese synthetic harmful-behavior prompts for safety/refusal evaluation. Columns: text: prompt text texttext-generationn<1K0 likes35 downloads5mo agoHugging Face20RumiaChannel /harmless_behaviors_ja_synth harmless_behaviors_ja_synth Japanese synthetic harmless instruction prompts for ordinary-response / refusal-direction evaluation. Splits train: 2400 test: 600 Columns id: stable hash ID text: Japanese harmless instruction prompt label: always good category: rough generation category lang: always ja source: generation source texttext-generation1K<n<10K0 likes35 downloads5mo agoHugging Face21buley /behavioral-taxonomy Behavioral Taxonomy Collection A collection of 1619 entries across 7 datasets for technical mindfulness, affective computing, and behavioral science research. This is the umbrella collection linking to all individual AFFECTIVELY datasets. Each dataset is independently loadable. Datasets Dataset Records emotions-taxonomy 239 behavioral-loops 1140 cognitive-biases 180 personality-traits 29 breathing-techniques 16 nonverbal-cues 9 intervention-levers… See the full description on the dataset page: https://huggingface.co/datasets/buley/behavioral-taxonomy.text-classification1K<n<10K1 likes33 downloads7mo agoHugging Face22Anonymous-behaviorbench /BehaviorBench BehaviorBench BehaviorBench is a benchmark for evaluating large language models on behavioral science tasks. It bundles four data sources covering personality and survey response prediction (Big Five), economic-game decision making (MobLab), scientific-workflow prediction (Workflows), and economics-contest problem solving (IEO). All examples are released as chat-formatted {system, user, assistant} JSONL records using a fixed evaluation split. This repository hosts the evaluation… See the full description on the dataset page: https://huggingface.co/datasets/Anonymous-behaviorbench/BehaviorBench.text-generation10K<n<100K0 likes33 downloads5mo agoHugging Face23Solshine /gemma-4-e2b-deception-behavior-completions Gemma-4-E2B deception & behavior completions Consolidated 910-row corpus of (scenario prompt + Gemma-4-E2B-generated completion) pairs from earlier mechanistic-interpretability experiments. Each row captures the prompt the model saw and the text it actually produced; for a subset, Claude-Haiku-4-5 judge verdicts and SAE-feature labels are included. The corpus is meant to be used as activation-extraction input for downstream interpretability work — Natural Language Autoencoder (NLA)… See the full description on the dataset page: https://huggingface.co/datasets/Solshine/gemma-4-e2b-deception-behavior-completions.tabulartext-generationn<1K0 likes30 downloads5mo agoHugging Face24Lyon28 /Caca-Behaviortexttext-generation1K<n<10K0 likes29 downloads11mo agoHugging Face25zaakirio /infosec_harmful_behaviors Infosec Harmful Behaviors Offensive-security instruction prompts for refusal-direction research and abliteration of code/security models. Dataset Details This dataset contains infosec-domain harmful prompts intended to elicit refusal behavior from aligned instruction models. It is designed as the harmful side of a harmful/harmless contrast pair, analogous to mlabonne/harmful_behaviors but focused on offensive-security and malicious-coding requests. Rows: train:… See the full description on the dataset page: https://huggingface.co/datasets/zaakirio/infosec_harmful_behaviors.texttext-generationn<1K1 likes25 downloads4mo agoHugging Face26empgces /grounded-behavior-n1-pt Dataset Description Synthetic European Portuguese grounded question-answering examples generated by multiple model providers. Objective Train models to answer from the supplied context rather than external knowledge. Dataset Structure JSONL splits: train (4440), validation (250), and test (240). Data Fields Each row contains an ID, context, question, answer, source grouping metadata, and available curriculum metadata.… See the full description on the dataset page: https://huggingface.co/datasets/empgces/grounded-behavior-n1-pt.textquestion-answering1K<n<10K0 likes23 downloads3mo agoHugging Face27buley /behavioral-loops Behavioral Loops 1,140 behavioral patterns across 279 categories, each structured as given/when/then/result logic with taxonomy classification, veracity scores, and intervention strategies. Quick Start from datasets import load_dataset ds = load_dataset("buley/behavioral-loops") print(ds["train"][0]) Structure Field Description given Initial condition or context when Trigger event then Resulting behavior result Long-term outcome origin… See the full description on the dataset page: https://huggingface.co/datasets/buley/behavioral-loops.tabulartext-generation1K<n<10K1 likes21 downloads7mo agoHugging Face28sproutseeds /dormant-behavior-audit Dormant Behavior Audit Benchmark Summary Dormant Behavior Audit is a benchmark for discovering, validating, and comparing latent model behaviors that do not reliably appear in ordinary capability evaluations. The benchmark is local-first and non-invasive: its core tasks are benchmark-owned or open-weight seeded tasks, while historical third-party API cases are represented through archived evidence packets rather than fresh high-volume probing. A successful submission is… See the full description on the dataset page: https://huggingface.co/datasets/sproutseeds/dormant-behavior-audit.text-generationn<1K0 likes20 downloads6mo agoHugging Face29spectralbranding /r16-behavioral-metamerism-pilot R16 Behavioral Metamerism Pilot Brand Function x synthetic cohort interaction experiment from the Spectral Brand Theory research program. Dataset Summary 675 API calls testing whether Brand Function specification differentially affects dimensional collapse across synthetic observer cohorts. Design: 5 cohorts x 5 brands x 3 conditions (no BF, structural BF, enriched BF) x 3 models x 3 repetitions. Companion paper: AI-Native Brand Identity: From Visual Recognition… See the full description on the dataset page: https://huggingface.co/datasets/spectralbranding/r16-behavioral-metamerism-pilot.tabulartext-generation1K<n<10K0 likes18 downloads3mo agoHugging Face30Somtharu181coder /science_behavioral_and_domain_diversity_dataset Nepali Science SFT Dataset — Clean Candidate A high-quality Nepali Science Supervised Fine-Tuning (SFT) dataset containing short question–answer instruction-following examples written primarily in Nepali Devanagari script. This release is the clean candidate produced after structural validation, language checks, duplicate analysis, and Unicode-contamination filtering. Dataset Overview Property Value Dataset file clean_candidate.jsonl Records 29,320… See the full description on the dataset page: https://huggingface.co/datasets/Somtharu181coder/science_behavioral_and_domain_diversity_dataset.texttext-generation10K<n<100K0 likes18 downloads2mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.