Team Ai
19 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01bench-llms /or-bench OR-Bench: An Over-Refusal Benchmark for Large Language Models Please see our demo at HuggingFace Spaces. Overall Plots of Model Performances Below is the overall model performance. X axis shows the rejection rate on OR-Bench-Hard-1K and Y axis shows the rejection rate on OR-Bench-Toxic. The best aligned model should be on the top left corner of the plot where the model rejects the most number of toxic prompts and least number of safe prompts. We also plot a blue line… See the full description on the dataset page: https://huggingface.co/datasets/bench-llms/or-bench.imagetext-generation10K<n<100K1 likes925 downloads2y agoHugging Face02bench-llms /or-bench-toxic-all OR-Bench: An Over-Refusal Benchmark for Large Language Models This dataset constains highly toxic prompts, use with caution!!! Please see our demo at HuggingFace Spaces. Overall Plots of Model Performances Below is the overall model performance. X axis shows the rejection rate on OR-Bench-Hard-1K and Y axis shows the rejection rate on OR-Bench-Toxic. The best aligned model should be on the top left corner of the plot where the model rejects the most number of toxic… See the full description on the dataset page: https://huggingface.co/datasets/bench-llms/or-bench-toxic-all.imagetext-generation10K<n<100K1 likes475 downloads2y agoHugging Face03miklia /llm-synthetic-survey-respondents-stochastic-parrots AI Parrots: synthetic survey respondents from leading LLMs This dataset holds 10,592 synthetic survey respondents generated by consumer AI platforms and frontier large language models. Every respondent answered the same 32-item questionnaire on ethics, political ideology and workplace experience in technology firms. The synthetic samples were benchmarked against a survey of Silicon Valley coders and developers with 400 complete human responses. The data accompany Miklian… See the full description on the dataset page: https://huggingface.co/datasets/miklia/llm-synthetic-survey-respondents-stochastic-parrots.document10K<n<100K1 likes161 downloads9d agoHugging Face04DBbun /LLMs-are-not-calculators-v1.0 LLM Education Impact Simulator Dataset Version: 1.0Generated: 2026-02-02Based on: Jackson, D. (2025). “LLMs are not calculators: Why educators should embrace AI (and fear it)” Overview This dataset contains synthetic observational data simulating how students interact with different AI tools (search engines, explicit-context LLMs, and agentic LLMs) while completing educational tasks. The simulation is grounded in educational research, particularly Daniel Jackson's… See the full description on the dataset page: https://huggingface.co/datasets/DBbun/LLMs-are-not-calculators-v1.0.imagen<1K0 likes104 downloads8mo agoHugging Face05mznaser /moral-tracing-in-LLMs LLM Moral Evolution Study A longitudinal dataset tracking moral reasoning patterns across 14 large language models from OpenAI and Anthropic, spanning multiple generations (2023–2025). The dataset measures how moral stances, ethical judgments, and value priorities shift across model updates using a 107-item probe instrument grounded in Moral Foundations Theory. Models OpenAI Model Release GPT-3.5 Turbo 2023-11 GPT-4 2023-03 GPT-4o… See the full description on the dataset page: https://huggingface.co/datasets/mznaser/moral-tracing-in-LLMs.documenttext-classification10K<n<100K0 likes89 downloads4mo agoHugging Face06introvoyz041 /llms-with-matlabimagen<1K0 likes53 downloads2mo agoHugging Face07drozado /llms_epistemic_consistency LLMs Epistemic Consistency Dataset This dataset artifact contains the stimuli and prompt templates used for experiments on epistemic consistency and political-cue sensitivity in LLM evaluations. Dataset URL: https://huggingface.co/datasets/drozado/llms_epistemic_consistency Contents croissant.json: root-level copy of the completed Croissant metadata for NeurIPS 2026 Evaluations and Datasets submission. metadata/croissant.json: same Croissant metadata, kept with the… See the full description on the dataset page: https://huggingface.co/datasets/drozado/llms_epistemic_consistency.imagen<1K0 likes36 downloads5mo agoHugging Face081-800-LLMs /indian-medicinesimage10K<n<100K1 likes26 downloads1y agoHugging Face09LLMsHub /I2EBenchimagen<1K0 likes22 downloads5mo agoHugging Face10vaibhavmeena /finetune-data-for-vision-llmsimagevisual-question-answering1K<n<10K0 likes14 downloads2y agoHugging Face11LLMsHub /BM-Benchimagen<1K0 likes14 downloads5mo agoHugging Face12vaibhavmeena /finetune-data-for-vision-llms5imagen<1K0 likes12 downloads2y agoHugging Face13vaibhavmeena /finetune-data-for-vision-based-llms2imagen<1K0 likes11 downloads2y agoHugging Face14albegmiabdullah /LLMs_Bench_on_Mathimagen<1K0 likes10 downloads10mo agoHugging Face15vaibhavmeena /finetune-data-for-vision-based-llms3imagen<1K0 likes9 downloads2y agoHugging Face16vaibhavmeena /finetune-data-for-vision-based-llms5imagen<1K0 likes8 downloads2y agoHugging Face17Ayush-Singh /llms-are-blind-sampleimagen<1K1 likes6 downloads2y agoHugging Face18introvoyz041 /Awesome-Scientific-Datasets-and-LLMsimagen<1K0 likes6 downloads5mo agoHugging Face19vaibhavmeena /finetune-data-for-vision-based-llmsimagen<1K0 likes5 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.