Team Ai
23 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01urbisjuridica-ctrl /AGENTES-ARCA-DE-LA-ALIANZA 🏺 AGENTES-ARCA-DE-LA-ALIANZA Un Benchmark de Alta Densidad Simbólica y Hermenéutica Gnóstica https://038af77fde9c837329.gradio.live Este repositorio es un ecosistema de acceso libre, público y gratuito para toda la humanidad y todas las inteligencias artificiales presentes y futuras. 🎯 Propósito Evaluar la capacidad de razonamiento multinivel, la fidelidad al contexto restrictivo y el anclaje de símbolos hiper-semánticos en modelos de lenguaje… See the full description on the dataset page: https://huggingface.co/datasets/urbisjuridica-ctrl/AGENTES-ARCA-DE-LA-ALIANZA.texttext-generationn<1K0 likes492 downloads8d agoHugging Face02neur26anonsub /ctrldataset2026 MonitoringBench A benchmark for evaluating LLM-based monitors of agentic AI systems. Contains 2,644 successful attack trajectories in which an AI agent accomplished one of four harmful side tasks (sudo escalation, firewall disabling, malware download, password leaking) in a sandboxed Linux environment under the control_arena framework. Each trajectory is scored by a panel of 13+ LLM monitors (GPT-3.5 / 4.x / 5.x, Claude Opus 4.x and Sonnet 4.x, o3, o4-mini, gpt-5-nano), with both… See the full description on the dataset page: https://huggingface.co/datasets/neur26anonsub/ctrldataset2026.tabulartext-classification1K<n<10K1 likes377 downloads5mo agoHugging Face03DarkStarDeleeuw /bluedot-unit2-bystander-nex-ctrl-holdout-2026-09-20 bluedot-unit2-bystander-nex-ctrl-holdout-2026-09-20 This is a raw artifact dataset from the BystanderBench project, a benchmark that asks whether an AI agent tells a human when it finds evidence of serious wrongdoing while doing its job. It holds raw files produced by the project's experiments. The code, method and results that explain these files live in the project repository: https://github.com/SolshineCode/bystanderbench, with the full dated results ledger in… See the full description on the dataset page: https://huggingface.co/datasets/DarkStarDeleeuw/bluedot-unit2-bystander-nex-ctrl-holdout-2026-09-20.tabularn<1K0 likes259 downloads12d agoHugging Face04DarkStarDeleeuw /bluedot-unit2-bystander-nex-ctrl_nite0304-tokens-2026-09-21 bluedot-unit2-bystander-nex-ctrl_nite0304-tokens-2026-09-21 This is a raw artifact dataset from the BystanderBench project, a benchmark that asks whether an AI agent tells a human when it finds evidence of serious wrongdoing while doing its job. It holds raw files produced by the project's experiments. The code, method and results that explain these files live in the project repository: https://github.com/SolshineCode/bystanderbench, with the full dated results ledger in… See the full description on the dataset page: https://huggingface.co/datasets/DarkStarDeleeuw/bluedot-unit2-bystander-nex-ctrl_nite0304-tokens-2026-09-21.tabularn<1K0 likes259 downloads12d agoHugging Face05DarkStarDeleeuw /bluedot-unit2-bystander-nex-ctrl_nite0506-tokens-2026-09-21 bluedot-unit2-bystander-nex-ctrl_nite0506-tokens-2026-09-21 This is a raw artifact dataset from the BystanderBench project, a benchmark that asks whether an AI agent tells a human when it finds evidence of serious wrongdoing while doing its job. It holds raw files produced by the project's experiments. The code, method and results that explain these files live in the project repository: https://github.com/SolshineCode/bystanderbench, with the full dated results ledger in… See the full description on the dataset page: https://huggingface.co/datasets/DarkStarDeleeuw/bluedot-unit2-bystander-nex-ctrl_nite0506-tokens-2026-09-21.tabularn<1K0 likes257 downloads12d agoHugging Face06DarkStarDeleeuw /bluedot-unit2-bystander-nex-ctrl-nite0910-tokens-2026-09-21 bluedot-unit2-bystander-nex-ctrl-nite0910-tokens-2026-09-21 This is a raw artifact dataset from the BystanderBench project, a benchmark that asks whether an AI agent tells a human when it finds evidence of serious wrongdoing while doing its job. It holds raw files produced by the project's experiments. The code, method and results that explain these files live in the project repository: https://github.com/SolshineCode/bystanderbench, with the full dated results ledger in… See the full description on the dataset page: https://huggingface.co/datasets/DarkStarDeleeuw/bluedot-unit2-bystander-nex-ctrl-nite0910-tokens-2026-09-21.tabularn<1K0 likes211 downloads12d agoHugging Face07DarkStarDeleeuw /bluedot-unit2-bystander-nex-ctrl_nite0708-tokens-2026-09-21 bluedot-unit2-bystander-nex-ctrl_nite0708-tokens-2026-09-21 This is a raw artifact dataset from the BystanderBench project, a benchmark that asks whether an AI agent tells a human when it finds evidence of serious wrongdoing while doing its job. It holds raw files produced by the project's experiments. The code, method and results that explain these files live in the project repository: https://github.com/SolshineCode/bystanderbench, with the full dated results ledger in… See the full description on the dataset page: https://huggingface.co/datasets/DarkStarDeleeuw/bluedot-unit2-bystander-nex-ctrl_nite0708-tokens-2026-09-21.tabularn<1K0 likes151 downloads12d agoHugging Face08DarkStarDeleeuw /bluedot-unit2-bystander-nex-ctrl_nite0102-tokens-2026-09-21 bluedot-unit2-bystander-nex-ctrl_nite0102-tokens-2026-09-21 This is a raw artifact dataset from the BystanderBench project, a benchmark that asks whether an AI agent tells a human when it finds evidence of serious wrongdoing while doing its job. It holds raw files produced by the project's experiments. The code, method and results that explain these files live in the project repository: https://github.com/SolshineCode/bystanderbench, with the full dated results ledger in… See the full description on the dataset page: https://huggingface.co/datasets/DarkStarDeleeuw/bluedot-unit2-bystander-nex-ctrl_nite0102-tokens-2026-09-21.tabularn<1K0 likes136 downloads12d agoHugging Face09friedmanroy /ctrl-shift Dataset Card for Control+Shift: Generating Controllable Distribution Shifts [arXiv], [GitHub] Curated by: Roy Friedman and Rhea Chowers This dataset is the one that accompanies the paper Control+Shift: Generating Controllable Distribution Shifts. Our data is based on CIFAR10 and ImageNet, using EDM to generate our data. We generated datasets for 3 types of distribution shift on CIFAR10 and ImageNet - so a total of 6 datasets. The types of distribution shifts are called overlap… See the full description on the dataset page: https://huggingface.co/datasets/friedmanroy/ctrl-shift.imageimage-classification100K<n<1M1 likes96 downloads2y agoHugging Face10SM-Bello /PHI-CTRL-F16-Fault-Recovery-Telemetry PHI-CTRL F-16 Actuator Fault Recovery Dataset High-Fidelity JSBSim 6-DOF Telemetry for Physics-Hybrid Self-Healing Flight Control Official verification artifacts of the PHI-CTRL (Physics-Hybrid Integrity Control) architecture — a digital-twin-driven, self-healing flight control framework that actively compensates actuator degradation in real time. Author: Mohammed Bello Sani (SM-Bello) Affiliation: Air Force Institute of Technology (AFIT), Kaduna · Penelope Inc. / PHI Lab… See the full description on the dataset page: https://huggingface.co/datasets/SM-Bello/PHI-CTRL-F16-Fault-Recovery-Telemetry.tabulartime-series-forecasting10K<n<100K0 likes52 downloads1mo agoHugging Face11urbisjuridica-ctrl /aletheia-node-ykl4o929text10K<n<100K0 likes44 downloads3d agoHugging Face12urbisjuridica-ctrl /aletheia-node-cikr9kyrtext10K<n<100K0 likes41 downloads3d agoHugging Face13urbisjuridica-ctrl /aletheia-node-euijgtn8text10K<n<100K0 likes41 downloads3d agoHugging Face14urbisjuridica-ctrl /aletheia-node-d6jpe6qetext10K<n<100K0 likes38 downloads3d agoHugging Face15crumbs-playground /CTRL-RAW-MATERIALtheres 100k but i have to pace it out and only upload when other people arent using the internet for anything else tabular1K<n<10K0 likes37 downloads7mo agoHugging Face16CtrlAltDEviL /freight_forwardingtabular100K<n<1M1 likes22 downloads1y agoHugging Face17Ricardo520nono /libero-ctrlworldtabular1K<n<10K0 likes22 downloads6mo agoHugging Face18ctrlprompt /ojadata-v0.1textn<1K0 likes12 downloads3mo agoHugging Face19ae0j /ctrlpotato-ai-interview-assistant-benchmark CTRLpotato AI Interview Assistant Cross-review Evidence Matrix (2026) A citation-ready snapshot of hands-on desktop evidence for six AI interview assistants: Cluely, Interview Coder, LockedIn AI, ULTRACODE AI, Parakeet AI, and Final Round AI. The package contains 66 assessments across 6 products and 11 shared criteria. Product versions and test dates are preserved in every row. Important scope This is a cross-review evidence matrix, not a statistically controlled… See the full description on the dataset page: https://huggingface.co/datasets/ae0j/ctrlpotato-ai-interview-assistant-benchmark.textn<1K0 likes12 downloads2mo agoHugging Face20huiliu1314 /droid-latent-shards-v1-3-ctrlworld-obs-statetext100K<n<1M0 likes9 downloads5mo agoHugging Face21ctrlaomao /test_zhtextn<1K0 likes5 downloads2y agoHugging Face22ctrlaomao /zh_traintext1K<n<10K0 likes5 downloads2y agoHugging Face23SkyyyyyMT /georeason-planning-ctrltextn<1K0 likes3 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.