datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
AetherCode
AetherCode: Evaluating LLMs' Ability to Win In Premier Programming Competitions
Introduction
Competitive programming has emerged as a critical benchmark for evaluating the reasoning and coding capabilities of Large Language Models (LLMs). Despite impressive progress on existing benchmarks, we argue that current evaluations overstate model proficiency, masking a substantial gap between LLMs and elite human programmers. This gap arises… See the full description on the dataset page: https://huggingface.co/datasets/m-a-p/AetherCode.aether-cyber-sft
enosislabs/aether-cyber-sft
Version: surface-clean-20260620
Generated UTC: 2026-06-20T03:36:32.011333+00:00
Source path: artifacts/aether-cyber-sft-surface-candidate.jsonl
Git commit: 6fadc0ba97b4c49a7b644e70afacddc8ba98029e
Curated Aether PRISM SFT dataset for authorized cybersecurity training.
Each source shard follows 5-row review discipline before publish.
Records
Total examples: 1794
Domains
vulnerability_research: 757
red_team_ops: 334… See the full description on the dataset page: https://huggingface.co/datasets/enosislabs/aether-cyber-sft.aether-0.8b-cyber-sft
enosislabs/aether-0.8b-cyber-sft
Version: aether-0.8b-cyber-20260618
Generated UTC: 2026-06-19T02:53:44.849450+00:00
Source path: data/curated
Git commit: 3ac5f50de88e43122cfda43d65f1ad01b5febff6
Curated Aether PRISM SFT dataset for authorized cybersecurity training.
Each source shard follows 5-row review discipline before publish.
Records
Total examples: 1900
Domains
vulnerability_research: 715
red_team_ops: 335
detection_engineering: 257… See the full description on the dataset page: https://huggingface.co/datasets/enosislabs/aether-0.8b-cyber-sft.
