Akahsizrr/kernelswarm
kernelswarm — multi-agent kernel-optimization episodes (RL) Synthetic dataset of multi-agent long-horizon optimization campaigns on AI-inference kernels. Each row is one whole episode: a lead orchestrator plus 2–20 specialist agents split the work, exchange dispatches/statuses/handoffs, submit full kernel candidates, and receive simulated compile/verify/bench feedback — with a reward attached to every environment result. Inference-focused only. All ops are inference primitives… See the full description on the dataset page: https://huggingface.co/datasets/Akahsizrr/kernelswarm.
055
