openenv
Datasets
All datasets matching “openenv”openenv-pr-review-benchmark
OpenEnv Code Review Environment
This project is now an OpenEnv-style reinforcement learning environment for code review.
An external AI agent receives a static PR task (diff + context), submits a review as an action, and gets a reward based on planted ground-truth issues.
What Changed
Added GitHub Actions integration endpoint for live PR grading.
Removed Nova Act and Bedrock from the active model path.
Added OpenAI-backed analyzer utilities.
Added deterministic OpenEnv… See the full description on the dataset page: https://huggingface.co/datasets/adityam/openenv-pr-review-benchmark.3ambench
3amBench: can your agent write alerts that page the right human at 3 a.m., and only then?
3amBench (package alertforge) is an RL environment and benchmark for a job SRE teams do every week:
owning Prometheus alerting rules and Alertmanager routing as code. Each task drops the agent into a realistic
monitoring/ repo for a fictional company with a handful of change requests: onboard a service onto
multi-window burn-rate SLO alerts, fix the alert that paged 30 times last night… See the full description on the dataset page: https://huggingface.co/datasets/openenvforge/3ambench.openenv-examples-trl-2026-07-06openenv-pathway-analysis-env
OpenEnv Pathway Analysis Environment
This repository packages pathway_analysis_env for Hugging Face Hub publication.
Contents
envs/pathway_analysis_env/ environment code
reproducible GEO benchmark inputs for 3 tasks
task expansion scripts:
create_geo_task.py
append_task_to_manifest.py
Run locally
uv sync --all-extras
PYTHONPATH=src:envs uv run python envs/pathway_analysis_env/scripts/run_agent_eval_suite.py --manifest… See the full description on the dataset page: https://huggingface.co/datasets/aparulsarma/openenv-pathway-analysis-env.openenv-scalingWorldAtlas
