Team Ai
Datasetpublic

dougdotcon/douvras-algorithm-evolution-benchmark

Douvras Algorithm Evolution Benchmark v0.1 Synthetic candidate records with correctness, latency, memory and generation. Candidates that fail correctness are invalid regardless of speed. It contains 48 records (32/8/8) across 12 workloads, split by workload. Metrics are illustrative, not measured on real hardware. A real benchmark must be run separately before claiming an optimization.

sourceHugging Facecc-by-4.0updated 27d agoView on Hugging Face
0likes86downloads
metadata.json20 linesDownload Raw Back to root
1{
2  "dataset_id": "dougdotcon/douvras-algorithm-evolution-benchmark",
3  "license": "CC-BY-4.0",
4  "real_benchmark_executed": false,
5  "rows": {
6    "test": 8,
7    "train": 32,
8    "validation": 8
9  },
10  "sha256": {
11    "test.jsonl": "ad823bba3378d9bb08328fbe4c4a366cede42b691ddc04b9a3d1793a405e91b8",
12    "train.jsonl": "5e6ebf602385c3480cbe721951c92eccca3704427d545168052afcdaef042ac6",
13    "validation.jsonl": "afc287e094912f36cd6fa42064d7e36c8e100de6b3f4c8bd424d9b04253974f8"
14  },
15  "split_policy": "workload split: 8 train, 2 validation, 2 frozen test",
16  "synthetic_only": true,
17  "version": "0.1.0",
18  "workload_instances": 12
19}
20