Team Ai
Datasetpublic

CompilingThings/compile-benchmark

CompilingThings Compile Benchmark for MQL5® This release evaluates compile success of generated MQL5 on two private held-out sets. The first is 300 Expert Advisor prompts, run on four arms: the base model, two tuned local models and one frontier API model. The second is 200 non-EA prompts (include files, custom indicators, scripts and services, 50 each), run on the three local arms, plus a stability re-run of one of them. The holdout results are attested, not fully verifiable:… See the full description on the dataset page: https://huggingface.co/datasets/CompilingThings/compile-benchmark.

sourceHugging Faceotherupdated 24d agoView on Hugging Face
1likes170downloads
.gitattributes5 linesDownload Raw Back to root
1# The manifest hashes these files by byte. Pin the line endings to the tree so a2# checkout cannot rewrite them and break every hash in SHA256SUMS.txt.3* text eol=lf4corpus_row_hashes.json filter=lfs diff=lfs merge=lfs -text5