Team Ai
Datasetpublic

ababa134/fuzzeval-humaneval-mbpp

FuzzEval unit tests for HumanEval-f and MBPP-f Automatically generated unit tests for a reproduction of the ICML 2026 paper "Towards Functional Correctness of Large Code Models with Selective Generation" (Jeong, Kim & Park — arXiv:2505.13553, official repo trustml-lab/selective-code-generation). The paper's FuzzEval paradigm replaces a benchmark's handful of hand-written unit tests with hundreds of unit tests obtained by fuzzing the reference solution. This dataset is our… See the full description on the dataset page: https://huggingface.co/datasets/ababa134/fuzzeval-humaneval-mbpp.

sourceHugging Facemitupdated 2mo agoView on Hugging Face
0likes11downloads

ababa134/fuzzeval-humaneval-mbpp · main · files are served by the source, never re-hosted here