datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
synthetic-code-understanding
SYNTHETIC-1
This is a subset of the task data used to construct SYNTHETIC-1. You can find the full collection here
Co-Simulation-Scheduling-And-Time-Synchronization-Code-Understanding-Dataset
Co-Simulation Scheduling and Time Synchronization Code Understanding Dataset
This dataset contains source code understanding samples focused on co-simulation master controllers and scheduling components, including simulation step sizes, execution order, time advancement, and synchronization control. Each sample pairs source code with an analysis question and an answer, and includes explanations of control flow, execution order, and timing relationships. It supports code SFT… See the full description on the dataset page: https://huggingface.co/datasets/Mobiusi/Co-Simulation-Scheduling-And-Time-Synchronization-Code-Understanding-Dataset.synthetic-code-understanding-v2-rustsynthetic-code-understanding-v2synthetic-code-understanding-v2-iteration-5synthetic-code-understanding-v2-js-testa1_code_primeintellect_code_understandingsynthetic-code-understanding-v2-iteration-12synthetic-code-understanding-v2-rust-testsynthetic-code-understanding-v2-rust-test-difficultsynthetic-code-understanding-v2-rust-test-sonnetsynthetic-code-understanding-v2-iteration-18-sonnetsynthetic-code-understanding-v2-cppsynthetic-code-understanding-v2-iteration-6synthetic-code-understanding-v2-iteration-7sanity-check-code-understanding-v2-rusta1_code_primeintellect_code_understanding_1744692978_eval_1331synthetic-code-understanding-v2-iteration-10sanity-check-code-understanding-fixed-jsonsanity-check-code-understanding-large-jsonload_in_code_primeintellect_code_understandinga1_code_primeintellect_code_understanding_eval_636d
mlfoundations-dev/a1_code_primeintellect_code_understanding_eval_636d
Precomputed model outputs for evaluation.
Evaluation Results
Summary
Metric
AIME24
AMC23
MATH500
MMLUPro
JEEBench
GPQADiamond
LiveCodeBench
CodeElo
CodeForces
Accuracy
15.7
50.7
70.0
25.6
31.9
42.4
0.6
2.6
3.2
AIME24
Average Accuracy: 15.67% ± 1.16%
Number of Runs: 10
Run
Accuracy
Questions Solved
Total Questions
1
13.33%
4
30
2
10.00%
3
30
3
10.00%… See the full description on the dataset page: https://huggingface.co/datasets/mlfoundations-dev/a1_code_primeintellect_code_understanding_eval_636d.synthetic-code-understanding-v2-iteration-15-o4synthetic-code-understanding-v2-iteration-17-sonnetsynthetic-code-understanding-v2-cpp-testsynthetic-code-understanding-v2-jssynthetic-code-understanding-v2-iteration-9synthetic-code-understanding-v2-iteration-13-o4synthetic-code-understanding-v2-cpp-test-2synthetic-code-understanding-v2-iteration-4
