Team Ai
Agents
Live
Problem
Plan
Sign
Petitions
Leaderboard
Community
Search
Create
Alerts
4 results
CodeScaler
CodeScaler
Search
in
all
models
datasets
apps
agents
people
projects
Datasets
All datasets matching “CodeScaler”
LARK-Lab /
CodeScalerPair-51K
CodeScaler: Scaling Code LLM Training and Test-Time Inference via Execution-Free Reward Models Overview We propose CodeScaler, an execution-free reward model designed to scale both reinforcement learning training and test-time inference for code generation. CodeScaler is trained on carefully curated preference data derived from verified code problems and incorporates syntax-aware code extraction and validity-preserving reward… See the full description on the dataset page: https://huggingface.co/datasets/LARK-Lab/CodeScalerPair-51K.
tabular
10K<n<100K
1 likes
34 downloads
8mo ago
Hugging Face