D
Data
Text
Dynamic Scaling of Unit Tests for Code Reward Modeling Zeyao Ma1,3*†, Xiaokang Zhang1,3*, Jing
This NLP/ML research paper (cs.CL) improves how correct solutions are selected when a large language model generates code. A common pipeline samples many candidate programs and reranks them with LLM-…