Unit-test generation focus
According to the model card, the model generates Python unit tests with detailed comments and without testing exception-throwing behavior.
Open Source Model Profile · KAKA22
CodeRM-8B is an 8.03B-parameter Llama text-generation model from KAKA22. According to the model card, it generates high-quality Python unit tests.
CodeRM-8B is published by KAKA22 as a Llama text-generation model. The captured configuration identifies LlamaForCausalLM with model type llama and Safetensors metadata reports 8,030,261,248 parameters. According to the model card, it is built on Llama3.1-8B-Instruct for Python unit-test generation.
According to the model card, the model generates Python unit tests with detailed comments and without testing exception-throwing behavior.
According to the model card, 60K synthetic tests were synthesized with Llama3.1-70B-Instruct from CodeFeedback-Filtered-Instruction and TACO, openly released as CodeRM-UnitTest.
According to the model card, the publisher reports HumanEval Plus, MBPP Plus, and LiveCodeBench gains comparable to Llama3.1-70B-Instruct; these are publisher claims.
Source: KAKA22/CodeRM-8B
Captured: Unknown. Processed: 2026-09-07T19:34:32.402121+00:00.
Introduction CodeRM-8B is a small yet powerful model designed to enable efficient and high-quality unit test generation. It is trained on a dataset of 60k high-quality synthetic Python unit tests using Llama3.1-70B-Instruct. These unit tests are synthesized based on two well-regarded code instruction tuning datasets: CodeFeedback-Filtered-Instruction and the training set of TACO . The training dataset used for unit test generation is openly available under CodeRM-UnitTest . For further information and details of training, refer to our paper: "Dynamic Scaling of Unit Tests for Code Reward Modeling" available on arXiv . You can also v…
F001F002F003F004F005F006F007F010F011F012F013F014F015F017F018F019