rlsamplingJF/evolm-4B-160BT-finemath_part1_part2-rm-lr1e-6-constant-warmup_0.05-bs16-gc1.0-cc0.01-ls0-step180 4B • Updated about 13 hours ago • 15
rlsamplingJF/Qwen2.5-7B-Instruct-finemath_part1-rm-lr1e-6-constant-warmup_0.05-bs8-gc1.0-cc0.01-ls0.1-step45 7B • Updated 1 day ago • 18
rlsamplingJF/Llama-3.2-3B-finemath_part1_part2-rm-lr1e-6-constant-warmup_0.05-bs16-gc1.0-cc0.01-ls0-step375 3B • Updated 2 days ago • 40
rlsamplingJF/Llama-3.2-3B-finemath_part1_part2-rm-lr1e-6-constant-warmup_0.05-bs16-gc1.0-step345 3B • Updated 3 days ago • 19
rlsamplingJF/evolm-4B-160BT-finemath_part1_part2-rm-lr1e-6-constant-warmup_0.05-bs16-gc1.0-cc0.01-ls0-initial 4B • Updated 3 days ago • 20
rlsamplingJF/evolm-4B-160BT-finemath_part1_part2-rm-lr1e-6-constant-warmup_0.05-bs16-gc1.0-cc0.01-ls0-step120 4B • Updated 3 days ago • 24
rlsamplingJF/Llama-3.2-1B-finemath_part1_part2-rm-lr1e-6-constant-warmup_0.05-bs16-gc1.0-cc0.01-ls0.1-step150 1B • Updated 4 days ago • 15
rlsamplingJF/Llama-3.2-1B-finemath_part1_part2-rm-lr1e-6-constant-warmup_0.05-bs16-gc1.0-cc0-ls0-initial 1B • Updated 4 days ago • 21
rlsamplingJF/Llama-3.2-1B-finemath_part1_part2-rm-lr1e-6-constant-warmup_0.05-bs16-gc1.0-cc0-ls0-step405 1B • Updated 4 days ago • 18
rlsamplingJF/Qwen2.5-7B-Instruct-finemath_part1-rm-lr1e-6-constant-warmup_0.05-bs16-gc1.0-cc0.01-ls0.1-step30 7B • Updated 5 days ago • 28