llama3.2-3B-Math-GRPO / model-00001-of-00002.safetensors

Commit History

(Trained with Unsloth)
6923d9b
verified

Harsha901 commited on