Qwen3-8B-LQA-3e-GRPO-100s-save2 / model-00003-of-00004.safetensors

Commit History

(Trained with Unsloth)
34ea21a
verified

lmq1909 commited on