qwen2-math-1_5b-step-dpo / trainer_state.json

Commit History

Model save
40e5cfe
verified

rasdani commited on