HYU-NLP-EVAL/qwen3-4b-rar-medicine-static-r0-matched-seed11-step-003 Text Generation • 4B • Updated 4 days ago • 165 • 1
Allenda/ScaleSeek-Qwen3.5-9B-GRPOv3-step160 Reinforcement Learning • 9B • Updated 7 days ago • 17 • 1