public-knowledge-project/agentic-jats-annotation-qwen3.5-9b-lora-v4-rl-step25 Updated May 19 • 19 • 1
public-knowledge-project/agentic-jats-annotation-qwen3.5-9b-lora-v8-rl-step70 Updated May 22 • 21 • 1
mradermacher/Back2Struct-Image2SVG-7B-GGUF Reinforcement Learning • 8B • Updated 4 days ago • 588 • 1
Inventors-Hub/Falcon3-10B-Instruct-BehaviorTree-3-epochs-GGUF Text Generation • 10B • Updated Jun 21, 2025
Rakushaking/Qwen4b-SFT-d9-merged-after-dpo-toml-xml-yaml-dpo Text Generation • 4B • Updated Feb 8 • 17 •