mradermacher/A2Search-3B-Instruct-i1-GGUF Reinforcement Learning • 3B • Updated 7 days ago • 1.46k • 1