mradermacher/Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled-GGUF 27B • Updated 14 days ago • 13.5k • 17
In-Context Reinforcement Learning for Tool Use in Large Language Models Paper • 2603.08068 • Published 14 days ago • 40