mradermacher/AgentHijack-Agent-GGUF Reinforcement Learning • 8B • Updated about 11 hours ago • 244 • 1