Running 229 The ultimate guide to multi-harness RL 🔀 229 Train open models with RL inside real agent harnesses
MMDuet2: Enhancing Proactive Interaction of Video MLLMs with Multi-Turn Reinforcement Learning Paper • 2512.06810 • Published Dec 7, 2025
Revisiting Multimodal Positional Encoding in Vision-Language Models Paper • 2510.23095 • Published Oct 27, 2025 • 23