Lego-RL Collection Harness-native RL for coding agents: the trained policy and the training task index. • 3 items • Updated 4 days ago • 11
SWE-Review: Closing the Loop on Issue Resolution with Agentic Code Review Paper • 2607.06065 • Published Jul 7 • 10
LEGO-RL: Harness-Native Reinforcement Learning for Coding Agents Paper • 2608.17393 • Published 6 days ago • 24
Lego-RL Collection Harness-native RL for coding agents: the trained policy and the training task index. • 3 items • Updated 4 days ago • 11
Dream-VL & Dream-VLA: Open Vision-Language and Vision-Language-Action Models with Diffusion Language Model Backbone Paper • 2512.22615 • Published Dec 27, 2025 • 51
SWE-Lego: Pushing the Limits of Supervised Fine-tuning for Software Issue Resolving Paper • 2601.01426 • Published Jan 4 • 25
Beyond Distillation: Pushing the Limits of Medical LLM Reasoning with Minimalist Rule-Based RL Paper • 2505.17952 • Published May 23, 2025 • 21