MOPD: Multi-Teacher On-Policy Distillation for Capability Integration in LLM Post-Training Paper • 2606.30406 • Published Jun 29 • 18
Running on A100 Agents 8 Nemotron-Labs-Audio-Visual Flamingo 🎬 8 Analyze videos and answer questions about their content
SWE-Pruner Pro: The Coder LLM Already Knows What to Prune Paper • 2607.18213 • Published 12 days ago • 78
DeepSearch-World: Self-Distillation for Deep Search Agents in a Verifiable Environment Paper • 2607.07820 • Published 24 days ago • 91
SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning Paper • 2607.14777 • Published 16 days ago • 103
RAGU: A Multi-Step GraphRAG Engine with a Compact Domain-Adapted LLM Paper • 2607.11683 • Published 19 days ago • 148
AREX: Towards a Recursively Self-Improving Agent for Deep Research Paper • 2607.21461 • Published 9 days ago • 149
Program-as-Weights: A Programming Paradigm for Fuzzy Functions Paper • 2607.02512 • Published 30 days ago • 238
view article Article One Adapter, Both Modalities: Field Notes from Building and Serving a Multimodal Reranker lightonai • 15 days ago • 18
view article Article Bringing Nunchaku 4-bit Diffusion Inference to Diffusers rootonchair, sayakpaul • 9 days ago • 60