MOPD: Multi-Teacher On-Policy Distillation for Capability Integration in LLM Post-Training Paper • 2606.30406 • Published Jun 29 • 19
Running on A100 Agents 8 Nemotron-Labs-Audio-Visual Flamingo 🎬 8 Analyze videos and answer questions about their content
SWE-Pruner Pro: The Coder LLM Already Knows What to Prune Paper • 2607.18213 • Published 14 days ago • 78
DeepSearch-World: Self-Distillation for Deep Search Agents in a Verifiable Environment Paper • 2607.07820 • Published 26 days ago • 91
SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning Paper • 2607.14777 • Published 18 days ago • 103
RAGU: A Multi-Step GraphRAG Engine with a Compact Domain-Adapted LLM Paper • 2607.11683 • Published 21 days ago • 148
AREX: Towards a Recursively Self-Improving Agent for Deep Research Paper • 2607.21461 • Published 11 days ago • 151
Program-as-Weights: A Programming Paradigm for Fuzzy Functions Paper • 2607.02512 • Published Jul 2 • 309
view article Article One Adapter, Both Modalities: Field Notes from Building and Serving a Multimodal Reranker lightonai • 18 days ago • 19
view article Article Bringing Nunchaku 4-bit Diffusion Inference to Diffusers rootonchair, sayakpaul • 11 days ago • 62