AI4AI at Test-Time: Strong-to-Weak Capability Transfer via Harnesses Paper • 2608.12307 • Published 1 day ago • 71
Co-Evolution in Agentic Systems: Toward Self-Directed Evolution Beyond Human Design Paper • 2608.10299 • Published 3 days ago • 123
JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Paper • 2607.23588 • Published 18 days ago • 125
SceneActBench: Can Agents Act on the 3D Scenes They See? Paper • 2607.22393 • Published 20 days ago • 7
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Paper • 2607.21653 • Published 22 days ago • 32
jrtmp/flash_attn-2.8.3.post1-cuda12.9-torch2.10-cp312-cp312-linux_x86_64.whl Updated 19 days ago • 32
jrtmp/flash_attn-2.8.3.post1-cuda12.9-torch2.10-cp312-cp312-linux_x86_64.whl Updated 19 days ago • 32
MeanFlowNFT: Bringing Forward-Process RL to Average-Velocity Generators Paper • 2607.15273 • Published 28 days ago • 17
Single-Rollout Asynchronous Optimization for Agentic Reinforcement Learning Paper • 2607.07508 • Published Jul 8 • 29
DSpark: Confidence-Scheduled Speculative Decoding with Semi-Autoregressive Generation Paper • 2607.05147 • Published Jul 6 • 43
Hierarchical Sparse Attention Done Right: Toward Infinite Context Modeling Paper • 2607.02980 • Published Jul 3 • 84
Nemotron-Labs-Diffusion: A Tri-Mode Language Model Unifying Autoregressive, Diffusion, and Self-Speculation Decoding Paper • 2607.05722 • Published Jul 7 • 13