LLMRouter: Unified Infrastructure for Developing, Evaluating, and Deploying LLM Routers Paper • 2608.06867 • Published 15 days ago • 109
Intern-S2-Preview: Scientific Agentic Foundation Model Paper • 2608.13505 • Published 9 days ago • 68
DarwinX: Evolving Agent Harnesses Through Natural Selection Paper • 2608.07545 • Published 22 days ago • 111
Alaya-EVOKE: From Linear-Scaling Supervision to Endless World Paper • 2608.13546 • Published 9 days ago • 132
Can We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social Dissemination Paper • 2608.14391 • Published 8 days ago • 277
SkillZip: Contract-Preserving Graph Compression for Scalable Agent Skill Libraries Paper • 2608.05604 • Published 16 days ago • 77
Mechanist: AI as a Scientific Instrument for Discovering the Mechanisms of Intelligence Paper • 2608.12036 • Published 10 days ago • 86
AI4AI at Test-Time: Strong-to-Weak Capability Transfer via Harnesses Paper • 2608.12307 • Published 10 days ago • 113
Co-Evolution in Agentic Systems: Toward Self-Directed Evolution Beyond Human Design Paper • 2608.10299 • Published 12 days ago • 134
ComBodied Agents: a New Paradigm of Human-Centric Agentic AI Paper • 2608.10915 • Published 11 days ago • 194
Spark-to-Paper: End-to-End Research Paper Generation as a Composable Skill Paper • 2608.11924 • Published 10 days ago • 285
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution Paper • 2608.00677 • Published 21 days ago • 261
Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA Paper • 2608.09819 • Published 12 days ago • 338
Ouroboros: A Self-Developing Frontier Coding Agent with Reviewed Core Evolution Paper • 2608.08311 • Published 14 days ago • 89
SWE-Bench ProMax: Benchmarking Agents on Large-Scale Multilingual Code Refactoring Paper • 2608.09802 • Published 12 days ago • 133
BDH-CQ: In-Context Learning with Recurrent Latent Reasoning Paper • 2608.09888 • Published 12 days ago • 704