nanoMuse: An Open-Source Personal Agent for Every Device You Own Paper • 2610.08699 • Published 4 days ago • 98
In-Context Learning for Robots: Methods and Applications Paper • 2609.36012 • Published 12 days ago • 322
Disaggregated Quantization: Specializing LLM Prefill and Decode Paper • 2609.26333 • Published 18 days ago • 93
YuE2: Unifying Symbolic and Audio Music Generation at Frontier Quality Paper • 2609.33757 • Published 13 days ago • 234
OmniEdu: Open Foundation Models for Learning and Teaching Paper • 2609.23088 • Published 21 days ago • 239
Realtime-Venus: A full-duplex interaction system with asynchronous delegation Paper • 2609.13814 • Published 28 days ago • 199
ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search Paper • 2609.13356 • Published 29 days ago • 266
Unlocking Lossless Speedups in LLMs via Discrete Diffusion Paper • 2609.04010 • Published Sep 3 • 114
AuK Technical Report: An Open-Source Foundational Model for Speech Generation and Editing Paper • 2609.08936 • Published Sep 8 • 161
Beyond Data Scaling: Representation-Centric Continued Pre-training for Vision-Language-Action Models Paper • 2608.27550 • Published Aug 27 • 83
Qwen-Drive-1.0: An Initial Step towards a Vision-Language Foundation Model for Autonomous Driving Paper • 2609.00111 • Published Aug 31 • 316
SimpleOPD: Simple Tokenizer-Agnostic On-Policy Distillation for Long-Context Reasoning Paper • 2608.14277 • Published Aug 14 • 36
BDH-CQ: In-Context Learning with Recurrent Latent Reasoning Paper • 2608.09888 • Published Aug 10 • 799
SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks Paper • 2608.02023 • Published Aug 3 • 161 • 8
SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks Paper • 2608.02023 • Published Aug 3 • 161