Zero-WAM: In-Context World-Action Modeling from Human Videos for Open-Ended Task Generalization Paper • 2608.26103 • Published 2 days ago • 13
Zero-WAM: In-Context World-Action Modeling from Human Videos for Open-Ended Task Generalization Paper • 2608.26103 • Published 2 days ago • 13
Native Video-Action Pretraining for Generalizable Robot Control Paper • 2607.08639 • Published Jul 9 • 1
Next Forcing: Causal World Modeling with Multi-Chunk Prediction Paper • 2606.11187 • Published Jun 9 • 7
4DAnyone: Create Anyone in 4D from a Casual Monocular Video Paper • 2608.20335 • Published 8 days ago • 77
Next Forcing: Causal World Modeling with Multi-Chunk Prediction Paper • 2606.11187 • Published Jun 9 • 7
MemEye: A Visual-Centric Evaluation Framework for Multimodal Agent Memory Paper • 2605.15128 • Published May 14 • 65