Zero-WAM: In-Context World-Action Modeling from Human Videos for Open-Ended Task Generalization Paper • 2608.26103 • Published 5 days ago • 18
Zero-WAM: In-Context World-Action Modeling from Human Videos for Open-Ended Task Generalization Paper • 2608.26103 • Published 5 days ago • 18
Native Video-Action Pretraining for Generalizable Robot Control Paper • 2607.08639 • Published Jul 9 • 1
Next Forcing: Causal World Modeling with Multi-Chunk Prediction Paper • 2606.11187 • Published Jun 9 • 7
4DAnyone: Create Anyone in 4D from a Casual Monocular Video Paper • 2608.20335 • Published 11 days ago • 81
Next Forcing: Causal World Modeling with Multi-Chunk Prediction Paper • 2606.11187 • Published Jun 9 • 7
MemEye: A Visual-Centric Evaluation Framework for Multimodal Agent Memory Paper • 2605.15128 • Published May 14 • 65