What Makes World Action Models Generalize? An Empirical Study of Test-Time Future Modeling Paper • 2609.34981 • Published 2 days ago • 103
OmniTaskonomy Collection OmniTaskonomy probes when generation helps understanding. Check our project page here: https://omni-taskonomy.github.io/ • 3 items • Updated about 24 hours ago • 2
OmniTaskonomy: When Does Visual Generation Improve Visual Understanding? Paper • 2609.38079 • Published 2 days ago • 44
OmniTaskonomy Collection OmniTaskonomy probes when generation helps understanding. Check our project page here: https://omni-taskonomy.github.io/ • 3 items • Updated about 24 hours ago • 2
OmniTaskonomy Collection OmniTaskonomy probes when generation helps understanding. Check our project page here: https://omni-taskonomy.github.io/ • 3 items • Updated about 24 hours ago • 2
Render-of-Thought: Rendering Textual Chain-of-Thought as Images for Visual Latent Reasoning Paper • 2601.14750 • Published Jan 21 • 18
Vision-as-Inverse-Graphics Agent via Interleaved Multimodal Reasoning Paper • 2601.11109 • Published Jan 16 • 3
Emu3.5: Native Multimodal Models are World Learners Paper • 2510.26583 • Published Oct 30, 2025 • 117