OmniTaskonomy: When Does Visual Generation Improve Visual Understanding? Paper • 2609.38079 • Published 6 days ago • 55
What Makes World Action Models Generalize? An Empirical Study of Test-Time Future Modeling Paper • 2609.34981 • Published 6 days ago • 134
OmniTaskonomy Collection OmniTaskonomy probes when generation helps understanding. Check our project page here: https://omni-taskonomy.github.io/ • 3 items • Updated 5 days ago • 2
OmniTaskonomy: When Does Visual Generation Improve Visual Understanding? Paper • 2609.38079 • Published 6 days ago • 55
OmniTaskonomy Collection OmniTaskonomy probes when generation helps understanding. Check our project page here: https://omni-taskonomy.github.io/ • 3 items • Updated 5 days ago • 2
OmniTaskonomy Collection OmniTaskonomy probes when generation helps understanding. Check our project page here: https://omni-taskonomy.github.io/ • 3 items • Updated 5 days ago • 2
Render-of-Thought: Rendering Textual Chain-of-Thought as Images for Visual Latent Reasoning Paper • 2601.14750 • Published Jan 21 • 18
Vision-as-Inverse-Graphics Agent via Interleaved Multimodal Reasoning Paper • 2601.11109 • Published Jan 16 • 3
Emu3.5: Native Multimodal Models are World Learners Paper • 2510.26583 • Published Oct 30, 2025 • 117