SenseNova-U1.5: Towards Native Unified Visual Intelligence Paper • 2609.11929 • Published 7 days ago • 259
Uncovering Understanding-Generation Synergy in Native Unified Multimodal Models: From Representation, Task to System Paper • 2609.01607 • Published 16 days ago • 24
VBVR-Pro: A Scalable and Verifiable Suite for Native Visual Collection 23 items • Updated 10 days ago • 5
V-Rubrics: Visual Faithfulness via Rubric-Based Reinforcement Learning Paper • 2608.25580 • Published 22 days ago • 16
VBVR-Pro: A Scalable and Verifiable Suite for Native Visual Reasoning Paper • 2608.26105 • Published 22 days ago • 271
AURORA-LM: Autoencoding Unified Representation for Continuous-Latent Diffusion Language Modeling Paper • 2608.02602 • Published Aug 3 • 83
ACE-Data-0: Human-Centric Ambient Capture as Embodied Data Engine Paper • 2607.28625 • Published Jul 30 • 47
HumanCLAW: Can Vision-Language Models Act Through a Body? Paper • 2607.27180 • Published Jul 29 • 78
Apple-π: Benchmarking Thinking with Video Towards Law-Grounded Physical Intelligence Paper • 2607.16401 • Published Jul 17 • 45
VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding Paper • 2607.14935 • Published Jul 16 • 173
Running Agents Featured 50 SenseNova Vision 📚 50 Analyze or generate images with AI-powered vision tasks