Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

H-EmbodVis

university
https://github.com/H-EmbodVis
Activity Feed

AI & ML interests

None defined yet.

Recent Activity

CFYao  updated a model about 2 hours ago
H-EmbodVis/TurboVLA
CFYao  published a model 1 day ago
H-EmbodVis/TurboVLA
dkliang  submitted a paper 2 days ago
TurboVLA: Real-Time Vision-Language-Action Model at 32 Hz on an RTX 4090 with <1 GB VRAM
View all activity

Papers

TurboVLA: Real-Time Vision-Language-Action Model at 32 Hz on an RTX 4090 with <1 GB VRAM

HERMES++: Toward a Unified Driving World Model for 3D Scene Understanding and Generation

View all Papers

Dingkang Liang's profile picture Xin Zhou's profile picture Cheng's profile picture Cheng Zhang's profile picture Xianjin-Wu's profile picture HENG FANG's profile picture Ellery Kant's profile picture Chenfei Yao's profile picture
H-EmbodVis 's papers 7
Submitted by
Dingkang Liang
123

TurboVLA: Real-Time Vision-Language-Action Model at 32 Hz on an RTX 4090 with <1 GB VRAM

H-EmbodVis H-EmbodVis
162 2
Submitted by
Xin Zhou
74

HERMES++: Toward a Unified Driving World Model for 3D Scene Understanding and Generation

H-EmbodVis H-EmbodVis
66 2
Submitted by
Dingkang Liang
116

When Numbers Speak: Aligning Textual Numerals and Visual Instances in Text-to-Video Diffusion Models

H-EmbodVis H-EmbodVis
68 4
Submitted by
Dingkang Liang
157

Out of Sight but Not Out of Mind: Hybrid Memory for Dynamic Video World Models

H-EmbodVis H-EmbodVis
268 4
Submitted by
Dingkang Liang
95

Generation Models Know Space: Unleashing Implicit 3D Priors for Scene Understanding

H-EmbodVis H-EmbodVis
421 5
Submitted by
Dingkang Liang
3

Towards Generalizable Robotic Manipulation in Dynamic Environments

H-EmbodVis H-EmbodVis
237 2
Submitted by
Dingkang Liang
7

Cook and Clean Together: Teaching Embodied Agents for Parallel Task Execution

H-EmbodVis H-EmbodVis
365 2
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs