TG-VLA
Full-size, whole-body Vision-Language-Action framework for humanoid robots
A full-size, whole-body Vision-Language-Action framework for humanoid robots, described as comprising HEX, HAF-VLA, and DSRL-DCT components for mobile manipulation.
Model Intelligence
Benchmarkable
No
Model level
family
Recent stories
1 linked story