Robotic Vision-Language Action Models

23 repos

Vision-language-action (VLA) models and learning frameworks that enable robots to understand and execute physical tasks from visual observations and language instructions. The cluster covers foundation models for robotic manipulation and navigation, reinforcement learning approaches for skill acquisition, and simulation environments for training embodied AI agents. Key themes include behavior cloning from demonstrations, world model learning, and real-to-sim transfer for continuous control in kitchen and household manipulation tasks.

Python · 9
robotics ·5,751
vla ·5,211
vision-language-action ·3,712
physical-ai ·3,168
autonomous-driving ·2,783
self-driving-car ·2,783
reasoning ·2,783
nvidia ·2,783
alpamayo ·2,783
autonomous-vehicles ·2,783