Robotics Manipulation & Vision-Language Learning

5 repos

Systems and models for teaching robots to manipulate objects and navigate environments through vision-language grounding and learned action policies. The cluster centers on simulated and real-world robotic manipulation tasks, vision-language-action models that ground natural language instructions in continuous control, and datasets/benchmarks like RoboCasa for training embodied AI agents. Python dominates the implementation landscape, reflecting the prevalence of deep learning frameworks and simulation environments for this research area.

Python · 5
vla ·2,538
robotics ·2,521
cogact ·1,904
simulation ·1,904
codebase ·1,904
pi0 ·1,904
openvla-oft ·1,904
memoryvla ·1,904
real-world ·1,904
toolbox ·1,904