OpenHelix-robot/awesome-dual-system-vla

A comprehensive list of papers about dual-system VLA models, including papers, codes, and related websites.

124

12 commits

updated Nov 21, 2025

See the code

README

Awesome-Dual-System-VLA

To address the challenges in vision-language-action (VLA) models, such as the difficulty of achieving efficient real-time performance, the high cost of pre-training, and the complexity of end-to-end fine-tuning on embodied data due to domain shift and catastrophic forgetting, Dual-System VLA models were introduced.

The development and architectural details of Dual-System VLA models are discussed in the paper.

This repository will be continuously updated, and we warmly welcome contributions from the community. If you have papers, projects, or resources that are not yet included, please feel free to submit them via a pull request or open an issue for discussion.

Current Results

CALVIN ABC→D

Method12345Avg. Len.
Single-System
OpenVLA91.377.862.052.143.53.27
UniVLA95.585.875.466.956.53.80
Seer94.487.279.972.264.33.98
Dual-System
LCB73.650.228.516.09.91.78
RationalVLA74.358.342.330.020.72.26
Robodual94.482.772.162.454.43.66
OpenHelix97.191.482.872.664.14.08

LIBERO

MethodLIBERO-SpatialLIBERO-ObjectLIBERO-GoalLIBERO-LongAvg.
Single-System
OpenVLA84.788.479.253.776.5
π096.898.895.885.294.2
OpenVLA-OFT97.698.497.994.597.1
GR00T N194.497.690.693.993.9
UniVLA96.596.895.692.095.2
Seer---87.7-
Dual-System
DexVLA97.299.195.6--
Hume98.699.899.498.698.6

✅ Dual-System VLA

Robot Manipulation

Humanoid Robot

❌ Not a Dual-System VLA

Robot Manipulation

Humanoid Robot

dual-system
vision-language-action-model

Contributors

BaiShuanghao

11 commits

OpenHelix-robot/awesome-dual-system-vla

A comprehensive list of papers about dual-system VLA models, including papers, codes, and related websites.

124

12 commits

updated Nov 21, 2025

See the code

README

Awesome-Dual-System-VLA

To address the challenges in vision-language-action (VLA) models, such as the difficulty of achieving efficient real-time performance, the high cost of pre-training, and the complexity of end-to-end fine-tuning on embodied data due to domain shift and catastrophic forgetting, Dual-System VLA models were introduced.

The development and architectural details of Dual-System VLA models are discussed in the paper.

This repository will be continuously updated, and we warmly welcome contributions from the community. If you have papers, projects, or resources that are not yet included, please feel free to submit them via a pull request or open an issue for discussion.

Current Results

CALVIN ABC→D

Method12345Avg. Len.
Single-System
OpenVLA91.377.862.052.143.53.27
UniVLA95.585.875.466.956.53.80
Seer94.487.279.972.264.33.98
Dual-System
LCB73.650.228.516.09.91.78
RationalVLA74.358.342.330.020.72.26
Robodual94.482.772.162.454.43.66
OpenHelix97.191.482.872.664.14.08

LIBERO

MethodLIBERO-SpatialLIBERO-ObjectLIBERO-GoalLIBERO-LongAvg.
Single-System
OpenVLA84.788.479.253.776.5
π096.898.895.885.294.2
OpenVLA-OFT97.698.497.994.597.1
GR00T N194.497.690.693.993.9
UniVLA96.596.895.692.095.2
Seer---87.7-
Dual-System
DexVLA97.299.195.6--
Hume98.699.899.498.698.6

✅ Dual-System VLA

Robot Manipulation

Humanoid Robot

❌ Not a Dual-System VLA

Robot Manipulation

Humanoid Robot

dual-system
vision-language-action-model

Contributors

BaiShuanghao

11 commits