yyyang/FocusUI-3B

Model

FocusUI-3B

1

3 commits

2 linked in READMEs

updated Feb 9, 2026

See the code

README

FocusUI-3B

This model was introduced in the paper:

FocusUI: Efficient UI Grounding via Position-Preserving Visual Token Selection

Model Zoo

ModelBackboneπŸ€— HuggingFace
FocusUI-3BQwen2.5-VL-3Bhttps://huggingface.co/yyyang/FocusUI-3B
FocusUI-7BQwen2.5-VL-7Bhttps://huggingface.co/yyyang/FocusUI-7B
FocusUI-2BQwen3-VL-2Bhttps://huggingface.co/yyyang/FocusUI-Qwen3-VL-2B

Dataset & Benchmarks

For the training and evaluation data, see FocusUI-Training-Data and UI-Grounding-Benchmarks.

Citation

@article{ouyang2025focusui,
  title   = {FocusUI: Efficient UI Grounding via Position-Preserving Visual Token Selection},
  author  = {Ouyang, Mingyu and Lin, Kevin Qinghong and Shou, Mike Zheng and Ng, Hwee Tou},
  year    = {2025},
  journal = {arXiv preprint},
}
qwen2_5_vl
safetensors

yyyang/FocusUI-3B

Model

FocusUI-3B

1

3 commits

2 linked in READMEs

updated Feb 9, 2026

See the code

README

FocusUI-3B

This model was introduced in the paper:

FocusUI: Efficient UI Grounding via Position-Preserving Visual Token Selection

Model Zoo

ModelBackboneπŸ€— HuggingFace
FocusUI-3BQwen2.5-VL-3Bhttps://huggingface.co/yyyang/FocusUI-3B
FocusUI-7BQwen2.5-VL-7Bhttps://huggingface.co/yyyang/FocusUI-7B
FocusUI-2BQwen3-VL-2Bhttps://huggingface.co/yyyang/FocusUI-Qwen3-VL-2B

Dataset & Benchmarks

For the training and evaluation data, see FocusUI-Training-Data and UI-Grounding-Benchmarks.

Citation

@article{ouyang2025focusui,
  title   = {FocusUI: Efficient UI Grounding via Position-Preserving Visual Token Selection},
  author  = {Ouyang, Mingyu and Lin, Kevin Qinghong and Shou, Mike Zheng and Ng, Hwee Tou},
  year    = {2025},
  journal = {arXiv preprint},
}
qwen2_5_vl
safetensors