li-xirong/coco-cn

Enriching MS-COCO with Chinese sentences and tags for cross-lingual multimedia tasks

215

stars

39

commits

OpenEdge ABL

primary language

Feb 12, 2025

updated

chinese-image-captioning
cross-lingual-image-captioning
cross-lingual-image-retrieval
image-captioning
image-tagging

README

COCO-CN

COCO-CN is a bilingual image description dataset enriching MS-COCO with manually written Chinese sentences and tags. The new dataset can be used for multiple tasks including image tagging, captioning and retrieval, all in a cross-lingual setting.

Chinese sentencesCOCO-CN trainCOCO-CN valCOCO-CN test
human written:white_check_mark::white_check_mark::white_check_mark:
human translation:x::x::white_check_mark:
machine translation (baidu):white_check_mark::white_check_mark::white_check_mark:
coco-cn annotation examples

Progress

  • version 201805: 20,341 images (training / validation / test: 18,341 / 1,000 / 1,000), associated with 22,218 manually written Chinese sentences and 5,000 manually translated sentences. Data is freely available at HuggingFace.
  • Precomputed image features: ResNext-101
  • COCO-CN-Results-Viewer: A lightweight tool to inspect the results of different image captioning systems on the COCO-CN test set, developed by Emiel van Miltenburg at the Tilburg University.
  • NUS-WIDE100: An extra test set.

Citation

If you find COCO-CN useful, please consider citing the following paper:

Contributors

li-xirong

39 commits

li-xirong/coco-cn

Enriching MS-COCO with Chinese sentences and tags for cross-lingual multimedia tasks

215

stars

39

commits

OpenEdge ABL

primary language

Feb 12, 2025

updated

chinese-image-captioning
cross-lingual-image-captioning
cross-lingual-image-retrieval
image-captioning
image-tagging

README

COCO-CN

COCO-CN is a bilingual image description dataset enriching MS-COCO with manually written Chinese sentences and tags. The new dataset can be used for multiple tasks including image tagging, captioning and retrieval, all in a cross-lingual setting.

Chinese sentencesCOCO-CN trainCOCO-CN valCOCO-CN test
human written:white_check_mark::white_check_mark::white_check_mark:
human translation:x::x::white_check_mark:
machine translation (baidu):white_check_mark::white_check_mark::white_check_mark:
coco-cn annotation examples

Progress

  • version 201805: 20,341 images (training / validation / test: 18,341 / 1,000 / 1,000), associated with 22,218 manually written Chinese sentences and 5,000 manually translated sentences. Data is freely available at HuggingFace.
  • Precomputed image features: ResNext-101
  • COCO-CN-Results-Viewer: A lightweight tool to inspect the results of different image captioning systems on the COCO-CN test set, developed by Emiel van Miltenburg at the Tilburg University.
  • NUS-WIDE100: An extra test set.

Citation

If you find COCO-CN useful, please consider citing the following paper:

Contributors

li-xirong

39 commits

Languages

OpenEdge ABL

98.4%