toshi456/LLaVA-CC3M-Pretrain-595K-JA

Dataset

11

stars

5

commits

4

linked in READMEs

Apr 12, 2024

updated

Browse cluster: Japanese Vision-Language Models

README

Dataset Card for "LLaVA-CC3M-Pretrain-595K-JA"

Dataset Details

Dataset Type: Japanese LLaVA CC3M Pretrain 595K is a localized version of the original LLaVA Visual Instruct CC3M 595K dataset. This version is translated into Japanese using cyberagent/calm2-7b-chat and is aimed at serving similar purposes in the context of Japanese language.

Resources for More Information: For information on the original dataset: liuhaotian/LLaVA-CC3M-Pretrain-595K

License: Must comply with license of CC-3M, BLIP (if you use their synthetic caption).

CC-3M The dataset may be freely used for any purpose, although acknowledgement of Google LLC ("Google") as the data source would be appreciated. The dataset is provided "AS IS" without any warranty, express or implied. Google disclaims all liability for any damages, direct or indirect, resulting from the use of the dataset.

Questions or Comments: For questions or comments about the original model, you can go to LLaVA GitHub Issues.

Intended use

Primary intended uses: The primary use of LLaVA is research on large multimodal models and chatbots.

Primary intended users: The primary intended users of the model are researchers and hobbyists in computer vision, natural language processing, machine learning, and artificial intelligence.

Contributors

toshi456

5 commits

toshi456/LLaVA-CC3M-Pretrain-595K-JA

Dataset

11

stars

5

commits

4

linked in READMEs

Apr 12, 2024

updated

Browse cluster: Japanese Vision-Language Models

README

Dataset Card for "LLaVA-CC3M-Pretrain-595K-JA"

Dataset Details

Dataset Type: Japanese LLaVA CC3M Pretrain 595K is a localized version of the original LLaVA Visual Instruct CC3M 595K dataset. This version is translated into Japanese using cyberagent/calm2-7b-chat and is aimed at serving similar purposes in the context of Japanese language.

Resources for More Information: For information on the original dataset: liuhaotian/LLaVA-CC3M-Pretrain-595K

License: Must comply with license of CC-3M, BLIP (if you use their synthetic caption).

CC-3M The dataset may be freely used for any purpose, although acknowledgement of Google LLC ("Google") as the data source would be appreciated. The dataset is provided "AS IS" without any warranty, express or implied. Google disclaims all liability for any damages, direct or indirect, resulting from the use of the dataset.

Questions or Comments: For questions or comments about the original model, you can go to LLaVA GitHub Issues.

Intended use

Primary intended uses: The primary use of LLaVA is research on large multimodal models and chatbots.

Primary intended users: The primary intended users of the model are researchers and hobbyists in computer vision, natural language processing, machine learning, and artificial intelligence.

Contributors

toshi456

5 commits