PixMo-CapQA is a synthetic dataset of question/answer pairs about images. The data was generated by using the Claude large language model to build Q/A pairs from dense captions of images (the model did not see the actual images).
PixMo-CapQA is a part of the PixMo dataset collection and was used to train the Molmo family of models
Quick links:
data = datasets.load_dataset("allenai/pixmo-cap-qa", split="train")
Images are stored as URLs that will need to be downloaded separately. The image URLs can be repeated since many of the images have multiple Q/A pairs.
question field contains the input text, it includes "[USER]" and "[ASSISTANT]" tagsanswer field contains the final target output textmessages field contains the same data in a list-of-messages formats. The first message is from the
user, then messages alternative between user and assistant. This text does not contain "[USER]" and "[ASSISTANT]" tagsThis dataset is licensed under ODC-BY-1.0. It is intended for research and educational use in accordance with Ai2's Responsible Use Guidelines. This dataset includes data generated from Claude which are subject to Anthropic terms of service and usage policy.
24 commits
1 commits
PixMo-CapQA is a synthetic dataset of question/answer pairs about images. The data was generated by using the Claude large language model to build Q/A pairs from dense captions of images (the model did not see the actual images).
PixMo-CapQA is a part of the PixMo dataset collection and was used to train the Molmo family of models
Quick links:
data = datasets.load_dataset("allenai/pixmo-cap-qa", split="train")
Images are stored as URLs that will need to be downloaded separately. The image URLs can be repeated since many of the images have multiple Q/A pairs.
question field contains the input text, it includes "[USER]" and "[ASSISTANT]" tagsanswer field contains the final target output textmessages field contains the same data in a list-of-messages formats. The first message is from the
user, then messages alternative between user and assistant. This text does not contain "[USER]" and "[ASSISTANT]" tagsThis dataset is licensed under ODC-BY-1.0. It is intended for research and educational use in accordance with Ai2's Responsible Use Guidelines. This dataset includes data generated from Claude which are subject to Anthropic terms of service and usage policy.
24 commits
1 commits