Model type: ShareCaptioner-Video is an open-source captioner fine-tuned on GPT4V-assisted ShareGPT4Video detailed caption data with supporting various durations, aspect ratios, and resolutions of videos. ShareCaptioner-Video is based on the InternLM-Xcomposer2-4KHD model.
ShareCaptaioner-Video features 4 roles:
Model date: ShareCaptioner was trained in May 2024.
Paper or resources for more information: [Project] [Paper] [Code]
Primary intended uses: The primary use of ShareCaptioner-Video is about producing high-quality video captions.
Primary intended users: The primary intended users of the model are researchers and hobbyists in computer vision, natural language processing, machine learning, and artificial intelligence.
arxiv.org/abs/2406.04325
Model type: ShareCaptioner-Video is an open-source captioner fine-tuned on GPT4V-assisted ShareGPT4Video detailed caption data with supporting various durations, aspect ratios, and resolutions of videos. ShareCaptioner-Video is based on the InternLM-Xcomposer2-4KHD model.
ShareCaptaioner-Video features 4 roles:
Model date: ShareCaptioner was trained in May 2024.
Paper or resources for more information: [Project] [Paper] [Code]
Primary intended uses: The primary use of ShareCaptioner-Video is about producing high-quality video captions.
Primary intended users: The primary intended users of the model are researchers and hobbyists in computer vision, natural language processing, machine learning, and artificial intelligence.
arxiv.org/abs/2406.04325