Dataset type: LLaVA-Plus-v1-117K is a set of GPT-generated multimodal tool-augmented instruction-following data. It is constructed for tool use to build large multimodal agents with GPT-4-plus vision/language capability.
Dataset date: LLaVA-Plus-v1-117K was collected in Sep 2023, by prompting ChatGPT/GPT-4-0314 API.
Paper or resources for more information: https://llava-vl.github.io/llava-plus
License: Attribution-NonCommercial 4.0 International It should abide by the policy of OpenAI: https://openai.com/policies/terms-of-use
Where to send questions or comments about the model: https://github.com/LLaVA-VL/LLaVA-Plus-Codebase/issues
Primary intended uses: The primary use of LLaVA-Plus is research on large multimodal agents, and chatbots.
Primary intended users: The primary intended users of the model are researchers and hobbyists in computer vision, natural language processing, machine learning, and artificial intelligence.
3 commits
Dataset type: LLaVA-Plus-v1-117K is a set of GPT-generated multimodal tool-augmented instruction-following data. It is constructed for tool use to build large multimodal agents with GPT-4-plus vision/language capability.
Dataset date: LLaVA-Plus-v1-117K was collected in Sep 2023, by prompting ChatGPT/GPT-4-0314 API.
Paper or resources for more information: https://llava-vl.github.io/llava-plus
License: Attribution-NonCommercial 4.0 International It should abide by the policy of OpenAI: https://openai.com/policies/terms-of-use
Where to send questions or comments about the model: https://github.com/LLaVA-VL/LLaVA-Plus-Codebase/issues
Primary intended uses: The primary use of LLaVA-Plus is research on large multimodal agents, and chatbots.
Primary intended users: The primary intended users of the model are researchers and hobbyists in computer vision, natural language processing, machine learning, and artificial intelligence.
3 commits