Downloads · 30 days
0
Efficient-Large-Model/ShareGPT4V-fp8
ShareGPT4V-fp8 is a machine learning model from Efficient-Large-Model. Use it for the machine learning task on the model card, and read the license before you ship it in a product. The card lists the license as cc-by-nc-4.0.
Downloads · 30 days
0
Access
Public
Updated Sep 23, 2024
Repo size
44.5 GB
Likes
0
Public
Click a slice to open those files.
.zip27.4 GB · 59%
From the Hugging Face model README
Dataset type:
Use filter-sharegpt4v_instruct_gpt4-vision_cap100k.json and filter-share-captioner_coco_lcs_sam_1246k_1107.json for experiments.
Dataset type: ShareGPT4V Captions 1.2M is a set of GPT4-Vision-powered multi-modal captions data.
It is constructed to enhance modality alignment and fine-grained visual concept perception in Large Multi-Modal Models (LMMs) during both the pre-training and supervised fine-tuning stages. This advancement aims to bring LMMs towards GPT4-Vision capabilities.
Dataset date: ShareGPT4V Captions 1.2M was collected in 11.07 2023.
Paper or resources for more information: [Project] [Paper] [Code]
License: Attribution-NonCommercial 4.0 International It should abide by the policy of OpenAI: https://openai.com/policies/terms-of-use
Primary intended uses: The primary use of ShareGPT4V Captions 1.2M is research on large multimodal models and chatbots.
Primary intended users: The primary intended users of this dataset are researchers and hobbyists in computer vision, natural language processing, machine learning, and artificial intelligence.