Expanding Performance Boundaries of Open-Source MLLM
OpenGVLab
community
AI & ML interests
Computer Vision
Organization Card
OpenGVLab
Welcome to OpenGVLab! We are a research group from Shanghai AI Lab focused on Vision-Centric AI research. The GV in our name, OpenGVLab, means general vision, a general understanding of vision, so little effort is needed to adapt to new vision-based tasks.
Models
- InternVL: a pioneering open-source alternative to GPT-4V.
- InternImage: a large-scale vision foundation models with deformable convolutions.
- InternVideo: large-scale video foundation models for multimodal understanding.
- VideoChat: an end-to-end chat assistant for video comprehension.
- All-Seeing-Project: towards panoptic visual recognition and understanding of the open world.
Datasets
- ShareGPT4o: a groundbreaking large-scale resource that we plan to open-source with 200K meticulously annotated images, 10K videos with highly descriptive captions, and 10K audio files with detailed descriptions.
- InternVid: a large-scale video-text dataset for multimodal understanding and generation.
Benchmarks
- MVBench: a comprehensive benchmark for multimodal video understanding.
Collections
11
A Pioneering Open-Source Alternative to GPT-4V
-
OpenGVLab/InternVL-Chat-V1-5
Image-Text-to-Text • Updated • 12.9k • 399 -
OpenGVLab/InternVL-Chat-V1-5-AWQ
Image-Text-to-Text • Updated • 470 • 10 -
OpenGVLab/Mini-InternVL-Chat-4B-V1-5
Image-Text-to-Text • Updated • 1.75k • 57 -
OpenGVLab/Mini-InternVL-Chat-2B-V1-5
Image-Text-to-Text • Updated • 3.1k • 63
models
88
OpenGVLab/VideoChat2_HD_stage4_Mistral_7B_hf
Updated
•
23
OpenGVLab/InternVL2-8B-Pretrain
Updated
OpenGVLab/InternVideo2_Chat_8B_InternLM2_5
Video-Text-to-Text
•
Updated
•
494
•
2
OpenGVLab/InternVideo2_chat_8B_HD
Video-Text-to-Text
•
Updated
•
531
•
7
OpenGVLab/ViCLIP-L-14-hf
Updated
•
39
OpenGVLab/ViCLIP-B-16-hf
Updated
•
23
OpenGVLab/GUI-Odyssey
Updated
OpenGVLab/InternVideo2-Chat-8B
Video Classification
•
Updated
•
971
•
14
OpenGVLab/InternVL-Chat-ViT-6B-Vicuna-7B
Visual Question Answering
•
Updated
•
36
•
8
OpenGVLab/InternVL-Chat-ViT-6B-Vicuna-13B
Visual Question Answering
•
Updated
•
19
•
7
datasets
25
OpenGVLab/InternVL-SA-1B-Caption
Viewer
•
Updated
•
8.63M
•
5
OpenGVLab/InternVL-Chat-V1-2-SFT-Data
Viewer
•
Updated
•
573k
•
46
•
9
OpenGVLab/InternVL-LaionCOCO-OCR
Updated
OpenGVLab/InternVL-WuKong-OCR
Updated
OpenGVLab/GMAI-MMBench
Preview
•
Updated
•
4
•
10
OpenGVLab/GUI-Odyssey
Viewer
•
Updated
•
7.74k
•
8
•
6
OpenGVLab/ScaleVLN
Updated
OpenGVLab/OmniCorpus-CC-210M
Viewer
•
Updated
•
208M
•
27
•
8
OpenGVLab/ShareGPT-4o
Viewer
•
Updated
•
59.4k
•
229
•
128
OpenGVLab/MVBench
Viewer
•
Updated
•
4k
•
52k
•
21