Expanding Performance Boundaries of Open-Source MLLM
OpenGVLab
community
AI & ML interests
Computer Vision
Organization Card
About org cards
OpenGVLab
Welcome to OpenGVLab! We are a research group from Shanghai AI Lab focused on Vision-Centric AI research. The GV in our name, OpenGVLab, means general vision, a general understanding of vision, so little effort is needed to adapt to new vision-based tasks.
Models
- InternVL: a pioneering open-source alternative to GPT-4V.
- InternImage: a large-scale vision foundation models with deformable convolutions.
- InternVideo: large-scale video foundation models for multimodal understanding.
- VideoChat: an end-to-end chat assistant for video comprehension.
- All-Seeing-Project: towards panoptic visual recognition and understanding of the open world.
Datasets
- ShareGPT4o: a groundbreaking large-scale resource that we plan to open-source with 200K meticulously annotated images, 10K videos with highly descriptive captions, and 10K audio files with detailed descriptions.
- InternVid: a large-scale video-text dataset for multimodal understanding and generation.
Benchmarks
- MVBench: a comprehensive benchmark for multimodal video understanding.
Collections
10
A Pioneering Open-Source Alternative to GPT-4V
-
OpenGVLab/InternVL-Chat-V1-5
Image-Text-to-Text • Updated • 32.9k • 377 -
OpenGVLab/Mini-InternVL-Chat-4B-V1-5
Image-Text-to-Text • Updated • 22.9k • 51 -
OpenGVLab/Mini-InternVL-Chat-2B-V1-5
Image-Text-to-Text • Updated • 24.1k • 51 -
OpenGVLab/InternVL-Chat-V1-5-Int8
Image-Text-to-Text • Updated • 4.26k • 58
models
68
OpenGVLab/InternVL2-40B
Image-Text-to-Text
•
Updated
•
14
•
6
OpenGVLab/InternVL2-1B
Image-Text-to-Text
•
Updated
•
8
•
1
OpenGVLab/InternVL-Chat-V1-5-AWQ
Image-Text-to-Text
•
Updated
•
2.75k
•
9
OpenGVLab/InternVL-Chat-V1-5-Int8
Image-Text-to-Text
•
Updated
•
4.26k
•
58
OpenGVLab/InternVL-Chat-V1-5
Image-Text-to-Text
•
Updated
•
32.9k
•
377
OpenGVLab/Mini-InternVL-Chat-4B-V1-5
Image-Text-to-Text
•
Updated
•
22.9k
•
51
OpenGVLab/Mini-InternVL-Chat-2B-V1-5
Image-Text-to-Text
•
Updated
•
24.1k
•
51
OpenGVLab/InternVL2-26B
Image-Text-to-Text
•
Updated
•
1.41k
•
46
OpenGVLab/InternVL2-8B
Image-Text-to-Text
•
Updated
•
1.3k
•
18
OpenGVLab/InternVL2-4B
Image-Text-to-Text
•
Updated
•
526
•
6
datasets
17
OpenGVLab/VideoChat2-IT
Viewer
•
Updated
•
1.82M
•
132
•
33
OpenGVLab/MVBench
Viewer
•
Updated
•
4k
•
238
•
15
OpenGVLab/GUI-Odyssey
Viewer
•
Updated
•
7.74k
•
9
•
2
OpenGVLab/MMT-Bench
Viewer
•
Updated
•
30k
•
6
•
1
OpenGVLab/MM-NIAH
Viewer
•
Updated
•
3.52k
•
5
•
9
OpenGVLab/InternVid-Full
Viewer
•
Updated
•
47.6M
•
33
•
7
OpenGVLab/ShareGPT-4o
Viewer
•
Updated
•
59.4k
•
183
•
93
OpenGVLab/CRPE
Viewer
•
Updated
•
544
•
2
•
5
OpenGVLab/Region-Evaluation-Data
Preview
•
Updated
•
3
•
1
OpenGVLab/AS-Core
Preview
•
Updated
•
3
•
5