-
OpenGVLab/InternVideo2_5_Chat_8B
Video-Text-to-Text • Updated • 13k • 40 -
OpenGVLab/InternVL_2_5_HiCo_R16
Video-Text-to-Text • Updated • 481 • 2 -
OpenGVLab/InternVL_2_5_HiCo_R64
Video-Text-to-Text • Updated • 212 • 1 -
InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling
Paper • 2501.12386 • Published • 1
![](https://cdn-avatars.huggingface.co/v1/production/uploads/64006c09330a45b03605bba3/FvdxiTkTqH8rKDOzGZGUE.jpeg)
OpenGVLab
community
AI & ML interests
Computer Vision
Recent Activity
View all activity
Organization Card
OpenGVLab
Welcome to OpenGVLab! We are a research group from Shanghai AI Lab focused on Vision-Centric AI research. The GV in our name, OpenGVLab, means general vision, a general understanding of vision, so little effort is needed to adapt to new vision-based tasks.
Models
- InternVL: a pioneering open-source alternative to GPT-4V.
- InternImage: a large-scale vision foundation models with deformable convolutions.
- InternVideo: large-scale video foundation models for multimodal understanding.
- VideoChat: an end-to-end chat assistant for video comprehension.
- All-Seeing-Project: towards panoptic visual recognition and understanding of the open world.
Datasets
- ShareGPT4o: a groundbreaking large-scale resource that we plan to open-source with 200K meticulously annotated images, 10K videos with highly descriptive captions, and 10K audio files with detailed descriptions.
- InternVid: a large-scale video-text dataset for multimodal understanding and generation.
- MMPR: a high-quality, large-scale multimodal preference dataset.
Benchmarks
- MVBench: a comprehensive benchmark for multimodal video understanding.
- CRPE: a benchmark covering all elements of the relation triplets (subject, predicate, object), providing a systematic platform for the evaluation of relation comprehension ability.
- MM-NIAH: a comprehensive benchmark for long multimodal documents comprehension.
- GMAI-MMBench: a comprehensive multimodal evaluation benchmark towards general medical AI.
Collections
20
spaces
11
Running
426
InternVL
⚡
Chat with an AI that understands text and images
Runtime error
VideoChat Flash
💬
Hierarchical Compression for Long-Context Video Modeling
Running
33
MVBench Leaderboard
🐨
Submit model evaluation and view leaderboard
Running
on
Zero
16
InternVideo2 Chat 8B HD
👁
Upload a video to chat about its contents
Running
10
ControlLLM
🚀
Display maintenance message for ControlLLM
Running
on
Zero
93
VideoMamba
🐍
Classify video and image content
models
149
![](https://cdn-avatars.huggingface.co/v1/production/uploads/64006c09330a45b03605bba3/FvdxiTkTqH8rKDOzGZGUE.jpeg)
OpenGVLab/VideoChat-Flash-Qwen2_5-7B_InternVideo2-1B
Video-Text-to-Text
•
Updated
![](https://cdn-avatars.huggingface.co/v1/production/uploads/64006c09330a45b03605bba3/FvdxiTkTqH8rKDOzGZGUE.jpeg)
OpenGVLab/VideoChat-Flash-Qwen2_5-7B-1M_res224
Video-Text-to-Text
•
Updated
•
1
![](https://cdn-avatars.huggingface.co/v1/production/uploads/64006c09330a45b03605bba3/FvdxiTkTqH8rKDOzGZGUE.jpeg)
OpenGVLab/VideoChat-Flash-Qwen2_5-2B_res448
Video-Text-to-Text
•
Updated
•
1.81k
•
12
![](https://cdn-avatars.huggingface.co/v1/production/uploads/64006c09330a45b03605bba3/FvdxiTkTqH8rKDOzGZGUE.jpeg)
OpenGVLab/VideoChat-Flash-Qwen2-7B_res224
Video-Text-to-Text
•
Updated
•
797
•
3
![](https://cdn-avatars.huggingface.co/v1/production/uploads/64006c09330a45b03605bba3/FvdxiTkTqH8rKDOzGZGUE.jpeg)
OpenGVLab/VideoChat-Flash-Qwen2-7B_res448
Video-Text-to-Text
•
Updated
•
716
•
8
![](https://cdn-avatars.huggingface.co/v1/production/uploads/64006c09330a45b03605bba3/FvdxiTkTqH8rKDOzGZGUE.jpeg)
OpenGVLab/InternVideo2_5_Chat_8B
Video-Text-to-Text
•
Updated
•
13k
•
40
![](https://cdn-avatars.huggingface.co/v1/production/uploads/64006c09330a45b03605bba3/FvdxiTkTqH8rKDOzGZGUE.jpeg)
OpenGVLab/VBench_Used_Models
Updated
•
1
![](https://cdn-avatars.huggingface.co/v1/production/uploads/64006c09330a45b03605bba3/FvdxiTkTqH8rKDOzGZGUE.jpeg)
OpenGVLab/InternVL_2_5_HiCo_R16
Video-Text-to-Text
•
Updated
•
481
•
2
![](https://cdn-avatars.huggingface.co/v1/production/uploads/64006c09330a45b03605bba3/FvdxiTkTqH8rKDOzGZGUE.jpeg)
OpenGVLab/InternVL_2_5_HiCo_R64
Video-Text-to-Text
•
Updated
•
212
•
1
![](https://cdn-avatars.huggingface.co/v1/production/uploads/64006c09330a45b03605bba3/FvdxiTkTqH8rKDOzGZGUE.jpeg)
OpenGVLab/InternVideo2-Stage2_6B
Video Classification
•
Updated
•
143
datasets
30
OpenGVLab/MMPR-v1.1
Preview
•
Updated
•
541
•
39
OpenGVLab/MMPR
Preview
•
Updated
•
116
•
46
OpenGVLab/GMAI-MMBench
Preview
•
Updated
•
112
•
15
OpenGVLab/V2PE-Data
Preview
•
Updated
•
335
•
6
OpenGVLab/InternVL-Domain-Adaptation-Data
Preview
•
Updated
•
242
•
8
OpenGVLab/GUI-Odyssey
Viewer
•
Updated
•
7.74k
•
24.7k
•
11
OpenGVLab/OmniCorpus-YT
Updated
•
574
•
12
OpenGVLab/OmniCorpus-CC-210M
Viewer
•
Updated
•
208M
•
108
•
19
OpenGVLab/OmniCorpus-CC
Viewer
•
Updated
•
986M
•
4.48k
•
12
OpenGVLab/MVBench
Viewer
•
Updated
•
4k
•
14.4k
•
29