OpenGVLab

community

https://github.com/opengvlab

opengvlab

OpenGVLab

Activity Feed Request to join this org

AI & ML interests

Computer Vision

Recent Activity

czczup updated a dataset about 1 hour ago

OpenGVLab/InternVL-Data

Weiyun1025 updated a dataset about 1 hour ago

OpenGVLab/InternVL-Data

wzk1015 updated a model 1 day ago

OpenGVLab/PIIP-LLaVA-Plus_ConvNeXt-L_CLIP-L_1024-336_7B

View all activity

Organization Card

Community About org cards

OpenGVLab

Welcome to OpenGVLab! We are a research group from Shanghai AI Lab focused on Vision-Centric AI research. The GV in our name, OpenGVLab, means general vision, a general understanding of vision, so little effort is needed to adapt to new vision-based tasks.

Models

InternVL: a pioneering open-source alternative to GPT-4V.
InternImage: a large-scale vision foundation models with deformable convolutions.
InternVideo: large-scale video foundation models for multimodal understanding.
VideoChat: an end-to-end chat assistant for video comprehension.
All-Seeing-Project: towards panoptic visual recognition and understanding of the open world.

Datasets

ShareGPT4o: a groundbreaking large-scale resource that we plan to open-source with 200K meticulously annotated images, 10K videos with highly descriptive captions, and 10K audio files with detailed descriptions.
InternVid: a large-scale video-text dataset for multimodal understanding and generation.
MMPR: a high-quality, large-scale multimodal preference dataset.

Benchmarks

MVBench: a comprehensive benchmark for multimodal video understanding.
CRPE: a benchmark covering all elements of the relation triplets (subject, predicate, object), providing a systematic platform for the evaluation of relation comprehension ability.
MM-NIAH: a comprehensive benchmark for long multimodal documents comprehension.
GMAI-MMBench: a comprehensive multimodal evaluation benchmark towards general medical AI.

Collections 24

spaces 11

InternVideo2.5

Hierarchical Compression for Long-Context Video Modeling

InternVL

Chat with an AI that understands text and images

MVBench Leaderboard

Submit model evaluation and view leaderboard

Running on Zero

InternVideo2 Chat 8B HD

Upload a video to chat about its contents

ControlLLM

Display maintenance message for ControlLLM

Running on Zero

VideoMamba

Classify video and image content

models 216

OpenGVLab/PIIP-LLaVA-Plus_ConvNeXt-L_CLIP-L_1024-336_7B

Image-Text-to-Text • Updated 1 day ago • 4

OpenGVLab/clip-vit-large-patch14to16-224

Updated 1 day ago • 6

OpenGVLab/PIIP-LLaVA_CLIP-BL_512-256_7B

Image-Text-to-Text • Updated 1 day ago • 4

OpenGVLab/PIIP-LLaVA_ConvNeXt-B_CLIP-L_1024-336_7B

Image-Text-to-Text • Updated 1 day ago • 6

OpenGVLab/PIIP-LLaVA_ConvNeXt-L_CLIP-L_1024-336_7B

Image-Text-to-Text • Updated 1 day ago • 5

OpenGVLab/clip-vit-large-patch14to16-336

Updated 1 day ago • 85

OpenGVLab/PIIP-LLaVA_CLIP-BL_512-448_7B

Image-Text-to-Text • Updated 1 day ago • 18

OpenGVLab/PIIP-LLaVA_ConvNeXt-L_CLIP-L_1024-336_13B

Image-Text-to-Text • Updated 1 day ago • 4

OpenGVLab/PIIP-LLaVA_ConvNeXt-B_CLIP-L_640-224_7B

Image-Text-to-Text • Updated 1 day ago • 5

OpenGVLab/PIIP-LLaVA_ConvNeXt-B_CLIP-L_1024-336_13B

Image-Text-to-Text • Updated 1 day ago • 5

datasets 41

OpenGVLab/InternVL-Data

Updated about 1 hour ago • 31 • 10

OpenGVLab/VisualPRM400K-v1.1

Preview • Updated 4 days ago • 98 • 4

OpenGVLab/VisualPRM400K-v1.1-Raw

Preview • Updated 6 days ago • 46 • 2

OpenGVLab/VisualPRM400K

Preview • Updated 6 days ago • 378 • 7

OpenGVLab/MMPR-v1.2-prompts

Updated 6 days ago • 234 • 1

OpenGVLab/MMPR-v1.2

Updated 6 days ago • 344 • 9

OpenGVLab/MMPR-v1.1

Preview • Updated 9 days ago • 260 • 45

OpenGVLab/MMPR

Preview • Updated 10 days ago • 123 • 48

OpenGVLab/LongVid

Preview • Updated 24 days ago • 27 • 2

OpenGVLab/NIAH-Video

Viewer • Updated 26 days ago • 629 • 109