Open MLLM family (1B-78B) from OpenGVLab. Excels at vision, reasoning, long context & agents via native multimodal pre-training. Outperforms base LLMs on text tasks.
Open MLLMs excelling in vision, reasoning & long context
Open MLLM family (1B-78B) from OpenGVLab. Excels at vision, reasoning, long context & agents via native multimodal pre-training. Outperforms base LLMs on text tasks.
Hi everyone! Check out InternVL3 from OpenGVLab – a new family of open vision-language models. They used a training approach mixing vision and text data from the start, which reportedly leads to strong performance in both understanding images/video and handling text tasks well. These models show good reasoning abilities and can handle long inputs. The weights and code are openly available. You can experience these model capabilities directly on their Chat Web and HF Space .
This is a powerful addition to the multimodal AI ecosystem. Vision, reasoning, and long context handling are so crucial — well done to the team!
The Open MLLM family is truly impressive! What stands out is how well these models handle vision and reasoning tasks while outperforming base LLMs even on text benchmarks. The native multimodal pre-training approach seems to be a game-changer. Can't wait to see what the community will build with these models. Wishing the OpenGVLab team continued success with this project!
Exciting for AI researchers working with multimodal tasks! 😄
Hey Zac Zuo & the OpenGVLab team (congrats on the hunt/launch!), this looks like a significant step forward for open vision-language models. Exciting to see strong performance reported from native multimodal pre-training, especially in reasoning and handling long context alongside vision tasks. As we're building AI experiences ( @UNI AI ), having powerful, open models like InternVL3 available is fantastic for the ecosystem. The ability to handle both image/video and text tasks well from the
Categories come from the product's launch tags. Most products appear in 2-3 categories. The primary category is listed first.
The scores reflect launch-period engagement. Historical data is preserved and doesn't change retroactively. The build date at the bottom shows when the index was last refreshed.
Check the similar products section on this page, or browse the category pages linked in the tags above. Each category page shows all products for a given year, sorted by engagement.
A measure of community engagement at launch. Higher means more people noticed and interacted with the product. It's a traction signal, not a quality rating.