The 14b version of the InternVL3 series. An advanced multimodal large language model (MLLM) series that demonstrates superior overall performance. Compared to InternVL 2.5, InternVL3 exhibits superior multimodal perception and reasoning capabilities, while further extending its multimodal capabilities to encompass tool usage, GUI agents, industrial image analysis, 3D vision perception, and more.
Modalities
Context
32K
Released
Apr 30, 2025
Knowledge Cutoff
Jan 2025
Token volume and request traffic to this model over time.
The 14b version of the InternVL3 series. An advanced multimodal large language model (MLLM) series that demonstrates superior overall performance. Compared to InternVL 2.5, InternVL3 exhibits superior multimodal perception and reasoning capabilities, while further extending its multimodal capabilities to encompass tool usage, GUI agents, industrial image analysis, 3D vision perception, and more.
InternVL3 14B has a 32,000 token context window.
InternVL3 14B accepts images and text as input and returns text.
InternVL3 14B was released on April 30, 2025. Its knowledge cutoff is January 31, 2025.