The Yi Vision is a complex visual task models provide high-performance understanding and analysis capabilities based on multiple images.
It's ideal for scenarios that require analysis and interpretation of images and charts, such as image question answering, chart understanding, OCR, visual reasoning, education, research report understanding, or multilingual document reading.
Modalities
Context
16K
Released
Aug 2, 2024
Knowledge Cutoff
Mar 2024
The Yi Vision is a complex visual task models provide high-performance understanding and analysis capabilities based on multiple images. It's ideal for scenarios that require analysis and interpretation of images and charts, such as image question answering, chart understanding, OCR, visual reasoning, education, research report understanding, or multilingual document reading.
Yi Vision has a 16,384 token context window.
Yi Vision accepts text and images as input and returns text.
Yi Vision was released on August 2, 2024. Its knowledge cutoff is March 31, 2024.
Token volume and request traffic to this model over time.