Skip to content
OpenRouterOpenRouter
© 2026 OpenRouter, Inc

Product

  • Chat
  • Rankings
  • Benchmarks
  • Apps
  • Discover
  • Models
  • Collections
  • Providers
  • Pricing
  • Enterprise
  • Labs

Company

  • About
  • Blog
  • Careers
    Hiring
  • Privacy
  • Terms of Service
  • Trust Center
  • Support
  • Works With OR
  • Data
  • Brand

Developer

  • Documentation
  • API Reference
  • Developer Platform
  • Status

Connect

  • Discord
  • GitHub
  • LinkedIn
  • X
  • YouTube
Favicon for liuhaotian

LLaVA 13B

liuhaotian/llava-13b

Model weights

LLaVA is a large multimodal model that combines a vision encoder and Vicuna for general-purpose visual and language understanding, achieving impressive chat capabilities and setting a new state-of-the-art accuracy on Science QA.

#multimodal

Modalities

Context

2K

Released

Nov 16, 2023

Knowledge Cutoff

Jun 2023

ActivityFAQExplore

Activity

Token volume and request traffic to this model over time.

Explore more models

AI Models with Vision: Multimodal LLMs for Image UnderstandingCollectionAI Model RankingsRanking

Frequently asked questions

LLaVA is a large multimodal model that combines a vision encoder and Vicuna for general-purpose visual and language understanding, achieving impressive chat capabilities and setting a new state-of-the-art accuracy on Science QA. #multimodal

LLaVA 13B has a 2,048 token context window.

LLaVA 13B accepts text and images as input and returns text.

LLaVA 13B was released on November 16, 2023. Its knowledge cutoff is June 30, 2023.