Llama Nemotron Embed VL 1B V2 and Embed V1 4B are available through the OpenRouter API, so switching between them takes a model slug change rather than a new integration.
Llama Nemotron Embed VL 1B V2, from Nvidia, has a 131,072-token context window and is priced by the provider serving it on OpenRouter.
Embed V1 4B, from Perplexity, has a 32,000-token context window and is priced at $0.03/M tokens on OpenRouter.