Llama Nemotron Embed VL 1B V2 and GTE-Large are available through the OpenRouter API, so switching between them takes a model slug change rather than a new integration.
Llama Nemotron Embed VL 1B V2, from Nvidia, has a 131,072-token context window and is priced by the provider serving it on OpenRouter.
GTE-Large, from thenlper, has a 512-token context window and is priced at $0.01/M tokens on OpenRouter.