nvidia · Embeddings
llama-nemotron-embed-vl-1b-v2 API
llama-nemotron-embed-vl-1b-v2 by nvidia is listed in ServerNet for Embeddings. Review current availability and pricing before connecting.
PreparingToman per million tokens
Sale prices will appear after provider configuration is complete.
- Context window
- 10,240
- Gateway output cap
- 4,096
Model features
10% tax is added to usage charges. Prices are frozen at request admission and settlement uses reported consumption. Only available models are exposed by the API. Availability requires a configured provider, valid prices and a supported execution path.