llama-nemotron-embed-vl-1b-v2
ActiveEmbedding model for semantic search, retrieval, clustering, and ranking pipelines
View current API pricing for this model
Input, output, and cached rates with source and verification time, plus same-workload cost examples.
Input / 1M tokens
$0.0000
Output / 1M tokens
$0.0000
Record updated 2026-10-11
Modality:modality_embed
Context:32,768 tokens
Release:Feb 10, 2026
Pricing
Platform
Scope
Input
Output
Details
Official
Standard
$0.0000
$0.0000
33k context · 2k max output
Latest changes
No recent changes.
Benchmarks
No benchmark data available.
⚡llama-nemotron-embed-vl-1b-v2 Comparisons & Scenario Guides
Popular Price & Spec Comparisons
Related Scenario Rankings
❓llama-nemotron-embed-vl-1b-v2 Frequently Asked Questions
How much does llama-nemotron-embed-vl-1b-v2 cost per 1M tokens?▼
llama-nemotron-embed-vl-1b-v2 is priced at $0.0000 per 1M input tokens and $0.0000 per 1M output tokens on official. Check provider terms before deployment.
What is the context window for llama-nemotron-embed-vl-1b-v2?▼
llama-nemotron-embed-vl-1b-v2 supports up to 32,768 tokens context with a max completion limit of 2,048 tokens.
What are cost-saving alternatives to llama-nemotron-embed-vl-1b-v2?▼
ModelPriceLab automatically calculates cheaper alternatives in the same model class, showing monthly spend differences and migration guides.
AI Cost Optimization
Does this change your best option?
Check your models and usage for free. See the relevant official changes, estimated cost, and a practical next step.