Skip to content
ModelPriceLab

Search model prices

Find source-linked rates by model, vendor, or platform.

nvidia

esm2-650m

Active

Open Llama instruction model for multilingual chat, reasoning, and coding

View current API pricing for this model
Input, output, and cached rates with source and verification time, plus same-workload cost examples.
Go to pricing page →
Input / 1M tokens
$0.0000
Output / 1M tokens
$0.0000
Record updated 2026-10-11
Modality:Specialized
Context:128,000 tokens
Release:Aug 29, 2024

Pricing

Platform
Scope
Input
Output
Details
Official
Standard
$0.0000
$0.0000
128k context · 8k max output

Latest changes

No recent changes.

Benchmarks

No benchmark data available.

⚡esm2-650m Comparisons & Scenario Guides

❓esm2-650m Frequently Asked Questions

How much does esm2-650m cost per 1M tokens?▼

esm2-650m is priced at $0.0000 per 1M input tokens and $0.0000 per 1M output tokens on official. Check provider terms before deployment.

What is the context window for esm2-650m?▼

esm2-650m supports up to 128,000 tokens context with a max completion limit of 8,192 tokens.

What are cost-saving alternatives to esm2-650m?▼

ModelPriceLab automatically calculates cheaper alternatives in the same model class, showing monthly spend differences and migration guides.

AI Cost Optimization

Does this change your best option?

Check your models and usage for free. See the relevant official changes, estimated cost, and a practical next step.