Skip to content
ModelPriceLab

Search model prices

Find source-linked rates by model, vendor, or platform.

Model Services/Microsoft Azure/GPT-3.5 Turbo 0125
azure

GPT-3.5 Turbo 0125

Deprecated

Compact GPT model for low-latency assistance and high-volume workloads

Deprecated

Deprecated endpoints may still be callable. Migrate before retirement.

View current API pricing for this model
Input, output, and cached rates with source and verification time, plus same-workload cost examples.
Go to pricing page →
Input / 1M tokens
$0.50
Output / 1M tokens
$1.50
Record updated 2026-10-11
Modality:LLM
Context:16,384 tokens
Release:Jan 25, 2024

Pricing

Platform
Scope
Input
Output
Details
Azure
Standard
$0.50
$1.50
16k context · 4k max output

How much are you overpaying?

Enter your monthly spend on this model to see what the same workload could cost on cheaper alternatives.

$
Enter a monthly spend to see the listed-rate comparison.

Latest changes

No recent changes.

Benchmarks

No benchmark data available.

⚡GPT-3.5 Turbo 0125 Comparisons & Scenario Guides

❓GPT-3.5 Turbo 0125 Frequently Asked Questions

How much does GPT-3.5 Turbo 0125 cost per 1M tokens?▼

GPT-3.5 Turbo 0125 is priced at $0.50 per 1M input tokens and $1.50 per 1M output tokens on azure. Check provider terms before deployment.

What is the context window for GPT-3.5 Turbo 0125?▼

GPT-3.5 Turbo 0125 supports up to 16,384 tokens context with a max completion limit of 4,096 tokens.

What are cost-saving alternatives to GPT-3.5 Turbo 0125?▼

ModelPriceLab automatically calculates cheaper alternatives in the same model class, showing monthly spend differences and migration guides.

AI Cost Optimization

Does this change your best option?

Check your models and usage for free. See the relevant official changes, estimated cost, and a practical next step.