azure_ai/Llama-4-Maverick-17B-128E-Instruct-FP8

Azure OpenAI text

azure_ai/Llama-4-Maverick-17B-128E-Instruct-FP8

Input

$0.2500

per 1M input tokens

Output

$1.00

per 1M output tokens

Cache read

n/a

per 1M cached read tokens

Cache write

n/a

per 1M cache write tokens

Context window

1,000,000

Max output

16,384

Effective date

Sep 17, 2026

Estimate a workload

Enter token counts to see the cost at this model's rates.

Estimated cost: $0.00

Price history

blended $ per 1M tokens (3:1)
$0.35 $0.65 $0.94 $1.23 Jun 15, 2026: $1.15 Sep 17, 2026: $0.44 Jun 15, 2026 Sep 17, 2026
Effective Input Output Blended
Sep 17, 2026 $0.25 $1.00 $0.44
Jun 15, 2026 $1.41 $0.35 $1.15

Source: litellm. Confirm against the provider's official pricing before relying on these figures.