grok-4-1-fast-reasoning - overcharge

Company MG 0 Reputation points
2026-07-27T01:40:32.1733333+00:00

Hi there,

I'm reaching out to see if someone can provide some clarification or guidance regarding the pricing for grok-4-1-fast-reasoning.

I was using grok-4-1-fast-reasoning to run some experiments. According to the pricing documentation: https://azure.microsoft.com/en-us/pricing/details/ai-foundry-models/grok/

the model is priced at $0.20 per million input tokens and $0.50 per million output tokens.

My usage metrics show:

Total token count: 1.29M (9,548 average per request)

Input tokens: 1.25M (9,285 average per request)

Output tokens: 35.52K

Estimated total cost: $9.42

Based on the documented pricing, I would have expected the estimated cost to be significantly lower than what is being reported in the metrics. The reported estimate of $9.42 doesn't seem to align with the published input and output token rates.

Could someone help me understand how the estimated cost is calculated? Are there any additional charges, different pricing tiers, or other factors that could explain the discrepancy?

Thanks in advance for any clarification.

Foundry Models
Foundry Models

A catalog of AI models in Microsoft Foundry that you can discover, compare, and deploy using Azure’s built‑in tools for evaluation, fine‑tuning, and inference

0 comments No comments

Your answer

Answers can be marked as 'Accepted' by the question author and 'Recommended' by moderators, which helps users know the answer solved the author's problem.