{T}
TokenMath.net
Verified against xAI Official Pricing
xAIbudgetGenerally AvailableActive Production

Grok 4.6 Mini Token Counter & Cost Calculator

High-efficiency xAI model for classification and fast agent loops

Standard Input / 1M
$0.40
Prompt tokens
Cached Input / 1M
$0.10
Prompt caching rate
Output / 1M
$1.60
Generation completion
Context Window
512k
Max out: 64k
API & Developer Integration Details
generally available
API Model Stringgrok-4.6-mini
Provider & FamilyxAI (o200k_base)
Prompt Caching Rules1,024 min tokens (75% off)
Long-Context TiersStandard flat rate throughout full window
First-party audit timestamp: 2026-09-23 04:43 UTCOfficial Documentation

Interactive Cost & Token Simulator for Grok 4.6 Mini

Live Calculation
Input Tokens2,500
Output Tokens800
Requests / Day5,000
Prompt Caching Ratio (50%)$0.1/M cached rate
Cost / Request$0.00191
Daily Spend (5,000 reqs)$9.525
Monthly Run-Rate (30d)$285.75

Standard Workload Cost Scenarios

ScenarioInput TokensOutput TokensUncached CostWith Prompt Caching
Short Chat Query1,000500$0.00120$0.00090
Document Summarization10,0002,000$0.00720$0.00420
Codebase & Context Analysis100,00020,000$0.0720$0.0420
Batch Corpus Processing1,000,000100,000$0.5600$0.2600
Technical Architecture & Pricing VerificationVerified: 2026-09-23 04:43 UTC
Model ArchitectureGrok 4.6 Mini
Provider OrganizationxAI
Tokenizer EncodingOpenAI o200k_base legacy (200k vocabulary)
API ConnectivityxAI API (/v1/chat/completions).
Batch API 50% DiscountSupported (50% off input & output)
Tier 1 Rate Quota (Free-Tier Reference)1,000 RPM · 1M TPM

Tier 1 (free tier) reference limits — your account's actual quota may be higher.

Blended 3:1 Cost / 1M Tokens$0.70
Budget companion to Grok 4.6 Flagship. Automatic prompt caching reads discount 75% ($0.10/1M) with no cache write surcharge.xAI Official Pricing

Verified Pricing History

September 2026: $0.4/M input · $1.6/M output ($0.1/M cached)
Grok 4.6 Mini launch pricing.

When to Choose Grok 4.6 Mini

Ideal for production workloads demanding budget capabilities, deep context depth (512k tokens), and reliability from xAI. Excellent when predictable tokenomics and prompt caching support are paramount.

When Another Model May Be Better

If your use-case requires sub-second streaming latency or ultra-high frequency classification at micro-cent pricing, consider lighter budget options such as Gemini Flash-Lite or Claude Haiku. For deep formal logic, consider dedicated reasoning models like o3.

Frequently Asked Questions About Grok 4.6 Mini

How much does 1 million tokens cost with Grok 4.6 Mini?

For Grok 4.6 Mini, 1 million input tokens costs $0.40, while 1 million output tokens costs $1.60. If using prompt caching, repetitive input prefixes are discounted to $0.10 per million.

What is the context window for Grok 4.6 Mini?

Grok 4.6 Mini features a maximum context window of 512,000 tokens (~384,000 words), with a maximum output limit of 65,536 tokens per completion.

Which tokenizer does Grok 4.6 Mini use?

Grok 4.6 Mini utilizes the OpenAI o200k_base legacy (200k vocabulary). Token counting on TokenMath runs client-side to ensure maximum privacy.

Does Grok 4.6 Mini support prompt caching discounts?

Yes. Grok 4.6 Mini supports prompt caching with a cached input rate of $0.10/1M (saving up to 75% on repeated input context).