grok-4-fast-reasoning

by Grok

Grok-4-fast is a cost-effective inference model developed by xAI that delivers cutting-edge performance with excellent token efficiency. The model features a 2 million token context window, advanced Web and X search capabilities, and a unified architecture supporting both "inference" and "non-inference" modes. Compared to Grok 4, it reduces thinking tokens by an average of 40% and lowers the price by 98% while achieving the same performance.

API Pricing

Input$0.2 / 1M tokens
Output$0.5 / 1M tokens
Cache read$0.05 / 1M tokens

Specifications

Context window2,000,000 tokens
Modalitiestext, image
Featuresthinking, tool calling, function calling, structured outputs

FAQ

What is grok-4-fast-reasoning?

Grok-4-fast is a cost-effective inference model developed by xAI that delivers cutting-edge performance with excellent token efficiency. The model features a 2 million token context window, advanced Web and X search capabilities, and a unified architecture supporting both "inference" and "non-inference" modes. Compared to Grok 4, it reduces thinking tokens by an average of 40% and lowers the price by 98% while achieving the same performance.

What is the context length of grok-4-fast-reasoning?

grok-4-fast-reasoning has a 2,000,000 token context window.

How much does grok-4-fast-reasoning cost?

On AIHubMix, grok-4-fast-reasoning costs $0.2 per million input tokens and $0.5 per million output tokens. Cached input reads are billed at $0.05 per million tokens.

What modalities does grok-4-fast-reasoning support?

grok-4-fast-reasoning accepts text and image input.

What features does grok-4-fast-reasoning support?

grok-4-fast-reasoning supports thinking, tool calling, function calling and structured outputs. Per-protocol parameter support is listed in the capability table on this page.

How do I call grok-4-fast-reasoning via API?

grok-4-fast-reasoning is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to grok-4-fast-reasoning — no other code changes needed.

Who develops grok-4-fast-reasoning?

grok-4-fast-reasoning is developed by Grok. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

More from Grok

grok-4.5

by Grok

Grok 4.5 was trained on datasets spanning knowledge in coding, science, engineering, and…

$2/1M in · $6/1M out
1,000,000 tokens context

grok-build-0.1

by Grok

Fast coding model trained specifically for agentic coding workflows.

$1/1M in · $2/1M out
256,000 tokens context

grok-4.3

by Grok

Grok 4.3 is amongst the leading models in intelligence and well priced when comparing to…

$1.25/1M in · $2.5/1M out
1,000,000 tokens context

grok-4-20-non-reasoning

by Grok

Grok 4.2 is xAI’s latest large language model, built for strong reasoning, multimodal…

$2/1M in · $6/1M out
2,000,000 tokens context

grok-4-20-reasoning

by Grok

Grok 4.2 is xAI’s latest large language model, built for strong reasoning, multimodal…

$2/1M in · $6/1M out
2,000,000 tokens context

grok-4.20-multi-agent-0309

by Grok

Grok 4.20 is our newest flagship model with industry-leading speed and agentic tool…

$2/1M in · $6/1M out
2,000,000 tokens context

Use grok-4-fast-reasoning via the AIHubMix unified API — one interface for every major LLM.