Grok 4 Fast (reasoning)

grok-4-fast-reasoning · Grok

Grok-4-fast is a cost-effective inference model developed by xAI that delivers cutting-edge performance with excellent token efficiency. The model features a 2 million token context window, advanced Web and X search capabilities, and a unified architecture supporting both "inference" and "non-inference" modes. Compared to Grok 4, it reduces thinking tokens by an average of 40% and lowers the price by 98% while achieving the same performance.

API Pricing

Input$0.2 / 1M tokens
Output$0.5 / 1M tokens
Cache read$0.05 / 1M tokens

Specifications

Context2M tokens
Modalitiestext, image
Featuresthinking, tool calling, function calling, structured outputs

Frequently asked questions

What is Grok 4 Fast (reasoning)?

Grok-4-fast is a cost-effective inference model developed by xAI that delivers cutting-edge performance with excellent token efficiency. The model features a 2 million token context window, advanced Web and X search capabilities, and a unified architecture supporting both "inference" and "non-inference" modes. Compared to Grok 4, it reduces thinking tokens by an average of 40% and lowers the price by 98% while achieving the same performance.

What is the context length of Grok 4 Fast (reasoning)?

Grok 4 Fast (reasoning) has a 2,000,000 token context window.

How much does Grok 4 Fast (reasoning) cost?

On AIHubMix, Grok 4 Fast (reasoning) costs $0.2 per million input tokens and $0.5 per million output tokens. Cached input reads are billed at $0.05 per million tokens.

What modalities does Grok 4 Fast (reasoning) support?

Grok 4 Fast (reasoning) accepts text and image input.

What capabilities does Grok 4 Fast (reasoning) support?

Grok 4 Fast (reasoning) supports thinking, tool calling, function calling and structured outputs. Per-protocol parameter support is listed in the capability table on this page.

How do I call Grok 4 Fast (reasoning) via API?

Grok 4 Fast (reasoning) is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to grok-4-fast-reasoning — no other code changes needed.

Who created Grok 4 Fast (reasoning)?

Grok 4 Fast (reasoning) is developed by Grok. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

More models from Grok

See all Grok models →

Grok 4.6

by Grok

Grok 4.6 is xAI’s (SpaceXAI) flagship multimodal reasoning model for coding, long-running…

$2/1M in · $6/1M out
500,000 tokens context

Grok 4.5

by Grok

Grok 4.5 was trained on datasets spanning knowledge in coding, science, engineering, and…

$2/1M in · $6/1M out
500,000 tokens context

Grok Build 0.1

by Grok

Fast coding model trained specifically for agentic coding workflows.

$1/1M in · $2/1M out
256,000 tokens context

Grok 4.3

by Grok

Grok 4.3 is amongst the leading models in intelligence and well priced when comparing to…

$1.25/1M in · $2.5/1M out
1,000,000 tokens context

grok-4-20-non-reasoning

by Grok

Grok 4.2 is xAI’s latest large language model, built for strong reasoning, multimodal…

$2/1M in · $6/1M out
2,000,000 tokens context

Grok 4 20 (reasoning)

by Grok

Grok 4.2 is xAI’s latest large language model, built for strong reasoning, multimodal…

$2/1M in · $6/1M out
2,000,000 tokens context

Use Grok 4 Fast (reasoning) via the AIHubMix unified API — one interface for every major LLM.