Grok 4.1 Fast (reasoning)

grok-4-1-fast-reasoning · Grok

Grok 4.1 is a new conversational model with significant improvements in real-world usability, delivering exceptional performance in creative, emotional, and collaborative interactions. It is more perceptive to nuanced user intent, more engaging to converse with, and more coherent in personality, while fully preserving its core intelligence and reliability. Built on large-scale reinforcement learning infrastructure, the model is optimized for style, personality, helpfulness, and alignment, and leverages frontier agentic reasoning models as reward evaluators to autonomously assess and iterate on responses at scale, significantly enhancing overall interaction quality.

API Pricing

Input$0.2 / 1M tokens
Output$0.5 / 1M tokens
Cache read$0.05 / 1M tokens

Specifications

Context2M tokens
Modalitiestext, image
Featuresthinking, tool calling, function calling, structured outputs

Frequently asked questions

What is grok-4-1-fast-reasoning?

Grok 4.1 is a new conversational model with significant improvements in real-world usability, delivering exceptional performance in creative, emotional, and collaborative interactions. It is more perceptive to nuanced user intent, more engaging to converse with, and more coherent in personality, while fully preserving its core intelligence and reliability. Built on large-scale reinforcement learning infrastructure, the model is optimized for style, personality, helpfulness, and alignment, and leverages frontier agentic reasoning models as reward evaluators to autonomously assess and iterate on responses at scale, significantly enhancing overall interaction quality.

What is the context length of grok-4-1-fast-reasoning?

grok-4-1-fast-reasoning has a 2,000,000 token context window.

How much does grok-4-1-fast-reasoning cost?

On AIHubMix, grok-4-1-fast-reasoning costs $0.2 per million input tokens and $0.5 per million output tokens. Cached input reads are billed at $0.05 per million tokens.

What modalities does grok-4-1-fast-reasoning support?

grok-4-1-fast-reasoning accepts text and image input.

What capabilities does grok-4-1-fast-reasoning support?

grok-4-1-fast-reasoning supports thinking, tool calling, function calling and structured outputs. Per-protocol parameter support is listed in the capability table on this page.

How do I call grok-4-1-fast-reasoning via API?

grok-4-1-fast-reasoning is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to grok-4-1-fast-reasoning — no other code changes needed.

Who created grok-4-1-fast-reasoning?

grok-4-1-fast-reasoning is developed by Grok. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

More models from Grok

grok-4.6

by Grok

Grok 4.6 is xAI’s (SpaceXAI) flagship multimodal reasoning model for coding, long-running…

$2/1M in · $6/1M out
500,000 tokens context

grok-4.5

by Grok

Grok 4.5 was trained on datasets spanning knowledge in coding, science, engineering, and…

$2/1M in · $6/1M out
500,000 tokens context

grok-build-0.1

by Grok

Fast coding model trained specifically for agentic coding workflows.

$1/1M in · $2/1M out
256,000 tokens context

grok-4.3

by Grok

Grok 4.3 is amongst the leading models in intelligence and well priced when comparing to…

$1.25/1M in · $2.5/1M out
1,000,000 tokens context

grok-4-20-non-reasoning

by Grok

Grok 4.2 is xAI’s latest large language model, built for strong reasoning, multimodal…

$2/1M in · $6/1M out
2,000,000 tokens context

grok-4-20-reasoning

by Grok

Grok 4.2 is xAI’s latest large language model, built for strong reasoning, multimodal…

$2/1M in · $6/1M out
2,000,000 tokens context

Use grok-4-1-fast-reasoning via the AIHubMix unified API — one interface for every major LLM.