xAI
Grok 4.1 Fast

Grok 4.1 Fast is xAI's best agentic tool calling model that shines in real-world use cases like customer support and deep research. 2M context window.

Reasoning can be enabled/disabled using the reasoning enabled parameter in the API. Learn more in our docs

Model details

Context window2,000,000 tokens

Max completion size37 tokens

Prompt cost / 1K tokens$0.0000002

Completion cost / 1K tokens$0.0000005

Accepts

Produces

Benchmark performance

Overall

score

3rd

placement

Cost

score

4th

placement

Logic

score

3rd

placement

Speed

score

22nd

placement

Scoring

score

14th

placement

Tool Use

score

3rd

placement

Hallucination

score

5th

placement

Classification

score

1st

placement

Structured Output

score

2nd

placement

Pricing

Usage pricing
Prompt	$0.0000002
Completion	$0.0000005
Request	FREE
Image	FREE
Web Search	FREE
Internal Reasoning	FREE

Best Overall scoring LLMs

xAI

Grok 4 Fast

score

1st

placement

Qwen

Qwen3 VL 235B A22B Instruct

score

2nd

placement

xAI

Grok 4.1 Fast

score

3rd

placement

OpenAI

GPT-5.1 Chat

score

4th

placement

OpenAI

GPT-5.1-Codex

score

4th

placement

Anthropic

Claude Haiku 4.5

score

5th

placement

Browse all LLMs

Model details

Benchmark performanceAll scores have maximum of 100 points.

Overall

Cost

Logic

Speed

Scoring

Tool Use

Hallucination

Classification

Structured Output

Pricing

Grok 4 Fast

Qwen3 VL 235B A22B Instruct

Grok 4.1 Fast

GPT-5.1 Chat

GPT-5.1-Codex

Claude Haiku 4.5

Benchmark performance