All AI models · LLM benchmarks · Methodology

xAI

Grok 4.1 Fast

Fast Grok model for responsive chat, reasoning, and tool-assisted work

Model facts

Context window
2M tokens
Maximum output
3K tokens
Cheapest paid input
$0.2 per 1M tokens
Cheapest paid output
$0.5 per 1M tokens
Open weights
no
Providers
6

Capabilities and modalities

reasoning, tool calling, structured output, attachments, temperature control, text, image.

Artificial Analysis benchmarks

Compare this model on the full LLM benchmark leaderboard.

ConfigurationAAICodingMathOutput tokens/s
Reasoning31.3not measured89.3not measured
Non-reasoning17.0not measured34.3not measured

Observed price history

  • 2026-08-24: $0.2 input / $0.5 output per 1M tokens

API providers

  • Abacus (Non-Reasoning) — model id grok-4-1-fast-non-reasoning; $0.2 input / $0.5 output per 1M tokens; provider documentation
  • Azure (Non-Reasoning) — model id grok-4-1-fast-non-reasoning; $0.2 input / $0.5 output per 1M tokens; provider documentation
  • FrogBot (Non-Reasoning) — model id grok-4-1-fast-non-reasoning; $0.2 input / $0.5 output per 1M tokens; provider documentation
  • Ofox — model id x-ai/grok-4.1-fast; $0.2 input / $0.5 output per 1M tokens; provider documentation
  • Perplexity Agent (Non-Reasoning) — model id xai/grok-4-1-fast-non-reasoning; $0.2 input / $0.5 output per 1M tokens; provider documentation
  • ZenMux — model id x-ai/grok-4.1-fast; $0.2 input / $0.5 output per 1M tokens; provider documentation