LLM List: AI models compared
Browse all LLMs in one comprehensive list: 1,781 large language models and AI models across 214 API providers. Compare token prices, context windows, capabilities, benchmark scores, open-weight availability, and local hardware requirements. Updated every six hours.
Explore the LLM benchmark leaderboard · How prices, benchmarks and model identities are calculated
- GPT-5.4 mini by OpenAI — 4K context; $0.375 input / $2.25 output per 1M tokens
- GPT-5.4 nano by OpenAI — 4K context; $0.1 input / $0.625 output per 1M tokens
- GPT-5.4 Pro by OpenAI — 1.05M context; $15 input / $90 output per 1M tokens
- gpt-5.4-high — not published context; not published input / not published output per 1M tokens
- gpt-5.4-medium — not published context; not published input / not published output per 1M tokens
- gpt-5.4-mini-2026-03-17 — 4K context; $0.75 input / $4.5 output per 1M tokens
- gpt-5.4-mini-high — not published context; not published input / not published output per 1M tokens
- gpt-5.4-nano-2026-03-17 — 4K context; $0.2 input / $1.25 output per 1M tokens
- gpt-5.4-nano-high — not published context; not published input / not published output per 1M tokens
- gpt-5.4-search — not published context; not published input / not published output per 1M tokens
- GPT-5.5 by OpenAI — 1.05M context; $0.188 input / $1.12 output per 1M tokens
- GPT-5.5 Instant by OpenAI — 4K context; $5 input / $30 output per 1M tokens
- GPT-5.5 Pro by OpenAI — 1.05M context; $15 input / $90 output per 1M tokens
- gpt-5.5-high — not published context; not published input / not published output per 1M tokens
- gpt-5.5-search — not published context; not published input / not published output per 1M tokens
- gpt-5.5-xhigh — not published context; not published input / not published output per 1M tokens
- GPT-5.6 — 1.05M context; $1.5 input / $12 output per 1M tokens
- GPT-5.6 Luna by OpenAI — 1.05M context; $0.1 input / $0.6 output per 1M tokens
- GPT-5.6 Luna Pro by OpenAI — 1.05M context; $0.1 input / $0.6 output per 1M tokens
- GPT-5.6 Sol by OpenAI — 1.05M context; $1 input / $5 output per 1M tokens
- GPT-5.6 Sol Pro by OpenAI — 1.05M context; $1 input / $5 output per 1M tokens
- GPT-5.6 Terra by OpenAI — 1.05M context; $1.5 input / $2 output per 1M tokens
- GPT-5.6 Terra Pro by OpenAI — 1.05M context; $1 input / $6 output per 1M tokens
- gpt-5.6-luna-low — not published context; not published input / not published output per 1M tokens
- gpt-5.6-luna-max — not published context; not published input / not published output per 1M tokens
- gpt-5.6-luna-medium — not published context; not published input / not published output per 1M tokens
- gpt-5.6-luna-xhigh — not published context; not published input / not published output per 1M tokens
- gpt-5.6-sol-low — not published context; not published input / not published output per 1M tokens
- gpt-5.6-sol-max — not published context; not published input / not published output per 1M tokens
- gpt-5.6-sol-medium — not published context; not published input / not published output per 1M tokens
- gpt-5.6-sol-search-xhigh — not published context; not published input / not published output per 1M tokens
- gpt-5.6-sol-xhigh — not published context; not published input / not published output per 1M tokens
- gpt-5.6-terra-low — not published context; not published input / not published output per 1M tokens
- gpt-5.6-terra-max — not published context; not published input / not published output per 1M tokens
- gpt-5.6-terra-medium — not published context; not published input / not published output per 1M tokens
- gpt-5.6-terra-xhigh — not published context; not published input / not published output per 1M tokens
- gpt-image-1-mini — not published context; not published input / not published output per 1M tokens
- GPT-OSS 120B TEE — 131K context; $2 input / $2 output per 1M tokens
- GPT-OSS 20B TEE — 131K context; $0.2 input / $0.8 output per 1M tokens
- GPT-OSS-120B-CS — 128K context; $0.35 input / $0.75 output per 1M tokens
- GPT-Realtime mini by OpenAI — not published context; $0.6 input / $2.4 output per 1M tokens
- gpt-realtime-2 by OpenAI — not published context; $4 input / $24 output per 1M tokens
- GPT-Realtime-2.1 by OpenAI — 128K context; $4 input / $24 output per 1M tokens
- gpt-realtime-whisper by OpenAI — not published context; not published input / not published output per 1M tokens
- Granite 4.0 H Micro by IBM — 131K context; $0.017 input / $0.112 output per 1M tokens
- Granite 4.0 Micro by IBM — 131K context; $0.017 input / $0.112 output per 1M tokens
- Granite 4.1 8B by IBM — 131K context; $0.05 input / $0.1 output per 1M tokens
- Granite 4.2 8B by IBM — 131K context; $0.1 input / $0.15 output per 1M tokens
- Granite-4.0-H-Small by IBM — 131K context; $0.064 input / $0.265 output per 1M tokens
- Grayline Qwen3 8B — 33K context; $0.3 input / $0.3 output per 1M tokens
- Green L — 128K context; $0.285 input / $0.912 output per 1M tokens
- Green L Raw — 128K context; $0.285 input / $0.912 output per 1M tokens
- Green R — 131K context; $0.399 input / $1.08 output per 1M tokens
- Green R Raw — 131K context; $0.399 input / $1.08 output per 1M tokens
- Green S — not published context; $0.0044 input / free output per 1M tokens
- Green S Pro — not published context; $0.0044 input / free output per 1M tokens
- Greg (Roleplay) — 229K context; $0.1 input / $0.3 output per 1M tokens
- Greg 1 Mini — 229K context; $0.07 input / $0.15 output per 1M tokens
- Greg 2 Super — 229K context; $1.5 input / $5 output per 1M tokens
- Greg 2 Ultra — 229K context; $3 input / $10 output per 1M tokens
- Grok 3 by xAI — 131K context; $3 input / $15 output per 1M tokens
- Grok 3 Mini by xAI — 131K context; $0.3 input / $0.5 output per 1M tokens
- Grok 4 by xAI — 256K context; $3 input / $15 output per 1M tokens
- Grok 4 Fast by xAI — 2M context; $0.2 input / $0.5 output per 1M tokens
- Grok 4.1 Fast by xAI — 2M context; $0.2 input / $0.5 output per 1M tokens
- Grok 4.1 Fast Non-Reasoning by xAI — 2M context; $0.18 input / $0.45 output per 1M tokens
- Grok 4.1 Fast Reasoning by xAI — 2M context; $0.18 input / $0.45 output per 1M tokens
- Grok 4.2 Fast by xAI — 2M context; $3 input / $9 output per 1M tokens
- Grok 4.2 Fast Non Reasoning by xAI — 2M context; $3 input / $9 output per 1M tokens
- Grok 4.20 by xAI — 2M context; $1.25 input / $2.5 output per 1M tokens
- Grok 4.20 (Non-Reasoning) by xAI — 2M context; $1.25 input / $2.5 output per 1M tokens
- Grok 4.20 (Reasoning) by xAI — 2M context; $1.25 input / $2.5 output per 1M tokens
- Grok 4.20 Beta Non-Reasoning — 2M context; $1.25 input / $2.5 output per 1M tokens
- Grok 4.20 Beta Non-Reasoning (0309) (xAI) by xAI — 2M context; $2 input / $6 output per 1M tokens
- Grok 4.20 Beta Reasoning — 2M context; $1.25 input / $2.5 output per 1M tokens
- Grok 4.20 Beta Reasoning (0309) (xAI) by xAI — 2M context; $2 input / $6 output per 1M tokens
- Grok 4.20 Multi Agent Beta — 2M context; $1.25 input / $2.5 output per 1M tokens
- Grok 4.20 Multi-Agent by xAI — 2M context; $1.25 input / $2.5 output per 1M tokens
- Grok 4.20 Non-Reasoning (Vertex AI (OpenAI-compatible)) — 2M context; $1.25 input / $2.5 output per 1M tokens
- Grok 4.20 Reasoning (Vertex AI (OpenAI-compatible)) — 2M context; $1.25 input / $2.5 output per 1M tokens
- Grok 4.3 by xAI — 1M context; $1.25 input / $2.5 output per 1M tokens
- Grok 4.5 by xAI — 5K context; $2 input / $6 output per 1M tokens
- Grok 4.6 by xAI — 5K context; $2 input / $6 output per 1M tokens
- Grok Build 0.1 by xAI — 256K context; $1 input / $2 output per 1M tokens
- Grok Code Fast 1 by xAI — 256K context; $0.18 input / $1.35 output per 1M tokens
- Grok Latest by xAI — 5K context; $2 input / $6 output per 1M tokens
- Grok STT — not published context; not published input / not published output per 1M tokens
- Grok Voice Think Fast 1.0 — not published context; not published input / not published output per 1M tokens
- Grok Voice Think Fast 2.0 — not published context; not published input / not published output per 1M tokens
- grok-4-0709 — 256K context; $2.7 input / $13.5 output per 1M tokens
- grok-4-1-fast-search — not published context; not published input / not published output per 1M tokens
- grok-4-fast-non-reasoning by xAI — 2M context; $0.18 input / $0.45 output per 1M tokens
- grok-4-fast-reasoning by xAI — 2M context; $0.18 input / $0.45 output per 1M tokens
- grok-4-search — not published context; not published input / not published output per 1M tokens
- grok-4.1 — 2K context; $2 input / $10 output per 1M tokens
- grok-4.2-beta — 2M context; $2 input / $6 output per 1M tokens
- grok-4.20-beta-0309-non-reasoning — 2M context; $2 input / $6 output per 1M tokens
- grok-4.20-beta-0309-reasoning — 2M context; $2 input / $6 output per 1M tokens
- grok-4.20-multi-agent-beta-0309 — 2M context; $2 input / $6 output per 1M tokens
- grok-4.3-high — not published context; not published input / not published output per 1M tokens