LLM List: AI models compared
Browse all LLMs in one comprehensive list: 1,781 large language models and AI models across 214 API providers. Compare token prices, context windows, capabilities, benchmark scores, open-weight availability, and local hardware requirements. Updated every six hours.
Explore the LLM benchmark leaderboard · How prices, benchmarks and model identities are calculated
- Gemini 2.5 Flash 0520 Thinking — 1.05M context; $0.15 input / $3.5 output per 1M tokens
- Gemini 2.5 Flash Image by Google — 33K context; $0.3 input / $2.5 output per 1M tokens
- Gemini 2.5 Flash Lite Preview by Google — 1.05M context; $0.15 input / $0.6 output per 1M tokens
- Gemini 2.5 Flash Lite Preview (09/2025) – Thinking — 1.05M context; $0.1 input / $0.4 output per 1M tokens
- Gemini 2.5 Flash Preview by Google — 1.05M context; $0.15 input / $0.6 output per 1M tokens
- Gemini 2.5 Flash Preview (09/2025) — 1.05M context; $0.3 input / $2.5 output per 1M tokens
- Gemini 2.5 Flash Preview (09/2025) – Thinking — 1.05M context; $0.3 input / $2.5 output per 1M tokens
- Gemini 2.5 Flash Preview Thinking — 1.05M context; $0.15 input / $3.5 output per 1M tokens
- Gemini 2.5 Flash-Lite by Google — 1.05M context; $0.1 input / $0.1 output per 1M tokens
- Gemini 2.5 Pro by Google — 1.05M context; $0.625 input / $5 output per 1M tokens
- Gemini 2.5 Pro Experimental 0325 — 1.05M context; $2.5 input / $10 output per 1M tokens
- Gemini 2.5 Pro Preview 0325 — 1.05M context; $2.5 input / $10 output per 1M tokens
- Gemini 2.5 Pro Preview 05-06 by Google — 1.05M context; $1.25 input / $10 output per 1M tokens
- Gemini 2.5 Pro Preview 06-05 by Google — 1.05M context; $1.12 input / $9 output per 1M tokens
- Gemini 3 Flash by Google — 1.05M context; $0.4 input / $2.4 output per 1M tokens
- Gemini 3 Flash (Preview) (Google AI Studio) — 1.05M context; $0.5 input / $3 output per 1M tokens
- Gemini 3 Flash (Preview) (Google Vertex AI) — 1.05M context; $0.5 input / $3 output per 1M tokens
- Gemini 3 Flash Preview by Google — 1.05M context; $0.07 input / $0.43 output per 1M tokens
- Gemini 3 Flash Thinking by Google — 1.05M context; $0.5 input / $3 output per 1M tokens
- Gemini 3 Pro by Google — 1.05M context; $1.6 input / $9.6 output per 1M tokens
- Gemini 3 Pro Image by Google — 1.05M context; $2 input / $12 output per 1M tokens
- Gemini 3 Pro Preview by Google — 1.05M context; $0.57 input / $3.43 output per 1M tokens
- Gemini 3.0 Flash Preview — 1M context; not published input / not published output per 1M tokens
- Gemini 3.0 Pro Image Preview — 33K context; not published input / not published output per 1M tokens
- Gemini 3.0 Pro Preview — 1M context; not published input / not published output per 1M tokens
- Gemini 3.1 Flash Image by Google — 33K context; $0.5 input / $3 output per 1M tokens
- Gemini 3.1 Flash Image Preview (Nano Banana 2) by Google — 131K context; $0.5 input / $3 output per 1M tokens
- Gemini 3.1 Flash Lite by Google — 1.05M context; $0.125 input / $0.75 output per 1M tokens
- Gemini 3.1 Flash Lite Image (Nano Banana 2 Lite) by Google — 66K context; $0.25 input / $1.5 output per 1M tokens
- Gemini 3.1 Flash Lite Preview by Google — 1.05M context; $0.125 input / $0.75 output per 1M tokens
- Gemini 3.1 Flash Live Preview — 131K context; $0.75 input / $4.5 output per 1M tokens
- Gemini 3.1 Pro (Preview High) by Google — 1.05M context; $2 input / $12 output per 1M tokens
- Gemini 3.1 Pro (Preview Low) by Google — 1.05M context; $2 input / $12 output per 1M tokens
- Gemini 3.1 Pro (Preview) (Google AI Studio) — 1.05M context; $2 input / $12 output per 1M tokens
- Gemini 3.1 Pro (Preview) (Google Vertex AI) — 1.05M context; $2 input / $12 output per 1M tokens
- Gemini 3.1 Pro (Preview) (Quartz) — 1.05M context; $2 input / $12 output per 1M tokens
- Gemini 3.1 Pro Preview by Google — 1.05M context; $1 input / $6 output per 1M tokens
- Gemini 3.1 Pro Preview Custom Tools by Google — 1.05M context; $2 input / $12 output per 1M tokens
- Gemini 3.5 Flash by Google — 1.05M context; $0.186 input / $1.11 output per 1M tokens
- Gemini 3.5 Flash Lite by Google — 1.05M context; $0.15 input / $1.25 output per 1M tokens
- Gemini 3.5 Flash Thinking by Google — 1.05M context; $1.5 input / $9 output per 1M tokens
- Gemini 3.5 Live Translate Preview — 16K context; $3.5 input / $21 output per 1M tokens
- Gemini 3.5 Transcribe by Google — not published context; $2 input / $12 output per 1M tokens
- Gemini 3.5 Transcribe Live by Google — not published context; not published input / not published output per 1M tokens
- Gemini 3.6 Flash by Google — 1.05M context; $0.375 input / $1.88 output per 1M tokens
- Gemini 3.7 Flash by Google — 1.05M context; $0.375 input / $1.88 output per 1M tokens
- Gemini 3.8 Flash by Google — 1.05M context; $0.375 input / $1.88 output per 1M tokens
- Gemini Embedding 001 by Google — 2K context; $0.15 input / free output per 1M tokens
- Gemini Embedding 2 by Google — 8K context; $0.2 input / free output per 1M tokens
- Gemini Flash Latest by Google — 1.05M context; $0.5 input / $3 output per 1M tokens
- Gemini Flash-Lite Latest by Google — 1.05M context; $0.25 input / $1.5 output per 1M tokens
- Gemini Omni Flash Preview by Google — 1M context; $1.5 input / $9 output per 1M tokens
- Gemini Pro Latest by Google — 1.05M context; $2 input / $12 output per 1M tokens
- Gemini Robotics-ER 1.6 Preview by Google — 131K context; $1 input / $5 output per 1M tokens
- gemini-2.0-flash-001 — not published context; not published input / not published output per 1M tokens
- gemini-2.5-flash-lite-preview-06-17 — 1.05M context; $0.09 input / $0.36 output per 1M tokens
- gemini-2.5-flash-lite-preview-09-2025 — 1.05M context; $0.09 input / $0.36 output per 1M tokens
- gemini-2.5-flash-nothink — 1M context; $0.3 input / $2.5 output per 1M tokens
- gemini-2.5-flash-preview-05-20 — 1.05M context; $0.135 input / $3.15 output per 1M tokens
- gemini-2.5-pro-grounding — not published context; not published input / not published output per 1M tokens
- gemini-3-flash (thinking-minimal) — not published context; not published input / not published output per 1M tokens
- gemini-3-flash-grounding — not published context; not published input / not published output per 1M tokens
- gemini-3-pro-image-preview — 33K context; $2 input / $120 output per 1M tokens
- gemini-3.1-flash-image-preview — 131K context; $0.5 input / $60 output per 1M tokens
- Gemini-3.1-Pro by Google — 1.05M context; $2 input / $12 output per 1M tokens
- gemini-3.1-pro-grounding — not published context; not published input / not published output per 1M tokens
- gemini-3.5-flash-high — not published context; not published input / not published output per 1M tokens
- gemini-3.8-flash-high — not published context; not published input / not published output per 1M tokens
- gemini-3.8-flash-low — not published context; not published input / not published output per 1M tokens
- gemini-3.8-flash-medium — not published context; not published input / not published output per 1M tokens
- gemini-deep-research by Google — 1.05M context; $1.6 input / $9.6 output per 1M tokens
- Gemma 2 — 8K context; $0.01 input / $0.03 output per 1M tokens
- Gemma 2 27B by Google — 8K context; $0.65 input / $0.65 output per 1M tokens
- Gemma 2 2b It by Google — 128K context; free input / free output per 1M tokens
- Gemma 3 by Google — 125K context; $0.15 input / $0.3 output per 1M tokens
- Gemma 3 12B by Google — 131K context; $0.05 input / $0.1 output per 1M tokens
- Gemma 3 27B by Google — 131K context; $0.08 input / $0.16 output per 1M tokens
- Gemma 3 4B by Google — 131K context; $0.04 input / $0.08 output per 1M tokens
- Gemma 3 4B (Pretrained) by Google — 33K context; $0.15 input / $0.15 output per 1M tokens
- Gemma 3n E2b It by Google — 128K context; free input / free output per 1M tokens
- Gemma 3N E4B Instruct by Google — 33K context; $0.06 input / $0.12 output per 1M tokens
- Gemma 3n E4b It by Google — 128K context; free input / free output per 1M tokens
- Gemma 4 — 262K context; $0.18 input / $0.5 output per 1M tokens
- Gemma 4 12B Instruct by Google — 262K context; $0.05 input / $0.25 output per 1M tokens
- Gemma 4 12B IT by Google — 33K context; $0.25 input / $0.25 output per 1M tokens
- Gemma 4 26B A4B by Google — 262K context; $0.042 input / $0.22 output per 1M tokens
- Gemma 4 26B A4B IT — 262K context; $0.07 input / $0.34 output per 1M tokens
- Gemma 4 26B A4B MeroMero — 262K context; $0.12 input / $0.38 output per 1M tokens
- Gemma 4 26B A4B MeroMero Thinking — 262K context; $0.12 input / $0.38 output per 1M tokens
- Gemma 4 26B A4B Thinking by Google — 262K context; $0.13 input / $0.4 output per 1M tokens
- Gemma 4 26B A4B Uncensored — 262K context; $0.12 input / $0.38 output per 1M tokens
- Gemma 4 26B A4B Uncensored TEE — 66K context; $0.15 input / $0.7 output per 1M tokens
- Gemma 4 26B A4B Uncensored Thinking — 262K context; $0.12 input / $0.38 output per 1M tokens
- Gemma 4 31B by Google — 262K context; $0.102 input / $0.297 output per 1M tokens
- Gemma 4 31B Claude 4.6 Opus Reasoning Distilled — 262K context; $0.306 input / $0.306 output per 1M tokens
- Gemma 4 31B Cognitive Unshackled — 262K context; $0.306 input / $0.306 output per 1M tokens
- Gemma 4 31B DarkIdol — 262K context; $0.306 input / $0.306 output per 1M tokens
- Gemma 4 31B Garnet V2 — 262K context; $0.306 input / $0.306 output per 1M tokens
- Gemma 4 31B IT — 262K context; $0.102 input / $0.297 output per 1M tokens
- Gemma 4 31B IT FP8 — 262K context; free input / free output per 1M tokens