LLM List: AI models compared
Browse all LLMs in one comprehensive list: 1,781 large language models and AI models across 214 API providers. Compare token prices, context windows, capabilities, benchmark scores, open-weight availability, and local hardware requirements. Updated every six hours.
Explore the LLM benchmark leaderboard · How prices, benchmarks and model identities are calculated
- grok-4.5-search — not published context; not published input / not published output per 1M tokens
- grok-4.6-high — not published context; not published input / not published output per 1M tokens
- Groq-Llama-4-Maverick-17B-128E-Instruct — 128K context; free input / free output per 1M tokens
- GTE Large (v1.5) — 8K context; $0.09 input / free output per 1M tokens
- Hermes 2 Pro Llama 3 8B by Nous Research — 131K context; $0.14 input / $0.14 output per 1M tokens
- Hermes 3 405B Instruct by Nous Research — 131K context; $1 input / $1 output per 1M tokens
- Hermes 3 70B by Nous Research — 131K context; $0.408 input / $0.408 output per 1M tokens
- Hermes 3 Llama 3.1 405b — 128K context; $1.1 input / $3 output per 1M tokens
- Hermes 4 (Thinking) by Nous Research — 128K context; $0.201 input / $0.4 output per 1M tokens
- Hermes 4 405B by Nous Research — 131K context; $0.996 input / $2.99 output per 1M tokens
- Hermes 4 70B by Nous Research — 131K context; $0.129 input / $0.399 output per 1M tokens
- Hermes 4 Large by Nous Research — 128K context; $0.3 input / $1.2 output per 1M tokens
- Hermes 4 Large (Thinking) by Nous Research — 128K context; $0.3 input / $1.2 output per 1M tokens
- Hermes High — 1.05M context; $1 input / $3.2 output per 1M tokens
- Hermes Low — 1.05M context; $1 input / $3.2 output per 1M tokens
- Hermes Medium — 1.05M context; $1 input / $3.2 output per 1M tokens
- Holo2 30B A3B — 22K context; $0.399 input / $0.969 output per 1M tokens
- Holo3-35B-A3B — 66K context; $0.25 input / $1.8 output per 1M tokens
- Holo3-35B-A3B Thinking — 66K context; $0.25 input / $1.8 output per 1M tokens
- Hunyuan A13B Instruct by Tencent — 131K context; $0.14 input / $0.57 output per 1M tokens
- Hunyuan-T1 — 131K context; free input / free output per 1M tokens
- Hunyuan-TurboS — 131K context; free input / free output per 1M tokens
- hunyuan-vision-1.5-thinking — not published context; not published input / not published output per 1M tokens
- Hy-MT2-1.8B by Tencent — 8K context; $0.044 input / $0.177 output per 1M tokens
- Hy-MT2-30B-A3B by Tencent — 8K context; $0.074 input / $0.295 output per 1M tokens
- Hy-MT2-7B by Tencent — 8K context; $0.074 input / $0.295 output per 1M tokens
- Hy3 by Tencent — 262K context; $0.132 input / $0.528 output per 1M tokens
- Hy3 Free — 19K context; free input / free output per 1M tokens
- Hy3 preview by Tencent — 262K context; $0.066 input / $0.26 output per 1M tokens
- Hy3 preview Free — 256K context; free input / free output per 1M tokens
- Hy4 preview by Tencent — 1.02M context; $0.67 input / $2 output per 1M tokens
- ibm-granite-h-small — not published context; not published input / not published output per 1M tokens
- InclusionAI Ling 3.0 Flash — 262K context; $0.06 input / $0.18 output per 1M tokens
- Inkling by Thinking Machines — 1.05M context; $0.95 input / $4.05 output per 1M tokens
- Inkling (256K) — 262K context; $1.87 input / $4.68 output per 1M tokens
- Inkling Small by Thinking Machines — 524K context; $0.45 input / $1.2 output per 1M tokens
- Inkling Small Thinking by Thinking Machines — 524K context; $0.5 input / $1.2 output per 1M tokens
- Inkling Thinking by Thinking Machines — 1.05M context; $1 input / $4.05 output per 1M tokens
- inkling-low — not published context; not published input / not published output per 1M tokens
- inkling-medium — not published context; not published input / not published output per 1M tokens
- inkling-small-low — not published context; not published input / not published output per 1M tokens
- inkling-small-medium — not published context; not published input / not published output per 1M tokens
- intellect-3 by Prime Intellect — not published context; not published input / not published output per 1M tokens
- Interfaze Beta — 1M context; $1.5 input / $3.5 output per 1M tokens
- Isometry — 262K context; $0.1 input / $0.45 output per 1M tokens
- K2-Think — 128K context; $0.17 input / $0.68 output per 1M tokens
- Kat Coder Air V2.5 by Kwaipilot — 256K context; $0.15 input / $0.6 output per 1M tokens
- Kat Coder Pro by Kwaipilot — 256K context; $0.3 input / $1.2 output per 1M tokens
- Kat Coder Pro V2 by Kwaipilot — 262K context; $0.3 input / $1.2 output per 1M tokens
- KAT-Coder-Pro V1 by Kwaipilot — 256K context; $0.3 input / $1.2 output per 1M tokens
- KAT-Coder-Pro V2.5 by Kwaipilot — 262K context; $0.74 input / $2.96 output per 1M tokens
- KB Whisper — 448 context; $0.0023 input / $0.0023 output per 1M tokens
- Kimi (latest) by Moonshot AI — 1.05M context; $1.79 input / $8.94 output per 1M tokens
- Kimi For Coding HighSpeed — 262K context; free input / free output per 1M tokens
- Kimi K2 by Moonshot AI — 131K context; $0.4 input / $1.8 output per 1M tokens
- Kimi K2 0711 by Moonshot AI — 131K context; $0.4 input / $1.8 output per 1M tokens
- Kimi K2 0711 Fast — 131K context; $0.4 input / $1.8 output per 1M tokens
- Kimi K2 0711 Instruct FP4 — 131K context; $0.4 input / $1.8 output per 1M tokens
- Kimi K2 0905 by Moonshot AI — 262K context; $0.4 input / $1.8 output per 1M tokens
- Kimi K2 Thinking by Moonshot AI — 262K context; $0.47 input / $2 output per 1M tokens
- Kimi K2 Thinking Turbo by Moonshot AI — 262K context; $1.15 input / $8 output per 1M tokens
- Kimi K2 Turbo — 262K context; $2.4 input / $10 output per 1M tokens
- Kimi K2 Turbo Preview — 256K context; $0.15 input / $8 output per 1M tokens
- Kimi K2.5 by Moonshot AI — 262K context; $0.3 input / $1.9 output per 1M tokens
- Kimi K2.5 Free — 262K context; free input / free output per 1M tokens
- Kimi K2.5 Thinking by Moonshot AI — 256K context; $0.3 input / $1.9 output per 1M tokens
- Kimi K2.6 by Moonshot AI — 262K context; $0.22 input / $1.14 output per 1M tokens
- Kimi K2.6 Fast — 262K context; $1.66 input / $8.78 output per 1M tokens
- Kimi K2.6 Nitro — 2K context; $0.275 input / $1.1 output per 1M tokens
- Kimi K2.6 TEE by Moonshot AI — 262K context; $0.58 input / $3.4 output per 1M tokens
- Kimi K2.6 Thinking by Moonshot AI — 256K context; $0.5 input / $2.6 output per 1M tokens
- Kimi K2.7 Code by Moonshot AI — 262K context; $0.275 input / $1.1 output per 1M tokens
- Kimi K2.7 Code Fast — 262K context; $0.95 input / $4 output per 1M tokens
- Kimi K2.7 Code Flex — 262K context; $0.618 input / $2.6 output per 1M tokens
- Kimi K2.7 Code Highspeed by Moonshot AI — 262K context; $1.9 input / $8 output per 1M tokens
- Kimi K2.7 Code Nitro — 2K context; $0.275 input / $1.1 output per 1M tokens
- Kimi K2.7 Code TEE — 262K context; $0.95 input / $4 output per 1M tokens
- Kimi K3 by Moonshot AI — 1.05M context; $2 input / $8 output per 1M tokens
- Kimi K3 Eco — 1M context; $1 input / $4 output per 1M tokens
- Kimi K3 Fast by Moonshot AI — 1M context; $3 input / $15 output per 1M tokens
- Kimi K3 Flex — 1.05M context; $1.95 input / $9.75 output per 1M tokens
- Kimi K3 TEE by Moonshot AI — 1.05M context; $3 input / $15 output per 1M tokens
- Kimi K3-256K — 262K context; free input / free output per 1M tokens
- kimi-k2-0711-preview — not published context; not published input / not published output per 1M tokens
- kimi-k2-0905-preview — 262K context; $0.632 input / $2.53 output per 1M tokens
- Kimi-K2-Instruct-0905 by Moonshot AI — 262K context; $1 input / $3 output per 1M tokens
- Kimi-K2.5-FW — 262K context; free input / free output per 1M tokens
- kimi-k2.5-instant — not published context; not published input / not published output per 1M tokens
- kimi/kimi-k2.5 — 262K context; $0.6 input / $3 output per 1M tokens
- Kloker — 2K context; $0.23 input / $1.16 output per 1M tokens
- Kloker Integration Architect — 2K context; $0.23 input / $1.16 output per 1M tokens
- Kloker Integration Developer — 2K context; $0.23 input / $1.16 output per 1M tokens
- L3 70B Euryale V2.1 by Sao10K — 8K context; $1.48 input / $1.48 output per 1M tokens
- L3 8B Stheno V3.2 by Sao10K — 8K context; $0.05 input / $0.05 output per 1M tokens
- L31 70B Euryale V2.2 by Sao10K — 8K context; $1.48 input / $1.48 output per 1M tokens
- Laguna M.1 — 262K context; free input / free output per 1M tokens
- Laguna S 2.1 by Poolside — 1.05M context; $0.09 input / $0.18 output per 1M tokens
- Laguna S 2.1 Free by Poolside — 256K context; free input / free output per 1M tokens
- Laguna S 2.1 Thinking by Poolside — 1.05M context; $0.1 input / $0.2 output per 1M tokens
- Laguna XS 2.1 by Poolside — 262K context; $0.06 input / $0.12 output per 1M tokens