LLM List: AI models compared
Browse all LLMs in one comprehensive list: 1,781 large language models and AI models across 214 API providers. Compare token prices, context windows, capabilities, benchmark scores, open-weight availability, and local hardware requirements. Updated every six hours.
Explore the LLM benchmark leaderboard · How prices, benchmarks and model identities are calculated
- The Drummer Cydonia 24B v2 by TheDrummer — 33K context; $0.1 input / $0.121 output per 1M tokens
- The Drummer Cydonia 24B v4 by TheDrummer — 33K context; $0.201 input / $0.241 output per 1M tokens
- The Drummer Cydonia 24B v4.1 by TheDrummer — 131K context; $0.35 input / $0.55 output per 1M tokens
- The Drummer Cydonia 24B v4.3 by TheDrummer — 33K context; $0.12 input / $0.15 output per 1M tokens
- The Drummer Magidonia 24B v4.3 by TheDrummer — 33K context; $0.1 input / $0.121 output per 1M tokens
- thinkingcap-qwen3.6-27b — 262K context; $0.4 input / $3 output per 1M tokens
- thinkingcap-qwen3.6-27b@eu — 262K context; $0.4 input / $3 output per 1M tokens
- TIM-Qwen3.6 27B — 8K context; $0.3 input / $3 output per 1M tokens
- Titan Text Embeddings V2 by Amazon — 8K context; not published input / not published output per 1M tokens
- Tongyi Intent Detect V3 — 8K context; $0.058 input / $0.144 output per 1M tokens
- Transcribe-1 by Fish Audio — not published context; not published input / not published output per 1M tokens
- Trendyol Asure 12B — 41K context; $0.1 input / $0.5 output per 1M tokens
- Trinity Large Preview — 131K context; free input / free output per 1M tokens
- Trinity Large Thinking by Arcee AI — 262K context; $0.25 input / $0.8 output per 1M tokens
- Trinity Mini — 131K context; $0.045 input / $0.15 output per 1M tokens
- UI-TARS 7B by ByteDance — 128K context; $0.1 input / $0.2 output per 1M tokens
- Umans Coder — 262K context; $0.95 input / $4 output per 1M tokens
- Umans Flash — 262K context; $0.15 input / $1 output per 1M tokens
- Uncensored by Cognitive Computations — 128K context; $0.2 input / $0.9 output per 1M tokens
- Universal Summarizer — 33K context; $30 input / $30 output per 1M tokens
- UnslopNemo 12B by TheDrummer — 1.02M context; $0.4 input / $0.4 output per 1M tokens
- UnslopNemo 12b v4 by TheDrummer — 33K context; $0.493 input / $0.493 output per 1M tokens
- usdcode — 128K context; free input / free output per 1M tokens
- usdvalidate — not published context; free input / free output per 1M tokens
- v0-1.0-md — 128K context; $3 input / $15 output per 1M tokens
- v0-1.5-lg — 512K context; $15 input / $75 output per 1M tokens
- v0-1.5-md — 128K context; $3 input / $15 output per 1M tokens
- Veiled Calla 12B — 33K context; $0.3 input / $0.3 output per 1M tokens
- Venice Role Play Uncensored — 128K context; $0.5 input / $2 output per 1M tokens
- Venice Uncensored by Cognitive Computations — 128K context; $0.2 input / $0.9 output per 1M tokens
- Venice Uncensored 1.2 — 128K context; $0.2 input / $0.9 output per 1M tokens
- Voxtral Mini (latest) — not published context; not published input / not published output per 1M tokens
- Voxtral Mini 3B — 32K context; $0.0046 input / free output per 1M tokens
- Voxtral Mini 3B 2507 — 128K context; $0.04 input / $0.04 output per 1M tokens
- Voxtral Small (latest) by Mistral — 32K context; $0.1 input / $0.3 output per 1M tokens
- Voxtral Small 24B by Mistral — 33K context; $0.0023 input / $0.0023 output per 1M tokens
- Voxtral Small 24B 2507 by Mistral — 33K context; $0.1 input / $0.3 output per 1M tokens
- voxtral-small-2507 — 32K context; $0.111 input / $0.334 output per 1M tokens
- Voyage Rerank 2.5 by Voyage AI — 32K context; not published input / not published output per 1M tokens
- Voyage Rerank 2.5 Lite by Voyage AI — 32K context; not published input / not published output per 1M tokens
- voyage-3-large by Voyage AI — 8K context; not published input / not published output per 1M tokens
- voyage-3.5 by Voyage AI — 8K context; not published input / not published output per 1M tokens
- voyage-3.5-lite by Voyage AI — 8K context; not published input / not published output per 1M tokens
- voyage-4 by Voyage AI — 32K context; not published input / not published output per 1M tokens
- voyage-4-large by Voyage AI — 32K context; not published input / not published output per 1M tokens
- voyage-4-lite by Voyage AI — 32K context; not published input / not published output per 1M tokens
- voyage-code-2 by Voyage AI — 8K context; not published input / not published output per 1M tokens
- voyage-code-3 by Voyage AI — 8K context; not published input / not published output per 1M tokens
- voyage-finance-2 by Voyage AI — 8K context; not published input / not published output per 1M tokens
- voyage-law-2 by Voyage AI — 8K context; not published input / not published output per 1M tokens
- Weaver (alpha) — 8K context; $0.4 input / $0.75 output per 1M tokens
- Web Answer — 33K context; $7.5 input / $7.5 output per 1M tokens
- Whisper by OpenAI — not published context; not published input / not published output per 1M tokens
- Whisper 3 Large by OpenAI — 448 context; $0.0023 input / $0.0023 output per 1M tokens
- Whisper Large v3 by OpenAI — 448 context; $0.003 input / free output per 1M tokens
- Whisper Large v3 Turbo by OpenAI — 448 context; $0.0023 input / $0.0023 output per 1M tokens
- WizardLM-2 8x22B by Microsoft — 66K context; $0.493 input / $0.493 output per 1M tokens
- Writer: Palmyra X5 — 1.04M context; $0.6 input / $6 output per 1M tokens
- X-Ai/Grok 4.1 Fast Non Reasoning by xAI — 2M context; not published input / not published output per 1M tokens
- X-Ai/Grok 4.1 Fast Reasoning by xAI — 20M context; not published input / not published output per 1M tokens
- x-AI/Grok-4-Fast by xAI — 2M context; not published input / not published output per 1M tokens
- X-Ai/Grok-4-Fast-Non-Reasoning by xAI — 2M context; not published input / not published output per 1M tokens
- X-Ai/Grok-4-Fast-Reasoning by xAI — 2M context; not published input / not published output per 1M tokens
- x-AI/Grok-4.1-Fast by xAI — 2M context; not published input / not published output per 1M tokens
- x-AI/Grok-Code-Fast 1 by xAI — 256K context; not published input / not published output per 1M tokens
- Xiaomi MiMo-V2.5 — 1.05M context; free input / free output per 1M tokens
- Xiaomi MiMo-V2.5-Pro — 1.05M context; free input / free output per 1M tokens
- XiaomiMiMo/MiMo-V2-Flash by Xiaomi — 262K context; $0.1 input / $0.3 output per 1M tokens
- Xpersona Frieren 1 — 1M context; $1.5 input / $6 output per 1M tokens
- Yi Large — 32K context; $3.2 input / $3.2 output per 1M tokens
- Yi Medium 200k — 2K context; $2.5 input / $2.5 output per 1M tokens
- Z-Ai/Autoglm Phone 9b by Z.AI — 13K context; not published input / not published output per 1M tokens
- Z-AI/GLM 4.6 by Z.AI — 205K context; $0.45 input / $1.5 output per 1M tokens
- Z-Ai/GLM 4.7 by Z.AI — 2K context; not published input / not published output per 1M tokens
- Z-Ai/GLM 5 by Z.AI — 2K context; not published input / not published output per 1M tokens
- zai-org/GLM-4.5-Air by Z.AI — 131K context; $0.14 input / $0.86 output per 1M tokens
- zai-org/GLM-5 by Z.AI — 205K context; $0.95 input / $2.55 output per 1M tokens
- zai-org/GLM-5.1 by Z.AI — 205K context; $1.4 input / $4.4 output per 1M tokens
- zai-org/GLM-5V-Turbo by Z.AI — 2K context; $1.2 input / $4 output per 1M tokens
- ZDev — 1M context; free input / free output per 1M tokens
- Zero Data Retention — 1M context; not published input / not published output per 1M tokens