LLM List: AI models compared
Browse all LLMs in one comprehensive list: 1,781 large language models and AI models across 214 API providers. Compare token prices, context windows, capabilities, benchmark scores, open-weight availability, and local hardware requirements. Updated every six hours.
Explore the LLM benchmark leaderboard · How prices, benchmarks and model identities are calculated
- Laguna XS.2 — 33K context; free input / free output per 1M tokens
- leanstral-1-5 — 262K context; free input / free output per 1M tokens
- leanstral-1-5@eu — 262K context; free input / free output per 1M tokens
- LFM2 24B A2B by Liquid AI — 33K context; $0.03 input / $0.12 output per 1M tokens
- LFM2.5 2.6B by Liquid AI — 128K context; $0.1 input / $0.2 output per 1M tokens
- Ling 2.6 Flash Free — 262K context; free input / free output per 1M tokens
- Ling 3.0 Flash by InclusionAI — 262K context; $0.021 input / $0.063 output per 1M tokens
- Ling 3.0 Flash Fin by InclusionAI — 262K context; $0.06 input / $0.18 output per 1M tokens
- Ling 3.0 Flash Fin (Free) — 262K context; free input / free output per 1M tokens
- Ling 3.0 Flash Thinking by InclusionAI — 262K context; $0.075 input / $0.22 output per 1M tokens
- Ling-1T by InclusionAI — 128K context; $0.56 input / $2.24 output per 1M tokens
- ling-2.5-1t — not published context; not published input / not published output per 1M tokens
- Ling-2.6-1T by InclusionAI — 262K context; $0.3 input / $2.5 output per 1M tokens
- Ling-2.6-flash by InclusionAI — 262K context; $0.1 input / $0.3 output per 1M tokens
- Ling-3.0-flash Free — 262K context; free input / free output per 1M tokens
- ling-3.0-tiny by InclusionAI — 262K context; free input / free output per 1M tokens
- Ling-3.0-tiny Free — 262K context; free input / free output per 1M tokens
- Ling-flash-2.0 by InclusionAI — 131K context; $0.14 input / $0.57 output per 1M tokens
- LiquidAI: LFM2.5-2.6B (free) by Liquid AI — 66K context; free input / free output per 1M tokens
- Llama 3 70B abliterated — 8K context; $0.7 input / $0.7 output per 1M tokens
- Llama 3 70B Instruct by Meta — 8K context; $0.51 input / $0.74 output per 1M tokens
- Llama 3 8B Instruct by Meta — 8K context; $0.04 input / $0.04 output per 1M tokens
- Llama 3 8B Lunaris by Sao10K — 8K context; $0.04 input / $0.05 output per 1M tokens
- Llama 3.05 Storybreaker Ministral 70b — 16K context; $0.493 input / $0.493 output per 1M tokens
- Llama 3.1 405B Instruct by Meta — 128K context; $1.95 input / $1.95 output per 1M tokens
- Llama 3.1 405B Instruct Turbo by Meta — 128K context; $3.5 input / $3.5 output per 1M tokens
- Llama 3.1 70B by Meta — 128K context; $0.4 input / $0.4 output per 1M tokens
- Llama 3.1 70B Celeste v0.1 — 33K context; $0.493 input / $0.493 output per 1M tokens
- Llama 3.1 70B Dracarys 2 — 33K context; $0.493 input / $0.493 output per 1M tokens
- Llama 3.1 70B Euryale by Sao10K — 2K context; $0.306 input / $0.357 output per 1M tokens
- Llama 3.1 70B Hanami by Sao10K — 33K context; $0.493 input / $0.493 output per 1M tokens
- Llama 3.1 70B Instruct — 128K context; $0.72 input / $0.72 output per 1M tokens
- Llama 3.1 8B by Meta — 131K context; $0.025 input / $0.025 output per 1M tokens
- Llama 3.1 8B (decentralized) — 128K context; $0.02 input / $0.03 output per 1M tokens
- Llama 3.1 8b (uncensored) by Aion Labs — 33K context; $0.8 input / $1.6 output per 1M tokens
- Llama 3.1 8B Instruct fp8 by Meta — 32K context; $0.152 input / $0.287 output per 1M tokens
- Llama 3.1 Euryale 70B v2.2 by Sao10K — 131K context; $0.85 input / $0.85 output per 1M tokens
- Llama 3.1 Nemotron 70B Instruct by NVIDIA — 131K context; $0.6 input / $0.6 output per 1M tokens
- Llama 3.1 Nemotron Nano 8B v1 — 131K context; free input / free output per 1M tokens
- Llama 3.1 Nemotron Nano VL 8B v1 — 33K context; free input / free output per 1M tokens
- Llama 3.1 Nemotron Ultra 253B — 128K context; free input / free output per 1M tokens
- Llama 3.2 11B Instruct by Meta — 128K context; $0.07 input / $0.33 output per 1M tokens
- Llama 3.2 11B Vision Instruct by Meta — 128K context; $0.055 input / $0.055 output per 1M tokens
- Llama 3.2 1B by Meta — 6K context; $0.01 input / $0.01 output per 1M tokens
- Llama 3.2 3B by Meta — 131K context; $0.02 input / $0.02 output per 1M tokens
- Llama 3.2 3B Instruct — 33K context; $0.03 input / $0.05 output per 1M tokens
- Llama 3.2 90B Vision Instruct by Meta — 128K context; $0.35 input / $0.4 output per 1M tokens
- Llama 3.3 70B by Meta — 128K context; $0.05 input / $0.23 output per 1M tokens
- Llama 3.3 70B Cu Mai — 33K context; $0.493 input / $0.493 output per 1M tokens
- Llama 3.3 70B Euryale by Sao10K — 131K context; $0.493 input / $0.493 output per 1M tokens
- Llama 3.3 70B Instruct — 131K context; $0.135 input / $0.4 output per 1M tokens
- Llama 3.3 70B Instruct abliterated by Huihui AI — 33K context; $0.7 input / $0.7 output per 1M tokens
- Llama 3.3 70B Instruct fp8 Fast by Meta — 24K context; $0.293 input / $2.25 output per 1M tokens
- Llama 3.3 70B Turbo by Meta — 131K context; $0.1 input / $0.32 output per 1M tokens
- Llama 3.3 70B Versatile — 131K context; $0.59 input / $0.79 output per 1M tokens
- Llama 3.3 70B Wayfarer — 33K context; $0.7 input / $0.7 output per 1M tokens
- Llama 3.3 Nemotron Super 49B v1 by NVIDIA — 131K context; free input / free output per 1M tokens
- Llama 3.3 Nemotron Super 49B v1.5 by NVIDIA — 131K context; $0.4 input / $0.4 output per 1M tokens
- Llama 4 Maverick by Meta — 1.05M context; $0.15 input / $0.6 output per 1M tokens
- Llama 4 Maverick 17B 128E Instruct by Meta — 524K context; $0.15 input / $0.6 output per 1M tokens
- Llama 4 Maverick 17B 128E Instruct FP8 by Meta — 1M context; $0.25 input / $1 output per 1M tokens
- Llama 4 Maverick 17B FP8 by Meta — 1.05M context; $0.2 input / $0.8 output per 1M tokens
- Llama 4 Maverick 17B Instruct by Meta — 1.05M context; $0.14 input / $0.59 output per 1M tokens
- Llama 4 Scout by Meta — 1.31M context; $0.1 input / $0.3 output per 1M tokens
- Llama 4 Scout 17B by Meta — 3.5M context; $0.1 input / $0.3 output per 1M tokens
- Llama 4 Scout 17B 16E Instruct by Meta — 128K context; $0.2 input / $0.78 output per 1M tokens
- Llama 4 Scout 17B Instruct — 3.5M context; $0.18 input / $0.59 output per 1M tokens
- Llama Guard 4 12B by Meta — 164K context; $0.18 input / $0.18 output per 1M tokens
- Llama Prompt Guard 2 22M by Meta — 512 context; $0.01 input / $0.01 output per 1M tokens
- Llama-3.1-8B-CS — 128K context; $0.1 input / $0.1 output per 1M tokens
- llama-3.1-nemotron-safety-guard-8b-v3 — 128K context; free input / free output per 1M tokens
- llama-3.1-nemotron-ultra-253b-v1 by NVIDIA — 128K context; $0.598 input / $1.79 output per 1M tokens
- llama-3.3-70b-cs — not published context; not published input / not published output per 1M tokens
- Llama-3.3-8B-Instruct — 128K context; free input / free output per 1M tokens
- llama-3_2-nemoretriever-300m-embed-v1 — 33K context; free input / free output per 1M tokens
- Llama-4-Scout-17B-16E-Instruct-FP8 by Meta — 128K context; free input / free output per 1M tokens
- Llama-Guard-3-8B by Meta — 131K context; $0.055 input / $0.055 output per 1M tokens
- llama-nemotron-embed-vl-1b-v2 — 33K context; free input / free output per 1M tokens
- llama-nemotron-rerank-vl-1b-v2 — 128K context; free input / free output per 1M tokens
- Llama-xLAM-2 70B fc-r — 128K context; $2.5 input / $2.5 output per 1M tokens
- LongCat 2.0 Thinking — 1.05M context; $0.75 input / $3 output per 1M tokens
- LongCat-2.0 by Meituan — 1.05M context; $0.3 input / $1.2 output per 1M tokens
- LongCat-2.0 Free — 1M context; free input / free output per 1M tokens
- Longcat-Flash-Chat by Meituan — 131K context; not published input / not published output per 1M tokens
- LucidNova RF1 100B — 12K context; $2 input / $5 output per 1M tokens
- LucidQuery Nexus Coder — 25K context; $2 input / $5 output per 1M tokens
- Lumimaid v0.2 — 16K context; $1 input / $1.5 output per 1M tokens
- Luminous Mirror — 262K context; $0.12 input / $0.38 output per 1M tokens
- Lynkr Auto (complexity routing) — 128K context; free input / free output per 1M tokens
- Lyria 3 Clip Preview by Google — 1.05M context; free input / free output per 1M tokens
- Lyria 3 Pro Preview by Google — 1.05M context; free input / free output per 1M tokens
- Mag Mell R1 — 16K context; $0.493 input / $0.493 output per 1M tokens
- Magibu 11B v8 — 8K context; $0.1 input / $0.5 output per 1M tokens
- Magistral Medium (latest) by Mistral — 128K context; $2 input / $5 output per 1M tokens
- Magistral Small by Mistral — 128K context; $0.5 input / $1.5 output per 1M tokens
- Magistral Small 1.2 by Mistral — 128K context; $0.5 input / $1.5 output per 1M tokens
- Magistral Small 2506 by Mistral — 128K context; $0.5 input / $1.5 output per 1M tokens
- Magnum V2 72B by Anthracite — 16K context; $2.01 input / $2.99 output per 1M tokens
- Magnum v4 72B by Anthracite — 33K context; $2.01 input / $2.99 output per 1M tokens
- MAI-Code-1-Flash — 256K context; $0.75 input / $4.5 output per 1M tokens