LLM List: AI models compared
Browse all LLMs in one comprehensive list: 1,781 large language models and AI models across 214 API providers. Compare token prices, context windows, capabilities, benchmark scores, open-weight availability, and local hardware requirements. Updated every six hours.
Explore the LLM benchmark leaderboard · How prices, benchmarks and model identities are calculated
- Mistral Medium 3.5 128B — 256K context; $1.5 input / $7.5 output per 1M tokens
- Mistral Medium 3.5 Thinking by Mistral — 256K context; $1.5 input / $7.5 output per 1M tokens
- Mistral Nemo by Mistral — 128K context; $0.019 input / $0.03 output per 1M tokens
- Mistral Nemo 12B Instruct by Mistral — 16K context; $0.038 input / $0.1 output per 1M tokens
- Mistral Nemo Instruct 2407 by Mistral — 128K context; $0.02 input / $0.04 output per 1M tokens
- Mistral Nemo Instruct 2407 TEE by Unsloth — 131K context; $0.025 input / $0.098 output per 1M tokens
- Mistral Nemo Starcannon 12b v1 — 16K context; $0.493 input / $0.493 output per 1M tokens
- Mistral Small (latest) by Mistral — 256K context; $0.1 input / $0.3 output per 1M tokens
- Mistral Small 3 by Mistral — 33K context; $0.05 input / $0.08 output per 1M tokens
- Mistral Small 3.1 by Mistral — 128K context; $0.1 input / $0.3 output per 1M tokens
- Mistral Small 3.1 24B by Mistral — 128K context; $0.1 input / $0.3 output per 1M tokens
- Mistral Small 3.2 by Mistral — 128K context; $0.075 input / $0.2 output per 1M tokens
- Mistral Small 3.2 24B by Mistral — 128K context; $0.075 input / $0.2 output per 1M tokens
- Mistral Small 3.2 24B Instruct (2506) by Mistral — 131K context; $0.1 input / $0.31 output per 1M tokens
- Mistral Small 4 by Mistral — 256K context; $0.143 input / $0.568 output per 1M tokens
- Mistral Small 4 119B by Mistral — 262K context; $0.4 input / $1.4 output per 1M tokens
- Mistral Small 4 119B Thinking by Mistral — 262K context; $0.4 input / $1.4 output per 1M tokens
- mistral-7b-instruct-v0.2 — 32K context; $0.159 input / $0.219 output per 1M tokens
- mistral-large-2402 — 32K context; $4.28 input / $12.95 output per 1M tokens
- mistral-large-2512 — 128K context; $1.1 input / $3.3 output per 1M tokens
- mistral-medium-2505 — not published context; not published input / not published output per 1M tokens
- mistral-medium-2508 — not published context; not published input / not published output per 1M tokens
- mistral-medium-3-5@eu — 262K context; $1.65 input / $8.25 output per 1M tokens
- mistral-nemotron by Mistral — 128K context; free input / free output per 1M tokens
- mistral-small-2503 — 128K context; $0.111 input / $0.334 output per 1M tokens
- mistral-small-2506 — not published context; not published input / not published output per 1M tokens
- mistral-small-2603 — not published context; not published input / not published output per 1M tokens
- mistral-small-3.1-24b-instruct-2503 — not published context; not published input / not published output per 1M tokens
- mistral-small-4-119b-2603 by Mistral — 128K context; free input / free output per 1M tokens
- Mistral: Mixtral 8x7B Instruct by Mistral — 33K context; free input / free output per 1M tokens
- mistralai--mistral-medium-instruct — 128K context; $0.36 input / $1.22 output per 1M tokens
- mistralai--mistral-small — 128K context; $0.07 input / $0.28 output per 1M tokens
- Mixtral 8x22B by Mistral — 66K context; $2 input / $6 output per 1M tokens
- Mixtral 8x7B by Mistral — 32K context; $0.7 input / $0.7 output per 1M tokens
- Mixtral 8x7B Instruct v0.1 — 32K context; $0.488 input / $0.758 output per 1M tokens
- MM Poly 8B — 33K context; $0.658 input / $1.11 output per 1M tokens
- MN-LooseCannon-12B-v1 — 16K context; $0.493 input / $0.493 output per 1M tokens
- Model Router — 2K context; $0.14 input / free output per 1M tokens
- Moonlight Dusk — 262K context; $0.12 input / $0.38 output per 1M tokens
- Moonshot Kimi K2 Instruct — 131K context; $0.574 input / $2.29 output per 1M tokens
- Moonshot Kimi K2 Thinking — 262K context; $0.574 input / $2.29 output per 1M tokens
- Moonshot Kimi K2.5 — 262K context; $0.574 input / $2.41 output per 1M tokens
- Moonshot Kimi K2.6 — 262K context; $0.929 input / $3.86 output per 1M tokens
- MoonshotAI Kimi Latest by Moonshot AI — 1.05M context; $2.5 input / $14 output per 1M tokens
- Moonshotai/Kimi-K2.5 by Moonshot AI — 262K context; $0.45 input / $2.25 output per 1M tokens
- moonshotai/Kimi-K2.6 by Moonshot AI — 262K context; $0.77 input / $4 output per 1M tokens
- MoonshotAI: Kimi K2 0711 by Moonshot AI — 131K context; $0.57 input / $2.3 output per 1M tokens
- MoonshotAI: Kimi K2 0905 by Moonshot AI — 262K context; $0.6 input / $2.5 output per 1M tokens
- Morph V3 Fast by Morph — 82K context; $0.8 input / $1.2 output per 1M tokens
- Morph V3 Large by Morph — 262K context; $0.9 input / $1.9 output per 1M tokens
- Motif 3 by Motif Technologies — 262K context; $0.5 input / $2 output per 1M tokens
- MS3.2 24B Magnum Diamond — 33K context; $0.493 input / $0.493 output per 1M tokens
- Multi-QA-mpnet-base-dot-v1 — 512 context; $0.009 input / free output per 1M tokens
- Muse Glimmer 30B by Meta — 131K context; $0.2 input / $0.8 output per 1M tokens
- Muse Glimmer 30B TEE — 131K context; $0.35 input / $1.5 output per 1M tokens
- Muse Spark 1.1 by Meta — 1.05M context; $1.25 input / $4.25 output per 1M tokens
- Muse Spark 1.2 by Meta — 1.05M context; $1.25 input / $4.25 output per 1M tokens
- Muse Spark 1.2 Contributor by Meta — 1.05M context; $0.1 input / $0.2 output per 1M tokens
- Muse Spark 1.2 Contributor (Data Used for Training) by Meta — 1M context; $0.1 input / $0.2 output per 1M tokens
- Muse Spark 1.2 Free — 1.05M context; free input / free output per 1M tokens
- Muse Spark 1.3 by Meta — 1.05M context; $1.25 input / $4.25 output per 1M tokens
- Muse Spark 1.3 Contributor by Meta — 1.05M context; $0.1 input / $0.2 output per 1M tokens
- Muse Spark 1.3 Free — 1.05M context; free input / free output per 1M tokens
- muse-glimmer by Meta — not published context; not published input / not published output per 1M tokens
- muse-spark-1.2-low — not published context; not published input / not published output per 1M tokens
- muse-spark-1.2-medium — not published context; not published input / not published output per 1M tokens
- muse-spark-1.2-xhigh — not published context; not published input / not published output per 1M tokens
- muse-spark-1.3-xhigh — not published context; not published input / not published output per 1M tokens
- muse-spark-1.3-xhigh-agent — not published context; not published input / not published output per 1M tokens
- Musica — 262K context; $0.12 input / $0.38 output per 1M tokens
- MythoMax 13B by Gryphe — 4K context; $0.06 input / $0.06 output per 1M tokens
- Mythomax L2 13B by Gryphe — 4K context; $0.09 input / $0.09 output per 1M tokens
- Nano Banana by Google — 33K context; $0.15 input / $1.25 output per 1M tokens
- Nano Banana 2 by Google — 66K context; $0.5 input / $2 output per 1M tokens
- Nano Banana 2 Lite by Google — 66K context; $0.25 input / $1.5 output per 1M tokens
- Nano Banana Pro by Google — 66K context; $1 input / $6 output per 1M tokens
- NanoGPT Help — 6K context; free input / free output per 1M tokens
- NemoMix 12B Unleashed — 33K context; $0.493 input / $0.493 output per 1M tokens
- Nemotron 3 Nano 30B A3B by NVIDIA — 262K context; $0.05 input / $0.2 output per 1M tokens
- Nemotron 3 Nano 30B A3B FP8 by NVIDIA — 1M context; $0.06 input / $0.25 output per 1M tokens
- Nemotron 3 Nano Omni by NVIDIA — 256K context; $0.059 input / $0.237 output per 1M tokens
- Nemotron 3 Nano Omni 30B A3B Reasoning by NVIDIA — 262K context; $0.2 input / $0.8 output per 1M tokens
- Nemotron 3 Nano Omni 30B TEE — 131K context; $0.025 input / $0.098 output per 1M tokens
- Nemotron 3 Super by NVIDIA — 262K context; $0.2 input / $0.8 output per 1M tokens
- Nemotron 3 Super 120B by NVIDIA — 1M context; $0.05 input / $0.25 output per 1M tokens
- Nemotron 3 Super 120B (Public Preview) — 1M context; $0.3 input / $0.65 output per 1M tokens
- Nemotron 3 Super 120B A12B by NVIDIA — 262K context; $0.085 input / $0.4 output per 1M tokens
- Nemotron 3 Super 120B Thinking by NVIDIA — 262K context; $0.05 input / $0.25 output per 1M tokens
- Nemotron 3 Super Free — 205K context; free input / free output per 1M tokens
- Nemotron 3 Ultra by NVIDIA — 1M context; $0.9 input / $1.7 output per 1M tokens
- Nemotron 3 Ultra 550B by NVIDIA — 1M context; $0.5 input / $2.5 output per 1M tokens
- Nemotron 3 Ultra 550B (DeepInfra) — 262K context; $0.5 input / $2.2 output per 1M tokens
- Nemotron 3 Ultra 550B A55B by NVIDIA — 1.05M context; $0.1 input / $0.1 output per 1M tokens
- Nemotron 3 Ultra 550B Thinking by NVIDIA — 1M context; $0.5 input / $2.5 output per 1M tokens
- Nemotron 3 Ultra Free — 1M context; free input / free output per 1M tokens
- Nemotron 3.5 Content Safety by NVIDIA — 131K context; $0.2 input / $0.2 output per 1M tokens
- Nemotron 3.5 Lightning by NVIDIA — 1M context; $0.05 input / $0.2 output per 1M tokens
- Nemotron 3.5 Lightning 30B by NVIDIA — 262K context; $0.05 input / $0.2 output per 1M tokens
- Nemotron 3.5 Lightning 30B A3B by NVIDIA — 262K context; $0.05 input / $0.15 output per 1M tokens
- Nemotron 3.5 Lightning Free — 262K context; free input / free output per 1M tokens