LLM List: AI models compared
Browse all LLMs in one comprehensive list: 1,781 large language models and AI models across 214 API providers. Compare token prices, context windows, capabilities, benchmark scores, open-weight availability, and local hardware requirements. Updated every six hours.
Explore the LLM benchmark leaderboard · How prices, benchmarks and model identities are calculated
- Nemotron 3.5 Lightning Thinking by NVIDIA — 1M context; $0.05 input / $0.2 output per 1M tokens
- Nemotron 70b by NVIDIA — 16K context; $0.357 input / $0.408 output per 1M tokens
- Nemotron Cascade 2 by NVIDIA — 262K context; $0.15 input / $0.6 output per 1M tokens
- Nemotron Nano 12B v2 VL by NVIDIA — 128K context; $0.2 input / $0.6 output per 1M tokens
- Nemotron Nano 12B v2 VL BF16 — 128K context; $0.2 input / $0.6 output per 1M tokens
- Nemotron Nano 3 30B — 128K context; $0.06 input / $0.24 output per 1M tokens
- Nemotron Nano 9B by NVIDIA — 128K context; $0.06 input / $0.23 output per 1M tokens
- Nemotron Nano 9B V2 by NVIDIA — 131K context; $0.06 input / $0.23 output per 1M tokens
- Nemotron Super by NVIDIA — 203K context; $0.3 input / $0.75 output per 1M tokens
- Nemotron Super 49B by NVIDIA — 128K context; $0.15 input / $0.15 output per 1M tokens
- Nemotron Tenyxchat Storybreaker 70b — 16K context; $0.493 input / $0.493 output per 1M tokens
- Nemotron Ultra by NVIDIA — 203K context; $0.6 input / $2.4 output per 1M tokens
- nemotron-3-content-safety — 128K context; free input / free output per 1M tokens
- nemotron-3-nano-30b-a3b-bf16 — not published context; not published input / not published output per 1M tokens
- nemotron-3-nano-omni@eu — 3K context; $0.06 input / $0.24 output per 1M tokens
- nemotron-3-nano:30b by NVIDIA — 1.05M context; $0.075 input / $0.3 output per 1M tokens
- nemotron-3-ultra-550b-a55b-nvfp4 — not published context; not published input / not published output per 1M tokens
- nemotron-3-ultra-nvfp4 — 262K context; $0.6 input / $2.4 output per 1M tokens
- nemotron-3.5-lightning-30b-a3b-nvfp4 — not published context; not published input / not published output per 1M tokens
- nemotron-content-safety-reasoning-4b — 128K context; free input / free output per 1M tokens
- nemotron-mini-4b-instruct — 128K context; free input / free output per 1M tokens
- nemotron-nano-v2-12b — 128K context; $0.24 input / $0.707 output per 1M tokens
- nemotron-voicechat — 128K context; free input / free output per 1M tokens
- NeoSmith Basic — 1M context; $1.17 input / $4.37 output per 1M tokens
- NeoSmith Maestro — 1M context; $2.4 input / $12 output per 1M tokens
- NeoSmith NeoLite — 512K context; $0.6 input / $2.4 output per 1M tokens
- NeoSmith Pro — 1M context; $1.81 input / $8.39 output per 1M tokens
- Neural Daredevil 8B abliterated — 8K context; $0.44 input / $0.44 output per 1M tokens
- Nex AGI: Nex-N2-Mini (retires Sep 8) by NEX AGI — 262K context; $0.025 input / $0.1 output per 1M tokens
- Nex AGI: Nex-N2-Pro (retires Sep 8) by NEX AGI — 262K context; $0.25 input / $1 output per 1M tokens
- Nex N2 Mini by NEX AGI — 262K context; $0.025 input / $0.1 output per 1M tokens
- Nex N2 Pro by NEX AGI — 262K context; $0.25 input / $1 output per 1M tokens
- Nomic Embed Text v1.5 — 8K context; $0.05 input / free output per 1M tokens
- North Mini Code by Cohere — 256K context; free input / free output per 1M tokens
- North Mini Code Free — 256K context; free input / free output per 1M tokens
- Nous: Hermes 3 405B Instruct by Nous Research — 131K context; $1 input / $1 output per 1M tokens
- Nous: Hermes 3 70B Instruct by Nous Research — 131K context; $0.7 input / $0.7 output per 1M tokens
- Nous: Hermes 4 405B by Nous Research — 131K context; $1 input / $3 output per 1M tokens
- Nous: Hermes 4 70B by Nous Research — 131K context; $0.13 input / $0.4 output per 1M tokens
- Nova 2 Lite by Amazon — 1M context; $0.3 input / $2.5 output per 1M tokens
- Nova 2 Pro — 1M context; free input / free output per 1M tokens
- Nova Lite by Amazon — 3K context; $0.06 input / $0.24 output per 1M tokens
- Nova Lite 1.0 by Amazon — 3K context; $0.059 input / $0.238 output per 1M tokens
- Nova Micro by Amazon — 128K context; $0.03 input / $0.1 output per 1M tokens
- Nova Micro 1.0 by Amazon — 128K context; $0.035 input / $0.14 output per 1M tokens
- Nova Premier 1.0 by Amazon — 1M context; $2.5 input / $12.5 output per 1M tokens
- Nova Pro by Amazon — 3K context; $0.56 input / $2.13 output per 1M tokens
- Nova Pro 1.0 by Amazon — 3K context; $0.799 input / $3.2 output per 1M tokens
- nova-lite-v1 — 3K context; $0.069 input / $0.275 output per 1M tokens
- nova-micro-v1 — 128K context; $0.04 input / $0.159 output per 1M tokens
- Novelist — 262K context; $0.1 input / $0.45 output per 1M tokens
- nv-embed-v1 — 33K context; free input / free output per 1M tokens
- nv-embedcode-7b-v1 — 33K context; free input / free output per 1M tokens
- nvidia--llama-3.2-nv-embedqa-1b — 8K context; $0.07 input / free output per 1M tokens
- NVIDIA: Nemotron 3 Nano Omni by NVIDIA — 256K context; free input / free output per 1M tokens
- NVIDIA: Nemotron 3 Ultra by NVIDIA — 1M context; free input / free output per 1M tokens
- NVIDIA: Nemotron 3.5 Lightning by NVIDIA — 1M context; free input / free output per 1M tokens
- o1 by OpenAI — 2K context; $14 input / $54 output per 1M tokens
- o1-pro by OpenAI — 2K context; $140 input / $540 output per 1M tokens
- o3 by OpenAI — 2K context; $1 input / $4 output per 1M tokens
- o3 Mini High by OpenAI — 2K context; $0.99 input / $4 output per 1M tokens
- o3-2025-04-16 — not published context; not published input / not published output per 1M tokens
- o3-deep-research by OpenAI — 2K context; $9 input / $36 output per 1M tokens
- o3-mini by OpenAI — 2K context; $0.55 input / $2.2 output per 1M tokens
- o3-pro by OpenAI — 2K context; $20 input / $40 output per 1M tokens
- o3-search — not published context; not published input / not published output per 1M tokens
- o4 Mini High by OpenAI — 2K context; $1.1 input / $4.4 output per 1M tokens
- o4-mini by OpenAI — 2K context; $0.55 input / $2.2 output per 1M tokens
- o4-mini-2025-04-16 — not published context; not published input / not published output per 1M tokens
- o4-mini-deep-research by OpenAI — 2K context; $1.8 input / $7.2 output per 1M tokens
- Omega Directive 24B Unslop v2.0 — 33K context; $0.5 input / $0.5 output per 1M tokens
- Open Mistral Nemo — 128K context; $0.15 input / $0.15 output per 1M tokens
- OpenAI ChatGPT-4o — 128K context; $5 input / $20 output per 1M tokens
- OpenAI o1 by OpenAI — 2K context; $15 input / $60 output per 1M tokens
- OpenAI o1 Pro by OpenAI — 2K context; $150 input / $600 output per 1M tokens
- OpenAI o3 by OpenAI — 2K context; $2 input / $8 output per 1M tokens
- OpenAI o3 Pro — 2K context; $20 input / $80 output per 1M tokens
- OpenAI o3-mini by OpenAI — 2K context; $1.1 input / $4.4 output per 1M tokens
- OpenAI o3-pro (2025-06-10) by OpenAI — 2K context; $22 input / $88 output per 1M tokens
- OpenAI o4 Mini by OpenAI — 2K context; $1.1 input / $4.4 output per 1M tokens
- OpenAI o4-mini high by OpenAI — 2K context; $1.1 input / $4.4 output per 1M tokens
- OpenAI: GPT-4 Turbo Preview by OpenAI — 128K context; $10 input / $30 output per 1M tokens
- OpenAI: GPT-5 Image by OpenAI — 4K context; $10 input / $10 output per 1M tokens
- OpenAI: GPT-5.6 Sol by OpenAI — 1.05M context; $2.5 input / $15 output per 1M tokens
- OpenAI: o1-mini — 128K context; $1.1 input / $4.4 output per 1M tokens
- OpenReasoning Nemotron 32B — 131K context; $0.1 input / $0.4 output per 1M tokens
- OpenRouter Free Models Router — 2K context; free input / free output per 1M tokens
- Opus Distill — 262K context; $0.12 input / $0.38 output per 1M tokens
- OrcaRouter Auto — 128K context; free input / free output per 1M tokens
- OrcaRouter Free — 66K context; free input / free output per 1M tokens
- OrcaRouter Fusion — 1M context; not published input / not published output per 1M tokens
- OrcaRouter Fusion Flash — 2K context; not published input / not published output per 1M tokens
- OrcaRouter Fusion Mini — 1M context; not published input / not published output per 1M tokens
- Ornith 1.5 35B — 262K context; $0.1 input / $0.4 output per 1M tokens
- Ornith 1.5 35B A3B — 328K context; $0.1 input / $0.4 output per 1M tokens
- Ornith 1.5 35B Thinking — 262K context; $0.1 input / $0.4 output per 1M tokens
- Ornith 1.5 9B — 262K context; $0.1 input / $0.2 output per 1M tokens
- Ornith 1.5 9B Thinking — 262K context; $0.1 input / $0.2 output per 1M tokens
- Ornith-1.0-35B-FP8 — 262K context; free input / free output per 1M tokens
- Osmosis Structure 0.6B — 4K context; $0.1 input / $0.5 output per 1M tokens