LLM List: AI models compared
Browse all LLMs in one comprehensive list: 1,781 large language models and AI models across 214 API providers. Compare token prices, context windows, capabilities, benchmark scores, open-weight availability, and local hardware requirements. Updated every six hours.
Explore the LLM benchmark leaderboard · How prices, benchmarks and model identities are calculated
- Rnj-1 Instruct — 33K context; $0.15 input / $0.15 output per 1M tokens
- roc — 262K context; $2.88 input / $11.52 output per 1M tokens
- Rocinante 12b by TheDrummer — 16K context; $0.408 input / $0.595 output per 1M tokens
- RouteLLM — 128K context; $3 input / $15 output per 1M tokens
- Saba by Mistral — 33K context; $0.199 input / $0.595 output per 1M tokens
- Safety GPT OSS 20B by OpenAI — 131K context; $0.075 input / $0.3 output per 1M tokens
- Sakana Namazu by Sakana AI — 262K context; $0.95 input / $4 output per 1M tokens
- Sao10k L3 8B Lunaris by Sao10K — 8K context; $0.05 input / $0.05 output per 1M tokens
- Sao10K Stheno 8b by Sao10K — 16K context; $0.201 input / $0.201 output per 1M tokens
- sap-abap-1 — 33K context; $0.48 input / $1.7 output per 1M tokens
- Sarvam 105B by Sarvam — 131K context; $0.04 input / $0.16 output per 1M tokens
- Sarvam 30B by Sarvam — 128K context; $0.02 input / $0.1 output per 1M tokens
- sarvam-m by Sarvam AI — 128K context; free input / free output per 1M tokens
- Seed 1.6 by ByteDance — 256K context; $0.12 input / $0.29 output per 1M tokens
- Seed 1.6 (250615) — 256K context; $0.25 input / $2 output per 1M tokens
- Seed 1.6 (250615) (ByteDance) by ByteDance — 256K context; $0.25 input / $2 output per 1M tokens
- Seed 1.6 (250915) — 256K context; $0.25 input / $2 output per 1M tokens
- Seed 1.6 (250915) (ByteDance) by ByteDance — 256K context; $0.25 input / $2 output per 1M tokens
- Seed 1.6 Flash by ByteDance — 256K context; $0.022 input / $0.223 output per 1M tokens
- Seed 1.6 Flash (250715) — 256K context; $0.07 input / $0.3 output per 1M tokens
- Seed 1.6 Flash (250715) (ByteDance) by ByteDance — 256K context; $0.07 input / $0.3 output per 1M tokens
- Seed 1.6 Vision — 256K context; $0.12 input / $1.15 output per 1M tokens
- Seed 1.8 by ByteDance — 256K context; $0.12 input / $0.29 output per 1M tokens
- Seed 1.8 (251228) — 256K context; $0.25 input / $2 output per 1M tokens
- Seed 1.8 (251228) (ByteDance) by ByteDance — 256K context; $0.25 input / $2 output per 1M tokens
- Seed 2.0 Code by ByteDance — 256K context; $0.4 input / $2.4 output per 1M tokens
- Seed 2.0 Lite by ByteDance — 256K context; $0.089 input / $0.534 output per 1M tokens
- Seed 2.0 Mini by ByteDance — 256K context; $0.03 input / $0.297 output per 1M tokens
- Seed 2.0 Pro by ByteDance — 256K context; $0.475 input / $2.37 output per 1M tokens
- Seed 2.1 Pro — 256K context; $0.707 input / $3.54 output per 1M tokens
- Seed 2.1 Turbo by ByteDance — 256K context; $0.354 input / $1.77 output per 1M tokens
- Seed Character — 256K context; $0.119 input / $0.297 output per 1M tokens
- Seed Evolving — 256K context; $0.884 input / $4.42 output per 1M tokens
- seed-2.1-pro-preview — not published context; not published input / not published output per 1M tokens
- SenseNova 6.8 Flash Lite — 262K context; free input / free output per 1M tokens
- Shadow Siren — 262K context; $0.12 input / $0.38 output per 1M tokens
- Shisa V2 Llama 3.3 70B — 128K context; $0.5 input / $0.5 output per 1M tokens
- Shisa V2.1 Llama 3.3 70B — 33K context; $0.5 input / $0.5 output per 1M tokens
- siliconflow/deepseek-r1-0528 — 164K context; $0.5 input / $2.18 output per 1M tokens
- siliconflow/deepseek-v3-0324 — 164K context; $0.25 input / $1 output per 1M tokens
- siliconflow/deepseek-v3.1-terminus — 164K context; $0.27 input / $1 output per 1M tokens
- siliconflow/deepseek-v3.2 — 164K context; $0.27 input / $0.42 output per 1M tokens
- Skyfall 36B V2 by TheDrummer — 33K context; $0.55 input / $0.8 output per 1M tokens
- SmolLM3 3B Base — 33K context; $0.15 input / $0.15 output per 1M tokens
- Solar Pro 2 by Upstage — 66K context; $0.25 input / $0.25 output per 1M tokens
- Solar Pro 3 by Upstage — 131K context; $0.25 input / $0.25 output per 1M tokens
- Solar Pro 4 by Upstage — 524K context; $0.03 input / $0.12 output per 1M tokens
- Solar Pro 4 Thinking by Upstage — 524K context; $0.03 input / $0.12 output per 1M tokens
- solar-10.7b-instruct by Upstage — 128K context; free input / free output per 1M tokens
- solar-mini by Upstage — 33K context; $0.15 input / $0.15 output per 1M tokens
- Sonar by Perplexity — 128K context; $1 input / $1 output per 1M tokens
- Sonar Deep Research by Perplexity — 128K context; $2 input / $8 output per 1M tokens
- Sonar Pro by Perplexity — 2K context; $3 input / $15 output per 1M tokens
- Sonar Pro Search by Perplexity — 2K context; $3 input / $15 output per 1M tokens
- Sonar Reasoning Pro by Perplexity — 128K context; $2 input / $8 output per 1M tokens
- SpaceXAI: Grok 4.20 by xAI — 2M context; $1.25 input / $2.5 output per 1M tokens
- SpaceXAI: Grok 4.20 Multi-Agent by xAI — 2M context; $1.25 input / $2.5 output per 1M tokens
- sparsedrive — 128K context; free input / free output per 1M tokens
- Standard Compute — 1M context; free input / free output per 1M tokens
- Stealth: Claude Opus 4.6 (20% off) — 1M context; $4 input / $20 output per 1M tokens
- Stealth: Claude Opus 4.7 (20% off) — 1M context; $4 input / $20 output per 1M tokens
- Stealth: Claude Opus 4.8 (20% off) — 1M context; $4 input / $20 output per 1M tokens
- Stealth: Claude Sonnet 4.6 (20% off) — 1M context; $2.4 input / $12 output per 1M tokens
- Stealth: Qwen3.6 Plus (50% off) — 1M context; $0.25 input / $1.5 output per 1M tokens
- Steelskull Electra R1 70b — 33K context; $0.7 input / $0.7 output per 1M tokens
- Steelskull Nevoria 70b — 33K context; $0.493 input / $0.493 output per 1M tokens
- Steelskull Nevoria R1 70b — 33K context; $0.493 input / $0.493 output per 1M tokens
- Step 1 (32K) — 33K context; $2.05 input / $9.59 output per 1M tokens
- Step 2 (16K) — 16K context; $5.21 input / $16.44 output per 1M tokens
- Step 3.5 Flash by StepFun — 256K context; $0.09 input / $0.3 output per 1M tokens
- Step 3.5 Flash 2603 by StepFun — 256K context; $0.1 input / $0.3 output per 1M tokens
- Step 3.7 Flash by StepFun — 256K context; $0.185 input / $1.11 output per 1M tokens
- Step 3.7 Flash Thinking by StepFun — 262K context; $0.2 input / $1.15 output per 1M tokens
- Step Router v1 — 256K context; not published input / not published output per 1M tokens
- Step-3 by StepFun — 66K context; $0.21 input / $0.57 output per 1M tokens
- StepAudio 2.5 ASR — not published context; not published input / not published output per 1M tokens
- StepFun 3.5 Flash by StepFun — 262K context; $0.09 input / $0.3 output per 1M tokens
- Stepfun-Ai/Gelab Zero 4b Preview by StepFun — 8K context; not published input / not published output per 1M tokens
- stepfun-ai/Step-3.5-Flash by StepFun — 262K context; $0.1 input / $0.3 output per 1M tokens
- streampetr — 128K context; free input / free output per 1M tokens
- studiovoice — 128K context; free input / free output per 1M tokens
- Synth — 1M context; not published input / not published output per 1M tokens
- Synth Code — 1M context; not published input / not published output per 1M tokens
- synthetic-video-detector — not published context; free input / free output per 1M tokens
- Tako — 2K context; not published input / not published output per 1M tokens
- Tencent HY 2.0 Instruct — 131K context; free input / free output per 1M tokens
- Tencent HY 2.0 Think — 131K context; free input / free output per 1M tokens
- Tencent Hy-MT2-Lite by Tencent — 8K context; $0.044 input / $0.177 output per 1M tokens
- Tencent Hy-MT2-Plus by Tencent — 8K context; $0.074 input / $0.295 output per 1M tokens
- Tencent Hy-MT2-Pro by Tencent — 8K context; $0.074 input / $0.295 output per 1M tokens
- Tencent Hy3 by Tencent — 262K context; $0.066 input / $0.26 output per 1M tokens
- Tencent Hy4 Preview by Tencent — 1.05M context; $0.834 input / $2.5 output per 1M tokens
- Text Embedding 005 by Google — 8K context; not published input / not published output per 1M tokens
- Text Max — 1M context; not published input / not published output per 1M tokens
- Text Multilingual Embedding 002 by Google — 8K context; not published input / not published output per 1M tokens
- Text Prime — 197K context; not published input / not published output per 1M tokens
- Text Standard — 128K context; not published input / not published output per 1M tokens
- text-embedding-3-large by OpenAI — 8K context; $0.09 input / free output per 1M tokens
- text-embedding-3-small by OpenAI — 8K context; $0.02 input / free output per 1M tokens
- text-embedding-ada-002 by OpenAI — 8K context; $0.1 input / free output per 1M tokens