llm-list.com

LLM List: AI models compared

Browse all LLMs in one comprehensive list: 1,781 large language models and AI models across 214 API providers. Compare token prices, context windows, capabilities, benchmark scores, open-weight availability, and local hardware requirements. Updated every six hours.

Explore the LLM benchmark leaderboard · How prices, benchmarks and model identities are calculated

  1. Nemotron 3.5 Lightning Thinking by NVIDIA — 1M context; $0.05 input / $0.2 output per 1M tokens
  2. Nemotron 70b by NVIDIA — 16K context; $0.357 input / $0.408 output per 1M tokens
  3. Nemotron Cascade 2 by NVIDIA — 262K context; $0.15 input / $0.6 output per 1M tokens
  4. Nemotron Nano 12B v2 VL by NVIDIA — 128K context; $0.2 input / $0.6 output per 1M tokens
  5. Nemotron Nano 12B v2 VL BF16 — 128K context; $0.2 input / $0.6 output per 1M tokens
  6. Nemotron Nano 3 30B — 128K context; $0.06 input / $0.24 output per 1M tokens
  7. Nemotron Nano 9B by NVIDIA — 128K context; $0.06 input / $0.23 output per 1M tokens
  8. Nemotron Nano 9B V2 by NVIDIA — 131K context; $0.06 input / $0.23 output per 1M tokens
  9. Nemotron Super by NVIDIA — 203K context; $0.3 input / $0.75 output per 1M tokens
  10. Nemotron Super 49B by NVIDIA — 128K context; $0.15 input / $0.15 output per 1M tokens
  11. Nemotron Tenyxchat Storybreaker 70b — 16K context; $0.493 input / $0.493 output per 1M tokens
  12. Nemotron Ultra by NVIDIA — 203K context; $0.6 input / $2.4 output per 1M tokens
  13. nemotron-3-content-safety — 128K context; free input / free output per 1M tokens
  14. nemotron-3-nano-30b-a3b-bf16 — not published context; not published input / not published output per 1M tokens
  15. nemotron-3-nano-omni@eu — 3K context; $0.06 input / $0.24 output per 1M tokens
  16. nemotron-3-nano:30b by NVIDIA — 1.05M context; $0.075 input / $0.3 output per 1M tokens
  17. nemotron-3-ultra-550b-a55b-nvfp4 — not published context; not published input / not published output per 1M tokens
  18. nemotron-3-ultra-nvfp4 — 262K context; $0.6 input / $2.4 output per 1M tokens
  19. nemotron-3.5-lightning-30b-a3b-nvfp4 — not published context; not published input / not published output per 1M tokens
  20. nemotron-content-safety-reasoning-4b — 128K context; free input / free output per 1M tokens
  21. nemotron-mini-4b-instruct — 128K context; free input / free output per 1M tokens
  22. nemotron-nano-v2-12b — 128K context; $0.24 input / $0.707 output per 1M tokens
  23. nemotron-voicechat — 128K context; free input / free output per 1M tokens
  24. NeoSmith Basic — 1M context; $1.17 input / $4.37 output per 1M tokens
  25. NeoSmith Maestro — 1M context; $2.4 input / $12 output per 1M tokens
  26. NeoSmith NeoLite — 512K context; $0.6 input / $2.4 output per 1M tokens
  27. NeoSmith Pro — 1M context; $1.81 input / $8.39 output per 1M tokens
  28. Neural Daredevil 8B abliterated — 8K context; $0.44 input / $0.44 output per 1M tokens
  29. Nex AGI: Nex-N2-Mini (retires Sep 8) by NEX AGI — 262K context; $0.025 input / $0.1 output per 1M tokens
  30. Nex AGI: Nex-N2-Pro (retires Sep 8) by NEX AGI — 262K context; $0.25 input / $1 output per 1M tokens
  31. Nex N2 Mini by NEX AGI — 262K context; $0.025 input / $0.1 output per 1M tokens
  32. Nex N2 Pro by NEX AGI — 262K context; $0.25 input / $1 output per 1M tokens
  33. Nomic Embed Text v1.5 — 8K context; $0.05 input / free output per 1M tokens
  34. North Mini Code by Cohere — 256K context; free input / free output per 1M tokens
  35. North Mini Code Free — 256K context; free input / free output per 1M tokens
  36. Nous: Hermes 3 405B Instruct by Nous Research — 131K context; $1 input / $1 output per 1M tokens
  37. Nous: Hermes 3 70B Instruct by Nous Research — 131K context; $0.7 input / $0.7 output per 1M tokens
  38. Nous: Hermes 4 405B by Nous Research — 131K context; $1 input / $3 output per 1M tokens
  39. Nous: Hermes 4 70B by Nous Research — 131K context; $0.13 input / $0.4 output per 1M tokens
  40. Nova 2 Lite by Amazon — 1M context; $0.3 input / $2.5 output per 1M tokens
  41. Nova 2 Pro — 1M context; free input / free output per 1M tokens
  42. Nova Lite by Amazon — 3K context; $0.06 input / $0.24 output per 1M tokens
  43. Nova Lite 1.0 by Amazon — 3K context; $0.059 input / $0.238 output per 1M tokens
  44. Nova Micro by Amazon — 128K context; $0.03 input / $0.1 output per 1M tokens
  45. Nova Micro 1.0 by Amazon — 128K context; $0.035 input / $0.14 output per 1M tokens
  46. Nova Premier 1.0 by Amazon — 1M context; $2.5 input / $12.5 output per 1M tokens
  47. Nova Pro by Amazon — 3K context; $0.56 input / $2.13 output per 1M tokens
  48. Nova Pro 1.0 by Amazon — 3K context; $0.799 input / $3.2 output per 1M tokens
  49. nova-lite-v1 — 3K context; $0.069 input / $0.275 output per 1M tokens
  50. nova-micro-v1 — 128K context; $0.04 input / $0.159 output per 1M tokens
  51. Novelist — 262K context; $0.1 input / $0.45 output per 1M tokens
  52. nv-embed-v1 — 33K context; free input / free output per 1M tokens
  53. nv-embedcode-7b-v1 — 33K context; free input / free output per 1M tokens
  54. nvidia--llama-3.2-nv-embedqa-1b — 8K context; $0.07 input / free output per 1M tokens
  55. NVIDIA: Nemotron 3 Nano Omni by NVIDIA — 256K context; free input / free output per 1M tokens
  56. NVIDIA: Nemotron 3 Ultra by NVIDIA — 1M context; free input / free output per 1M tokens
  57. NVIDIA: Nemotron 3.5 Lightning by NVIDIA — 1M context; free input / free output per 1M tokens
  58. o1 by OpenAI — 2K context; $14 input / $54 output per 1M tokens
  59. o1-pro by OpenAI — 2K context; $140 input / $540 output per 1M tokens
  60. o3 by OpenAI — 2K context; $1 input / $4 output per 1M tokens
  61. o3 Mini High by OpenAI — 2K context; $0.99 input / $4 output per 1M tokens
  62. o3-2025-04-16 — not published context; not published input / not published output per 1M tokens
  63. o3-deep-research by OpenAI — 2K context; $9 input / $36 output per 1M tokens
  64. o3-mini by OpenAI — 2K context; $0.55 input / $2.2 output per 1M tokens
  65. o3-pro by OpenAI — 2K context; $20 input / $40 output per 1M tokens
  66. o3-search — not published context; not published input / not published output per 1M tokens
  67. o4 Mini High by OpenAI — 2K context; $1.1 input / $4.4 output per 1M tokens
  68. o4-mini by OpenAI — 2K context; $0.55 input / $2.2 output per 1M tokens
  69. o4-mini-2025-04-16 — not published context; not published input / not published output per 1M tokens
  70. o4-mini-deep-research by OpenAI — 2K context; $1.8 input / $7.2 output per 1M tokens
  71. Omega Directive 24B Unslop v2.0 — 33K context; $0.5 input / $0.5 output per 1M tokens
  72. Open Mistral Nemo — 128K context; $0.15 input / $0.15 output per 1M tokens
  73. OpenAI ChatGPT-4o — 128K context; $5 input / $20 output per 1M tokens
  74. OpenAI o1 by OpenAI — 2K context; $15 input / $60 output per 1M tokens
  75. OpenAI o1 Pro by OpenAI — 2K context; $150 input / $600 output per 1M tokens
  76. OpenAI o3 by OpenAI — 2K context; $2 input / $8 output per 1M tokens
  77. OpenAI o3 Pro — 2K context; $20 input / $80 output per 1M tokens
  78. OpenAI o3-mini by OpenAI — 2K context; $1.1 input / $4.4 output per 1M tokens
  79. OpenAI o3-pro (2025-06-10) by OpenAI — 2K context; $22 input / $88 output per 1M tokens
  80. OpenAI o4 Mini by OpenAI — 2K context; $1.1 input / $4.4 output per 1M tokens
  81. OpenAI o4-mini high by OpenAI — 2K context; $1.1 input / $4.4 output per 1M tokens
  82. OpenAI: GPT-4 Turbo Preview by OpenAI — 128K context; $10 input / $30 output per 1M tokens
  83. OpenAI: GPT-5 Image by OpenAI — 4K context; $10 input / $10 output per 1M tokens
  84. OpenAI: GPT-5.6 Sol by OpenAI — 1.05M context; $2.5 input / $15 output per 1M tokens
  85. OpenAI: o1-mini — 128K context; $1.1 input / $4.4 output per 1M tokens
  86. OpenReasoning Nemotron 32B — 131K context; $0.1 input / $0.4 output per 1M tokens
  87. OpenRouter Free Models Router — 2K context; free input / free output per 1M tokens
  88. Opus Distill — 262K context; $0.12 input / $0.38 output per 1M tokens
  89. OrcaRouter Auto — 128K context; free input / free output per 1M tokens
  90. OrcaRouter Free — 66K context; free input / free output per 1M tokens
  91. OrcaRouter Fusion — 1M context; not published input / not published output per 1M tokens
  92. OrcaRouter Fusion Flash — 2K context; not published input / not published output per 1M tokens
  93. OrcaRouter Fusion Mini — 1M context; not published input / not published output per 1M tokens
  94. Ornith 1.5 35B — 262K context; $0.1 input / $0.4 output per 1M tokens
  95. Ornith 1.5 35B A3B — 328K context; $0.1 input / $0.4 output per 1M tokens
  96. Ornith 1.5 35B Thinking — 262K context; $0.1 input / $0.4 output per 1M tokens
  97. Ornith 1.5 9B — 262K context; $0.1 input / $0.2 output per 1M tokens
  98. Ornith 1.5 9B Thinking — 262K context; $0.1 input / $0.2 output per 1M tokens
  99. Ornith-1.0-35B-FP8 — 262K context; free input / free output per 1M tokens
  100. Osmosis Structure 0.6B — 4K context; $0.1 input / $0.5 output per 1M tokens