All AI models · LLM benchmarks · Methodology

Ornith 1.5 35B A3B

Mixture-of-experts coding-reasoning model for agentic software tasks, tool use, and image understanding

Model facts

Context window
328K tokens
Maximum output
66K tokens
Cheapest paid input
$0.1 per 1M tokens
Cheapest paid output
$0.4 per 1M tokens
Open weights
yes
Providers
2

Capabilities and modalities

reasoning, tool calling, structured output, attachments, temperature control, open weights, text, image.

Local hardware estimate

At 4K context: Q4_K_M 20.2 GiB, Q8_0 36.9 GiB, F16 68.3 GiB working memory.

Parameters
36.0 billion
Architecture
qwen3_5moe
Model license
mit

Planning estimate adapted from the MIT-licensed whichllm estimator using metadata from Hugging Face; not a vendor minimum requirement.

Observed price history

  • 2026-08-28: $0.1 input / $0.4 output per 1M tokens

API providers

  • RunInfra — model id ornith-ai/Ornith-1.5-35B-A3B; $0.1 input / $0.4 output per 1M tokens; provider documentation
  • IteraCompute — model id iteracompute/ornith-1.5-35b-a3b; $0.3 input / $3 output per 1M tokens; provider documentation