All AI models · LLM benchmarks · Methodology

InclusionAI

ling-3.0-tiny

Ling-3.0-tiny is an efficient 7.9B parameter MoE model from inclusionAI with only 1.3B active parameters per token. Built for responsive agents, reliable instruction following and multi turn conversation, with a 256K context window, native function calling, prompt caching and switchable Thinking and Instant modes.

Model facts

Context window
262K tokens
Maximum output
33K tokens
Cheapest paid input
free per 1M tokens
Cheapest paid output
free per 1M tokens
Open weights
no
Providers
1

Capabilities and modalities

reasoning, tool calling, text.

Artificial Analysis benchmarks

Compare this model on the full LLM benchmark leaderboard.

ConfigurationAAICodingMathOutput tokens/s
default24.526.5not measured206.2

Observed price history

  • 2026-08-24: free input / free output per 1M tokens

API providers