Ling 3.0 Flash

LLM
inclusionAI
New

Ling 3.0 Flash is a cost-effective MoE model built for long-horizon tasks, tool calling, and coding. It has 124B total parameters with 5.1B activated, a 256K context window, and a hybrid reasoning mode.

Context tokens

262,144

Output tokens

32,768

Released

Jul 22, 2026

Capabilities

Tool use
Structured output
Thinking

Supported tools

No tools enabled.

Pricing

TypeCreditsUnits
InputCredits per 1k tokens
OutputCredits per 1k tokens

Variants

No variants available for this model.

Try This Model

Write a prompt and experiment with Ling 3.0 Flash in the model experiments page. You can compare it with other models side by side.