Ling 2.6 Flash

LLM
inclusionAI

Ling 2.6 Flash is a 104B-parameter (7.4B active) instruct model designed for agent workflows requiring fast responses, strong task execution, and high token efficiency, with a focus on coding, document processing, and lightweight reasoning.

Context tokens

262,144

Output tokens

8,192

Released

Apr 24, 2026

Schema

This model is optimized to produce concise outputs by design, so responses may be notably shorter than expected. It is particularly tuned for agentic and multi-step tool-use workflows.

Schema documentation

Capabilities

Tool use
Structured output

Supported tools

No tools enabled.

Pricing

TypeCreditsUnits
Input1.33Credits per 1k tokens
Output3.99Credits per 1k tokens
Cache hit0.27Credits per 1k tokens

Variants

No variants available for this model.

Try This Model

Write a prompt and experiment with Ling 2.6 Flash in the model experiments page. You can compare it with other models side by side.