I

Ling-2.6-flash

LLMby inclusionAI·Model page

inclusionAI's fast Ling-2.6 LLM with a 262K-token context for text generation tasks.

Max output

Most tokens the model can return in a single response.

33Ktokens
Share:

Pricing

per 1M tokens
Input$0.01 /1M
Output$0.03 /1M
Cache read$0.002 /1M

Model Description

Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency....

Author
I
inclusionAI
Organization
inclusionai
Details
Downloads
Likes
AccessClosed Source
Context262K tokens
Input price$0.01 /1M
Output price$0.03 /1M
CreatedApr 21, 2026
Updated
View on Hugging Face
Benchmarks
Intelligence14.1
Coding25.3
Agentic2.3
Get the full context.

Sign up to read complete case studies, access detailed metrics, and unlock all use cases.

Ling-2.6-flash — AI Model Details | Applied