I

Ling 3.0 Flash

LLMby inclusionAI·Model page

No description available for this model yet.

Max output

Most tokens the model can return in a single response.

33Ktokens
Share:

Pricing

per 1M tokens
Input$0.021 /1M
Output$0.063 /1M
Cache read$0.004 /1M

Model Description

Ling-3.0-flash is a 124B-parameter Mixture-of-Experts (MoE) model, with approximately 5.1B parameters activated per token. The model is designed with token efficiency and production-scale agentic inference as key priorities, enabling developers...

Author
I
inclusionAI
Organization
inclusionai
Details
Downloads
Likes
AccessOpen Source
Context262K tokens
Input price$0.021 /1M
Output price$0.063 /1M
CreatedJul 23, 2026
Updated
View on Hugging Face
Benchmarks
Intelligence27.4
Coding50.6
Agentic21.1
Get the full context.

Sign up to read complete case studies, access detailed metrics, and unlock all use cases.

Ling 3.0 Flash — AI Model Details | Applied