Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency. It delivers performance comparable to state-of-the-art models at a similar scale while significantly reducing token usage across coding, document processing, and lightweight agent workflows.
Token volume and request traffic to this model over time.
Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency.
Ling-2.6-flash has a 262,144 token context window.
Ling 3.0 Flash VL, Ling 3.0 Flash Sante (free), Ling 3.0 Flash Fin and 1 more are other text models from inclusionAI.
Ling-2.6-flash was released on April 21, 2026.