inclusionAI: Ling-2.6-flash by inclusionai | Mume AI
Ling-2.6-flash
Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency. It delivers performance comparable to state-of-the-art models at a similar scale while significantly reducing token usage across coding, document processing, and lightweight agent workflows.
by Inclusionai|262K context|$0.08/M input tokens|$0.24/M output tokens
Endpoints
Available providers for this model, with details on pricing, context limits, and real-time health metrics.