Ling-3.0-flash is a 124B-parameter Mixture-of-Experts (MoE) model, with approximately 5.1B parameters activated per token. The model is designed with token efficiency and production-scale agentic inference as key priorities, enabling developers...
by Inclusionai|262K context|$0.02/M input tokens|$0.06/M output tokens
Endpoints
Available providers for this model, with details on pricing, context limits, and real-time health metrics.