InclusionAIFastFastActive

inclusionAI: Ling 3.0 Flash

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

Specification
Model ID
inclusionai/ling-3.0-flash
Modality
Text >text
Context
262.1K tokens
Input
$0.021/1M
Output
$0.063/1M
Updated
Sep 20, 2026
Scores
Capabilities
  • Supports tool calling
  • No vision support listed
  • Supports long-context workflows
  • Supports deeper reasoning tasks