inclusionAI
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
Available via OpenRouter: create a free account at openrouter.ai, generate an API key, then call this model as inclusionai/ling-3.0-flash through the OpenRouter API (or any OpenRouter-compatible app/client).
Official: openrouter.ai/inclusionai/ling-3.0-flash