Ling-3.0-flash
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
Anyone in the Project can @-mention Ling-3.0-flash with the team's shared context - pooled credits, one chat, one memory.
Starter is free forever - 1 Project, 100 credits/month, 1 MCP. No card.
Specifications
- Provider
- inclusionai
- Category
- llm
- Context length
- 262,144 tokens
- Max output
- 32,768 tokens
- Modalities
- text
- License
- proprietary
- Released
- 2026-07-23
Pricing
- Input
- $0.02/Mtok
- Output
- $0.06/Mtok
- Model ID
inclusionai/ling-3.0-flash
Per-token prices show what the model costs upstream. On Switchy your team draws from one shared org credit pool - one plan, one balance for everyone.
Team cost calculator
5 seats · 80 msgs/day
Switchy meters this against your org's shared credit pool - one plan, one balance for everyone.
Providers
Performance
Benchmarks
Works well with
Top MCPs
Compatibility data comes from first-party telemetry; once we have enough co-usage signal, top MCPs for this model will appear here.
How Switchy teams use it
Starter prompts
Compare with
More language models
- Ling 3.0 Flash Fininclusionai
- Ling 3.0 Flash Fin (free)inclusionai
- LiquidAI: LFM2.5-2.6B (free)liquid
- Magnum v4 72Banthracite-org
- Mancer: Weaver (alpha)mancer
- Meituan: LongCat 2.0meituan
- Meta: Llama 3.1 70B Instructmeta-llama
- Meta: Llama 3.1 8B Instructmeta-llama
- Meta: Llama 3.2 1B Instructmeta-llama
- Meta: Llama 3.2 3B Instructmeta-llama
- Meta: Llama 3.3 70B Instructmeta-llama
- Meta: Llama 4 Maverickmeta-llama