*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
| Type | Model |
| Section | Models & platforms / text->text |
| Pricing | paid (от $0.02/mo) |
| Platform | Self-hosted |
| Systems | api, python, self-hosted |
| Hosting | cloud |
| Site language | en |
| Vendor | Inclusionai |
| Launched | 2026-07-23 |