*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
Details
| Field | Value |
|---|---|
| id | inclusionai/ling-3.0-flash |
| family | ling |
| modality | text->text |
| context_length | 262144 |
| max_output_tokens | 32768 |
| release_date | 2026-07-23 |
| is_open_weights | no |
| supports_tool_call | yes |
| supports_reasoning | yes |
| supports_structured_output | no |
| supports_attachment | no |
| input_modalities | text |
| output_modalities | text |