Efficient Qwen model for fast chat, extraction, and high-volume workloads
Details
| Field | Value |
|---|---|
| id | qwen-flash |
| family | qwen |
| modality | text->text |
| context_length | 1000000 |
| max_output_tokens | 32768 |
| knowledge_cutoff | 2024-04 |
| release_date | 2025-07-28 |
| is_open_weights | no |
| supports_tool_call | yes |
| supports_reasoning | yes |
| supports_structured_output | no |
| supports_attachment | no |
| input_modalities | text |
| output_modalities | text |