Fast GLM vision model for screenshots, documents, and multimodal agent tasks

Details

FieldValue
idglm-5v-turbo
familyglm
modalitytext+image+video+pdf->text
context_length200000
max_output_tokens131072
release_date2026-04-01
is_open_weightsno
supports_tool_callyes
supports_reasoningyes
supports_structured_outputno
supports_attachmentyes
input_modalitiestext, image, video, pdf
output_modalitiestext