Z.ai releases open-source GLM-OCR multimodal document model
Original titlezai-org/GLM-OCR
AISummary
Z.ai has released GLM-OCR, a 0.9B-parameter multimodal OCR model for complex document understanding, under the MIT License. The model scores 94.62 on OmniDocBench V1.5 and supports deployment through vLLM, SGLang, and Ollama, with an official SDK for document parsing.
AIWhy it matters
The page gives benchmark scores, a 0.9B parameter size, and supported serving frameworks, which help readers weigh OCR deployment options against heavier alternatives.
Source: Z.ai (GLM) · new models on Hugging Face · huggingface.coPublished · added here