Lead story
Models & availability
Latest
Lead story
Models & availability
Latest
Z.ai launched GLM-OCR, a compact and high-performance optical character recognition model using a self-developed CogViT and GLM-0.5B encoder-decoder architecture with CLIP pre-training for robust visual semantic understanding.
From the source
We’ve launched GLM-OCR, a compact and high-performance optical character recognition model powered by the self-developed CogViT and GLM-0.5B encoder-decoder architecture, enabling efficient cross-modal alignment through its dedicated connection layer.
docs.z.ai