Read about our latest product features, solutions, and updates.
GLM 5.2 is text-only: no image input or vision. The real multimodal model from Z.ai is GLM-5V-Turbo. What each does, how to add vision, and what's coming.