GLM-4.5V: vision-language model from Z.ai
Z.ai published the GLM-4.5V image-text-to-text model on Hugging Face. The model is a fine-tuned variant of GLM-4.5-Air-Base, supports conversational tasks in Chinese and English, and is released under an MIT license.
PUBLISHED2025-08-10
OBSERVED2026-08-11
AGE1y
SOURCES1
- Pipeline tag: image-text-to-text
- Library: transformers
- Fine-tuned from: zai-org/GLM-4.5-Air-Base
- Tags: glm4v_moe, conversational, zh, en
- Arxiv reference: 2507.01006
- License: MIT
- Endpoints compatible
COMMUNITY
No curated reactions recorded for this event. Facts and takes are kept in separate layers — community context is added by hand, never blended into the record above.