GLM-4.7-Flash: Z.ai's MoE model on Hugging Face
Z.ai published GLM-4.7-Flash on Hugging Face, a text-generation model using a MoE lite architecture (glm4_moe_lite) under the MIT license. It is deployed via transformers and supports English and Chinese conversation. The repository is tagged as Azure-deployable and endpoint-compatible.
PUBLISHED2026-01-19
OBSERVED2026-08-11
AGE7mo
SOURCES1
- Pipeline: text-generation (conversational)
- Library: transformers
- Architecture: glm4_moe_lite (Mixture of Experts)
- Languages: en, zh
- License: MIT
- Deployment: endpoints_compatible, deploy:azure
- Referenced paper: arxiv:2508.06471
- Repository created 2026-01-19, last modified 2026-01-29
COMMUNITY
No curated reactions recorded for this event. Facts and takes are kept in separate layers — community context is added by hand, never blended into the record above.