GLM is a family of neural networks from the Chinese company Zhipu AI, known outside China as Z.ai. Its defining trait is a focus on development: the models edit code carefully, hold the structure of a project, and can follow a multi-step plan, calling search and tools along the way.
The top model in the family holds a million tokens in context — a large repository or a whole folder of documents at once. Z.ai publishes the weights under a free licence, so GLM can also be run on your own hardware, and the family costs noticeably less than Western models of a comparable tier.
GLM models are available in GPTunneL — worldwide, without a VPN, with one balance and billing for the tokens you actually spend.