Text model

GLM by Z.ai — strong at code for a fraction of the price

GLM models by Z.ai in GPTunneL: context up to 1M tokens, tuned for code and tool use. Pay per token, no subscription and no VPN.

What the GLM family is good at

The traits Z.ai keeps from generation to generation — you can count on them whichever model you pick.

Tuned for code

The main reason people pick this family: Z.ai tunes its models for development first — they edit files carefully and hold the structure of a project.

Works with tools

The models can search the web and call tools, which is why they get picked for multi-step agentic work rather than plain conversation.

Below market price

At comparable quality on code, the family costs noticeably less than Western models of the same tier.

Open weights

Z.ai publishes its model weights under a free licence, so they can be run on your own hardware. In GPTunneL the same models are available without any hardware of your own.

Which model to pick

Three tiers in the family — for different tasks and budgets.

GLM 5 TurboThroughput and speedGLM 5.2The default pickGLM 5.1The proven generation
Best forSimple requests, classification, auto-repliesCode, agentic runs, large projectsEveryday tasks where the result is already dialled in
Context190K tokens1M tokens190K tokens
Tool useYesYesYes
Price per tokenLowLowLow

Specs come from the platform catalogue. Exact per-model prices are on the Pricing

Every GLM model

Release dates follow Z.ai's official announcements.

June 2026

GLM 5.2

The current generation: context jumped straight to a million tokens, and the biggest gains landed on code — which is what the family gets picked for.

March — April 2026

GLM 5.1

An interim update to the fifth generation — it follows instructions more precisely and is steadier on multi-step plans.

February 2026

GLM 5GLM 5 Turbo

The generation where the family became a serious choice for development. A lighter Turbo variant sits alongside it — the same generation, noticeably faster.

2025

GLM 4

The previous generation, the one the company took international with, renaming itself Z.ai along the way.

How billing works

There is no subscription: you top up one balance and spend it on any model on the platform.

Pay per token

You are charged for exactly what the request and the answer used. Not using it costs nothing: no limits, no monthly fee.

One balance for every model

GLM, ChatGPT, Claude, image and video generation — all out of the same wallet. No separate subscription per service.

Top up the way you prefer

International cards, Apple Pay, Google Pay or crypto. The minimum top-up is $5.

How to start with GLM

Sign in to GPTunneL
One account for every model on the platform.
Read this module and suggest how to move the database work into its own layer
orders.pyGLM 5.2
Module review

Database queries are spread across three handlers, so the same selection has to be edited in three places.

Move them into a repository with methods per entity — the handlers then stop knowing about SQL at all.

The transaction is easier to raise one level up into the service: right now it opens inside the loop.

Sign in whichever way suits you.

The model switches right inside the input — you can change it mid-conversation.

The answer arrives in the chat and the context is kept until the conversation ends.

Try GLM in GPTunneL

Signing up takes a minute, and $5 on the balance is enough to see whether the model fits your task.

Frequently asked questions

No. Requests go through GPTunneL's infrastructure, so the models open like any ordinary website — no VPN and no proxy.

GLM is a family of neural networks from the Chinese company Zhipu AI, known outside China as Z.ai. Its defining trait is a focus on development: the models edit code carefully, hold the structure of a project, and can follow a multi-step plan, calling search and tools along the way.

The top model in the family holds a million tokens in context — a large repository or a whole folder of documents at once. Z.ai publishes the weights under a free licence, so GLM can also be run on your own hardware, and the family costs noticeably less than Western models of a comparable tier.

GLM models are available in GPTunneL — worldwide, without a VPN, with one balance and billing for the tokens you actually spend.