Z.ai: GLM 5.2

z-ai/glm-5.2 · Z.aiTokenizer lineage measured
Tokenizer group (measured)
Z.ai
Evidence confidence
Moderate
Why moderate: one template-boundary shift from the group rather than an exact signature member.
Declared tokenizer tag
Other
Entered catalogue
16 June 2026
Last tested
21 August 2026
Measurement
Measured
Evidence
  • tokenizer signature one template-boundary shift from 3 models
Compare with Z.ai: GLM 5.3

Identity statement

Z.ai: GLM 5.2 is Z.ai's model, and its measured tokenizer sits in Z.ai's own signature group alongside its sibling models. Measured models from Undisclosed (stealth) carry the same vocabulary, which reads as reuse of Z.ai's tokenizer line and questions nobody's authorship; they are listed under "Shares this fingerprint" below.

Evidence stack

Signals of different kinds, weighed together
SignalResultReference
Tokenizer signatureClean measurement; no exact match yettk_1d05586d
API surfaceNo other entry shares this exact contractapi_fc7114a8
Context and output1,048,576 context · 131,072 max outputdeclared
Reasoning contractOptional · efforts xhigh, high · default highdeclared
Serving providers observedAlibaba, Crusoe, Decart, DigitalOcean, Z.AI, Baidu, BaseTen, GMICloud, Sail Researchmeasured 21 August 2026

Shares this fingerprint

Exact signature first; template-boundary shifts of the same signature beneath

Closest measured models

Agreement across the mutually clean test strings

Click a row to open the full comparison.

The measured fingerprint

Marginal prompt-token cost of each test string, grouped by script
All 50 rows

English and whitespace

en-prose14
en-long16
spaces-203
spaces-603
tabs-203
newlines-204
mixed-ws5

Digits

digits-95
digits-127
digits-3010
digits-sep9
float-long14

CJK

zh-common9
zh-long10
zh-rare18
ja-kana10
ja-kanji14
ko17

Other scripts

ru13
ar13
he24
hi27
th24
el14

Emoji

emoji-basic5
emoji-skin11
emoji-zwj-family7
emoji-zwj-x321
emoji-flags24
emoji-prof15

Rare Unicode

math27
boxdraw28
combining10
cjk-ext-b13
surrogates40
zalgo22
rtl-mix8

Code

py-code25
py-indent15
json21
html16
regex44
camel5
snake6

Repetition and encodings

rare-word-x531
repeat-tok22
base6435
hex14
url17
uuid28

An x marks a row excluded as corrupt (caching or provider interference detected during measurement). Overhead subtracted: 12 prompt tokens.

Declared record

What the catalogue claims about this model

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...

Modalities
text->text
Declared tokenizer
Other
Prompt price
$0.97 / M tokens
Completion price
$3.04 / M tokens

Supported parameters

frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_pparallel_tool_callspresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p

defaults: {"temperature":1,"top_p":0.95,"top_k":null,"frequency_penalty":null,"presence_penalty":null,"repetition_penalty":null}

Catalogue entryWeights on Hugging Face

History

Every observation, kept as taken
  • 16 June 2026Enters the OpenRouter catalogue with no declared family.
  • 21 August 2026Fingerprinted in the YFarmX catalogue sweep · 50 of 50 strings measured clean.
  • 21 August 2026YFarmX assessment: consistent with the Z.ai family, moderate confidence.

Cite this page as the evidence record for Z.ai: GLM 5.2: the URL is stable, measurements are dated, and revisions append to the history above. Method: tokenizer fingerprinting, v1, 50 strings. The raw data behind every figure is on the data page.