YFarmX

AI Desk · Model Identity · Compare

DeepSeek: DeepSeek V4 Pro 0813 · WizardLM-2 8x22B

Identity similarity: lowDivergent fingerprints · 32/50These models tokenise differently. Divergence on the high-information strings (CJK, emoji, digits) separates unrelated vocabularies within a handful of samples.

DeepSeek: DeepSeek V4 Pro 0813 recordWizardLM-2 8x22B recordOpen in the interactive bench

Tokenizer fingerprint

32 of 50 mutually clean strings agree
Test stringDeepSeek: DeepSeek V4 Pro 0813WizardLM-2 8x22BVerdict
en-prose1414match
en-long1615differs
spaces-2033match
spaces-6033match
tabs-2044match
newlines-2044match
mixed-ws77match
digits-934differs
digits-1245differs
digits-301011differs
digits-sep78differs
float-long910differs
zh-common910differs
zh-long1011differs
zh-rare1819differs
ja-kana910differs
ja-kanji1112differs
ko1312differs
ru1312differs
ar1010match
he1413differs
hi1818match
th1111match
el1616match
emoji-basic1010match
emoji-skin1414match
emoji-zwj-family1111match
emoji-zwj-x33333match
emoji-flags1616match
emoji-prof1919match
math2222match
boxdraw2727match
combining1010match
cjk-ext-b1616match
surrogates3333match
zalgo3636match
rtl-mix66match
py-code2525match
py-indent1515match
json2020match
html1617differs
regex4748differs
camel76differs
snake88match
rare-word-x52626match
repeat-tok4141match
base643434match
hex1212match
url1717match
uuid2728differs

Values are the marginal prompt-token cost of each string with the model's fixed overhead subtracted. An x marks a row excluded as corrupt for that model.

API surface

The declared serving contracts differ
FieldDeepSeek: DeepSeek V4 Pro 0813WizardLM-2 8x22BVerdict
Context length104857665535× differs
Max output tokens3932168000× differs
Supported parametersfrequency_penalty, include_reasoning, logit_bias, logprobs, max_tokens, min_p, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_pfrequency_penalty, max_tokens, presence_penalty, repetition_penalty, response_format, seed, stop, temperature, top_k, top_p× differs
Defaults{"temperature":1,"top_p":1}{}× differs
Reasoning contract{"mandatory":false,"supported_efforts":["max","high","low"],"default_effort":"high"}× differs
Modalitiestext->texttext->text✓ match

Method: tokenizer fingerprinting, v1, 50 strings. This page is static and dated; the interactive bench compares any two of the 486 catalogue models.