AI Desk · Model Identity · Compare

MoonshotAI: Kimi K3 · Tencent: Hy-MT2-1.8B

Identity similarity: lowDivergent fingerprints · 30/49These models tokenise differently. Divergence on the high-information strings (CJK, emoji, digits) separates unrelated vocabularies within a handful of samples.

MoonshotAI: Kimi K3 recordTencent: Hy-MT2-1.8B recordOpen in the interactive bench

Tokenizer fingerprint

30 of 49 mutually clean strings agree
Test stringMoonshotAI: Kimi K3Tencent: Hy-MT2-1.8BVerdict
en-prose1414match
en-long1617differs
spaces-2033match
spaces-6033match
tabs-2044match
newlines-2044match
mixed-ws67differs
digits-933match
digits-1244match
digits-301010match
digits-sep77match
float-long99match
zh-common98differs
zh-long1010match
zh-rare1818match
ja-kana1111match
ja-kanji1412differs
ko1313match
ru1616match
ar1515match
he1616match
hi1424differs
th1817differs
el5722differs
emoji-basic710differs
emoji-skin2025differs
emoji-zwj-family1111match
emoji-zwj-x33333match
emoji-flags1424differs
emoji-prof2023differs
math2626match
boxdraw2430differs
combiningx10
cjk-ext-b1316differs
surrogates3140differs
zalgo3636match
rtl-mix88match
py-code2525match
py-indent1515match
json2020match
html1616match
regex4045differs
camel66match
snake36differs
rare-word-x52831differs
repeat-tok4141match
base643035differs
hex1112differs
url1717match
uuid2727match

Values are the marginal prompt-token cost of each string with the model's fixed overhead subtracted. An x marks a row excluded as corrupt for that model.

API surface

The declared serving contracts differ
FieldMoonshotAI: Kimi K3Tencent: Hy-MT2-1.8BVerdict
Context length10485768192× differs
Max output tokens4096× differs
Supported parametersfrequency_penalty, include_reasoning, logit_bias, logprobs, max_tokens, min_p, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_k, top_logprobs, top_pmax_completion_tokens, max_tokens, stop, temperature× differs
Defaults{"temperature":null,"top_p":0.95,"top_k":null,"frequency_penalty":null,"presence_penalty":null,"repetition_penalty":null}{}× differs
Reasoning contract{"mandatory":false,"default_enabled":true,"supported_efforts":["max","high","low"],"default_effort":"max"}× differs
Modalitiestext+image+video->texttext->text× differs

Method: tokenizer fingerprinting, v1, 50 strings. Raw data on the data page. This page is static and dated; the interactive bench compares any two of the 422 catalogue models.