YFarmX logoYFarmX

Identity statement

Pareto has a clean measured fingerprint that currently sits alone: no other measured model shares its signature.

Verification note

Checked against the primary artefacts, 5 October 2026

Measured on 5 October 2026 through the endpoint Unbiased serves itself, with a fixed overhead of zero. The prompt counts it reports sit at about half of what any tokenizer on record produces, and on the strings that separate tokenizers they fall to one: a Chinese sentence of 18 characters is billed as one token, a 30-digit number as two, 63 characters of English prose as six. That is close to a word count. No vocabulary produces it, so the reading places nothing and agrees with nothing. The earlier Pareto 26.9 investigation, run while the model was listed as Union Alpha, read real token counts from three tokenizers behind one endpoint; the 26.10 preview, measured the same day as this record, reads as Llama 3 on 48 of 49 strings.

Pareto on OpenRouterPareto 26.10 preview on OpenRouter

What we checked

Four different kinds of evidence, and what each one showed
We checkedWhat we foundWhere it came from
How it counts tokensCounted cleanly, and no other model in the catalogue counts the same waytk_7d8ce940
How its API is set upSet up identically to 1 other listing, field for fieldapi_7c06b7a0
How much it can read and writeReads up to 262,144 tokens at once, writes up to 131,072its own listing
Who served our requestsUnbiasedwe measured it, 5 October 2026

Closest measured models

Agreement across the mutually clean test strings

Click a row to open the full comparison.

The measured fingerprint

What each test string cost this model, in tokens

We sent Pareto fifty short pieces of text and recorded what each one cost it in tokens. The bars below are those costs. Two models built on the same tokenizer produce the same bars; a model built on a different one produces a different set, which is what makes this a fingerprint.

All 50 rows

English and whitespace

en-prose6
en-long8
spaces-201
spaces-601
tabs-201
newlines-201
mixed-ws1

Digits

digits-91
digits-121
digits-302
digits-sep1
float-long1

CJK

zh-common1
zh-long2
zh-rare10
ja-kana1
ja-kanji3
ko5

Other scripts

ru5
ar2
he6
hi10
th3
el8

Emoji

emoji-basic2
emoji-skin6
emoji-zwj-family3
emoji-zwj-x325
emoji-flags8
emoji-prof11

Rare Unicode

math14
boxdraw19
combining2
cjk-ext-b8
surrogates25
zalgo25
rtl-mix1

Code

py-code17
py-indent7
json12
html8
regex39
camel8
snake9

Repetition and encodings

rare-word-x518
repeat-tok33
base6426
hex4
url9
uuid19

Overhead subtracted: 0 prompt tokens.

Declared record

What the catalogue claims about this model

Pareto is a multimodal composite model built for research, coding, and agentic workflows, while delivering frontier-level performance across a broad range of general-purpose tasks.

Modalities
text+image->text
Declared tokenizer
Other
Prompt price
$2.50 / M tokens
Completion price
$7.50 / M tokens

Supported parameters

max_tokensresponse_formattemperaturetool_choicetoolstop_p

defaults: {}

Catalogue entry

History

Every observation, kept as taken
  • 17 September 2026Enters the OpenRouter catalogue with no declared family.
  • 5 October 2026Fingerprinted in the YFarmX catalogue sweep · 50 of 50 strings measured clean.