YFarmX logoYFarmX

Meta: Llama 4 Maverick

meta-llama/llama-4-maverick · MetaDeclared and measured agree
Tokenizer lineage (measured)
Llama4
Declared tokenizer tag
Llama4
Entered catalogue
5 April 2025
Last tested
5 October 2026
Evidence confidence
High

Why high: an exact fingerprint shared with 1 model that declares the same family, without a second signal of a different kind yet.

Evidence
  • ✓ exact tokenizer signature shared with 1 model

Identity statement

Meta: Llama 4 Maverick is Meta's model. Its measured tokenizer sits with the Llama4 group, which is evidence that it builds on the Llama4 tokenizer. Tokenizer reuse is normal engineering, open vocabularies travel between labs, and this says nothing against Meta's authorship of the model itself.

What we checked

Four different kinds of evidence, and what each one showed
We checkedWhat we foundWhere it came from
How it counts tokensCounts every one of the 50 test strings exactly like 1 other modeltk_59e2520c
How its API is set upA combination of settings no other listing in the catalogue usesapi_2b53aac3
How much it can read and writeReads up to 1,048,576 tokens at once, writes up to 16,384its own listing
Who served our requestsDigitalOceanwe measured it, 5 October 2026
Measured twiceThe 5 October 2026 reading agreed with the 21 August 2026 reading on every one of the 50 strings clean in both: the tokenizer behind this name held.50/50 agree

Shares this fingerprint

Exact signature first; template-boundary shifts of the same signature beneath

Closest measured models

Agreement across the mutually clean test strings

Click a row to open the full comparison.

The measured fingerprint

What each test string cost this model, in tokens

We sent Meta: Llama 4 Maverick fifty short pieces of text and recorded what each one cost it in tokens. The bars below are those costs. Two models built on the same tokenizer produce the same bars; a model built on a different one produces a different set, which is what makes this a fingerprint.

All 50 rows

English and whitespace

en-prose14
en-long17
spaces-203
spaces-603
tabs-203
newlines-204
mixed-ws6

Digits

digits-93
digits-124
digits-3010
digits-sep7
float-long9

CJK

zh-common11
zh-long12
zh-rare18
ja-kana9
ja-kanji12
ko10

Other scripts

ru10
ar11
he10
hi12
th8
el14

Emoji

emoji-basic10
emoji-skin20
emoji-zwj-family11
emoji-zwj-x333
emoji-flags16
emoji-prof19

Rare Unicode

math24
boxdraw21
combining10
cjk-ext-b16
surrogates37
zalgo31
rtl-mix6

Code

py-code25
py-indent15
json20
html16
regex42
camel6
snake6

Repetition and encodings

rare-word-x530
repeat-tok40
base6435
hex11
url17
uuid27

Overhead subtracted: 10 prompt tokens.

Declared record

What the catalogue claims about this model

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...

Modalities
text+image->text
Declared tokenizer
Llama4
Prompt price
$0.19 / M tokens
Completion price
$0.65 / M tokens

Supported parameters

frequency_penaltylogit_biaslogprobsmax_tokenspresence_penaltyrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p

defaults: {}

Catalogue entryWeights on Hugging Face

History

Every observation, kept as taken
  • 5 April 2025Enters the OpenRouter catalogue declared as Llama4.
  • 21 August 2026Fingerprinted in the YFarmX catalogue sweep · 50 of 50 strings measured clean.
  • 5 October 2026Fingerprinted again, to check the reading held · 50 of 50 strings measured clean.
  • 5 October 2026YFarmX assessment: consistent with the Llama4 family, high confidence.