OpenAI: GPT-4o-mini

openai/gpt-4o-mini · OpenAIDeclared and measured agree
Tokenizer lineage (measured)
GPT
Evidence confidence
High
Why high: an exact fingerprint shared with 9 models that declare the same family, without a second signal of a different kind yet.
Declared tokenizer tag
GPT
Entered catalogue
18 July 2024
Last tested
21 August 2026
Measurement
Measured
Evidence
  • exact tokenizer signature shared with 9 models

Identity statement

OpenAI: GPT-4o-mini is OpenAI's model. Its measured tokenizer sits with the GPT group, which is evidence that it builds on the GPT tokenizer. Tokenizer reuse is normal engineering, open vocabularies travel between labs, and this says nothing against OpenAI's authorship of the model itself.

Evidence stack

Signals of different kinds, weighed together
SignalResultReference
Tokenizer signatureExact match with 9 other modelstk_d1411c9d
API surfaceIdentical contract to 2 other entriesapi_12e59f90
Context and output128,000 context · 16,384 max outputdeclared
Reasoning contractNone declareddeclared
Serving providers observedOpenAImeasured 21 August 2026

Shares this fingerprint

Exact signature first; template-boundary shifts of the same signature beneath

Closest measured models

Agreement across the mutually clean test strings

Click a row to open the full comparison.

The measured fingerprint

Marginal prompt-token cost of each test string, grouped by script
All 50 rows

English and whitespace

en-prose14
en-long17
spaces-203
spaces-603
tabs-203
newlines-204
mixed-ws6

Digits

digits-93
digits-124
digits-3010
digits-sep7
float-long9

CJK

zh-common12
zh-long17
zh-rare20
ja-kana10
ja-kanji13
ko11

Other scripts

ru11
ar11
he10
hi11
th11
el14

Emoji

emoji-basic8
emoji-skin13
emoji-zwj-family11
emoji-zwj-x333
emoji-flags16
emoji-prof20

Rare Unicode

math31
boxdraw27
combining11
cjk-ext-b13
surrogates40
zalgo35
rtl-mix5

Code

py-code25
py-indent15
json20
html16
regex44
camel6
snake6

Repetition and encodings

rare-word-x531
repeat-tok22
base6434
hex11
url17
uuid27

An x marks a row excluded as corrupt (caching or provider interference detected during measurement). Overhead subtracted: 7 prompt tokens.

Declared record

What the catalogue claims about this model

GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...

Modalities
text+image+file->text
Declared tokenizer
GPT
Prompt price
$0.15 / M tokens
Completion price
$0.60 / M tokens

Supported parameters

frequency_penaltylogit_biaslogprobsmax_completion_tokensmax_tokenspredictionpresence_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_logprobstop_pweb_search_options

defaults: {}

Catalogue entry

History

Every observation, kept as taken
  • 18 July 2024Enters the OpenRouter catalogue declared as GPT.
  • 21 August 2026Fingerprinted in the YFarmX catalogue sweep · 50 of 50 strings measured clean.
  • 21 August 2026YFarmX assessment: consistent with the GPT family, high confidence.

Cite this page as the evidence record for OpenAI: GPT-4o-mini: the URL is stable, measurements are dated, and revisions append to the history above. Method: tokenizer fingerprinting, v1, 50 strings. The raw data behind every figure is on the data page.