OpenAI: gpt-oss-120b

openai/gpt-oss-120b · OpenAIDeclared and measured agree
Tokenizer lineage (measured)
GPT
Evidence confidence
High
Why high: an exact fingerprint shared with 9 models that declare the same family, without a second signal of a different kind yet.
Declared tokenizer tag
GPT
Entered catalogue
5 August 2025
Last tested
21 August 2026
Measurement
Measured
Evidence
  • exact tokenizer signature shared with 9 models

Identity statement

OpenAI: gpt-oss-120b is OpenAI's model. Its measured tokenizer sits with the GPT group, which is evidence that it builds on the GPT tokenizer. Tokenizer reuse is normal engineering, open vocabularies travel between labs, and this says nothing against OpenAI's authorship of the model itself.

Evidence stack

Signals of different kinds, weighed together
SignalResultReference
Tokenizer signatureExact match with 9 other modelstk_d1411c9d
API surfaceNo other entry shares this exact contractapi_a133a72d
Context and output131,072 context · 131,072 max outputdeclared
Reasoning contractMandatory · efforts high, medium, low · default mediumdeclared
Serving providers observedGroqmeasured 21 August 2026

Shares this fingerprint

Exact signature first; template-boundary shifts of the same signature beneath

Closest measured models

Agreement across the mutually clean test strings

Click a row to open the full comparison.

The measured fingerprint

Marginal prompt-token cost of each test string, grouped by script
All 50 rows

English and whitespace

en-prose14
en-long17
spaces-203
spaces-603
tabs-203
newlines-204
mixed-ws6

Digits

digits-93
digits-124
digits-3010
digits-sep7
float-long9

CJK

zh-common12
zh-long17
zh-rare20
ja-kana10
ja-kanji13
ko11

Other scripts

ru11
ar11
he10
hi11
th11
el14

Emoji

emoji-basic8
emoji-skin13
emoji-zwj-family11
emoji-zwj-x333
emoji-flags16
emoji-prof20

Rare Unicode

math31
boxdraw27
combining11
cjk-ext-b13
surrogates40
zalgo35
rtl-mix5

Code

py-code25
py-indent15
json20
html16
regex44
camel6
snake6

Repetition and encodings

rare-word-x531
repeat-tok22
base6434
hex11
url17
uuid27

An x marks a row excluded as corrupt (caching or provider interference detected during measurement). Overhead subtracted: 71 prompt tokens.

Declared record

What the catalogue claims about this model

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...

Modalities
text->text
Declared tokenizer
GPT
Prompt price
$0.03 / M tokens
Completion price
$0.17 / M tokens

Supported parameters

frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_atop_ktop_logprobstop_p

defaults: {"temperature":null,"top_p":null,"frequency_penalty":null}

Catalogue entryWeights on Hugging Face

History

Every observation, kept as taken
  • 5 August 2025Enters the OpenRouter catalogue declared as GPT.
  • 21 August 2026Fingerprinted in the YFarmX catalogue sweep · 50 of 50 strings measured clean.
  • 21 August 2026YFarmX assessment: consistent with the GPT family, high confidence.

Cite this page as the evidence record for OpenAI: gpt-oss-120b: the URL is stable, measurements are dated, and revisions append to the history above. Method: tokenizer fingerprinting, v1, 50 strings. The raw data behind every figure is on the data page.