NVIDIA: Nemotron Nano 9B V2 (free)

nvidia/nemotron-nano-9b-v2:free · NVIDIATokenizer lineage measured
Tokenizer lineage (measured)
Mistral
Evidence confidence
High
Why high: one template-boundary shift from a group whose members all declare the same family.
Declared tokenizer tag
Other
Entered catalogue
5 September 2025
Last tested
21 August 2026
Measurement
Measured
Evidence
  • ✓ tokenizer signature one template-boundary shift from 14 models

Identity statement

NVIDIA: Nemotron Nano 9B V2 (free) is NVIDIA's model. Its measured tokenizer sits with the Mistral group, which is evidence that it builds on the Mistral tokenizer. Tokenizer reuse is normal engineering, open vocabularies travel between labs, and this says nothing against NVIDIA's authorship of the model itself.

Evidence stack

Signals of different kinds, weighed together
SignalResultReference
Tokenizer signatureClean measurement; no exact match yettk_b0a747db
API surfaceNo other entry shares this exact contractapi_add7a805
Context and output128,000 context · – max outputdeclared
Reasoning contractOptionaldeclared
Serving providers observedNvidiameasured 21 August 2026

Shares this fingerprint

Exact signature first; template-boundary shifts of the same signature beneath

Closest measured models

Agreement across the mutually clean test strings

Click a row to open the full comparison.

The measured fingerprint

Marginal prompt-token cost of each test string, grouped by script
All 50 rows

English and whitespace

en-prose13
en-long17
spaces-203
spaces-603
tabs-204
newlines-207
mixed-ws7

Digits

digits-99
digits-1212
digits-3030
digits-sep12
float-long22

CJK

zh-common16
zh-long20
zh-rare20
ja-kana11
ja-kanji13
ko13

Other scripts

ru12
ar9
he10
hi14
th17
el13

Emoji

emoji-basic20
emoji-skin40
emoji-zwj-family19
emoji-zwj-x357
emoji-flags32
emoji-prof35

Rare Unicode

math33
boxdraw30
combining14
cjk-ext-b16
surrogates53
zalgo36
rtl-mix6

Code

py-code24
py-indent15
json24
html15
regex48
camel6
snake7

Repetition and encodings

rare-word-x535
repeat-tok40
base6438
hex19
url19
uuid36

An x marks a row excluded as corrupt (caching or provider interference detected during measurement). Overhead subtracted: 13 prompt tokens.

Declared record

What the catalogue claims about this model

NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and...

Modalities
text->text
Declared tokenizer
Other
Prompt price
$0 / M tokens
Completion price
$0 / M tokens

Supported parameters

include_reasoningmax_tokensreasoningresponse_formatseedstructured_outputstemperaturetool_choicetoolstop_p

defaults: {"temperature":null,"top_p":null,"frequency_penalty":null}

Catalogue entryWeights on Hugging Face

History

Every observation, kept as taken
  • 5 September 2025Enters the OpenRouter catalogue with no declared family.
  • 21 August 2026Fingerprinted in the YFarmX catalogue sweep · 50 of 50 strings measured clean.
  • 21 August 2026YFarmX assessment: consistent with the Mistral family, high confidence.

Cite this page as the evidence record for NVIDIA: Nemotron Nano 9B V2 (free): the URL is stable, measurements are dated, and revisions append to the history above. Method: tokenizer fingerprinting, v1, 50 strings. The raw data behind every figure is on the data page.