YFarmX logoYFarmX

Meta: Llama 3.1 8B Instruct

meta-llama/llama-3.1-8b-instruct · MetaDeclared and measured agree
Tokenizer lineage (measured)
Llama3
Declared tokenizer tag
Llama3
Entered catalogue
23 July 2024
Last tested
21 August 2026
Evidence confidence
High

Why high: an exact fingerprint shared with 8 models that declare the same family, without a second signal of a different kind yet.

Evidence
  • ✓ exact tokenizer signature shared with 8 models

Identity statement

Meta: Llama 3.1 8B Instruct is Meta's model. Its measured tokenizer sits with the Llama3 group, which is evidence that it builds on the Llama3 tokenizer. Tokenizer reuse is normal engineering, open vocabularies travel between labs, and this says nothing against Meta's authorship of the model itself.

What we checked

Four different kinds of evidence, and what each one showed
We checkedWhat we foundWhere it came from
How it counts tokensCounts every one of the 50 test strings exactly like 8 other modelstk_6588ed5b
How its API is set upA combination of settings no other listing in the catalogue usesapi_457c69d0
How much it can read and writeReads up to 131,072 tokens at onceOpenRouter's listing shows 117,964 output tokens, which is 90% of the context window, the figure it carries where the provider declares no maximum.its own listing
Who served our requestsDeepInfrawe measured it, 21 August 2026

Shares this fingerprint

Exact signature first; template-boundary shifts of the same signature beneath

Closest measured models

Agreement across the mutually clean test strings

Click a row to open the full comparison.

The measured fingerprint

What each test string cost this model, in tokens

We sent Meta: Llama 3.1 8B Instruct fifty short pieces of text and recorded what each one cost it in tokens. The bars below are those costs. Two models built on the same tokenizer produce the same bars; a model built on a different one produces a different set, which is what makes this a fingerprint.

All 50 rows

English and whitespace

en-prose14
en-long16
spaces-203
spaces-603
tabs-203
newlines-204
mixed-ws5

Digits

digits-93
digits-124
digits-3010
digits-sep7
float-long9

CJK

zh-common12
zh-long18
zh-rare18
ja-kana11
ja-kanji12
ko14

Other scripts

ru14
ar12
he24
hi15
th14
el13

Emoji

emoji-basic10
emoji-skin30
emoji-zwj-family15
emoji-zwj-x345
emoji-flags24
emoji-prof26

Rare Unicode

math31
boxdraw23
combining11
cjk-ext-b13
surrogates40
zalgo35
rtl-mix8

Code

py-code25
py-indent15
json20
html16
regex44
camel5
snake6

Repetition and encodings

rare-word-x530
repeat-tok21
base6435
hex11
url17
uuid27

Overhead subtracted: 10 prompt tokens.

Declared record

What the catalogue claims about this model

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 8B instruct-tuned version is fast and efficient. It has demonstrated strong performance compared to...

Modalities
text->text
Declared tokenizer
Llama3
Prompt price
$0.05 / M tokens
Completion price
$0.08 / M tokens

Supported parameters

frequency_penaltylogit_biaslogprobsmax_tokensmin_ppresence_penaltyrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p

defaults: {}

Catalogue entryWeights on Hugging Face

History

Every observation, kept as taken
  • 23 July 2024Enters the OpenRouter catalogue declared as Llama3.
  • 21 August 2026Fingerprinted in the YFarmX catalogue sweep · 50 of 50 strings measured clean.
  • 21 August 2026YFarmX assessment: consistent with the Llama3 family, high confidence.