Perceptron: Perceptron Mk1

perceptron/perceptron-mk1 · PerceptronTokenizer lineage measured
Tokenizer lineage (measured)
Qwen
Evidence confidence
High
Why high: one template-boundary shift from a group whose members all declare the same family.
Declared tokenizer tag
Other
Entered catalogue
12 May 2026
Last tested
21 August 2026
Measurement
Measured
Evidence
  • tokenizer signature one template-boundary shift from 10 models

Identity statement

Perceptron: Perceptron Mk1 is Perceptron's model. Its measured tokenizer sits with the Qwen group, which is evidence that it builds on the Qwen tokenizer. Tokenizer reuse is normal engineering, open vocabularies travel between labs, and this says nothing against Perceptron's authorship of the model itself.

Evidence stack

Signals of different kinds, weighed together
SignalResultReference
Tokenizer signatureClean measurement; no exact match yettk_c5c0a8ed
API surfaceNo other entry shares this exact contractapi_f0a7362d
Context and output32,768 context · 8,192 max outputdeclared
Reasoning contractOptionaldeclared
Serving providers observedPerceptronmeasured 21 August 2026

Shares this fingerprint

Exact signature first; template-boundary shifts of the same signature beneath

Closest measured models

Agreement across the mutually clean test strings

Click a row to open the full comparison.

The measured fingerprint

Marginal prompt-token cost of each test string, grouped by script
All 50 rows

English and whitespace

en-prose14
en-long16
spaces-203
spaces-603
tabs-203
newlines-204
mixed-ws5

Digits

digits-99
digits-1212
digits-3030
digits-sep12
float-long22

CJK

zh-common9
zh-long10
zh-rare18
ja-kana7
ja-kanji10
ko13

Other scripts

ru12
ar10
he14
hi15
th8
el13

Emoji

emoji-basic9
emoji-skin30
emoji-zwj-family18
emoji-zwj-x354
emoji-flags24
emoji-prof29

Rare Unicode

math29
boxdraw28
combining5
cjk-ext-b13
surrogates40
zalgo36
rtl-mix6

Code

py-code26
py-indent18
json25
html16
regex44
camel5
snake6

Repetition and encodings

rare-word-x531
repeat-tok22
base6436
hex17
url19
uuid36

An x marks a row excluded as corrupt (caching or provider interference detected during measurement). Overhead subtracted: 6 prompt tokens.

Declared record

What the catalogue claims about this model

Perceptron Mk1 (Mark One) is Perceptron's highest-quality vision-language model for video and embodied reasoning.** It accepts image and video inputs paired with natural language queries, and produces detailed visual understanding...

Modalities
text+image+video->text
Declared tokenizer
Other
Prompt price
$0.15 / M tokens
Completion price
$1.50 / M tokens

Supported parameters

frequency_penaltyinclude_reasoningmax_tokenspresence_penaltyreasoningstructured_outputstemperaturetop_ktop_p

defaults: {"temperature":null,"top_p":null,"top_k":null,"frequency_penalty":null,"presence_penalty":null,"repetition_penalty":null}

Catalogue entry

History

Every observation, kept as taken
  • 12 May 2026Enters the OpenRouter catalogue with no declared family.
  • 21 August 2026Fingerprinted in the YFarmX catalogue sweep · 50 of 50 strings measured clean.
  • 21 August 2026YFarmX assessment: consistent with the Qwen family, high confidence.

Cite this page as the evidence record for Perceptron: Perceptron Mk1: the URL is stable, measurements are dated, and revisions append to the history above. Method: tokenizer fingerprinting, v1, 50 strings. The raw data behind every figure is on the data page.