YFarmX logoYFarmX

Inference.net: Schematron V2 Turbo

inference-net/schematron-v2-turbo · Inference NetTokenizer lineage measured
Tokenizer lineage (measured)
IBM
Entered catalogue
12 September 2026
Last tested
5 October 2026
Evidence confidence
Moderate

Why moderate: an exact fingerprint match with one sibling only, and no second signal of a different kind yet.

Evidence
  • ✓ exact tokenizer signature shared with 1 model

Identity statement

Inference.net: Schematron V2 Turbo is Inference Net's model. Its measured tokenizer sits with the IBM group, which is evidence that it builds on the IBM tokenizer. Tokenizer reuse is normal engineering, open vocabularies travel between labs, and this says nothing against Inference Net's authorship of the model itself.

What we checked

Four different kinds of evidence, and what each one showed
We checkedWhat we foundWhere it came from
How it counts tokensCounts every one of the 50 test strings exactly like 1 other modeltk_b6cc7610
How its API is set upA combination of settings no other listing in the catalogue usesapi_f359f852
How much it can read and writeReads up to 128,000 tokens at once, writes up to 8,192its own listing
Who served our requestsInferenceNetwe measured it, 5 October 2026

Shares this fingerprint

Exact signature first; template-boundary shifts of the same signature beneath

Closest measured models

Agreement across the mutually clean test strings

Click a row to open the full comparison.

The measured fingerprint

What each test string cost this model, in tokens

We sent Inference.net: Schematron V2 Turbo fifty short pieces of text and recorded what each one cost it in tokens. The bars below are those costs. Two models built on the same tokenizer produce the same bars; a model built on a different one produces a different set, which is what makes this a fingerprint.

All 50 rows

English and whitespace

en-prose14
en-long16
spaces-203
spaces-603
tabs-204
newlines-205
mixed-ws5

Digits

digits-93
digits-125
digits-3015
digits-sep7
float-long10

CJK

zh-common18
zh-long25
zh-rare23
ja-kana14
ja-kanji17
ko19

Other scripts

ru16
ar25
he24
hi35
th26
el29

Emoji

emoji-basic10
emoji-skin30
emoji-zwj-family18
emoji-zwj-x354
emoji-flags24
emoji-prof29

Rare Unicode

math32
boxdraw28
combining14
cjk-ext-b13
surrogates39
zalgo36
rtl-mix11

Code

py-code28
py-indent18
json23
html17
regex47
camel5
snake11

Repetition and encodings

rare-word-x531
repeat-tok22
base6435
hex11
url21
uuid27

Overhead subtracted: 30 prompt tokens.

Declared record

What the catalogue claims about this model

Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...

Modalities
text->text
Declared tokenizer
Other
Prompt price
$0.03 / M tokens
Completion price
$0.15 / M tokens

Supported parameters

frequency_penaltylogit_biasmax_tokensmin_ppresence_penaltyrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetop_ktop_p

defaults: {}

Catalogue entryWeights on Hugging Face

History

Every observation, kept as taken
  • 12 September 2026Enters the OpenRouter catalogue with no declared family.
  • 5 October 2026Fingerprinted in the YFarmX catalogue sweep · 50 of 50 strings measured clean.
  • 5 October 2026YFarmX assessment: consistent with the IBM family, moderate confidence.