Skip to content

Inference.net: Schematron V2 Turbo

Inference NetNew
inference-net/schematron-v2-turbo

About this model

Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...

Make a request

OpenAI-compatible: set the base URL and your Promptix API key, keep your existing SDK.

curl https://promptix.tn/api/v1/chat/completions \  -H "Authorization: Bearer $PROMPTIX_API_KEY" \  -H "Content-Type: application/json" \  -d '{    "model": "inference-net/schematron-v2-turbo",    "messages": [      {"role": "user", "content": "Hello! What can you do?"}    ]  }'

Pricing

In TND per 1M tokens. VAT and the platform fee are added at top-up.

Input tokens

0.116 TND

per 1M tokens

Output tokens

0.581 TND

per 1M tokens

1,000 requests with 1K input and 500 output tokens cost about 0.406 TND.

Specifications

Provider
Inference Net
Context window
128,000 tokens
Input modality
Text
Added
Sep 12, 2026
Supported parameters
frequency_penaltylogit_biasmax_tokensmin_ppresence_penaltyresponse_formatseedstop