s1-pro and s1-vision run on Surogate Rune 26B-A4B v3. Winnow-12B Q8 is the fallback for s1-pro (text); s1-vision has no fallback on the EU tier or on Peer-to-Peer without worldwide opt-in: if Rune is unavailable, s1-vision requests get a retryable 503 response. Peer-to-Peer keys with worldwide processing opt-in may be served by on-demand worldwide GPUs running the Gemma 4 12B image implementation when capacity is busy. Peer-to-Peer-tier requests may use Rune when spare capacity is available. s1-fast runs Plumb-4B on both tiers. Launch prices are per 1 million input tokens; output tokens are free. Free allowance by invitation. Pay as you go; top up from EUR 1. Private individuals and businesses worldwide can sign up and top up. Taxes are shown in checkout before you pay.
Launch pricing
These launch rates exclude applicable taxes. Top-ups are prepaid credit via Stripe with proper invoices from productivity-boost.com Betriebs UG (haftungsbeschränkt) & Co. KG. Private individuals and businesses worldwide can sign up and top up. Taxes are shown in checkout before you pay. Jev’s published list price is $0.042 per 1 million input tokens, with output tokens free. Source: TypeSafe, checked 27 September 2026. EUR rates are fixed using the ECB reference of €1 = $1.1403 on 25 September 2026, then rounded to €0.001 per million tokens.
Requests that contain an image are billed at the image rate for all their input tokens (text and image positions), because image processing takes about 3× the GPU time per token. Text-only s1-vision requests use the s1-pro rate.
The rate schedule and conversion basis are summarized in launch pricing sources. The calculator converts Jev’s published USD rate at the same ECB reference only to estimate a like-currency comparison; it is not a Jev EUR tariff. Peer-to-Peer first uses spare EU capacity, with EU requests taking priority. When busy, Peer-to-Peer may use on-demand GPU capacity rented through Lium (lium.io), operated by independent third-party providers in various countries, including outside the EU/EEA. This happens only for keys with “Allow worldwide processing” enabled. Without opt-in, Peer-to-Peer stays on EU capacity and may be queued or rejected when busy. Existing customers keep EU-only processing unless they opt in. Request and response content is not stored on any node. Rates use each request’s actual model input-token count.
Two approaches
Choose the approach that fits your data.
Same model profiles and API request shape. Implementations depend on available capacity. Choose per API key or per request with one header.
PEER-TO-PEERElastic capacity
Peer-to-Peer — capacity that scales with your workload
Built to scale worldwide. Peer-to-Peer starts on spare EU capacity. When it is busy, keys you opted in can use on-demand GPU capacity rented through Lium, operated by independent third-party providers in various countries, including outside the EU/EEA.
Today
EU requests have priority. Enable “Allow worldwide processing” for each key that may leave the EU. Without opt-in, Peer-to-Peer stays on EU capacity and may be queued or rejected when busy. Existing customers keep EU-only processing unless they opt in. Request and response content is not stored on any node.
s1-fast $0.032 · s1-pro $0.038 per 1M input tokens
EU — for the strictest data-protection requirements
For teams with the highest data-protection demands or strict EU regulatory requirements. EU requests run only on contracted operators inside the EU, and inference data never leaves the EU. EU requests have priority, even if a key allows worldwide Peer-to-Peer processing.
s1-fast $0.034 · s1-pro $0.040 per 1M input tokens
Inference data stays in the EU
Contracted EU operators only — EU requests have priority
Sign-in (Google) and payments (Stripe) are handled by the sub-processors listed on our sub-processor page; they never receive your prompts.
Workload calculator
Estimate monthly usage
Choose a model rate, then enter your estimated volume. This is a cost estimate, not an invoice.
Model tokenizers can count the same text differently. Tax and other fees are not included.
Estimated monthly cost
—
Enter your workload to calculate an estimate.
Compared with Jev list priceEnter workload
Assumes 30 days per month and the selected launch rate. Estimates only; your invoice follows actual input tokens.
ACTUAL MODEL USAGE
Input tokens
Usage is based on the selected model’s actual token count. s1-pro accepts up to three text questions; s1-fast and s1-vision accept one. Read the request limits →.
NO OUTPUT CHARGE
Output tokens
The API returns decisions rather than generated prose. TypeSafe’s published Jev list price also lists output tokens as free.