hostjev

Pay for input, not output.

Billed per input token. Output is free: answers are read from the model's probabilities, not written as text. Figures may change before general availability.

Every account

Preview pricing

€0.042

per 1M input tokens

Output tokens are free

  • Prepaid credits, top up from €25
  • Credits never expire
  • No plans, no monthly minimum

Input tokens per month

€0.42

a month for 10M input tokens

≈ 20,000 decisions at about 500 tokens each

€4.20

a month for 100M input tokens

≈ 200,000 decisions at about 500 tokens each

€42.00

a month for 1B input tokens

≈ 2,000,000 decisions at about 500 tokens each

€420.00

a month for 10B input tokens

≈ 20,000,000 decisions at about 500 tokens each

≈ 2.4B input tokens per €100

€5 trial credits for new accounts · Illustrative

Join the waitlist

What's included

The same service for every account, from the first decision. There is no plan to upgrade to.

EU GPU regions
GPU inference runs only in EU-RO-1, EU-SE-1 and EUR-IS-2.
Packed and separate modes
Every question shares one prefilled state in packed mode, the default. In separate mode each question is answered alone.
Calibration published per model version
Temperature scaling is fitted per primitive, and figures are published with the model version they belong to. No figures published yet.
No state retained by default
Counts, latency and token totals are kept so that usage can be billed.
DPA on request
The DPA names every sub-processor, including RunPod Secure Cloud. Request the DPA
Email support
Every account can write to the team directly. hello@hostjev.com
Dedicated capacity
A worker of your own or custom rate limits, arranged directly rather than sold as a plan. Talk to us

Questions about the bill