PRICING

Free to start. Pay per token, or take a plan.

Models on our CPU servers answer free. Priced models answer on GPUs at their price per million input tokens, which an audit sets at no more than $0.04, and output tokens are free. Plans include GPU tokens every month, and a Dedicated GPU keeps one for your organization.

Plans

Free

$0 no card needed

Try every served model, and pay as you go on GPUs.

  • 500 decisions a day on our CPU servers, free
  • $5 of credit once, when you verify your email, for GPU answers (90 days)
  • 60 requests a minute
  • GPU answers from your credit, at each model's price

Open to everyone

Start free

Team

$49 a month, per organization

For an organization: its owners, admins and developers.

  • 2B GPU input tokens a month for the organization
  • Its owners, admins and developers use them on every model they call
  • Every member uses them on its private models, on GPUs
  • 1,000,000 decisions a month on CPU, 1,200 requests a minute
  • Then pay as you go: its private models from its payer's credit, members' other calls from their own

Coming soon: buy online

Contact us

Dedicated GPU

$299 a month, 24 GB

$599 a month for 48 GB

A GPU reserved for your organization, set up by our team.

  • One GPU for your organization's requests, once our team has set it up with you
  • Its private models and the public ones
  • No per-token charges from credit: billed by agreement
  • 3,000 requests a minute

Coming soon: buy online

Contact us

Plans are granted by our team until buying online opens; your plan, its tokens this month and your credit are on Settings → Billing. Larger volumes or terms of your own: [email protected].

Pay as you go on GPUs

$0.04per million input tokens, the most an audit prices a model at. Each model's price is what its GPU costs, measured by an audit on our GPUs, plus our margin.
$0for output tokens. System One models read and decide; they never generate text.
256tokens, the least a request is billed as: per request, not per question. A failed request costs nothing.

Prices appear here as models are priced. GPU answers are paid from a plan's included tokens first, then from prepaid credit; the API's usage.billed_tokens says what each request was billed as. How the API works

The free tier

Every model our CPU servers hold answers free, within 500 decisions a day. When you verify your email you get $5 of credit once, for 90 days, which pays for GPU answers in the API and the playground. Refer a friend and you both get more.

Fine-tuning

Fine-tuning in System One Studio, as it opens, costs what its GPU costs plus 50%, by the started minute. It is paid from bought or granted credit, never from the free credit, and each job holds a cap you choose, so it never costs more.

Questions

How is a GPU request billed?

By its input tokens, at the model's price per million, and at least 256 tokens a request: a request with eight questions is still one request. Output tokens are free, because System One models read and decide without writing text. A request that fails costs nothing.

What does a model cost?

Each model's price is what its GPU costs per million input tokens, measured by an audit on our GPUs, plus our margin; an audit never prices a model above $0.04. Models on our CPU servers answer free within your plan's decisions.

What happens when my plan’s tokens run out?

GPU answers carry on, paid from prepaid credit at each model’s price: yours, or for an organization’s private models its payer’s. Without credit, models on our CPU servers still answer within your decisions. The included tokens renew on the 1st of each month (UTC).

Can I buy a plan today?

Not online yet: buying online is coming soon. Until then our team grants plans by hand; write to [email protected] and say which plan you want. Credit packs can be bought on Settings → Billing, and every verified account gets $5 of free credit.

What is a Dedicated GPU?

A GPU with 24 GB or 48 GB of memory that our team sets up for one organization’s requests: once it is set up, nobody else’s requests queue in front of yours, and until then they run on our shared GPUs. Nothing is charged per token from credit; it is billed monthly, by agreement. Working alone? We set it up on an organization of your own.