Pricing

A GPU's hourly price is shown before you launch it, and model prices are on the Models page.
GPU capacity and inference, with one BRL ledger

GPUs: per hour

The hourly price covers the GPU instance in the configuration listed. You are charged by the minute while it runs.

Models: per million tokens

Input and output tokens have separate prices, per million, and each request is charged for its own usage.

What the price includes

  • Taxes are included in the price shown.
  • For services priced in US dollars upstream, the price in reais already includes the exchange rate, the IOF tax and the cost of paying our providers in dollars.
  • Fees for paying by Pix, card, or boleto are not part of the price, and nothing is deducted from your balance: it is credited with the full amount you pay.

When you are charged

GPUs: per hour

GPU compute is billed by the minute while running. Stopped or suspended instances accrue no new compute usage; any unpaid recorded usage or separately accepted plan commitment remains due.

Models: per million tokens

Inference: when a request is served.

Prepaid

Prepaid: usage is debited from your balance as it happens.

Postpaid

Postpaid: usage adds up over the month and is invoiced when the month closes.

Plans

Every plan published now, by billing mode. You choose a plan in the portal, then accept its terms or sign its contract, as required by the plan.

Build what comes next.

Request an invitation and tell the team what you want to build.