Early access · Paid API credit is live. New accounts still get free starter credit.
PricingDocsModelsFAQEarnBlogAboutSign in →

Home / Pricing

One live rate for every model.

No plans or seats. Pay per token at a platform-set per-model rate, the same no matter which Mac serves the request.

pay per token · from $0.03/M input · 12 models live

Up to 10M free tokens

$1 promotional credit on signup. No card required.

The real token count depends on the model and your input/output mix, so treat 10 million as an estimate. See the complete calculation.

Start free

Why idle Macs change the cost structure

Typical API pricing includes several layers between silicon and your app: capacity is bought, rented, repackaged, and metered. Umbra routes demand to Apple Silicon that is already paid for. The marginal cost is mostly electricity, which is how the network can undercut commodity hosted rates on the long-tail catalog without inventing subsidies.

Providers earn a majority share of metered on-demand revenue at the same platform-set rates you pay. Cash-out eligibility is in the provider agreement.

Live per-model rates

Priced per million tokens. Output is typically higher than input.

GET /v1/models

Alpha honesty: prices may change during the alpha. Pricing is platform-set and identical for every provider serving a model; providers do not set their own rates. How free credit works.

For developers

Keep your OpenAI or Anthropic SDK, change the base URL and key, pay only for tokens used. Free credit on signup, no card required. Read the quickstart.

Create an account

For providers

Earn per token from these same platform-set prices, a majority share of metered on-demand revenue. Cash out via the provider wallet when available; eligibility is in the provider agreement.

See host economics