Nexith Core is now serving on our own sovereign GPU cluster — cache hits save up to 90%, OpenAI-compatible.Learn more
Nexith
Sign Up
Pricing

One account.
Two ways to build.

Choose a Token Plan for your monthly capacity, then pay API usage transparently by the tokens your models actually process.

Start buildingSee API rates ↓
01 / Token Plans

Capacity that fits your work.

Monthly token capacity and rate limits are configured in the Nexith control plane.

Pro

POPULAR
Unlimited / month
600 requests / min
400,000 tokens / min
No daily spend cap
Choose plan

Enterprise

Unlimited / month
6,000 requests / min
4,000,000 tokens / min
No daily spend cap
Talk to sales
02 / API Pricing

Pay for the tokens you use.

Public model rates, in USD per 1M tokens. Cached input reflects eligible cache hits.

Model
Input
Cached input
Output
Nexith Core
nexith-core
$4.00
$0.40
$20.00
How usage works

Every token has a line item.

Input tokens

Your prompt, documents and images sent to the model.

Cached input

Eligible repeated context is charged at the lower cached rate—not advertised as free.

Output tokens

The generated response, streamed or returned in one response.

Frequently asked questions

A Token Plan sets your monthly capacity and API rate limits.

Need capacity beyond standard plans?

Private deployments, custom limits and dedicated support for production teams.

Contact sales
Nexith — Frontier AI Models | Nexith