Pricing

Usage-based — pay for the tokens you use, priced per model.sample data

Pay per token

Billed on input + output tokens at each model's rate. No subscription, no minimums.

Per-key spend caps

Put a monthly USD cap on any API key; live metering stops spend at the limit.

Prepaid wallet

Top up your Neviri AI wallet — a signup credit covers your first requests.

ModelProviderContextInput / 1MOutput / 1M
Embed LargeSelf-hosted8K$0.02—
Neviri Open 8BSelf-hosted32K$0.05$0.10
Vision Open 12BSelf-hosted128K$0.10$0.10
Fast MiniResold64K$0.15$0.60
Coder 32BSelf-hosted128K$0.18$0.18
Neviri Open MoESelf-hosted64K$0.24$0.24
Neviri Open 70BSelf-hosted32K$0.35$0.40
Reasoning MiniResold128K$0.60$2.40
Flagship ChatResold128K$2.50$10.00
Vision MultimodalResold128K$2.50$10.00
Reasoning ProResold200K$3.00$15.00
Frontier MaxResold200K$5.00$15.00

Prices shown are illustrative sample data. Final per-model rates are set in the gateway catalogue at launch; embeddings bill on input only.