Skip to content
Neox Platform

Pay-as-you-go API

OpenAI- and Anthropic-compatible. Prepaid balance, billed on actual tokens. Works with any client or program that supports a custom endpoint.

Three steps

1Top up

Add funds in the console. Credited 1:1, never expires.

2Create a key

Generate an sk-neox- key, optionally scoped to models and a spend limit.

3Set the base URL

Point your client or SDK at the base URL below.

https://gateway.neox-dev.com/v1
from openai import OpenAI

client = OpenAI(
    base_url="https://gateway.neox-dev.com/v1",
    api_key="sk-neox-...",
)
reply = client.chat.completions.create(
    model="deepseek-v4-flash",
    messages=[{"role": "user", "content": "hi"}],
)
print(reply.choices[0].message.content)

Models & pricing

Billed on actual tokens, USD per million tokens. Failed requests are not charged.

ModelInputCached inputOutput

Loading…

Browse all models

Errors and limits

Errors come back in an OpenAI-compatible body: {"error": {"code", "message", "requestId"}}. Branch on error.code, not on the message text.

HTTPerror.codeMeaningWhat to do
401platform.invalid_keyInvalid or revoked keyCheck the key or create a new one in the console
402platform.insufficient_balanceBalance too lowTop up in the console and retry
403platform.key_limitThis key hit its spend limitRaise the key limit or use another key
403platform.wallet_frozenBalance is frozenContact support@neox-dev.com
404platform.model_not_offeredModel not offeredList models with GET /models
429quota.rate_limitRate limited: 600 requests/min and 10 concurrent per accountLower concurrency and retry with backoff
429upstream.rate_limitedThe model provider is throttlingRetry later or switch models

How it differs from Neox Studio

Neox Studio

A subscription for the Neox desktop app, CLI and Android, with an included allowance.

Neox Platform

Pay as you go, from any client or your own programs. Separate from your subscription.

FAQ

Does my balance expire?

No. Top-ups are credited 1:1 and drawn down as you use them.

Am I charged for failed requests?

No. A hold estimated from the input and the maximum output (max_tokens) is placed when a request starts, settled on actual tokens when it succeeds, and released in full if it fails.

Which endpoints are supported?

OpenAI-compatible /v1/chat/completions (including streaming), /v1/responses, /v1/embeddings and /v1/models, plus Anthropic-compatible /v1/messages (including streaming and tool use).

Can I cap what a key spends?

Yes. Each key can have its own spend limit and allowed models; it is refused once the limit is reached.

Do requests pass through Neox servers?

Yes. Platform requests are relayed to model providers through the Neox gateway. Prompt content is not retained by default.

Is the Anthropic format (/v1/messages) supported?

Yes. Set the base URL of your Anthropic SDK or tool to https://gateway.neox-dev.com (without /v1) and use your sk-neox- key. Both x-api-key and Authorization: Bearer are accepted. /v1/messages/count_tokens returns an estimate.

Can I run my own agent on it?

Yes. The open-source Neox Agent SDK connects with an openai-compatible provider. See the Neox Agent SDK tab under Examples, and /sdk.

Can my platform balance be used in the Neox apps?

Yes. Your account has a single platform balance: besides API calls, subscribers who use up their plan allowance in the Neox desktop app, CLI or Android can keep going on the balance (extra usage, which you can turn off anytime).

Start with Neox Platform

Top up, create a key, and start calling.