Cheap tokens for open-weight Qwen models, served on an OpenAI-compatible API. From $0.014 per million tokens. Billed by usage, no minimums.
One key for inference and the no-code app. Read the docs →
Qwen instruct for chat and tool use. Qwen3 Coder for agentic coding. Qwen embedding for retrieval. Served from dedicated GPUs that stay warm: no cold starts, no pool sizing, no warmup pings. OpenAI-compatible, a drop-in for the closed providers you already pay too much for.
Prices are USD per million tokens, deducted from your credit balance. 1,000 credits = $1, so $0.06 per Mtok is about 60 credits.
Drop our base URL into your OpenAI client. Or skip the SDK and hit a live model right here, no signup needed.
Ten free requests per day, paid for by us, just to prove the latency claim.
API key for the standard flow. x402 for one-off and agent-to-agent calls. Pay per request in USDC, no account needed.
Sign up, generate an API key, drop it into any OpenAI SDK. The same key works for inference and the data platform.
No signup. Send a request, get a 402 back with payment details, pay with USDC on Base, retry. Works with x402-fetch or any x402-compatible client.
No setup, no infra, no calls with a sales engineer. Three steps and you are streaming.
Free account, no card. Monthly credits included. Takes about thirty seconds.
Generate an API key in your account. The same key works for inference and data collection.
Point any OpenAI client at api.napu.ai/v1, pick a model, ship.
Frontier APIs charge dollars per million tokens. We charge cents. Same OpenAI client, smaller bill.
The questions everyone sends us before signing up.
Billed by usage from your credit balance. Cockpits and the no-code app are not included.