Every request, itemized
A per-request ledger of model, tokens and exact cost, with spend by key and by model. Know what each feature costs.
A drop-in, OpenAI-compatible API. Keep your SDK, your prompts and your code. Change the base URL and key, and every token costs 30% less than list.
The customer was charged twice for the March invoice after updating their card, and wants the duplicate $49 payment refunded to the original card before the next billing date.
Illustrative request on gpt-4o-mini, priced from the live rate card. OpenAI list shown struck through; Blizon is 30% less.
Works with anything that speaks the OpenAI API
Quickstart
Blizon speaks the Chat Completions API. Point your client at a new base URL, swap the key, and everything else stays exactly where it is.
import osfrom openai import OpenAIclient = OpenAI(removed:api_key=os.environ["OPENAI_API_KEY"],added:base_url="https://api.blizon.tech/v1",added:api_key=os.environ["BLIZON_API_KEY"],)stream = client.chat.completions.create(model="gpt-4.1-mini",messages=[{"role": "user", "content": "Hello"}],stream=True,)
Platform
Requests go straight to the model. Everything around them, from keys to costs to failover, is handled for you.
A per-request ledger of model, tokens and exact cost, with spend by key and by model. Know what each feature costs.
Streaming responses are passed through as server-sent events, chunk by chunk. No buffering.
If an upstream errors out, Blizon automatically retries the request on a healthy one.
One key per app or environment. Track spend per key and revoke any of them instantly.
We log token counts and cost for billing. Prompt and completion contents are never stored.
1 credit = $1. Credits are assigned to your account after approval and each request deducts its exact cost. No plans, no seats, no monthly fee.
Savings
Same models, same tokens, 30% off the bill. Drag to your current monthly OpenAI spend.
Type any amount, or use the slider from $100 to $100,000.
You save
Pricing
One flat discount on input, cached input and output, for every model. No tiers, no commitments.
| Model | Input | Cached input | Output | You save |
|---|---|---|---|---|
| GPT-5.x7 | ||||
gpt-5 | $0.875OpenAI list $1.25 | $0.0875OpenAI list $0.125 | $7.00OpenAI list $10.00 | −30% |
gpt-5-mini | $0.175OpenAI list $0.25 | $0.0175OpenAI list $0.025 | $1.40OpenAI list $2.00 | −30% |
gpt-5-nano | $0.035OpenAI list $0.05 | $0.0035OpenAI list $0.005 | $0.28OpenAI list $0.40 | −30% |
gpt-5.1 | $0.875OpenAI list $1.25 | $0.0875OpenAI list $0.125 | $7.00OpenAI list $10.00 | −30% |
gpt-5.3-codex | $1.23OpenAI list $1.75 | $0.1225OpenAI list $0.175 | $9.80OpenAI list $14.00 | −30% |
gpt-5.4 | $1.75OpenAI list $2.50 | $0.175OpenAI list $0.25 | $10.50OpenAI list $15.00 | −30% |
gpt-5.5-instant | $3.50OpenAI list $5.00 | $0.35OpenAI list $0.50 | $21.00OpenAI list $30.00 | −30% |
| GPT-4.13 | ||||
gpt-4.1 | $1.40OpenAI list $2.00 | $0.35OpenAI list $0.50 | $5.60OpenAI list $8.00 | −30% |
gpt-4.1-mini | $0.28OpenAI list $0.40 | $0.07OpenAI list $0.10 | $1.12OpenAI list $1.60 | −30% |
gpt-4.1-nano | $0.07OpenAI list $0.10 | $0.0175OpenAI list $0.025 | $0.28OpenAI list $0.40 | −30% |
| GPT-4o2 | ||||
gpt-4o | $1.75OpenAI list $2.50 | $0.875OpenAI list $1.25 | $7.00OpenAI list $10.00 | −30% |
gpt-4o-mini | $0.105OpenAI list $0.15 | $0.0525OpenAI list $0.075 | $0.42OpenAI list $0.60 | −30% |
| GPT-3.51 | ||||
gpt-3.5-turbo | $0.35OpenAI list $0.50 | —not available | $1.05OpenAI list $1.50 | −30% |
How it works
Tell us what you're building. Every request is reviewed by hand, and we reply by email.
You get an invite to set a password, create API keys, and we load credits to your account.
Point your OpenAI client at Blizon with your new key. That's the whole migration.
Request access
Access is request-only. Every request is read by a person, and approved accounts get an invite and credits to start sending requests.
Everything except company is required.
FAQ
Something else on your mind? Mention it in your access request and we'll answer when we reply.
base_url to https://api.blizon.tech/v1 with your Blizon key.stream: true and you get standard server-sent events, passed through as they arrive. The gateway adds under 1 ms of overhead, and token usage is still recorded for billing.402 insufficient_quota error until credits are added. Nothing is charged beyond your balance.Request access, get approved, change two lines. Every token after that costs 30% less.