Cheaper on every token.
Credits are prepaid inference, spent at your tier's rates on any model in the catalog. There is no subscription and no commitment beyond the first $5.
| Per 1M tokens | Provider list | Base (−20%) | Scale (−25%) |
|---|---|---|---|
| GPT-5.6 Solgpt-5.6-sol | |||
| Input | $5.00 | $4.00 | $3.75 |
| Cached input | $0.50 | $0.40 | $0.375 |
| Output | $30.00 | $24.00 | $22.50 |
| GPT-5.6 Terragpt-5.6-terra | |||
| Input | $2.50 | $2.00 | $1.88 |
| Cached input | $0.25 | $0.20 | $0.188 |
| Output | $15.00 | $12.00 | $11.25 |
| GPT-5.6 Lunagpt-5.6-luna | |||
| Input | $1.00 | $0.80 | $0.75 |
| Cached input | $0.10 | $0.08 | $0.075 |
| Output | $6.00 | $4.80 | $4.50 |
| GPT-5.4 Progpt-5.4-pro | |||
| Input | $30.00 | $24.00 | $22.50 |
| Cached input | — | — | — |
| Output | $180 | $144 | $135 |
| Grok 4.3grok-4.3 | |||
| Input | $1.25 | $1.00 | $0.938 |
| Cached input | $0.20 | $0.16 | $0.15 |
| Output | $2.50 | $2.00 | $1.88 |
| Kimi K2.7 Codekimi-k2.7-code | |||
| Input | $0.95 | $0.76 | $0.713 |
| Cached input | $0.19 | $0.152 | $0.142 |
| Output | $4.00 | $3.20 | $3.00 |
| Kimi K2.7 Code Fastkimi-k2.7-code-fast | |||
| Input | $1.05 | $0.84 | $0.787 |
| Cached input | $0.18 | $0.144 | $0.135 |
| Output | $4.40 | $3.52 | $3.30 |
| GPT Image 2gpt-image-2 · Scale tier only | |||
| Text input | $5.00 | $4.00 | $3.75 |
| Cached text input | $1.25 | $1.00 | $0.938 |
| Image input | $8.00 | $6.40 | $6.00 |
| Cached image input | $2.00 | $1.60 | $1.50 |
| Image output | $30.00 | $24.00 | $22.50 |
| GPT Realtime 2.1gpt-realtime-2.1 · Scale tier only | |||
| Text input | $4.00 | $3.20 | $3.00 |
| Cached text input | $0.40 | $0.32 | $0.30 |
| Text output | $24.00 | $19.20 | $18.00 |
| Audio input | $32.00 | $25.60 | $24.00 |
| Cached audio input | $0.40 | $0.32 | $0.30 |
| Audio output | $64.00 | $51.20 | $48.00 |
| Image input | $5.00 | $4.00 | $3.75 |
The discount applies to every model, no exceptions. Context windows and versions are on the models page. GPT-5.4 Pro has no cached input rate.
Scale tier
Unlocks automatically once lifetime purchases reach $5,000, applies to all subsequent usage, and never expires. There is no form to fill in and nobody to call.
Credits
One credit is one dollar of inference at your tier's rates. The minimum purchase is $5, and unused credits are refundable within 30 days.
Estimate your monthly cost.
Drag the sliders to your expected volume.
Provider list
$550.00
Base
$440.00
−$110.00/mo
Scale
$412.50
−$137.50/mo
Assumes uncached input. Cached input for gpt-5.6-sol is billed at $0.40 per 1M on Base, $0.38 on Scale.
Common questions.
- Can I get a refund?
- Unused credits are refundable within 30 days of purchase. Contact support with your account email and the unused balance goes back to your original payment method.
- What are the rate limits?
- Fair use, with defaults most workloads never hit. If you run sustained high-volume traffic and need reserved throughput, contact us.
- What counts as cached input?
- Parts of your prompt the servers have seen recently, such as a repeated system prompt. Cache hits are reported in usage.prompt_tokens_details.cached_tokens and billed at the lower rate automatically. Nothing to configure.
- Are these the real models?
- Yes. Each model reports the exact version it is, for example gpt-5.6-sol-2026-07-09, and every response is cryptographically signed so you can prove it. Check any response on the verify page.
- What payment methods do you accept?
- Cards, processed by Stripe. Purchases are one-time credit top-ups starting at $5. Nothing recurring is stored against your account.
Start with $5.Priced like infrastructure.
Unused credits are refundable for 30 days.