Pricing
Your choice of frontier model — at 20% less
Run Claude Opus 4.8, GPT-5.5 and more through one API, billed per million tokens at a flat 20% under each provider's list price. No subscriptions, no per-seat fees, no minimums.
Your choice of frontier model
Run the frontier models you already trust — billed per million tokens at 20% below each provider's list price. No tiers, no seats.
| Model | Context | Input / M | Output / M |
|---|---|---|---|
Claude Opus 4.8 Anthropic | 200k | $4.00 $5.00 | $20.00 $25.00 |
GPT-5.5 OpenAI | 400k | $3.20 $4.00 | $12.80 $16.00 |
Claude Sonnet 4.6 Anthropic | 200k | $2.40 $3.00 | $12.00 $15.00 |
Gemini 3 Pro Google | 1M | $2.00 $2.50 | $8.00 $10.00 |
How billing works
Pre-load credits and pay only for the input and output tokens you use. Spend and budgets are visible in your dashboard.
Every model is billed at 20% under its provider list price; your invoice is the sum of metered usage across the models you route to.
Enterprise & sovereign
Sovereign lanes (regional data residency), pinned models, dedicated capacity, and committed-use discounts are available under an enterprise agreement.
Questions, answered plainly
How am I billed?
Can I cap what a workspace spends?
Why is Graphene cheaper than going direct?
Do you support data residency requirements?
Get started
Create a workspace, issue your first API key, and route inference through Graphene with the same contracts you see in the docs.