List prices
| Input | Short context | $10 / 1M tok |
| Cached input | Short context | $1 / 1M tok |
| Cache writes | Short context | $12.5 / 1M tok |
| Output | Short context | $50 / 1M tok |
| Input | Long context | $20 / 1M tok |
| Cached input | Long context | $2 / 1M tok |
| Cache writes | Long context | $25 / 1M tok |
| Output | Long context | $75 / 1M tok |
Estimate your monthly cost
Based on GPT-6 Astra's current list prices. You pay your provider directly - Vevee meters this usage against your own plans.
Repeated prompts can be cheaper: cached input is $1.00 / 1M tokens.
Frequently asked questions
How much does GPT-6 Astra cost?
GPT-6 Astra costs $10 per 1 million input tokens and $50 per 1 million output tokens, with cached input at $1 per 1 million tokens. These are OpenAI's list rates for the Short context tier.
How much does a typical GPT-6 Astra request cost?
A request with 1,500 input tokens and 500 output tokens costs about $0.040. At 1,000 such requests a month that is roughly $40.00.
How many GPT-6 Astra tokens do I get for $10?
$10 buys roughly 1.0 million input tokens or 0.2 million output tokens at list price.
How do I track GPT-6 Astra costs per end user?
Use a metering layer. With Vevee you call track() or reserve()/commit() around each GPT-6 Astra request with the end user's ID, define plan limits in the dashboard, and Vevee enforces them and shows per-user usage and cost - no backend to build.
Track GPT-6 Astra spend per user
Knowing the list price is half the problem - the other half is knowing which of your users consume it and stopping the ones who blow past their plan. Vevee meters every GPT-6 Astra call per end user, enforces your plan limits, and shows you per-user cost. Free tier, no card.