List prices
| Input | Short context | $0.1 / 1M tok |
| Cached input | Short context | $0.01 / 1M tok |
| Cache writes | Short context | $0.125 / 1M tok |
| Output | Short context | $0.5 / 1M tok |
| Input | Long context | $0.2 / 1M tok |
| Cached input | Long context | $0.02 / 1M tok |
| Cache writes | Long context | $0.25 / 1M tok |
| Output | Long context | $0.75 / 1M tok |
Estimate your monthly cost
Based on GPT-6 Luna's current list prices. You pay your provider directly - Vevee meters this usage against your own plans.
Repeated prompts can be cheaper: cached input is $0.010 / 1M tokens.
Frequently asked questions
How much does GPT-6 Luna cost?
GPT-6 Luna costs $0.1 per 1 million input tokens and $0.5 per 1 million output tokens, with cached input at $0.01 per 1 million tokens. These are OpenAI's list rates for the Short context tier.
How much does a typical GPT-6 Luna request cost?
A request with 1,500 input tokens and 500 output tokens costs about $0.0004. At 1,000 such requests a month that is roughly $0.400.
How many GPT-6 Luna tokens do I get for $10?
$10 buys roughly 100 million input tokens or 20 million output tokens at list price.
How do I track GPT-6 Luna costs per end user?
Use a metering layer. With Vevee you call track() or reserve()/commit() around each GPT-6 Luna request with the end user's ID, define plan limits in the dashboard, and Vevee enforces them and shows per-user usage and cost - no backend to build.
Track GPT-6 Luna spend per user
Knowing the list price is half the problem - the other half is knowing which of your users consume it and stopping the ones who blow past their plan. Vevee meters every GPT-6 Luna call per end user, enforces your plan limits, and shows you per-user cost. Free tier, no card.