THE COST OF BUILDING WITH AI
Lower-cost GPT API access
Compare published GPT input, output and cached-input prices. Find the exact model ID, supported endpoint and output limit before migrating your OpenAI SDK.
THE SHORT ANSWER
Supported GPT models are available with a Tokely API key at https://api.tokely.me/v1. Use the exact model page to choose Chat Completions or Responses. Price comparisons must match the GPT version and token mix; ChatGPT subscriptions are separate products.
Match the endpoint as well as the model
Some models use Chat Completions, others use the supported Responses subset. Tokely Responses is stateless: include history in each request and use store=false. Provider-hosted tools and every OpenAI endpoint are not implied by SDK compatibility.
Budget for reasoning and output bounds
Reasoning can count toward output usage. Set an explicit output limit within the Tokely model's published bound, and leave sufficient available balance for the reservation. A low requested limit can stop a useful answer before it completes.
GPT API price comparison
| Model | Tokely input / output | Tokely example | OpenRouter example | Difference vs OpenRouter |
|---|---|---|---|---|
| GPT-5.4 MiniOpenAI | $0.07425 / $0.4455Cache read: $0.007425 | $0.185625 | $1.875$0.75 / $4.5 per 1M | 90.1% lower |
| GPT-6 LunaOpenAI | $0.09 / $0.45Cache read: $0.009 | $0.2025 | $0.225$0.1 / $0.5 per 1M | 10.0% lower |
| GPT-5 CodexOpenAI | $0.126397 / $1.011175Cache read: $0.01264 | $0.379191 | Not matched | — |
| GPT-5.6 LunaOpenAI | $0.18 / $1.08Cache read: $0.018 | $0.45 | $0.5$0.2 / $1.2 per 1M | 10.0% lower |
| GPT-6 SolOpenAI | $0.505588 / $1.516763Cache read: $0.040447 | $0.884779 | $4.5$2 / $10 per 1M | 80.3% lower |
| GPT-6.1 SolOpenAI | $0.505588 / $1.516763Cache read: $0.020224 | $0.884779 | $4.5$2 / $10 per 1M | 80.3% lower |
| GPT-5.6 TerraOpenAI | $0.505588 / $1.820115Cache read: $0.040447 | $0.960617 | $5$2 / $12 per 1M | 80.8% lower |
| GPT-5.2 CodexOpenAI | $0.641667 / $5.133334Cache read: $0.064167 | $1.925001 | $5.25$1.75 / $14 per 1M | 63.3% lower |
| GPT-5.5OpenAI | $1.011175 / $4.550288Cache read: $0.101118 | $2.148747 | $12.5$5 / $30 per 1M | 82.8% lower |
| GPT-5.6 SolOpenAI | $1.263969 / $4.550288Cache read: $0.101118 | $2.401541 | $4.5$2 / $10 per 1M | 46.6% lower |
| GPT-5.1 CodexOpenAI | $1.008334 / $8.066667Cache read: $0.100834 | $3.025001 | $3.75$1.25 / $10 per 1M | 19.3% lower |
| GPT-5.3 CodexOpenAI | $1.05875 / $8.47Cache read: $0.105875 | $3.17625 | $5.25$1.75 / $14 per 1M | 39.5% lower |
| GPT-6 AstraOpenAI | $2.527938 / $7.583813Cache read: $0.202235 | $4.423891 | $22.5$10 / $50 per 1M | 80.3% lower |
| GPT-5.2OpenAI | $1.575 / $12.6Cache read: $0.1575 | $4.725 | $5.25$1.75 / $14 per 1M | 10.0% lower |
| GPT-5.4OpenAI | $2.25 / $13.5Cache read: $0.225 | $5.625 | $6.25$2.5 / $15 per 1M | 10.0% lower |
Questions before you switch
Can I keep the OpenAI Python or JavaScript SDK?
Yes for the supported endpoints and parameters. Set the Tokely base URL and key, then use a supported Tokely model ID. Check the migration guide for differences.
Does a ChatGPT subscription cover API usage?
No. Tokely API usage spends your Tokely balance; a ChatGPT subscription does not fund that balance.
Keep exploring
AI API pricing. Know what you pay. ↗
Published Tokely rates for text, images, video and audio. Pay from one USD balance, compare exact models and estimate your workload before you switch.
A cost-focused OpenRouter alternative ↗
Compare exact-model rates, supported endpoints and migration requirements before moving to Tokely. One balance for supported text, image, video and audio models.
AI API price index ↗
An exportable first-party dataset of enabled Tokely models, exact-model OpenRouter token references and supported media variants. Sources, units and dates are included.
START WITH YOUR OWN WORKLOAD
Know the price.
Then make the call.
Check the model, compare the rate, and connect with a Tokely key.