THE COST OF BUILDING WITH AI

Lower-cost GPT API access

Compare published GPT input, output and cached-input prices. Find the exact model ID, supported endpoint and output limit before migrating your OpenAI SDK.

THE SHORT ANSWER

Supported GPT models are available with a Tokely API key at https://api.tokely.me/v1. Use the exact model page to choose Chat Completions or Responses. Price comparisons must match the GPT version and token mix; ChatGPT subscriptions are separate products.

Match the endpoint as well as the model

Some models use Chat Completions, others use the supported Responses subset. Tokely Responses is stateless: include history in each request and use store=false. Provider-hosted tools and every OpenAI endpoint are not implied by SDK compatibility.

Budget for reasoning and output bounds

Reasoning can count toward output usage. Set an explicit output limit within the Tokely model's published bound, and leave sufficient available balance for the reservation. A low requested limit can stop a useful answer before it completes.

GPT API price comparison

USD per 1M tokens. Example: 1M uncached input + 250K output. Tokely rates dated 2026-10-09; OpenRouter references checked 2026-10-09. Platform/checkout fees, taxes, special units and context tiers excluded. Different models are not quality equivalents. Methodology.
ModelTokely input / outputTokely exampleOpenRouter exampleDifference vs OpenRouter
GPT-5.4 MiniOpenAI$0.07425 / $0.4455Cache read: $0.007425$0.185625$1.875$0.75 / $4.5 per 1M90.1% lower
GPT-6 LunaOpenAI$0.09 / $0.45Cache read: $0.009$0.2025$0.225$0.1 / $0.5 per 1M10.0% lower
GPT-5 CodexOpenAI$0.126397 / $1.011175Cache read: $0.01264$0.379191Not matched—
GPT-5.6 LunaOpenAI$0.18 / $1.08Cache read: $0.018$0.45$0.5$0.2 / $1.2 per 1M10.0% lower
GPT-6 SolOpenAI$0.505588 / $1.516763Cache read: $0.040447$0.884779$4.5$2 / $10 per 1M80.3% lower
GPT-6.1 SolOpenAI$0.505588 / $1.516763Cache read: $0.020224$0.884779$4.5$2 / $10 per 1M80.3% lower
GPT-5.6 TerraOpenAI$0.505588 / $1.820115Cache read: $0.040447$0.960617$5$2 / $12 per 1M80.8% lower
GPT-5.2 CodexOpenAI$0.641667 / $5.133334Cache read: $0.064167$1.925001$5.25$1.75 / $14 per 1M63.3% lower
GPT-5.5OpenAI$1.011175 / $4.550288Cache read: $0.101118$2.148747$12.5$5 / $30 per 1M82.8% lower
GPT-5.6 SolOpenAI$1.263969 / $4.550288Cache read: $0.101118$2.401541$4.5$2 / $10 per 1M46.6% lower
GPT-5.1 CodexOpenAI$1.008334 / $8.066667Cache read: $0.100834$3.025001$3.75$1.25 / $10 per 1M19.3% lower
GPT-5.3 CodexOpenAI$1.05875 / $8.47Cache read: $0.105875$3.17625$5.25$1.75 / $14 per 1M39.5% lower
GPT-6 AstraOpenAI$2.527938 / $7.583813Cache read: $0.202235$4.423891$22.5$10 / $50 per 1M80.3% lower
GPT-5.2OpenAI$1.575 / $12.6Cache read: $0.1575$4.725$5.25$1.75 / $14 per 1M10.0% lower
GPT-5.4OpenAI$2.25 / $13.5Cache read: $0.225$5.625$6.25$2.5 / $15 per 1M10.0% lower

Change the token mix in the calculator →

Questions before you switch

Can I keep the OpenAI Python or JavaScript SDK?

Yes for the supported endpoints and parameters. Set the Tokely base URL and key, then use a supported Tokely model ID. Check the migration guide for differences.

Does a ChatGPT subscription cover API usage?

No. Tokely API usage spends your Tokely balance; a ChatGPT subscription does not fund that balance.

Keep exploring

OpenRouter migration checklist · Billing documentation

START WITH YOUR OWN WORKLOAD

Know the price.
Then make the call.

Check the model, compare the rate, and connect with a Tokely key.

Read the quickstart ↗Explore all models →