GET STARTED
Find the right model.
Choose an available model and use its exact ID in your request.
Explore the catalog
Browse Models to compare providers, capabilities, context limits and token prices. A catalog entry alone does not mean the model can be called. Check its availability and supported endpoint on the model page.
Some models require streaming. Some support Responses only. Structured output and function calling depend on the model.
GET /v1/models
This authenticated endpoint lists enabled public model IDs. Use an exact ID from the list in the model field. The response does not include full capability or pricing metadata.
curl https://api.tokely.me/v1/models \
-H "Authorization: Bearer $TOKELY_API_KEY"Context and output
The service accepts an explicit output limit up to the lower of the model’s published limit and 65,536 tokens. Models without an established output limit use a 4,096-token service cap.
The default output budget is 256 tokens, including reasoning. Increase it for longer answers and reasoning models, and ensure your balance covers the reservation.