All models

Qwen

qwen3-30b-a3b-think

Qwen3's training dataset is significantly larger than Qwen2.5's. While Qwen2.5 was pretrained on 18 trillion tokens, Qwen3 uses nearly twice that amount — approximately 36 trillion tokens — covering 119 languages and dialects.

textAPI availablestreaming
qwen3-30b-a3b-think

Tokely price / 1M tokens

$0.25input
$0.25cached input
$3output

USD · Published rate · Updated 2026-10-06

Context window

UnconfirmedNo verified context specification

First request

Available
curl https://tokely.me/v1/chat/completions \
  -H "Authorization: Bearer $TOKELY_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "qwen3-30b-a3b-think",
    "messages": [{"role": "user", "content": "Hello!"}],
    "max_tokens": 256
  }'

Set TOKELY_API_KEY to your dashboard key. Requests require a funded Tokely balance. Base URL: https://tokely.me/v1.

Compare nearby prices

USD per million tokens. Compare price and context; these are not quality benchmarks.
ModelInputOutputContext
qwen3-30b-a3b-thinkQwen · this model$0.25$3Unconfirmed
gpt-5.4-nanoOpenAI$0.45$2.8125400,000
deepseek-v4-flashDeepSeek$0.825$2.4751,048,576
qwen3-235b-a22b-instruct-2507Qwen$0.6325$2.53262,144
gemini-flash-latestGoogle$0.330885$2.76289Unconfirmed

Capabilities & limits

Provider
Qwen
API model ID
qwen3-30b-a3b-think
Category
Text
Endpoints
POST /v1/chat/completions
Tokely max output
4,096 tokens / request, including reasoning tokens
Default output
256 tokens. Set an explicit limit for longer responses.
Request limits
256 KB JSON body · 128 messages · 32 function tools
Rate limit
60 requests / minute / account, across its keys
Supported input
Text messages and function tool results
Model reference

API reference

Authenticate with Authorization: Bearer YOUR_TOKELY_API_KEY. Use an OpenAI SDK with the Tokely base URL.

model
Required string: qwen3-30b-a3b-think
messages
An array of system, developer, user, assistant and tool messages.
max_tokens
Optional integer, 1–4,096. Defaults to 256.
stream
Set true for Server-Sent Events (SSE).

The response contains choices with message or tool_calls, finish_reason, and usage (prompt_tokens, completion_tokens, total_tokens). Streaming sends chat completion chunks.

400 / 415
Invalid request, parameter, endpoint for this model, or content type.
401 / 403
Invalid, revoked or restricted API key.
402
Insufficient account balance or API key budget.
429
Request limit reached. Retry after the indicated delay.
502 / 503 / 504
Upstream failure, temporary service unavailability or request timeout.

Frequently asked questions

How much does qwen3-30b-a3b-think cost?

$0.25 per million input tokens, $3 per million output tokens and $0.25 per million cached input tokens. Usage is charged to your Tokely balance.

Do I need a separate Qwen account?

Use your Tokely API key and balance. A separate provider key is not required.

How do I switch to this model?

Set model to qwen3-30b-a3b-think and use /v1/chat/completions. Keep your Tokely base URL and API key.

Related versions