qwen3-235b-a22b API
qwen3-235b-a22b is a text-generation model from Alibaba. In ServerNet, its chat route accepts conversation messages and returns a generated answer; the price separates input and output tokens.
PreparingToman per million tokens
Prices exclude tax. The current Iran usage tax rate is 10%. The rate applicable to your account is fixed at request admission.
- API model ID
- gapgpt-qwen3-235b-a22b-b4441d4fdd
- Model developer
- Alibaba
- Service type
- Text chat
- Current availability
- Preparing
- Configured context limit
- 4,096 tokens
- Gateway output cap
- 1,024 tokens
Model features
Streaming ✓
What is qwen3-235b-a22b?
Use the public identifier gapgpt-qwen3-235b-a22b-b4441d4fdd for qwen3-235b-a22b in ServerNet. The gateway uses this identifier to select the configured route; the developer is Alibaba.
For qwen3-235b-a22b, verify the exact generation and task variant. Text, coding, vision and speech models in the Qwen family have different input contracts.
Where to use qwen3-235b-a22b
- Add qwen3-235b-a22b to a support assistant that receives the relevant conversation as text.
- Test qwen3-235b-a22b on summaries and extraction using your own documents, with a small output limit first.
- Compare answer quality, total token usage and cost per completed task before choosing qwen3-235b-a22b for production.
Build your integration
Send system and user messages, set max_tokens and store X-Request-Id with your usage record. Stream only when this route supports it. Function calls describe actions for your application to run; the gateway does not execute your tools.
Inputs and service limits
The current chat route accepts text only and one completion per request. Images, audio and hosted paid tools are not enabled by the source model’s capabilities. The configured context and output limits may be below the developer’s maximum. Test language quality on your own dataset.
A request cost example
For 1,000 input tokens and 500 output tokens, without a cache discount. This is a catalogue estimate, not a fixed request fee. Reported consumption, applicable tax and the price frozen for your request determine settlement.
- Usage before tax, Toman
- 112
- Tax (10%), Toman
- 12
- Estimated total, Toman
- 124
Connect to qwen3-235b-a22b API
Enable ai:chat on your key and ai:models:read to check availability. Set SERVERNET_API_KEY in your server environment. Change the idempotency key for each new request; keep it unchanged only for an identical retry.
This example documents the request format. Send it only after this model appears in GET /v1/models.
curl https://servernet.cloud/v1/chat/completions \
-H "Authorization: Bearer $SERVERNET_API_KEY" \
-H 'Content-Type: application/json' \
-H 'Idempotency-Key: CHANGE_FOR_EACH_NEW_REQUEST' \
-d '{
"model": "gapgpt-qwen3-235b-a22b-b4441d4fdd",
"messages": [
{
"role": "user",
"content": "Hello"
}
],
"max_tokens": 128
}'
Use the same Idempotency-Key when retrying an identical request. Successful responses are replayed without a second charge for 24 hours. Do not automatically retry when x-should-retry:false is present. The key is shown once. Store it in a server environment variable. Keep it out of browser code, GitHub and logs.
Keys, budgets and usage reportsFrequently asked questions
Is qwen3-235b-a22b available through ServerNet?
qwen3-235b-a22b is currently unavailable for API use. A catalogue listing does not activate billing or access. Choose an available related model or check this page again after the route is verified.
Which model ID and API endpoint should I use for qwen3-235b-a22b?
Use gapgpt-qwen3-235b-a22b-b4441d4fdd as the model field, the endpoint shown below and the matching key scope. ServerNet keys are project-specific; an SDK also needs https://servernet.cloud/v1 as its base URL.
How is qwen3-235b-a22b usage charged?
Input and output tokens use separate tariffs. Reported reasoning tokens, where supported, count towards output consumption. Cached input receives a discount only when a verified cached tariff exists. Usage and applicable tax appear separately in the dashboard.
What should I check before using qwen3-235b-a22b in production?
The current chat route accepts text only and one completion per request. Images, audio and hosted paid tools are not enabled by the source model’s capabilities. The configured context and output limits may be below the developer’s maximum. Test language quality on your own dataset.
Compare models for the same task
claude-fable-5
Available Text chatConfigured context limit: 4,096 tokens
claude-fable-5
Available Text chatConfigured context limit: 1,000,000 tokens
claude-fable-5-1
Available Text chatConfigured context limit: 4,096 tokens
claude-haiku-4-5
Available Text chatConfigured context limit: 200,000 tokens
claude-haiku-4-5-20251001
Available Text chatConfigured context limit: 4,096 tokens
claude-haiku-5-5
Available Text chatConfigured context limit: 4,096 tokens