Qwen3 VL 235B A22B Instruct
Qwen3 VL 235B A22B Instruct is Alibaba's general-purpose model. It reads images. It takes up to 256K tokens of context and writes up to 129.02K per response.
Output price
$1.008
USD / 1M tokens
37% below list price
Add $5, get $5 more on your first top-up.
Same Qwen3 VL 235B A22B Instruct. 37% less.
Specs.
- Model ID
- qwen3-vl-235b-a22b-instruct
- Made by
- Alibaba
- Context window
- 256K tokens
- Max output
- 129.02K tokens
- Endpoint
- /v1/chat/completions
- Billing
- Per token, from one balance
Edge gateways
Requests enter at the gateway nearest your app and go straight to the model, so routing adds very little to each call.
One key, every model
Switch models by changing a string. One prepaid balance covers all of them, billed per token.
Nothing stored
We keep token counts for billing, never your prompts or responses.
Call Qwen3 VL 235B A22B Instruct in a minute.
Point any OpenAI-compatible client at our base URL, keep your key in an environment variable, and use the model ID qwen3-vl-235b-a22b-instruct.
import os
from openai import OpenAI
client = OpenAI(
base_url="https://api.promptsforless.com/v1",
api_key=os.environ["PFL_API_KEY"],
)
response = client.chat.completions.create(
model="qwen3-vl-235b-a22b-instruct",
messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)More from Alibaba.
Questions.
How much does Qwen3 VL 235B A22B Instruct cost?
$1.008 / 1M tokens output and $0.252 / 1M tokens input, 37% below Alibaba's list price. You pay per token from a prepaid balance, with no subscription or minimum. We match your first $5 top-up, so you start with $10 of credit.
Is it the real Qwen3 VL 235B A22B Instruct?
Yes. Your requests run on Alibaba's Qwen3 VL 235B A22B Instruct, not a distilled copy or a substitute. The discount comes from how we buy capacity, not from changing the model.
Which tools can use Qwen3 VL 235B A22B Instruct?
Anything that speaks the OpenAI API: set the base URL to our endpoint and use the model ID qwen3-vl-235b-a22b-instruct. We have setup guides for Cursor, Codex, OpenCode, the OpenAI and Vercel AI SDKs, LangChain and more.
How fast is it?
Requests go through our edge gateways, so the network hop adds very little on top of Qwen3 VL 235B A22B Instruct's own generation time. Responses stream back token by token.