Qwen3 235B A22B Instruct 2507

Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following, logical reasoning, math, code, and tool usage. The model supports a native 262K context length and does not implement "thinking mode" (<think> blocks). Compared to its base variant, this version delivers significant gains in knowledge coverage, long-context reasoning, coding benchmarks, and alignment with open-ended tasks. It is particularly strong on multilingual understanding, math reasoning (e.g., AIME, HMMT), and alignment evaluations like Arena-Hard and WritingBench.

ChatQwen262k tokens$0.0981 / $0.5995 · 1M

qwen/qwen3-235b-a22b-2507

Context
262k tokens
Completion cap
16,384
Tools
Yes
JSON
Yes
Released
2025-07-21

Call it from Geek Hub

Same OpenAI SDK. Change the base URL and the model id.

curl https://api.geekhub.mx/v1/chat/completions \
  -H "Authorization: Bearer $GEEKHUB_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "qwen/qwen3-235b-a22b-2507",
    "messages": [{"role": "user", "content": "Hola"}]
  }'
Get an API key

Frequently asked questions

What is Qwen3 235B A22B Instruct 2507?
Qwen3-235B-A22B-Instruct-2507 is a multilingual, instruction-tuned mixture-of-experts language model based on the Qwen3-235B architecture, with 22B active parameters per forward pass. It is optimized for general-purpose text generation, including instruction following, logical reasoning, math, code, and tool usage. The model supports a native 262K context length and does not implement "thinking mode" (<think> blocks). Compared to its base variant, this version delivers significant gains in knowledge coverage, long-context reasoning, coding benchmarks, and alignment with open-ended tasks. It is particularly strong on multilingual understanding, math reasoning (e.g., AIME, HMMT), and alignment evaluations like Arena-Hard and WritingBench. Qwen3 235B A22B Instruct 2507 runs on the Geek Hub API (OpenAI-compatible). Model id: qwen/qwen3-235b-a22b-2507.
Is Qwen3 235B A22B Instruct 2507 free?
No. Input is $0.0981 per 1M tokens and output is $0.5995 per 1M tokens on Geek Hub (markup included).
What is the context length of Qwen3 235B A22B Instruct 2507?
Qwen3 235B A22B Instruct 2507 has a 262k tokens context window. It supports up to 16,384 completion tokens.
Does Qwen3 235B A22B Instruct 2507 support tool calling and structured outputs?
Qwen3 235B A22B Instruct 2507 accepts tools and tool_choice for function calling. It also supports structured outputs via a JSON schema in response_format.
When was Qwen3 235B A22B Instruct 2507 released?
Qwen3 235B A22B Instruct 2507 was released on 2025-07-21.

More models from Qwen