Qwen3 30B A3B Instruct 2507

Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference. It operates in non-thinking mode and is designed for high-quality instruction following, multilingual understanding, and agentic tool use. Post-trained on instruction data, it demonstrates competitive performance across reasoning (AIME, ZebraLogic), coding (MultiPL-E, LiveCodeBench), and alignment (IFEval, WritingBench) benchmarks. It outperforms its non-instruct variant on subjective and open-ended tasks while retaining strong factual and coding performance.

ChatQwen262k tokens$0.0525 / $0.2104 · 1M

qwen/qwen3-30b-a3b-instruct-2507

Context
262k tokens
Completion cap
32,000
Tools
Yes
JSON
Yes
Released
2025-07-29

Call it from Geek Hub

Same OpenAI SDK. Change the base URL and the model id.

curl https://api.geekhub.mx/v1/chat/completions \
  -H "Authorization: Bearer $GEEKHUB_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "qwen/qwen3-30b-a3b-instruct-2507",
    "messages": [{"role": "user", "content": "Hola"}]
  }'
Get an API key

Frequently asked questions

What is Qwen3 30B A3B Instruct 2507?
Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference. It operates in non-thinking mode and is designed for high-quality instruction following, multilingual understanding, and agentic tool use. Post-trained on instruction data, it demonstrates competitive performance across reasoning (AIME, ZebraLogic), coding (MultiPL-E, LiveCodeBench), and alignment (IFEval, WritingBench) benchmarks. It outperforms its non-instruct variant on subjective and open-ended tasks while retaining strong factual and coding performance. Qwen3 30B A3B Instruct 2507 runs on the Geek Hub API (OpenAI-compatible). Model id: qwen/qwen3-30b-a3b-instruct-2507.
Is Qwen3 30B A3B Instruct 2507 free?
No. Input is $0.0525 per 1M tokens and output is $0.2104 per 1M tokens on Geek Hub (markup included).
What is the context length of Qwen3 30B A3B Instruct 2507?
Qwen3 30B A3B Instruct 2507 has a 262k tokens context window. It supports up to 32,000 completion tokens.
Does Qwen3 30B A3B Instruct 2507 support tool calling and structured outputs?
Qwen3 30B A3B Instruct 2507 accepts tools and tool_choice for function calling. It also supports structured outputs via a JSON schema in response_format.
When was Qwen3 30B A3B Instruct 2507 released?
Qwen3 30B A3B Instruct 2507 was released on 2025-07-29.

More models from Qwen