Qwen3.6 35B A3B

Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated DeltaNet linear attention with standard gated attention layers, enabling efficient inference at a fraction of the compute cost. The model supports a 262K token native context window (extensible to 1M via YaRN) and accepts text, image, and video inputs. It includes integrated thinking mode with reasoning traces preserved across multi-turn conversations, function calling, and structured output. Released under the Apache 2.0 license.

ChatQwen262k tokens$0.1526 / $1.0901 · 1M

qwen/qwen3.6-35b-a3b

Contexto
262k tokens
Máx. completion
262,144
Tools
JSON
Lanzamiento
2026-04-27

Llámalo desde Geek Hub

El mismo SDK de OpenAI. Cambia el base URL y el id del modelo.

curl https://api.geekhub.mx/v1/chat/completions \
  -H "Authorization: Bearer $GEEKHUB_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "qwen/qwen3.6-35b-a3b",
    "messages": [{"role": "user", "content": "Hola"}]
  }'
Consigue tu API key

Preguntas frecuentes

¿Qué es Qwen3.6 35B A3B?
Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated DeltaNet linear attention with standard gated attention layers, enabling efficient inference at a fraction of the compute cost. The model supports a 262K token native context window (extensible to 1M via YaRN) and accepts text, image, and video inputs. It includes integrated thinking mode with reasoning traces preserved across multi-turn conversations, function calling, and structured output. Released under the Apache 2.0 license. Qwen3.6 35B A3B corre en el API de Geek Hub (compatible con OpenAI). Id: qwen/qwen3.6-35b-a3b.
¿Qwen3.6 35B A3B es gratis?
No. El input cuesta $0.1526 / 1M tokens y el output $1.0901 / 1M tokens en Geek Hub (markup incluido).
¿Cuál es el contexto de Qwen3.6 35B A3B?
Qwen3.6 35B A3B tiene una ventana de 262k tokens. Soporta hasta 262,144 tokens de completion.
¿Qwen3.6 35B A3B soporta tool calling y structured outputs?
Qwen3.6 35B A3B acepta tools y tool_choice para function calling. También soporta structured outputs con un JSON schema en response_format.
¿Cuándo se lanzó Qwen3.6 35B A3B?
Qwen3.6 35B A3B se lanzó el 27 de abril de 2026.

Más modelos de Qwen