Inkling (batch)

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems, retrieval-augmented generation, instruction following, and multilingual conversational applications. Its native image and audio understanding supports multimodal analysis alongside text.

Speech-to-textThinkingmachines524k tokens$1.0901 / min

thinkingmachines/inkling:batch

Contexto
524k tokens
Máx. completion
Tools
JSON
No
Lanzamiento
2026-07-17

Llámalo desde Geek Hub

El mismo SDK de OpenAI. Cambia el base URL y el id del modelo.

curl https://api.geekhub.mx/v1/chat/completions \
  -H "Authorization: Bearer $GEEKHUB_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "thinkingmachines/inkling:batch",
    "messages": [{"role": "user", "content": "Hola"}]
  }'
Consigue tu API key

Preguntas frecuentes

¿Qué es Inkling (batch)?
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems, retrieval-augmented generation, instruction following, and multilingual conversational applications. Its native image and audio understanding supports multimodal analysis alongside text. Inkling (batch) corre en el API de Geek Hub (compatible con OpenAI). Id: thinkingmachines/inkling:batch.
¿Inkling (batch) es gratis?
No. $1.0901 por minuto en Geek Hub.
¿Cuál es el contexto de Inkling (batch)?
Inkling (batch) tiene una ventana de 524k tokens.
¿Cuándo se lanzó Inkling (batch)?
Inkling (batch) se lanzó el 17 de julio de 2026.

Más modelos de Thinkingmachines