# Thinking Machines: Inkling

Provider: thinkingmachines  
Category: llm  
Model ID: `thinkingmachines/inkling`

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

## Specs

- Context length: 524288 tokens
- Max output: 471859 tokens
- Modalities: text, image, audio
- Released: 2026-07-17

## Pricing

- Input: $1.00 per million tokens
- Output: $4.05 per million tokens

## Verdict

Inkling offers a massive 524K token context window at aggressive pricing — $1 input / $4.05 output per million tokens — making it a strong candidate for long-document workflows where cost matters. Multimodal support (text, image, audio) adds flexibility for mixed-media tasks. The catch: no public benchmarks yet, so you're flying blind on reasoning quality and accuracy relative to established models. Reach for this when context length and budget are your top constraints and you can validate outputs internally.

---
Last verified: 2026-09-04T02:00:56.012Z  
Canonical URL: https://switchy.build/directory/models/inkling