
Inkling is a 975B-param (41B active) multimodal released under Apache 2.0 with 1M-token context and same-day fine-tuning on Tinker, giving builders a permissively licensed U.S. open base model to customize rather than a locked-down API. Its novel architecture (relative positional bias, scaled short-conv layers, dual shared-expert MoE) makes it a distinct alternative to Llama, DeepSeek, and GLM/Kimi for teams needing open weights they can adapt.
“Thinky only seems to come up for air once every few months; most recently with Interaction models - but each time they do they impress, showing both taste and depth.”
“Our model, called Inkling, is a Mixture-of-Experts transformer with 975B total parameters, 41B active. It supports a context window of up to 1M tokens. It was pretrained on 45 trillion tokens of text, images, audio and video.”
“Thinking Machines Lab launched Inkling, its first fully released open-weights foundation model family entry, positioning it as a customizable multimodal base model rather than a benchmark-maxed flagship.”
Checking sign-in…
Loading comments…