← Back to the wire

Thinking Machines Lab Releases Inkling: A 975B-Parameter Open-Weights Multimodal MoE With 41B Active Parameters And Controllable Thinking Effort

AnnouncementModelJul 15, 2026

Thinking Machines Lab released Inkling, a 975B-parameter Mixture-of-Experts model with open weights and a 1M-token context window. The model activates 41B parameters per token and was pretrained on 45 trillion tokens across text, images, audio, and video. Its MoE architecture largely follows DeepSeek-V3. A smaller variant, Inkling-Small, matches the larger model on many benchmarks and will release after testing. Inkling supports fine-tuning on Tinker and is deployable via multiple runtimes.

Receipt № 6931 source · awaiting confirmation ◐

Evidence

1source· awaiting independent confirmation

No score is assigned. Sources and their independence are shown in the citation chain below.

Citation chain · 1 source

Thinking Machines LabCompanyInklingModelInkling-SmallModelDeepSeek-V3Model
Canonical: https://www.marktechpost.com/2026/07/15/thinking-machines-lab-releases-inkling-a-975b-parameter-open-weights-multimodal-moe-with-41b-active-parameters-and-controllable-thinking-effort/