← Back to the wire

Thinking Machines Lab Releases Inkling: A 975B-Parameter Open-Weights Multimodal MoE With 41B Active Parameters And Controllable Thinking Effort

AnnouncementModelJul 15, 2026

Thinking Machines Lab released Inkling, a 975B-parameter Mixture-of-Experts model with open weights and a 1M-token context window. The model activates 41B parameters per token and was pretrained on 45 trillion tokens across text, images, audio, and video. Its MoE architecture largely follows DeepSeek-V3. A smaller variant, Inkling-Small, matches the larger model on many benchmarks and will release after testing. Inkling supports fine-tuning on Tinker and is deployable via multiple runtimes.

Evidence

1source· awaiting independent confirmation

No score is assigned. Sources and their independence are shown in the citation chain below.

Citation chain · 1 source

Thinking Machines LabCompanyInklingModelInkling-SmallModelDeepSeek-V3Model
Canonical: https://www.marktechpost.com/2026/07/15/thinking-machines-lab-releases-inkling-a-975b-parameter-open-weights-multimodal-moe-with-41b-active-parameters-and-controllable-thinking-effort/