Skip to content
Product Hunt · AI

Thinking Machines releases Inkling, a 975B open-weights multimodal model built for fine-tuning

Inkling

Thinking Machines launched Inkling on Product Hunt, a 975B MoE open-weights model with 41B active parameters and 1M context window. It handles text, images, and audio natively, with controllable reasoning effort, under Apache 2.0. The companion Tinker API handles LoRA fine-tuning without infrastructure overhead, aimed at researchers who want full control over data and algorithms. The post does not disclose benchmark scores or pricing.

Why it matters: Thinking Machines dropped a 975B MoE open model with only 41B active params — inference cost should be low. 1M context + native multimodal is a strong spec sheet. No benchmarks or real-world latency numbers yet, so holding below 85.

Read the original ↗Export Markdown