Thinking Machines Releases Native Multimodal Interaction Model for Real-Time Human-AI Collaboration
Thinking Machines发布原生多模态"交互模型",实现实时人机协作
Thinking Machines released an interaction model that natively receives audio, video, and text input, processes foreground interaction at 200-millisecond intervals, and uses a background reasoning model for long-horizon planning and tool calls.
Why it matters: HKR-H/K/R all pass: this is more than a model notice, with a two-layer foreground/background interaction design. Pricing, access scope, and benchmarks are missing, so it sits at the lower end of 85-94.