Thinking Machines Lab's sparse mixture-of-experts flagship: 975B total parameters with 41B active per token across 66 layers, and a 1M-token context window. Built for reasoning, coding, tool use, and multimodal work.
Thinking Machines Lab's sparse mixture-of-experts flagship: 975B total parameters with 41B active per token across 66 layers, and a 1M-token context window. Built for reasoning, coding, tool use, and multimodal work.