Switch language한국어
Back to the list

Thinking Machines Lab Releases Inkling-Small: A 276B Total, 12B Active Open Weights Multimodal MoE Model

TL;DR AI

Key summary

2 min read
  1. Thinking Machines Lab has released Inkling-Small, an Apache 2.0 open-weights multimodal MoE model.

  2. It handles text, images, and audio, with 276B total parameters and a 1M-token context window.

  3. Trained on NVIDIA GB300 NVL72 systems, it can also run when quantized on a single B300 or dual H200 setup.

  4. Benchmark results show gains on several reasoning and coding tasks, though some knowledge and epistemic measures lag.

  5. The release expands practical self-hosted options for startups and enterprises needing private multimodal AI.

Read the original