415.tech
AI & tech, from the frontlines of Silicon Valley
Inkling-Small: 276B sparse MoE, 12B active, Apache 2.0

Inkling-Small: 276B sparse MoE, 12B active, Apache 2.0

Thinking Machines released Inkling-Small under Apache 2.0 — a 276B sparse mixture-of-experts model with 12B active parameters that scores 40 on the Artificial Analysis Intelligence Index, one point behind the 975B Inkling from two weeks ago. Near-frontier reasoning (95.5% AIME 2026, 80.2% SWE-bench Verified) now runs on open weights with 12B active parameters, and 25 quantized variants ship across llama.cpp, Ollama, LM Studio and Jan.

Source: huggingface.co

Post on XEmail

Inkling-Small is a general-purpose multimodal model that accepts text, image and audio inputs and generates text outputs.

Thinking Machines

Why this matters

  • → Frontier reasoning (95.5% AIME, 80.2% SWE-bench) now runs on 12B active parameters under open weights.
  • → Apache 2.0 license enables research, fine-tuning, and third-party integration without restrictions.
  • → 25 quantized variants ship across llama.cpp, Ollama, LM Studio, Jan for immediate local deployment.
Open frontier reasoning