Thinking Machines Inkling: What Developers Need to Know
Blog post from Eden AI
Inkling is a 975 billion parameter Mixture-of-Experts (MoE) model developed by Thinking Machines Lab, led by Mira Murati, and released in July 2026 under the Apache 2.0 license. It stands out by activating only 41 billion parameters per request, balancing inference cost with extensive knowledge capacity, and allows users to adjust the "thinking" effort for varied tasks, from simple lookups to complex analyses. Unlike many 2026 models focused on single benchmarks, Inkling is designed for broad competence across text, code, multimodal, and audio tasks. It achieves top performance as a US-origin open-weight model in benchmarks like AIME 2026 and GPQA Diamond and offers benefits like self-hosting, fine-tuning, and provider diversification, reducing dependency on proprietary models. Accessible through platforms like Eden AI, Inkling can be integrated alongside models like Claude and GPT without changing code, offering flexibility in AI strategies. While it requires significant GPU resources for self-hosting, managed hosting options are available, making it a versatile tool for developers seeking competitive, open-weight AI solutions.
No tracked trend matches for this post yet.
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.