Meet Inkling: Thinking Machines Lab's new customizable model
Blog post from Baseten
Inkling, developed by Thinking Machines Lab, is a multimodal, autoregressive transformer model with 975 billion parameters, designed to process text, images, and audio inputs and generate text outputs. It features a mixture-of-experts architecture, activating only a portion of its parameters per task, which balances performance with efficiency. Inkling is supported from day one on the Baseten Platform, where it can be accessed through Model APIs and Dedicated Inference deployments. This model, built for breadth and optimized for developers creating AI-powered applications, offers open weights for customization and deployment. Despite its substantial infrastructure requirements, Baseten's autoscaling and multi-cloud capacity management enable reliable and scalable deployment. Inkling aims to extend human will and judgment, making it ideal for tasks ranging from coding assistance to chatbots and retrieval-augmented generation systems.
| Trend | Post Mentions | Total Month Mentions | Posts | Companies | MoM |
|---|---|---|---|---|---|
| AI Model Fine-tuning | 2 | 402 | 99 | 46 | -46% |
| RAG | 1 | 619 | 146 | 64 | -38% |
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.