Home / Companies / Baseten / Blog / Post Details
Content Deep Dive

Meet Inkling: Thinking Machines Lab's new customizable model

Blog post from Baseten

Post Details
Company
Date Published
Author
Albert Lee
Word Count
736
Company Posts That Month
7
Language
English
Hacker News Points
-
Post removed?
No
Summary

Inkling, developed by Thinking Machines Lab, is a multimodal, autoregressive transformer model with 975 billion parameters, designed to process text, images, and audio inputs and generate text outputs. It features a mixture-of-experts architecture, activating only a portion of its parameters per task, which balances performance with efficiency. Inkling is supported from day one on the Baseten Platform, where it can be accessed through Model APIs and Dedicated Inference deployments. This model, built for breadth and optimized for developers creating AI-powered applications, offers open weights for customization and deployment. Despite its substantial infrastructure requirements, Baseten's autoscaling and multi-cloud capacity management enable reliable and scalable deployment. Inkling aims to extend human will and judgment, making it ideal for tasks ranging from coding assistance to chatbots and retrieval-augmented generation systems.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
AI Model Fine-tuning 2 402 99 46 -46%
RAG 1 619 146 64 -38%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.