| DeepSeek V4.1 Flash is live on Baseten Model APIs, day 0:
- Smarter, faster, and more efficient than DeepSeek v4 Pro … |
@baseten |
Company |
Original |
2026-09-10 |
1,390 |
23 |
2 |
1 |
0 |
4 |
| Thanks Chris Metinko and @axios for covering the story!
https://t.co/3wV6cjOh3F |
@baseten |
Company |
Quote |
2026-09-10 |
1,337 |
8 |
0 |
0 |
0 |
2 |
| https://t.co/JZXABlmahG |
@baseten |
Company |
Original |
2026-09-10 |
47,925 |
113 |
16 |
9 |
22 |
34 |
| A conversation about open-source models in production, plus a Mario Kart tournament.
Oct 7 at @Techweek_ in SF by a1… |
@baseten |
Company |
Quote |
2026-09-09 |
2,213 |
20 |
1 |
3 |
0 |
2 |
| Speaker diarization is one of the hardest problems in voice AI and the best-performing models can be too slow and expen… |
@baseten |
Company |
Original |
2026-09-09 |
131,170 |
15 |
1 |
3 |
0 |
7 |
| We're excited to bring Mercury 2.5 to Baseten:
- Sub-200ms p50 latency, great for real-time voice use cases
- 1,100+… |
@baseten |
Company |
Quote |
2026-09-09 |
5,774 |
64 |
3 |
3 |
1 |
14 |
| We're proud to define the quality-latency Pareto frontier for STT in Coval's benchmarks.
Voice AI is an inference prob… |
@baseten |
Company |
Original |
2026-09-08 |
1,422 |
17 |
4 |
3 |
1 |
3 |
| AI labs and frontier research teams run RL managed rollouts on Baseten. They publish new weights, and we pick them up a… |
@baseten |
Company |
Quote |
2026-09-08 |
1,753 |
29 |
1 |
1 |
0 |
6 |
| We partnered with @harvey to train agents for M&A diligence and showed that model-harness co-optimization can improve a… |
@baseten |
Company |
Quote |
2026-09-08 |
1,548 |
34 |
3 |
1 |
0 |
10 |
| RT @philipkiely: Tomorrow at 11 AM PT, I'll be live on air with Declan Jackson of Artificial Analysis to discuss GLM-5.… |
@baseten |
Company |
Repost |
2026-09-07 |
110 |
0 |
7 |
0 |
0 |
0 |
| GLM-5.3 is now live with vision, only on Baseten Model APIs.
Generate and convert images to code thanks to our post-tr… |
@baseten |
Company |
Original |
2026-09-04 |
2,585 |
50 |
0 |
2 |
0 |
8 |
| RT @DannieHerz: Notion helps us think. We help them listen. Notion's AI Meeting Notes serve countless teams (ours inclu… |
@baseten |
Company |
Repost |
2026-09-04 |
111 |
0 |
2 |
0 |
0 |
0 |
| Notion's AI Meeting Notes run on Baseten, and Baseten's knowledge base runs on Notion. Our teams have been working toge… |
@baseten |
Company |
Original |
2026-09-04 |
25,159 |
98 |
10 |
8 |
7 |
28 |
| "I've been through the lifecycle of building a product that people love. And I should be using that information to make… |
@baseten |
Company |
Quote |
2026-09-04 |
2,214 |
22 |
0 |
0 |
0 |
4 |
| We're proud to support companies like OpenEvidence that are pushing the frontier of AI in their specialized domains.
T… |
@baseten |
Company |
Quote |
2026-09-04 |
5,734 |
65 |
0 |
2 |
0 |
14 |
| RT @DannieHerz: Baseten was built on the belief that the health of the open-source AI ecosystem matters. Base Labs is o… |
@baseten |
Company |
Repost |
2026-09-03 |
129 |
0 |
2 |
0 |
0 |
0 |
| RT @tuhinone: We're very excited to launch Base Labs to democratize open-source AI research.
We'll push the frontier i… |
@baseten |
Company |
Repost |
2026-09-03 |
107 |
0 |
11 |
0 |
0 |
0 |
| RT @amiruci: We have clearly benefited a ton from the ecosystem of ever-improving open models and the published researc… |
@baseten |
Company |
Repost |
2026-09-03 |
103 |
0 |
5 |
0 |
0 |
0 |
| We’re excited to keep pushing the open-source AI ecosystem forward. |
@baseten |
Company |
Quote |
2026-09-03 |
14,696 |
165 |
8 |
2 |
0 |
51 |
| We released a remote MCP server and skill so coding agents can operate the Baseten platform faster and more efficiently… |
@baseten |
Company |
Quote |
2026-09-03 |
4,310 |
61 |
2 |
2 |
1 |
13 |
| Today, we're releasing GLM-5.3 Fast: one of the most intelligent open-weight models ever at an even higher TPS.
Design… |
@baseten |
Company |
Original |
2026-09-03 |
17,155 |
210 |
11 |
13 |
1 |
59 |
| RT @LangChain: Happening today!
An evening of technical talks, meeting builders, and end-of-summer vibes with @PrimeIn… |
@baseten |
Company |
Repost |
2026-09-02 |
95 |
0 |
3 |
0 |
0 |
0 |
| RT @philipkiely: https://t.co/ESeIBZP3zR |
@baseten |
Company |
Repost |
2026-09-02 |
116 |
0 |
43 |
0 |
0 |
0 |
| We're excited to sponsor PACT (Pressure-Applied Compliance Testing), a new benchmark for enterprise assistant rule-foll… |
@baseten |
Company |
Quote |
2026-09-01 |
3,314 |
43 |
1 |
5 |
0 |
8 |
| We're honored to be named a 2026 IA40 winner!
The IA40 recognizes the 40 most important private companies in applied A… |
@baseten |
Company |
Original |
2026-09-01 |
2,771 |
73 |
2 |
4 |
0 |
6 |
| RT @philipkiely: Teach your agents inference.
The full text of Inference Engineering is now available to LLMs everywh… |
@baseten |
Company |
Repost |
2026-08-31 |
120 |
0 |
60 |
0 |
0 |
0 |
| RT @LangChain: 🌉 Happening 9/2 in SF!
Join LangChain, @baseten, and @primeintellect for a dive into continual learning… |
@baseten |
Company |
Repost |
2026-08-31 |
109 |
0 |
3 |
0 |
0 |
0 |
| "If you're not taking care of yourself, you're not writing good code."
We were honored to host @bryan_johnson and @sar… |
@baseten |
Company |
Quote |
2026-08-31 |
3,603 |
29 |
0 |
2 |
0 |
2 |
| RT @amiruci: Even hackernews is raving about this model! |
@baseten |
Company |
Repost |
2026-08-29 |
118 |
0 |
1 |
0 |
0 |
0 |
| Our kernel engineers built an agentic framework to automatically find, build, validate, and ship optimized kernels into… |
@baseten |
Company |
Quote |
2026-08-29 |
13,406 |
140 |
5 |
3 |
1 |
91 |
| RT @thealexker: https://t.co/YPKWwqEQ3F |
@baseten |
Company |
Repost |
2026-08-28 |
121 |
0 |
50 |
0 |
0 |
0 |
| RT @tuhinone: GLM 5.3 is amazing, the rate of progress of open-weights is clearly continuing to accelerate, and we'll h… |
@baseten |
Company |
Repost |
2026-08-28 |
108 |
0 |
5 |
0 |
0 |
0 |
| GLM-5.3 is live on Baseten Model APIs, day 0.
- The smartest open-weight model at 743B params
- 1M token context
- US … |
@baseten |
Company |
Original |
2026-08-28 |
15,660 |
64 |
5 |
4 |
7 |
18 |
| RT @philipkiely: Should you use GLM-5.3 or GLM-5.3-Flash (Ox Alpha)?
- 753B/40B vs 320B/18B params
- 60 vs 57 AA score… |
@baseten |
Company |
Repost |
2026-08-28 |
107 |
0 |
4 |
0 |
0 |
0 |
| Post-train GLM-5.3 and GLM-5.3-Flash on Baseten Loops.
Inference + training support on day 0.
https://t.co/mbIxzhP7l… |
@baseten |
Company |
Original |
2026-08-28 |
7,635 |
81 |
5 |
6 |
3 |
28 |
| We're proud to be the fastest inference provider on Artificial Analysis, OpenRouter, and Hugging Face for GLM-5.3-Flash… |
@baseten |
Company |
Original |
2026-08-27 |
21,510 |
210 |
9 |
13 |
2 |
45 |
| 👀 |
@baseten |
Company |
Quote |
2026-08-27 |
5,885 |
77 |
0 |
2 |
0 |
4 |
| GLM-5.3-Flash is live, day 0, on Baseten Model APIs.
- Useful for general intelligence + agentic coding
- More intelli… |
@baseten |
Company |
Original |
2026-08-26 |
9,162 |
123 |
4 |
9 |
2 |
21 |
| RT @thealexker: glm-5.3-flash is a story of efficiency. top gems:
> outperforms glm-5.2 across domains at 1/10 th… |
@baseten |
Company |
Repost |
2026-08-26 |
116 |
0 |
5 |
0 |
0 |
0 |
| RT @amiruci: GLM 5.3 Flash post-training is supported in Baseten Loops on day 0. And there's good reason for it. DM @on… |
@baseten |
Company |
Repost |
2026-08-26 |
102 |
0 |
6 |
0 |
0 |
0 |
| Our engineers are cooking. Stay tuned. |
@baseten |
Company |
Quote |
2026-08-26 |
9,387 |
193 |
3 |
8 |
2 |
4 |
| RT @oneill_c: If you want to RL (or sft/opd) glm-5.3-flash, DM me to get access to it on @baseten loops. We're very exc… |
@baseten |
Company |
Repost |
2026-08-26 |
106 |
0 |
17 |
0 |
0 |
0 |
| RT @gambhir_amit: https://t.co/WCLrBFAHQX |
@baseten |
Company |
Repost |
2026-08-25 |
117 |
0 |
4 |
0 |
0 |
0 |
| RT @bryan_johnson: Doing a private longevity event with @baseten in SF on Wednesday.
> biological age testing
>… |
@baseten |
Company |
Repost |
2026-08-24 |
100 |
0 |
11 |
0 |
0 |
0 |
| RT @saranormous: Doing a private conversation with @bryan_johnson on longevity + tech on Aug 26 (with bioage testing on… |
@baseten |
Company |
Repost |
2026-08-21 |
92 |
0 |
8 |
0 |
0 |
0 |
| Some questions demand the most accurate answers. @youdotcom built its Answer API to help users get grounded, cited answ… |
@baseten |
Company |
Original |
2026-08-21 |
6,623 |
33 |
8 |
3 |
2 |
9 |
| GLM-5.2 now supports vision in production, only on Baseten.
Ingest images and convert them to code, thanks to our post… |
@baseten |
Company |
Original |
2026-08-21 |
25,654 |
65 |
0 |
5 |
4 |
17 |
| More evidence of a many-model future.
Congrats @OpenRouter! |
@baseten |
Company |
Quote |
2026-08-20 |
2,690 |
27 |
2 |
0 |
0 |
2 |
| RT @hwchase17: top tier webinar tmrw!
@Vtrivedy10 (@LangChain) @willcb (@PrimeIntellect) @AEllisBloor (@baseten) and I… |
@baseten |
Company |
Repost |
2026-08-20 |
107 |
0 |
14 |
0 |
0 |
0 |
| Last week, the entire Baseten team came together in Chicago for our biannual company offsite.
Since the beginning, we'… |
@baseten |
Company |
Original |
2026-08-19 |
4,114 |
77 |
9 |
1 |
3 |
11 |
| RT @amiruci: DeepSeek-V4-Pro-0813 came out a few days ago. Baseten is the fastest provider according to @OpenRouter, @A… |
@baseten |
Company |
Repost |
2026-08-18 |
97 |
0 |
3 |
0 |
0 |
0 |
| RT @p0: Next Thursday, August 20th, @baseten is hosting @paraga for a live discussion about how agents are changing the… |
@baseten |
Company |
Repost |
2026-08-18 |
98 |
0 |
1 |
0 |
0 |
0 |
| We're proud to be the fastest inference provider on Artificial Analysis for DeepSeek V4 Pro 0813, at 147 TPS.
Our eng… |
@baseten |
Company |
Original |
2026-08-18 |
11,792 |
69 |
1 |
3 |
1 |
20 |
| Inference Engineering is free for everyone to read:
https://t.co/sF1pX4LSQB |
@baseten |
Company |
Quote |
2026-08-18 |
167,562 |
1,853 |
125 |
26 |
11 |
2,514 |
| Qwen3.8-27B ranks #7 on Artificial Analysis' Agentic Index at only 27B parameters, a fraction of the size (and cost) of… |
@baseten |
Company |
Quote |
2026-08-18 |
2,654 |
22 |
1 |
2 |
0 |
7 |
| We are big fans of @NotionHQ.
Proud to power the inference behind the product our own team can't work without. |
@baseten |
Company |
Quote |
2026-08-18 |
4,963 |
20 |
0 |
2 |
0 |
6 |
| RT @julien_c: Own your models https://t.co/qOVTxeflTk |
@baseten |
Company |
Repost |
2026-08-17 |
112 |
0 |
10 |
0 |
0 |
0 |
| Run SFT or RL on Qwen3.8-27B, the most intelligent small model to date.
Day 0 support on the Baseten Loops SDK: https:… |
@baseten |
Company |
Original |
2026-08-15 |
4,319 |
40 |
5 |
0 |
1 |
11 |
| DeepSeek V4 Pro 0813 is live day 0 on Baseten Model APIs.
- Frontier-level Artificial Analysis Intelligence score
- Ze… |
@baseten |
Company |
Original |
2026-08-14 |
910,196 |
73 |
0 |
5 |
1 |
15 |
| RT @philipkiely: https://t.co/fg4jUePnls |
@baseten |
Company |
Repost |
2026-08-14 |
102 |
0 |
7 |
0 |
0 |
0 |
| RT @DannieHerz: Bryan Johnson is truly an obsessive. On August 26th, he'll join us at our HQ for dinner, a private conv… |
@baseten |
Company |
Repost |
2026-08-13 |
117 |
0 |
2 |
0 |
0 |
0 |
| At Baseten, we have a thesis: obsessives move the world forward. Few people embody that more completely than @bryan_joh… |
@baseten |
Company |
Original |
2026-08-13 |
132,231 |
134 |
5 |
9 |
5 |
27 |
| Coming to Baseten Model APIs today. https://t.co/kUNeDbdqx1 |
@baseten |
Company |
Original |
2026-08-13 |
2,752 |
46 |
0 |
2 |
1 |
10 |
| We are strong believers in a healthy frontier ecosystem and are dedicated to building the inference and training to sup… |
@baseten |
Company |
Quote |
2026-08-11 |
3,916 |
22 |
0 |
2 |
0 |
3 |
| A clear example of how fine-tuning a model on your unique data can yield a much cheaper, specialized model.
Thanks to … |
@baseten |
Company |
Quote |
2026-08-11 |
2,917 |
18 |
1 |
1 |
0 |
9 |
| NVIDIA Nemotron 3.5 Lightning is live on Baseten day 0!
This is the fastest open model in its class, built for long-ru… |
@baseten |
Company |
Original |
2026-08-11 |
3,079 |
50 |
1 |
3 |
1 |
8 |
| Muse Glimmer, Meta's new open-weight agentic model, is available on Baseten day 0!
- 30B parameters, including a 1.8B … |
@baseten |
Company |
Original |
2026-08-11 |
3,216 |
44 |
1 |
4 |
0 |
4 |
| We're proud to be @arcprize's inference partner as they test the impact of models like DeepSeek V4 Flash and Kimi K3.
… |
@baseten |
Company |
Quote |
2026-08-07 |
3,878 |
35 |
2 |
1 |
0 |
5 |
| Baseten is now an official inference provider on @huggingface 🤗
Run Kimi K3, DeepSeek V4 Flash, and GLM-5.2 on Baseten… |
@baseten |
Company |
Original |
2026-08-06 |
50,640 |
192 |
15 |
11 |
5 |
26 |
| We often see smaller, specialized models outperform the frontier on specific tasks. Fine-tuning DeepSeek V4 Flash achie… |
@baseten |
Company |
Original |
2026-08-05 |
3,856 |
53 |
6 |
2 |
1 |
19 |
| "A really strong open-weight ecosystem will lead to an even stronger closed-weight ecosystem. It's necessary for both p… |
@baseten |
Company |
Original |
2026-08-05 |
2,821 |
37 |
3 |
7 |
0 |
11 |
| Open-source models have reached a new level of intelligence; our own engineers use models like GLM-5.2 for tasks like k… |
@baseten |
Company |
Quote |
2026-08-04 |
5,016 |
51 |
1 |
2 |
0 |
12 |
| One of our engineers asked @poolsideai's Laguna S 2.1 to transform a 715-file C++ game from neon cyberpunk into an Anci… |
@baseten |
Company |
Original |
2026-08-03 |
3,541 |
33 |
2 |
4 |
0 |
4 |
| Yes, we will have Qwen3.8-Max. |
@baseten |
Company |
Quote |
2026-08-03 |
5,684 |
109 |
3 |
8 |
0 |
1 |
| Try DeepSeek V4 Flash on our Model APIs today.
- 80–98% cheaper than other frontier models
- Comparable intelligence
-… |
@baseten |
Company |
Original |
2026-08-01 |
23,232 |
139 |
5 |
9 |
4 |
21 |
| Qwen3-TTS supervised fine-tuning (SFT) for voice cloning is now available on Baseten Training.
We see 16% faster time … |
@baseten |
Company |
Quote |
2026-07-31 |
3,841 |
20 |
2 |
1 |
0 |
8 |
| Thinking Machines Lab’s new Inkling-Small is now available on our Model APIs!
At 276B parameters (12B active), Inkling… |
@baseten |
Company |
Original |
2026-07-30 |
2,807 |
45 |
2 |
3 |
0 |
4 |
| The easiest way to run open (Kimi K3, GLM-5.2) and closed models in your harness: Baseten Switch! |
@baseten |
Company |
Quote |
2026-07-30 |
4,056 |
56 |
3 |
1 |
0 |
17 |
| Today, we're announcing Baseten for Model Labs.
We believe the AI landscape will be made up of a diverse ecosystem of … |
@baseten |
Company |
Original |
2026-07-29 |
49,081 |
219 |
25 |
11 |
19 |
74 |
| Thrilled to be partnering with companies like @harvey that value research as much as we do. |
@baseten |
Company |
Quote |
2026-07-29 |
4,188 |
37 |
1 |
2 |
0 |
4 |
| We shipped 18x faster tokenization for Kimi K3 on day 0.
Michael Feil is such a cracked engineer, we even let him nam… |
@baseten |
Company |
Quote |
2026-07-29 |
7,931 |
73 |
1 |
1 |
0 |
18 |
| Proud to support the Cursor team with day-0 support for Kimi K3! |
@baseten |
Company |
Quote |
2026-07-28 |
13,898 |
129 |
1 |
1 |
0 |
10 |
| Happening today! Link to register in thread. |
@baseten |
Company |
Quote |
2026-07-28 |
3,805 |
10 |
1 |
1 |
0 |
1 |
| "The central change is not scale alone. Each architectural step changes what the model stores, how it updates that stat… |
@baseten |
Company |
Quote |
2026-07-28 |
7,191 |
66 |
4 |
1 |
1 |
19 |
| Tomorrow we're hosting a virtual discussion and Q&A on how to evaluate Kimi K3 for enterprise workflows.
Join our … |
@baseten |
Company |
Original |
2026-07-27 |
11,658 |
29 |
1 |
2 |
3 |
6 |
| Serving a 2.8T parameter frontier model like Kimi K3 at scale, on day 0, takes a village. |
@baseten |
Company |
Quote |
2026-07-27 |
5,334 |
90 |
1 |
4 |
0 |
11 |
| Kimi K3 is now live on our Model APIs, day 0. https://t.co/Itz5SDiEYz |
@baseten |
Company |
Original |
2026-07-27 |
684,174 |
612 |
48 |
33 |
28 |
111 |
| Jeanne DeWitt Grosser, the COO of Vercel, took a 10-person SDR team down to one person and an AI agent.
At Stripe, she… |
@baseten |
Company |
Original |
2026-07-24 |
3,197 |
15 |
1 |
2 |
1 |
5 |
| Kimi K3 drops Monday, and our team is already cooking.
Get on the Kimi K3 drop list and get $25 of free Model API cred… |
@baseten |
Company |
Original |
2026-07-24 |
56,897 |
811 |
26 |
16 |
6 |
201 |
| You can fine-tune GLM-5.2 on the Baseten Loops SDK and easily deploy your checkpoint to a production inference endpoint… |
@baseten |
Company |
Original |
2026-07-24 |
4,424 |
88 |
6 |
4 |
0 |
16 |
| Today, we're introducing GLM-5.2 Fast: our GLM-5.2 Model API designed for the most demanding real-time use cases.
The … |
@baseten |
Company |
Original |
2026-07-23 |
114,682 |
538 |
29 |
27 |
14 |
169 |
| We added vision support to GLM-5.2 with just a 2-layer MLP, <50M parameters.
Shout out to @part_harry_ for his wor… |
@baseten |
Company |
Quote |
2026-07-23 |
5,669 |
70 |
4 |
0 |
0 |
9 |
| In case you missed it:
@part_harry_ didn't just add vision support to GLM-5.2, he also made it open source.
We love… |
@baseten |
Company |
Quote |
2026-07-23 |
3,196 |
33 |
0 |
2 |
0 |
4 |
| The future will be built on many models, open and closed, general and specialized at different levels of granularity.
… |
@baseten |
Company |
Quote |
2026-07-22 |
1,949 |
19 |
0 |
1 |
0 |
2 |
| The American open-weight ecosystem is growing. We're thrilled to support Poolside as they push open-weight coding model… |
@baseten |
Company |
Quote |
2026-07-21 |
4,248 |
39 |
3 |
1 |
0 |
5 |
| You can now generate 5 seconds of video in under 2.5 seconds, courtesy of the Baseten kernels team. |
@baseten |
Company |
Quote |
2026-07-20 |
3,773 |
25 |
0 |
0 |
0 |
3 |
| "So, can a model learn facts continually in its weights?
Creating usable knowledge: solvable
Preserving capability: so… |
@baseten |
Company |
Quote |
2026-07-17 |
5,175 |
27 |
1 |
3 |
0 |
14 |
| We're building in Canada!
We opened two new offices in Toronto and Montreal, and we're growing fast. If you’re excited… |
@baseten |
Company |
Original |
2026-07-16 |
41,424 |
147 |
17 |
13 |
7 |
45 |
| NVIDIA Nemotron 3 Embed 8B and 1B just dropped, and we're excited to offer them day 0 in our Model Library!
We're also… |
@baseten |
Company |
Original |
2026-07-16 |
1,907 |
27 |
1 |
1 |
0 |
3 |
| GLM 5.2 is one of the best open models available, but it can't support image inputs natively.
Until now. |
@baseten |
Company |
Quote |
2026-07-16 |
6,660 |
42 |
0 |
2 |
0 |
11 |
| Inkling is live on Baseten.
We’re proud to partner with @thinkymachines to provide day 0 support. Try it here: https:/… |
@baseten |
Company |
Original |
2026-07-15 |
4,702,745 |
87 |
6 |
10 |
4 |
17 |
| Join us for Built on Baseten: a showcase for early-stage founders and technical builders who are owning their intellige… |
@baseten |
Company |
Original |
2026-07-15 |
1,616 |
14 |
1 |
4 |
0 |
4 |
| Some workloads demand either the highest throughput or the lowest latency. Embedding workloads need both.
We built Bas… |
@baseten |
Company |
Quote |
2026-07-14 |
1,984 |
17 |
1 |
0 |
0 |
2 |
| Research has always been core to Baseten, from training to model performance and beyond.
Charlie, who leads part of o… |
@baseten |
Company |
Quote |
2026-07-14 |
3,321 |
25 |
2 |
1 |
0 |
6 |
| Step 3.7 Flash is now in the Baseten Model Library!
This is a 198B-parameter sparse MoE model (11B active per token) w… |
@baseten |
Company |
Original |
2026-07-14 |
7,899 |
17 |
0 |
2 |
1 |
2 |
| We're seeing it too: companies of all sizes are adopting more open-source AI, including alongside closed models.
The f… |
@baseten |
Company |
Quote |
2026-07-14 |
4,555 |
31 |
0 |
2 |
0 |
6 |
| Agents will use the web more than humans ever have. High-accuracy AI web search is what @p0 is built for.
We're proud … |
@baseten |
Company |
Quote |
2026-07-10 |
7,481 |
46 |
5 |
6 |
0 |
19 |
| https://t.co/rVIxR6TkDK |
@baseten |
Company |
Original |
2026-07-10 |
18,080 |
48 |
3 |
3 |
4 |
29 |
| Excited to be featured in Notion's newest customer story on how we're scaling expertise with Custom Agents!
Meet katzg… |
@baseten |
Company |
Quote |
2026-07-08 |
4,473 |
36 |
0 |
3 |
0 |
6 |
| If you're at RAISE Paris today, come see @tuhinone on the master stage at 4:40 p.m. with @CorinneMRiley! https://t.co/c… |
@baseten |
Company |
Original |
2026-07-08 |
3,612 |
17 |
3 |
2 |
1 |
1 |
| Thank you, @tbpn, for having our CEO @tuhinone on to talk about how companies of all sizes are investing in owning thei… |
@baseten |
Company |
Original |
2026-07-07 |
6,684 |
57 |
2 |
6 |
0 |
22 |
| We partnered with Braintrust to evaluate how GLM-5.2 performs on long-context retrieval tasks.
Not only is its retrie… |
@baseten |
Company |
Quote |
2026-07-07 |
5,849 |
44 |
1 |
3 |
0 |
19 |
| The future will be built on many models.
Open-source: good enough for many use cases at a small fraction of the price.… |
@baseten |
Company |
Quote |
2026-07-07 |
9,768 |
62 |
2 |
7 |
1 |
24 |
| The future will be built on many models, each intentionally chosen for specific and specialized tasks.
Great read fro… |
@baseten |
Company |
Quote |
2026-07-06 |
8,793 |
39 |
1 |
1 |
0 |
23 |
| We're thrilled to power Laguna XS 2.1 from Poolside via the Baseten Frontier Gateway!
Laguna XS 2.1 supports a 256K co… |
@baseten |
Company |
Quote |
2026-07-02 |
4,097 |
20 |
1 |
2 |
0 |
2 |
| Catch our CEO @tuhinone live on @tbpn at 1 p.m. PT! https://t.co/cCuTfB58X8 |
@baseten |
Company |
Original |
2026-07-02 |
1,288 |
21 |
2 |
2 |
0 |
3 |
| We made an early bet on multi-cloud infrastructure. Our Multi-cloud Capacity Manager (MCM) abstracts 18+ clouds, which:… |
@baseten |
Company |
Original |
2026-07-01 |
1,571 |
20 |
2 |
2 |
1 |
3 |
| We partnered with Braintrust to highlight the real-world strength of the latest open-source model.
For long-context re… |
@baseten |
Company |
Quote |
2026-06-30 |
6,436 |
52 |
4 |
1 |
0 |
27 |
| Plus:
- You own the performance optimizations.
- You can specialize them for your exact use case.
We love open-source! |
@baseten |
Company |
Quote |
2026-06-30 |
3,824 |
43 |
3 |
2 |
0 |
5 |
| Speechify was built to make learning accessible to everyone. Simba 3.0 powers its voice-first productivity platform for… |
@baseten |
Company |
Original |
2026-06-29 |
7,138 |
36 |
10 |
2 |
3 |
9 |
| Voice AI is unforgiving: any delay ruins the user experience. We're honored to be SpeechifyAI's sole provider powering … |
@baseten |
Company |
Quote |
2026-06-29 |
2,893 |
26 |
4 |
2 |
0 |
3 |
| NYC team repping at Nasdaq tower this week! 💚
We're growing fast, and we're hiring across the board. If you're excited… |
@baseten |
Company |
Original |
2026-06-26 |
4,050 |
73 |
5 |
7 |
2 |
10 |
| Live draft model training is now part of our Speculation Engine in the Baseten Inference Stack.
Where rolled out, we s… |
@baseten |
Company |
Quote |
2026-06-26 |
3,879 |
32 |
0 |
3 |
0 |
12 |
| Excited to power GLM-5.2 on @cline!
How to use it in about 10 seconds: https://t.co/MOrWq9qtda |
@baseten |
Company |
Original |
2026-06-25 |
8,892 |
86 |
7 |
4 |
1 |
18 |
| "Frontier models for the hardest general intelligence and post-trained open source for high-volume and specialized work… |
@baseten |
Company |
Quote |
2026-06-24 |
4,571 |
42 |
1 |
5 |
0 |
12 |
| You can now access our GLM-5.2 API through the Merge Gateway!
GLM-5.2 matches frontier model intelligence while runnin… |
@baseten |
Company |
Original |
2026-06-24 |
4,716 |
90 |
10 |
9 |
0 |
19 |
| "That's when they come to open-source models, that's when they come to Baseten, that's when they come to post-train mod… |
@baseten |
Company |
Quote |
2026-06-23 |
4,774 |
41 |
8 |
3 |
1 |
7 |
| Excited to be a day 0 launch partner for BioNeMo, NVIDIA's new, fully-open agent toolkit for scientific workflows!
Al… |
@baseten |
Company |
Quote |
2026-06-23 |
4,952 |
42 |
5 |
4 |
0 |
5 |
| The best open model with the best performance: GLM-5.2 runs at >280 TPS and <0.8s TTFT on Baseten.
Try it here: … |
@baseten |
Company |
Quote |
2026-06-22 |
30,442 |
129 |
5 |
12 |
6 |
46 |
| We’re excited to announce our $1.5B Series F.
Baseten exists to help companies own their intelligence and run AI produ… |
@baseten |
Company |
Quote |
2026-06-22 |
95,001 |
280 |
23 |
28 |
10 |
70 |
| Most supervised fine-tuning (SFT) studies run on generic data. Ours run on production tasks, paired with evals our team… |
@baseten |
Company |
Quote |
2026-06-18 |
3,544 |
20 |
2 |
3 |
0 |
10 |
| Moving from closed to open models can cut costs significantly.
See how much you can save by using GLM-5.2:
https://t.c… |
@baseten |
Company |
Quote |
2026-06-17 |
3,404 |
25 |
0 |
4 |
0 |
7 |
| Ever wonder how much money you'd save by switching to open-source models?
We just launched a cost calculator for that… |
@baseten |
Company |
Original |
2026-06-17 |
5,816,328 |
230 |
14 |
20 |
0 |
30 |
| Kimi K2.7 Code is live on Baseten Model APIs.
Kimi K2.7 Code has 30% lower reasoning-token usage compared to K2.6, wit… |
@baseten |
Company |
Original |
2026-06-17 |
910,406 |
58 |
4 |
6 |
1 |
6 |
| We 💚 @cursor_ai, thanks for inviting some of our post-training experts to give chalk talks at Cursor Compile! |
@baseten |
Company |
Quote |
2026-06-17 |
2,376 |
20 |
2 |
2 |
0 |
5 |
| Baseten is proud to partner with the Boltz team on this incredible launch. |
@baseten |
Company |
Quote |
2026-06-16 |
4,221 |
31 |
3 |
0 |
0 |
8 |
| I spy ... GLM 5.2 on @NotionHQ!
Check out https://t.co/e4br5JXagu’s latest advanced agentic and reasoning model, power… |
@baseten |
Company |
Original |
2026-06-16 |
2,981 |
36 |
1 |
1 |
0 |
6 |
| GLM 5.2 is live on Baseten.
5.2 is built for agentic engineering: stronger coding, sharper agentic reasoning, and a lo… |
@baseten |
Company |
Original |
2026-06-16 |
1,927,023 |
141 |
8 |
8 |
5 |
33 |
| The new AgentPerf benchmark by @ArtificialAnlys shows that @NVIDIAAI Blackwell delivers best performance for demanding … |
@baseten |
Company |
Original |
2026-06-12 |
2,476 |
16 |
3 |
0 |
1 |
5 |
| We're thrilled to be working with the Harvey team to push open models to frontier-level performance for legal AI.
Sho… |
@baseten |
Company |
Quote |
2026-06-12 |
2,578 |
20 |
2 |
3 |
0 |
3 |
| Congrats to the MiniMax team on the open-source launch of M3!
There are very few <500bn parameter models that can t… |
@baseten |
Company |
Original |
2026-06-12 |
21,638 |
43 |
3 |
5 |
1 |
10 |
| Join Baseten, Lovable, and ElevenLabs to hack on the future of healthcare. |
@baseten |
Company |
Quote |
2026-06-12 |
932 |
9 |
0 |
2 |
0 |
1 |
| We've heard from customers that they ship model updates >50% more often with rolling deploys than their previous sol… |
@baseten |
Company |
Quote |
2026-06-12 |
3,450 |
23 |
1 |
3 |
0 |
15 |
| Great to see @Baseten’s own @oneill_c and @part_harry_ sitting down with @cursor_ai’s @sjwhitmore to talk about the man… |
@baseten |
Company |
Quote |
2026-06-11 |
3,182 |
16 |
1 |
3 |
0 |
5 |
| We are excited to announce that we have partnered with @_inception_ai to make Mercury 2 available on Baseten. This mak… |
@baseten |
Company |
Quote |
2026-06-11 |
26,337 |
46 |
5 |
3 |
5 |
12 |
| https://t.co/HGVpRPEhBo |
@baseten |
Company |
Original |
2026-06-11 |
34,159 |
43 |
6 |
2 |
10 |
11 |
| The longer the context, the more memory your LLM needs. We introduce research techniques to compress that memory 200x o… |
@baseten |
Company |
Quote |
2026-06-10 |
3,666 |
39 |
0 |
5 |
1 |
9 |
| Baseten is live on the Respan Gateway.
Congratulations to the @RespanAI team on their Gateway launch as they bring obs… |
@baseten |
Company |
Original |
2026-06-09 |
1,324 |
14 |
7 |
3 |
0 |
2 |
| Join Charlie for a conversation with @thatsjonsense
and @sarahmsachs on how @GammaApp and @NotionHQ think about model s… |
@baseten |
Company |
Quote |
2026-06-08 |
2,663 |
15 |
0 |
4 |
0 |
1 |
| GLM 5.1 now achieves 160+ TPS and <2-second TTFT on Baseten.
Ideal for agentic workloads that need high throughput … |
@baseten |
Company |
Original |
2026-06-05 |
6,909 |
91 |
2 |
8 |
1 |
15 |
| Are you tired of waiting 17 minutes for an AI agent to finish a code change?
As an agent’s context grows, standard tra… |
@baseten |
Company |
Quote |
2026-06-04 |
1,517 |
20 |
0 |
2 |
0 |
3 |
| We are excited to welcome Gabe Stern as General Counsel.
Welcome, Gabe! https://t.co/fSpxLhZpCx |
@baseten |
Company |
Original |
2026-06-04 |
7,526 |
14 |
0 |
2 |
2 |
2 |
| Agents append to their own context. But attention is quadratic, so 2x context = 4x work per step.
Nemotron 3 Ultra swa… |
@baseten |
Company |
Quote |
2026-06-04 |
1,986 |
32 |
4 |
3 |
1 |
13 |
| 10M developers use @opencode every month. This means the experience has to feel the same every hour of every day; slow … |
@baseten |
Company |
Original |
2026-06-03 |
14,266 |
98 |
4 |
6 |
0 |
23 |
| Today, Baseten and @MicrosoftAI are excited to announce that MAI-Thinking-1 is coming to Baseten.
MAI-Thinking-1 is a … |
@baseten |
Company |
Original |
2026-06-02 |
15,919 |
65 |
5 |
3 |
2 |
18 |
| We're hosting an Inference Café for a16z tech week NYC at Verci Flatiron!
For the folks experimenting with frontier op… |
@baseten |
Company |
Original |
2026-06-01 |
2,763 |
27 |
1 |
1 |
0 |
6 |
| Philip worked with @rimelabs to create an enterprise-grade voice clone for an ambitious project: an AI-narrated audiobo… |
@baseten |
Company |
Quote |
2026-06-01 |
3,133 |
21 |
2 |
2 |
1 |
5 |
| Proud to partner with JetBrains. Congrats on the launch! |
@baseten |
Company |
Quote |
2026-06-01 |
2,681 |
23 |
0 |
1 |
0 |
1 |
| Want to teach a robot to open a door without shattering the glass, crushing the handle, or clipping through reality? If… |
@baseten |
Company |
Original |
2026-06-01 |
1,984 |
23 |
0 |
5 |
0 |
9 |
| We’re honored to be featured in the @Redpoint Infrared 100 list alongside our incredible customers.
As we enter the mu… |
@baseten |
Company |
Quote |
2026-05-29 |
3,328 |
28 |
2 |
4 |
0 |
2 |
| Our kernels team shipped 2.5x faster FLUX.2 image gen without visible quality loss.
How they did it: Distribution Matc… |
@baseten |
Company |
Quote |
2026-05-29 |
1,697 |
16 |
0 |
3 |
0 |
2 |
| Our research team partnered with @Harvey and showed that post-trained open models can compete at the frontier on LAB, t… |
@baseten |
Company |
Quote |
2026-05-27 |
6,043 |
37 |
6 |
4 |
0 |
18 |
| Excited to support the @trajectorylabs team as their inference partner.
Powering live, continuous post-training for f… |
@baseten |
Company |
Quote |
2026-05-27 |
5,925 |
42 |
4 |
4 |
1 |
3 |
| Most enterprise AI models stay static after deployment, forcing teams to rely on endless prompt tweaking.
This is the … |
@baseten |
Company |
Original |
2026-05-27 |
7,268 |
92 |
4 |
12 |
0 |
27 |
| If you haven’t built AI automation that relies on documents, it’s easy to underestimate how hard it is. Parsing perform… |
@baseten |
Company |
Quote |
2026-05-26 |
2,830 |
21 |
1 |
5 |
0 |
3 |
| Biotech R&D is generating more scientific AI models than ever, from protein structure prediction to molecular docking t… |
@baseten |
Company |
Original |
2026-05-20 |
2,313 |
35 |
10 |
1 |
0 |
10 |
| “Intensity plus joy — an aha moment for me at Baseten is that those two things are not on opposite sides of a spectrum.… |
@baseten |
Company |
Quote |
2026-05-19 |
11,264 |
24 |
1 |
0 |
1 |
8 |
| Sub-second image generation with Flux.2 [dev] and Qwen-Image:
Flux.2 [dev]: 2.3x faster, 0.98s latency (B200)
Qwen-Ima… |
@baseten |
Company |
Quote |
2026-05-18 |
3,324 |
24 |
2 |
1 |
1 |
5 |
| Last week, we launched Baseten Frontier Gateway. This week, Marylise sits down with Bola to talk about why. |
@baseten |
Company |
Quote |
2026-05-15 |
4,247 |
19 |
0 |
1 |
0 |
3 |
| We serve Qwen3-TTS on vLLM-Omni at $3 per 1M characters. That's 90% lower in cost than comparable closed-source TTS API… |
@baseten |
Company |
Quote |
2026-05-14 |
76,017 |
97 |
8 |
3 |
5 |
80 |
| “The question every app layer company is now asking is no longer ‘how do we use AI?’ It is ‘how do we resist commodific… |
@baseten |
Company |
Quote |
2026-05-13 |
1,865 |
21 |
0 |
0 |
0 |
12 |
| Intelligence should be defined by the people closest to the work. Intelligence should be owned by all of us.
Let’s bui… |
@baseten |
Company |
Quote |
2026-05-13 |
35,848 |
47 |
4 |
3 |
0 |
20 |
| We're proud to share our partnership story with @SpeechifyAI.
Speechify just announced SIMBA 3.0, ranked top 10 global… |
@baseten |
Company |
Quote |
2026-05-12 |
1,676 |
16 |
7 |
1 |
1 |
0 |
| https://t.co/D0Oo41Bgnn |
@baseten |
Company |
Original |
2026-05-12 |
5,225 |
23 |
2 |
0 |
3 |
6 |
| EliseAI is the leading AI startup automating complex housing and healthcare systems, which are two of the most operatio… |
@baseten |
Company |
Original |
2026-05-11 |
4,693 |
28 |
3 |
0 |
1 |
13 |
| DFlash vs EAGLE on a single B200 with Qwen3-8B:
EAGLE: ~2x speedup, autoregressive drafting (one forward pass per toke… |
@baseten |
Company |
Quote |
2026-05-08 |
44,163 |
160 |
13 |
8 |
4 |
133 |
| Open-source RL libraries break at frontier scale. We built Baseten Loops to fix this.
Loops is a training SDK that tak… |
@baseten |
Company |
Quote |
2026-05-08 |
67,291 |
110 |
7 |
2 |
6 |
110 |
| We are excited to share our partnership story with @poolsideai and proud of what we've built together.
Poolside has be… |
@baseten |
Company |
Original |
2026-05-07 |
1,808 |
37 |
1 |
0 |
0 |
7 |
| "No post-training Pre-PMF"
Our CEO and Co-Founder @tuhinone sat down w/ @saranormous and @eladgil on @nopriorspod to d… |
@baseten |
Company |
Quote |
2026-05-06 |
9,080 |
36 |
4 |
1 |
0 |
20 |
| https://t.co/sGB4D3iKel |
@baseten |
Company |
Original |
2026-05-06 |
24,370 |
59 |
9 |
3 |
9 |
27 |
| Harvey is raising the bar in legal AI. Congratulations on launching the Legal Agent Benchmark! Baseten is proud to supp… |
@baseten |
Company |
Quote |
2026-05-06 |
2,007 |
25 |
2 |
0 |
0 |
2 |
| Introducing @NVIDIA Nemotron 3 Nano Omni.
NVIDIA Nemotron 3 Nano Omni is an open multimodal foundation model that unif… |
@baseten |
Company |
Original |
2026-04-28 |
4,968 |
32 |
8 |
1 |
1 |
5 |
| Laguna XS.2 has landed, and it’s live on Baseten. We’ve baked in inference optimizations so you can deploy @poolsideai'… |
@baseten |
Company |
Original |
2026-04-28 |
3,833 |
52 |
7 |
2 |
0 |
4 |
| The strongest open-source agentic model is live on Baseten!
DeepSeek V4 is a preview of two powerful MoE models: V4-Pr… |
@baseten |
Company |
Original |
2026-04-24 |
3,701 |
70 |
2 |
4 |
0 |
14 |
| We made an early bet that our permissions model needed to handle complex, many-to-many relationships, and orgs of any s… |
@baseten |
Company |
Quote |
2026-04-23 |
3,023 |
13 |
1 |
0 |
0 |
7 |
| Our engineers developed RadixMLP, a technique that exploits the position-wise nature of MLPs, LayerNorms, linear projec… |
@baseten |
Company |
Quote |
2026-04-22 |
13,308 |
86 |
6 |
0 |
2 |
77 |
| Kimi K2.6 has landed, and it is live on Baseten!
We have baked in multiple inference optimizations so that you can lev… |
@baseten |
Company |
Original |
2026-04-20 |
116,549 |
141 |
8 |
12 |
4 |
46 |
| Most people installing AI coding agents are optimizing the wrong things.
Harnesses are the new IDE. Here are the three… |
@baseten |
Company |
Quote |
2026-04-17 |
5,788 |
36 |
1 |
1 |
0 |
56 |
| Honored to be featured in the Forbes 2026 AI 50 List alongside so many of our great customers!
https://t.co/gMsw9PrPM… |
@baseten |
Company |
Original |
2026-04-16 |
1,801 |
24 |
6 |
1 |
0 |
6 |
| Speculative decoding is one of the highest-leverage inference optimizations out there.
Here are tactical steps on how … |
@baseten |
Company |
Quote |
2026-04-13 |
7,414 |
66 |
3 |
0 |
0 |
58 |
| HumanX was a blast! Some highlights:
-> @tuhinone and @saranormous took the main stage to discuss what it takes to … |
@baseten |
Company |
Original |
2026-04-13 |
1,311 |
11 |
1 |
0 |
0 |
0 |
| We recently launched BDN for 2–3x faster cold starts for large models at massive scale.
This is how it works. |
@baseten |
Company |
Quote |
2026-04-09 |
4,996 |
31 |
1 |
0 |
1 |
25 |
| Come jam on open-source agents with us!
We're hosting a NYC hackathon with our friends at @veris_ai on April 18 in SoH… |
@baseten |
Company |
Original |
2026-04-09 |
1,232 |
13 |
1 |
1 |
1 |
4 |
| In finance, speed is a competitive edge.
@hebbia powers the AI that gives the world's leading financial institutions b… |
@baseten |
Company |
Original |
2026-04-08 |
7,092 |
44 |
7 |
3 |
3 |
17 |
| We're at HumanX all week!
Here's what to expect today:
-> Lightning talk at the Vultr booth with @philipkiely (3 PM)
… |
@baseten |
Company |
Original |
2026-04-07 |
812 |
12 |
1 |
0 |
0 |
3 |
| We are excited to be the day 0 launch partner for Rime Mist v3!
Mist v3 is a significant step forward for production v… |
@baseten |
Company |
Original |
2026-04-07 |
1,256 |
15 |
1 |
0 |
0 |
5 |
| Our engineers just shipped the fastest named entity recognition (NER) inference on the market: 1 ms P50 and 3 ms P99 se… |
@baseten |
Company |
Quote |
2026-04-06 |
13,074 |
114 |
6 |
4 |
1 |
74 |
| What if training LLMs didn't require rebuilding your entire product as a sandbox?
RL training forces companies to repa… |
@baseten |
Company |
Quote |
2026-04-06 |
14,147 |
101 |
5 |
5 |
0 |
111 |
| Gemma 4 is live on Baseten and available to all customers on day 0 via the Baseten model library.
All models in the Ge… |
@baseten |
Company |
Original |
2026-04-02 |
2,467 |
42 |
7 |
3 |
0 |
12 |
| What if LLMs could remember as humans do?
LLM memory is either perfect and lossless or ultra-compressed. What does a s… |
@baseten |
Company |
Quote |
2026-04-01 |
5,704 |
53 |
9 |
1 |
0 |
45 |
| We've had a great month of March! A brief recap:
-> NVIDIA GTC, featuring books, ice cream, and swag
-> KubeCon EMEA, … |
@baseten |
Company |
Original |
2026-04-01 |
922 |
3 |
0 |
0 |
1 |
1 |
| Our researchers have been running autoresearch loops on Baseten Training for months. It's now a core part of how we wor… |
@baseten |
Company |
Quote |
2026-03-30 |
20,385 |
67 |
7 |
0 |
0 |
87 |
| Initially, we believed that open-source models could only accomplish a fraction of the tasks of closed-source models. S… |
@baseten |
Company |
Quote |
2026-03-30 |
23,092 |
55 |
3 |
0 |
1 |
37 |
| We're thrilled to be featured in @fastcompany's 2026 list of most innovative companies in applied AI.
This is the year… |
@baseten |
Company |
Original |
2026-03-27 |
1,708 |
18 |
2 |
0 |
2 |
4 |
| You've seen the Pied Piper memes, but if you want to understand how TurboQuant works you should read this article from … |
@baseten |
Company |
Quote |
2026-03-27 |
10,370 |
56 |
6 |
2 |
2 |
60 |
| Check out GLM-5 on Baseten Model API live in @Linkup_platform, making agents both fast and accurate: |
@baseten |
Company |
Quote |
2026-03-25 |
1,945 |
14 |
2 |
2 |
0 |
1 |
| Zed built its editor from scratch because performance is non-negotiable for a responsive IDE. When your product lives o… |
@baseten |
Company |
Original |
2026-03-24 |
19,543 |
57 |
4 |
2 |
3 |
9 |
| Open-source LLM training is fragmented, and there's no clear guide.
@stefanopopoulos on our Post-Training team mapped… |
@baseten |
Company |
Quote |
2026-03-24 |
9,542 |
67 |
8 |
3 |
3 |
59 |
| We are thrilled to welcome Sameer Paranjpye to lead our engineering organization.
Welcome, Sameer!
https://t.co/ysqoU… |
@baseten |
Company |
Original |
2026-03-23 |
8,538 |
40 |
6 |
2 |
2 |
5 |
| Cold starts for large models are one of the hardest problems in AI inference infrastructure. Today we're launching the … |
@baseten |
Company |
Quote |
2026-03-19 |
2,648 |
33 |
2 |
2 |
0 |
9 |
| Developers used to argue about programming languages; now they argue about harnesses.
NemoClaw is NVIDIA's answer to y… |
@baseten |
Company |
Quote |
2026-03-19 |
2,043 |
12 |
3 |
0 |
0 |
5 |
| Want to enjoy a sweet treat in between sessions and meetings?
Come by for a scoop of our ice cream, created exclusivel… |
@baseten |
Company |
Original |
2026-03-18 |
969 |
9 |
3 |
1 |
0 |
2 |
| If you've signed up for a physical copy of Inference Engineering and are at GTC, come to our booth for a free copy fres… |
@baseten |
Company |
Original |
2026-03-18 |
1,701 |
14 |
0 |
1 |
1 |
2 |
| Travel in style. Get rides to and from GTC in our cars. 👀 https://t.co/x4sMwXCw9S |
@baseten |
Company |
Original |
2026-03-18 |
1,814 |
14 |
2 |
1 |
1 |
4 |
| Don't miss @philipkiely's talk on high-performance inference for frontier AI models at 4 PM today! https://t.co/3YkHXJk… |
@baseten |
Company |
Original |
2026-03-16 |
1,115 |
14 |
3 |
1 |
0 |
4 |
| Live from Jensen's keynote remarks at GTC:
"The inflection point of inference has arrived.
AI now has to think. In or… |
@baseten |
Company |
Original |
2026-03-16 |
1,924 |
19 |
4 |
3 |
0 |
5 |
| Happy GTC week!
Need to hail a ride to or from the conference? Reply or repost with a picture of our Baseten-wrapped r… |
@baseten |
Company |
Original |
2026-03-16 |
2,480 |
35 |
3 |
1 |
2 |
1 |
| In this piece, @maxkirkby and @oneill_c use constitutional alignment as a testbed to evaluate the importance of on- ver… |
@baseten |
Company |
Quote |
2026-03-13 |
970 |
13 |
0 |
0 |
0 |
1 |
| We're excited for NVIDIA GTC next week!
Here's an overview of what to expect:
-> @philipkiely's speaking session on h… |
@baseten |
Company |
Original |
2026-03-13 |
2,900 |
16 |
0 |
0 |
1 |
2 |
| We love seeing our customer and partner friends IRL, and we've got a lot more events coming up! https://t.co/bnv5j6FJXc… |
@baseten |
Company |
Original |
2026-03-12 |
734 |
14 |
1 |
1 |
0 |
0 |
| We are thrilled to welcome Matt Slagle to lead our global revenue organization!
Welcome, Matt.
https://t.co/32eBpGME5… |
@baseten |
Company |
Original |
2026-03-12 |
4,068 |
22 |
2 |
1 |
1 |
0 |
| We're excited to be day-0 launch partners for NVIDIA Nemotron 3 Super!
You can try it now on Baseten, or read @rachelr… |
@baseten |
Company |
Original |
2026-03-11 |
4,933 |
64 |
7 |
2 |
2 |
9 |
| The Posit team is behind some of the most widely used data science tools, including RStudio and Positron. Posit uses Ba… |
@baseten |
Company |
Original |
2026-03-10 |
1,447 |
14 |
2 |
0 |
0 |
2 |
| you're probably worried about the wrong 9's |
@baseten |
Company |
Quote |
2026-03-09 |
2,296 |
24 |
0 |
1 |
0 |
6 |
| We've launched the fastest GLM 5 API available at 190 TPS and 0.79 sec TTFT with the Baseten Inference Stack.
Ready fo… |
@baseten |
Company |
Original |
2026-03-06 |
19,658 |
101 |
8 |
16 |
3 |
26 |
| Long-running agents accumulate context while model memory stays fixed. This leads to a tradeoff: either discard older i… |
@baseten |
Company |
Original |
2026-03-05 |
2,496 |
39 |
5 |
6 |
2 |
15 |
| Earlier this month, we hosted our biannual company-wide offsite and gathered 180 teammates in Austin, TX. Highlights in… |
@baseten |
Company |
Original |
2026-03-04 |
5,190 |
44 |
1 |
6 |
3 |
10 |
| A few months ago, we trained a specialist model that beats Gemini on emergency medicine documentation and runs 6–8x fas… |
@baseten |
Company |
Original |
2026-03-03 |
1,209 |
21 |
2 |
1 |
0 |
6 |
| We painted San Francisco green and pink, and the message is clear — you need to own your inference.
If you spot us ar… |
@baseten |
Company |
Original |
2026-03-02 |
3,345 |
44 |
9 |
11 |
5 |
13 |
| What is a core advantage of open-source? The community and economics.
" If you need large manpower for something you'r… |
@baseten |
Company |
Original |
2026-02-27 |
747 |
13 |
0 |
0 |
0 |
3 |
| Large world models (LWMs) are the latest modality in AI. These foundation models push the boundaries of what AI can do … |
@baseten |
Company |
Quote |
2026-02-25 |
1,609 |
17 |
0 |
0 |
0 |
6 |
| Introducing RadixMLP: intra-batch prefix deduplication for 1.4–5x faster prefill.
Tokens with identical prefixes (like… |
@baseten |
Company |
Original |
2026-02-24 |
2,951 |
27 |
3 |
0 |
2 |
18 |
| New in the Baseten ML Cookbook: a multi-node training recipe for @MiniMax_AI M2.5, the open MoE model that's SOTA in co… |
@baseten |
Company |
Original |
2026-02-24 |
943 |
13 |
0 |
0 |
0 |
7 |
| we've learned from the best! |
@baseten |
Company |
Quote |
2026-02-24 |
887 |
19 |
0 |
1 |
0 |
1 |
| Every developer has the opportunity to own their intelligence by learning how to do inference, but today the field is s… |
@baseten |
Company |
Quote |
2026-02-23 |
2,082 |
29 |
1 |
2 |
0 |
4 |
| "Medicine doesn’t tolerate “remarkably capable.” It requires specifically correct."
We're honored to partner with Open… |
@baseten |
Company |
Quote |
2026-02-23 |
795 |
7 |
0 |
1 |
0 |
3 |
| "No other product lets you launch ten different training jobs on four different datasets." –Head of Clinical NLP, OpenE… |
@baseten |
Company |
Original |
2026-02-20 |
16,895 |
30 |
7 |
0 |
2 |
24 |
| GLM 5 is live on Baseten.
Opus 4.6 level performance at 10% of the cost.
It takes the intelligence from its predecess… |
@baseten |
Company |
Original |
2026-02-19 |
2,034 |
34 |
1 |
2 |
0 |
6 |
| Generational AI companies are powered by Baseten.
Why? We obsess over the milliseconds, so they can ship the future.
… |
@baseten |
Company |
Original |
2026-02-19 |
4,098 |
37 |
7 |
4 |
4 |
11 |
| coming soon 👀 |
@baseten |
Company |
Quote |
2026-02-18 |
2,277 |
30 |
1 |
0 |
0 |
7 |
| The Gamma team uses the Baseten Inference Stack to:
-> Generate millions of images per day
-> Reduce latency per image… |
@baseten |
Company |
Original |
2026-02-17 |
756 |
16 |
1 |
0 |
0 |
4 |
| Just because it's a federal holiday doesn't mean we're slacking.
MiniMax M2.5 is live on our Model APIs.
Try it here:… |
@baseten |
Company |
Original |
2026-02-16 |
5,995 |
44 |
3 |
1 |
0 |
3 |
| RL often throws away useful signal at intermediate steps, or as @karpathy put it, it's like "sucking supervision throug… |
@baseten |
Company |
Original |
2026-02-13 |
14,582 |
34 |
3 |
0 |
1 |
12 |
| We replicated Microsoft Research's Generative Adversarial Distillation (GAD) to distill Qwen3-4B from GPT-5.2.
Standar… |
@baseten |
Company |
Original |
2026-02-13 |
4,089 |
30 |
4 |
1 |
3 |
10 |
| Sully.ai is transforming healthcare efficiency with Baseten’s Model APIs, running frontier open models like gpt-oss-12… |
@baseten |
Company |
Original |
2026-02-12 |
1,324 |
24 |
4 |
1 |
0 |
5 |
| Following up on yesterday's release 🚨
How did we build the fastest Kimi K2.5 inference?
• Custom EAGLE-3 speculator… |
@baseten |
Company |
Quote |
2026-02-11 |
16,860 |
47 |
7 |
1 |
2 |
5 |
| Continuing this week with a case study ☕️
How did @sullyai return 30M+ clinical minutes to doctors? By ditching closed… |
@baseten |
Company |
Original |
2026-02-10 |
4,219 |
30 |
3 |
2 |
4 |
10 |
| Introducing Kimi K2.5 on Baseten’s Model APIs with the most performant TTFT (0.26 sec) and TPS (340) on Artificial Anal… |
@baseten |
Company |
Original |
2026-02-10 |
15,323 |
98 |
8 |
11 |
5 |
21 |
| We're living in the era of metric obsession. (How is your sleep score after Super Bowl weekend?) 🏈
Now your metric obs… |
@baseten |
Company |
Original |
2026-02-09 |
1,081 |
28 |
0 |
2 |
0 |
4 |
| LLMs are amnesiacs. Once context fills up, they forget everything. To fight this means grappling with a core question: … |
@baseten |
Company |
Original |
2026-02-06 |
7,722 |
79 |
11 |
8 |
1 |
29 |
| What’s the connection between LLM fine-tuning and Hollywood? It turns out that there are many, from VFX tooling to bran… |
@baseten |
Company |
Original |
2026-02-05 |
797 |
16 |
4 |
2 |
0 |
0 |
| LLMs display human-like behavior, with Karpathy once describing them as stochastic "people spirits."
This makes them n… |
@baseten |
Company |
Original |
2026-02-05 |
702 |
13 |
0 |
0 |
0 |
3 |
| The best OpenClaw🦞 setup is fully open-source.
Kimi K2.5 on Baseten outperforms Opus 4.5 on agentic benchmarks at 8x … |
@baseten |
Company |
Original |
2026-02-04 |
7,724 |
55 |
6 |
2 |
2 |
34 |
| MARS-Flash is now available on Baseten.
If you know Baseten, you know we’re obsessed with speed. Enter, MARS-Flash. MA… |
@baseten |
Company |
Original |
2026-02-04 |
690 |
12 |
0 |
1 |
0 |
0 |
| Ready to cook? 🍳
New training recipe in the Baseten ML Cookbook: GLM 4.7 and 4.7 Flash
Fine tune the leading multimod… |
@baseten |
Company |
Original |
2026-02-03 |
889 |
10 |
0 |
0 |
1 |
6 |
| Thanks @NVIDIAAI for inviting us to Dynamo Day! We're active users of Dynamo, iterating on it in production for perform… |
@baseten |
Company |
Original |
2026-02-03 |
3,952 |
16 |
1 |
0 |
1 |
2 |
| Nemotron 3 Nano NVFP4 is now available on Baseten + NVIDIA B200
BF16-level accuracy, up to 4× higher throughput vs FP8… |
@baseten |
Company |
Original |
2026-01-28 |
2,362 |
16 |
2 |
0 |
1 |
2 |
| Thank you @BloombergTV for having our CEO and co-founder @tuhinone and day 1 investor @saranormous yesterday to discuss… |
@baseten |
Company |
Original |
2026-01-27 |
2,625 |
29 |
3 |
2 |
0 |
4 |
| We boosted acceptance rate by up to 40% with the Baseten Speculation Engine.
How? By combining Multi-Token Prediction … |
@baseten |
Company |
Original |
2026-01-27 |
13,959 |
29 |
5 |
1 |
2 |
6 |
| LIVE
Tune in to hear @tuhinone discuss our series E, open source, and the multi-model future on CNBC |
@baseten |
Company |
Quote |
2026-01-26 |
1,396 |
6 |
1 |
1 |
1 |
0 |
| Last week, we announced our $300M series E at a $5B valuation.
This week, we’re unpacking it all (and more!), starting… |
@baseten |
Company |
Original |
2026-01-26 |
2,815 |
31 |
2 |
0 |
0 |
2 |
| ANNOUNCEMENT: @tuhinone IS COMING ON @tbpn RIGHT NOW
WE REPEAT: @tuhinone IS COMING ON TBPN RIGHT NOW |
@baseten |
Company |
Original |
2026-01-23 |
1,514 |
26 |
1 |
0 |
0 |
0 |
| We’re thrilled to announce that we have raised $300M at a $5B valuation. The round is led by IVP and CapitalG, both dou… |
@baseten |
Company |
Original |
2026-01-23 |
300,560 |
344 |
24 |
41 |
24 |
164 |
| Tired of waiting for video generation? Say less.
We've optimized the Wan 2.2 runtime to hit: 3x faster inference on NV… |
@baseten |
Company |
Original |
2026-01-22 |
1,898 |
19 |
3 |
4 |
0 |
3 |
| We’re thrilled to be working with @LangChain to power the fastest way to generate production-ready agents without code.… |
@baseten |
Company |
Original |
2026-01-21 |
5,446 |
23 |
4 |
4 |
1 |
8 |
| Want to learn about how to run high performance LLM inference at scale?Our Head of DevRel @philipkiely has the perfect … |
@baseten |
Company |
Original |
2026-01-20 |
2,718 |
26 |
3 |
0 |
2 |
6 |
| 🚀 We're thrilled to introduce the fastest, most accurate, and cost-efficient Whisper-powered transcription and diarizat… |
@baseten |
Company |
Original |
2026-01-16 |
2,734 |
28 |
2 |
7 |
2 |
4 |
| "Engineers who are product-minded and know what they want to ship, because they understand their users, now they have i… |
@baseten |
Company |
Original |
2026-01-13 |
1,079 |
10 |
1 |
0 |
0 |
1 |
| Inference performance isn’t just about the model. It relies on the entire inference stack.
In our Inference Stack whit… |
@baseten |
Company |
Original |
2026-01-09 |
9,097 |
42 |
5 |
2 |
2 |
20 |
| "It's going to let you take the world as you see it, or let the code that you have that works today [...] and just go."… |
@baseten |
Company |
Original |
2026-01-07 |
1,002 |
5 |
1 |
0 |
1 |
2 |
| Here’s the part people don’t like to say out loud: in most practical settings, RL is the wrong starting point.
The cri… |
@baseten |
Company |
Original |
2026-01-06 |
1,084 |
11 |
3 |
0 |
1 |
1 |
| the fastest GLM 4.7 available to try today on Baseten
https://t.co/yYLe0iKxNv |
@baseten |
Company |
Quote |
2025-12-29 |
2,082 |
13 |
1 |
1 |
0 |
3 |
| Just in time for the new year!
Awesome job by our model performance team to hit the top of @ArtificialAnlys for GLM 4… |
@baseten |
Company |
Quote |
2025-12-28 |
1,723 |
20 |
3 |
1 |
0 |
1 |
| 2025 redefined what it takes to get an AI application production-ready. Here are 3 infra learnings that will help you b… |
@baseten |
Company |
Original |
2025-12-22 |
888 |
11 |
2 |
1 |
0 |
3 |
| @thealexker from Baseten recently spoke at AI Dev x NYC by @DeepLearningAI.
In this talk, he covers the rise of open-… |
@baseten |
Company |
Original |
2025-12-19 |
580 |
7 |
0 |
1 |
0 |
0 |
| Turns out we know a lot of people that like to go fast.
Congrats to Stewart Noll and Alex Choy for taking the top tw… |
@baseten |
Company |
Original |
2025-12-17 |
1,219 |
28 |
1 |
1 |
0 |
2 |
| Baseten supports @nvidia Nemotron 3 Nano on day zero
Up to 4× faster token generation, high accuracy, and predictable … |
@baseten |
Company |
Original |
2025-12-15 |
2,733 |
21 |
2 |
1 |
0 |
1 |
| RL is a critical step towards serving production-ready AI applications with very specific needs requirements. In this g… |
@baseten |
Company |
Original |
2025-12-12 |
3,038 |
17 |
4 |
2 |
1 |
6 |
| "We want people to own their own intelligence, and we now see a really straight shot to get there."
@amiruci sits down… |
@baseten |
Company |
Original |
2025-12-11 |
1,768 |
16 |
3 |
3 |
1 |
4 |
| Today we’re welcoming the @parsedlabs team to Baseten!
With their RL and post-training expertise, Baseten is enabling … |
@baseten |
Company |
Original |
2025-12-10 |
22,627 |
47 |
5 |
2 |
4 |
10 |
| "These video gen models are getting pretty good. The real Hollywood use case is they messed up one of the shots and the… |
@baseten |
Company |
Original |
2025-12-09 |
3,754 |
25 |
5 |
1 |
1 |
6 |
| "Baseten's [training] solution is not point and click. Baseten's solution caters to people who are more hands on. When … |
@baseten |
Company |
Original |
2025-12-08 |
2,099 |
32 |
3 |
1 |
1 |
2 |
| Are you wondering how to get GPT-5-level performance at the fraction of the cost? We walk you through the top 3 directi… |
@baseten |
Company |
Original |
2025-12-05 |
779 |
11 |
0 |
0 |
0 |
2 |
| We're excited to partner with @getstream_io to help developers build fast, production-ready Vision Agents.
Together, … |
@baseten |
Company |
Quote |
2025-12-05 |
1,379 |
9 |
2 |
1 |
0 |
1 |
| If you need an adrenaline rush to wake up from your post-Thanksgiving stupor… we got you.
@deepseek_ai V3.2 dropped th… |
@baseten |
Company |
Original |
2025-12-04 |
4,760 |
43 |
12 |
10 |
4 |
2 |
| Agents that don't hallucinate? Meet APT: @ScaledCognition's Agentic Pretrained Transformer — the only frontier model fo… |
@baseten |
Company |
Original |
2025-12-04 |
1,449 |
16 |
2 |
3 |
2 |
5 |
| Baseten is proud to support training jobs for The LLM Data Company 💪 |
@baseten |
Company |
Quote |
2025-12-03 |
2,040 |
17 |
4 |
0 |
1 |
1 |
| Voice AI should feel like talking to a human. @rimelabs gets this and their models capture nuance, tone, and personali… |
@baseten |
Company |
Original |
2025-12-02 |
851 |
24 |
1 |
3 |
0 |
4 |
| Congrats to the team at Mistral AI on the new Apache-licensed Mistral Large 3, a foundation model with frontier-class i… |
@baseten |
Company |
Quote |
2025-12-02 |
952 |
10 |
0 |
2 |
0 |
0 |
| Cool 🚀 Launch alert!
We’ve officially opened applications for Baseten for Startups
> Up to $25K credits for inferenc… |
@baseten |
Company |
Original |
2025-12-01 |
33,429 |
38 |
7 |
3 |
6 |
14 |
| Enterprises are adopting AI way faster than anyone expected and @Get_WRITER is leading the charge.
From ROI-driven tra… |
@baseten |
Company |
Original |
2025-11-21 |
39,666 |
38 |
6 |
3 |
3 |
7 |
| Congrats to our friends at Deep Cogito on launching the most powerful US-based OSS model. It turns out LLM self play pr… |
@baseten |
Company |
Quote |
2025-11-19 |
1,265 |
10 |
0 |
0 |
0 |
0 |
| @tuhinone sits down on the Gradient Dissent podcast by @wandb
They discuss all things inference and what sets Baseten… |
@baseten |
Company |
Original |
2025-11-19 |
413 |
7 |
0 |
1 |
0 |
0 |
| Shoutout to the incredible team at @oxen_ai! Turning datasets → deployed models like it’s light work. They build fast. … |
@baseten |
Company |
Original |
2025-11-18 |
1,691 |
20 |
4 |
4 |
3 |
2 |
| Happy Monday 👋 We're pleased to welcome some new Baseten crew members to the team. Say hello to Tom Berger, Paulina Pev… |
@baseten |
Company |
Original |
2025-11-17 |
762 |
5 |
0 |
0 |
0 |
0 |
| @thealexker just spoke on stage at @DeepLearningAI AI Dev 25 x NYC about how Baseten helped Sourcegraph optimize their … |
@baseten |
Company |
Original |
2025-11-14 |
174 |
8 |
0 |
0 |
0 |
0 |
| We're here at AI Dev 25 x NYC with @DeepLearningAI and @AndrewYNg
Come say hi! https://t.co/CPVi2GFsCR |
@baseten |
Company |
Original |
2025-11-14 |
1,649 |
17 |
2 |
1 |
1 |
2 |
| Working with the @GammaApp team never quite feels like work, and that’s how their product feels. "Criminally fun."
We … |
@baseten |
Company |
Original |
2025-11-13 |
22,867 |
44 |
4 |
1 |
3 |
7 |
| Baseten used @nvidia Dynamo to double inference speed for long-context code generation and increased throughput by 1.6x… |
@baseten |
Company |
Quote |
2025-11-13 |
1,214 |
16 |
0 |
1 |
0 |
3 |
| From cricket to the earliest days of Baseten to where we are today, @adityaag and @tuhinone go deep on the Minus One po… |
@baseten |
Company |
Quote |
2025-11-13 |
2,619 |
18 |
2 |
0 |
0 |
1 |
| Welcome to the new age Defense Against the Dark Arts.
It's called fast inference! (& Harry Potter would be jealous). … |
@baseten |
Company |
Original |
2025-11-12 |
1,922 |
26 |
2 |
4 |
2 |
4 |
| Congrats to the World Labs team on the launch today! Marble lets you create 3D worlds from just a single image, text pr… |
@baseten |
Company |
Quote |
2025-11-12 |
3,070 |
33 |
2 |
6 |
4 |
3 |
| We had a great time at our Baseten rooftop happy hour at KubeCon with our friends at @OpsLevelHQ, @Sentry, @JellyFish_W… |
@baseten |
Company |
Original |
2025-11-12 |
963 |
21 |
0 |
0 |
1 |
1 |
| At @KubeCon_ ? Swing by Booth #631 to test your inference knowledge and earn some swag! https://t.co/rLCW9dvr1o |
@baseten |
Company |
Original |
2025-11-11 |
913 |
17 |
0 |
0 |
0 |
0 |
| It’s Monday, and we could all use a little help thinking. Thankfully we have the new Kimi K2 Thinking to do it for us. … |
@baseten |
Company |
Original |
2025-11-11 |
60,790 |
97 |
11 |
7 |
10 |
25 |
| Excited to share this piece from @VentureBeat spotlighting how Baseten is redefining the AI infrastructure game:
“Bas… |
@baseten |
Company |
Original |
2025-11-10 |
1,350 |
17 |
7 |
2 |
0 |
0 |
| Congratulations on the launch, this will be big! Excited to growth the partnership 🤝 |
@baseten |
Company |
Quote |
2025-11-10 |
1,540 |
21 |
0 |
1 |
0 |
1 |
| Congratulations to @GammaApp on this amazing milestone!
We're happy to continue our support of your amazing growth.
… |
@baseten |
Company |
Quote |
2025-11-10 |
1,850 |
13 |
2 |
0 |
0 |
2 |
| Heading to KubeCon next week? Come visit the team at Booth #631 to test your AI knowledge. Top of the leaderboard gets … |
@baseten |
Company |
Original |
2025-11-07 |
673 |
9 |
0 |
0 |
0 |
1 |
| Fun fact - we asked people to describe their favorite agent in SF. We got suggestions for a bunch of new agentic apps t… |
@baseten |
Company |
Original |
2025-11-05 |
653 |
13 |
0 |
1 |
0 |
1 |
| @Madisonkanna sits down with @thdxr from @opencode to talk about agents, code gen, and open source. Check out the full… |
@baseten |
Company |
Original |
2025-11-05 |
12,435 |
50 |
4 |
1 |
0 |
14 |
| Our team grows when our customers grow, and our customers are on a tear. Please welcome Zane, Allen, Brooke, and Katie … |
@baseten |
Company |
Original |
2025-11-04 |
835 |
13 |
0 |
0 |
0 |
2 |
| After months of feedback from our early customers and thousands of jobs completed, Baseten Training is officially ready… |
@baseten |
Company |
Original |
2025-10-30 |
14,573 |
42 |
5 |
4 |
8 |
8 |
| We had a great weekend at @CalHacks 12.0 in SF
Congratulations to the winners of the Baseten prize for best use of ope… |
@baseten |
Company |
Original |
2025-10-29 |
698 |
10 |
1 |
0 |
0 |
1 |
| We are so excited to be a launch partner for @nvidia Nemotron Nano 2 VL today and offer day-zero support for this high… |
@baseten |
Company |
Original |
2025-10-28 |
1,900 |
21 |
3 |
2 |
0 |
1 |
| Another week, another group of new Baseten colleagues to introduce! Please welcome Victoria Jones, Analisa Ruff, Aryan … |
@baseten |
Company |
Original |
2025-10-27 |
1,401 |
20 |
0 |
2 |
0 |
1 |
| This week, Baseten's model performance team unlocked the fastest TPS and TTFT for gpt-oss 120b on @nvidia hardware. Whe… |
@baseten |
Company |
Original |
2025-10-24 |
44,544 |
99 |
15 |
14 |
15 |
26 |
| DeepSeek-OCR stunned the internet this week with 10x more efficient compression, unlocking faster and cheaper intellige… |
@baseten |
Company |
Original |
2025-10-24 |
2,817 |
42 |
3 |
3 |
1 |
17 |
| We’re thrilled to welcome four new team members: Ben Levitt, Tisha Garza, Aaryam Sharma, and Narek Amirbekian.
Ben jo… |
@baseten |
Company |
Original |
2025-10-22 |
982 |
18 |
1 |
1 |
0 |
0 |
| We're seeing a lot of usage around DeepSeek's new OCR model. Alex packaged it so you can deploy and test it yourself - … |
@baseten |
Company |
Quote |
2025-10-22 |
3,779 |
26 |
5 |
2 |
0 |
17 |
| We see the massive AWS outage. Baseten web app is down but inference, new deploys, training jobs, and the model managem… |
@baseten |
Company |
Original |
2025-10-20 |
6,789 |
33 |
8 |
3 |
1 |
0 |
| We unleashed our model performance team on GLM 4.6 and we’re very excited to be the fastest provider available today on… |
@baseten |
Company |
Original |
2025-10-17 |
11,582 |
101 |
4 |
14 |
9 |
22 |
| Powering inference for the fastest growing AI companies like OpenEvidence, Writer, and Clay means being the first to us… |
@baseten |
Company |
Original |
2025-10-16 |
6,328 |
21 |
3 |
1 |
2 |
3 |
| For all the engineers who have ever dreamt of winning a championship belt - you can now give up your Brazilian Jiu Jits… |
@baseten |
Company |
Original |
2025-10-15 |
1,707 |
18 |
2 |
2 |
2 |
1 |
| Fast Company named Baseten one of the 5 Next Big Things in Tech 2025!
We’re proud to be recognized for powering the fa… |
@baseten |
Company |
Original |
2025-10-14 |
1,665 |
19 |
5 |
1 |
0 |
0 |
| From sketch to a 3D model in under 5 seconds with a 1B parameter model
We built a flower card generator using Autodesk… |
@baseten |
Company |
Original |
2025-10-13 |
818 |
13 |
0 |
1 |
0 |
2 |
| We caught up with the one and only @thdxr on Opencode's newly launched Zen and his hot takes
“Zen isn’t a for-profit t… |
@baseten |
Company |
Original |
2025-10-10 |
11,372 |
35 |
4 |
3 |
0 |
10 |
| If you're in London, catch Rachel Rapp with our friends from Tavily and cognee at Redis Released.
From building and de… |
@baseten |
Company |
Original |
2025-10-08 |
823 |
10 |
0 |
2 |
1 |
1 |
| Fast models for our fast friends at Factory! |
@baseten |
Company |
Quote |
2025-10-07 |
2,988 |
25 |
1 |
3 |
0 |
0 |
| Being fast for one customer isn't enough.
Low-latency inference at scale requires the ability to recruit every GPU in … |
@baseten |
Company |
Original |
2025-10-07 |
939 |
11 |
4 |
2 |
0 |
0 |
| Video generation workloads are long-running, have high memory consumption, and require stable scaling.
Using the Base… |
@baseten |
Company |
Quote |
2025-10-06 |
971 |
11 |
2 |
1 |
0 |
1 |
| Today we wrapped up our latest company-wide offsite in beautiful San Diego.
From nearly 20 hackathon projects (soon to… |
@baseten |
Company |
Original |
2025-10-03 |
827 |
18 |
0 |
1 |
0 |
0 |
| Embeddings power search, RecSys, and agents, but making them performant in production requires satisfying two different… |
@baseten |
Company |
Original |
2025-10-02 |
549 |
8 |
1 |
1 |
0 |
0 |
| From document processing and image recognition to drug discovery, healthcare use cases are at the forefront of AI adopt… |
@baseten |
Company |
Original |
2025-10-01 |
572 |
10 |
0 |
1 |
0 |
1 |
| We’re thrilled to welcome four new team members: William Gao, Corina Fitzgerald Wagle, Joseph Sandler, and Christopher … |
@baseten |
Company |
Original |
2025-09-30 |
695 |
6 |
0 |
0 |
0 |
0 |
| We’re hosting our friends at @openrouter for a SF Tech Week breakfast talk!
Join us at Baseten HQ on October 8 at 10AM… |
@baseten |
Company |
Original |
2025-09-29 |
733 |
7 |
1 |
2 |
0 |
1 |
| What do Superhuman, Baseten, and Ricky Bobby all have in common?
An obsession with speed. If you’re a Superhuman user,… |
@baseten |
Company |
Original |
2025-09-26 |
971 |
13 |
1 |
1 |
0 |
2 |
| Baseten is growing! If you're looking for your next opportunity, take a look at our nearly 40 open roles across enginee… |
@baseten |
Company |
Original |
2025-09-25 |
972 |
11 |
0 |
1 |
0 |
1 |
| Join us at @Get_Writer HQ on Tuesday, October 7 for a happy hour during SF Tech Week!
We’re teaming up with Writer, Ro… |
@baseten |
Company |
Original |
2025-09-24 |
693 |
6 |
1 |
1 |
0 |
0 |
| "If you look at the Zed codebase, there's a ton of handwritten custom code—shaders written in languages like Metal... T… |
@baseten |
Company |
Original |
2025-09-23 |
30,282 |
58 |
7 |
3 |
5 |
16 |
| We’ll be at SigSum SF this Thursday, Sept 25!
Catch:
- @philip_kiely's talk "Inference Engineering for Hypergrowth" (1… |
@baseten |
Company |
Original |
2025-09-22 |
644 |
6 |
0 |
1 |
0 |
1 |
| Don’t miss our Head of Infra, Colin McGrath, on the AI Meets Reliability panel in San Francisco on September 23!
He's… |
@baseten |
Company |
Original |
2025-09-19 |
754 |
7 |
0 |
1 |
0 |
1 |
| We’re excited to welcome four new team members: Christopher Frost, Eskil Olsen, Nidhi Hiremath, and Jimmy Whitaker.
Es… |
@baseten |
Company |
Original |
2025-09-18 |
629 |
7 |
0 |
0 |
2 |
0 |
| If you see a doctor today, chances are they're using OpenEvidence for trustworthy, up-to-date medical information at th… |
@baseten |
Company |
Original |
2025-09-17 |
3,455 |
30 |
9 |
3 |
1 |
5 |
| “The key is having good intuition, being willing to go out on a limb, building fast, learning fast, and killing things … |
@baseten |
Company |
Original |
2025-09-16 |
1,199 |
16 |
4 |
1 |
0 |
4 |
| Qwen3 Next 80B A3B Thinking outperforms higher-cost and closed models like Gemini 2.5 Flash Thinking on benchmarks, nea… |
@baseten |
Company |
Original |
2025-09-15 |
5,048 |
16 |
1 |
3 |
2 |
2 |
| Tonight we’re at the AWS Builder Loft for AI Infra Night.
Our Head of Infrastructure, Colin McGrath, will be on stage … |
@baseten |
Company |
Original |
2025-09-11 |
780 |
6 |
0 |
3 |
0 |
2 |
| When products have the potential to help patients, speed and reliability matter.
Latent provides the largest health s… |
@baseten |
Company |
Original |
2025-09-10 |
12,009 |
19 |
2 |
1 |
1 |
1 |
| In 90 days, we’ve increased our daily inference requests 8x. And we’re not slowing down.
Our Head of Infra, Colin McGr… |
@baseten |
Company |
Original |
2025-09-09 |
645 |
6 |
0 |
1 |
1 |
0 |
| We just raised a $150M Series D, and we’re growing!
If you're looking for your next opportunity, take a look at our 30… |
@baseten |
Company |
Original |
2025-09-08 |
706 |
8 |
0 |
2 |
0 |
1 |
| AI everywhere = Inference everywhere = Baseten everywhere |
@baseten |
Company |
Quote |
2025-09-05 |
7,588 |
13 |
3 |
2 |
0 |
2 |
| Come see @tuhinone live at 1:50pm (PT) - don't miss it! |
@baseten |
Company |
Quote |
2025-09-05 |
3,520 |
11 |
0 |
0 |
0 |
0 |
| We raised a $150M Series D! Thank you to all of our customers who trust us to power their inference.
We're grateful t… |
@baseten |
Company |
Quote |
2025-09-05 |
36,756 |
72 |
7 |
3 |
3 |
29 |
| We're excited to be a day 0 partner for EmbeddingGemma, Google’s new open-source embedding model!
You can deploy it di… |
@baseten |
Company |
Quote |
2025-09-04 |
1,988 |
23 |
2 |
2 |
0 |
3 |
| We're thrilled to be close partners with the NVIDIA team.
We use the latest accelerated compute, tools like Dynamo and… |
@baseten |
Company |
Quote |
2025-09-04 |
1,684 |
31 |
2 |
1 |
0 |
2 |
| Today's the day: Join our lead DevRel Philip Kiely and Agustín Bernardo, Senior AI Engineer at @Superhuman, live today … |
@baseten |
Company |
Original |
2025-09-04 |
867 |
9 |
2 |
1 |
0 |
1 |
| We're excited to announce Fall into Inference: a multi-month deep dive into our cloud ecosystem and how we use Multi-cl… |
@baseten |
Company |
Original |
2025-09-03 |
2,302 |
30 |
7 |
6 |
1 |
0 |
| Join us for our first ever SF Tech Breakfast with @MorganBarrettX on Wednesday, September 10th!
If you're an AI Leader… |
@baseten |
Company |
Original |
2025-09-02 |
2,311 |
12 |
1 |
2 |
1 |
2 |
| Our team met Parsed a few months ago, and we could not be more excited to see the inflection point they are a part of -… |
@baseten |
Company |
Quote |
2025-08-28 |
3,946 |
26 |
3 |
0 |
0 |
0 |
| Welcome to Baseten @DannieHerz!
We’re thrilled to announce that Dannie Herzberg has joined as our new President to lea… |
@baseten |
Company |
Original |
2025-08-27 |
97,746 |
104 |
8 |
15 |
9 |
17 |
| DeepSeek V3 + R1 (now V3.1 and R1 0528) proved that open-source models can rival closed ones at lower cost and with gre… |
@baseten |
Company |
Original |
2025-08-27 |
2,215 |
18 |
2 |
2 |
1 |
1 |
| Embedding models are powering the next wave of AI products, from search to personalized experiences. But how do you mov… |
@baseten |
Company |
Original |
2025-08-25 |
2,196 |
23 |
3 |
3 |
3 |
0 |
| Join us at the AWS Builder Loft on Sept. 11 for AI Infra Night!
Our Head of Infrastructure, Colin McGrath, is presenti… |
@baseten |
Company |
Original |
2025-08-22 |
606 |
11 |
0 |
2 |
0 |
2 |
| DeepSeek v3.1 is live on our Model APIs! https://t.co/OF70dcIdwd |
@baseten |
Company |
Original |
2025-08-22 |
1,519 |
21 |
4 |
2 |
0 |
0 |
| No spoilers on what's going on under the hood, but let's say we're looking forward to Oxen's Fine-Tune Friday tomorrow 🐄 |
@baseten |
Company |
Quote |
2025-08-21 |
1,092 |
14 |
1 |
0 |
0 |
0 |
| We’re thrilled to welcome four new team members: Shounak Ray, Matte Lim, Xiaoyi Yu, and Fred Liu.
Shounak joins our Pe… |
@baseten |
Company |
Original |
2025-08-21 |
635 |
8 |
0 |
0 |
0 |
2 |
| Our first ever Inference Invitational is tomorrow, 8/21!
Join us for friendly pickleball & padel tournaments, grea… |
@baseten |
Company |
Original |
2025-08-20 |
654 |
10 |
0 |
3 |
0 |
0 |
| Want to fine-tune gpt-oss-120b?
We teamed up with Axolotl to launch a new recipe to run fine-tuning out of the box — … |
@baseten |
Company |
Original |
2025-08-19 |
5,078 |
20 |
7 |
3 |
1 |
7 |
| When it comes to open-source text-to-speech, Orpheus should be your go-to model.
We're seeing so much demand around r… |
@baseten |
Company |
Original |
2025-08-18 |
1,080 |
14 |
2 |
2 |
0 |
2 |
| Qwen 3 instruct is now on Baseten Model APIs.
Our model performance team has worked quite a bit of magic to reach ~95t… |
@baseten |
Company |
Original |
2025-08-15 |
896 |
6 |
0 |
1 |
0 |
0 |
| When AI education needs to feel human, latency matters.
Praktika hit <300 ms transcription with 50% cost savings… |
@baseten |
Company |
Original |
2025-08-14 |
756 |
11 |
2 |
1 |
0 |
0 |
| If you know Baseten, you know we love a good code gen use case. We're excited to deepen our IDE ties with our integrati… |
@baseten |
Company |
Quote |
2025-08-13 |
2,345 |
22 |
1 |
4 |
0 |
1 |
| Day 2 of Ai4 Vegas is here!
Did you catch our lead DevRel Philip Kiely’s talk "Inference in the Wild: Lessons from S… |
@baseten |
Company |
Original |
2025-08-13 |
603 |
6 |
0 |
2 |
0 |
1 |
| The Qwen bug is here, and we’ve caught it. With the release of use case specific models in July, Qwen users can now sel… |
@baseten |
Company |
Original |
2025-08-12 |
755 |
10 |
0 |
2 |
0 |
0 |
| We're thrilled to welcome Joey Zwicker as our new Head of Forward Deployed Engineering!
We've grown rapidly over the l… |
@baseten |
Company |
Original |
2025-08-11 |
10,343 |
32 |
2 |
5 |
3 |
2 |
| If you're at Ai4 in Vegas next week, don't miss Philip Kiely's talk "Inference in the Wild: Lessons from Scaling Real-t… |
@baseten |
Company |
Original |
2025-08-08 |
1,071 |
5 |
1 |
1 |
1 |
0 |
| "The reality is that for each customer it’s a different story. Migrating to a new model isn’t a small effort, there are… |
@baseten |
Company |
Original |
2025-08-07 |
6,427 |
19 |
5 |
3 |
2 |
1 |
| In the time from when this blog was initially written, to it hitting #1 on @hackernews, to this tweet being written, GP… |
@baseten |
Company |
Quote |
2025-08-07 |
4,810 |
10 |
1 |
1 |
1 |
0 |
| Our model performance team has been working... |
@baseten |
Company |
Quote |
2025-08-06 |
931 |
18 |
1 |
0 |
0 |
2 |
| New projects already being built on GPT OSS!
Build your own with our Model APIs here
-> https://t.co/CjWiopF4qo |
@baseten |
Company |
Quote |
2025-08-05 |
3,799 |
15 |
2 |
0 |
0 |
4 |
| We're excited to be an OpenAI launch partner for the release of GPT OSS 120B and 20B!
Model APIs coming shortly, with … |
@baseten |
Company |
Original |
2025-08-05 |
8,399 |
50 |
8 |
2 |
2 |
3 |
| TEI doesn't run on B200s — but BEI does. BEI achieves 3.6x higher embeddings throughput than TEI and 3.3x that of vLLM … |
@baseten |
Company |
Original |
2025-08-04 |
1,214 |
21 |
5 |
2 |
0 |
2 |
| Pickleball fan? Meet us in SF on August 21st for pickleball and padel and some great food and drinks!
Join our tournam… |
@baseten |
Company |
Original |
2025-08-01 |
1,536 |
11 |
3 |
2 |
0 |
3 |
| Congrats to the @producer_ai team on this massive launch! |
@baseten |
Company |
Quote |
2025-07-31 |
1,217 |
13 |
0 |
0 |
1 |
0 |
| We're thrilled to be working with you and the entire @Sourcegraph team! |
@baseten |
Company |
Quote |
2025-07-31 |
2,385 |
31 |
1 |
0 |
1 |
1 |
| Deep Cogito just dropped 4 new open LLMs -- each one is SOTA for its size.
We're excited to be a launch partner for th… |
@baseten |
Company |
Quote |
2025-07-31 |
769 |
9 |
0 |
1 |
1 |
1 |
| Building reliable agents requires a different tech stack: one that natively supports compound AI systems and evaluates … |
@baseten |
Company |
Original |
2025-07-29 |
1,194 |
18 |
3 |
2 |
0 |
2 |
| Tutorial: Transcribe audio in real-time with Whisper Large V3 and WebSockets.
WebSockets are persistent, bidirectional… |
@baseten |
Company |
Original |
2025-07-28 |
1,056 |
15 |
2 |
3 |
0 |
6 |
| When @zeddotdev set out to build Edit Prediction, they knew they wanted it to feel instantaneous. But their previous in… |
@baseten |
Company |
Original |
2025-07-25 |
1,214 |
12 |
5 |
2 |
0 |
0 |
| Forget AI writing your code. AI can now control your home through voice.
We’ve had a blast putting Voxtral through the… |
@baseten |
Company |
Original |
2025-07-24 |
1,795 |
16 |
2 |
5 |
0 |
3 |
| Only a handful of models dominated the ASR space—until now.
Voxtral has a 30-minute transcription range, a 40-minute … |
@baseten |
Company |
Original |
2025-07-23 |
1,363 |
11 |
2 |
1 |
0 |
3 |
| Another week, another model drop! Voxtral was released last week and you can now deploy it on Baseten.
Transcription … |
@baseten |
Company |
Original |
2025-07-22 |
1,287 |
15 |
3 |
3 |
0 |
2 |
| We're thrilled to welcome four new team members: Marcel Chacon, Yana Chen, Alex Ker, and Yikai Zhu!
Marcel and Yana ar… |
@baseten |
Company |
Original |
2025-07-21 |
2,466 |
16 |
2 |
0 |
0 |
2 |
| Wondering why Kimi K2 is all the buzz this week? Or, looking for a place to test it?
Check out our blog on Kimi K2—we … |
@baseten |
Company |
Original |
2025-07-19 |
2,590 |
9 |
0 |
1 |
0 |
1 |
| Baseten is growing! We’re always looking for determined, humble people to join our team.
Catch us at the Greylock Tec… |
@baseten |
Company |
Original |
2025-07-17 |
1,081 |
6 |
0 |
1 |
0 |
0 |
| Kimi K2 has arrived. You can deploy it on Baseten.
Join as we briefly dig into why K2 is generating so much buzz.
If … |
@baseten |
Company |
Original |
2025-07-16 |
1,031 |
12 |
0 |
2 |
0 |
0 |
| Confession. Kimi K2 is one of our new favorite models for agentic use cases.
Baseten is powering the fastest Kimi K2 a… |
@baseten |
Company |
Original |
2025-07-15 |
7,983 |
48 |
5 |
4 |
3 |
12 |
| 👀 |
@baseten |
Company |
Quote |
2025-07-15 |
975 |
9 |
0 |
1 |
0 |
0 |
| We're excited to welcome four new team members: Deepak Nagaraj, Aghilan Nathan, Madison Kanna, and Jose Roman (Jojo) Or… |
@baseten |
Company |
Original |
2025-07-14 |
838 |
10 |
0 |
0 |
0 |
1 |
| Excited to partner with @ZeroEntropy_AI to power their new state-of-the-art, open-source reranker model, zerank! |
@baseten |
Company |
Quote |
2025-07-14 |
1,353 |
9 |
1 |
0 |
0 |
1 |
| Join us at our office for Founders Friday on July 18!
We’re hosting a morning of great conversation, strong coffee, an… |
@baseten |
Company |
Original |
2025-07-11 |
1,082 |
3 |
0 |
1 |
0 |
0 |
| Speed isn’t optional—it’s survival. IVP’s Shravan puts it plainly:
“Anything you can do to offload non-core activiti… |
@baseten |
Company |
Original |
2025-07-02 |
479 |
5 |
1 |
1 |
0 |
0 |
| Catch our CEO Tuhin on the keynote stage at RAISE Paris next week! He’ll cover the AI landscape, the rise of open-sourc… |
@baseten |
Company |
Original |
2025-07-01 |
553 |
3 |
0 |
0 |
0 |
2 |
| We’re thrilled to welcome four new team members: Marylise Tauzia, Kaushik Chatterjee, Shivain Vij, and Justin Schmitt!
… |
@baseten |
Company |
Original |
2025-06-30 |
534 |
7 |
0 |
0 |
0 |
0 |
| Baseten is growing! If you're looking for your next opportunity, take a look at our nearly 30 open roles across enginee… |
@baseten |
Company |
Original |
2025-06-26 |
1,264 |
11 |
0 |
1 |
0 |
2 |
| Friendly reminder from @willreed (Spark Capital): Your team's time is best spent on your product, not the infrastructur… |
@baseten |
Company |
Original |
2025-06-25 |
714 |
3 |
0 |
1 |
0 |
1 |
| Have you checked out the views from our new office?
Join us here on July 18th for our first Founders Friday. Enjoy a… |
@baseten |
Company |
Original |
2025-06-24 |
780 |
6 |
1 |
2 |
0 |
0 |
| If your workload goes down, so does your product.
Workloads running on a single cloud are constrained — cloud provider… |
@baseten |
Company |
Original |
2025-06-23 |
1,160 |
11 |
2 |
2 |
0 |
1 |
| Catch @tuhinone on the keynote stage at RAISE Paris on July 9th!
And grab some swag from Tuhin and the Baseten team at… |
@baseten |
Company |
Original |
2025-06-19 |
727 |
5 |
0 |
0 |
0 |
0 |
| "Inference is more than just vibes" — catch us on the Bay Bridge! https://t.co/JbjtgOSkli |
@baseten |
Company |
Original |
2025-06-17 |
1,356 |
16 |
1 |
0 |
1 |
0 |
| A huge welcome to four of our new team members: Beverly Rivas, Ke Bao, Saptarshi Bhattacherya, and Tri Dao!
Beverly jo… |
@baseten |
Company |
Original |
2025-06-17 |
1,685 |
21 |
1 |
1 |
1 |
0 |
| Hot take: “Inference is not a commodity. There is strong complexity and differences between inference providers.”
Coul… |
@baseten |
Company |
Original |
2025-06-16 |
673 |
8 |
0 |
1 |
0 |
4 |
| We're excited to introduce the Baseten Performance Client, a new open-source Python library for up to 12x higher throug… |
@baseten |
Company |
Original |
2025-06-13 |
2,961 |
23 |
6 |
4 |
3 |
4 |
| green is our brand color for a reason |
@baseten |
Company |
Quote |
2025-06-12 |
2,057 |
22 |
2 |
0 |
0 |
0 |
| We're excited to be on the InfraRed 100 list again this year! And we're in great company with so many of our customers … |
@baseten |
Company |
Original |
2025-06-11 |
520 |
2 |
0 |
1 |
0 |
0 |
| Forward deployed engineers (FDEs) are core to our company. They work directly with customers, contribute to product dev… |
@baseten |
Company |
Original |
2025-06-10 |
1,149 |
23 |
3 |
2 |
0 |
7 |
| We're at the AWS Summit in Washington, D.C. today, come meet Andrew, Kerrick, and Danny at the booth!
Rumor has it we… |
@baseten |
Company |
Original |
2025-06-10 |
572 |
5 |
0 |
1 |
0 |
0 |
| Our customers run AI products where every millisecond and request matter.
Over the years, we found fundamental limitat… |
@baseten |
Company |
Original |
2025-06-09 |
6,510 |
20 |
6 |
2 |
1 |
2 |
| If you’re in San Francisco tomorrow, don’t miss our lunch workshop on unlocking open-source models in production!
Enjo… |
@baseten |
Company |
Original |
2025-06-09 |
629 |
5 |
0 |
2 |
0 |
0 |
| Two events, three talks, a happy hour with @oxen_ai, and a ton of merch — thank you to everyone who came to see us at t… |
@baseten |
Company |
Original |
2025-06-06 |
2,275 |
11 |
3 |
0 |
1 |
2 |
| We're psyched for our San Francisco AI Breakfast on Tuesday, June 10th!
If you're a ML Engineering Leader in SF, come … |
@baseten |
Company |
Original |
2025-06-05 |
561 |
4 |
0 |
1 |
0 |
0 |
| Join us for breakfast, lunch, or both on June 10th in San Francisco!
AI Breakfast (Tuesday, June 10 - 9AM)
If you’re … |
@baseten |
Company |
Original |
2025-06-04 |
541 |
4 |
0 |
1 |
0 |
1 |
| We’re excited to partner with @oxen_ai on their fine-tuning launch. It’s almost too easy — zero-code fine-tuning, from … |
@baseten |
Company |
Quote |
2025-06-03 |
3,163 |
19 |
9 |
2 |
0 |
3 |
| Meet us at AI Engineer World’s Fair this week! We'll be at booth G5. Get a demo from our engineers and grab some Artifi… |
@baseten |
Company |
Original |
2025-06-02 |
1,336 |
4 |
1 |
1 |
0 |
1 |
| Impressed by these ultra-realistic, multilingual AI actors — a huge unlock for creative teams scaling content.
Congra… |
@baseten |
Company |
Quote |
2025-06-02 |
4,240 |
28 |
3 |
1 |
1 |
4 |
| “People think that GPUs + vLLM = production grade inference. We know that to not be true. With this you can get 80% of … |
@baseten |
Company |
Original |
2025-05-31 |
1,243 |
13 |
2 |
2 |
0 |
6 |
| UPDATE: We crossed 100 tokens per second. https://t.co/Mr6jHZhuOE |
@baseten |
Company |
Quote |
2025-05-29 |
1,300 |
13 |
0 |
2 |
0 |
0 |
| New DeepSeek just dropped.
Proud to serve the fastest DeepSeek R1 0528 inference on OpenRouter (#1 on TTFT and TPS) wi… |
@baseten |
Company |
Quote |
2025-05-29 |
4,410 |
19 |
9 |
4 |
2 |
2 |
| Congrats to our friends at @retool! Agents are a game-changer for automating repetitive tasks (and Retool has automated… |
@baseten |
Company |
Quote |
2025-05-28 |
1,152 |
13 |
1 |
0 |
0 |
1 |
| Our secret sauce? The Baseten Inference Stack.
It consists of two core layers: the Inference Runtime and Inference-op… |
@baseten |
Company |
Original |
2025-05-27 |
6,277 |
27 |
6 |
4 |
1 |
6 |
| We launched a bunch of stuff - come see us and talk about it IRL:
• June 3-5 - AI Engineer World’s Fair (SF)
• June 4 … |
@baseten |
Company |
Original |
2025-05-23 |
1,134 |
12 |
3 |
3 |
0 |
0 |
| let there be inference https://t.co/OxinWizB6d |
@baseten |
Company |
Original |
2025-05-22 |
332,417 |
356 |
48 |
54 |
23 |
114 |
| 🚀 Our "technical" marketer might not be looped in, but today is our biggest launch day yet.
We're introducing two new … |
@baseten |
Company |
Original |
2025-05-21 |
35,344 |
90 |
19 |
7 |
4 |
26 |
| Inference is everywhere.
Come find us in San Francisco! |
@baseten |
Company |
Original |
2025-05-20 |
1,690 |
24 |
2 |
1 |
0 |
2 |
| 🚀 We've been heads down for months, and now it's finally launch week.
Today, we’re releasing our new brand. We believe… |
@baseten |
Company |
Original |
2025-05-19 |
7,294 |
65 |
10 |
4 |
2 |
9 |
| An American, an Aussie, and a founder all walk into the Louvre...
(Plot twist: it's the same person)
Grab a croissant… |
@baseten |
Company |
Original |
2025-05-16 |
798 |
7 |
0 |
2 |
1 |
0 |
| Your AI app is too slow. At least Sarah Guo thinks so.
Your customers deserve better—but it's not your fault. Baseten … |
@baseten |
Company |
Original |
2025-05-16 |
813 |
11 |
1 |
1 |
0 |
5 |
| Congrats to our friends at Patronus AI on the new AI agent launch, Percival!
Percival can fix other agents across 20+… |
@baseten |
Company |
Quote |
2025-05-14 |
1,025 |
14 |
2 |
0 |
0 |
0 |
| If you’re going to the AI Engineer World’s Fair in San Francisco, don’t miss our CTO Amir’s talk “The Rise of Open Mode… |
@baseten |
Company |
Original |
2025-05-14 |
422 |
3 |
0 |
1 |
0 |
0 |
| Founder walks into an investor meeting… “So… what should we talk about?”
We may freestyle our meetings, but we know wh… |
@baseten |
Company |
Original |
2025-05-13 |
577 |
6 |
1 |
1 |
1 |
1 |
| Join us in San Francisco on Wednesday, May 21st for Inference on Tap!
Drink beer and wine, dig into tapas, enjoy great… |
@baseten |
Company |
Original |
2025-05-12 |
420 |
5 |
0 |
1 |
0 |
0 |
| "The best companies in AI don't care about what model is under the hood..."
Check out our CEO @tuhinone's interview wi… |
@baseten |
Company |
Original |
2025-05-09 |
1,419 |
7 |
2 |
0 |
0 |
2 |
| We're excited to welcome two new team members: Brian Hartrick and Phillippe Siclait!
Phillippe joins our team of Model… |
@baseten |
Company |
Original |
2025-05-08 |
417 |
10 |
0 |
2 |
0 |
0 |
| We’re thrilled to partner with Canopy Labs to offer production-grade real-time inference for Orpheus TTS! |
@baseten |
Company |
Quote |
2025-05-06 |
2,674 |
25 |
3 |
2 |
1 |
1 |
| If you're an AI leader living in NYC, come join us this Wednesday (5/7) for conversation about model deployment infrast… |
@baseten |
Company |
Original |
2025-05-05 |
799 |
3 |
1 |
1 |
0 |
0 |
| Get started with the new image-to-image model by @FotographerAI with a one-click deploy on Baseten. |
@baseten |
Company |
Quote |
2025-05-04 |
3,260 |
10 |
2 |
1 |
3 |
3 |
| If you're in NYC on Thursday, May 8th, don't miss our Voice AI Panel for ML Engineers!
Join us for an afternoon panel … |
@baseten |
Company |
Original |
2025-05-02 |
311 |
3 |
0 |
2 |
0 |
0 |
| Early benchmarks of Qwen 3 with SGLang show promising initial results and key avenues for improvement.
We're seeing:
-… |
@baseten |
Company |
Original |
2025-04-30 |
1,877 |
8 |
1 |
1 |
1 |
2 |
| "This is the thing about AI — you gotta burn the boats.”
Our CEO Tuhin Srivastava sat down with Emma Cosgrove and the… |
@baseten |
Company |
Original |
2025-04-29 |
618 |
7 |
3 |
1 |
1 |
0 |
| We have day 0 support for #Qwen3 by Alibaba Qwen on Baseten using SGLang.
Qwen 3 235B's architecture benefits from bot… |
@baseten |
Company |
Original |
2025-04-28 |
14,986 |
46 |
12 |
4 |
5 |
17 |
| We've seen a ton of Llama 4 usage in the space, especially with Scout's full 10M context. We're thrilled to be joining … |
@baseten |
Company |
Original |
2025-04-28 |
398 |
5 |
0 |
1 |
0 |
0 |
| Back by popular demand: We’ve teamed up with NYC Tech Breakfast again to host a breakfast for CTOs and AI engineers in … |
@baseten |
Company |
Original |
2025-04-25 |
364 |
3 |
0 |
1 |
0 |
0 |
| Congrats to our friends at @rimelabs on the new text-to-speech model drop!
It’s very realistic, capturing nuances of h… |
@baseten |
Company |
Quote |
2025-04-24 |
2,198 |
17 |
8 |
1 |
1 |
3 |
| We're thrilled to welcome Mahmoud Hassan and Tammie Vu to our team! 🎊
Tammie is bringing her expertise to our people o… |
@baseten |
Company |
Original |
2025-04-23 |
340 |
5 |
0 |
0 |
0 |
0 |
| We're just 1 week out from AWS Summit London.
If you plan to be at the ExCeL London on Wednesday, April 30, come say h… |
@baseten |
Company |
Original |
2025-04-22 |
384 |
4 |
0 |
1 |
0 |
0 |
| Baseten is hiring!
We’re always looking for motivated, humble people to join our team. We have 24 open roles across E… |
@baseten |
Company |
Original |
2025-04-21 |
296 |
4 |
0 |
1 |
0 |
1 |
| We’ve seen a lot of interest in B200s after our launch.
Our lead DevRel, @philip_kiely, wrote a blog explaining some o… |
@baseten |
Company |
Original |
2025-04-18 |
482 |
6 |
2 |
1 |
0 |
1 |
| A huge welcome to two of our new team members: Ashlee Flagg and Yizhou Guo! 🎉
Yizhou and Ashlee are joining the team i… |
@baseten |
Company |
Original |
2025-04-17 |
340 |
4 |
0 |
0 |
0 |
0 |
| 🚀 You can now use NVIDIA B200s on Baseten and get higher throughput, lower latency, and better cost per token! 🚀
From … |
@baseten |
Company |
Original |
2025-04-15 |
2,921 |
15 |
6 |
9 |
2 |
2 |
| Thanks to everyone who stopped by to say hi at Google Cloud Next in Vegas last week! If you didn’t catch us there, we’r… |
@baseten |
Company |
Original |
2025-04-14 |
535 |
5 |
0 |
1 |
0 |
0 |
| You can now use embedding models on Baseten as part of @trychroma's Python SDK!
Check out the guide by @philip_kiely … |
@baseten |
Company |
Original |
2025-04-11 |
3,739 |
12 |
3 |
2 |
2 |
3 |
| If you're at Google Next, don't miss Bola Malek's talk "Secure and Optimize AI and ML Workloads with the Cross-Cloud Ne… |
@baseten |
Company |
Original |
2025-04-10 |
861 |
5 |
1 |
1 |
0 |
0 |
| We're thrilled to be included in the #ForbesAI50! 🎉
Congratulations to everyone who made it, it's great to see so many… |
@baseten |
Company |
Quote |
2025-04-10 |
1,999 |
19 |
5 |
3 |
2 |
0 |
| It's Day 1 of Google Cloud Next! If you're attending, stop by the booth (3341) for some ice cream and swag.
Say hi to … |
@baseten |
Company |
Original |
2025-04-09 |
651 |
9 |
0 |
2 |
0 |
1 |
| New bots for Llama 4 Scout and Maverick are now live on Poe! Get started with an 8M token context window for Scout (yes… |
@baseten |
Company |
Original |
2025-04-07 |
966 |
14 |
4 |
1 |
1 |
0 |
| Llama 4 is here! 🦙🚀
Scout | 109B Parameters | 10M Context
Maverick | 400B Parameters | 1M Context
Llama 4 mode… |
@baseten |
Company |
Original |
2025-04-05 |
5,385 |
27 |
9 |
6 |
1 |
0 |
| Thanks to everyone at Kubecon London who swung by yesterday to chat with us.
If you haven't had time to talk to the te… |
@baseten |
Company |
Original |
2025-04-03 |
339 |
5 |
1 |
0 |
0 |
0 |
| Meet up with the Baseten team at KubeCon London tomorrow!
Swing by booth # N651 for demos, Baseten cupcakes, an "Artif… |
@baseten |
Company |
Original |
2025-04-01 |
274 |
3 |
0 |
1 |
0 |
0 |
| The first Baseten bot is live on Poe! It's very fast, you can ask questions in your language of choice and get instant … |
@baseten |
Company |
Original |
2025-03-28 |
6,337 |
46 |
9 |
3 |
1 |
2 |
| 🚀 We’re thrilled to introduce Baseten Embeddings Inference (BEI), the most performant embeddings solution available! 🚀
… |
@baseten |
Company |
Original |
2025-03-27 |
1,669 |
22 |
6 |
3 |
1 |
5 |
| Thank you @Wing_VC and @EricNewcomer for recognizing Baseten in the Enterprise Tech 30!
Shout out to our incredible c… |
@baseten |
Company |
Quote |
2025-03-25 |
656 |
12 |
1 |
1 |
0 |
0 |
| Thanks to everyone who came by to see us at @NVIDIAGTC! If you didn't catch us then (or even if you did), you can meet … |
@baseten |
Company |
Original |
2025-03-24 |
405 |
7 |
0 |
1 |
1 |
0 |
| Baseten is growing! If you're looking for your next opportunity, take a look at our 18 open roles across engineering an… |
@baseten |
Company |
Original |
2025-03-21 |
1,092 |
4 |
2 |
1 |
1 |
1 |
| Thanks to everyone who came to @defpan and @philip_kiely's talk on inference optimization yesterday—we had a full house… |
@baseten |
Company |
Original |
2025-03-20 |
423 |
5 |
1 |
0 |
0 |
1 |
| GTC day one was a blast. If you haven't caught us yet, stop by the booth (#1233) for some Baseten Blend coffee, or watc… |
@baseten |
Company |
Original |
2025-03-19 |
426 |
3 |
0 |
1 |
0 |
0 |
| We're excited to announce our partnership with @nvidia to provide inference for NVIDIA NIM models on dedicated endpoint… |
@baseten |
Company |
Quote |
2025-03-18 |
9,291 |
28 |
9 |
4 |
1 |
4 |
| Meet up with the Baseten team at @NVIDIAGTC! Swing by booth #1233 for demos, Baseten Blend coffee beans, coffee with ou… |
@baseten |
Company |
Original |
2025-03-14 |
340 |
5 |
0 |
1 |
0 |
0 |
| Gemma 3 just dropped from @GoogleAI. 🔥 If you want to try it out, you can deploy it in two clicks from our model librar… |
@baseten |
Company |
Original |
2025-03-12 |
1,409 |
15 |
4 |
2 |
1 |
1 |
| If you're at #HumanX2025, @philip_kiely and Andy are ready for you at booth #701. Stop by and grab some swag! https://t… |
@baseten |
Company |
Original |
2025-03-11 |
347 |
4 |
0 |
0 |
0 |
0 |
| Meet the Baseten crew at #HumanX!
Visit us at booth #701 and chat with @tuhinone, @philip_kiely, @mj_bilodeau, and And… |
@baseten |
Company |
Original |
2025-03-10 |
495 |
11 |
0 |
1 |
0 |
0 |
| You can start using the new @Alibaba_Qwen Qwen QwQ-32B from our model library in two clicks 👇 https://t.co/uV4y19pucA |
@baseten |
Company |
Original |
2025-03-06 |
1,192 |
7 |
2 |
1 |
0 |
1 |
| We're celebrating our first podcast episode by streaming it live this Thursday, March 6th at 11 a.m. PT!
Join our host… |
@baseten |
Company |
Original |
2025-03-04 |
1,117 |
4 |
0 |
1 |
1 |
0 |
| "We want to work with our customers. We enjoy working with our customers. The thing you will hear time and time again i… |
@baseten |
Company |
Original |
2025-03-03 |
1,029 |
9 |
2 |
1 |
0 |
0 |
| Nobody knows what inference means but it's provocative https://t.co/hsfGdkhEv4 |
@baseten |
Company |
Original |
2025-02-28 |
497,400 |
981 |
82 |
92 |
46 |
491 |
| Better model efficiency is leading to increased AI adoption, especially in the enterprise. @dee_bosa and the team at @C… |
@baseten |
Company |
Original |
2025-02-27 |
2,058 |
10 |
2 |
1 |
0 |
1 |
| Last week was crazy. Thank you to everyone who celebrated our Series C with us, and met us in person at @aiDotEngineer … |
@baseten |
Company |
Original |
2025-02-24 |
1,904 |
8 |
3 |
1 |
0 |
1 |
| It was great talking with @mims about how some of our customers like @descript and @PicnicHealth are shaping the AI spa… |
@baseten |
Company |
Original |
2025-02-22 |
1,244 |
8 |
2 |
2 |
1 |
0 |
| "In this market, your No. 1 differentiation is how fast you can move. That is the core benefit for our customers. You c… |
@baseten |
Company |
Original |
2025-02-20 |
1,330 |
22 |
4 |
1 |
0 |
0 |
| 2025 is the year of inference.
We're thrilled to announce our $75m Series C co-led by @IVP and @sparkcapital with pa… |
@baseten |
Company |
Original |
2025-02-19 |
162,221 |
833 |
130 |
449 |
27 |
79 |
| Meet the Baseten team at the @aiDotEngineer Summit in NYC this week!
📍 Booth G3, get a demo and grab some swag https:/… |
@baseten |
Company |
Original |
2025-02-18 |
917 |
4 |
1 |
0 |
0 |
1 |
| Congrats @zeddotdev on the new open-source model drop! It was a pleasure customizing Zeta's inference performance to hi… |
@baseten |
Company |
Original |
2025-02-14 |
5,169 |
33 |
2 |
0 |
0 |
1 |
| How can you run DeepSeek-R1 on H100s when it doesn’t fit on a single node?
With multi-node inference you can split Dee… |
@baseten |
Company |
Original |
2025-02-13 |
1,156 |
17 |
6 |
1 |
0 |
2 |
| We're psyched to welcome two new team members: @lucas_dehaas and Kenzie Amack 🎉
Kenzie and Luke are joining our market… |
@baseten |
Company |
Original |
2025-02-12 |
717 |
7 |
0 |
0 |
1 |
0 |
| Back by popular demand: Join us for the next NYC Tech Breakfast with @MorganBarrettX on Wednesday, February 19th!
If y… |
@baseten |
Company |
Original |
2025-02-11 |
303 |
2 |
0 |
1 |
0 |
0 |
| "There are big implications from DeepSeek for highly regulated industries. Companies that have strict data compliance r… |
@baseten |
Company |
Quote |
2025-02-10 |
714 |
8 |
1 |
0 |
0 |
0 |
| "We've been working closely with the DeepSeek AI and SGLang teams for months to get these models running well." - @tuhi… |
@baseten |
Company |
Original |
2025-02-08 |
401 |
3 |
0 |
1 |
0 |
0 |
| What LLM inference optimizations work best on the
@nvidia GH200? It turns out KV cache reuse benefits from the high CPU… |
@baseten |
Company |
Original |
2025-02-07 |
389 |
10 |
0 |
2 |
0 |
1 |
| 🚀 We’re thrilled to announce that Baseten Chains is now GA for production compound AI! 🚀
In 5 years, every app will ha… |
@baseten |
Company |
Original |
2025-02-06 |
1,128 |
16 |
4 |
1 |
1 |
1 |
| Ever wanted to spin up an AI-powered game or tool without wrangling complex code? Now you can in our DeepSeek playgroun… |
@baseten |
Company |
Original |
2025-02-04 |
530 |
6 |
0 |
1 |
0 |
0 |
| It was great chatting with @dee_bosa today for @CNBC.
More to come soon! https://t.co/PfqdJqdmgN |
@baseten |
Company |
Original |
2025-02-03 |
1,874 |
34 |
5 |
1 |
1 |
3 |
| If you're in SF this Thursday, don't miss our Whisper & Whisky tasting event for AI experts in the Bay! 🥃
Enjoy Ja… |
@baseten |
Company |
Original |
2025-02-03 |
330 |
7 |
0 |
1 |
0 |
0 |
| Have you tried DeepSeek-R1 yet? You can run it in our DeepSeek playground—no setup required. https://t.co/0aTMLKP0kU |
@baseten |
Company |
Original |
2025-02-02 |
590 |
7 |
0 |
1 |
1 |
1 |
| "H200s are the only widely available Nvidia chip that can run the full DeepSeek models on a single node.
You can also… |
@baseten |
Company |
Original |
2025-02-01 |
1,439 |
17 |
5 |
2 |
0 |
1 |
| Thanks @dee_bosa and @CNBC for chatting with Baseten CEO @tuhinone about DeepSeek-R1! We've heard from hundreds of comp… |
@baseten |
Company |
Original |
2025-01-31 |
1,500 |
19 |
3 |
1 |
1 |
3 |
| Huge congrats to the @riffusionai team on the launch of their new generative music model!
Try it out 👇 |
@baseten |
Company |
Quote |
2025-01-30 |
640 |
4 |
1 |
0 |
1 |
0 |
| Live today at 11 a.m. PT. Don't miss it. |
@baseten |
Company |
Quote |
2025-01-30 |
942 |
11 |
1 |
0 |
1 |
0 |
| Probably nothing https://t.co/5EgQ9J8PJI |
@baseten |
Company |
Original |
2025-01-30 |
1,964 |
27 |
4 |
4 |
0 |
4 |
| We're happy to be supporting @GregKamradt and @mikeknoop with their work for the @arcprize ! |
@baseten |
Company |
Quote |
2025-01-29 |
3,233 |
18 |
0 |
2 |
0 |
0 |
| happy to help! |
@baseten |
Company |
Quote |
2025-01-29 |
403 |
5 |
0 |
0 |
0 |
1 |
| "Panicking about privacy for DeepSeek only makes sense if you don't run the model in the US or EU. Models can't magical… |
@baseten |
Company |
Original |
2025-01-29 |
11,259 |
142 |
14 |
10 |
5 |
4 |
| DeepSeek-R1 is blowing up right now, but we're not surprised. And not just because we’ve been working closely with @dee… |
@baseten |
Company |
Original |
2025-01-28 |
59,335 |
531 |
58 |
19 |
4 |
22 |
| We're thrilled to welcome Deepak Bhandarkar and Nikhil Narayen to the Baseten team! 🎉
Nikhil is the newest Software En… |
@baseten |
Company |
Original |
2025-01-27 |
1,239 |
9 |
0 |
1 |
0 |
0 |
| Thanks to everyone who joined our NYC Tech Breakfast with @MorganBarrettX—we love a full house! 💚 Shout out to @JulienT… |
@baseten |
Company |
Original |
2025-01-26 |
1,994 |
7 |
1 |
2 |
1 |
0 |
| Baseten is proud to announce our partnership with @AthenaIntell to power their AI employees, enabling organizations wit… |
@baseten |
Company |
Original |
2025-01-24 |
1,528 |
6 |
2 |
1 |
1 |
0 |
| Drink whisky, dig into great food, get a new headshot, and network with the AI community in the Bay at our Whisper &… |
@baseten |
Company |
Original |
2025-01-24 |
653 |
5 |
0 |
1 |
0 |
1 |
| If you read the Technology section of the @nytimes this morning, you might have noticed some familiar names.
Thanks @C… |
@baseten |
Company |
Original |
2025-01-23 |
1,688 |
22 |
2 |
2 |
2 |
2 |
| The open-source community is still reeling from @deepseek_ai's new R1 drop, the new best-in-class reasoning model on pa… |
@baseten |
Company |
Original |
2025-01-22 |
5,301 |
26 |
2 |
2 |
2 |
12 |
| The new DeepSeek-R1 drop by @deepseek_ai isn’t just an open-source model that rivals o1. It’s a massive upgrade to ever… |
@baseten |
Company |
Original |
2025-01-21 |
7,872 |
17 |
3 |
1 |
2 |
5 |
| We're psyched to team up with @MorganBarrettX on the first Tech Breakfast Club of 2025 in NYC!
If you're an AI leade… |
@baseten |
Company |
Original |
2025-01-20 |
1,195 |
4 |
0 |
2 |
1 |
0 |
| Our Co-founder @amiruci and Model Performance Engineer @zhyncs42 sat down with @latentspacepod to dive deep into DeepSe… |
@baseten |
Company |
Quote |
2025-01-19 |
1,754 |
10 |
3 |
0 |
0 |
4 |
| Looking for a new job in 2025? Baseten is hiring! We're looking to fill roles across:
𝗘𝗻𝗴𝗶𝗻𝗲𝗲𝗿𝗶𝗻𝗴: Support, Infra, Mod… |
@baseten |
Company |
Original |
2025-01-17 |
501 |
5 |
0 |
1 |
0 |
0 |
| We're thrilled to welcome @feilsystem and Bryan Zhang to our team! 🎊
Bryan is bringing his expertise to our infrastru… |
@baseten |
Company |
Original |
2025-01-16 |
442 |
7 |
2 |
0 |
0 |
0 |
| In 2024, we pushed the boundaries of highly performant, scalable, and reliable production inference. Check out our co-f… |
@baseten |
Company |
Quote |
2025-01-09 |
667 |
7 |
5 |
0 |
0 |
0 |
| Join @philip_kiely and Jordan from @EverydayAI_ tomorrow morning to learn about how enterprises are embracing fast, cos… |
@baseten |
Company |
Quote |
2025-01-07 |
798 |
3 |
0 |
0 |
0 |
0 |