| Manu (@manuaero) joined Jason Calacanis on This Week in AI alongside @sarahookr (Adaption Labs) and @spirosx (Resolve A… |
@labelbox |
Company |
Original |
2026-07-16 |
78,714 |
12 |
1 |
4 |
0 |
2 |
| Do AI models become less honest when they self-report their own misbehavior?
Our Applied Research team studied a key q… |
@labelbox |
Company |
Original |
2026-07-07 |
396,103 |
21 |
5 |
5 |
0 |
2 |
| 1/ Today, we’re introducing Recursion: the RL platform for building, evaluating, and deploying specialist agents.
The … |
@labelbox |
Company |
Original |
2026-06-25 |
8,700,075 |
306 |
12 |
12 |
2 |
49 |
| Where do models change their minds?
Natural Language Autoencoders (NLAs) offer a promising way to translate a model’s … |
@labelbox |
Company |
Original |
2026-06-16 |
3,932,013 |
77 |
4 |
9 |
1 |
25 |
| When AI benchmarks saturate, what comes next?
Historically, leaderboard saturation leads to two paths: hyper-specializ… |
@labelbox |
Company |
Original |
2026-05-20 |
10,696,848 |
115 |
11 |
42 |
2 |
31 |
| This week, we had the pleasure of hosting 50+ researchers and builders from leading AI companies to meet, talk and soci… |
@labelbox |
Company |
Original |
2026-04-23 |
1,611,060 |
73 |
5 |
5 |
3 |
9 |
| Interrupt a voice agent mid-sentence and most models struggle to stay aligned with the original objective.
We built Ec… |
@labelbox |
Company |
Original |
2026-03-05 |
4,762 |
5 |
2 |
2 |
0 |
5 |
| Voice agents are moving beyond rigid turn based systems toward real time, natural conversation, streaming understanding… |
@labelbox |
Company |
Original |
2026-03-04 |
2,713,533 |
885 |
138 |
18 |
166 |
121 |
| AI safety is often judged by refusal rates on adversarial benchmarks. But what if we are measuring keyword sensitivity,… |
@labelbox |
Company |
Original |
2026-02-20 |
3,760,026 |
1,485 |
251 |
47 |
184 |
134 |
| Dario (CEO of @AnthropicAI) x @dwarkesh_sp just unpacked where AI is headed since their last chat 3 years ago, covering… |
@labelbox |
Company |
Original |
2026-02-13 |
1,145,093 |
1,124 |
92 |
19 |
3 |
126 |
| We're excited to share that we’ve acquired @upcraftai to bring AI agents to the heart of how we scale human expertise … |
@labelbox |
Company |
Original |
2026-02-10 |
1,745,426 |
1,126 |
107 |
39 |
7 |
89 |
| A few takeaways from the must-watch episode from @elonmusk x @dwarkesh_sp x @collision that dropped today. An almost … |
@labelbox |
Company |
Original |
2026-02-05 |
1,516,949 |
1,329 |
190 |
59 |
19 |
133 |
| Just @ilyasut x @dwarkesh_sp dishing out AI insights in time for thanksgiving. Key takeaways:
- LLMs are benchmark spe… |
@labelbox |
Company |
Original |
2025-11-25 |
419,304 |
391 |
34 |
10 |
3 |
156 |
| Today, we’re launching Labelbox Applied Research, along with its three flagship pillars:
- Evals: A unified framework … |
@labelbox |
Company |
Original |
2025-11-20 |
455 |
3 |
0 |
0 |
0 |
2 |
| This week, @satyanadella gave @dwarkesh_sp and Dylan Patel from @SemiAnalysis_ an exclusive look at Microsoft’s Fairwat… |
@labelbox |
Company |
Original |
2025-11-12 |
175,712 |
164 |
17 |
2 |
1 |
45 |
| Essential weekend reading. The Scaling Era: an oral history of AI by @dwarkesh_sp and thank you @stripepress! https://t… |
@labelbox |
Company |
Original |
2025-10-27 |
129,802 |
195 |
17 |
4 |
1 |
77 |
| Highly recommend tuning into @dwarkesh_sp's episode today with @karpathy. They dive deep into why RL is so information-… |
@labelbox |
Company |
Original |
2025-10-17 |
1,149,903 |
1,549 |
134 |
25 |
1 |
113 |
| Thrilled to be featured in Dwarkesh’s latest episode with Richard Sutton, widely regarded as the father of reinforcemen… |
@labelbox |
Company |
Original |
2025-09-26 |
293,705 |
675 |
71 |
13 |
2 |
40 |
| We recently invited @dwarkesh_sp to stop by our SF robotics lab. World-class podcaster, rookie robotics intern. https:/… |
@labelbox |
Company |
Original |
2025-09-12 |
1,004,340 |
1,572 |
140 |
17 |
4 |
90 |
| We’ve always admired how @dwarkesh_sp sparks conversations with top thinkers in AI, academia, and tech.
Now we’re tea… |
@labelbox |
Company |
Original |
2025-09-05 |
1,041,024 |
1,256 |
101 |
32 |
0 |
98 |
| Introducing ConstraintBench: a new benchmark for evaluating LLM reasoning on realistic resource-constrained project sch… |
@labelbox |
Company |
Original |
2025-08-22 |
417,987 |
476 |
50 |
10 |
4 |
28 |
| As AI advances, so do the human skills required to shape and align it.
Full report: https://t.co/pmiGSriNU6 https://t.… |
@labelbox |
Company |
Original |
2025-07-17 |
769 |
4 |
0 |
0 |
0 |
0 |
| Grok-4 just landed on our Complex Reasoning leaderboard, and it’s impressive💥
- Math: 81.8%
- Pure Math: 84.8%
- Applie… |
@labelbox |
Company |
Original |
2025-07-10 |
1,280 |
12 |
2 |
11 |
0 |
0 |
| obligatory AI company SF billboard https://t.co/OFmCcuBPHD |
@labelbox |
Company |
Original |
2025-07-09 |
610 |
12 |
1 |
3 |
0 |
0 |
| Catch @manuaero at 41:20 https://t.co/Fh3EMEkeuR |
@labelbox |
Company |
Quote |
2025-07-01 |
747 |
6 |
0 |
0 |
0 |
0 |
| these boxes aren’t going to label themselves |
@labelbox |
Company |
Original |
2025-06-29 |
671 |
5 |
0 |
0 |
0 |
0 |