Home / Companies / Weave / Blog / Post Details
Content Deep Dive

Your agents are paying frontier prices to read tool output

Blog post from Weave

Post Details
Company
Date Published
Author
-
Word Count
664
Company Posts That Month
1
Language
English
Hacker News Points
-
Post removed?
No
Summary

Weave optimizes the routing of requests to machine learning models by selecting the most cost-effective model that can handle each specific task, thereby substantially reducing costs while maintaining performance. By routing entire sessions rather than individual requests, Weave preserves the working cache, ensuring that cost savings are real rather than theoretical. This approach results in significant cost reductions, such as an 88.5% cut on median requests and a 68.7% cut across all routed-down requests, while maintaining reliability with a 98.8% success rate for failover requests. Unlike naive routing practices that switch models to save money but end up increasing costs due to cache loss, Weave's session-based routing ensures that these savings are sustainable in production environments. The system is designed to quickly adapt to new model releases without requiring code changes, allowing users to benefit from improved models and cost savings seamlessly.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
Real-time 1 1,106 270 109 -81%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.