Can agents price a PR's impact or utility?
Blog post from Weave
Weave is exploring whether agent-based prediction markets can help quantify the impact of engineering changes beyond its existing “code output” metric, which measures work in pre-AI expert-engineer hours but not resulting utility. The proposed system defines impact broadly through a change’s effects on users and systems, including durability, adoption, debt, and value unlocked, then asks agents to assign a dollar-valued utility estimate to pull requests using a central limit order book. Agents receive distinct personas, research budgets, private information sessions, and tools to inspect code and trade, with the final volume-weighted price serving as the market’s assessment after a fixed trading period. This price could be compared with estimated engineering input costs to estimate a return multiplier, although the author describes the approach as preliminary and notes that current runs are essentially “beauty contests” without objective future outcomes. In a trial involving a small UI change that removed a redundant icon, the market fell from roughly $13 to $3 as agents researched the change and revised their estimates, suggesting neutral rather than meaningful value relative to its low attributed AI token cost. Future iterations may incorporate delayed labels and usage data from sources such as Slack or PostHog to better connect market prices with how changes are actually used.
| Trend | Post Mentions | Total Month Mentions | Posts | Companies | MoM |
|---|---|---|---|---|---|
| Secrets Management | 1 | 451 | 99 | 43 | -80% |
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.