Home / Companies / Surge AI / Blog / Post Details
Content Deep Dive

DeepSeek V4 Pro: 59.7 on the Tuesday Work Index

Blog post from Surge AI

Post Details
Company
Date Published
Author
-
Word Count
1,355
Company Posts That Month
4
Language
English
Hacker News Points
-
Post removed?
No
Summary

DeepSeek V4 Pro achieved 59.7 on Surge AI’s Tuesday Work Index, improving substantially over its preview versions and surpassing several competing models, although Fable 5 and GPT 5.6 Sol Max remain ahead. Across benchmarks measuring professional instruction following, enterprise agents, research mathematics, and writing, it was presented as competitive with frontier systems, with particular emphasis on cost efficiency. On ComplexConstraints, it scored 42.1%, or roughly 83% of the leading score, for about $34 per run—under 10% of the leading system’s cost—while on Riemann-bench it matched Grok 4.5’s 38.4% score for $11.04 compared with $122.38, placing it on the cost-performance Pareto frontier. The assessment characterizes DeepSeek V4 Pro as a strong open-weight option for workloads requiring complex reasoning at lower cost, while noting that it does not match the highest absolute scores of leading frontier models.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
Gemini 3.7 Flash 4 82 12 8 -
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.