DeepSeek V4 Pro: 59.7 on the Tuesday Work Index
Blog post from Surge AI
DeepSeek V4 Pro achieved 59.7 on Surge AI’s Tuesday Work Index, improving substantially over its preview versions and surpassing several competing models, although Fable 5 and GPT 5.6 Sol Max remain ahead. Across benchmarks measuring professional instruction following, enterprise agents, research mathematics, and writing, it was presented as competitive with frontier systems, with particular emphasis on cost efficiency. On ComplexConstraints, it scored 42.1%, or roughly 83% of the leading score, for about $34 per run—under 10% of the leading system’s cost—while on Riemann-bench it matched Grok 4.5’s 38.4% score for $11.04 compared with $122.38, placing it on the cost-performance Pareto frontier. The assessment characterizes DeepSeek V4 Pro as a strong open-weight option for workloads requiring complex reasoning at lower cost, while noting that it does not match the highest absolute scores of leading frontier models.
| Trend | Post Mentions | Total Month Mentions | Posts | Companies | MoM |
|---|---|---|---|---|---|
| Gemini 3.7 Flash | 4 | 82 | 12 | 8 | - |
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.