Step 3.7 Flash is Live on DeepInfra: An Agentic, Multimodal Model Built for Production
Blog post from Deepinfra
DeepInfra has announced the release of Step 3.7 Flash, a 198-billion-parameter sparse Mixture-of-Experts vision-language model optimized for agentic workflows, now available on their platform. This model is designed to execute complex tasks such as parsing financial reports and managing multi-step search loops with a focus on execution reliability over raw model quality. It supports a 256K context window and offers three reasoning levels to balance speed, cost, and depth per request. Step 3.7 Flash integrates a language backbone with a vision encoder for native image understanding, performing well in benchmarks like ClawEval-1.1 and SimpleVQA. The model is accessible through DeepInfra's OpenAI-compatible API, maintaining competitive pricing and ease of use for developers familiar with the platform.
| Trend | Post Mentions | Total Month Mentions | Posts | Companies | MoM |
|---|---|---|---|---|---|
| LLM | 1 | 6,237 | 1,165 | 246 | -31% |
| Multi-agent systems | 1 | 538 | 169 | 80 | -1% |
| Observability | 1 | 4,230 | 776 | 198 | +24% |
| RAG | 1 | 1,000 | 260 | 106 | -52% |
| Real-time | 1 | 5,758 | 1,361 | 266 | +0% |
| Vector Search | 1 | 1,897 | 384 | 134 | -16% |
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.