Step 3.7 Flash is Live on DeepInfra: An Agentic, Multimodal Model Built for Production
Blog post from Deepinfra
DeepInfra has announced the release of Step 3.7 Flash, a 198-billion-parameter sparse Mixture-of-Experts vision-language model optimized for agentic workflows, now available on their platform. This model is designed to execute complex tasks such as parsing financial reports and managing multi-step search loops with a focus on execution reliability over raw model quality. It supports a 256K context window and offers three reasoning levels to balance speed, cost, and depth per request. Step 3.7 Flash integrates a language backbone with a vision encoder for native image understanding, performing well in benchmarks like ClawEval-1.1 and SimpleVQA. The model is accessible through DeepInfra's OpenAI-compatible API, maintaining competitive pricing and ease of use for developers familiar with the platform.
| Trend | Post Mentions | Total Month Mentions | Posts | Companies | MoM |
|---|---|---|---|---|---|
| LLM | 1 | 6,292 | 1,205 | 252 | -36% |
| Multi-agent systems | 1 | 556 | 175 | 81 | -7% |
| Observability | 1 | 4,261 | 791 | 201 | +16% |
| RAG | 1 | 1,005 | 263 | 108 | -56% |
| Real-time | 1 | 6,055 | 1,444 | 270 | -11% |
| Vector Search | 1 | 1,918 | 398 | 137 | -21% |
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.