Home / Companies / Deepinfra / Blog / Post Details
Content Deep Dive

Step 3.7 Flash is Live on DeepInfra: An Agentic, Multimodal Model Built for Production

Blog post from Deepinfra

Post Details
Company
Date Published
Author
Deep
Word Count
910
Company Posts That Month
6
Language
English
Hacker News Points
-
Post removed?
No
Summary

DeepInfra has announced the release of Step 3.7 Flash, a 198-billion-parameter sparse Mixture-of-Experts vision-language model optimized for agentic workflows, now available on their platform. This model is designed to execute complex tasks such as parsing financial reports and managing multi-step search loops with a focus on execution reliability over raw model quality. It supports a 256K context window and offers three reasoning levels to balance speed, cost, and depth per request. Step 3.7 Flash integrates a language backbone with a vision encoder for native image understanding, performing well in benchmarks like ClawEval-1.1 and SimpleVQA. The model is accessible through DeepInfra's OpenAI-compatible API, maintaining competitive pricing and ease of use for developers familiar with the platform.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
LLM 1 6,237 1,165 246 -31%
Multi-agent systems 1 538 169 80 -1%
Observability 1 4,230 776 198 +24%
RAG 1 1,000 260 106 -52%
Real-time 1 5,758 1,361 266 +0%
Vector Search 1 1,897 384 134 -16%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.