Home / Companies / Vast.ai / Blog / Post Details
Content Deep Dive

Mistral Small 4 Just Dropped — Run It on Affordable H200s with Vast.ai

Blog post from Vast.ai

Post Details
Company
Date Published
Author
Team Vast
Word Count
753
Company Posts That Month
4
Language
English
Hacker News Points
-
Post removed?
No
Summary

Mistral Small 4, the latest model from Mistral AI, consolidates capabilities for instruction, reasoning, and coding into a single, efficient framework. This 119-billion-parameter mixture-of-experts model activates just 6.5 billion parameters per token, offering significant improvements in latency and throughput over its predecessor, Mistral Small 3. With a Pixtral vision encoder, a 256K-token context window, and the ability to toggle between fast responses and detailed reasoning, the model is designed to handle diverse tasks, including vision and complex analysis. Deployed on affordable hardware like 2x H200 GPUs via Vast.ai, it offers flexibility in deployment without the need for quantization, and its weights are stored in FP8 format under an Apache 2.0 license. The guide details the process of setting up and deploying Mistral Small 4, emphasizing its efficient resource use and broad application potential.

Trends Found in this Post

No tracked trend matches for this post yet.

Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.