Product updates: Multi-node training clusters, B200 and H200s, and Client 1.0 release
Blog post from Modal
Modal announced beta support for multi-node GPU training through a clustered decorator that co-schedules GPUs across hosts with 3.2 Tbps InfiniBand RDMA networking, aiming to enable near-linear training scalability. The platform also introduced serverless NVIDIA B200 and H200 GPUs, priced at $6.25 and $4.54 per hour respectively, with claimed LLM inference gains of two to four times over H100s. Modal Client 1.0 emphasizes API stability and includes migration guidance, while recent client updates add read-only volumes, secret support in shell sessions, cron timezones, and timestamped application logs. Additional resources cover benchmark comparisons between SGLang and vLLM, examples for running and optimizing FLUX image-generation models, and Quora’s use of Modal Sandboxes for high-volume code execution on its Poe chatbot platform.
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.