Announcing the general availability of Llama 4 as MaaS on Vertex AI
Blog post from Google Cloud
Llama 4, the latest iteration of Meta's large language models, is now available as a fully managed API endpoint on Google Cloud's Vertex AI, addressing infrastructure challenges and enabling users to focus on application development. This release includes Llama 4 Scout, optimized for single-GPU environments with advanced reasoning and multimodal task efficiency, and Llama 4 Maverick, designed for complex tasks like image understanding. The Llama 4 Model-as-a-Service (MaaS) offering provides zero infrastructure management, guaranteed performance, and enterprise-grade security, all accessed through a simple API endpoint. Users can begin by accepting the Llama Community License Agreement and selecting their desired model via the Vertex AI Model Garden. The cost model is pay-as-you-go, with pricing details and quotas available on the Vertex AI pricing page. The service aims to streamline AI application development and invites users to explore the model while providing feedback through the Google Cloud community forum.
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.