Gemini 1.5 Pro Now Available in 180+ Countries; with Native Audio Understanding, System Instructions, JSON Mode and more
Blog post from Google Cloud
Google AI has launched the Gemini 1.5 Pro model in public preview across 180+ countries via the Gemini API, introducing new features such as native audio understanding and a File API for easier file handling. Developers can now utilize system instructions and JSON mode for more controlled model output, and benefit from improvements in function calling. The model's input modalities have been expanded to include both audio and visual reasoning capabilities for video content. Additionally, a new text embedding model, "text-embedding-004," is available, offering superior retrieval performance compared to existing models. Developers are encouraged to access these advancements through Google AI Studio, explore the new Gemini API Cookbook for guidance, and participate in the community on Discord.
No tracked trend matches for this post yet.
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.