Home / Companies / Google Cloud / Blog / Post Details
Content Deep Dive

Gemini 1.5 Pro Now Available in 180+ Countries; with Native Audio Understanding, System Instructions, JSON Mode and more

Blog post from Google Cloud

Post Details
Company
Date Published
Author
Jaclyn Konzelmann, and Megan Li
Word Count
510
Company Posts That Month
6
Language
English
Hacker News Points
-
Post removed?
No
Summary

Google AI has launched the Gemini 1.5 Pro model in public preview across 180+ countries via the Gemini API, introducing new features such as native audio understanding and a File API for easier file handling. Developers can now utilize system instructions and JSON mode for more controlled model output, and benefit from improvements in function calling. The model's input modalities have been expanded to include both audio and visual reasoning capabilities for video content. Additionally, a new text embedding model, "text-embedding-004," is available, offering superior retrieval performance compared to existing models. Developers are encouraged to access these advancements through Google AI Studio, explore the new Gemini API Cookbook for guidance, and participate in the community on Discord.

Trends Found in this Post

No tracked trend matches for this post yet.

Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.