Gemini Omni
谷歌发布了新一代多模态AI模型Gemini Omni,该模型能够无缝处理文本、图像、音频和视频等多种输入形式,并在复杂推理和多模态理解方面展现出显著提升。Gemini Omni旨在为用户提供更加自然和全面的交互体验,进一步推动AI在跨模态应用场景中的发展。
谷歌发布了新一代多模态AI模型Gemini Omni,该模型能够无缝处理文本、图像、音频和视频等多种输入形式,并在复杂推理和多模态理解方面展现出显著提升。Gemini Omni旨在为用户提供更加自然和全面的交互体验,进一步推动AI在跨模态应用场景中的发展。
At Google I/O, Google announced Gemini Spark, a personal AI agent running on Gemini 3.5 Flash and Antigravity. The open source Gemini CLI will stop working with AI subscriptions on June 18th, replaced by a closed source Antigravity CLI.
Google released Gemini 3.5 Flash directly to general availability, rolling it out across its Gemini app, Search, and developer platforms. The model is 3x the price of 3 Flash Preview, fitting a broader trend of AI labs raising API prices. A new beta Interactions API with server-side history management was also announced.
Google announced Gemini Spark, a new AI personal agent capable of navigating users' digital lives and acting on their behalf across Google products. The agent has been tested with limited users and will be available next week to subscribers of AI Ultra, a new $100-per-month tier.
llm-gemini 0.32a0 has been released. It is compatible with the llm>=0.32a0 alpha and adds the ability to stream reasoning tokens.
The release of llm-gemini 0.32 adds support for the new gemini-3.5-flash model, enabling users to access Gemini 3.5 Flash through the plugin.