Gemini Omni
Googleが発表したGemini Omniは、テキスト、画像、音声、動画をシームレスに処理・統合するマルチモーダルAIモデルです。単一のモデルで複数のデータタイプを同時に扱えるため、より自然で人間らしいインタラクションを実現します。この技術は生成AIの新たな可能性を切り拓き、検索やアシスタント機能の進化に貢献すると期待されています。
Googleが発表したGemini Omniは、テキスト、画像、音声、動画をシームレスに処理・統合するマルチモーダルAIモデルです。単一のモデルで複数のデータタイプを同時に扱えるため、より自然で人間らしいインタラクションを実現します。この技術は生成AIの新たな可能性を切り拓き、検索やアシスタント機能の進化に貢献すると期待されています。
At Google I/O, Google announced Gemini Spark, a personal AI agent running on Gemini 3.5 Flash and Antigravity. The open source Gemini CLI will stop working with AI subscriptions on June 18th, replaced by a closed source Antigravity CLI.
Google released Gemini 3.5 Flash directly to general availability, rolling it out across its Gemini app, Search, and developer platforms. The model is 3x the price of 3 Flash Preview, fitting a broader trend of AI labs raising API prices. A new beta Interactions API with server-side history management was also announced.
Google announced Gemini Spark, a new AI personal agent capable of navigating users' digital lives and acting on their behalf across Google products. The agent has been tested with limited users and will be available next week to subscribers of AI Ultra, a new $100-per-month tier.
llm-gemini 0.32a0 has been released. It is compatible with the llm>=0.32a0 alpha and adds the ability to stream reasoning tokens.
The release of llm-gemini 0.32 adds support for the new gemini-3.5-flash model, enabling users to access Gemini 3.5 Flash through the plugin.