Googleがリアルタイム会話向けのGemini 3.5 Live Translateを発表

Googleは、多言語会話中にほぼ瞬時の音声翻訳を可能にするAIモデル「Gemini 3.5 Live Translate」を発表しました。このツールは70以上の言語に対応しており、従来のシステムで一般的だった遅延の削減を目指しています。火曜日より開発者向けに公開されました。

このモデルは、音声を順番に処理するのではなく、継続的なストリーミング翻訳を実行します。このアプローチにより、話者の話すペース、イントネーション、感情的なトーンを維持しながら、わずか数秒の遅延で会話を進めることが可能です。Googleによると、同システムは騒がしい環境や声の重なり、日常会話にも対応します。言語は自動的に検出され、一つの会話の中で数千通りの言語ペアをサポートします。開発者は、Gemini Live APIおよびAI Studioのパブリックプレビューを通じてこのモデルにアクセスできます。一部のエンタープライズ顧客は今月中にGoogle Meetで利用可能となり、その後、順次拡大される予定です。また、このツールはAndroidおよびiOS向けのGoogle翻訳アプリにも近日中に導入されます。すべてのオーディオストリームには、AIによって生成されたことを示すSynthIDの電子透かしが含まれます。同社は、この技術がカスタマーサポートやツアー、教室などの実用的な環境向けに設計されていることを強調しています。

関連記事

Illustration of a woman using ChatGPT voice features on a tablet in a living room setting.
AIによって生成された画像

OpenAI rolls out new GPT-Live voice models for ChatGPT

AIによるレポート AIによって生成された画像

OpenAI has begun releasing two updated voice models for ChatGPT that allow simultaneous listening and speaking. The changes aim to create more natural conversations.

Google announced new artificial intelligence models in May during Google I/O 2026. The Gemini 3.5 and Gemini Omni tools aim to handle tasks proactively.

AIによるレポート

Google is rolling out its Gemini AI model more widely, with new features for smart home devices and on-device use in Chrome.

Google Workspace is expanding its AI capabilities with improved inbox assistance and the ability to have conversations directly with important work documents.

AIによるレポート

Apple is preparing to introduce a redesigned Siri powered by Google's Gemini models at its upcoming Worldwide Developers Conference. The update, expected with iOS 27, will feature major interface changes and enhanced AI capabilities. Release is scheduled for mid-September.

このウェブサイトはCookieを使用します

サイトを改善するための分析にCookieを使用します。詳細については、プライバシーポリシーをお読みください。
拒否