// AI NEWS · FEATURE

Google Extends Gemini's Voice Experience Into Gmail, Keep, and Docs With 'Live' Conversations

Announced at Google I/O 2026 and rolling out in the US this summer, Gmail Live, Keep Live, and Docs Live bring Gemini's real-time voice AI into Workspace — while the core Gemini Live voice, camera, and screen-share experience is now free for every Android and iOS user.

2026-08-14 · by Moe Ameen

What happened

At Google I/O 2026 on May 19, 2026, Google announced it is bringing Gemini Live — its real-time, natural voice conversation mode — into three Workspace apps: Gmail Live, Keep Live, and Docs Live. Instead of typing prompts, you hold a spoken back-and-forth with your inbox, notes, and documents: ask Gmail to find and summarize a thread, have Keep turn spoken notes into organized lists, or talk Docs through a draft. Google said the features begin rolling out in the US this summer to Google AI Pro and Google AI Ultra subscribers, in English, with a preview for Google Workspace business customers.

The Workspace push sits on top of a broader expansion of Gemini's voice experience through 2026. Gemini Live — voice, live camera, and screen sharing — is now free for everyone on Android and iOS, with no subscription required for the core conversation mode. You can interrupt, change the subject mid-answer, point your camera at something and ask about it, or share your screen, and it plugs into Gmail, Calendar, Maps, Tasks, Keep, and YouTube so you can act hands-free while the conversation continues.

Underneath is a new model. On March 26, 2026, Google DeepMind released Gemini 3.1 Flash Live, an audio-to-audio model that skips the usual speech-to-text-to-text-to-speech pipeline and generates spoken responses directly — lower latency, better preservation of tone and pacing, and the ability to follow a conversation for longer. Google says it powers both Gemini Live and Search Live across more than 200 countries and territories, and its audio output is watermarked with SynthID. Over the summer Google has also been adding finer voice controls, including real-time speech-rate adjustment, selectable accents and character voices, and new voice options. Treat exact rollout timing and subscriber requirements as a moving target and confirm current details against Google.

Why it matters for creators

  • Voice becomes a capture surface. You can brainstorm hooks, dictate a rough script, or talk out a newsletter while walking — hands-free — and Gemini keeps up in real time instead of one typed prompt at a time.
  • The camera and screen-share turn Gemini into a research and feedback tool: point it at a competitor's post, a product, or your own draft on screen and ask what to change.
  • It is free at the core. The voice, camera, and screen-share conversation mode no longer needs a subscription on Android or iOS — though the new Workspace "Live" features and some of the newer voice controls are gated to AI Pro/Ultra in the US at launch.
  • Gemini still stops at ideas and text. It talks, reasons, and organizes — it does not produce a captioned vertical video, a brand-exact carousel, or a scheduled multi-platform calendar. The hands-free session ends with a transcript, not published content.
  • That gap between "talked it out" and "shipped it" is exactly where a production engine earns its place in the workflow.

How to act on this with Kompozy

The natural way to use this as a creator is as a capture tool. Talk out ten hook variations, a rough script, or a week of angles into Gemini Live on a walk or between meetings, and you come back with a clean transcript instead of a blinking cursor. The problem is what happens next: that transcript is raw text, and turning it into finished, on-brand, multi-platform content by hand is the part that never gets done. That is the exact handoff [Kompozy](/) is built for. Drop the transcript or the idea in as a source, and Kompozy's copy engine — Claude and OpenAI, governed by your [Persona Brief](/glossary/persona-brief) and banned-word filters — rewrites it in your real voice rather than Gemini's generic one.

From there Kompozy generates the formats a voice assistant can't: [Persona Shorts](/glossary/persona-shorts) and HeyGen avatar video, Carousel Posts and Persona Tweets rendered pixel-exact through [HyperFrames](/glossary/hyperframes), Photo Posts, Quote Graphics, a blog article, and an email newsletter — then schedules and publishes the whole set across the eight social platforms plus blog and email on [Autopilot](/glossary/autopilot), with a per-post review pipeline. Gemini Live is the mouth — the fastest way yet to get an idea out of your head. Kompozy is the studio and the distribution that turns what you said into content people actually see.

Quick takeaways

  • Google is extending Gemini Live's voice mode into Gmail, Keep, and Docs as Gmail Live, Keep Live, and Docs Live — announced at Google I/O 2026 (May 19, 2026).
  • The Workspace "Live" features roll out in the US this summer to Google AI Pro and Ultra subscribers, in English, with a Workspace business preview.
  • Gemini Live's core voice, camera, and screen-sharing conversation mode is now free for all Android and iOS users, no subscription required.
  • The experience runs on Gemini 3.1 Flash Live, an audio-to-audio model released March 26, 2026 that generates speech directly for lower latency; its output is SynthID-watermarked.
  • Gemini voice is a capture and ideation surface — Kompozy is the layer that turns the resulting transcript into captioned video, carousels, blogs, and a scheduled multi-platform release.

Frequently asked questions

What is Gemini Live?

Gemini Live is Google's real-time voice conversation mode in the Gemini app. You talk with Gemini naturally — interrupting, changing the subject, sharing your camera or screen — and it responds hands-free while pulling from Gmail, Calendar, Maps, Tasks, Keep, and YouTube. It runs on the Gemini 3.1 Flash Live audio model.

Is Gemini voice free?

The core Gemini Live experience — voice conversation, live camera, and screen sharing — is now free for all Android and iOS users with no subscription. The newer Workspace features (Gmail Live, Keep Live, Docs Live) and some of the latest voice controls are gated to Google AI Pro and Ultra subscribers in the US at launch.

What are Gmail Live, Keep Live, and Docs Live?

They extend Gemini Live's spoken back-and-forth into Google Workspace: Gmail Live lets you talk to your inbox to find and summarize threads, Keep Live turns spoken notes into organized lists, and Docs Live helps turn spoken ideas into drafts. Announced at Google I/O 2026, they roll out in the US this summer to AI Pro and Ultra subscribers.

Can Gemini voice create and publish social content?

No. Gemini voice brainstorms, dictates, reasons, and organizes, but it produces text and answers — not captioned vertical video, brand-exact carousels, or a scheduled multi-platform calendar. To turn a Gemini voice transcript into finished, on-brand content and publish it across platforms, you pair it with a content engine like Kompozy.

Related news

← All AI news · Get started →