Technology

Google Vids Adds AI Avatars From a Selfie, Bringing Generative Video Into the Productivity Suite

Martin HollowayPublished 6d ago5 min readBased on 8 sources
Reading level
Google Vids Adds AI Avatars From a Selfie, Bringing Generative Video Into the Productivity Suite

Google announced on July 16, 2026 that Google Vids now lets users create a custom digital avatar from a selfie and a short voice recording. The avatar is tied to the account holder's Google account, meaning it cannot be freely shared or transferred to someone else. The update is part of a broader integration of Google's Gemini Omni model into Vids, and it also introduces text-prompt-based video editing, image-to-video generation, background swapping, lighting correction, visual effects, and step-by-step editing that lets users change one element without regenerating an entire clip (TechCrunch).

The model behind these features is Gemini Omni Flash, which Google detailed in a separate Workspace blog post the same day (Google Workspace Blog). Gemini Omni Flash can generate video from a written prompt combined with reference images a user uploads, and it supports text-based editing commands for modifying existing footage.

The personal avatar feature is limited to users 18 or older in certain regions. Google watermarks each avatar invisibly using SynthID, its system for tracking the origin of AI-generated content, and ties the avatar to the account holder's likeness and Google account to prevent third-party misuse (TechCrunch).

Gemini Omni and personal avatars are available to Google AI Pro and Ultra subscribers and to Google Workspace business customers (Google Blog). Google Vids also offers a no-cost tier for creating, editing, and sharing videos without paying, a tier Google introduced in August 2025 and reiterated in April 2026 product updates (Google Blog; Google Blog).

Google Vids lives inside Google Workspace and is positioned as a business tool for company updates and training videos. Beyond avatars and generative video, its feature set includes automatic transcript trimming, AI-powered voiceover enhancements, and a feature called Workspace Drops for distributing on-brand videos at scale (Google Workspace Blog; Google Workspace Blog; Google Workspace Blog).

The update puts Google Vids in direct competition with dedicated AI video platforms including HeyGen, Synthesia, Captions, and D-ID (TechCrunch). Those companies have built their businesses around avatar-driven video creation. Google's entry differs less in raw capability than in distribution: Vids is bundled into Workspace, which puts it in front of an enterprise user base that already relies on Docs, Sheets, and Slides as part of their daily workflow.

The SynthID watermarking and account-bound avatar design suggest Google is being deliberate about provenance — the ability to trace AI-generated content back to its source. Restricting avatars to the account holder's likeness and Google account, rather than allowing open avatar creation, limits the surface area for deepfake misuse within the product itself. Whether SynthID watermarks survive downstream recompression or platform re-encoding is an open technical question outside the scope of Google's announcement.

The competitive picture for the dedicated AI video companies is now more challenging. HeyGen and Synthesia have competed on feature depth and specialization; Google's advantage is that Vids requires no separate procurement, no new vendor review, and no additional cost for Workspace business customers already on AI Pro or Ultra plans. The step-by-step editing model also addresses a practical pain point in AI video generation, where a single prompt change traditionally meant regenerating an entire clip rather than adjusting one element.

The no-cost tier lowers the barrier further. A team lead who wants to produce a training video can do so without budget approval, using their own likeness as an avatar, inside the same environment where the team already collaborates. That reduction in friction matters more than any single feature in the update.

Google has not disclosed specific latency figures, output resolution caps, or maximum clip durations for Gemini Omni Flash in Vids. The company also has not detailed which regions are excluded from the personal avatar feature beyond the 18-plus age gate. These are details that will determine how useful Vids is in production workflows, and they are absent from today's announcement.

The broader context here is that Google is folding generative video into the productivity suite where knowledge workers already spend their day, the same way it folded generative text into Docs and generative images into Slides. The capability set, taken individually, matches what specialist tools already offer. The capability set, taken as a bundled Workspace feature with a free tier, is what shifts the competitive calculation.