Google Vids Adds Personal AI Avatars and Gemini Omni Flash Video Generation

Google announced on July 16, 2026 that Google Vids now lets users create custom digital avatars from a selfie and a voice recording, tying the generated likeness to the account holder's Google account. The update, part of a broader integration of Google's multi-modal model Gemini Omni into Vids, also introduces text-prompt-driven video editing, image-to-video generation, background swapping, lighting correction, visual effects, and step-by-step incremental editing that spares users from regenerating clips from scratch (TechCrunch).
The specific model powering these generative features is Gemini Omni Flash, which Google detailed in a separate Workspace blog post on the same day (Google Workspace Blog). Gemini Omni Flash in Vids can generate video from a combination of a written prompt and reference images uploaded by the user, and it supports text-based editing commands for modifying existing footage.
The personal avatar feature is restricted to users aged 18 or older in certain regions. Google watermarks avatars invisibly using SynthID, its provenance-tracking system for AI-generated content, and ties each avatar to the account holder's likeness and Google account to prevent portability or misuse by third parties (TechCrunch).
Gemini Omni and personal avatars are available to Google AI Pro and Ultra subscribers and to Google Workspace business customers (Google Blog). Google Vids also offers a no-cost tier that allows users to create, edit, and share videos without paying, a tier Google introduced in August 2025 and reiterated in April 2026 product updates (Google Blog; Google Blog).
Google Vids sits inside Google Workspace and is positioned as a business tool for company updates and training videos. The feature set extends beyond avatars and generative video: it includes automatic transcript trimming, AI-powered voiceover enhancements, and a feature called Workspace Drops for delivering on-brand videos at scale (Google Workspace Blog; Google Workspace Blog; Google Workspace Blog).
The update places Google Vids in direct competition with dedicated AI video platforms including HeyGen, Synthesia, Captions, and D-ID (TechCrunch). Those companies have built their businesses around avatar-driven video creation, and Google's entry differs less in raw capability than in distribution: Vids is bundled into Workspace, which puts it in front of an enterprise install base that already uses Docs, Sheets, and Slides as part of their daily workflow.
The SynthID watermarking and account-bound avatar design signal that Google is being deliberate about provenance. The decision to restrict avatars to the account holder's likeness and Google account, rather than allowing open avatar creation, limits the surface area for deepfake misuse within the product itself. Whether SynthID watermarks survive downstream recompression or platform re-encoding remains an open technical question outside the scope of Google's announcement.
Looking at what this means for the competitive landscape, the dedicated AI video companies now face a platform incumbent with a free tier and an existing enterprise distribution channel. HeyGen and Synthesia have competed on feature depth and specialization; Google's advantage is that Vids requires no separate procurement, no new vendor review, and no additional cost for Workspace business customers already on AI Pro or Ultra plans. The step-by-step editing model also addresses a practical pain point in AI video generation, where a single prompt change traditionally meant regenerating an entire clip rather than adjusting one element.
The no-cost tier lowers the barrier further. A team lead who wants to produce a training video can do so without budget approval, using their own likeness as an avatar, inside the same environment where the team already collaborates. That friction reduction matters more than any single feature in the update.
Google has not disclosed specific latency figures, output resolution caps, or maximum clip durations for Gemini Omni Flash in Vids. The company also has not detailed which regions are excluded from the personal avatar feature beyond the 18-plus age gate. These are details that will determine how useful Vids is in production workflows, and they are absent from today's announcement.
The broader arc here is straightforward. Google is folding generative video into the productivity suite where knowledge workers already spend their day, the same way it folded generative text into Docs and generative images into Slides. The capability set, taken individually, matches what specialist tools already offer. The capability set, taken as a bundled Workspace feature with a free tier, is what shifts the competitive calculation.


