The Big Shift: From Slide Deck Helper to AI Avatar Video Studio
Google Vids is an AI avatar video creation tool inside Google Workspace that turns typed prompts, selfies, and short voice recordings into personalized videos where a digital version of you presents the content on screen, with built-in editing powered by Gemini Omni for background, lighting, and post-production changes.
The key takeaway: Google is no longer flirting with video; it is walking straight into the enterprise avatar studio market. Two new Google Vids features—personalized video avatars and Gemini Omni multimodal editing—let you star in camera-free videos generated from text prompts, reference photos, a selfie, and a ten‑second voice clip. That combination pushes Vids beyond “AI-assisted workplace tool” into a full-fledged video creation platform that competes head-on with AI video generation tools already used in many companies. If you already pay for Workspace, this could be the moment separate avatar subscriptions start to look like unnecessary overhead.

How Google’s Personalized Video Avatars Work in Practice
The personal avatar feature is designed for people who hate cameras but still need to show up in video. The first time you use it, you submit a selfie and record about ten seconds of your voice; Vids uses that to build a custom avatar that “will look and sound just like you.” After that, your employer can spin up training videos, onboarding modules, or company announcements with your digital stand‑in, without asking you to clear your calendar or memorize a script.
From a viewer’s perspective, this matters more than stock avatars. Personal avatars make workplace content more conversational and relatable, which can hold attention far better than faceless slides. The catch: these avatars are tied to a person’s Google account and currently restricted to people aged 18 or older in unspecified “specified regions,” so global rollouts will have adoption gaps and compliance questions from day one.
Gemini Omni Video Editing: The Missing Link in AI Avatar Workflows
Avatar generators have always had a weak point: editing. Change your mind about the background or lighting, and you often have to regenerate the whole video. Gemini Omni is Google’s answer to that problem. Integrated directly into Vids, this multimodal model “mashes inputs together” from prompts and reference photos to create personalized video, then accepts step‑by‑step edits instead of forcing you to start over.
In practice, that means you can ask Omni to swap a background, adjust lighting on phone-recorded footage, or add effects within a single session. You can walk forward and backward through edits instead of scrapping a project. For non-technical users, this turns what used to be a specialist workflow into something that feels like editing a slide deck: type what you want, iterate, and stay inside Workspace the whole time.
Why This Directly Threatens HeyGen, Synthesia, and Friends
The strategic play is obvious. Existing AI video generation tools like HeyGen, Synthesia, Captions, and D‑ID built their businesses helping companies produce avatar-based videos at scale, without film crews or on‑camera executives. They sell subscriptions that start around USD 50 (approx. RM230) per month for individuals and can exceed USD 500 (approx. RM2,300) per month at enterprise tier. Meanwhile, large organizations already pay about USD 12 to USD 22 (approx. RM55–RM100) per user per month for Google Workspace.
The uncomfortable question for incumbents is simple: why keep a separate AI avatar line item when a similar tool is bundled with tools staff already use every day? History suggests bundled chat killed many standalone enterprise chat tools, and bundled Meet reshaped the video conferencing landscape. The same consolidation pressure is now coming for avatar video platforms, especially in departments that mainly need onboarding, policy explainers, and routine internal comms.
Friction, Risks, and a Practical Workflow for Teams
From a workflow perspective, the appeal is strong. Inside Workspace, you draft a script, feed it to Vids, add reference images, and let Gemini Omni assemble a first cut. You then tweak the background, adjust lighting, and apply post‑production effects until the video “pops,” all without ever booking a studio. For HR, learning teams, or product marketing, that is end-to-end AI avatar video creation with far less friction than traditional production.
But the gaps are real. Google has not explained what happens to the selfie and voice clips that generate these personalized video avatars—how long biometric data is stored, who can access it, or whether it is reused for training. Personal avatars are also restricted to adults in unnamed regions, with enforcement in enterprises effectively delegated to IT admins. Add regulatory scrutiny of bundled AI features and open questions about Gemini’s training data, and any serious deployment will need legal and compliance sign‑off, not just an enthusiastic comms team.






