Google has quietly rolled out a new feature in its Vids platform that lets users create AI-generated videos starring their own digital avatars. The tool, which is now available to Google Workspace users, lets anyone upload a selfie and a voice recording to generate a custom avatar that can speak, gesture, and move in real time, all powered by Google’s Gemini Omni model.
This isn’t just about making a talking head. The platform includes advanced editing tools like background swapping, lighting correction, and step-by-step video editing. That means you can take a simple recording and turn it into a polished, professional video, no camera crew or studio required. For businesses, this could be a game changer.
Google positions Vids as a direct competitor to platforms like Synthesia and HeyGen, which have been popular for corporate training, marketing, and internal communications. But Google’s approach is different. It’s integrated into the Google Workspace ecosystem, which means teams can use it alongside Docs, Sheets, and Meet without leaving the platform. That’s a big advantage for companies already using Google’s tools.
The Gemini Omni model is the backbone of this feature. It’s designed to handle multi-modal inputs, meaning it can process both visual and audio data simultaneously. That’s why you can upload a selfie and a voice clip and get a coherent, expressive avatar. The model also handles lighting and background adjustments automatically, which is a huge time-saver for users who don’t have video editing experience.
For corporate use cases, this could mean massive efficiency gains. Imagine a company training new hires on complex software. Instead of recording a live demo with a presenter, they could use Vids to create a series of short, personalized videos starring a digital avatar that walks through the software step by step. Or a marketing team could produce multiple versions of a product launch video, each with a different avatar, without hiring multiple actors or producers.
There’s also a clear cost advantage. Live talent is expensive, especially for repetitive or scalable content. With Vids, companies can produce hundreds of videos with the same avatar, each tailored to a different audience or message. That’s not just cheaper, it’s more flexible.
Google’s move into AI video creation is part of a broader trend toward automation in corporate media. As we’ve seen in other areas, AI tools are reducing the need for human labor in repetitive tasks, and video production is no exception. This is especially relevant for companies that need to produce content quickly and at scale, like e-commerce brands, educational institutions, or SaaS companies.
One thing to note: Google is not the only player in this space. Synthesia and HeyGen have been around for years and have built strong user bases. But Google’s advantage lies in its integration with existing enterprise tools and its ability to scale across Google Workspace. That could give it an edge in the corporate market.
As first reported by TechCrunch, Google’s Vids is still in early stages, but the potential is clear. It’s not just about making videos, it’s about making them faster, cheaper, and more personalized.
For teams looking to automate their video workflows without losing control, this is a tool worth watching. As we’ve discussed in our own posts, building workflow automation without losing control is a key challenge for modern businesses, and Vids offers a compelling solution.
The future of corporate video may look a lot like this: no live talent, no expensive studios, just AI avatars and smart editing tools, all managed from within your existing software stack. Google’s Vids is a step toward that future.