Narration that knows your shots
Every clip arrives with the prompt that created it. Studio uses those prompts to write one coherent script across all your scenes — no re-describing, and your images are never uploaded to write it.
Google Automator
GFA Studio turns everything you made in Google Flow into a finished video — narration, captions and motion included — in one click. Everything runs on your own computer: your media, your API keys and the rendering never leave it. Then it hands you every asset so you can take it anywhere.
Included with a Pro membership during the beta. Windows only.
Generating the clips is the fast part. Assembling them is where the day goes.
Studio reads the prompt behind every clip, writes a script that connects them, voices it, cuts the captions to the audio and adds motion. You get a complete video first — then you change whatever you want.
Every clip arrives with the prompt that created it. Studio uses those prompts to write one coherent script across all your scenes — no re-describing, and your images are never uploaded to write it.
Zoom, pan, drift and reveal presets give stills real movement, and captions are cut to the narration automatically. Adjust any scene individually, or restyle every scene at once.
Use your own OpenAI or ElevenLabs key, run the local model, record directly, or upload your own narration.
Generate BGM locally that always covers the full project length. No monthly limits and no credit counter.
Click to seek, drag to select, trim from either edge, delete a section, undo. The original recording is never modified.
Take the finished video and publish it. Or take the pieces and finish the job in whatever editor you already use. Studio exports both, always.
Not a proprietary project format you can only open here. Real files: the rendered video, every scene clip, narration audio per scene and as one track, the background music, and standard SRT captions any editor accepts.
Files are numbered with zero padding and ship with a manifest listing scene order, start time and duration — so three hundred clips land in the right order in any tool, not in the order your file browser feels like.
And the part no other tool can give you: the original prompt for every scene, exported as a CSV. Regenerate a single shot in Flow and swap the file. Reuse the set on your next project. Turn it into descriptions and chapters.
✓ Your work is never trapped in StudioImages, recordings, project files and rendered videos stay in the local folder you choose. Rendering runs on your computer. Your API keys are stored by your operating system — not in the extension, and not in a GFA database.
GFA never receives your media or your AI keys. Network access happens only when you deliberately use a connected AI provider, and narration is written from your saved prompts, so your images are never uploaded to write a script.
GFA Studio is included with a Pro membership during the beta. Join while the beta is open and Studio stays included for as long as you keep your membership — even after it becomes a paid product on its own.
This membership is a licence for the Google Flow Automator extension; Studio is a benefit included during the beta, not a separate purchase. We'll announce the end of the beta at least 30 days in advance. You can cancel your membership at any time.
No Terminal, no developer tools, no manual extension setup. One guided installer puts everything Studio needs in place.
Official installers are published only on this page.
Securely signed by Microsoft, so Windows can verify the installer came from us and has not been tampered with. Credentials are stored locally by your operating system, and the official GFA extension connects automatically. macOS is not supported in this release.
Beta installer coming soonEvery release shows its version, file size and checksum. Do not download GFA Studio from third-party websites.
Better to know now than to find out after you've paid.
There is no macOS build, and none is planned for this release. If you're on a Mac, Studio will not run for you.
Local AI needs 12 GB of disk (20 GB recommended) and a dedicated GPU with at least 6 GB of VRAM. Integrated graphics will not run it — without local AI the requirement drops to 2 GB and no GPU.
Some things are still rough and will change. Keep your own backups, check what it generates before you publish it, and tell us what breaks.
Supporter and Pro members only. Studio launches from the Google Flow Automator extension and checks your membership, so a Starter account cannot open it during the beta.
Windows PCs only. There is no macOS build. Studio also does real work on your hardware rather than on a server, so check the minimum specification before you install it — an underpowered machine is the single most common reason people have a bad first experience.
It takes the images and videos you generated in Google Flow and turns them into a finished video: one scene per clip, a narration script written from your original prompts, spoken audio, captions timed to that audio, motion on stills, and background music. That's the one-click result. Everything after that is you adjusting whatever you want.
Video rendering, audio processing and local AI models are far too heavy for a browser extension, and Chrome deliberately limits what an extension may do with your file system and credentials. A desktop companion keeps the extension light and lets your computer handle media and keys properly.
They automate generating and downloading — and to be fair, several do it well. All of them stop at the download. Studio is only about what happens next: assembling those files into something publishable.
Studio is built for bulk work rather than single clips. Long projects render in segments and are extended locally so audio and music still cover the whole timeline.
When the extension sends media to Studio, it recovers the prompt that generated each item. That's how narration gets written without uploading your images anywhere, and it's why you can regenerate one bad scene later without hunting for what you originally typed. No other tool in this workflow has that information.
The rendered MP4; every scene clip as a separate file; narration audio per scene and as one continuous track; the background music track; captions as standard SRT; the original prompt for every scene as a CSV; and a manifest describing scene order, start times and durations.
Yes. Everything exports as standard media files plus SRT captions, which every major editor imports. The manifest and zero-padded file names keep hundreds of clips in the right order wherever you drop them.
File names are zero padded — scene_001, not scene_1 — so alphabetical sorting matches scene order in every tool. The manifest carries the exact timing if you need it.
No, and this is deliberate. Projects live in a local folder you choose, and the export is ordinary media files rather than a proprietary project format. There is nothing to trap you here.
Studio analyzes each image together with its original prompt to understand the full context and content. It's reading not just what's in the picture, but matching it against what you asked for, so everything makes sense when the narration is written. The more detailed your prompts and the more complex your images, the longer this step takes.
The "one click" creates a finished video by running several steps inside: image analysis (reading your prompts and media), story generation (writing a script that connects all your scenes), narration generation (turning that script into audio), caption generation (timing captions to match the audio), and background music generation. Each step is happening on your computer, and all of it is heavily influenced by your computer's processor, RAM, and GPU. A faster machine will finish noticeably sooner. This is not a network delay — your hardware is doing the work.
Performance scales with your CPU, GPU, and available RAM. Upgrading to a newer processor, adding more RAM, or using a computer with a dedicated graphics card can cut generation time significantly. The AI models are the bottleneck, not the network or our servers.
You have two options. First, you can regenerate it — select a scene, change the narration text if you want, and click regenerate. If you're using a local AI model and it's not meeting your needs, you can switch to cloud narration: use your own OpenAI, Anthropic, or ElevenLabs API key and Studio will send the prompt directly to those services instead. Your keys stay in your OS credential store and are never stored by GFA.
Yes, completely. The Studio audio editor lets you click to seek, drag to select, trim from either edge, delete a section, and undo. The original recording is never modified — all edits are non-destructive. You can also delete the narration track entirely and record your own voice directly into Studio.
Local text-to-speech models are smaller than cloud services to keep file sizes reasonable on your machine. If you need more voices or higher quality narration, use a cloud provider: connect your own OpenAI, Anthropic, or ElevenLabs key through Studio. That will draw on their much larger model and voice collections. Be aware that using cloud services consumes your API tokens — check the pricing on those services to understand the cost per character narrated.
You can edit the narration text directly in the Studio editor. Highlight any scene, change the text, and regenerate just that scene's narration. If you want to rewrite the whole script, you can also copy the entire script to an LLM like Claude, ask it to rewrite it, then paste the new version back into Studio and regenerate all narration at once. That approach gives you full creative control while keeping the videos and timing in sync.
No. The local AI models work well for most cases and cost nothing beyond the initial download. Many users get excellent results without ever touching API keys. Try local first — if the quality meets your standard, you're done. Use cloud services only if you need more, not by default.
A Pro membership on the Google Flow Automator extension, signed in with the same account. Studio launches from the extension.
If you join while the beta is open, Studio stays included with your Pro membership for as long as you keep that membership — including after Studio becomes a separately paid product.
We haven't set the price yet, and we'd rather decide it after seeing how people actually use the beta than announce a number we might have to change. What we can commit to: if you join during the beta and keep your membership, it costs you nothing extra.
Your Studio access ends with your membership. If you rejoin later, whatever terms apply at that time will apply to you — the beta benefit isn't held open indefinitely.
No. A failed payment isn't treated as leaving. While the payment is being retried your access continues, and there's an additional grace period after the billing period ends. The benefit is tied to cancelling, not to a bank hiccup.
Your membership pays for the extension; Studio is a benefit included during the beta rather than a separate purchase, so we don't offer refunds on the basis of Studio. You can cancel your membership at any time, and if something has genuinely gone wrong, talk to us on the support page — we'd rather sort it out than argue about it.
We don't have a fixed date yet. We'll announce it at least 30 days in advance so nobody is caught out.
Minimum is what Studio needs in order to run at all. Recommended is what it needs to run comfortably — local AI and video rendering competing for the same hardware. Meeting only the minimum means it will work, not that it will be quick.
With Local AI
| Component | Minimum | Recommended |
|---|---|---|
| OS | Windows 10 64-bitversion 2004 / build 19041 | Windows 11 64-bit |
| Processor | 2 cores / 4 threadsIntel Core i3-8100 · AMD Ryzen 3 2200G | 6 cores / 12 threadsIntel Core i5-12400 · AMD Ryzen 5 5600 |
| Memory | 16 GB RAM | 16 GB RAM |
| Graphics | Dedicated GPU, 6 GB VRAMNVIDIA GTX 1660 · AMD RX 6600 | Dedicated GPU, 8 GB VRAMNVIDIA RTX 3060 Ti · RTX 4060 |
| Storage | 12 GB available space | 20 GB available space (SSD) |
| Additional | Chrome or Edge with the Google Flow Automator extension | — |
⚠ Integrated graphics are not supported (Intel UHD / Iris Xe, AMD Radeon Graphics). Local AI requires a dedicated GPU.
Without Local AI
Using cloud AI or Copy-for-LLM instead? The requirements are much lower.
| Component | Minimum |
|---|---|
| Processor | 2 cores / 4 threads |
| Memory | 8 GB RAM |
| Graphics | Not required |
| Storage | 2 GB available space |
Running an AI model locally is genuinely demanding work, and the part that decides whether it is pleasant or painful is GPU memory (VRAM). The model has to fit in it. When it does, your graphics card does the work and things move quickly. When it does not, the work spills onto the CPU and system RAM, and the same job can take many times longer.
This is why the graphics row deserves more attention than the others. A fast CPU does not compensate for a card that is short on VRAM, and integrated graphics (Intel UHD / Iris Xe, AMD Radeon Graphics) will not run local AI at all — it needs a dedicated GPU with 6 GB of VRAM at minimum, 8 GB to be comfortable.
You are not shut out. Turn on Low VRAM mode, which stops Studio from loading a heavy local model and lets you use ChatGPT or Claude through their ordinary chat interface instead — you copy the script across rather than running a model on your own hardware.
There is no API key and no extra cost involved in that route. It uses the same free chat you would use in a browser, so a modest PC can still produce the whole video.
Yes. If you already have an OpenAI, Anthropic or ElevenLabs key you can connect it and Studio will call those services directly, which is faster and needs nothing from your GPU. Keys are kept in your operating system's credential store and never sent to a GFA server.
Be aware that this spends your own tokens. Every narration or script request is billed to your account by that provider, so check their pricing before you run a large project. The local and Low VRAM routes remain free — an API key is a convenience, never a requirement.
Yes. The Windows installer is securely signed by Microsoft, which means Windows can confirm it came from us and has not been altered since it was built. You will see the publisher named in the install prompt rather than an "unknown publisher" warning.
Download it only from this page. Copies from third-party download sites are not covered by that guarantee.
No. This release is Windows 10 and Windows 11 only, and there is no macOS build planned for it. Please don't subscribe for Studio if you're on a Mac.
12 GB with local AI, and 20 GB on an SSD is the comfortable figure — most of that is the local model files. If you skip local AI and use cloud AI or Copy-for-LLM instead, 2 GB is enough.
To generate the media in the first place, yes — Studio works with what Flow produces, and Flow's own limits apply to generation. Studio itself doesn't consume Flow credits.
Not to us. Selected media is copied into your local project folder and rendering happens on your machine. Narration is written from your saved prompts, so the images themselves aren't sent to an AI provider to write the script either.
In your operating system's credential store. They are never sent to a GFA server. They travel over TLS only to the provider you chose, when you use that provider.
Yes. Local narration, local captioning and local background music run entirely on your computer once installed. Cloud voices need a connection, and the extension checks your membership periodically.
It's a beta and we're not going to pretend otherwise. Core workflow is solid; some rendering and import edge cases are still being fixed, and details will change between builds. Keep backups of anything you care about and check output before publishing.
Yes, please do. Studio has a feedback and feature request form built in — open Help inside the application and submit it from there. Tell us what you actually want, in your own words.
Every request gets read. If it is something that can be built, it gets added — beta requests are what shape the release version.
Join while the beta is open and Studio stays included with your membership. Install once, send your clips from the extension, and get a finished video plus every asset behind it.