Storyboard to video is no longer a thin wrapper around stills. Generate scenes end to end, keep characters consistent with entity-aware references, and move from outline to playable video with fewer manual patch-ups. The storyboard flow itself got a dedicated UX pass so planning and generation feel like one product.
The script to video workflow got another UX pass focused on clarity and momentum: cleaner writer layout, better floating controls, and less friction between drafting, styling, and generate. Small copy and navigation fixes make restart and remix paths obvious when you want a second take.
The editor is much more agent-friendly. Floating prompts support reference-image attachments for edit and animate. Generation history sits on the canvas as a first-class action. Direct side-panel asset editing routes into the same floating prompt UX. Project-level audio gets its own layer, timeline insertion between clips is clearer, and multi-select drag is less surprising.
Section transition duration is editable from the side panel. Generative sound effects can ride along with transitions, so cuts feel produced instead of silent. Timeline performance was refactored so dense projects stay responsive while you scrub and rearrange.
Browse and add royalty-free sound effects without leaving VideoGen. Pair them with transitions, accents, or timeline moments the same way you already place music and voiceover.
Slideshow workflows can generate slides without forcing a PDF upload first. Motion graphic assets can carry arbitrary Remotion code for richer animated compositions. Lottie support lands for motion-friendly assets in the editor stack.
Exports and AI video clip generation can include an end-screen mode, so finished videos can close with a polished final beat instead of cutting to black.
Avatars now stay visible for the full video duration when they should, and avatar quality lives in the picker where you expect it. Entity-aware avatar tooling makes character setup more consistent across workflows.
Voiceover flows support transcript upload and scene-aware handling, with temporary-output cleanup so intermediate media does not clutter your library. Audio trim timings are honored on export, and waveform rendering respects gaps and playback speed.
A What's new experience surfaces recent product updates inside the app, so you do not have to hunt the website changelog to learn what changed.
Hand-written TypeScript and Python SDKs replace the older generated clients, with clearer resource namespaces and wait helpers. The VideoGen CLI mirrors the same operations from the terminal. MCP connects Cursor, Claude Desktop, and other agents to the live VideoGen API tools. Production API add-on customers get a larger included credit grant for shipping at scale.
The dashboard is rebuilt around how people actually start videos. Browse workflow shelves, jump into AI media tools from the home grid, and get back to recent projects faster. Loading states, empty states, and navigation are tighter so the home screen feels like a launchpad instead of a waiting room.
Script to video and voiceover to video got a full UX pass: clearer style steps, better defaults for post-production, logo controls on workflows, and a smoother path from prompt to finished preview. Prompt-and-media generation is now a first-class workflow, so you can start from a prompt plus reference media without bolting steps together by hand.
Storyboard to video is available as a full workflow. Plan scenes, generate visuals with character consistency in mind, and move into the editor when you are ready. Character grids and sequential generation help recurring people stay looking like themselves across scenes.
The timeline foundation is new. Sections use explicit layer types for audio, background visuals, and visual overlays, so arrangements stay predictable as projects get complex. You can add asset-to-asset transitions, audio transitions, and finer post-production transition controls without fighting the old single-arrangement model.
After generation, you land in a dedicated preview experience with cleaner playback controls, clearer edit/remix/restart actions, and a feedback assistant for iterating on the result. Preview analytics and mobile preview flows were rebuilt so the first watch is easier on every device.
AI quality is unified across image, video, and avatar generation. Pick the tier that fits the job, save your preference across workflows, and use Max when you want the strongest agent and generation quality. Try Max prompts make the top tier easier to discover without burying default settings.
Create a project directly from an AI playground result. Generative tools handle slides and PDFs as image inputs more reliably, and you can generate AI thumbnails on the public view page. Replace actions show up on generatable assets so iteration stays one click away.
The credit system is easier to understand day to day. The dashboard shows daily usage broken down by feature, team credit charts are available in the ledger, and you can export usage to CSV. Included-credit refills, top-up purchase tracking, and clearer included-vs-additional credit copy reduce surprise when you scale up.
Presets are workflow-aware: save caption style, logo position, and the settings that actually matter for the flow you use. Public video templates and remix-template entry points make it easier to start from a finished look instead of a blank timeline.
The developer API matured into a stable workflows-and-tools surface with better examples, guides, and rate limiting. Install the VideoGen CLI to drive the same API from your terminal and CI. Webhooks, quality parameters on workflows, and broader remix-action coverage make automation paths production-ready.

Generative video, AI avatars, AI images, and every other AI feature are now available on every paid plan. If you're on Pro, you get access to everything. If you're on Business, you get all of that plus the Premium stock library included at no extra cost, along with a higher monthly credit allotment.
We've overhauled the script-to-video workflow with improved AI image generation, more visual style presets, smoother section transitions, built-in effects, and better defaults across the board. The result is higher-quality videos straight from a prompt, with less manual editing needed afterward.
You can now choose between faster, lower-cost drafts and higher-quality finals when generating AI images. Pick the tier that fits the job — quick iterations or polished results. We'll be expanding quality tiers to more features over time.
Your plan now comes with a monthly credit allotment that refills at the start of each billing cycle. A rate card shows what each action costs, a usage ledger lets you see where your credits went, and you can buy additional credits or enable automatic top-ups at any time from the Usage tab.
You can now create videos, images, voices, and avatars programmatically — or run full workflows like script-to-video from your own apps. Get an API key from the Developers tab, install the TypeScript or Python SDK, and generate your first video in minutes.
When you generate a video, the AI agent shows its reasoning as it works. You can follow along as it makes decisions about visuals, pacing, and structure — so you always know what's happening and why.
Set colors, fonts, and logo for the entire project from a dedicated Theme tab. Every section picks up the theme automatically, keeping your videos visually consistent without editing each one individually.
Lock assets in place while keeping their content editable. Drag assets between layers. Crop images inline without leaving the canvas. Edit SVG colors directly. Add borders to images and shapes. Duplicate assets with one click, and hold Alt while dragging to disable snapping.
A redesigned share popover lets you set viewer, editor, or admin access for each collaborator. Your team's email domain is auto-detected so adding colleagues is faster. Projects from other teams now open in read-only mode to prevent accidental edits.
Projects are now fully collaborative. Invite teammates to edit the same project simultaneously - changes sync in real-time across all editors. You can see who's currently working on a project and where they're editing, so you never overwrite each other's work.
When creating a video from a script, you can now choose a visual style to guide how images and media are generated. Select from styles like anime, cinematic, minimalist, or corporate to give your video a consistent look without manually editing each asset.
We've rebuilt the script to video workflow to give you more control. You can now regenerate your script at any point - tweak the tone, adjust the length, or take it in a new direction without starting over. The new flow makes it easier to iterate until your script is exactly right.
We’ve added an "Image to video" option to the "Generate AI clip" workflow: start with an image and a prompt, and VideoGen picks the best model to animate it. You can choose to include audio in the generated video clip as well.
If you have an existing image that you would like to animate, you can select the image and use the "Animate to video" action in the right side panel.
Each new video now starts with a workflow. We are launching with 7 core workflows, with plans to add more in the future.
Additionally, users can request a new workflow!
We've made significant performance improvements across the platform to deliver a faster and smoother editing experience:
The mobile preview now loads faster and runs more smoothly across all devices:
We've enhanced the export pipeline to make video exports more reliable and consistent. Exports now complete more successfully, with better error handling and recovery for edge cases.
We've fixed a bug in our AI avatar generation that was blocking some users from seeing their avatar in the editor. We also made it easier to decide whether or not an avatar should be generated for a video by adding a "Narration mode" option in the "Overview" page.
VideoGen 3.0 transforms our platform into a full-featured video editor powered by AI. This release introduces a redesigned three-stage creation flow (Overview, Outline, Editor), a brand-new interactive canvas, and an enhanced timeline editor. We've rebuilt our rendering pipeline for perfect preview-to-export accuracy, added a background task queue for reliable long-running operations, and expanded our stock library with over 12 million new assets. Together, these updates create a more visual, intuitive, and powerful editing experience.
We introduced a redesigned video creation flow built around three stages — Overview, Outline, and Editor — to make project setup and AI collaboration more structured and predictable.
In the Overview page, you can upload images, videos, and audio files that you want the AI agent to use when generating your video. These assets serve as context for the AI — they can appear directly as visuals, help guide topic understanding, or be referenced when building the script and outline.
You can also specify your media sources, such as Free Stock, Wikimedia, iStock, AI Images, or Music. The AI agent will draw from these sources during generation, combining your uploaded assets with external visuals and audio to produce the most relevant media for each scene.
You also have more fine-grained controls, allowing you to define your aspect ratio, duration range, and language.
After submitting your brief, the AI agent creates a structured outline that breaks your video into sections.
Each section is assigned a section type based on its audio handling:
You can review and edit these sections before moving into the editor.
Within each section, you can set featured media, which takes priority over AI-selected b-roll. Featured media ensures specific visuals (like brand clips, demo videos, or uploaded footage) always appear in the final render for that section.
This new three-stage workflow creates a clearer separation between planning, structure, and editing — while giving the AI a stronger context for generating accurate visuals and narration.
We introduced a new layout system that gives users more control over how text and visuals are arranged within each section.
Layouts determine the visual structure of a scene — how the title, subtitle, and media appear on screen — making it easier to match the presentation style to the content type.
The following layouts are now available in the editor:
We've added a brand-new interactive canvas that allows direct manipulation of elements in your video:
These controls are powered by our unified rendering engine, meaning you can see exact, real-time changes to your final composition as you work.
This brings a much more visual and intuitive editing experience — you can now fine-tune positioning, scaling, and animations directly on the canvas without manually entering numbers.
We've redesigned the timeline editor to give you more precise control over your video's timing and structure:
The timeline syncs in real-time with the canvas preview, so every change you make is immediately reflected in your composition. You can scrub through the timeline to preview specific moments, making it easy to fine-tune transitions and timing across your entire video.
We've overhauled our video rendering pipeline so that both preview and final export now run through the same underlying renderer. Previously, previews and exports used slightly different rendering code paths, which could occasionally lead to inconsistencies between what you saw while editing and the final output.
By consolidating them into a unified pipeline:
What you see is what you get – the export will now perfectly match your preview.
Rendering bugs are easier to track and fix because there’s only one rendering path to maintain.
We can introduce advanced editing features more quickly, since any improvements to the renderer apply to both preview and export automatically.
This foundation makes video editing more reliable today and faster to evolve in the future.
We implemented a new background task queue to make long tasks run more reliably, even if you close the tab before the process completes. The following actions will always be executed as background tasks:
With minimal latency, automatic retries, and multiple fallbacks, this new system was built from the ground up to make the video generation experience as seamless as possible for our users.
We’ve expanded the built-in stock media library with over 12 million new assets, including integrations with Pexels Images and Wikimedia Commons. This update provides broader visual coverage across topics, giving the AI agent access to both high-quality stock footage and educational media such as diagrams and public figures.
We overhauled our UX for dealing with failed subscription payments across the entire app. Now, when you attempt to use any paid feature while your subscription is inactive, a modal appears with clear instructions on how to reactivate your subscription. From here, you can view the incomplete invoice, manage your subscription, or contact our customer support team (with the relevant details of your account automatically included in the conversation). There is also a clear warning that your subscription is inactive on the main dashboard with a button to open this modal.
You can now share a copy of your project with your teammates. Click "Share" in top-right corner of the project editor, click "Share a copy", and then enter a comma-separated list of emails you'd like to share the project with. Each recipient will receive a full copy of your project in their inbox, allowing them to edit, generate, and export the video from their own account. Recipients who are not already part of your team will be added to your team upon acceptance of the invitation.
We introduced a new "Generate video clip" tool that fully synthesizes a 5-10 second video based on a prompt. It may take a few minutes to generate, and results are best for well-structured prompts with specific subjects, actions, and settings. We are currently offering this tool exclusively to Business subscribers.

We converted all personal workspaces to single-member teams, making it easier than ever to create videos alongside your teammates. To invite your teammates, simply click "Invite teammates" on the top-right corner of the dashboard and enter their emails. To see a list of all of your team members and modify their permissions, visit the Teams page.

Media tools are a set of flows to create and generate assets in the project editor. You can access these tools in the right side panel by clicking on the asset in the timeline. For a blank asset, the list of available tools will appear directly in the side bar. For a populated non-transcript asset, click "Replace" to replace the asset with the output of a media tool.
The following tools are currently available:
Many more generative AI tools are coming soon!
All videos are now generated with a background music track to complement the content of your video. To power this system, we built an AI music agent that intelligently analyzes your video outline and automatically selects the perfect track from our music library. We also enhanced our music library with many more tracks to cover a wide range of different genres, moods, and tempos.
We reimplemented our timeline and preview to only load what's necessary for the visible portion of your video, allowing for optimized playback of long videos in the project editor. Previously, videos over 10 minutes long could be somewhat laggy.

When you include your own media assets in the video generation form, VideoGen places each of these assets where they are most relevant to the voice-over script. We overhauled our system for this with a new AI agent that understands the content of each asset and intelligently edits together the entire b-roll track. The agent will also choose different animation styles depending on its categorization of the asset (e.g., screenshot, icon, infographic).
![]()
You can now generate an AI avatar on top of your video to present your voice-over script with matching lip movements. Choose from our library of over 100 lifelike presenters to make your videos more engaging and personal. Avatars are currently only available to Business and Enterprise subscribers.
To add an AI avatar to an existing AI voice section, click on the speaker name, click on the avatar button at the top of the popover, select your favorite avatar presenter, and then click generate. Your avatar will be ready to preview and export within a few minutes!
We extended the timeline to have multiple layers to allow for more flexibility and customization in your videos. The bottom layer shows the background assets, which you can trim, split, replace, and rearrange. The middle layer shows the script asset, which corresponds to your AI voice and/or avatar. Finally, the top layer shows your title screen overlay, which you can customize in the "Theme" tab on the left side panel. In the timeline, you can also click on an asset to select it and view more advanced editing capabilities in the right side panel.