First-and-last-frame video generation model for rapid creative iteration
veo31-fast is the Fast video model in the Google Veo 3.1 series, suitable for turning text concepts or static images into dynamic shots and quickly comparing different creative directions. It supports text-to-video, single-image first-frame guidance, two-image first-and-last-frame transitions, and visual extension of existing videos. With landscape and portrait aspect ratios and asynchronous task processing, it can be used for ad drafts, product showcases, and storyboard previews.
Clarify capacity, inputs and outputs, and invocation methods before choosing a model.
Creation methods
Text-to-video, image-to-video, existing video extension
Image control
1 image guides the first frame; 2 images guide the first and last frames
Aspect ratios
16:9 landscape, 9:16 portrait
Generation resolution
720p by default; supports 1080p and the get1080p operation
Extension resolution
720p, 1080p
Prompt assistance
Automatic translation can be enabled through translation
Result delivery
JSON task results and video links; supports asynchronous polling and callbacks
The above lists the actual creation and invocation specifications for veo31-fast on this platform; native capabilities of the series are not automatically considered all features available through this entry point.
Core Capabilities
Discover what veo31-fast can bring to your work.
From Descriptions to Dynamic Shots
Text-to-video is ideal for the creative stage when visual assets are not yet available. Enter the subject, environment, action, and camera changes to generate videos for review. You can keep the scene fixed while adjusting only camera movement or action descriptions to compare different creative directions; Fast is better suited to this workflow of exploring first and selecting afterward.
Constrain Transitions with Start and End Frames
Image-to-video is not simply about adding reference images: one image defines the start of a shot, while two images guide both the start and end. Combined with prompts that describe the intermediate action, it can create product reveals, scene changes, or storyboard transitions. It is suited to projects with existing key visuals and makes it more intuitive to compare different transition options.
Continue Developing Existing Shots
With video extension, you can continue actions or camera changes after an already generated clip and use new prompts to guide subsequent content. Extension results can also be extended again, making it suitable for building shots segment by segment. Both generation and extension support asynchronous processing, so applications can first save the task identifier, then receive the completed result and video link.
Use Cases
Start with specific tasks to find where the model can make an impact.
Advertising Concepts and Opening Shot Tests
Enter a product theme, usage environment, and opening action to create different shot drafts for the same advertising concept. Vertical videos can be used for mobile creative reviews, while horizontal videos can be used for presentations or storyboard discussions. The deliverables are playable video candidates that help teams assess visual direction instead of relying only on written scripts to imagine the final result.
Animating Product Images
Use a product image as the first frame, describe the presentation action and camera movement, and generate a dynamic presentation draft; if a target end image already exists, you can submit both the first and last images to explore the change between the two states. It is suitable for product displays, packaging reveals, and concept demonstrations, though product form and details should still be checked before final use.
Storyboard Previsualization and Shot Continuation
Turn storyboard keyframes into video to see whether scene transitions match the narrative intent; if a shot needs to continue moving forward, use the video ID from the generated result to call the extension feature. The deliverables can be used for director discussions, preliminary editing rhythm tests, or concept presentations, helping clarify how shots connect before formal production.
How to Choose This Model
Choose based on task complexity, input materials, and expected results.
Choose Fast for Creative Exploration, Compare with veo31 for Fine-Tuning
When the task focuses on comparing multiple prompts, opening actions, or first-to-last-frame transitions, prioritize veo31-fast. If you place greater emphasis on the refined presentation of the final visuals, use the same materials to compare with veo31 in Quality mode. The trade-off between the two should be based on specific shot results; do not interpret Fast as a version with fixed processing time or as better suited for every task.
Choose First-to-Last Frames and Multi-Image Fusion Separately
veo31-fast is suitable for text-only creation, as well as using one or two images to control the start and end of a shot. If the task is to combine elements from multiple images, choose veo31-fast-ingredients, whose creation method requires image input. The two solve different problems: the former emphasizes temporal start-to-end changes, while the latter emphasizes combining content from multiple images, so their image rules should not be interchanged.
Get Started
From a small-scale task to production integration.
01
Prepare the Task and Materials
Define the goal, required inputs, and output requirements, using real business examples as a starting point.
02
Try It in the API Testing Area
Open the trial page, confirm the parameters supported by this entry point, then submit a small-scale task to review the results.
03
Integrate According to the API Documentation
Keep the complete model ID, use the request format specified in the documentation, and confirm billing rules on the Pricing page.
Usage Limits
Before formal use, understand the output quality and capability scope.
Image inputs are used to guide the first frame or first-to-last frames, with a maximum of two images, and should not be used as multi-image fusion. When submitting two images, it is recommended that changes in the subject and scene have explainable continuity, and that you check whether intermediate frames show unexpected shape changes.
Extensions must use data[].id from the current generation or completed extension result, rather than any external video URL, and do not use task_id in place of the video ID. Video IDs from earlier generation records may not be extendable; you can regenerate them before continuing production.
Extension results can be extended further, but cannot be used again for shot motion modifications in /veo/reshoot or object additions and removals in /veo/objects. If your production workflow still requires these edits, retain the original generated video and plan the order of editing and extension in advance.
Frequently Asked Questions
Answers to common questions about using veo31-fast.
How do I start text-to-video with veo31-fast?
Submit model=veo31-fast, action=text2video, and prompt to /veo/videos. The prompt can be organized by subject, scene, action, and camera changes; set aspect_ratio=9:16 when you need portrait orientation. After completion, get the video from video_url in the result.
What is the difference between one and two images?
When using image2video, one image is used for first-frame guidance, while two images are used for first-and-last-frame guidance. Put the image links in image_urls, and use prompt to describe the desired actions or transitions in between. This model is not a blending mode that freely combines multiple images into a scene.
How do I get 1080p video?
You can select resolution=1080p during generation, or use the get1080p action on an already generated video and submit the corresponding video_id. Generation, obtaining a higher-resolution version, and extending are different actions; extension requests use the 720p or 1080p option of /veo/extend.
How do I extend a shot generated by veo31-fast?
Submit model=veo31-fast and data[].id from the generation result as video_id to /veo/extend, and you can optionally add prompt to guide the subsequent footage. Extension results can also be extended further; save the task ID and video ID separately to avoid using the task query identifier for extension.
How do I use Chinese prompts and asynchronous calls?
You can enter Chinese prompts and enable automatic translation with translation=true. For application integration, you can set async=true, obtain task_id first, and then query the result; you can also provide callback_url to receive completion notifications, then save the video ID and download link for subsequent production.
Model information · Updated: 2026-10-01. See the API and pricing sections for call parameters and billing rules.