Short video generation model for rapid previews and batch creation
Seedance 1.0 Pro Fast is a Pro variant in ByteDance's Seedance 1.0 series designed for rapid creation, suitable for turning text ideas or static images into short videos. doubao-seedance-1-0-pro-fast-251015 can be used for ad drafts, dynamic visual previews, and batch asset exploration, with a focus on validating motion, composition, and shot plans before moving on to editing and sound production.
Input parameters and result formats vary by service. Use the public API for this model and follow its guide for generation, task retrieval and editing operations.
Specifications and API features
Creation method
Text-to-video, image-to-video; images are controlled using the first frame or a combination of first and last frames
Video duration
2–12 seconds, in whole seconds
Output resolution
480p, 720p, 1080p
Aspect ratio options
16:9, 4:3, 1:1, 3:4, 9:16, 21:9, adaptive
Text input
Each text item in content can be up to 1000 characters
Camera and generation controls
Fixed camera, random seed, optional last-frame return
Task delivery
Supports asynchronous queries and callbacks; obtain the video link upon completion; native audio generation is not supported
Seedance 1.0 is natively designed for image- and text-driven video creation; the durations, aspect ratios, and controls above correspond to the calling scope of this model on this platform.
Core Capabilities
Validate shots first, then refine the final video
Pro Fast is suitable for exploring different expressions of the same idea in the early stages of production. Prompts can be organized around the subject, action, scene, lighting, and camera movement, such as comparing the visual difference between a fixed camera and a slow push-in. Use previews first to determine whether the action works, then adjust the aspect ratio and resolution for the selected option.
Bring static images into motion
Image-to-video is suitable for creating from existing product images, illustrations, or scene designs. The first frame determines the video's starting visual, the last frame specifies the ending image, and text describes the action and camera changes in between. It focuses on the starting and ending conditions of the image, rather than a character reference mode that only extracts a person's identity and arbitrarily changes scenes.
Integrate short-video generation into production workflows
With asynchronous tasks, generation can be embedded in asset management or creative tools without keeping the page waiting. Save the task_id after submission, use queries or callbacks to receive completed results, then retrieve the video link. When you need to preview the final composition, you can request the last-frame image to provide a basis for subsequent shot design.
Use Cases
Motion drafts for product advertising
Input a static product image and describe visual actions such as rising steam, swaying fabric, or slow camera movement to generate advertising drafts for discussion. Teams can compare lighting, background atmosphere, and the extent of motion, then pass selected clips to the editing workflow to add brand subtitles, music, and layouts required for formal placement.
Illustration and concept design previs
Use character illustrations or scene concept art as the first frame, paired with concise action descriptions, to observe the pacing and composition after the image is brought to life. It is suitable for early validation of game cutscenes, animated mood films, and visual proposals; the deliverable is a short video clip for discussion, not a production file that automatically completes an entire animation.
Explore short-video assets in batches
Prepare multiple sets of text descriptions around the same theme, changing the subject's actions, shot size, or camera direction to generate candidate clips. Vertical format can be used for mobile content previews, while horizontal format is suitable for presentations or editing assets. Associate each creative version with a task number to facilitate filtering, downloading, and recording directions for subsequent revisions.
How to choose this model
When to choose Pro Fast
When the task focuses on quick previews, concept selection, and batch exploration of short clips, Pro Fast can be prioritized. It is a different invocation model from the standard 1.0 Pro, and is suitable for generating separately and comparing key actions and visual details. Lite is divided into T2V and I2V entry points; when text and first-and-last-frame workflows need to be used together, Pro Fast makes it easier to organize tasks consistently.
When to switch to later versions
If delivery requires audio to be generated together with the video, choose Seedance 1.5 Pro or 2.x with audio support; if multimodal conditions such as character identity, reference video, or reference audio are needed, consider 2.0. For existing video editing, extension, or longer clips, 2.5 is suitable. Pro Fast is better suited to short videos driven by images and text, and does not cover the workflows of these later versions.
Getting started
Organize content and assets
Provide text in content, and use first_frame to define the first frame for images; when an ending image is needed, combine it with last_frame to create a first-and-last-frame input.
Select the correct version and camera settings
Specify model=doubao-seedance-1-0-pro-fast-251015 for /seedance/videos; first test with duration=5, resolution=720p, and a clearly defined aspect ratio. The 1.0 series does not generate native audio.
Save the final video and task records
First obtain the task_id asynchronously, then query /seedance/tasks or receive a callback; after completion, check the subject, action, and ending, and save the selected final video and task records. This model outputs no native audio, so add voiceover and music in post-production when sound is needed.
The first frame shows a camera on a table. The camera slowly circles to the right, highlighting reflections in the lens glass and the metal material, with the background unchanged and no text.
Acceptance and next steps
Validate the composition with a single camera movement, and compare 480p or 720p drafts; output speed cannot be guaranteed solely by the Fast name.
Usage Limitations
This model does not support generating audio via generate_audio. Descriptions of dialogue, music, or sound effects in prompts cannot replace audio-track production; when voice-over, ambient sound, or background music is needed, add them in post-production or use a model that supports audio generation, and avoid treating a silent preview directly as a complete audiovisual production.
Image input uses the first_frame first frame, or a combination of first_frame and last_frame as the first and last frames; do not use reference_image, reference_audio, or reference_video to organize multimodal reference tasks. First and last frames are used to constrain the beginning and ending visuals; they do not mean that every intermediate frame can be reproduced precisely, nor are they equivalent to character identity control in arbitrary scenes.
The creation range is 2–12 seconds and up to 1080p; do not use 4k, 30-second, or automatic duration settings. Long narratives should be split into multiple short clips and then edited together; action design should emphasize the main changes, avoiding cramming too many events, complex transitions, and conflicting shot requirements into a limited duration.
Frequently Asked Questions
How should I choose between Pro Fast and the standard 1.0 Pro?
Pro Fast is designed for quick previews and batch creation, while the standard 1.0 Pro is a separate model. If the main goal is to screen creative ideas, you can start with Fast; if you have strict requirements for specific visual details, it is recommended to compare the results of both using the same assets and prompts before deciding on a production approach.
How should images be submitted?
Add an image_url item to the content array, and place the image address in image_url.url; do not write image_url directly as a string. Use text to describe the action and set first_frame first; if an ending visual is needed, combine it with last_frame to form first-and-last-frame input; do not describe first-and-last-frame tasks as character multimodal reference tasks.
Can it directly generate voice-over or background music?
No, this model belongs to the 1.0 series that does not support generate_audio. You can generate the visuals first, then add voice-over, music, and sound effects during editing. If the task must generate sound at the same time, Seedance 1.5 Pro or a 2.x model that supports audio generation is more suitable.
Can it be used to edit or extend existing videos?
The main purpose of Pro Fast is to generate short videos from text or images; it does not use reference-video editing and extension modes. For such tasks, choose Seedance 2.5 and provide video assets according to the editing or extension requirements; you cannot turn a generation task into an editing task simply by adding related fields to this model.
How do I get the generated video after submission?
Submit model and content to /seedance/videos, and you can set async to true to obtain a task_id, then query the task result; you can also set callback_url to receive completion notifications. After completion, read the video link in data. Applications should distinguish between a submitted task and a generated video to avoid displaying a completed status prematurely.