Extend video duration
Deprecated: After October 26, 2026 at 11:59 PM UTC, requests to this endpoint will no longer work. See Migrate from V1 to V2.
Extend a video by generating additional frames at the beginning or end. The model uses context frames from the input to produce a seamless continuation with consistent motion and audio.
Audio is generated for the extended portion if the input video has audio.
Returns the generated video directly in the response. For an asynchronous version that returns a job to poll, see v2/extend.
Billed per second, based on the extended portion plus the context frames used from the input video. See Pricing.
Authentication
Request
Input video for extending. See Input Formats for supported formats and codecs.
- Supported aspect ratios: 16:9 and 9:16
- Maximum resolution: 3840x2160 (4K)
- Minimum frame count: 73 (around 3 seconds at 24fps)
The output video preserves the input video’s resolution.
Duration in seconds to extend the video. Minimum 2 seconds, maximum 20 seconds (480 frames at 24fps).
Where to extend the video:
end(default): Extends the video at the end.start: Extends the video at the beginning.
Advanced parameter: Number of seconds from the input video to use as context for the extension (maximum 20 seconds).
The model uses context frames from the input video to generate a more coherent extension.
The sum of context + duration (converted to frames using the input video’s FPS) cannot exceed 505 frames (~21 seconds at 24fps). For higher-FPS inputs, the maximum total duration in seconds will be proportionally lower; for lower-FPS inputs, it will be proportionally higher.
If not provided, defaults to maximize available context within the 505 frame limit while respecting the 20-second cap.