Skip to main content
Generate videos using ByteDance SeeDance 2.0 Pro RFR — the highest quality tier. Supports text-to-video, first/last frame, and multimodal reference workflows with image, video, and audio references. Output durations are 4–15 seconds at resolutions up to 4K (10-bit H.265).
RFR — Real Face Restriction. This model rejects reference images and videos that contain real human faces, unless the likeness is authorized on the provider side. Use the non-RFR SeeDance 2.0 Pro if your references include real people.

Model

Request types

Parameters

This model does not accept seed or camera_fixed. Output is mp4 at 24 fps.

Example - Text-to-Video

Example - First & Last Frame

Example - Multimodal Reference

Frame image_urls are fetched server-side, so they only need to be reachable by Unifically. Reference video/audio URLs are forwarded to the provider directly and must be publicly downloadable.

Response