Skip to main content
Generate videos using ByteDance SeeDance 2.0 Mini RFR — the smallest, most affordable 2.0 tier. Supports text-to-video, first/last frame, and multimodal reference workflows with image, video, and audio references. Output durations are 4–15 seconds at 480p or 720p.
RFR — Real Face Restriction. This model rejects reference images and videos that contain real human faces, unless the likeness is authorized on the provider side. Use the non-RFR SeeDance 2.0 Mini if your references include real people.

Model

Request types

Parameters

This model does not accept seed or camera_fixed. Output is mp4 at 24 fps.

Example - Text-to-Video

Example - First & Last Frame

Example - Multimodal Reference

Frame image_urls are fetched server-side, so they only need to be reachable by Unifically. Reference video/audio URLs are forwarded to the provider directly and must be publicly downloadable.

Response