ByteDance

How to Create AI Videos With a Reference Character with Seedance 2.0

Seedance 2 Omni adds character reference images to a generation that already accepts up to 12 reference assets — a unique combo of identity lock plus broad multimodal context (audio reference, location reference, palette reference). For an AI influencer producer running high-volume content where each episode varies wardrobe, location, and mood while identity stays anchored, Seedance Omni delivers strong per-clip Sutui economics. It is the pragmatic middle option between Vidu Q2 (densest reference) and Kling O3 Reference (tightest choreography).

Step-by-Step Guide

1

Stack character + context references in the Omni slot

Seedance Omni's 12-asset reference slot is what differentiates it. Build the slot: 4 character references (front, three-quarter, profile, expressive) + 1 wardrobe reference + 1 location reference + 1 palette swatch + 1 audio tone reference. The character references anchor identity; the others anchor episode-specific context. Vidu and Kling do not accept this multimodal density.

2

Pick Omni Pro for the densest reference work

Seedance 2.0 Omni Pro is the variant that combines all capabilities: text-to-video, image-to-video, video-to-video, and reference-driven generation. For a brand spokesperson running across multiple shot types in one campaign, Omni Pro is the right pick. Omni Premium is faster when batch turnaround matters; Standard for cost-aware drafts.

3

Vary wardrobe and location while keeping identity slot stable

For an episodic series, the 4 character references stay fixed across episodes. Only the wardrobe ref + location ref + palette + audio tone slots vary. Save the canvas with the character references locked, then swap only the context references per episode. This is the canvas-as-template pattern at its cleanest.

4

Write prompts in cinematography vocabulary

Seedance reads camera + motion language well. "Slow forward dolly toward Mia, soft golden hour rim light, 1:1 aspect, 6 seconds." Don't describe Mia's face — the references handle that. Describe the camera, the light, the action she takes, and the duration. Cinematography vocabulary unlocks Seedance's motion engine.

5

Use the audio tone reference for series-consistent ambience

Seedance Omni's audio reference slot biases the in-pass ambient sound. For a series with consistent sonic branding (always warm cafe ambience for the lifestyle vlog series), pin one audio tone reference and reuse across all episodes. Audio reads consistent across the series automatically.

6

Cut down to multiple aspects from the same canvas

Seedance supports six aspect ratios. Render the same character + context setup in 1:1 (Meta), 9:16 (TikTok/Reels), and 16:9 (YouTube) by duplicating the node and changing the aspect parameter. The references stay; the aspect changes. Identity reads identical across formats.

Prompt Examples

Identity-locked Meta cutdown. The references do all the visual heavy lifting; the prompt handles camera and timing.

Slow forward dolly toward Mia, soft golden hour rim light, 1:1 aspect, 6 seconds. Wardrobe and location from references.

Vertical TikTok cutdown using the audio tone reference for ambience continuity.

Mia walks left to right through the location, slight handheld breathing, ambient cafe sound from audio reference, 9:16 vertical, 5 seconds

Wide YouTube cut with palette guidance from the swatch reference.

Mia in profile, slow turn toward camera, palette and lighting from palette swatch reference, 16:9, 8 seconds

Parameter Tips

Stack 4 character references + context references (wardrobe, location, palette, audio tone) in the 12-slot Omni reference set.

Save the canvas with character references locked; swap only context refs per episode for the cleanest template pattern.

Don't describe the character's face in prompts — Seedance reads identity from the reference stack.

Use the audio tone reference for series-consistent ambience without a separate audio chain.

Render the same setup in multiple aspects by duplicating the node, not by re-prompting from scratch.

Pick Omni Pro for the densest multi-modal work; Omni Premium when batch speed matters; Standard for drafts.

What to Expect

Seedance 2 Omni outputs 4-15 second clips at 1080p across six aspect ratios with strong character identity from the 4-image reference + multimodal context. Render times: Standard 60-120s, Pro 90-180s. Best for high-volume episodic content where audio + palette + location must hold while identity locks. For tightest identity at maximum reference density use Vidu Q2; for choreographed action with native dialogue lip-sync use Kling O3 Reference; Seedance is the pragmatic middle pick.

Use Seedance 2.0 on Astorie

Connect Seedance 2.0 with other AI models on Astorie's infinite canvas. No GPU required — start free.

Get Started Free

Related features

Docs

Related reading

Try Other Models for This Task

How to Create AI Videos With a Reference Character