Chinese AI Tools
ProductsModelsIntegrationsRankingsLatest changes
TopicsAvailability TrackerUse casesSubmit a toolAccount
ZHSearch

Chinese AI Tools

Independent directory for Chinese AI products. Product availability, pricing and terms can change. Verify before commercial use.

Editorial standardsClaim productUpdate infoGet featuredAdvertise

Guide

Seedance 2.0 Prompt Guide: Multimodal Video Templates

A practical Seedance 2.0 prompt structure for text, image, video and audio references, with reusable templates for generation, editing and extension.

Published 2026-08-26 · Updated 2026-08-26

Verdict

Write one primary action, then specify the subject, environment, camera, lighting, audio and continuity constraints. Assign a clear role to every reference asset, and change one variable at a time during testing.

Ranking basis

This guide was checked on August 26, 2026 against ByteDance's official Seedance 2.0 capability page and the BytePlus plan page. It documents workflow structure, not a universal prompt formula. BytePlus currently presents conflicting 4K and 1080p resolution claims, so output limits must be verified in the live product before production.

ByteDance documents Seedance 2.0 as a multimodal video system that accepts text, image, audio and video inputs, supports joint audio-video generation, and exposes controls for performance, lighting, shadows, camera movement, reference-based editing and extension. The templates below are editorial starting points derived from those documented controls; they are not official Seedance prompt strings or independently benchmarked guarantees.

Prompt structure

Use a fixed order so failed generations are easier to diagnose. Keep the requested action physically coherent and avoid stacking unrelated scene changes into one short clip.

Subject, action and environment

Name the main subject, one observable action and the setting. Example structure: [subject] [single action] in [environment], with [important continuity constraint].

Camera and light

Add shot size, camera path, speed, focus behavior, lighting direction and shadow behavior only after the action is clear.

Audio and continuity

Describe dialogue, ambience and timing separately. State which identity, wardrobe, object, background or motion must remain consistent.

Text-to-video template

Template: [subject and appearance] [single action] in [environment]. [shot size], camera [movement] at [pace], focus on [visual priority]. [lighting direction and quality], [shadow behavior]. Audio: [dialogue or sound], timed to [event]. Preserve [identity, wardrobe, object and background constraints].

Short product shot

A brushed-steel coffee grinder rotates once on a clean studio table. Medium close-up, slow 30-degree orbit, focus locked on the front dial. Soft key light from camera left with a crisp contact shadow. Audio: one quiet mechanical click at the end. Preserve the logo, dial markings and table position.

Reference and editing templates

Tell the model what each uploaded asset controls. Do not assume that a reference image, video or audio clip will automatically be interpreted as identity, composition, motion or sound guidance.

Image reference

Use image 1 for the character identity and wardrobe; use image 2 for the room layout and color palette. Animate only [action]. Keep facial features, clothing details and background geometry unchanged. Camera [movement]; lighting [description]; audio [description].

Motion or audio reference

Use video 1 only for movement timing and camera rhythm; do not copy its subject or location. Use audio 1 for beat and event timing. Render [new subject and scene], with [continuity constraints].

Edit or extend

Continue from the final frame without a cut. Extend [action] for [requested duration], preserve direction of travel, camera velocity, lighting, ambience and subject identity, then end on [specific final state].

A controlled testing loop

Start with a short draft, inspect the failure mode, and revise one field. Keep reference assets and any exposed random seed fixed while comparing camera, motion or audio variants.

Score five dimensions

Review motion coherence, identity continuity, reference adherence, camera execution and audio synchronization separately instead of using one overall impression.

Verify live limits

Set duration and resolution in the product controls or API rather than embedding them only in prose. Recheck the live account because the checked BytePlus page contains conflicting resolution statements.

Sources

ByteDance Seedance 2.0 official model pageBytePlus Seedance 2.0 plan page

Next actions

  • - Run the same short scene with one camera variable changed per version.
  • - Check the live account for supported duration, resolution, token cost and commercial terms before batch production.
  • - Store the prompt, reference assets, settings and output review together so successful shots can be reproduced.

Related product profiles

ByteDance Seed / Doubao ArkSeedance 2.0