Seedance 2 Video Edit: AI Video Editor

SD 2.0 Video Edit modifies existing videos based on text prompts and optional reference images.

📝

Overview

About this model

SD 2.0 Video Edit is an advanced AI video modification tool from ByteDance. It unleashes multi-shot storytelling by allowing you to seamlessly edit and transform existing videos using natural language prompts. Whether you want to perform style transfers, change backgrounds, or modify specific elements within a scene, this model offers director-level control while maintaining original motion consistency and narrative structure.

1Style Transfer: Apply new visual styles (e.g., cyberpunk, anime, cinematic) to existing footage without losing physical movement.
2Background Replacement: Intelligently swap out environments behind your subjects.
3VFX and Enhancements: Add dynamic elements or enhance lighting and atmosphere in post-production.
4Character Modification: Subtly or dramatically alter the appearance of characters or objects in your scene.
5Creative Remixing: Transform stock footage into custom, branded content.
💰

Pricing & Value

Cost analysis

muapiapp$0.60 base price

muapiapp brings ByteDance's cutting-edge video editing to you early, providing premium video-to-video manipulation starting around $0.60 per edit depending on duration and quality.

CompetitorsVaries

Many standard APIs lack native video-to-video capabilities of this tier, making muapiapp an exclusive entry point to Hollywood-grade AI editing.

* Competitor pricing is estimated based on similar model architectures and usage tiers.

⚙️

Technical Details

Configuration schema

Promptstring

Text prompt describing the video edit.

Default ValueReplace the running man with @image1. Preserve the exact running motion, speed, and camera shake. Ensure the armor glows dynamically with the environment lighting and reflects passing car lights. Maintain realistic foot contact with the ground and motion blur consistency.
Video URLsarray

Upload up to 1 video URL. (Max size: 10MB, Max duration: 15s)

Default Valuehttps://d3adwkbyhxyrtq.cloudfront.net/webassets/videomodels/seedance-v2.0-video-edit-in.avif
Image URLsarray

Upload up to 9 image URLs.

Default Valuehttps://d3adwkbyhxyrtq.cloudfront.net/webassets/videomodels/seedance-v2.0-video-edit.jpg
Audio Reference URLsarray

Up to 3 reference audio clip URLs (MP3/WAV, total max 15s). Each Nth audio corresponds to @audioN in the prompt.

Default Valueundefined
Aspect RatioEnum (4 options)

-

Default Value16:9
QualityEnum (2 options)

-

Default Valuebasic
Duration (seconds)int

Output video duration in seconds (4–15).

Default Value5
📖

Implementation Guide

Developer documentation

How to Use SD 2.0 Video Edit

  1. Upload Your Source Video: Provide a valid URL to the target video via the video_urls array (limit 1 video).

  2. (Optional) Add Reference Images: Include up to 9 image URLs via images_list if you want the model to mimic a precise aesthetic, character, or object style during the edit.

  3. Write Your Edit Prompt: Clearly describe what should happen in the video. If you are uploading reference images, use @image1, @image2 to guide the model (e.g., 'Transform the environment into the city from @image1').

  4. Select Duration: Pick output video length (4–15s).

  5. Select Quality Settings: Choose basic ($0.21/sec output + $0.063/sec per input video second) for faster, cost-effective rendering, or high ($0.30/sec output + $0.09/sec per input video second) for maximum visual fidelity. Input source video incurs a 30% surcharge based on its duration.

  6. Submit Request: Once submitted, the model will output a seamlessly edited video matching your prompt.

Common Questions

Frequently asked

Can I edit the video's audio track?

Currently, the SD 2.0 video edit model focuses strictly on visual transformations. Audio manipulation is generally handled independently.

How long can the source video be?

For optimal results and to stay within API processing limits, it's recommended to provide source videos under 15 seconds.

What is the difference between 'basic' and 'high' quality?

'Basic' uses a faster rendering pipeline ideal for rapid prototyping, while 'high' utilizes the standard, deep-rendering pipeline for production-ready, cinematic quality.