ByteDance's next-generation audio-visual generation model with a 4.5B parameter Dual-Branch Diffusion Transformer architecture. Seedance 1.5 Pro generates video and audio simultaneously in a single unified pass — eliminating the timing issues of sequential audio dubbing. Supports multi-language lip-sync (English, Mandarin, Japanese, Korean, Spanish, and more), cinematic camera control (pan, tilt, zoom, orbit), multi-character dialogue, and character consistency across shots. Produces clips from 4–12 seconds at up to 1080p. The number of tokens is given by (height of output video * width of output video * duration * 24) / 1024
Modalities
Price
from $0.02306/second
Released
Mar 23, 2026
This model is hosted by one provider. OpenRouter forwards every request to it directly — no routing decisions to make.
Throughput is how fast the model writes (tokens per second — higher is better). Latency is total round-trip time (lower is better). TTFT is time-to-first-token — how long before you see anything appear (lower is better).
Uptime is the percentage of the past 3 days that at least one provider was responding to requests. Availability is the percentage of time that inference was successfully served. OpenRouter continuously monitors and uses the next-best provider when one returns an error.
Scores on standardized evaluations. Higher percentages are better — and rank percentile shows where this model lands among all models on OpenRouter.
| Source | Benchmark | Score |
|---|---|---|
| Artificial Analysis | Seedance 1.5 pro Elo | 1,173 |
Public apps that send the most traffic to this model. Good signal for what real production workloads look like — and a hint at which use cases this model is best suited for.
Token volume and request traffic to this model over time.
Drop-in code to call this model. OpenRouter's API is OpenAI-compatible — most SDKs work by just swapping the base URL. The only thing that changes between models is the model slug below.
| $2.40 | $0.05184 | $1.20 | $0.02592 | 52.72s |
End-to-end latency
52.72s
P50, best provider
100.00%
98.87%
When an error occurs in an upstream provider, we can recover by routing to another healthy provider, if your request filters allow it. You can access per-provider uptime data programmatically through the Endpoints API. Learn more about our load balancing and customization options.
ByteDance's next-generation audio-visual generation model with a 4.5B parameter Dual-Branch Diffusion Transformer architecture. Seedance 1.5 Pro generates video and audio simultaneously in a single unified pass — eliminating the timing issues of sequential audio dubbing.
Seedance 1.5 Pro costs from $0.02306/second at 480p to $0.4666/second at 4K for Video (with audio), from $0.01153/second at 480p to $0.2333/second at 4K for Video (no audio), $2.40/M tokens for Video Tokens (with audio) and $1.20/M tokens for Video Tokens (no audio).
Seedance 1.5 Pro supports 4–12 second clips, can be steered with a first frame and last frame image and produces a matching audio track.
Seedance 1.5 Pro accepts text and images as input and returns video.
Seedance 2.0 Mini, Seedance 2.5, Seedance 2.0 and 1 more are other video models from ByteDance.
Seedance 1.5 Pro was released on March 23, 2026.