docs
models / ByteDance / Seedance 2.0

Seedance 2.0

videobytedance/seedance-2.0

ByteDance's next-generation video model with a unified multimodal architecture. Generates high-quality video with synchronized audio from text, images, video clips, and audio inputs. Supports multimodal references (up to 9 images, 3 videos, 3 audio files), native audio generation, video editing, video extension, intelligent duration, and adaptive aspect ratio.

$0.15list priceper second of video
pricing
unitlist price
per second of video$0.15
per second of 1080p video (without video input)$0.37
per second of 1080p video (with video input)$0.914
per second of 480p video (without video input)$0.07
per second of 480p video (with video input)$0.172
per second of 4K video (without video input)$0.78
per second of 4K video (with video input)$1.87
per second of 720p video (without video input)$0.15
per second of 720p video (with video input)$0.372

Everyone pays the list price. Reaching Elite (hold or stake 1,337 $XAVA, or hold 1,337,000 $XINF) unlocks the Elite price and a bigger buyback of your token. How it works

curl
curl https://zinf.ai/v1/media/generations \
  -H "Authorization: Bearer $XINF_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"bytedance/seedance-2.0","input":{"prompt":"a pixel-art rocket lifting off","duration":4}}'
limits
providerByteDance
typevideo
context window—
max output—
prompt storagenone
provider retentionthe model provider's policy
livesoon
requests 24h—
p50 latency—
uptime 7d—