ByteDance released Seedance 2.5, a video model that generates clips of up to 30 seconds in a single run and accepts as many as 50 multimodal references — images, text, style frames, character shots — to steer the result. Audio and video are generated jointly, and the model keeps character appearance, lighting and motion style consistent across the whole clip. It is rolling out on ByteDance's Jimeng AI and Doubao Pro apps, with an API coming to the Volcano Engine platform.