Volcengine released Seedance 2.5, the next-generation video generation model from ByteDance. The biggest improvement: single-shot 30-second video generation with consistent quality, crossing the "production-usable" threshold for advertising and short-form content.

The technical details: Seedance 2.5 is a DiT-based video model with 3D full attention (spatial + temporal). The model is trained on a mixture of 5-second clips (for short-range consistency) and 60-second clips (for long-range consistency), with a "temporal curriculum" that gradually increases the training clip length. The result: 30 seconds of generated video with no quality degradation over time.

The benchmark: on the "long-video consistency" benchmark, Seedance 2.5 scores 87.4, compared to 71.2 for Seedance 2.0 (the previous version). The biggest improvement is in motion smoothness — the generated video has no "jitter" or "popping" artifacts over 30 seconds.

The "production-usable" threshold: previous video models had a "production-usable" length of ~5 seconds — anything longer had visible quality degradation. Seedance 2.5's 30 seconds opens up new use cases: short-form ads, social-media clips, product demos, and game cutscenes. The model is also available with a "multi-shot" mode that can generate 60+ seconds with explicit shot boundaries.

The commercial angle: Seedance 2.5 is available via Volcengine's API at $0.08 per second of 1080p video. The first batch of enterprise customers includes Bilibili (for creator tools), iQiyi (for drama production), and several advertising agencies.

The bigger takeaway: "30 seconds" is the new production threshold for video generation. The "5-second clip" era is ending, and "30-second production-usable" is the new baseline. The next round of competition will be in "60 seconds + multi-shot" — can a model generate a full short film in one go?