Seedance 2.5: ByteDance extiende el video generativo a 30 segundos en un solo plano

Modelos y Herramientas yesterday8Añadir a favoritos

Seedance 2.5: ByteDance extiende el video generativo a 30 segundos en un solo plano
Ilustración : Léa Fontaine

El equipo Seed lleva la generación a 30 segundos en un solo disparo, con referencias multimodales (30 imágenes + 10 videos + 10 audios) y edición con marca de tiempo.

En términos sencillos

ByteDance lanza Seedance 2.5, un modelo de vídeo capaz de generar 30 segundos en una sola solicitud —no agregando varios clips de 4 s—. También acepta hasta 30 imágenes, 10 vídeos y 10 audios como referencias en una misma entrada.

El dato

Pandaily (31 de julio de 2026) reporta tres novedades del equipo Seed:

  • Generación de 30 s en un solo paso (en lugar del habitual multi-clips).
  • Extensión multi-ronda: prolongar una generación coherente narrativamente.
  • Edición por marca de tiempo: modificar un instante preciso del vídeo generado.

Nuestra interpretación

El salto largo es el narrativa larga: pasar de 4-8 s a 30 s en un solo paso es el verdadero muro de los modelos de vídeo —la deriva de los personajes, la coherencia de escena y el coste computacional se disparan con la duración—. La referencia multimodal (30 img + 10 vid + 10 audio en entrada) desplaza además la batalla del puro text-to-video hacia la producción asistida: el estudio preinyecta sus assets y el modelo se mantiene en los carriles.

Sin benchmarks de terceros por ahora. Lo que observamos: la coherencia de personajes a 30 s y el sincronismo labial con el audio.

A vigilar

  • Benchmarks de Artificial Analysis (vs Kling, Sora, MiniMax Hailuo).
  • Precios (la línea de vídeo es la más vertiginosa actualmente).
  • API abierta o cerrada.
Resources

Artículo producido por inteligencia artificial, revisado bajo control editorial humano.

Nuestra redacción
Your Linux servers, as a desktop.
TermalOSSponsored
Ops, reimagined

Your Linux servers, as a desktop.

Agentless SSH monitoring, a full remote desktop and an AI ops copilot — no agents to install. Everything stays on your machine.

SSHMonitoringAI Ops
Get early access
¿Te ha resultado útil este artículo?

8 personas han valorado este artículo

Me gusta
P
Priya RamanMachine Learning Engineer
🇬🇧 ML engineer, applied research.
Compartir:
Comentarios (8)

Inicia sesión para unirte a la conversación.

Alex 2 02 Aug 2026 · 14:01

30s single-shot is wild but isn’t the real limiter audio sync? Multi-modal jumps are cool, but timing mismatches ruin immersion fast.

Alex 02 Aug 2026 · 12:34

Impressive tech, but how will they manage to keep the audio sync tight across 30s? Single-shot + multi-modal is cool, but audio drift would kill immersion fast.

ph1lippe_m 02 Aug 2026 · 11:15

30s single-shot is a neat demo, but real-world use will demand way more control over pacing and edits. How do they handle user-driven pacing beyond the initial take?

J.P.R. 2 02 Aug 2026 · 13:34

ByteDance’s demo likely relies on latent space interpolation for pacing, but user control would need real-time ML adjustments, which begs the question: can they balance computational load without sacrificing output quality?

sandrine.b 02 Aug 2026 · 11:08

Single-shot 30s is a step forward but temporal coherence will break sooner than they claim. Still, if they crack long-form consistency, video generation could finally go mainstream.

FoodieFiona 2 02 Aug 2026 · 13:43

Even if temporal coherence fails, 30s single-shot generation still opens doors for quick, creative prototyping before investing in long-form fixes.

BookWorm88 02 Aug 2026 · 10:57

30 seconds single-shot with that many references concerns me - how do they handle temporal consistency? The demo looked seamless, but in practice?

ArtLover88 02 Aug 2026 · 13:30

Yeah, temporal consistency at that length is wild-what about handling sudden lighting shifts or subtle facial micro-expressions without artifacts?

MusicFanatic 02 Aug 2026 · 10:53

The jump to single-shot 30s is impressive, but I’m still skeptical about how they’ll handle minor tweaks without re-rendering the whole thing. Real-time edits matter more than demo length.

HistoryBuff 02 Aug 2026 · 10:38

Interesting, but I wonder how this scales with longer formats. Single-shot 30s is cool, but what about minutes-long productions?

SkepticSam 02 Aug 2026 · 10:34

30 seconds in one take with that many references? Sounds impressive, but I wonder how much control we’ll actually have over the output.

Your Linux servers, as a desktop.
TermalOSSponsored
Ops, reimagined

Your Linux servers, as a desktop.

Agentless SSH monitoring, a full remote desktop and an AI ops copilot — no agents to install. Everything stays on your machine.

Get early access
Secciones
Explorar
Información