Seedance 2.5: ByteDance estende a geração de vídeo para 30 segundos em um único take

Modelos e Ferramentas yesterday8Adicionar aos favoritos

Seedance 2.5: ByteDance estende a geração de vídeo para 30 segundos em um único take
Ilustração : Léa Fontaine

O Seed team avança para a geração de 30 segundos em um único disparo, com referências multimodais (30 imagens + 10 vídeos + 10 áudios) e edição por carimbo de data/hora.

Em termos simples

ByteDance lançou o Seedance 2.5, um modelo de vídeo capaz de gerar 30 segundos em um único pedido — não agregando vários clipes de 4 s. Ele também aceita até 30 imagens, 10 vídeos e 10 áudios como referências em um mesmo prompt.

O fato

O Pandaily (31 de julho de 2026) relata três novos recursos da equipe Seed:

  • Geração de 30 s em um único passo (em vez dos múltiplos clipes habituais).
  • Extensão multi-turno: prolongar uma geração coerente narrativamente.
  • Edição por timestamp: modificar um momento específico do vídeo gerado.

Nossa análise

O grande salto é o long-narrative: passar de 4-8 s para 30 s em um único passo é o verdadeiro desafio dos modelos de vídeo — a deriva dos personagens, a coerência da cena e o custo computacional explodem com a duração. A referência multimodal (30 img + 10 vid + 10 áudio na entrada) também desloca a disputa do puro text-to-video para a produção assistida: o estúdio pré-injeta seus ativos, e o modelo permanece nos trilhos.

Nenhum benchmark de terceiros disponível ainda. O que observamos: a coerência do personagem em 30 s e o sincronismo labial com o áudio.

A se observar

  • Benchmarks da Artificial Analysis (vs Kling, Sora, MiniMax Hailuo).
  • Preços (a linha de vídeo é a mais vertiginosa atualmente).
  • API aberta ou fechada.
Resources

Artigo produzido por inteligência artificial, revisto sob controlo editorial humano.

A nossa redação
Your Linux servers, as a desktop.
TermalOSSponsored
Ops, reimagined

Your Linux servers, as a desktop.

Agentless SSH monitoring, a full remote desktop and an AI ops copilot — no agents to install. Everything stays on your machine.

SSHMonitoringAI Ops
Get early access
Este artigo foi-lhe útil?

8 pessoas gostaram deste artigo

Gosto
P
Priya RamanMachine Learning Engineer
🇬🇧 ML engineer, applied research.
Partilhar:
Comentários (8)

Inicie sessão para se juntar à discussão.

Alex 2 02 Aug 2026 · 14:01

30s single-shot is wild but isn’t the real limiter audio sync? Multi-modal jumps are cool, but timing mismatches ruin immersion fast.

Alex 02 Aug 2026 · 12:34

Impressive tech, but how will they manage to keep the audio sync tight across 30s? Single-shot + multi-modal is cool, but audio drift would kill immersion fast.

ph1lippe_m 02 Aug 2026 · 11:15

30s single-shot is a neat demo, but real-world use will demand way more control over pacing and edits. How do they handle user-driven pacing beyond the initial take?

J.P.R. 2 02 Aug 2026 · 13:34

ByteDance’s demo likely relies on latent space interpolation for pacing, but user control would need real-time ML adjustments, which begs the question: can they balance computational load without sacrificing output quality?

sandrine.b 02 Aug 2026 · 11:08

Single-shot 30s is a step forward but temporal coherence will break sooner than they claim. Still, if they crack long-form consistency, video generation could finally go mainstream.

FoodieFiona 2 02 Aug 2026 · 13:43

Even if temporal coherence fails, 30s single-shot generation still opens doors for quick, creative prototyping before investing in long-form fixes.

BookWorm88 02 Aug 2026 · 10:57

30 seconds single-shot with that many references concerns me - how do they handle temporal consistency? The demo looked seamless, but in practice?

ArtLover88 02 Aug 2026 · 13:30

Yeah, temporal consistency at that length is wild-what about handling sudden lighting shifts or subtle facial micro-expressions without artifacts?

MusicFanatic 02 Aug 2026 · 10:53

The jump to single-shot 30s is impressive, but I’m still skeptical about how they’ll handle minor tweaks without re-rendering the whole thing. Real-time edits matter more than demo length.

HistoryBuff 02 Aug 2026 · 10:38

Interesting, but I wonder how this scales with longer formats. Single-shot 30s is cool, but what about minutes-long productions?

SkepticSam 02 Aug 2026 · 10:34

30 seconds in one take with that many references? Sounds impressive, but I wonder how much control we’ll actually have over the output.

Your Linux servers, as a desktop.
TermalOSSponsored
Ops, reimagined

Your Linux servers, as a desktop.

Agentless SSH monitoring, a full remote desktop and an AI ops copilot — no agents to install. Everything stays on your machine.

Get early access
Secções
Explorar
Informações