ANTI-AI ARCHIVE
ART-HISTORY NODE / 319 · 2025-04-17

Frame Context Packing and Drift Prevention in Next-Frame-Prediction Video Diffusion Models

Frame Context Packing and Drift Prevention in Next-Frame-Prediction Video Diffusion Models

Zhang, Lvmin; Cai, Shengqu; Li, Muyang; Wetzstein, Gordon; Agrawala, Maneesh

Introduction

FramePack packs past-frame context by importance and studies generation drift, making memory and computation in longer video generation an independent research entry.

ORIGINAL DOCUMENT#319
Figure 1: FramePack structures with different historical-frame importance schedules.View full image ↗

Figure 1: FramePack structures with different historical-frame importance schedules. Source: v3; the image version is distinct from the initial submission date.

Zhang, Lvmin; Cai, Shengqu; Li, Muyang; Wetzstein, Gordon; Agrawala, Maneesh · original paper / arXiv · Rights in the paper and depicted works remain with their holders. Research quotation does not establish an open licence; republication rights await independent review.

Source · Frame Context Packing and Drift Prevention in Next-Frame-Prediction Video Diffusion Models ↗

RESEARCH ACCOUNT

FramePack packs past-frame context by importance and studies generation drift, making memory and computation in longer video generation an independent research entry.

ANTI-AI ARCHIVE · Revised 2026-10-03

What the paper investigates

Different historical frames receive different context budgets, allowing a fixed-length input to retain more frames. The paper discusses endpoints, sampling order and discrete history as drift-prevention strategies. The current version revises the original account, so its version date is retained separately.

Section sources: Frame Context Packing and Drift Prevention in Next-Frame-Prediction Video Diffusion Models

Its place in generative-art history

Editorial interpretation: longer sequences depend on what a model retains, forgets and accumulates as error. Duration becomes a production problem with memory and computation requirements rather than merely a number of seconds in a product announcement.

Section sources: Frame Context Packing and Drift Prevention in Next-Frame-Prediction Video Diffusion Models

Reading limits and versions

Fixed context cost does not imply indefinitely coherent video. This account reads the current version but dates the initial public submission.

Section sources: Frame Context Packing and Drift Prevention in Next-Frame-Prediction Video Diffusion Models

Sources for this account

Frame Context Packing and Drift Prevention in Next-Frame-Prediction Video Diffusion Models ↗

Read the arXiv abstract, authors and version history, plus the selected figure/page on PDF page 4; proofs and full experiments were not independently audited.

Read the arXiv abstract, authors and version history, plus the selected figure/page on PDF page 4; proofs and full experiments were not independently audited. The account and translation are AI-assisted, pending independent human review. Section references identify evidence without claiming independent verification of every historical statement.

Continue with a comparative question

Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion ↗

FramePack: compare methods, control or evaluation conditions with Self Forcing.

HunyuanVideo ↗

FramePack: compare methods, control or evaluation conditions with HunyuanVideo.

Sora ↗

FramePack: compare methods, control or evaluation conditions with From images to video.

Dates & version record

Date displayed for this node: 2025-04-17 · Historical date recorded for the source: 2025-04-17

The timeline uses the initial arXiv submission, distinct from conference publication, model release and the archive addition on 3 October 2026. The account and image refer to v3; its date is recorded in the source history.

These dates refer to the historical event or recorded version, not this page’s publication date. The original date precision and unresolved questions are retained.

Original sources & further reading

These links lead to the cited paper, article, institution or conference page. External texts retain their source languages.

01
Frame Context Packing and Drift Prevention in Next-Frame-Prediction Video Diffusion Models ↗

Frame Context Packing and Drift Prevention in Next-Frame-Prediction Video Diffusion Models

https://arxiv.org/abs/2504.12626

Current cited URL

Read the arXiv abstract, authors and version history, plus the selected figure/page on PDF page 4; proofs and full experiments were not independently audited.

Provenance, translation & verification

Archive node #319 · Initial research-paper submission; methods, evaluation and production conditions

Research materials & supplement references · 1

Recent research papers: additions and reconciled sources since 2022 · Read the arXiv abstract, authors and version history, plus the selected figure/page on PDF page 4; proofs and full experiments were not independently audited. · 031
Date as recorded: 2025-04-17
Frame Context Packing and Drift Prevention in Next-Frame-Prediction Video Diffusion Models ↗
Frame Context Packing and Drift Prevention in Next-Frame-Prediction Video Diffusion Models

Archive account based on the listed research materials, not a full translation of the linked work. AI-assisted translation; independent human review pending.

Official material was read within the stated scope; see the source note for reading limits and outstanding checks.

Cite this node

ANTI-AI ARCHIVE. Frame Context Packing and Drift Prevention in Next-Frame-Prediction Video Diffusion Models. Art-history node #319. https://salondesrefuses.cn/en/art-history/354

For specific historical claims, also cite the original sources above and include your access date. This account is not a full translation of the linked work.

Adjacent nodes follow chronological order; adjacency does not establish direct influence or causation.