Level-of-Token Diffusion

Supplementary Results

← Project website

↑ Index

3. Layout Adaptive Video Generation — Qualitative Results

We derive Level-of-Token layouts from depth, semantic masks, bounding boxes, and texture variance to generate videos. These examples show how different layout sources allocate detail over time, with full video playback complementing the still frames in the paper.

Loading selected videos…