Skip to content

Author

Zehuan Yuan

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Bernini: Latent Semantic Planning for Video Diffusion

To better handle multiple visual inputs, Segment-Aware 3D Rotary Positional Embedding (SA-3D RoPE) is introduced, and chain-of-thought reasoning in the planner is incorporated to better transfer understanding into generation.

Chenchen Liu, Junying Chen, Lei Li et al. · 10 citations · ⚡2

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.