Skip to content

RecEdit-Drive:3D Reconstruction-Guided Spatiotemporal Video Editing for Autonomous Driving Scenes

· 0 citations · 67 references

TL;DR

RecEdit-Drive is introduced, a framework that integrates Spatial Feature Warping and Spatiotemporal Collaborative Modeling to effectively control 3D object variations and enhance video consistency, and an inference strategy to reconstruct an accurate background structure through noise manipulation is designed.

View source

Similar papers

Preprint Aug 2026

SPVC: Structured and Panoptic Video Fixing for Cross-Dataset Driving Scene Rendering

Driving scene reconstruction and rendering, especially with 3D Gaussian Splatting, has become an important component of autonomous driving simulation. However, rendered views often degrade under extrapolated ego trajectories and scene edits, producing blurry structures, temporal flicker, and foreground-background misal...

Gen Li, Shu Han, Yun-Xi Qiao et al. · 0 citations
#computer vision Preprint Aug 2026

MaskFlow: Precise, Consistent and Seamless Regional Image Editing

The proposed MaskFlow, a training framework for precise localization, consistent background preservation, and seamless boundary transitions, incorporates the mask into the probability path and flow-matching objective, coordinating generation within the editable region with source preservation outside it.

Rui Xu, Yang Yong, Shun-Zi Yang et al. · 0 citations
Preprint Sep 2026

DecoGS: Adaptive Static-Dynamic Decoupling of 3D Gaussians for Free-Viewpoint Video Streaming

Streaming 3D reconstruction demands both speed and temporal fidelity, goals that existing methods undermine by updating every Gaussian every frame, even in static regions. We present DecoGS, a method for efficient online training of 3D Gaussians from streaming videos. Unlike prior methods that update the entire scene i...

Idil Sulo, Alexey Supikov, Ilke Demir et al. · 0 citations
Preprint Aug 2026

ScaleVid: Geometry-Aware Video Object Scaling with Mesh-Free Inference

This work presents a progressive two-stage training framework that decouples geometry-aware foreground transformation from background preservation and realistic video composition, without mesh-pixel alignment and explicit 3D reconstruction at inference.

Youze Huang, Peng-Hui Ruan, Bojia Zi et al. · 0 citations
Aug 2026

Dynamic View Synthesis from Monocular Videos via Motion-aware Gaussian Splatting.

This paper proposes a semantics-guided scene decoupling module that separates Gaussian primitives into static and dynamic components based on motion vectors, and introduces a motion-aware densification module for motion compensation, which alleviates the incomplete rendering of dynamic objects caused by insufficient sp...

Chulin Zhao, Xue Wang, Guo-Qing Zhou et al. · 0 citations
Preprint Aug 2026

EditaLive! Unified Character Video Editing for Live Streaming

Conventional video editing primarily focuses on scene-level content, whereas live streaming places greater emphasis on the human subject. However, directly applying existing video-editing methods to human-centric live streaming remains challenging, as they may introduce facial-expression inconsistencies and typically d...

Zhiyuan Li, Chi-Man Pun, Peng-Tao Jiang et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.