Heavy rainfall severely degrades outdoor videos by corrupting high-frequency details and introducing motion blur, critically undermining the reliability of visual tasks. Recently, State Space Models (SSMs), particularly Mamba, have emerged as efficient alternatives for vision tasks with their linear complexity and abil...
Kui Jiang, Yiang Chen, Yan Luo et al.· 0 citations
A novel framework, VFC-Net, which generates a uniformly distributed coarse point cloud to effectively guide dense reconstruction and introduces a lightweight VoxAttn module in both stages to efficiently capture missing geometric structures.
Guo-Qing Zhang, Wen-Bo Zhao, Yuanchao Bai et al.· IEEE Transactions on Image P...· 0 citations
ViP-Rig is a visual-prompted framework that supports both prompt-first rigging and result-guided editing by injecting features extracted from user-drawn or edited 2D skeletal and rigidity prompts into frozen pretrained backbones into a frozen pretrained autoregressive generator.
Zihan Qin, Ming-Ze Sun, Yifan Mao et al.· arXiv.org· 1 citation
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.