Skip to content
Preprint

SUMI: Scalable Unified Model for 3D Point Cloud Inference

Aug 2026 · 0 citations · 77 references
Computer Science

TL;DR

SUMI injects noisy geometric features into cross-attention with coarse structural features, enabling reverse denoising to refine local geometry while preserving global consistency in coarse-to-fine point cloud completion.

Abstract

Point cloud completion commonly follows a coarse-to-fine paradigm, where a low-density coarse shape is first predicted and then upsampled to the target resolution. Although recent methods have improved global structure recovery, the fine stage often remains limited by simple upsampling and insufficient interaction with coarse structural features, making local detail reconstruction challenging. We propose SUMI, a diffusion-enhanced refinement module for coarse-to-fine point cloud completion. Unlike prior diffusion-based completion methods that use diffusion as a standalone point generator, SUMI injects noisy geometric features into cross-attention with coarse structural features, enabling reverse denoising to refine local geometry while preserving global consistency. SUMI can also be integrated into existing coarse-to-fine models as a flexible refinement module. Experiments on PCN, ShapeNet-55/34, and MVP demonstrate consistent improvements over strong baselines. SUMI achieves the best overall CD and F1-score on PCN, reduces CD by up to 16.1% on ShapeNet-55, and obtains the best CD across all output densities on MVP.

View source

Similar papers

Sep 2026

VFC-Net: Point Cloud Completion via Voxel-based Transformer.

A novel framework, VFC-Net, which generates a uniformly distributed coarse point cloud to effectively guide dense reconstruction and introduces a lightweight VoxAttn module in both stages to efficiently capture missing geometric structures.

Guo-Qing Zhang, Wen-Bo Zhao, Yuanchao Bai et al. · 0 citations
Conference Open access Sep 2026

PointGP: Geometry-Primed Attention for Point Cloud Analysis

PointGP is proposed, a geometry-primed framework that uses rectified local geometric topology as the primary cue for attention generation and achieves competitive accuracy with strong parameter and computational efficiency compared with representative strong baselines.

Yong Yang, Jian-Min Huang, Meng-Yuan Ge et al. · 0 citations
Preprint Sep 2026

ScaleBlind: Point Cloud Completion under Unknown Scale

Point cloud completion aims to infer a complete 3D shape from a partial point cloud and serves as a fundamental building block for downstream tasks such as reconstruction, editing, and simulation. Despite the recent progress, existing learning-based methods often implicitly rely on access to the ground-truth shape scal...

Sheng-Hui Wu, Chen Wang, Yuan Feng et al. · 0 citations
Preprint Sep 2026

Point Diffusion Mamba: Unified Diffusion-State-Space Modeling for Single-View 3D Reconstruction under Data Scarcity

While single-view 3D reconstruction has seen significant progress, extrapolating complex 3D structures from inherently ambiguous 2D observations remains fundamentally ill-posed, particularly in the critically underexplored data-scarce regime. To address this challenge, we propose Point Diffusion Mamba (PDM), a method t...

Wei Zhou, Xin-Zhe Shi, Xingxing Hao et al. · 0 citations
Open access Sep 2026

GAPrompt++: Multi-Granular Geometry-Aware Point Cloud Prompt for 3D Vision Model.

Pre-trained 3D vision models have substantially advanced point cloud analysis, yet adapting them to downstream tasks via full fine-tuning is computationally expensive and storage-intensive. Parameter-Efficient Fine-Tuning (PEFT) offers a promising alternative by reducing both adaptation cost and storage burden. However...

Zi-Xiang Ai, Zhen-Yu Cui, Yufei Guo et al. · 0 citations
Preprint Sep 2026

Partition-Invariant Tuning for 3D Scene Understanding

PointPiT is proposed, a partition-invariant tuning framework for scene-level point clouds that integrates local geometric patterns with global scene context to mitigate partition-induced representation shifts, and achieves consistent state-of-the-art performance among representative PEFT methods.

Hong-Qiang Lin, Tian-Le Wang, Shui-Wang Li et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.