Skip to content

A Semantic Parsing Method for Indoor Scene Images Based on Prior Knowledge of Building Structure.

Aug 2026 · Journal of Visualized Experiments · Vol 234 · 0 citations
Medicine

TL;DR

This paper proposes a semantic parsing method that leverages building-structure priors that uses a shifted-window hierarchical transformer encoder to extract multi-scale visual features and combines a gradient-direction-consistency line segment detection algorithm to construct a Manhattan 3D bounding box.

View source

Similar papers

Open access Sep 2026

Vision–language guided semantic-geometric transformer for memory-efficient 3D scene understanding

This work proposes a segmentation framework guided by Contrastive Language–Image Pre-training (CLIP) that enriches sparse 3D tokens with vision–language semantic priors and introduces a decoupled CLIP-induced semantic residual that forms semantic-geometric attention biases for local window attention.

Li-Cheng Liu, Yu Li, Fu-Yong Liu · 0 citations
Preprint Aug 2026

GaussianDS: Depth-supervised Semantic Gaussian Splatting for Scene Understanding

GaussianDS, a depth-supervised semantic 3DGS framework that treats semantic lifting as a supervision-alignment problem and jointly optimizes RGB appearance, rendered depth, and compact semantics from scratch, is proposed.

Yu-Fei Zhang, Chen-Lu Zhan, Hong-Wei Wang · 0 citations
#artificial intelligence Preprint Oct 2026

RelationVGGT: Visual Geometry Transformers for 3D Spatial Relation Segmentation

Recent advances in 3D reconstruction have progressed from per-scene optimization to feed-forward inference, and semantic scene understanding has followed suit -- yet existing methods remain confined to object-centric perception, neglecting spatial relations between objects. We formulate 3D spatial relation segmentation...

Minsu Kim, Jaesung Choe, J. Lee et al. · 0 citations
Open access Sep 2026

Semantic-Guided Adaptive Gaussian Segmentation

3D Gaussian Splatting (3DGS) enables real-time photorealistic scene reconstruction, yet its segmentation tasks suffer from two critical flaws: poor 3D consistency (e.g., blurred instance boundaries and unstable cross-view semantic association) and insufficient structural awareness near ambiguous object boundaries. To a...

Ying-Han Zhou, Fan Zhou · 0 citations
Open access 2026

AnyRoad: A Frequency-Aware Adapter Framework for Road Segmentation With Segment Anything Model

Automated road extraction from high-resolution satellite imagery is critical for geospatial applications. However, accurate segmentation requires balancing global topological continuity with local boundary precision. Existing methods often struggle with this tradeoff, while directly adapting large-scale foundation mode...

Li-Lun Deng, Jiang-She Zhang, Hao-Wen Bai et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.