Skip to content

SemanticSlider3D: Training-Free Continuous Semantic Editing for 3D Objects

Aug 2026 · 0 citations · 79 references
Computer Science

TL;DR

SemanticSlider3D is presented, a technique for continuous semantic attribute editing of 3D objects that requires no per-attribute training and supported decision-making in 3D prototyping and was perceived as a valuable addition to existing workflows.

Abstract

Fine-grained control over continuous semantic attributes of 3D objects is essential for 3D content creation, but is not well supported by conventional 3D modeling workflows or prompt-based interaction with existing generative AI tools. While slider-based methods have proven effective for fine-grained semantic control in 2D image generation, no equivalent approach exists for 3D. Extending these 2D methods to 3D is non-trivial due to challenges unique to 3D, including geometric integrity and cross-view coherence. We present SemanticSlider3D, a technique for continuous semantic attribute editing of 3D objects that requires no per-attribute training. Given a user-specified attribute, our pipeline constructs a semantic editing direction in the latent space of a state-of-the-art 3D generation model, presenting a diverse and coherent spectrum of 3D variations. A technical validation on a dataset of 50 3D object-attribute pairs shows our method was preferred by all five human assessors across variation range, consistency, 3D object quality, and attribute disentanglement, over a baseline combining a 2D slider with an image-to-3D model. An exploratory study with six participants demonstrates that SemanticSlider3D supported decision-making in 3D prototyping and was perceived as a valuable addition to existing workflows.

View source

Similar papers

Preprint Aug 2026

ES3D: Embedding Semantics into 3D Space for Component-Aware Editing

Existing 3D editing methods have made notable progress in controllability, yet they remain limited in several important ways. Most approaches rely on text-driven editing, which struggles to express fine-grained visual changes intended by the user. Moreover, many methods require manually supplied 3D masks or introduce u...

Xuancheng Jin, Ren-Gan Xie, Jiayuan Lu et al. · 0 citations
Review Open access Aug 2026

Semantic 3D Gaussian Splatting: A State-of-the-Art Review

A unified multi-axis taxonomy is introduced that enables us to classify the available methods in 3D Gaussian splatting methods in terms of five complementary categories: semantic vocabulary space, representation form, functional role, knowledge source, and query mechanism.

J. Flotyński · 0 citations
Aug 2026

PointPDF V2: A Unified Framework for Continual Open-World 3D Semantic Segmentation.

This work proposes PointPDF V2, a unified framework that integrates open-set recognition (OSS) and incremental learning (IL) into a cohesive pipeline and introduces a more challenging continual OWSS protocol in 3D, where models must simultaneously preserve the known-class performance, acquire new knowledge, and still i...

Jinfeng Xu, Xianzhi Li, Yixue Hao et al. · 0 citations
Sep 2026

XMask3D++: Cross-Modal Mask Reasoning for Open Vocabulary 3D Segmentation.

Existing methodologies in open vocabulary 3D semantic segmentation primarily concentrate on establishing a unified feature space encompassing 3D, 2D, and textual modalities. Nevertheless, traditional techniques such as global feature alignment or vision-language model distillation tend to impose only approximate corres...

Ziyi Wang, Yan-Bo Wang, Xu-Min Yu et al. · 0 citations
Sep 2026

LDS3D: Language‐Driven Object Stylization based on 3DGS

3D Gaussian Splatting (3DGS) has emerged as a powerful method that offers significant improvements in both rendering quality and efficiency. However, the task of achieving open‐ended scene style transfer remains a considerable challenge, as existing methods typically focus on the entire scene and fail to address the...

F. Wang, J. Wang, Y. Tang et al. · 0 citations
Preprint Sep 2026

GraphWrit3R: End-to-End 3D Scene Graph Writing

This work presents GraphWrit3R, a simple end-to-end method that takes a 3D point cloud, Gaussian Splats, or a combination of both as input, and directly outputs a complete scene graph as a structured JSON script, achieving state-of-the-art performance on object class, predicate, and triplet recall on the 3DSSG benchmar...

Luka Milivojevic, Nikola Popovic, Sayan Deb Sarkar et al. · 0 citations

Related blog posts

Microsoft Research Blog Jul 30, 2026

Echoverse: Deep, evolving environments for computer-use agents

Computer-use AI agents struggle with multi-step workflows like email and customer support. Echoverse trains agents in realistic environments rather than simply providing more training tasks, helping them improve as the tasks, tests, and environments evolve. The post Echoverse: Deep, evolving environments for computer-use agents appeared first on Microsoft Research.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.