SemanticSlider3D is presented, a technique for continuous semantic attribute editing of 3D objects that requires no per-attribute training and supported decision-making in 3D prototyping and was perceived as a valuable addition to existing workflows.
Abstract
Fine-grained control over continuous semantic attributes of 3D objects is essential for 3D content creation, but is not well supported by conventional 3D modeling workflows or prompt-based interaction with existing generative AI tools. While slider-based methods have proven effective for fine-grained semantic control in 2D image generation, no equivalent approach exists for 3D. Extending these 2D methods to 3D is non-trivial due to challenges unique to 3D, including geometric integrity and cross-view coherence. We present SemanticSlider3D, a technique for continuous semantic attribute editing of 3D objects that requires no per-attribute training. Given a user-specified attribute, our pipeline constructs a semantic editing direction in the latent space of a state-of-the-art 3D generation model, presenting a diverse and coherent spectrum of 3D variations. A technical validation on a dataset of 50 3D object-attribute pairs shows our method was preferred by all five human assessors across variation range, consistency, 3D object quality, and attribute disentanglement, over a baseline combining a 2D slider with an image-to-3D model. An exploratory study with six participants demonstrates that SemanticSlider3D supported decision-making in 3D prototyping and was perceived as a valuable addition to existing workflows.
Existing 3D editing methods have made notable progress in controllability, yet they remain limited in several important ways. Most approaches rely on text-driven editing, which struggles to express fine-grained visual changes intended by the user. Moreover, many methods require manually supplied 3D masks or introduce u...
Xuancheng Jin, Ren-Gan Xie, Jiayuan Lu et al.· 0 citations
A unified multi-axis taxonomy is introduced that enables us to classify the available methods in 3D Gaussian splatting methods in terms of five complementary categories: semantic vocabulary space, representation form, functional role, knowledge source, and query mechanism.
This work proposes PointPDF V2, a unified framework that integrates open-set recognition (OSS) and incremental learning (IL) into a cohesive pipeline and introduces a more challenging continual OWSS protocol in 3D, where models must simultaneously preserve the known-class performance, acquire new knowledge, and still i...
Jinfeng Xu, Xianzhi Li, Yixue Hao et al.· IEEE Transactions on Pattern...· 0 citations
Existing methodologies in open vocabulary 3D semantic segmentation primarily concentrate on establishing a unified feature space encompassing 3D, 2D, and textual modalities. Nevertheless, traditional techniques such as global feature alignment or vision-language model distillation tend to impose only approximate corres...
Ziyi Wang, Yan-Bo Wang, Xu-Min Yu et al.· IEEE Transactions on Pattern...· 0 citations
3D Gaussian Splatting (3DGS) has emerged as a powerful method that offers significant improvements in both rendering quality and efficiency. However, the task of achieving open‐ended scene style transfer remains a considerable challenge, as existing methods typically focus on the entire scene and fail to address the...
F. Wang, J. Wang, Y. Tang et al.· Computer graphics forum (Pri...· 0 citations
This work presents GraphWrit3R, a simple end-to-end method that takes a 3D point cloud, Gaussian Splats, or a combination of both as input, and directly outputs a complete scene graph as a structured JSON script, achieving state-of-the-art performance on object class, predicate, and triplet recall on the 3DSSG benchmar...
Luka Milivojevic, Nikola Popovic, Sayan Deb Sarkar et al.· 0 citations
Computer scientist, entrepreneur, and philanthropist will collaborate with the MIT Schwarzman College of Computing to advance AI and scientific discovery.
Computer-use AI agents struggle with multi-step workflows like email and customer support. Echoverse trains agents in realistic environments rather than simply providing more training tasks, helping them improve as the tasks, tests, and environments evolve. The post Echoverse: Deep, evolving environments for computer-use agents appeared first on Microsoft Research.
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.