Flow matching is central to 3D generation, yet in practice its reinforcement learning (RL) methods are largely adapted from 2D visual generation. Representative DPO-, GRPO-, and NFT-style objectives, when applied to negative trajectories, mainly steer predicted velocities away from the corresponding directions without...
Zhen-Wu-Yong Zhou, Zhi-Wei Ning, Pu-Hua Jiang et al.· 0 citations
WorldClaw is presented, a fully agentic, coarse-to-fine framework for open-world 3D scene generation that produces large-scale scenes with coherent spatial organization, visually compelling local content, and editable instance-level assets while preserving a consistent global terrain structure.
Chunchao Guo, Jinpeng Li, Yang Li et al.· 5 citations
This work presents SceneActBench, a benchmark for visually conditioned action across five 3D tasks under a unified agent-environment loop, and evaluates each final output against hidden ground truth with task-specific geometric metrics.
Yifei Zhao, Xiangxin Zhou, Wenhao Yang et al.· arXiv.org· 2 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.