Skip to content
Preprint

LumiTokens: 3D Relighting via Token-Space Lighting Transformation

Aug 2026 · 0 citations · 41 references
Computer Science

TL;DR

LumiTokens is a framework that formulates 3D relighting as a direct transformation on latent scene tokens, without explicit 3D representations, rendering equations, or physics-based decomposition, and achieves comparable or superior relighting quality to other methods and supports progressive, composable lighting edits.

Abstract

Existing 3D relighting methods operate through either explicit material decomposition, diffusion-based view-space generation, or a combination of both, requiring full recomputation for each new lighting condition. We observe that recent latent scene representations, which encode multi-view images into a set of compact tokens with no fixed physical semantics, open up a novel design space for relighting. We present LumiTokens, a framework that formulates 3D relighting as a direct transformation on latent scene tokens, without explicit 3D representations, rendering equations, or physics-based decomposition. Our model introduces a Scene Token Editor that processes scene tokens jointly with light-ray tokens through self-attention, producing updated tokens that can be decoded into multi-view-consistent relit images. To support diverse lighting types through a unified interface, all lighting signals, including environment maps, point lights, and area lights, are parameterized as Plucker ray tokens, enabling native 3D user interaction with a representation that carries no explicit spatial structure. Crucially, this design supports progressive relighting: because the editor's output remains in the same latent space as its input, a user can incrementally build up illumination one light source at a time, with each edit composing in token space. Experiments demonstrate that LumiTokens achieves comparable or superior relighting quality to other methods and supports progressive, composable lighting edits. Project page: https://neu-vi.github.io/LumiTokens/

View source

Similar papers

#diffusion models Preprint Sep 2026

SceneHI: High-Resolution 3D-Consistent Scene Texturing with Controllable Illumination

This work introduces an exact analytical pixel-to-texel mapping that aligns diffusion trajectories across multiple viewpoints, and utilizes High-Resolution Latent Textures as a persistent canvas for gradually denoised textures, while camera views perform the denoising steps in latent pixel space.

Athanasios Tragakis, M. Aversa, D. Ivanova et al. · 0 citations
Preprint Sep 2026

LightBridge: Feed-Forward Generative Relighting for 3D Gaussian Splatting

LightBridge is presented, a feed-forward generative framework for controllable relighting of complete 3DGS assets in a single pass and efficient single-pass prediction of complete relit 3DGS assets without scene-specific optimization.

Heng Cao, Pan-Hao Cheng, Huang-Sheng Du et al. · 0 citations
Preprint Aug 2026

LightFuse: Relightable Interactive Gaussian Scene Reconstruction via Multi-Scan Fusion and 2D Gaussian Ray Tracing

Relightable interactive scene reconstruction aims to build an editable 3D model from scans of different object arrangements and render new layouts under novel illumination. Existing methods either bake lighting into appearance or recover material and illumination only for fixed scenes, leaving edited layouts with incon...

Haonan Zhou, Gaoxiang Linghu, You-Lin Jia et al. · 0 citations
Preprint Aug 2026

Luce: Relightable Gaussians for 3D Asset Generation

High-fidelity image-to-3D generation requires a 3D representation that captures both geometry and appearance. However, preserving fine detail across the physically based rendering (PBR) modalities needed for relighting remains challenging. To address this, we propose Luce, a 3D representation that unifies geometry and...

M. Singh, Michele Stoppa, Alvise Memo et al. · 0 citations
#artificial intelligence Preprint Sep 2026

RelightFormer: Feed-forward Generative Transformer for Multiview Object Relighting

Image relighting is traditionally tackled via complex inverse rendering pipelines, which suffer from ill-posed optimization, or single-image generative models that ignore crucial multi-view cues necessary for understanding 3D geometry and material interactions. To address these limitations, we introduce a feed-forward...

He-Jun Wang, Jin-Xi Li, Jun-Wei Jiang et al. · 0 citations
Preprint Sep 2026

OmniFabric: Coherent UV Space Texture Synthesis for 3D Garment Reconstruction

Automated generation of production-ready 3D garment assets from a single image is a central challenge in digital content creation. While recent generative models have significantly advanced 3D geometry reconstruction, synthesizing high-quality textures remains a bottleneck. Existing methods often bake environmental ill...

Di Huang, Yuan-Hao Wang, Cheng Zhang et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.