This work proposes QuantaFlow, the first method for dense optical flow directly from SPAD streams, which embeds SPAD representation construction into iterative flow refinement and constructs multi-scale representations containing intensity and structural cues.
Abstract
Optical flow remains challenging in high-speed and low-light scenes, where the limited frame rate and sensitivity of conventional cameras lead to motion blur and underexposure. Single-photon avalanche diode (SPAD) cameras offer single-photon sensitivity and extremely fine temporal sampling. However, individual slices in these high FPS binary photon streams are too sparse for dense correspondence. Temporal aggregation can provide the spatial cues required by optical flow, but accumulating photons at fixed coordinates blurs moving structures. Motion-aware aggregation can reduce this blur, yet it depends on the flow being estimated. To address this dependency, we propose QuantaFlow, the first method for dense optical flow directly from SPAD streams. Instead of constructing a fixed input representation, QuantaFlow embeds SPAD representation construction into iterative flow refinement. At each iteration, the current flow coarsely aligns the slices within the source and target sub-streams. A photon-flux transformation then constructs multi-scale representations containing intensity and structural cues, while adaptive multi-scale fusion balances photon noise and residual motion blur at each pixel. The fused representations drive a feature-warping flow update, and the refined flow guides representation construction in the next iteration. We further construct a synthetic dataset for SPAD optical-flow training and evaluation. Experiments on the synthetic dataset and real-world SPAD data demonstrate the effectiveness and generalization of QuantaFlow.
Single-photon LiDAR (SP-LiDAR), with its extremely high sensitivity, is well suited for imaging in photon-sparse conditions and under strong background noise. However, most existing approaches rely on histogram accumulation; in dynamic scenes, histogram accumulation mixes time-of-arrival (ToA) events across motion, joi...
Siao Cai, Shun Lv, Zeyu Chen et al.· IEEE Transactions on Computa...· 0 citations
Light-field microscopy enables snapshot volumetric imaging, but its information rate is constrained by both optical encoding and detector readout architecture. Here we develop a task-dependent Fisher-information framework that evaluates optical encoders relative to the detector resource limiting acquisition throughput....
Single-photon cameras based on single-photon avalanche diode (SPAD) technology are gaining popularity for 3D sensing, thanks to their extreme sensitivity and time resolution. There are two key challenges with single-photon cameras that limit their widespread use: (i) they suffer from non-linear distortions called''pile...
Kaustubh Sadekar, Vivek K. Goyal, David Maier et al.· 0 citations
Streaming 4D reconstruction has been demonstrated only indoors, on dense camera rigs surrounding subjects that move at human pace. Outdoor 4D reconstruction exists but relies either on cameras mounted on the moving vehicle itself, or on limited-coverage arrays observing quasi-static subjects offline. The case that actu...
Saswat Subhajyoti Mallick, Riu Cherdchusakulchai, Marc Ruiz Olle et al.· 0 citations
Flow matching enables efficient low-light image enhancement (LLIE) with very few sampling steps, yet standard architectures lack explicit scene understanding, causing structural degradation and artifacts in challenging regions. We propose DINO-guided Flow Matching, which leverages a frozen DINOv3 backbone to provide il...
Xiang-Rui Zeng, Ling-Yu Zhu, Jing-Ming He et al.· 2026 International Symposium...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.