Skip to content
Preprint

Optical Flow from Photons

Aug 2026 · 0 citations · 35 references
Computer Science

TL;DR

This work proposes QuantaFlow, the first method for dense optical flow directly from SPAD streams, which embeds SPAD representation construction into iterative flow refinement and constructs multi-scale representations containing intensity and structural cues.

Abstract

Optical flow remains challenging in high-speed and low-light scenes, where the limited frame rate and sensitivity of conventional cameras lead to motion blur and underexposure. Single-photon avalanche diode (SPAD) cameras offer single-photon sensitivity and extremely fine temporal sampling. However, individual slices in these high FPS binary photon streams are too sparse for dense correspondence. Temporal aggregation can provide the spatial cues required by optical flow, but accumulating photons at fixed coordinates blurs moving structures. Motion-aware aggregation can reduce this blur, yet it depends on the flow being estimated. To address this dependency, we propose QuantaFlow, the first method for dense optical flow directly from SPAD streams. Instead of constructing a fixed input representation, QuantaFlow embeds SPAD representation construction into iterative flow refinement. At each iteration, the current flow coarsely aligns the slices within the source and target sub-streams. A photon-flux transformation then constructs multi-scale representations containing intensity and structural cues, while adaptive multi-scale fusion balances photon noise and residual motion blur at each pixel. The fused representations drive a feature-warping flow update, and the refined flow guides representation construction in the next iteration. We further construct a synthetic dataset for SPAD optical-flow training and evaluation. Experiments on the synthetic dataset and real-world SPAD data demonstrate the effectiveness and generalization of QuantaFlow.

View source

Similar papers

2026

Single-Photon Neural Assumed-Density Filter for Dynamic Scene Reconstruction

Single-photon LiDAR (SP-LiDAR), with its extremely high sensitivity, is well suited for imaging in photon-sparse conditions and under strong background noise. However, most existing approaches rely on histogram accumulation; in dynamic scenes, histogram accumulation mixes time-of-arrival (ToA) events across motion, joi...

Siao Cai, Shun Lv, Zeyu Chen et al. · 0 citations
Preprint Aug 2026

Fisher-information limits of detector-bandwidth-efficient 3D light-field microscopy

Light-field microscopy enables snapshot volumetric imaging, but its information rate is constrained by both optical encoding and detector readout architecture. Here we develop a task-dependent Fisher-information framework that evaluates optical encoders relative to the detector resource limiting acquisition throughput....

Liang Gao · 0 citations
Preprint Aug 2026

High-Flux Count-Free Single-Photon 3D Cameras

Single-photon cameras based on single-photon avalanche diode (SPAD) technology are gaining popularity for 3D sensing, thanks to their extreme sensitivity and time resolution. There are two key challenges with single-photon cameras that limit their widespread use: (i) they suffer from non-linear distortions called''pile...

Kaustubh Sadekar, Vivek K. Goyal, David Maier et al. · 0 citations
Preprint Sep 2026

Racing in Volume with Flow Ensembles

Streaming 4D reconstruction has been demonstrated only indoors, on dense camera rigs surrounding subjects that move at human pace. Outdoor 4D reconstruction exists but relies either on cameras mounted on the moving vehicle itself, or on limited-coverage arrays observing quasi-static subjects offline. The case that actu...

Saswat Subhajyoti Mallick, Riu Cherdchusakulchai, Marc Ruiz Olle et al. · 0 citations
Jul 2026

Incorporating DINO Priors into Flow Matching for Low-Light Image Enhancement

Flow matching enables efficient low-light image enhancement (LLIE) with very few sampling steps, yet standard architectures lack explicit scene understanding, causing structural degradation and artifacts in challenging regions. We propose DINO-guided Flow Matching, which leverages a frozen DINOv3 backbone to provide il...

Xiang-Rui Zeng, Ling-Yu Zhu, Jing-Ming He et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.