Skip to content

Author

Yi Fang

We have 4 of 40 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Open access Jun 2025

Hierarchical Scoring With 3D Gaussian Splatting for Instance Image-Goal Navigation

Instance Image-Goal Navigation (IIN) asks an agent to locate the specific object instance shown in a goal image. Existing 3D Gaussian Splatting (3DGS) based methods rely on pose-centric search—sampling many viewpoints, rendering them, and comparing against the goal—which is inefficient in continuous 6-DoF space. We ins...

Yijie Deng, Shuaihang Yuan, Geeta Chandra Raju Bethala et al. · 4 citations · ⚡1
Preprint Sep 2026

CST-WM: A Causally Structured World Model for Embodied Visual Tracking

Embodied visual tracking requires a robot to choose actions that keep a moving target observable at a suitable distance, and to recover it after occlusion, out-of-view drift, or distractor crossings. We cast the task as planning over future target evidence with an action-conditioned world model. In logged tracking data...

Jun-Yi Hu, Shuaihang Yuan, Jia-Zhao Liang et al. · 0 citations
#computer vision Preprint Sep 2026

SignDino: Self-Supervised Sign Language Representation Learning via Temporal-Axis Self-Distillation

SignDino, a self-supervised sign-video encoder that moves the DINOv3 student--teacher recipe from the spatial domain of image crops to the temporal domain of tracked sign streams, provides a strong public self-supervised representation and shows competitive or state-of-the-art performance under matched downstream evalu...

Jun-Yi Hu, Zhe-Wen He, Hao Huang et al. · 0 citations
Jul 2026

VTaMo: Video-Text Alignment Model for Sign Language Translation

VTaMo is presented, a framework that introduces explicit multi-granularity alignment at three levels: local alignment via entropy-regularized optimal transport with a learnable null token for fine-grained frame-to-token correspondences; global alignment via a learnable orthogonal transformation that calibrates embeddin...

Junyi Hu, Zhewen He, Hao Huang et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.