Skip to content

Author

Jiajun Cheng

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Jul 2026

Latent-Action-Guided Video-Language Feature Learning for Surgical Instrument-Tissue Interaction Recognition

This work introduces \the authors', which compresses frame-to-frame changes into latent actions and predicts next-frame features during end-to-end video--language alignment and achieves competitive recognition with faster inference and smaller INT4 accuracy drops than V-JEPA2/2.1.

Jiajun Cheng, Sainan Liu, Subarna Tripathi et al. · 0 citations
Preprint Jul 2026

LAVIFT: Latent-Action-Guided Vision Fine-Tuning for Surgical Interaction Recognition

Understanding instrument-tissue interactions is essential for context-aware surgical AI and autonomous robotic surgery. Pretrained vision-language models (VLMs) and vision encoders offer an alternative to conventional interaction classifiers by transferring broad visual and semantic knowledge. However, adapting them to...

Jiajun Cheng, Subarna Tripathi, Sainan Liu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.