Skip to content

Author

Ivan Laptev

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

#small language model Preprint Oct 2026

Omni-Embed-Mini: Binding Modalities Without Forgetting via Dense Distillation

This work presents Omni-Embed-Mini, a 0.9B-parameter model that maps text, speech, audio, images, video, and visually-rich documents into a single shared cosine space without updating any text-side parameter, and is competitive with the closed gemini-embedding-2, edging ahead of it on the overall-modality average.

Mohammed Irfan Kurpath, Jaseel Muhammad Kaithakkodan, Sahal Shaji Mullappilly et al. · 0 citations

Long Story Short: Story-level Video Understanding from 20K Short Films

This work proposes Short-Films 20K (SF20K), the largest publicly available movie dataset, and accompanies this dataset with SF20K-Test, a manual, open-ended question answering benchmark, showing that instruction tuning on the large-scale dataset substantially improves model performance.

Ridouane Ghermi, Xi Wang, Vicky Kalogeiton et al. · 11 citations · ⚡1

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.