Skip to content

Similar papers

Preprint Aug 2026

Activation Outliers Matter: Robust Recovery for Quantized Multimodal LLMs

This work proposes Residual Fallback Quantization (RFQ), a lightweight activation reconstruction framework that supplements the primary ulta-low-bit activation representation with an auxiliary quantized residual pathway that improves activation fidelity while preserving the efficiency advantages of ultra-low-bit comput...

Tanzila Rahman, Mehran Taghian Jazi, Yunke Peng et al. · 0 citations
Jul 2026

C-PTQ: Fisher-weighted Channel-wise Sensitivity for Post-training Quantization of MLLMs

C-PTQ is proposed, a unified channel-wise PTQ method that harmonizes task-specific loss perturbation and quantization error and achieves state-of-the-art performance without auxiliary modules like LoRA, thereby maintaining high efficiency.

Jia-Meng Li, Han Zhou, M. Blaschko · 0 citations
Preprint Sep 2026

SCULPT: Training Edge Vision Models for Post-Training Quantization Readiness

Edge vision models are difficult to deploy on resource-constrained hardware, making low-bit post-training quantization (PTQ) attractive. In practice, standard FP32 training often produces heavy-tailed activation distributions whose outliers destabilize activation quantization: preserving the full range wastes quantizat...

Bharadwaj Kavuri, Sourav Babu-PK, Varadhraj Ellapan et al. · 0 citations
Jul 2026

SCALPEL: Semantic Cross-modal Alignment via LLM-Powered Encoder Learning for Medical Vision-Language Representation

Vision-language pre-training (VLP) serves as a cornerstone for medical multimodal representation learning. However, existing medical VLP frameworks are often constrained by the limited context windows and shallow representational capacities of lightweight text encoders when processing lengthy, terminology-dense clinica...

Yu Fu, En-Yu Bao, Xiangyu Shen et al. · 0 citations
Preprint Aug 2026

MVC-Bench: Benchmarking Calibration of Medical Vision-Language Models

Reliable evaluation of vision-language models (VLMs) and medical vision-language models (Medical-VLMs) requires calibrated confidence, particularly under realistic clinical conditions. However, existing efforts mainly focused on improving accuracy, leaving calibration in the medical domain underexplored. To this end, w...

Ashshak Sharifdeen, Shihab Aaqil Ahamed, Ufaq Khan et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.