Skip to content

Author

Caiqi Zhang

2 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

When Calibration Rankings Reverse: Accuracy-Controlled Evaluation for Fair Comparison of LLMs

It is shown that raw global calibration metrics are not robust for cross-model comparison, and that fair calibration comparison requires accuracy-aware evaluation, and proposed ACE, an accuracy-controlled evaluation framework, is proposed.

Zhichao Yang, Caiqi Zhang, Ruihan Yang et al. · 1 citation

Beyond the Final Layer: Intermediate Representations for Better Multilingual Calibration in Large Language Models

This work conducts the first large-scale, systematic studies of multilingual calibration across six model families and over 100 languages, revealing that non-English languages suffer from systematically worse calibration.

Ej Zhou, Caiqi Zhang, Tiancheng Hu et al. · 10 citations · ⚡1

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.