Skip to content

Author

Sowmya Vajjala

We have 2 of 72 papers

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Last Translation Benchmark

For scientific progress, we need benchmarks that test the limits of state-of-the-art models, and evaluation methods that inform us about failure cases. As models get stronger, standard benchmarks for machine translation are approaching saturation. Further, automatic translation metrics are unreliable, opaque, and vulne...

Vilém Zouhar, Niyati Bafna, Mukund Choudhary et al. · 0 citations
Jul 2026

LLM Judges Can Be Too Generous When There Is No Reference Answer

The results emphasize the need for calibrating the LLM judges with a sample with reference-aware evaluation before using them in reference-free setups reliably, and the methodology provides a blueprint for researchers and practitioners in doing such calibration of LLM judges for other tasks.

Chalamalasetti Kranti, Sowmya Vajjala · 1 citation

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.