Probing cross-lingual differences in how LLMs represent Natural Language Inference
Cross-lingual probe transfer and direct comparison of probe weights show that NLI representations are most alignable in the middle layers: probes transfer best there, and probes trained independently on different languages converge to similar weight vectors, peaking mid-network.
Nicolas Ramos Fernandez
· 0 citations