Multi-Hop Knowledge Composition is Bound by Pretraining Exposure
Large Language Models fail at implicit multi-hop reasoning: a model answers"When was $X$ born?"and"Who is $Y$'s closest friend?"correctly but fails on"When was $Y$'s closest friend born?"in a single forward pass, even when both facts are perfectly memorized and individually retrievable. We study this failure in a contr...