Language models often process long inputs sequentially in chunks, but continuing to read after sufficient evidence has been acquired wastes computation. Existing stopping mechanisms either learn sufficiency from internal activations or train an exit gate, while a simpler alternative asks the model whether it has read e...
Muath Alyobi, M. Eltahir, Almoayyad Abuljdail et al.· 0 citations
MathNet-Retrieve asks a retriever to find, for a math problem, a document stating the same problem. An LLM under one fixed prompt writes each gold document and its near-miss distractors; LLM judges filter them. We call this procedure the"recipe", training on pairs built the same way"recipe-matching", and ask how much s...
A. Habibullah, M. Alshiekh, Yazan Alshoibi et al.· 0 citations
Late-interaction retrieval is the state-of-the-art for visual document search, but it pays for its accuracy in storage. Existing compression methods retain a subset or local average of the N~1,000 vectors per page. Under aggressive storage budgets, however, these methods degrade sharply, and alternatives require retrai...
M. Eltahir, Talal Aloushan, Rose Khairoalsendi et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.