Large language models are increasingly applied to high-risk domains such as law, yet complex legal reasoning remains limited by two structural challenges. First, existing RAG and GraphRAG methods emphasize lexical or semantic similarity while overlooking normative relations among legal provisions. Second, vanilla Chain...
Qing-Jing Chen, Jun-Kai Zhang, Shao-Chun Wang et al.· 0 citations
Deep research requires models to retrieve, connect, and synthesize evidence from large-scale heterogeneous sources to answer complex queries and produce analytical reports. Existing benchmarks mainly evaluate final outcomes, such as answer correctness, report quality, or citation alignment, while providing limited visi...
Yubo Sun, Chunyi Peng, Yukun Yan et al.· arXiv.org· 0 citations
MemoryCard is a video-memory-based augmentation framework that organizes long videos into self-contained Memory Cards, each corresponding to a distinct topic or event, and consistently improves long-video QA performance under comparable visual-token budgets.
Qing Yang, Pengcheng Huang, Xinze Li et al.· arXiv.org· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.