Skip to content

URA-NER: A Unified Retrieval-Augmented Framework with Retrieval Alignment and Uncertainty Reduction for Low-Resource NER

Sep 2026 · 0 citations
Computer Science

TL;DR

A novel unified retrieval-augmented framework, URA-NER, including three key components: Progressive Granularity Retrieval, Model-aware Representation Enhancement, and Reason-aware Knowledge Verification is proposed, including three key components: Progressive Granularity Retrieval, Model-aware Representation Enhancement, and Reason-aware Knowledge Verification.

Abstract

In-context learning (ICL) based on large language models (LLMs) has shown promising potential in alleviating performance bottlenecks caused by the limited availability of annotated data in Named Entity Recognition (NER). However, existing methods still face issues of retrieval misalignment and generation uncertainty, making their performance heavily dependent on the LLM's capabilities. As the parameter scale of LLMs decreases, their performance in few-shot settings deteriorates significantly. In this paper, we propose a novel unified retrieval-augmented framework, URA-NER, including three key components: Progressive Granularity Retrieval (PGR), Model-aware Representation Enhancement (MaRE), and Reason-aware Knowledge Verification. PGR is a two-stage retrieval mechanism that achieves stage alignment. It first retrieves demonstrations for span detection based on the query's global semantics, and then for type classification based on the specific entity context, providing fine-grained local information. Moreover, MaRE employs entity pre-recognition to guide the construction of representations, ensuring the query and demonstrations are aligned within the LLM's semantic space and attention pattern. In addition, to mitigate generation uncertainty, we propose RaKV, a closed-loop"generation-retrieval-verification"process. It explicates the LLM's reasoning paths, leverages them for the retrieval of external knowledge, and reorganizes the knowledge into verification evidence aligned with the original reasoning paths. We conduct extensive experiments on multiple low-resource NER datasets. Results demonstrate that URA-NER significantly enhances the performance of LLMs under low-resource settings, with particularly pronounced gains for smaller LLMs, achieving new state-of-the-art results on several benchmarks.

View source

Similar papers

Aug 2026

STaR: a soft-labeling and triplet-aware retriever for efficient retrieval-augmented QA

This study proposes STaR, a novel retriever fine-tuning framework that integrates BM25 similarity graph-based soft labeling with a triplet similarity learning strategy based on Sentence-BERT (SBERT), and introduces a triplet-aware SBERT training architecture that explicitly models relative semantic distances between qu...

Jiali Jiang, Chih-Yung Chang, Youxi Li et al. · 0 citations
Open access 2026

A New Metadata-Aware Retrieval-Augmented Generation (RAG) Architecture for Trustworthy Legal Question Answering

Large Language Models (LLMs) offer strong capabilities for Natural Language Processing, yet their inherent uncertainty often produces hallucinations, confident but incorrect statements, which is critical in domains requiring precise knowledge representation. Retrieval-Augmented Generation (RAG) reduces this risk throug...

Alexandra V. Jove-Ticona, Luis J. Duarte-Coaquera, Israel N. Chaparro-Cruz et al. · 0 citations
Conference Aug 2026

Resource-Efficient Semantic Retrieval Optimization for Retrieval-Augmented Generation Using FAISS, Milvus, HNSW, and Product Quantization

Retrieval-Augmented Generation (RAG) enhances large language models (LLMs) by integrating external knowledge, but its performance depends heavily on efficient semantic vector search. This paper presents a deployment-oriented empirical study that systematically compares established backends and ANN configurations under...

Achmad Rizal, Jui-Tang Wang, Gerald Wijaya Tanubrata et al. · 0 citations
#artificial intelligence Preprint Aug 2026

Select, Don't Train: The Benefits of Modular Entity Disambiguation with LLM-Based Selection

A systematic comparison of retrieval strategies for candidate generation under a shared LLM-based selection stage, combining sparse retrieval (BM25), Web KB search, and a state-of-the-art trained dense retriever with several open- and closed-source LLMs is presented.

Fina Polat, Daniel Daza, Pengyu Zhang et al. · 0 citations
#artificial intelligence Preprint Sep 2026

STAIR (STructure Aware Information Retriever): A novel dataset and LLM based retriever for document structure augmentation

Retrieval Augmented Generation (RAG) is a key component for generating accurate and hallucination free answers using Large Language Models (LLMs). LLMs are improving at handling long context, but still suffer from"lost in the middle"problem. Thus, precise and accurate retrieval is important. Current retrievers chunk lo...

Vineet Kumar, Meghanadh Pulivarthi, Vishwajeet Kumar et al. · 1 citation

Related blog posts

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.