Low-resource hallucination detection in LLMs on multi-task datasets via iterative pseudo-labeling using confidence thresholding and active learning
Large Language Models have advanced natural language generation, but they often produce outputs that are grammatically correct yet factually incorrect or misleading. This issue, commonly known as hallucination, reduces the reliability of such systems, especially in domains such as law, medicine, journalism, and educati...