BiasScope: Inference-Time Bias Detection and Mitigation for Fair Natural Language Processing (NLP) Using Retrieval-Augmented Generation and Prompt Engineering
The appearance of social stereotypes in Large Language Models (LLMs) is a crucial concern in research on fairness in the context of Natural Language Processing (NLP). This study focuses on gender bias, occupational stereotypes, gender comparisons, and trait attributions. Numerous mitigation techniques depend on fine-tu...