MoMHa: Multi-Objective Optimization of LLM Harnesses over Accuracy, Safety, and Tokens
Most work on improving large language models treats accuracy as the sole objective. We argue that the harness, the Python code surrounding the model that constructs prompts, routes calls, and parses outputs, is a first-class design surface whose quality is inherently multi-objective: an accurate harness that refuses no...