Review
Jul 2026
Pre-Flight: A Benchmark for Evaluating Large Language Models on Aviation Operational Knowledge
It is argued that domain specific evaluation of this kind is a necessary precondition for responsible deployment of generative AI in non safety critical aviation operations.
Alex Brooker, T. Hughes
· 0 citations