When Is Complex Chunking Worth It? A Multi-Objective Evaluation of Chunking Methods at Scale
This work evaluates eight representative chunking strategies across two scalable corpora, three embedding models, and multiple corpus sizes, measuring both retrieval effectiveness and system-level costs and shows that computationally expensive methods rarely provide consistent gains over simpler chunking.