JigShape: Evaluating Visual-Geometric Reasoning in VLMs through Jigsaw Puzzles
A benchmark with tab-and-blank interlocking pieces where geometric constraints provide strong local compatibility requirements that, combined with visual content, yield unambiguous ground truth is introduced, establishing scalable geometric reasoning as an open challenge for vision-language models.