Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring

Open in new window