Researchers have developed CogArena, a tool for evaluating cognitive abilities in large language models (LLMs) across multiple methods, aiming to assess whether LLM scores reflect true cognitive structures that consistently emerge and generalize. This matters because it could improve the understanding of LLM capabilities and their alignment with human cognitive functions.