Skip to content
arXiv cs.CL · Papers

CogArena: A Multimethod Evaluation of Cognitive Ability Structure in Large Language Models

arXiv:2607.24999v1 Announce Type: new Abstract: LLM cognitive scores are increasingly summarized as per-ability profiles whose dimensions should converge across tasks, respond selectively to matched interventions, and generalize beyond the models used to define them. We introduce CogArena, a procedurally generated 13-p