Skip to main content
eScholarship
Open Access Publications from the University of California

Cognitive-Science--Inspired Evaluation of Large Language Models

Creative Commons 'BY' version 4.0 license
Abstract

Large language models (LLMs) can appear impressively capable across conversation, reasoning tasks, and standardized benchmarks. Yet merely appearing intelligent on traditional metrics is not the same as possessing robust, generalizable cognitive capacities. This disconnect highlights the need for more principled approaches to model evaluation. To address this gap, the symposium brings together four talks centered on developing a cognitive-science--inspired approach to assessing LLMs.