Skip to main content
Download PDF
- Main
Cognitive-Science--Inspired Evaluation of Large Language Models
- Fu, Shuhao;
- Beger, Claas;
- Yi, Ryan Cho;
- Mitchell, Melanie;
- Frank, Michael C.;
- Millière, Raphaël;
- Hawkins, Robert;
- Ullman, Tomer D.
© 2026 by the author(s). Learn more.
Abstract
Large language models (LLMs) can appear impressively capable across conversation, reasoning tasks, and standardized benchmarks. Yet merely appearing intelligent on traditional metrics is not the same as possessing robust, generalizable cognitive capacities. This disconnect highlights the need for more principled approaches to model evaluation. To address this gap, the symposium brings together four talks centered on developing a cognitive-science--inspired approach to assessing LLMs.