Centaur

A foundation model to predict and capture human cognition

Establishing a unified theory of cognition has been a major goal of psychology. While there have been previous attempts to instantiate such theories by building computational models, we currently do not have one model that captures the human mind in its entirety. A first step in this direction is to create a model that can predict human behavior in a wide range of settings. Here we introduce Centaur, a computational model that can predict and simulate human behavior in any experiment expressible in natural language. We derived Centaur by finetuning a state-of-the-art language model on a novel, large-scale data set called Psych-101. Psych-101 reaches an unprecedented scale, covering trial-by-trial data from over 60,000 participants performing over 10,000,000 choices in 160 experiments. Centaur not only captures the behavior of held-out participants better than existing cognitive models, but also generalizes to new cover stories, structural task modifications, and entirely new domains. Furthermore, we find that the model's internal representations become more aligned with human neural activity after finetuning. Taken together, our results demonstrate that it is possible to discover computational models that capture human behavior across a wide range of domains. We believe that such models provide tremendous potential for guiding the development of cognitive theories and present a case study to demonstrate this.

Contact: Marcel Binz


Model	Paper	Code

How to use Centaur?

Have access to at least one 80 GB GPU (e.g. A100)? You can run the model locally using unsloth.
Have neither GPUs nor money? You can explore a smaller version on Google Colab's free GPUs or the larger model via our Hugging Face Space.

How to prompt Centaur?

You can find prompt examples here or in the Appendix of our preprint.
We did not employ a particular prompt template – just phrase everything in natural language.
Human choices are encapsulated by "<<" and ">>" tokens.
Most experiments in the training data are framed in terms of button presses. If possible, it is recommended to use that style.