CogSci 2025

•

July 31, 2025

•

San Francisco, United States

keywords:

computational modeling

bayesian modeling

decision making

learning

psychology

Humans learn by interacting directly with their environments and by communicating via language. In this project, we explore this interaction between language and experiential learning through a novel sequential decision-making task, the "instructed bandit task" (IBT). In the IBT, agents make choices and receive rewards sampled from an unknown Gaussian distributions after being given linguistic hints. The IBT assesses how linguistic input and experienced reward values combine to determine choice behavior. We additionally propose a novel Bayesian reinforcement learning model that combines Bayesian updating from experience with propositional constraints that capture the meaning of the linguistic hints. As a point of comparison, we evaluate both human participants and Centaur, a LLaMA-based model fine-tuned to mimic human behavior, on the IBT. Our results show that all agents converge with the Bayesian model, and the granular difference in choice sequences reveal the varied role instruction plays in decision-making tasks.

Downloads

Paper

Next from CogSci 2025

Leveraging Machine Learning and Wearable Cameras to Analyze Children’s Social Interactions
poster

Leveraging Machine Learning and Wearable Cameras to Analyze Children’s Social Interactions

CogSci 2025

Manuel Bohn
+2
Nele-Pauline Suffo and 4 other authors

31 July 2025

Similar lecture

Post-hoc loss-calibration for Bayesian neural networks
lightning talk

Post-hoc loss-calibration for Bayesian neural networks

UAI 2021

+1
Soumya Ghosh and 3 other authors

28 July 2021