Observing others stay or switch : how social prediction errors are integrated into reward reversal learning

Ihssen, N, Mussweiler, T M and Linden, D E J (2016) Observing others stay or switch : how social prediction errors are integrated into reward reversal learning. Cognition, 153 (August). pp. 19-32. ISSN 0010-0277

Official URL: http://www.sciencedirect.com/science/article/pii/S...

DOI: https://doi.org/10.1016/j.cognition.2016.04.012

Abstract

Reward properties of stimuli can undergo sudden changes, and the detection of these 'reversals' is often made difficult by the probabilistic nature of rewards/punishments. Here we tested whether and how humans use social information (someone else's choices) to overcome uncertainty during reversal learning. We show a substantial social influence during reversal learning, which was modulated by the type of observed behavior. Participants frequently followed observed conservative choices (no switches after punishment) made by the (fictitious) other player but ignored impulsive choices (switches), even though the experiment was set up so that both types of response behavior would be similarly beneficial/detrimental (Study 1). Computational modeling showed that participants integrated the observed choices as a 'social prediction error' instead of ignoring or blindly following the other player. Modeling also confirmed higher learning rates for 'conservative' versus 'impulsive' social prediction errors. Importantly, this 'conservative bias' was boosted by interpersonal similarity, which in conjunction with the lack of effects observed in a non-social control experiment (Study 2) confirmed its social nature. A third study suggested that relative weighting of observed impulsive responses increased with increased volatility (frequency of reversals). Finally, simulations showed that in the present paradigm integrating social and reward information was not necessarily more adaptive to maximize earnings than learning from reward alone. Moreover, integrating social information increased accuracy only when conservative and impulsive choices were weighted similarly during learning. These findings suggest that to guide decisions in choice contexts that involve reward reversals humans utilize social cues conforming with their preconceptions more strongly than cues conflicting with them, especially when the other is similar.

More Details

[error in script]

Item Type:	Article
Subject Areas:	Organisational Behaviour
Additional Information:	© 2017 Elsevier B.V. This is an open access article published under the CC BY license (http://creativecommons.org/licenses/by/4.0/).
Funder Name:	Economic and Social Research Council
Date Deposited:	08 Sep 2017 09:59
Date of first compliant deposit:	08 Sep 2017
Subjects:	Choice Prediction
Last Modified:	20 Apr 2024 01:31
URI:	https://lbsresearch.london.edu/id/eprint/886

[error in script] More

Export and Share

Download

Published Version - Text

Download (2MB)

Available under License

Statistics

Altmetrics

View details on Dimensions' website

Downloads from LBS Research Online

View details

Actions (login required)

Edit Item