bioRxiv ScienceSearch

bioRxiv · 10.1101/067603

Habits without Values

Abstract

Habits form a crucial component of behavior. In recent years, key computational models have conceptualized habits as arising from model-free reinforcement learning (RL) mechanisms, which typically select between available actions based on the future value expected to result from each. Traditionally, however, habits have been understood as behaviors that can be triggered directly by a stimulus, without requiring the animal to evaluate expected outcomes. Here, we develop a computational model instantiating this traditional view, in which habits develop through the direct strengthening of recently taken actions rather than through the encoding of outcomes. We demonstrate that this model accounts for key behavioral manifestations of habits, including insensitivity to outcome devaluation and contingency degradation, as well as the effects of reinforcement schedule on the rate of habit formation. The model also explains the prevalent observation of perseveration in repeated-choice tasks as an additional behavioral manifestation of the habit system. We suggest that mapping habitual behaviors onto value-free mechanisms provides a parsimonious account of existing behavioral and neural data. This mapping may provide a new foundation for building robust and comprehensive models of the interaction of habits with other, more goal-directed types of behaviors and help to better guide research into the neural mechanisms underlying control of instrumental behavior more generally.

Source connections

Explore related subjects

Keep this discovery

BibTeXRIS

Kevin Miller, Amitai Shenhav, Elliot Ludvig. 2016-08-03. Habits without Values. https://doi.org/10.1101/067603

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

FUNCTIONAL MRI IN AWAKE DOGS PREDICTS SUITABILITY FOR ASSISTANCE WORK

The overall goal of this work was to measure the efficacy of fMRI for predicting whether a dog would be a successful service dog. The training and imaging were performed in 50 dogs entering advanced training at 17-21 months of age. FMRI responses were measured while each dog observed hand signals indicating either reward or no reward and given by both a familiar handler and a stranger. 49 dogs successfully completed fMRI training and scanning. Of these, 33 eventually completed service training and were matched with a person, while 10 were released for behavioral reasons. Using anatomically defined regions-of-interest in the ventral caudate, amygdala, and visual cortex, we developed a classifier based on the dogs' outcomes. We found that responses in the stranger condition were sufficient to develop an accurate brain-based classifier. On all data, the classifier had a positive predictive value of 96% with 10% false positives. The area under the receiver operating characteristic curve was 0.90 (0.79 with 4-fold cross-validation, P=0.02), indicating a significant diagnostic capability. Within the stranger condition, the differential response to [reward - no reward] in ventral caudate was positively correlated with a successful outcome, while the differential response in the amygdala was negatively correlated to outcome. These results show that successful service dogs transfer knowledge to strangers as indexed by ventral caudate activity without excessive arousal as measured in the amygdala.

Animal Behavior and Cognition

Automatic Head Tracking of The Common Marmoset.

New technologies for manipulating and recording the nervous system allow us to perform unprecedented experiments. However, the influence of our experimental manipulations on psychological processes must be inferred from their effects on behavior. Today, quantifying behavior has become the bottleneck for large-scale, high-throughput, experiments. The method presented here addresses this issue by using deep learning algorithms for video-based animal tracking. Here we describe a reliable automatic method for tracking head position and orientation from simple video recordings of the common marmoset (Callithrix jacchus). This method for measuring marmoset behavior allows for the estimation of gaze within foveal error, and can easily be adapted to a wide variety of similar tasks in biomedical research. In particular, the method has great potential for the simultaneous tracking of multiple marmosets to quantify social behaviors.

Animal Behavior and Cognition

Behavior of Caenorhabditis elegans in a nicotine gradient modulated by food

Nicotine decreases food intake, and smokers often report that they smoke to control their weight. To see whether similar phenomena could be observed in the model organism Caenorhabditis elegans, we challenged drug-naive nematodes with a chronic low (0.01 mM) and high (1 mM) nicotine concentration for 55 h (from hatching to adulthood). After that, we recorded changes in their behavior in a nicotine gradient, where they could choose a desired nicotine concentration. By using a combination of behavioral and morphometric methods, we found that both nicotine and food modulate worm behavior. In the presence of food the nematodes adapted to the low nicotine concentration, when placed in the gradient, chose a similar nicotine concentration like C. elegans adapted to the high nicotine concentration. However, in the absence of food, the nematodes adapted to the low nicotine concentration, when placed in the gradient of this alkaloid, chose a similar nicotine concentration like naive worms. The nematodes growing up in the presence of high concentrations of nicotine had a statistically smaller body size, compared to the control condition, and the presence of food did not cause any enhanced slowing movement. These results provide a platform for more detailed molecular and cellular studies of nicotine addiction and food intake in this model organism.

Animal Behavior and Cognition