bioRxiv ScienceSearch

bioRxiv · 10.1101/129007

Identification Of Animal Behavioral Strategy By Inverse Reinforcement Learning ~ Its Application To Thermotaxis In C. elegans ~

Abstract

Animals are able to reach a desired state in an environment by controlling various behavioral patterns. Identification of the behavioral strategy used for this control is important for understanding animals decision-making and is fundamental to dissect information processing done by the nervous system. However, methods for quantifying such behavioral strategies have not been fully established. In this study, we developed an inverse reinforcement-learning (IRL) framework to identify an animals behavioral strategy from behavioral time-series data. As a particular target, we applied this framework to C. elegans thermotactic behavior; after cultivation at a constant temperature with or without food, the fed and starved worms prefer and avoid from the cultivation temperature on a thermal gradient, respectively. Our IRL approach revealed that the fed worms used both absolute and temporal derivative of temperature and that their strategy comprised mixture of two strategies: directed migration (DM) and isothermal migration (IM). The DM is a strategy that the worms efficiently reach to specific temperature, which explained thermotactic behaviors of the fed worms. The IM is a strategy that the worms track along a constant temperature, which reflects isothermal tracking well observed in previous studies. We also showed the neural basis underlying the strategies, by applying our method to thermosensory neuron-deficient worms. In contrast to fed animals, the strategy of starved animals indicated that they escaped the cultivation temperature using only absolute, but not temporal derivative of temperature. Thus, our IRL-based approach is capable of identifying animal strategies from behavioral time-series data and will be applicable to wide range of behavioral studies, including decision-making of other organisms.\n\nAuthor SummaryUnderstanding animal decision-making has been a fundamental problem in neuroscience and behavioral ecology. Many studies analyze actions that represent decision-making in behavioral tasks, in which rewards are artificially designed with specific objectives. However, it is impossible to extend this artificially designed experiment to a natural environment, because in a natural environment, the rewards for freely-behaving animals cannot be clearly defined. To this end, we must reverse the current paradigm so that rewards are identified from behavioral data. Here, we propose a new reverse-engineering approach (inverse reinforcement learning) that can estimate a behavioral strategy from time-series data of freely-behaving animals. By applying this technique with thermotaxis in C. elegans, we successfully identified the reward-based behavioral strategy.

Source connections

Explore related subjects

Keep this discovery

BibTeXRIS

Yamaguchi, S., Naoki, H., Ikeda, M., Tsukada, Y., Nakano, S., Mori, I., Ishii, S.. 2017-04-20. Identification Of Animal Behavioral Strategy By Inverse Reinforcement Learning ~ Its Application To Thermotaxis In C. elegans ~. https://doi.org/10.1101/129007

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

The Unreasonable Effectiveness of Cell Types in Describing Neuronal Physiological Features

Single-cell RNA sequencing (scRNA-seq) captures detailed gene expression profiles at scale, while patch-clamp recordings measure intrinsic neuronal electrophysiological properties. Modeling the relations between these two modalities remains a challenge. Here, we compare how well electrophysiological features can be predicted by traditional transcriptomic cell type classification, representations derived from a foundational model (scGPT) pretrained on large-scale scRNA-seq datasets, ion channel-coding genes, and highly variable genes. Using paired transcriptomic and electrophysiological patch-sequencing data from 495 human neurons from neurosurgical tissue, we find that cluster-level cell type representations consistently outperform highly variable gene selection, ion channel gene selection, and context-enriched scGPT embeddings. Notably, performance varies across model architectures and initializations, and the best results are obtained by combining the outputs of separate cell type and scGPT-based models. Together, these findings suggest that traditional discrete cellular classification is highly effective in predicting physiological features. For maximum performance it can be complemented by pretrained transformer models.

neuroscience

A nonlinear inhibition pathway underlying cortical responses to tuned holographic optogenetic perturbations

Optogenetics enables causal manipulation of cortical activity. Perturbation responses can be counterintuitive due to network interactions, making theory essential for predicting them. Existing approaches often rely on linear approximations, which fail for many biologically relevant perturbations. Here we develop a nonlinear theory of responses to holographic perturbations in cell-type-specific recurrent networks with structured connectivity. We fit a nonlinear model to mouse V1 data, which shows cotuned-ensemble suppression: perturbing spatially clustered neurons with similar preferred orientations yields markedly stronger short-range suppression than perturbing untuned ensembles. We show that cotuned-ensemble suppression arises from a feature-tuned, nonlinear inhibition pathway implicating somatostatin-positive (SST) interneurons. The theory predicts that cotuned ensembles suppress parvalbumin-positive (PV) neurons but facilitate SST neurons, and links the degree of cotuned-ensemble suppression or facilitation to the variance of the SST response. This framework identifies mechanisms by which nonlinear inhibition sculpts cortical dynamics and establishes a predictive basis for targeted optogenetic interventions.

neuroscience

Proteomic signatures of APOE ε4 across human tissues and cell types in Alzheimers disease

The apolipoprotein E {varepsilon}4 (APOE {varepsilon}4) allele is the strongest genetic risk factor for late-onset Alzheimers disease (AD). However, the underlying molecular mechanisms remain unclear. This study included 1691 participants from the Religious Orders Study and Rush Memory and Aging Project (ROSMAP), 1226 participants from the Accelerating Medicines Partnership - Alzheimers Disease (AMP-AD) Diverse Cohorts Study, and 735 participants from the Alzheimers Disease Neuroimaging Initiative (ADNI). To characterise APOE {varepsilon}4 molecular effects, we analysed proteomic data from plasma, cerebrospinal fluid (CSF), and induced pluripotent stem cell (iPSC)-derived astrocytes and neurons, as well as transcriptomic and proteomic data from multiple brain regions. The association of APOE {varepsilon}4 with AD neuropathology was also examined. APOE {varepsilon}4 carriers shared a plasma proteomic signature enriched for immune processes, irrespective of AD diagnosis. A machine learning classifier trained on this signature discriminated APOE {varepsilon}4 carriers from non-carriers in an independent cohort using CSF proteomics. APOE {varepsilon}4 carriage was associated with higher Braak stages and Consortium to Establish a Registry for Alzheimers Disease (CERAD) score. However, only limited APOE {varepsilon}4-associated transcriptomic and proteomic changes were observed in bulk brain tissue, with poor cross-layer concordance. Proteomic analyses of iPSC-derived astrocytes and neurons further revealed cell-type-specific APOE {varepsilon}4-associated changes. APOE {varepsilon}4 is associated with a consistent proteomic signature across plasma and CSF. Its molecular effects in the brain differ across cell types, brain regions and molecular layers. These findings support the need for cell-type-resolved multi-omic studies to elucidate how APOE {varepsilon}4 confers AD risk.

neuroscience