bioRxiv Science⌕ Search

bioRxiv · 10.1101/2023.02.28.530443

Dimensionality and ramping: Signatures of sentence integration in the dynamics of brains and deep language models

Abstract

A sentence is more than the sum of its words: its meaning depends on how they combine with one another. The brain mechanisms underlying such semantic composition remain poorly understood. To shed light on the neural vector code underlying semantic composition, we introduce two hypotheses: First, the intrinsic dimensionality of the space of neural representations should increase as a sentence unfolds, paralleling the growing complexity of its semantic representation, and second, this progressive integration should be reflected in ramping and sentence-final signals. To test these predictions, we designed a dataset of closely matched normal and Jabberwocky sentences (composed of meaningless pseudo words) and displayed them to deep language models and to 11 human participants (5 men and 6 women) monitored with simultaneous magneto-encephalography and intracranial electro-encephalography. In both deep language models and electrophysiological data, we found that representational dimensionality was higher for meaningful sentences than Jabberwocky. Furthermore, multivariate decoding of normal versus Jabberwocky confirmed three dynamic patterns: (i) a phasic pattern following each word, peaking in temporal and parietal areas, (ii) a ramping pattern, characteristic of bilateral inferior and middle frontal gyri, and (iii) a sentence-final pattern in left superior frontal gyrus and right orbitofrontal cortex. These results provide a first glimpse into the neural geometry of semantic integration and constrain the search for a neural code of linguistic composition. Significance statementStarting from general linguistic concepts, we make two sets of predictions in neural signals evoked by reading multi-word sentences. First, the intrinsic dimensionality of the representation should grow with additional meaningful words. Second, the neural dynamics should exhibit signatures of encoding, maintaining, and resolving semantic composition. We successfully validated these hypotheses in deep Neural Language Models, artificial neural networks trained on text and performing very well on many Natural Language Processing tasks. Then, using a unique combination of magnetoencephalography and intracranial electrodes, we recorded high-resolution brain data from human participants while they read a controlled set of sentences. Time-resolved dimensionality analysis showed increasing dimensionality with meaning, and multivariate decoding allowed us to isolate the three dynamical patterns we had hypothesized.

Source connections

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Desbordes, T., Dehaene, S., Lakretz, Y., Chanoine, V., Benar, C., Badier, J.-M., Oquab, M., Caron, R., Trebuchon, A., King, J.-R.. 2023-03-01. Dimensionality and ramping: Signatures of sentence integration in the dynamics of brains and deep language models. https://doi.org/10.1101/2023.02.28.530443

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

Neurodegeneration-inducing macromolecules exit the brain via nanovascular conduits formed by reticular fibroblasts

Accumulation of proteins such as amyloid beta (Abeta), hyperphosphorylated tau and alpha-synuclein within the brain alters neural information processing and causes neurodegeneration(1-3), but how toxic solutes are cleared from the brain remains highly controversial(4,5). Proposed exit routes include efflux across endothelial cells into the blood(6,7), and movement to the pial surface via vasomotion-induced pumping along spaces within arteriolar smooth muscle(8) or via outflow along the perivascular space of ascending venules promoted by water flux through astrocytes (the glymphatic system(9)). From the pial surface of the brain, drainage may continue to dural lymphatics, along the outer sheaths of exiting cranial nerves and across the cribriform plate(10-14). We now report the presence, in mice and humans, of 2 micron diameter conduits that remove fluorescently labelled tau and Abeta from the brain. These conduits form a spatially-organised mesh within the walls of penetrating arterioles and pial arteries, and around the surface of ascending venules and deep cerebral and pial veins. They course through the pial and arachnoid layers to span the CSF space, wrapping the brain and cranial nerves. They are formed of reticular fibroblasts, which label for VE-cadherin(15) and PDGFRalpha(16), the lymphatic markers(17) podoplanin, VEGFR3 and Prox1, and reticular fibroblast extracellular matrix components collagen I and VI(16,18-20). Parenchymal tau drains from the brain at a similar rate via arteriolar conduits and via conduits around venules, arguing against preferential removal by a glymphatic mechanism. In Alzheimer's disease model mice, Abeta is seen traversing these lymph node-like conduits. Modulation of molecular transfer via this route may accelerate or delay cognitive decline, and slowed transfer from arteriolar to pial-arachnoid conduits may initiate cerebral amyloid angiopathy.

neuroscience↗

Analysis of the influence of gradual changes in matrix sentence similarity on neural envelope tracking

Neural tracking of speech is a well-established phenomenon in neuroscience. However, for speech signals with a fixed structure, significant correlations between speech envelopes and neurophysiological representations occur even for unheard sentences. We exploit a structured speech-in-noise matrix hearing test (Oldenburger Sentence Test, OLSA) to systematically quantify the relationship between acoustic sentence similarity and neural tracking. Simultaneous magnetoencephalography (MEG) and 76-channel electroencephalography (EEG) data, including 16 channels positioned directly around the ears (ear-EEG), were recorded from 21 young adults with normal hearing during the presentation of clean-speech audiobooks and OLSA sentences at six signal-to-noise ratios. A linear decoder trained on audiobooks reconstructed OLSA sentence envelopes. Reconstruction accuracies were compared using a linear mixed model across heard (matched) and unheard (mismatched) sentences of varying acoustic similarity. Significant reconstruction accuracies were achieved across MEG, EEG, and ear-EEG for both matched and mismatched sentences. For mismatched sentences, these accuracies gradually increased with their acoustic similarity to the heard speech data. The high similarity between sentences, which is especially prominent in matrix tests, can cause significant spurious tracking for mismatched stimuli. This effect can reach levels comparable to those of matched sentences and can be mistaken for true neural tracking. Robust neural tracking across modalities further supported the established viability of ear-EEG compared to whole-head systems.

neuroscience↗

Seizures and tauopathy following neurotrauma are mediated by prion protein and metabotropic glutamate receptor 5

Traumatic brain injury (TBI) is one of the world's leading causes of death and disability and a major risk factor for dementias. The primary dementia associated with TBI is chronic traumatic encephalopathy (CTE), a neurodegenerative disease classified as a tauopathy, in which toxic tau molecules lead to disease pathologies and degeneration. The processes that lead to tauopathy and subsequent dementia after TBI remain unclear. Here, we built upon the finding that seizures after TBI may be a mechanism leading to tauopathy, by dissecting the functions of the metabotropic glutamate receptor 5 - cellular prion protein (mGluR5-PrPC) pathway. We delivered TBI to larval in a blast paradigm, and quantified aggregation of Tau via a genetically-encoded Tau-GFP fusion reporter. Zebrafish larvae lacking prp2 (homolog of mammalian cellular Prion Protein, PrPC) displayed a 168% increase in post-traumatic seizures activity after TBI. An mGluR5 agonist (CHPG) reduced post-traumatic seizures, whereas an mGluR5 antagonist (MPEP) increased post-traumatic seizures. Moreover, agonizing mGluR5 reduced tau aggregation and antagonizing mGluR5 increased tau burden. Larvae seizing from convulsants, rather than TBI, were treated with CHPG/MPEP and provided a similar pattern of outcomes, suggesting seizures may be a factor needed for mGluR5 activity to influence tau aggregation. The PrPC-mGluR5 pathway is proposed as one candidate pathomechanism linking TBI to subsequent seizures and tauopathy, and thus it warrants investigation as a target for prophylactic interventions.

neuroscience↗