bioRxiv Science⌕ Search

bioRxiv · 10.64898/2026.03.20.712696

Impact of Kernel Dimensionality on the Generalizability and Efficiency of Convolutional Neural Networks to Decode Neural Drive from High-density Electromyography Signal

Abstract

Convolutional neural networks (CNNs) have been widely used to estimate neural drive from high-density surface electromyography (HD-sEMG) signals in neural machine interfaces owing to their real-time capability. Depending on kernel dimensionality (1D, 2D, or 3D), CNNs can extract temporal, spatial, or spatiotemporal features. Given that motor unit action potentials propagate across both space and time, architectures that exploit spatial features may offer advantages for neural drive estimation. Despite the potential importance of kernel dimensionality, its influence on neural drive estimation remains poorly understood. Existing studies have mainly evaluated CNN generalizability across participants, contraction intensities, or muscles within the same HD-sEMG dataset, while computational efficiency has seldom been considered. As a result, it remains unclear whether different kernel dimensionalities affect cross-dataset generalizability and computational efficiency. In this study, we implemented three CNN architectures--differing only in kernel dimensionality-- to investigate whether exploiting the spatial and spatiotemporal features of motor unit action potentials improves the generalizability and computational efficiency of neural drive estimation from HD-sEMG recorded during lower-limb isometric contractions. We trained the CNNs on one HD-sEMG dataset and evaluated them, without retraining, on two independent, unseen datasets recorded from different participants, sessions, and protocols--one spanning three contraction intensities and the other three muscles. All three architectures are generalized to both unseen datasets. The 2D and 3D CNNs marginally outperformed the 1D CNN with a 0.2% increase in R, while the 3D CNN showed no advantage over the 2D CNN. Computational efficiency depended on kernel dimensionality in a platform-dependent manner. On the CPU, the 3D CNN showed the slowest inference time, which was 2x slower than the 2D and 1D CNN, owing to the higher arithmetic cost of its spatiotemporal convolutions. On the GPU, all three architectures achieved similar inference times of about 1.36 ms/sample. These findings indicate that increased architectural complexity of CNN does not improve generalizability for neural drive estimation, and that a 2D CNN offers the best balance of accuracy and efficiency for a reliable, deployable CNN-based neural drive estimator--particularly on CPU-only or resource-constrained platforms.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Fu, J., Huang, H. J., Wen, Y.. 2026-03-24. Impact of Kernel Dimensionality on the Generalizability and Efficiency of Convolutional Neural Networks to Decode Neural Drive from High-density Electromyography Signal. https://doi.org/10.64898/2026.03.20.712696

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

Attention Across Scales: From Individual Variation to Social Hierarchies and Brain Networks in Semi-Free-Ranging Macaques

Attention is a fundamental brain function supporting perception, decision-making, and social behavior, and its dysfunction profoundly impairs daily life. It is both dynamic and stable, varying across observations and individuals, changing across the lifespan, and being shaped by social and environmental experience. Yet capturing this complexity remains a central challenge in neuroscience. Here, we integrated longitudinal behavioral assessments of semi-free-ranging macaques living in naturalistic social groups with resting-state fMRI. We quantified performance across days, ages, and social hierarchies and related it to intrinsic brain organization. Distinct attentional phenotypes emerged, including individuals with reduced attentional control. Performance followed an inverted-U lifespan trajectory, improving from childhood to adulthood before declining. Social status modulated attentional performance. Critically, nonlinear lifespan trajectories and associations with individual attentional differences were most clearly expressed in frontoparietal connectivity. Together, these findings reveal how sustained attention is organized across scales, providing a biological framework for its individual diversity, social modulation, and neural basis.

neuroscience↗

Decoding natural scenes from patterned optogenetic responses in mouse visual cortex

A central challenge in developing visual cortical prostheses is to determine how visual stimuli should be transformed into effective patterns of cortical stimulation. Although advances in stimulation technologies, including optogenetics, provide increasingly precise control over cortical activity, it remains unclear whether artificially evoked activity can reproduce the information content of naturally evoked visual representations. Here we establish a quantitative framework for evaluating visual encoding strategies by decoding cortical responses evoked by natural vision and patterned optogenetic stimulation. We developed a novel dual-modal paradigm in awake mice to bridge the gap between endogenous photostimulation and artificial network driving. By co-expressing the high-performance calcium indicator GCaMP6s and the red-shifted, ultra-sensitive opsin rsChRmine-oScarlet in the primary visual cortex (V1), we successfully translated dynamic natural movie frames into patterned, spatiotemporal optogenetic stimulation. Quantitative comparisons of macro-scale dynamics demonstrated that this patterned optogenetic injection evokes cortical states highly comparable and representationally aligned with those driven by actual visual photostimulation. To systematically evaluate the fidelity of these responses, we developed STAR, a deep learning model featuring spatial and temporal attention mechanisms, and successfully reconstructed the frames of natural movies from V1 signals under both experimental modalities. Collectively, our results demonstrate that complex sensory information can be both naturally encoded and synthetically injected into V1 circuits with high decoding fidelity. This work provides an empirical and computational proof-of-concept for intelligent, closed-loop biomimetic encoders, establishing a robust framework for next-generation cortical visual neuroprostheses and bidirectional brain-machine interfaces.

neuroscience↗

Why Is Spontaneous Blink Timing Informative? An Adaptive Scheduling Perspective

Spontaneous eye blinks have long been linked to cognitive processing, yet how task demands shape blink timing and its relationship to behavioral performance remains unclear. We examined spontaneous blink behavior in 576 adults performing two variants of the Continuous Performance Task (CPT). Blink occurrence and timing were most strongly modulated by the experimental condition in the more demanding CPT-AX task, whereas their association with response time was stronger in the CPT-X task, where more consistent blink timing predicted faster responses. This dissociation suggests that task structure changes not only blink behavior but also the behavioral relevance of blink timing. These findings are consistent with an adaptive scheduling account of spontaneous blinking and provide a conceptual framework for understanding when and why blink timing contains chronometric information about ongoing cognition.

neuroscience↗