Paper Overview
Field: Machine Learning Authors: Ximeng Mao, Nanda H. Krishna, Avery Hee-Woon Ryoo, Matthew G. Perich, Guillaume Lajoie Published: 2026-07-15 arXiv: 2607.14086
Summary
Robust and accurate neural decoders are integral to neurotechnologies such as brain-computer interfaces and closed-loop experiments. Recent work has shown that tokenizing neural data at the spike level facilitates multi-session pretraining and delivers state-of-the-art decoding performance. However, current spike-based models are restricted to supervised learning (SL), limiting training to datasets with paired behavioural labels.
To address this limitation, the authors introduce MOJO (Masked autOencoder-based JOint training), a training framework for spike-tokenizing models that jointly leverages self-supervised learning (SSL) via masked autoencoding and SL objectives.
Key Findings
- Evaluated on three spiking datasets, spanning monkey motor cortex activity during reaching tasks and multi-regional mouse recordings covering visual and decision-making tasks.
- MOJO outperforms purely SL-trained models, with the improvement being especially significant when trained with limited labelled data, and in few-shot fine-tuning where only small amounts of labelled data are available for new sessions.
- Incorporating SSL produces more interpretable neuron representations, improving brain-region classification and spike-statistics prediction without explicitly optimizing for these tasks.
- MOJO generalizes beyond spike data to human speech-cortex electrocorticography (ECoG), continuing to outperform pure SL models and achieving performance comparable to neural foundation models (NFMs) specifically designed for continuous signals.
Implications
Overall, augmenting spike-tokenized models with SSL improves performance in label-scarce settings, enables the use of unlabelled data across diverse tasks and species, and generalizes to other neural modalities. These results point the way toward more flexible and scalable data usage when training neural foundation models.
--- *Source: arXiv:2607.14086*