A discrete generative model of neuronal spiking activity on microelectrode arrays
cs.LG, eess.SP, q-bio.NC
Submitted: 2026-09-20
Updated: 2026-09-20
Code: https://github.com/tanveerderik/Organoid-Binary-Spike-Spatiotemporal-Data-Modeling
License: http://creativecommons.org/licenses/by/4.0/
The gist: Generative models of neural activity could help characterize tissue dynamics, compare experimental conditions, and simulate population activity for applications ranging from disease and drug-response
Terminology
Abstract
Generative models of neural activity could help characterize tissue dynamics, compare experimental conditions, and simulate population activity for applications ranging from disease and drug-response studies to closed-loop experimentation. Existing approaches, however, typically assume a fixed set of sorted neurons, whereas high-density microelectrode arrays produce extremely sparse, array-wide binary spike volumes in which the observed subset of electrodes varies across assays. We introduce a discrete generative model that represents this activity using a shared vocabulary of spatiotemporal motifs. A residual vector-quantized autoencoder learns the motif vocabulary, while a factorized masked transformer predicts where activity occurs and which motif appears at each active location. We evaluate the model on 31 assays spanning human brain organoids and acute ex vivo human hippocampal tissue. The learned motifs are broadly reused: assay identity explains only 9% of the entropy in motif use, and motif overlap across tissue types is comparable to overlap within them. When representation quality is evaluated independently of the generative prior, our approach achieves 5.2 times the voxel-level reconstruction average precision of a matched flat tokenizer. For masked completion and free generation, the full model achieves 1.4 -- 2.6 times the site-level average precision of the matched generative baseline and outperforms it across all four families of generation metrics. These results establish a compact, reusable representation for array-wide spiking activity without learned assay-specific parameters, providing a scalable foundation for generative modeling across diverse neural preparations.
Sources
- SpikeProphecy: A Large-Scale Benchmark for Autoregressive Neural Population Forecasting
- Implicit Behavioral Decoding from Next-Step Spike Forecasts at Population Scale
- A Computational Perspective on NeuroAI and Synthetic Biological Intelligence
Related papers
- Polynomial-Augmented Neural Networks (PANNs) with Weak Orthogonality Constraints for Enhanced Function and PDE Approximation
- AIRL-S: Unifying Reinforcement Learning and Search-Based Test-Time Scaling via Adversarial Inverse Reinforcement Learning
- Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation
- Convergence issues in Relational Concept Analysis based on AOC-posets
- Beliefs Beyond Posteriors: Local-Consistency Optimisation for Bayesian Neural Networks
- Understanding Diffusion Models via Ratio-Based Function Approximation with SignReLU Networks