Weight initialization shapes task organization in recurrent neural networks
Krause, R.; Mante, V.
Show abstract
Flexibly recombining computational modules is essential for biological and artificial neural networks to rapidly adapt to changing environments. This requires modules to be shared across tasks rather than rigidly segregated, yet what determines this organization remains unknown. Previous work suggests that weight initialization shapes whether networks learn task-specific or generic representations, but it is unclear whether this extends to recurrent networks and, more importantly, to network connectivity. Here, we systematically vary the initial weight variance of recurrent neural networks and study them using a framework that allows us to identify the functionally relevant connectivity subspaces for each computational module. We find that networks with low initial weight variance converge to solutions in which different subtasks rely on largely overlapping weight subspaces, whereas high-variance networks implement subtasks in higher-dimensional, more segregated weight subspaces. Our results also provide mechanistic insights with implications for interpreting biological neural circuits and for designing efficient recurrent architectures.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- From space to time: Spatial inhomogeneities lead to the emergence of spatio-temporal activity sequences in spiking neuronal networks 96%
- "Backpropagation and the brain" realized in cortical error neuron microcircuits 96%
- Learning compositional sequences with multiple time scales through a hierarchical network of spiking neurons 95%
Similar papers in this journal
- Local lateral connectivity is sufficient for replicating cortex-like topographical organization in deep neural networks 96%
- Modeling Attention and Binding in the Brain through Bidirectional Recurrent Gating 96%
- Taming the chaos gently: a Predictive Alignment learning rule in recurrent neural networks 95%
Similar papers in this journal
- Learning probabilistic representations with randomly connected neural circuits 96%
- Are place cells just memory cells? Memory compression leads to spatial tuning and history dependence 95%
- Orchestrated Excitatory and Inhibitory Learning Rules Lead to the Unsupervised Emergence of Self-sustained and Inhibition-stabilized Dynamics 95%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.