Back

The asymmetric transfers of visual perceptual learning determined by the stability of geometrical invariants

Yang, Y.; Zhuo, Y.; Zuo, Z.; Zhou, T.; Chen, L.

2024-12-28 neuroscience
10.1101/2024.01.02.573923 bioRxiv
Show abstract

We quickly and accurately recognize the dynamic world by extracting invariances from highly variable scenes, a process can be continuously optimized through visual perceptual learning (VPL). While it is widely accepted that the visual system prioritizes the perception of more stable invariants, the influence of the structural stability of invariants on VPL remains largely unknown. In this study, we designed three geometrical invariants with varying levels of stability for VPL: projective (e.g., collinearity), affine (e.g., parallelism), and Euclidean (e.g., orientation) invariants, following the Kleins Erlangen program. We found that learning to discriminate low-stability invariant transferred asymmetrically to those with higher stability, and that training on high-stability invariants enabled location transfer. To explore learning-associated plasticity in the visual hierarchy, we trained deep neural networks (DNNs) to model this learning procedure. We reproduced the asymmetric transfer between different invariants in DNN simulations and found that the distribution and time course of plasticity in DNNs suggested a neural mechanism similar to the reverse hierarchical theory (RHT), yet distinct in that invariant stability--not task difficulty or precision--emerged as the key determinant of learning and generalization. We propose that VPL for different invariants follows the Klein hierarchy of geometries, beginning with the extraction of high-stability invariants in higher-level visual areas, then recruiting lower-level areas for the further optimization needed to discriminate less stable invariants.

Matching journals

The top 1 journal accounts for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.