Contrast invariant tuning in human perception of image content
Fruend, I.; Patel, J.; Stalker, E. D.
Show abstract
Higher levels of visual processing are progressively more invariant to low-level visual factors such as contrast. Although this invariance trend has been well documented for simple stimuli like gratings and lines, it is difficult to characterize such invariances in images with naturalistic complexity. Here, we use a generative image model based on a hierarchy of learned visual features--a Generative Adversarial Network--to constrain image manipulations to remain within the vicinity of the manifold of natural images. This allows us to quantitatively characterize visual discrimination behaviour for naturalistically complex, non-linear image manipulations. We find that human tuning to such manipulations has a factorial structure. The first factor governs image contrast with discrimination thresholds following a power law with an exponent between 0.5 and 0.6, similar to contrast discrimination performance for simpler stimuli. A second factor governs image content with approximately constant discrimination thresholds throughout the range of images studied. These results support the idea that human perception factors out image contrast relatively early on, allowing later stages of processing to extract higher level image features in a stable and robust way.
Matching journals
The top 1 journal accounts for 50% of the predicted probability mass.
Similar papers in this journal
- Mechanisms of human dynamic object recognition revealed by sequential deep neural networks 96%
- How well do rudimentary plasticity rules predict adult visual object learning? 96%
- Arousal state affects perceptual decision-making by modulating hierarchical sensory processing in a large-scale visual system model 96%
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- Efficient coding of natural scene statistics predicts discrimination thresholds for grayscale textures 97%
- Gain, not concomitant changes in spatial receptive field properties, improves task performance in a neural network attention model 96%
- Multisensory integration operates on correlated input from unimodal transients channels 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.