Computational role of eccentricity dependent cortical magnification
Name
CBMM-Memo-017.pdf
Size
1.04 MB
Format
Adobe PDF
Checksum (MD5)
5c7ec3ab4e693e84aa6e9c28c72a3b54
Author(s) • •
Poggio, Tomaso
Mutch, Jim
Isik, Leyla
Date Issued
June 6, 2014
Publisher
Center for Brains, Minds and Machines (CBMM), arXiv
Citation
arXiv:1406.1770v1
Series/Report no.
CBMM Memo Series;017
Abstract
We develop a sampling extension of M-theory focused on invariance to scale and translation. Quite surprisingly, the theory predicts an architecture of early vision with increasing receptive field sizes and a high resolution fovea — in agreement with data about the cortical magnification factor, V1 and the retina. From the slope of the inverse of the magnification factor, M-theory predicts a cortical “fovea” in V1 in the order of 40 by 40 basic units at each receptive field size — corresponding to a foveola of size around 26 minutes of arc at the highest resolution, ≈6 degrees at the lowest resolution. It also predicts uniform scale invariance over a fixed range of scales independently of eccentricity, while translation invariance should depend linearly on spatial frequency. Bouma’s law of crowding follows in the theory as an effect of cortical area-by-cortical area pooling; the Bouma constant is the value expected if the signature responsible for recognition in the crowding experiments originates in V2. From a broader perspective, the emerging picture suggests that visual recognition under natural conditions takes place by composing information from a set of fixations, with each fixation providing recognition from a space-scale image fragment — that is an image patch represented at a set of increasing sizes and decreasing resolutions.
Subjects
Invariance
Theories for Intelligence
Machine Learning
Vision
Terms of Use
Attribution-NonCommercial 3.0 United States
Persistent DSpace Link