Joel Leibo | Massachusetts Institute of Technology (MIT) - Academia.edu

Skip to main content

Joel Leibo

Massachusetts Institute of Technology (MIT), Brain and Cognitive Sciences, Research affiliate

Followers

73

Following

54

Co-authors

6

Public Views

Interests

Uploads

Papers

View-Tolerant Face Recognition and Hebbian Learning Imply Mirror-Symmetric Neural Tuning to Head Orientation

Current biology : CB, Jan 9, 2017

The primate brain contains a hierarchy of visual areas, dubbed the ventral stream, which rapidly ... more The primate brain contains a hierarchy of visual areas, dubbed the ventral stream, which rapidly computes object representations that are both specific for object identity and robust against identity-preserving transformations, like depth rotations [1, 2]. Current computational models of object recognition, including recent deep-learning networks, generate these properties through a hierarchy of alternating selectivity-increasing filtering and tolerance-increasing pooling operations, similar to simple-complex cells operations [3-6]. Here, we prove that a class of hierarchical architectures and a broad set of biologically plausible learning rules generate approximate invariance to identity-preserving transformations at the top level of the processing hierarchy. However, all past models tested failed to reproduce the most salient property of an intermediate representation of a three-level face-processing hierarchy in the brain: mirror-symmetric tuning to head orientation [7]. Here, we...

Multi-agent Reinforcement Learning in Sequential Social Dilemmas

by Vinicius Zambaldi, Thore Graepel, and Joel Leibo

Matrix games like Prisoner's Dilemma have guided research on social dilemmas for decades. However... more Matrix games like Prisoner's Dilemma have guided research on social dilemmas for decades. However, they necessarily treat the choice to cooperate or defect as an atomic action. In real-world social dilemmas these choices are temporally extended. Cooperativeness is a property that applies to policies , not elementary actions. We introduce sequential social dilemmas that share the mixed incentive structure of matrix game social dilemmas but also require agents to learn policies that implement their strategic intentions. We analyze the dynamics of policies learned by multiple self-interested independent learning agents, each using its own deep Q-network, on two Markov games we introduce here: 1. a fruit Gathering game and 2. a Wolfpack hunting game. We characterize how learned behavior in each domain changes as a function of environmental factors including resource abundance. Our experiments show how conflict can emerge from competition over shared resources and shed light on how the sequential nature of real world social dilemmas affects cooperation.

Invariant Recognition Predicts Tuning of Neurons in Sensory Cortex

Cognitive Science and Technology, 2016

Unsupervised learning of invariant representations with low sample complexity: the magic of sensory cortex or a new framework for machine learning?

Unsupervised learning of clutter-resistant visual representations from natural videos

Populations of neurons in inferotemporal cortex (IT) maintain an explicit code for object identit... more Populations of neurons in inferotemporal cortex (IT) maintain an explicit code for object identity that also tolerates transformations of object appearance e.g., position, scale, viewing angle [1, 2, 3]. Though the learning rules are not known, recent results [4, 5, 6] suggest the operation of an unsupervised temporal-association-based method e.g., Foldiak's trace rule [7]. Such methods exploit the temporal continuity of the visual world by assuming that visual experience over short timescales will tend to have invariant identity content. Thus, by associating representations of frames from nearby times, a representation that tolerates whatever transformations occurred in the video may be achieved. Many previous studies verified that such rules can work in simple situations without background clutter, but the presence of visual clutter has remained problematic for this approach. Here we show that temporal association based on large class-specific filters (templates) avoids the pr...

The invariance hypothesis implies domain-specific regions in visual cortex

The dynamics of invariant object recognition in the human visual system

Journal of Neurophysiology, 2014

Can a biologically-plausible hierarchy effectively replace face detection, alignment, and recognition pipelines?

by Joel Leibo and Tomaso Poggio

Subtasks of Unconstrained Face Recognition

by Joel Leibo and Tomaso Poggio

Unsupervised Learning of Invariant Representations in Hierarchical Architectures

by Joel Leibo and Andrea Tacchetti

Representations that are invariant to translation, scale and other transformations, can considera... more Representations that are invariant to translation, scale and other transformations, can considerably reduce the sample complexity of learning, allowing recognition of new object classes from very few examples - a hallmark of human recognition. Empirical estimates of one-dimensional projections of the distribution induced by a group of affine transformations are proven to represent a unique and invariant signature associated with an image. We show how projections yielding invariant signatures for future images can be learned automatically, and updated continuously, during unsupervised visual experience. A module performing filtering and pooling, like simple and complex cells as proposed by Hubel and Wiesel, can compute such estimates. Under this view, a pooling stage estimates a one-dimensional probability distribution. Invariance from observations through a restricted window is equivalent to a sparsity property w.r.t. to a transformation, which yields templates that are a) Gabor for optimal simultaneous invariance to translation and scale or b) very specific for complex, class-dependent transformations such as rotation in depth of faces. Hierarchical architectures consisting of this basic Hubel-Wiesel module inherit its properties of invariance, stability, and discriminability while capturing the compositional organization of the visual world in terms of wholes and parts, and are invariant to complex transformations that may only be locally affine. The theory applies to several existing deep learning convolutional architectures for image and speech recognition. It also suggests that the main computational goal of the ventral stream of visual cortex is to provide a hierarchical representation of new objects which is invariant to transformations, stable, and discriminative for recognition - this representation may be learned in an unsupervised way from natural visual experience.

Learning invariant representations and applications to face verification

by Joel Leibo and Tomaso Poggio

The Invariance Hypothesis and the Ventral Stream

Does invariant recognition predict tuning of neurons in sensory cortex?

by Joel Leibo, Andrea Tacchetti, Fabio Anselmi, and Tomaso Poggio

The computational magic of the ventral stream: sketch of a theory (and why some deep architectures work)

by Joel Leibo and Andrea Tacchetti

Body-form and body-pose recognition with a hierarchical model of the ventral stream

by Joel Leibo and Tomaso Poggio

Learning and disrupting invariance in visual recognition with a temporal association rule

Preliminary MEG decoding results (2012)

Throwing Down the Visual Intelligence Gauntlet (2012)

How can cells in the anterior medial face patch be viewpoint invariant? (2011)

Learning to discount transformations as the computational goal of visual cortex (2011)

View-Tolerant Face Recognition and Hebbian Learning Imply Mirror-Symmetric Neural Tuning to Head Orientation

Current biology : CB, Jan 9, 2017

The primate brain contains a hierarchy of visual areas, dubbed the ventral stream, which rapidly ... more The primate brain contains a hierarchy of visual areas, dubbed the ventral stream, which rapidly computes object representations that are both specific for object identity and robust against identity-preserving transformations, like depth rotations [1, 2]. Current computational models of object recognition, including recent deep-learning networks, generate these properties through a hierarchy of alternating selectivity-increasing filtering and tolerance-increasing pooling operations, similar to simple-complex cells operations [3-6]. Here, we prove that a class of hierarchical architectures and a broad set of biologically plausible learning rules generate approximate invariance to identity-preserving transformations at the top level of the processing hierarchy. However, all past models tested failed to reproduce the most salient property of an intermediate representation of a three-level face-processing hierarchy in the brain: mirror-symmetric tuning to head orientation [7]. Here, we...

Multi-agent Reinforcement Learning in Sequential Social Dilemmas

by Vinicius Zambaldi, Thore Graepel, and Joel Leibo

Matrix games like Prisoner's Dilemma have guided research on social dilemmas for decades. However... more Matrix games like Prisoner's Dilemma have guided research on social dilemmas for decades. However, they necessarily treat the choice to cooperate or defect as an atomic action. In real-world social dilemmas these choices are temporally extended. Cooperativeness is a property that applies to policies , not elementary actions. We introduce sequential social dilemmas that share the mixed incentive structure of matrix game social dilemmas but also require agents to learn policies that implement their strategic intentions. We analyze the dynamics of policies learned by multiple self-interested independent learning agents, each using its own deep Q-network, on two Markov games we introduce here: 1. a fruit Gathering game and 2. a Wolfpack hunting game. We characterize how learned behavior in each domain changes as a function of environmental factors including resource abundance. Our experiments show how conflict can emerge from competition over shared resources and shed light on how the sequential nature of real world social dilemmas affects cooperation.

Invariant Recognition Predicts Tuning of Neurons in Sensory Cortex

Cognitive Science and Technology, 2016

Unsupervised learning of invariant representations with low sample complexity: the magic of sensory cortex or a new framework for machine learning?

Unsupervised learning of clutter-resistant visual representations from natural videos

Populations of neurons in inferotemporal cortex (IT) maintain an explicit code for object identit... more Populations of neurons in inferotemporal cortex (IT) maintain an explicit code for object identity that also tolerates transformations of object appearance e.g., position, scale, viewing angle [1, 2, 3]. Though the learning rules are not known, recent results [4, 5, 6] suggest the operation of an unsupervised temporal-association-based method e.g., Foldiak's trace rule [7]. Such methods exploit the temporal continuity of the visual world by assuming that visual experience over short timescales will tend to have invariant identity content. Thus, by associating representations of frames from nearby times, a representation that tolerates whatever transformations occurred in the video may be achieved. Many previous studies verified that such rules can work in simple situations without background clutter, but the presence of visual clutter has remained problematic for this approach. Here we show that temporal association based on large class-specific filters (templates) avoids the pr...

The invariance hypothesis implies domain-specific regions in visual cortex

The dynamics of invariant object recognition in the human visual system

Journal of Neurophysiology, 2014

Can a biologically-plausible hierarchy effectively replace face detection, alignment, and recognition pipelines?

by Joel Leibo and Tomaso Poggio

Subtasks of Unconstrained Face Recognition

by Joel Leibo and Tomaso Poggio

Unsupervised Learning of Invariant Representations in Hierarchical Architectures

by Joel Leibo and Andrea Tacchetti

Representations that are invariant to translation, scale and other transformations, can considera... more Representations that are invariant to translation, scale and other transformations, can considerably reduce the sample complexity of learning, allowing recognition of new object classes from very few examples - a hallmark of human recognition. Empirical estimates of one-dimensional projections of the distribution induced by a group of affine transformations are proven to represent a unique and invariant signature associated with an image. We show how projections yielding invariant signatures for future images can be learned automatically, and updated continuously, during unsupervised visual experience. A module performing filtering and pooling, like simple and complex cells as proposed by Hubel and Wiesel, can compute such estimates. Under this view, a pooling stage estimates a one-dimensional probability distribution. Invariance from observations through a restricted window is equivalent to a sparsity property w.r.t. to a transformation, which yields templates that are a) Gabor for optimal simultaneous invariance to translation and scale or b) very specific for complex, class-dependent transformations such as rotation in depth of faces. Hierarchical architectures consisting of this basic Hubel-Wiesel module inherit its properties of invariance, stability, and discriminability while capturing the compositional organization of the visual world in terms of wholes and parts, and are invariant to complex transformations that may only be locally affine. The theory applies to several existing deep learning convolutional architectures for image and speech recognition. It also suggests that the main computational goal of the ventral stream of visual cortex is to provide a hierarchical representation of new objects which is invariant to transformations, stable, and discriminative for recognition - this representation may be learned in an unsupervised way from natural visual experience.

Learning invariant representations and applications to face verification

by Joel Leibo and Tomaso Poggio

The Invariance Hypothesis and the Ventral Stream

Does invariant recognition predict tuning of neurons in sensory cortex?

by Joel Leibo, Andrea Tacchetti, Fabio Anselmi, and Tomaso Poggio

The computational magic of the ventral stream: sketch of a theory (and why some deep architectures work)

by Joel Leibo and Andrea Tacchetti

Body-form and body-pose recognition with a hierarchical model of the ventral stream

by Joel Leibo and Tomaso Poggio

Learning and disrupting invariance in visual recognition with a temporal association rule

Preliminary MEG decoding results (2012)

Throwing Down the Visual Intelligence Gauntlet (2012)

How can cells in the anterior medial face patch be viewpoint invariant? (2011)

Learning to discount transformations as the computational goal of visual cortex (2011)