Skip to content

Learning predictive neural representations by straightening natural videos

Xueyan Niu, Cristina Savin, Eero Simoncelli

Center for Neural Science, New York University; NYU; Center for Neural Science, Center for Data Science; New York University / Flatiron Institute; Neural Science

COSYNE 2023
Mar 10, 2023
Montreal, Canada

Poster

Learning predictive neural representations by straightening natural videos poster

Poster audio

Abstract

Recent experiments demonstrate that the brain transforms visual inputs into representations that follow
straighter temporal trajectories than their initial photoreceptor encoding, facilitating prediction by linear
extrapolation (Hénaff et al 2019; 2021). Can the brain use this principle to learn visual representations?
Here, we develop an objective that quantifies straightening, augment it with a regularizer to prevent collapse
to trivial solutions, and use it to train deep feedforward neural networks on video sequences. Decoding
of the learned representation with a separately trained readout network reveals that the representation
preserves visual information in video frames, and can make good next-frame predictions. Separate SVM
decoders reveal that the representation has isolated visual and physical identities in videos, including
object category, position, shape, and types of motion. When the straightening objective is applied at
different levels of the hierarchy and corresponding temporal scales, the same learning procedure yields
hierarchical temporal representations that can predict future inputs at multiple time scales. The local, fast
representations learned by our model encode and predict fine details of local motion, while the global and
slow representations encode visual aspects that persist for longer durations. Overall, our model provides a
potential mechanism by which the visual system can partition and represent features at different spatial
and temporal resolutions along the visual hierarchy.

Details

Session
Poster Session I
Cite
Xueyan Niu, Cristina Savin, Eero Simoncelli (2023). Learning predictive neural representations by straightening natural videos. COSYNE 2023. https://doi.org/10.57736/c676-9890 (opens in a new tab)

We use essential cookies to run the site. Optional analytics and public-page session replay help us improve World Wide. Learn more.