Vision Science seminars
February 2022
The effect of gravity on the perception of distance and self-motion: a multisensory perspective
Laurence Harris· Centre for Vision Research, York University, Toronto
Thu, Feb 10 · 16:00 UTC
Gravity is a constant in our lives. It provides an internalized reference to which all other perceptions are related. We can experimentally manipulate the relationship between physical gravity with other cues to the direction of “up” using virtual reality - with either HMDs or specially built tilting environments - to explore how gravity contributes to perceptual judgements. The effect of gravity can also be cancelled by running experiments on the International Space Station in low Earth orbit. Changing orientation relative to gravity - or even just perceived orientation – affects your perception of how far away things are (they appear closer when supine or prone). Cancelling gravity altogether has a similar effect. Changing orientation also affects how much visual motion is needed to perceive a particular travel distance (you need less when supine or prone). Adapting to zero gravity has the opposite effect (you need more). These results will be discussed in terms of their practical consequences and the multisensory processes involved, in particular the response to visual-vestibular conflict.
From natural scene statistics to multisensory integration: experiments, models and applications
Cesare Parise· Oculus VR
Wed, Feb 9 · 13:00 UTC
To efficiently process sensory information, the brain relies on statistical regularities in the input. While generally improving the reliability of sensory estimates, this strategy also induces perceptual illusions that help reveal the underlying computational principles. Focusing on auditory and visual perception, in my talk I will describe how the brain exploits statistical regularities within and across the senses for the perception space, time and multisensory integration. In particular, I will show how results from a series of psychophysical experiments can be interpreted in the light of Bayesian Decision Theory, and I will demonstrate how such canonical computations can be implemented into simple and biologically plausible neural circuits. Finally, I will show how such principles of sensory information processing can be leveraged in virtual and augmented reality to overcome display limitations and expand human perception.
Visual and cross-modal plasticity in adult humans
Claudia Lunghi· Laboratoire des Systèmes Perceptifs, Ecole Normale Supérieure & CNRS, Paris, France
Thu, Feb 3 · 16:00 UTC
Neuroplasticity is a fundamental property of the nervous system that is maximal early in life, within a specific temporal window called critical period. However, it is still unclear to which extent the plastic potential of the visual cortex is retained in adulthood. We have surprisingly revealed residual ocular dominance plasticity in adult humans by showing that short-term monocular deprivation unexpectedly boosts the deprived eye (both at the perceptual and at the neural level), reflecting homeostatic plasticity. This effect is accompanied by a decrease of GABAergic inhibition in the primary visual cortex and can be modulated by non-visual factors (motor activity and motor plasticity). Finally, we have found that cross-modal plasticity is preserved in adult normal-sighted humans, as short-term monocular deprivation can alter early visuo-tactile interactions. Taken together, these results challenge the classical view of a hard-wired adult visual cortex, indicating that homeostatic plasticity can be reactivated in adult humans.
Visual appearance is an important factor in product and lighting design, and depends on the combination of form, materials, context, and lighting. Such design spaces are seemingly endless and full of optical as well as perceptual interactions. A systematic approach to navigate this space and to predict the resulting appearance can support designers in their iterative work flow, avoiding losing time on trial and error and offering understanding of the optical and perceptual effects. It should also allow artistic freedom to interactively vary the design, and enable easy communication to team members and clients. I will present examples of such approaches via canonical sets, simplifying design spaces in perception-based manners to arrive at intuitive presentations, with a focus on light(ing) design and material appearance.
The pervasive role of visuospatial coding
Edward Silson· School of Philosophy, Psychology & Language Sciences, University of Edinburgh, UK
Tue, Feb 1 · 12:15 UTC
Historically, retinotopic organisation (the spatial mapping of the retina across the cortical surface) was considered the purview of early regions of visual cortex (V1-V4) only and that anterior, more cognitively involved regions abstracted this information away. The contemporary view is quite different. Here, with Advancing technologies and analysis methods, we see that retinotopic information is not simply thrown away by these regions but rather is maintained to the potential benefit of our broader cognition. This maintenance of visuospatial coding extends not only through visual cortex, but is present in parietal, frontal, medial and subcortical structures involved with coordinating-movements, mind-wandering and even memory. In this talk, I will outline some of the key empirical findings from my own work and the work of others that shaped this contemporary perspective.
January 2022
Separable pupillary signatures of perception and action during perceptual multistability
Jan Brascamp· Michigan State University
Wed, Jan 26 · 13:00 UTC
The pupil provides a rich, non-invasive measure of the neural bases of perception and cognition, and has been of particular value in uncovering the role of arousal-linked neuromodulation, which alters cortical processing as well as pupil size. But pupil size is subject to a multitude of influences, which complicates unique interpretation. We measured pupils of observers experiencing perceptual multistability -- an ever-changing subjective percept in the face of unchanging but inconclusive sensory input. In separate conditions the endogenously generated perceptual changes were either task-relevant or not, allowing a separation between perception-related and task-related pupil signals. Perceptual changes were marked by a complex pupil response that could be decomposed into two components: a dilation tied to task execution and plausibly indicative of an arousal-linked noradrenaline surge, and an overlapping constriction tied to the perceptual transient and plausibly a marker of altered visual cortical representation. Constriction, but not dilation, amplitude systematically depended on the time interval between perceptual changes, possibly providing an overt index of neural adaptation. These results show that the pupil provides a simultaneous reading on interacting but dissociable neural processes during perceptual multistability, and suggest that arousal-linked neuromodulation shapes action but not perception in these circumstances. This presentation covers work that was published in e-life
Did you see that hazard? Scanning and detection deficits of drivers with hemianopia
Alexandra Bowers· Harvard Ophthalmology
Tue, Jan 25 · 16:00 UTC
Synergy of color and motion vision for detecting approaching objects in Drosophila
Kit Longden· Janelia Research Campus, HHMI
Mon, Jan 24 · 16:00 UTC
I am working on color vision in Drosophila, identifying behaviors that involve color vision and understanding the neural circuits supporting them (Longden 2016). I have a long-term interest in understanding how neural computations operate reliably under changing circumstances, be they external changes in the sensory context, or internal changes of state such as hunger and locomotion. On internal state-modulation of sensory processing, I have shown how hunger alters visual motion processing in blowflies (Longden et al. 2014), and identified a role for octopamine in modulating motion vision during locomotion (Longden and Krapp 2009, 2010). On responses to external cues, I have shown how one kind of uncertainty in the motion of the visual scene is resolved by the fly (Saleem, Longden et al. 2012), and I have identified novel cells for processing translation-induced optic flow (Longden et al. 2017). I like working with colleagues who use different model systems, to get at principles of neural operation that might apply in many species (Ding et al. 2016, Dyakova et al. 2015). I like work motivated by computational principles - my background is computational neuroscience, with a PhD on models of memory formation in the hippocampus (Longden and Willshaw, 2007).
What does the primary visual cortex tell us about object recognition?
Tiago Marques· MIT
Mon, Jan 24 · 13:00 UTC
Object recognition relies on the complex visual representations in cortical areas at the top of the ventral stream hierarchy. While these are thought to be derived from low-level stages of visual processing, this has not been shown, yet. Here, I describe the results of two projects exploring the contributions of primary visual cortex (V1) processing to object recognition using artificial neural networks (ANNs). First, we developed hundreds of ANN-based V1 models and evaluated how their single neurons approximate those in the macaque V1. We found that, for some models, single neurons in intermediate layers are similar to their biological counterparts, and that the distributions of their response properties approximately match those in V1. Furthermore, we observed that models that better matched macaque V1 were also more aligned with human behavior, suggesting that object recognition is derived from low-level. Motivated by these results, we then studied how an ANN’s robustness to image perturbations relates to its ability to predict V1 responses. Despite their high performance in object recognition tasks, ANNs can be fooled by imperceptibly small, explicitly crafted perturbations. We observed that ANNs that better predicted V1 neuronal activity were also more robust to adversarial attacks. Inspired by this, we developed VOneNets, a new class of hybrid ANN vision models. Each VOneNet contains a fixed neural network front-end that simulates primate V1 followed by a neural network back-end adapted from current computer vision models. After training, VOneNets were substantially more robust, outperforming state-of-the-art methods on a set of perturbations. While current neural network architectures are arguably brain-inspired, these results demonstrate that more precisely mimicking just one stage of the primate visual system leads to new gains in computer vision applications and results in better models of the primate ventral stream and object recognition behavior.
Commonly used face cognition tests yield low reliability and inconsistent performance: Implications for test design, analysis, and interpretation of individual differences data
Anna Bobak, Alex Jones· University of Stirling & Swansea University
Thu, Jan 20 · 16:00 UTC
Unfamiliar face processing (face cognition) ability varies considerably in the general population. However, the means of its assessment are not standardised, and selected laboratory tests vary between studies. It is also unclear whether 1) the most commonly employed tests are reliable, 2) participants show a degree of consistency in their performance, 3) and the face cognition tests broadly measure one underlying ability, akin to general intelligence. In this study, we asked participants to perform eight tests frequently employed in the individual differences literature. We examined the reliability of these tests, relationships between them, consistency in participants’ performance, and used data driven approaches to determine factors underpinning performance. Overall, our findings suggest that the reliability of these tests is poor to moderate, the correlations between them are weak, the consistency in participant performance across tasks is low and that performance can be broadly split into two factors: telling faces together, and telling faces apart. We recommend that future studies adjust analyses to account for stimuli (face images) and participants as random factors, routinely assess reliability, and that newly developed tests of face cognition are examined in the context of convergent validity with other commonly used measures of face cognition ability.
A novel form of retinotopy in area V2 highlights location-dependent feature selectivity in the visual system
Madineh Sedigh-Sarvestani· Max Planck Florida Institute for Neuroscience
Wed, Jan 19 · 16:30 UTC
Topographic maps are a prominent feature of brain organization, reflecting local and large-scale representation of the sensory surface. Traditionally, such representations in early visual areas are conceived as retinotopic maps preserving ego-centric retinal spatial location while ensuring that other features of visual input are uniformly represented for every location in space. I will discuss our recent findings of a striking departure from this simple mapping in the secondary visual area (V2) of the tree shrew that is best described as a sinusoidal transformation of the visual field. This sinusoidal topography is ideal for achieving uniform coverage in an elongated area like V2 as predicted by mathematical models designed for wiring minimization, and provides a novel explanation for stripe-like patterns of intra-cortical connections and functional response properties in V2. Our findings suggest that cortical circuits flexibly implement solutions to sensory surface representation, with dramatic consequences for large-scale cortical organization. Furthermore our work challenges the framework of relatively independent encoding of location and features in the visual system, showing instead location-dependent feature sensitivity produced by specialized processing of different features in different spatial locations. In the second part of the talk, I will propose that location-dependent feature sensitivity is a fundamental organizing principle of the visual system that achieves efficient representation of positional regularities in visual input, and reflects the evolutionary selection of sensory and motor circuits to optimally represent behaviorally relevant information. The relevant papers can be found here: V2 retinotopy (Sedigh-Sarvestani et al. Neuron, 2021) Location-dependent feature sensitivity (Sedigh-Sarvestani et al. Under Review, 2022)
A Flash of Darkness within Dusk: Crossover inhibition in the mouse retina
Henrique Von Gersdorff· OHSU
Tue, Jan 18 · 13:00 UTC
To survive in the wild small rodents evolved specialized retinas. To escape predators, looming shadows need to be detected with speed and precision. To evade starvation, small seeds, grass, nuts and insects need to also be detected quickly. Some of these succulent seeds and insects may be camouflaged offering only low contrast targets.Moreover, these challenging tasks need to be accomplished continuously at dusk, night, dawn and daytime. Crossover inhibition is thought to be involved in enhancing contrast detectionin the microcircuits of the inner plexiform layer of the mammalian retina. The AII amacrine cells are narrow field cells that play a key role in crossover inhibition. Our lab studies the synaptic physiology that regulates glycine release from AII amacrine cellsin mouse retina. These interneurons receive excitation from rod and conebipolar cells and transmit excitation to ON-type bipolar cell terminals via gap junctions. They also transmit inhibition via multiple glycinergic synapses onto OFF bipolar cell terminals.AII amacrine cells are thus a central hub of synaptic information processing that cross links the ON and the OFF pathways. What are the functions of crossover inhibition? How does it enhance contrast detection at different ambient light levels? How is the dynamicrange, frequency response and synaptic gain of glycine release modulated by luminance levels and circadian rhythms? How is synaptic gain changed by different extracellular neuromodulators, like dopamine, and by intracellular messengers like cAMP, phosphateand Ca2+ ions from Ca2+ channels and Ca2+ stores? My talk will try to answer some of these questions and will pose additional ones. It will end with further hypothesis and speculations on the multiple roles of crossover inhibition.
What happens to our ability to perceive multisensory information as we age?
Fiona Newell· Trinity Collge Dublin
Thu, Jan 13 · 16:00 UTC
Our ability to perceive the world around us can be affected by a number of factors including the nature of the external information, prior experience of the environment, and the integrity of the underlying perceptual system. A particular challenge for the brain is to maintain a coherent perception from information encoded by the peripheral sensory organs whose function is affected by typical, developmental changes across the lifespan. Yet, how the brain adapts to the maturation of the senses, as well as experiential changes in the multisensory environment, is poorly understood. Over the past few years, we have used a range of multisensory tasks to investigate the role of ageing on the brain’s ability to merge sensory inputs. In particular, we have embedded an audio-visual task based on the sound-induced flash illusion (SIFI) into a large-scale, longitudinal study of ageing. Our findings support the idea that the temporal binding window (TBW) is modulated by age and reveal important individual differences in this TBW that may have clinical implications. However, our investigations also suggest the TWB is experience-dependent with evidence for both long and short term behavioural plasticity. An overview of these findings, including recent evidence on how multisensory integration may be associated with higher order functions, will be discussed.
If we can make computers play chess, why can't we make them see?
SP Arun· IISc, Bangalore
Mon, Jan 3 · 21:30 UTC
If we can make computers play chess and even Jeopardy and Go, then why can't we make them see like us? How does our brain solve the problem of seeing? I will describe some of our recent insights into understanding object recognition in the brain using behavioral, neuronal and computational methods.
December 2021
Wiring Minimization of Deep Neural Networks Reveal Conditions in which Multiple Visuotopic Areas Emerge
Dina Obeid· Harvard University
Wed, Dec 15 · 05:00 UTC
The visual system is characterized by multiple mirrored visuotopic maps, with each repetition corresponding to a different visual area. In this work we explore whether such visuotopic organization can emerge as a result of minimizing the total wire length between neurons connected in a deep hierarchical network. Our results show that networks with purely feedforward connectivity typically result in a single visuotopic map, and in certain cases no visuotopic map emerges. However, when we modify the network by introducing lateral connections, with sufficient lateral connectivity among neurons within layers, multiple visuotopic maps emerge, where some connectivity motifs yield mirrored alternations of visuotopic maps–a signature of biological visual system areas. These results demonstrate that different connectivity profiles have different emergent organizations under the minimum total wire length hypothesis, and highlight that characterizing the large-scale spatial organizing of tuning properties in a biological system might also provide insights into the underlying connectivity.
Spatial Integration in Normal Face Processing and Its Breakdown in Congenital Prosopagnosia
Galia Avidan· Ben Gurion U
Tue, Dec 14 · 16:00 UTC
Molecular recognition and the assembly of feature-selective retinal circuits
Arjun Krishnaswamy· Department of Physiology, McGill University
Tue, Dec 14 · 05:00 UTC
Roles of attention and consciousness in perceptual learning
Kazuhisa Shibata· RIKEN Center for Brain Science
Mon, Dec 13 · 22:00 UTC
Visual perceptual learning (VPL) is defined as improved performance on a visual task due to visual experience. It was once argued that attention to a visual feature is necessary for VPL of the feature to occur. Contrary to this view, a phenomenon called task-irrelevant VPL demonstrated that VPL can occur due to exposure to a feature which is sub-threshold and task-irrelevant, and therefore, unattended. A series of findings based on task-irrelevant VPL has indicated the following two mechanisms. First, attention to a feature facilitates VPL of the feature while inhibiting VPL of unattended and supra-threshold features. Second, reward paired with a feature enables VPL of the feature irrespective of whether the feature is attended or not. However, we recently found an additional twist; VPL of a task-irrelevant and supra-threshold feature embedded in a natural scene is not subject to the inhibition of attention. This new finding suggests a need to revise the current view or add a new mechanism as to how VPL occurs.