arXiv Science⌕ Search

arXiv · 2610.05419

Recurrent network dynamics explain the time course of perceptual grouping in natural scenes

Abstract

How the brain groups image elements into coherent objects in natural scenes remains unclear. Existing models for grouping in biological vision rely on simplified stimuli with explicit boundaries and cannot explain human behavior in naturalistic settings. We propose a mechanistic framework in which recurrent interactions propagate enhanced neuronal activity within and between cortical areas. The framework links computational principles, neural circuitry and perceptual psychology. Local boundary signals govern early grouping, whereas later stages integrate top-down feedback carrying information about object identity, grouping features across internal edges into coherent object representations. We instantiated this framework as a brain-inspired recurrent neural network trained to group features in natural images. The network learned to propagate enhanced activity across a cued object's representation, mirroring the brain's perceptual grouping mechanisms. The network's dynamics also predicted human reaction times. The framework links cortical dynamics to human vision and generates testable predictions for neuro-physiological and psychophysical experiments.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Sami Mollard, Alekh K. Ashok, Lore Goetschalckx, Drew Linsley, Thomas Serre, Sander M. Bohte, Pieter R. Roelfsema. 2026-10-04. Recurrent network dynamics explain the time course of perceptual grouping in natural scenes. https://arxiv.org/abs/2610.05419

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Dynamical incompatibilities in paced finger tapping experiments

Paced finger-tapping tasks are used to probe the error correction mechanism underlying sensorimotor synchronization. Despite their century-long history, fundamental contradictions persist in the literature. One such contradiction arises when comparing the two most common types of period perturbation: step change and phase shift. The stimulus sequence is exactly the same up to and including the (unexpected) perturbed stimulus. Why then would the timing of the next response be different between perturbation types, as observed? We show, both experimentally and theoretically, that responses to both types of perturbation are dynamically incompatible when recorded in separate experiments; that is, they cannot be described by a single underlying dynamical system due to the build-up of different temporal contexts. In contrast, when both types of perturbation are presented randomly within the same experiment, the responses become compatible and can be explained by a single mechanism. We conclude that a single underlying dynamical system can represent the response to all perturbation types, signs, and sizes, which is nevertheless calibrated by temporal context. Our results challenge the established idea of phase and period correction processes that are separately activated for different perturbation types.

q-bio.NC↗

From communication to computation in neurons-on-a-chip: an in silico study of neurotopomorphic computing

Living neuronal networks transform inputs through recurrent cellular and population dynamics, yet it is unknown which network architecture supports which computation. Neurons-on-a-chip turn this question into a design problem because microchannels guide axonal growth and set the network architecture. We introduce IC$^3$, an Integrated Characterisation of Communication-Driven Computation, which characterizes network state through neuronal dynamics, functional communication, and structural support. We implemented nine architectures \textit{in silico} as conductance-based spiking networks and tested each on frequency decoding, temporal-order discrimination, and fading memory. Predominantly feedforward circuits decoded best. Sequential Chain and Microchannel Diode had the lowest IC$^3$ and recruited a third of reachable neurons, yet achieved the two highest scores on both classification tasks. Across architectures, higher IC$^3$ went with lower classification scores. We term this new direction \emph{neurotopomorphic computing}, in which the physical organisation of neuronal connectivity is engineered as part of the computing substrate.

q-bio.NC↗

COMPASS 2.0: psychometric representational similarity analysis distinguishes symptom structure from personal signal

Language models can score psychiatric questionnaires from speech, but agreement with self-report may reflect the questionnaire rather than the person. We introduce psychometric representational similarity analysis, a framework for comparing the structure of speech-derived scores, self-report, item wording and theory, and implement it alongside person-level construct scoring in COMPASS 2.0. We show how similarly worded items induce covariance without psychological signal. In pre-registered discovery and confirmation analyses of clinical interviews from 275 participants, language-derived symptom geometry resembled wording more than self-report, with no structure beyond wording detected by the registered tests. Geometric agreement with self-report survived assigning participants someone else's answers, whereas person-paired scores captured distress more than specific symptoms. Complementary analyses examined counselling quality and wording structure across 34 instruments and the Research Domain Criteria (RDoC) framework. These findings distinguish agreement about psychological structure from evidence that language-derived assessments track individual people.

q-bio.NC↗