arXiv ScienceSearch

arXiv · 2210.12294

Bayesian inference is facilitated by modular neural networks with different time scales

Abstract

Various animals, including humans, have been suggested to perform Bayesian inferences to handle noisy, time-varying external information. In performing Bayesian inference, the prior distribution must be shaped by sampling noisy external inputs. However, the mechanism by which neural activities represent such distributions has not yet been elucidated. In this study, we demonstrated that the neural networks with modular structures including fast and slow modules effectively represented the prior distribution in performing accurate Bayesian inferences. Using a recurrent neural network consisting of a main module connected with input and output layers and a sub-module connected only with the main module and having slower neural activity, we demonstrated that the modular network with distinct time scales performed more accurate Bayesian inference compared with the neural networks with uniform time scales. Prior information was represented selectively by the slow sub-module, which could integrate observed signals over an appropriate period and represent input means and variances. Accordingly, the network could effectively predict the time-varying inputs. Furthermore, by training the time scales of neurons starting from networks with uniform time scales and without modular structure, the above slow-fast modular network structure spontaneously emerged as a result of learning wherein prior information was selectively represented in the slower sub-module. These results explain how the prior distribution for Bayesian inference is represented in the brain, provide insight into the relevance of modular structure with time scale hierarchy to information processing, and elucidate the significance of brain areas with slower time scales.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Kohei Ichikawa, Kunihiko Kaneko. 2022-10-21. Bayesian inference is facilitated by modular neural networks with different time scales. https://arxiv.org/abs/2210.12294

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

The Platonic brain bridge hypothesis: human brain networks as an architectural prior for multimodal large language models

Multimodal large language models predict brain activity, but brain alignment has been a measurement, not a design tool. We propose the Platonic brain bridge hypothesis: omni models, multimodal large language models that process video, audio and text jointly, converge on brain-like representations usable in both directions. From model to brain, brain-likeness of seven omni models is stable across participants, rises with every input channel in three bases, and our encoders lead the Algonauts 2025 out-of-distribution leaderboard. From brain to model, Brain-MoE fixes the expert partition of a frozen base to the seven networks of human cortex, trains experts on network-labelled Brain-AVQA questions, raises held-out accuracy in all 15 model-benchmark pairs by 6.42 percentage points on average and exceeds capacity-matched random experts in 14. Brain-Scope localizes the correspondence to sparse features whose removal weakens brain prediction. Human brain organization is therefore a usable architectural prior for multimodal large language models.

q-bio.NC

When Teachers Smile or Frown: A Profile-Based Analysis of Achievement Emotions

Achievement emotions shape how students engage with and learn from academic tasks, yet most studies examine individual emotions rather than co-occurring affective profiles and their dynamics. We examined latent achievement-emotion profiles and their transitions following exposure to different instructor facial expressions during a video lecture. Self-reported data from 78 Grade VII and VIII students revealed three profiles: enthusiastic, demotivated, and vulnerable. Profile transitions differed across instructor conditions, with happy expressions favouring more adaptive transitions and angry expressions favouring transitions toward demotivation. Exploratory factor analysis and Bayesian structural modelling further identified preparedness and cognitive restraint as regulatory dimensions associated with profile switching.

q-bio.NC

Nonlinear dynamics of random neural networks with second-order synaptic motifs

Classical theories of random neural networks typically assume independent connectivity, overlooking the local motif structures prevalent in biological circuits. Here, we investigate how four second-order synaptic motifs (chain, reciprocal, convergent, and divergent) shape the dynamics of nonlinear firing-rate networks. While previous studies have established that chain correlations generate outlier eigenvalues, we demonstrate that these motifs also jointly reshape the Jacobian eigenvalue bulk. Using the path-integral formalism, we derive a dynamic mean-field theory which reveals that the chain motif acts as a retarded feedback of the ensemble-mean activity through the response kernel, producing a rich repertoire of dynamical regimes, including ferromagnetic states and limit cycles. At sufficiently large magnitude, negative chain correlations produce a glassy, multistable regime that was previously mainly associated with partially symmetric networks. Our theory also distinguishes convergent from divergent motifs: divergent correlations primarily rescale temporal noise, while convergent correlations suppress temporal chaos by converting nonzero mean activity into quenched heterogeneity. Finally, analyses of the Lyapunov spectrum and participation-ratio dimension show that motif structure changes the geometry of chaotic activity, reducing entropy production and attractor dimensionality even when the effective spectral edge is held fixed. Together, these findings establish second-order motifs as a fundamental structural mechanism governing the dynamical regimes of local cortical circuits.

q-bio.NC