arXiv ScienceSearch

arXiv · 2609.07508

Homeostasis Revisited and Reformulated Through Hidden Markov Model Control

Abstract

A common formalization of homeostasis is the free energy principle, a framework that defines a set of desired observation values, or critical states, that the agent should reach or remain close to. Under the free energy principle, an agent should act to maximize the probability of receiving the desired observations. Here we revisit the common approach of solving the problem of maximizing the log probability of the desired observations by maximizing a variational lower bound, the so-called negative free energy. We show that, instead, an approach directly maximizing that probability under the agent's policy is better suited to, and provides a better solution for, the original homeostatic control problem. This is done using hidden Markov model (HMM) control by allowing the policy to act over hidden states or noisy versions thereof while trying to maximize the probability of repeatedly having the desired observations. HMM control largely improves performance over the variational, or free energy, approach. We also show that the optimal policy is strictly deterministic, while the variational approach leads to a stochastic policy approximation. We finally provide a homeostatic reinterpretation of the maximum occupancy principle -a principle proposing that agents ought to maximize the occupancy of action-state path space -by defining homeostatic states as any states that do not immediately entail the termination or death of the agent.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Rubén Moreno-Bote. 2026-09-07. Homeostasis Revisited and Reformulated Through Hidden Markov Model Control. https://arxiv.org/abs/2609.07508

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

The Platonic brain bridge hypothesis: human brain networks as an architectural prior for multimodal large language models

Multimodal large language models predict brain activity, but brain alignment has been a measurement, not a design tool. We propose the Platonic brain bridge hypothesis: omni models, multimodal large language models that process video, audio and text jointly, converge on brain-like representations usable in both directions. From model to brain, brain-likeness of seven omni models is stable across participants, rises with every input channel in three bases, and our encoders lead the Algonauts 2025 out-of-distribution leaderboard. From brain to model, Brain-MoE fixes the expert partition of a frozen base to the seven networks of human cortex, trains experts on network-labelled Brain-AVQA questions, raises held-out accuracy in all 15 model-benchmark pairs by 6.42 percentage points on average and exceeds capacity-matched random experts in 14. Brain-Scope localizes the correspondence to sparse features whose removal weakens brain prediction. Human brain organization is therefore a usable architectural prior for multimodal large language models.

q-bio.NC

When Teachers Smile or Frown: A Profile-Based Analysis of Achievement Emotions

Achievement emotions shape how students engage with and learn from academic tasks, yet most studies examine individual emotions rather than co-occurring affective profiles and their dynamics. We examined latent achievement-emotion profiles and their transitions following exposure to different instructor facial expressions during a video lecture. Self-reported data from 78 Grade VII and VIII students revealed three profiles: enthusiastic, demotivated, and vulnerable. Profile transitions differed across instructor conditions, with happy expressions favouring more adaptive transitions and angry expressions favouring transitions toward demotivation. Exploratory factor analysis and Bayesian structural modelling further identified preparedness and cognitive restraint as regulatory dimensions associated with profile switching.

q-bio.NC

Nonlinear dynamics of random neural networks with second-order synaptic motifs

Classical theories of random neural networks typically assume independent connectivity, overlooking the local motif structures prevalent in biological circuits. Here, we investigate how four second-order synaptic motifs (chain, reciprocal, convergent, and divergent) shape the dynamics of nonlinear firing-rate networks. While previous studies have established that chain correlations generate outlier eigenvalues, we demonstrate that these motifs also jointly reshape the Jacobian eigenvalue bulk. Using the path-integral formalism, we derive a dynamic mean-field theory which reveals that the chain motif acts as a retarded feedback of the ensemble-mean activity through the response kernel, producing a rich repertoire of dynamical regimes, including ferromagnetic states and limit cycles. At sufficiently large magnitude, negative chain correlations produce a glassy, multistable regime that was previously mainly associated with partially symmetric networks. Our theory also distinguishes convergent from divergent motifs: divergent correlations primarily rescale temporal noise, while convergent correlations suppress temporal chaos by converting nonzero mean activity into quenched heterogeneity. Finally, analyses of the Lyapunov spectrum and participation-ratio dimension show that motif structure changes the geometry of chaotic activity, reducing entropy production and attractor dimensionality even when the effective spectral edge is held fixed. Together, these findings establish second-order motifs as a fundamental structural mechanism governing the dynamical regimes of local cortical circuits.

q-bio.NC