arXiv ScienceSearch

arXiv · 2407.04680

Lost in Translation: The Algorithmic Gap Between LMs and the Brain

Abstract

Language Models (LMs) have achieved impressive performance on various linguistic tasks, but their relationship to human language processing in the brain remains unclear. This paper examines the gaps and overlaps between LMs and the brain at different levels of analysis, emphasizing the importance of looking beyond input-output behavior to examine and compare the internal processes of these systems. We discuss how insights from neuroscience, such as sparsity, modularity, internal states, and interactive learning, can inform the development of more biologically plausible language models. Furthermore, we explore the role of scaling laws in bridging the gap between LMs and human cognition, highlighting the need for efficiency constraints analogous to those in biological systems. By developing LMs that more closely mimic brain function, we aim to advance both artificial intelligence and our understanding of human cognition.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Tommaso Tosato, Pascal Jr Tikeng Notsawo, Saskia Helbling, Irina Rish, Guillaume Dumas. 2024-07-05. Lost in Translation: The Algorithmic Gap Between LMs and the Brain. https://arxiv.org/abs/2407.04680

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

The Platonic brain bridge hypothesis: human brain networks as an architectural prior for multimodal large language models

Multimodal large language models predict brain activity, but brain alignment has been a measurement, not a design tool. We propose the Platonic brain bridge hypothesis: omni models, multimodal large language models that process video, audio and text jointly, converge on brain-like representations usable in both directions. From model to brain, brain-likeness of seven omni models is stable across participants, rises with every input channel in three bases, and our encoders lead the Algonauts 2025 out-of-distribution leaderboard. From brain to model, Brain-MoE fixes the expert partition of a frozen base to the seven networks of human cortex, trains experts on network-labelled Brain-AVQA questions, raises held-out accuracy in all 15 model-benchmark pairs by 6.42 percentage points on average and exceeds capacity-matched random experts in 14. Brain-Scope localizes the correspondence to sparse features whose removal weakens brain prediction. Human brain organization is therefore a usable architectural prior for multimodal large language models.

q-bio.NC

When Teachers Smile or Frown: A Profile-Based Analysis of Achievement Emotions

Achievement emotions shape how students engage with and learn from academic tasks, yet most studies examine individual emotions rather than co-occurring affective profiles and their dynamics. We examined latent achievement-emotion profiles and their transitions following exposure to different instructor facial expressions during a video lecture. Self-reported data from 78 Grade VII and VIII students revealed three profiles: enthusiastic, demotivated, and vulnerable. Profile transitions differed across instructor conditions, with happy expressions favouring more adaptive transitions and angry expressions favouring transitions toward demotivation. Exploratory factor analysis and Bayesian structural modelling further identified preparedness and cognitive restraint as regulatory dimensions associated with profile switching.

q-bio.NC

Nonlinear dynamics of random neural networks with second-order synaptic motifs

Classical theories of random neural networks typically assume independent connectivity, overlooking the local motif structures prevalent in biological circuits. Here, we investigate how four second-order synaptic motifs (chain, reciprocal, convergent, and divergent) shape the dynamics of nonlinear firing-rate networks. While previous studies have established that chain correlations generate outlier eigenvalues, we demonstrate that these motifs also jointly reshape the Jacobian eigenvalue bulk. Using the path-integral formalism, we derive a dynamic mean-field theory which reveals that the chain motif acts as a retarded feedback of the ensemble-mean activity through the response kernel, producing a rich repertoire of dynamical regimes, including ferromagnetic states and limit cycles. At sufficiently large magnitude, negative chain correlations produce a glassy, multistable regime that was previously mainly associated with partially symmetric networks. Our theory also distinguishes convergent from divergent motifs: divergent correlations primarily rescale temporal noise, while convergent correlations suppress temporal chaos by converting nonzero mean activity into quenched heterogeneity. Finally, analyses of the Lyapunov spectrum and participation-ratio dimension show that motif structure changes the geometry of chaotic activity, reducing entropy production and attractor dimensionality even when the effective spectral edge is held fixed. Together, these findings establish second-order motifs as a fundamental structural mechanism governing the dynamical regimes of local cortical circuits.

q-bio.NC