arXiv Science⌕ Search

arXiv · 2609.32858

Improved Learning of Molecular Energetics Through an Electron-Wise Joint Charge Density and Energy Objective

Abstract

We present a graph neural network-based learning framework trained to jointly predict electron densities and molecular energies. The model is trained to predict electron densities calculated with a Generalized Gradient Approximation (GGA) functional while simultaneously learning to predict molecular energies obtained with hybrid functional calculations at the B3LYP level of theory. Our results indicate that the rich spatial information in electron density distribution can be used to improve and accelerate the learning of accurate energies. By sharing an equivariant molecular representation across density and energy prediction heads, the model learns complementary molecular quantities within a unified framework. This provides a new promising route for practical use cases of many recent charge density learning frameworks for atomic scale simulations.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Vadim Ionas, Jonas Elsborg, Felix Ærtebjerg, Arghya Bhowmik. 2026-09-26. Improved Learning of Molecular Energetics Through an Electron-Wise Joint Charge Density and Energy Objective. https://arxiv.org/abs/2609.32858

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Pre-registered tests of solid-state-physics-inspired LLM compression: a cluster-level negative result at small-language-model scale

We report a three-month autonomous research-agent program testing five solid-state-physics-inspired compression mappings on pretrained language models, with predictions committed to git before any pilot data and a 3-sigma gate deciding PASS or SHELVE. The common anchor -- area-law / Kohn-nearsighted decay of the one-particle density matrix -- has a distance face (P001 Wannier, P002 tight-binding) and a rank face (P003 DMRG-truncated MLPs, P005 Wilson-RG, P011 tensor-train embeddings). P005 was pre-empted at Phase 1; three of four Phase-3 pilots were falsified. On the attention face, GPT-2-medium attention-versus-distance is best fit by a stretched exponential in 12 of 16 median-layer heads once probe padding is excluded, and a tight-binding cutoff costs +96% perplexity (P002); on Pythia-160M the Wannier sparsity 0.054 +/- 0.004 is indistinguishable from PCA, random-Haar and identity baselines (P001). On the rank face, per-token tensor-train bond dimension does not track surprisal (r = 0.016 vs a pre-registered 0.65) and the format inflates rather than compresses (P011). P003 is mixed: its scaling claim shelved (r = -0.434), its MPO premise died at stage-0, and its cross-paper check, r = 0.523 as first written, collapses to 0.047 under the same correction, leaving both cross-paper checks null. The results invert the pre-registered prediction that most attention heads behave like Kohn-nearsighted insulators, pointing instead to critical, glassy or heavy-tailed regimes; the inversion is specific to the <= 350M scale tested, while the rank-face no-gain result held to 7-8B. We contribute the pre-registration + 3-sigma + cluster-framing + append-only-catalogue discipline -- including why our own enforcement gate was designed but not deployed -- four pre-registered negative results with full data release, and the inversion. The catalogue holds eighteen concluded studies, seventeen negative.

cond-mat.dis-nn↗

Localization in tight-binding models with power-law distributed couplings

We study the localization properties of 1D and 2D tight-binding models with power-law distributed couplings by comparing the spectrum and the localization properties of the eigenmodes of the Laplacian and the adjacency matrix, using numerical diagonalization of these matrices for different system sizes and connectivities. These two matrices are relevant for different types of dynamical processes. While all eigenmodes of the adjacency matrix are localized for sufficiently large system sizes, the Laplacian matrix always leads to a small proportion of system-spanning modes due to a conservation law, and therefore to power-law tails in the probability distribution of the participation ratio and its relation to the eigenvalues. In one dimension, the exponent of these power laws change continuously with the exponent that characterizes the distribution of couplings. In two dimensions, the modes with the largest relaxation times change from system-spanning to localized when the exponent of the distribution of couplings becomes larger than 0.75. We provide phenomenological explanations for all these findings.

cond-mat.dis-nn↗

Exact probability distributions of complex spacing ratios in non-Hermitian random matrices

The complex spacing ratio is the complex displacement from a reference eigenvalue to its nearest neighbor divided by the corresponding displacement to its next-to-nearest neighbor. Its statistics provide a useful diagnostic of spectral correlations and nonintegrability in open quantum systems. Here, starting from the exact joint eigenvalue probability densities, we derive finite-$N$ complex-spacing-ratio distributions for the Gaussian ensembles of non-Hermitian random matrices in classes AI$^†$ and AII$^†$, realized by complex symmetric and complex self-dual random matrices, respectively. In class AII$^†$, we obtain an exact algebraic expression for arbitrary $N$ and explicitly evaluate the distributions and representative moments for $N=3, 4, 5, 6$. In class AI$^†$, although the joint density retains a noncompact integral over nonunitary eigenvector degrees of freedom, we analytically derive a normalized one-dimensional integral representation for $N=3$ and determine the asymptotic behavior, including a logarithmic correction to the cubic level repulsion and a nonanalytic contribution to the angular density. We further confirm these analytical results through direct numerical diagonalization of non-Hermitian random matrices.

cond-mat.dis-nn↗