arXiv ScienceSearch

arXiv subjects

Ju Li

Publications and source records attributed to Ju Li.

At least 19 recordsLinked to original sources

Coupled-cluster molecular properties across the main group that extrapolate beyond training size

Coupled-cluster theory defines the accuracy standard for molecular electronic-structure properties but scales too steeply for routine application, whereas density-functional theory is affordable yet systematically biased. We resolve this trade-off with a single equivariant network, MEHnet-MG, that predicts an effective one-electron Hamiltonian from one inexpensive B3LYP/def2-SVP calculation and derives a broad suite of properties from it (energy, optical gap, dipole, quadrupole, polarizability, Mulliken atomic charges, and Mayer bond orders) at coupled-cluster accuracy across nine main-group elements, including the under-served phosphorus, sulfur, and chlorine chemistries. The model is trained on a new in-house dataset of multi-property labels computed at the CCSD(T) level for all nine elements. On a held-out test set, it reduces the error of every property by a factor of 3.8 to 230 relative to semi-local, hybrid, and double-hybrid DFT (referenced to composite CCSD(T)/cc-pVTZ; Methods), while adding only ~25 ms wall time per molecule, delivering coupled-cluster-quality predictions at the cost of a single DFT calculation. Critically, deriving every property from a predicted Hamiltonian rather than pooling per-atom features builds the correct size-scaling into the model architecture: on pi-conjugated oligothiophenes it matches finite-field CCSD polarizability and the EOM-CCSD optical gap to ~2% at the largest sizes where those references remain affordable (44 and 37 atoms, where a single CCSD field point already costs ~500x the model's entire inference) and extrapolates the corrected trends to 58-atom chains, a regime where pooling-based architectures fail by construction. Accurate extrapolation is therefore set by the model's inductive bias rather than by the training data.

physics.chem-ph

Nonresonant optomechanical control of structural phases

Optical tweezers demonstrate how light can exert forces to trap, repel, and manipulate microscopic particles without absorption. Recent theory has suggested that such forces can extend beyond particle manipulation to drive structural phase transitions in solids. Here we apply this optomechanical principle to tin selenide (SnSe), a material where proximity to several different structural phases gives rise to its high thermoelectric figure of merit and makes it a candidate for a switchable topological crystalline insulator. Whereas the force for standard optical tweezers arises from a gradient in the intensity of a light field, the optomechanical force is mediated by a gradient in the dielectric constant as a function of phonon coordinate. Unlike conventional methods that rely on resonant excitation and absorption through the imaginary part of the dielectric function, this approach operates dispersively through the real part and can be directly driven by Raman processes, enabling selective transitions with reduced energy cost and ultrafast response. Using time-domain Raman scattering, we show that above a critical mid-infrared field strength the $A_g$ Raman modes disappear abruptly without softening, signaling the formation of a new structural phase. This phase, distinct from those induced by heating or carrier excitation, exhibits large-amplitude and long-lived modulations in its optical response. Complementing this observation, we show also evidence for an equivalent DC-field-driven structural phase transformation to a higher symmetry phase, as observed by atom probe tomography. Our study demonstrates the concept of nonresonant optomechanical phase control and defines novel opportunities for synthesizing hidden structural phases with unique functional properties.

physics.optics

Atomistic Language Models Understand and Generate Materials

Atomistic structure and natural language have long been modeled separately, with language models either calling atomistic models as tools or being fine-tuned on lossy textual encodings that discard atomistic information. We introduce Atomistic Language Models (ALMs) to pursue native multimodality, in which a single language backbone understands atomistic structures, generates materials from natural language, and optimizes crystal structures as instructed by text. By unifying a pretrained atomistic encoder, large language model, and denoising diffusion model through purely continuous projectors and staged training, ALMs achieve state-of-the-art results on crystal structure prediction and de novo generation. ALMs are enabled by a continuous bridge that maps language model embeddings directly into the steering space of atomistic diffusion, and are assisted by Text-to-Crystal Feynman-Kac (T2C-FK), a particle-based sampler that scores partial denoising trajectories to enforce stoichiometric targets at inference time. To evaluate the ability of ALMs to optimize and generate materials from natural-language prompts and 3D atom-coordinate inputs, we introduce ALM Bench, the first benchmark for text-conditioned crystal generation and optimization. Code, training data, and model weights will be released soon.

cs.LG

Complementary Thermodynamic Mechanisms of Boron and Carbon Segregation at Grain Boundaries in Nickel Alloys

Grain boundary stabilization by light interstitials is central to the performance of Ni-based superalloys, yet the thermodynamic mechanisms governing their interactions with substitutional chemistry remain poorly resolved. Here, we use hybrid Monte Carlo molecular dynamics simulations to quantify how boron and carbon modify the thermodynamic, structural, and chemical ordering of grain boundaries in Ni--Cr alloys. By analyzing interfacial state variables, site-resolved segregation spectra, local chemical ordering, and structural evolution, we show that boron and carbon stabilize grain boundaries through complementary pathways. Carbon drives saturation-controlled stabilization by recruiting Cr and conditioning boundary chemistry, while suppressing temperature-driven structural transformations of the boundary. In contrast, boron stabilizes grain boundaries through a selective mechanism that lowers the interfacial grand potential via localized ordering while permitting gradual structural evolution. These effects arise from coupled interactions between interstitial segregation and Cr redistribution, which together regulate site accessibility, chemical competition, and the range of accessible interfacial states. This work provides a thermodynamic framework for grain boundary engineering and suggests design principles for leveraging interstitial-substitutional interactions in alloys.

cond-mat.mtrl-sci

Economy of Minds: Emerging Multi-Agent Intelligence with Economic Interactions

How can a population of agents self-orchestrate and self-adapt into stronger collective intelligence without centralized control? Inspired by Friedrich Hayek's economic theory of decentralized coordination in markets, we study this question through an agent economy in which agents compete via auctions for the right to act, exchange payments, and accumulate wealth from environmental rewards. These simple economic signals induce decentralized credit assignment, driving planning without global orchestration or explicit communication protocols. The population evolves through economic selection: effective agents accumulate wealth and are mutated via exploitation, while ineffective ones go bankrupt and are replaced via exploration. We show that, initialized with weak agents, the economy produces emergent multi-step reasoning strategies and outperforms stronger monolithic baselines across five agentic tasks, including mathematical reasoning, financial research, scientific research, accelerator design, and distributed-system optimization. We further provide theoretical insights into how economic dynamics shape agent behaviors, linking local incentives to long-term global performance. Our results suggest a new path to multi-agent intelligence: rather than engineering coordination, we can design decentralized incentive structures under which it automatically emerges.

cs.CL

Can LLMs extract scientific consensus? A case study in high-temperature superconductivity

Scientific knowledge is increasingly dispersed across vast and heterogeneous scientific literature, where important claims are often implicit, evolving, and internally debated. While large language models (LLMs) have shown impressive performance in information extraction and summarization, their ability to recover latent scientific consensus remains unclear. Here, we investigate this problem in the context of high-temperature superconductivity (HTS), a long-standing and highly debated topic in condensed matter physics, as a challenging testbed. Using near 18,000 highly-cited publications over the past seven decades, we construct a structured knowledge graph linking competing superconducting mechanisms, material families, evidential modalities, and citation relations. We find that LLM-extracted representations recover coherent and physically interpretable structures, including family-dependent mechanism profiles, evidence-specific correlations, and citation-mediated temporal evolution of scientific beliefs. Ablation studies on LLM further show that the global structure remains robust across prompting, decoding, and model variations. Our results suggest that LLMs can indeed serve as scalable tools for deciphering scientific knowledge in domains characterized by competing interpretations and evolving knowledge.

cs.DL

Harnessing AtomisticSkills for Agentic Atomistic Research

Computational materials science and chemistry span vast knowledge domains and fractured software ecosystems. Although large language models (LLMs) have demonstrated research capabilities, scaling monolithic agents to manage the rigor and complexity of atomistic research remains a challenge. Here, we introduce AtomisticSkills, an open-source harness framework that empowers general-purpose AI coding agents to conduct atomistic research across materials science, chemistry, and drug discovery. By hierarchically decomposing scientific workflows into agent skills and tools, AtomisticSkills provides agents with modular, extensible, and plug-and-play research capabilities. The framework integrates more than 100 human-curated multidisciplinary skills, including database access, thermodynamics and kinetics modeling, and diverse simulation engines employing machine learning interatomic potentials (MLIPs) and density functional theory (DFT). We validate its functional coverage against scientific literature and demonstrate robust orchestration capabilities across diverse scientific campaigns: generative design of Li-ion solid-state electrolytes, high-throughput screening of metal-organic frameworks for CO2 capture, autonomous MLIP benchmarking and fine-tuning, multi-stage structure-based virtual screening for drug design, multimodal X-ray diffraction pattern analysis, and screening of Fe-oxide catalysts for oxygen evolution reaction. AtomisticSkills provides a critical agent infrastructure towards building fully autonomous AI scientists.

physics.chem-ph

D-PACE: Dynamic Position-Aware Cross-Entropy for Parallel Speculative Drafting

Speculative decoding accelerates LLM inference by having a small drafter propose tokens that a larger target model verifies in parallel. Recent diffusion-based parallel drafters such as DFlash predict the full B-token block in one forward pass, enabling deeper drafters and longer accepted blocks. However, existing multi-token drafter objectives often use fixed position-dependent weighting schedules, such as head-dependent weights or block-position decays, which do not adapt as the positions limiting acceptance change during training. To address this, we derive per-position training weights from a differentiable surrogate of expected accepted draft length, matching the weight of each position to its log-probability gradient contribution. The resulting loss, D-PACE (Dynamic Position-Aware Cross-Entropy), shifts training signal toward positions that currently limit acceptance as the drafter improves. Across six benchmarks, two Qwen3-4B draft depths, two decoding temperatures, and two additional target models, D-PACE consistently improves both wall-clock speedup and average emitted length, with 2.3\% measured training-time overhead and no changes to the drafter architecture or inference procedure.

cs.LG

Multi-Persona Debate System for Automated Scientific Hypothesis Generation

Modern scientific discovery is bottlenecked not by data scarcity, but by the inability to synthesize fragmented knowledge into actionable hypotheses. This challenge is especially acute in battery materials research, where electrochemical performance, interfacial behavior, and manufacturing feasibility must be optimized simultaneously. Here, we present the Multi-Persona Debate System (MPDS), a literature-grounded framework for automated scientific hypothesis generation that combines literature retrieval, long-context large language model reasoning, corpus-driven persona induction, and structured multi-agent debate. MPDS constructs literature snapshots of up to 500 papers, grounds agents in role-specific evidence pools, and conducts a three-round citation-aware debate followed by moderator synthesis, enabling negotiation between personas while preserving evidence traceability. We evaluate MPDS using a temporally controlled protocol excluding direct access to target papers, including two held-out battery-materials case studies and a blinded comparison across 30 matched cases. In sodium-ion anode and all-solid-state battery cathode design tasks, MPDS recovered design logics aligned with experimentally validated solution spaces and generated more mechanistically explicit, process-aware proposals than simpler baselines. To assess the impact of personas and debate, we introduce Integrative Hypothesis Quality scoring. In ablation studies, MPDS achieved the highest mean score among five conditions, with its largest advantage in cross-perspective integration. A laboratory follow-up suggests utility as a diagnostic aid for identifying practical bottlenecks in workflows. These results indicate that structured debate over literature snapshots improves hypothesis formation under coupled engineering constraints and provides a reusable workflow for text-intensive scientific discovery.

cs.CL

Matlantis-PFP v8: Universal Machine Learning Interatomic Potential with Better Experimental Agreements via r2SCAN Functional

Universal Machine Learning Interatomic Potentials (uMLIPs) enable atomistic simulations and high-throughput screening at scales far beyond those accessible with density functional theory (DFT). However, most existing uMLIPs are trained on Perdew--Burke--Ernzerhof (PBE) generalized gradient approximation (GGA) data and are therefore fundamentally limited by PBE-level accuracy. In this paper, we argue that better zero-shot predictions versus experiments must be an explicit design target for uMLIPs and present PFP v8, a uMLIP available on the Matlantis service that overcomes the inherent limitations of the PBE functional by being trained to reproduce the regularized-restored strongly constrained and appropriately normed (r2SCAN) meta-GGA potential-energy surface across a wide range of chemical domains. Without requiring domain-specific fine-tuning, PFP v8 delivers systematically improved agreement with experimental data or high-accuracy references for crystals, molecules, and surfaces, outperforming PBE-based DFT calculations. Crucially, in long-time molecular dynamics simulations that are computationally impractical with DFT, PFP v8 predicts melting points with an average error of approximately 130 K, halving the error relative to PBE-trained models. These results establish that uMLIPs can move beyond the limitations of their training approximations and achieve substantially improved agreement with experiment across diverse chemical domains, further narrowing the gap between simulation and reality.

physics.chem-ph

Exotic Cooperative Quantum Optics of Moire Exciton Superlattices

The unique properties of two-dimensional moire systems have been widely studied from many perspectives. However, relatively little work has explored how the real space structure of the moire systems can directly engender novel properties and functionalities. In this work, we exploit the feature that moire excitons naturally form an ordered superlattice with a lattice constant comparable to the wavelength of the resonant light, which enables intriguing cooperative optical responses. Particularly, we show that the collective moire exciton states can have either strongly enhanced (superradiant) or suppressed (subradiant) radiative decay rate, depending on their in-plane wavevector. These super- and subradiant states can be efficiently switched by a gate-induced electric field gradient. Moreover, the cooperative transmittance $T$ of the nanometer-thick moire system can be switched from $T \approx 0$ (opaque) to $T \approx 1$ (transparent) with less than $2~\%$ heterostrain or a $1^{\circ}$ adjustment in the twist angle $\theta$. These features are robust against non-radiative losses and inhomogeneity, making the moire system a highly versatile platform for cooperative quantum optics with potential applications in e.g., single photon storage and switching.

cond-mat.mtrl-sci

Spectral Sampling of Boron Diffusion in Ni Alloys: Cr and Mo Effects on Bulk and Grain Boundary Transport

Understanding how light interstitials migrate in chemically complex alloys is essential for predicting defect dynamics and long-term stability. Here, we introduce a spectral sampling framework to quantify boron diffusion activation energies in Ni and demonstrate how substitutional solutes (Cr, Mo) reshape interstitial point defect transport in both the bulk and along crystallographic defects. In the bulk, boron migration energy distributions exhibit distinct modality tied to solute identity and spatial arrangement: both Cr and Mo raise barriers in symmetric cages but induce directional asymmetry in partially decorated environments. Extending this framework to a $\Sigma5\langle100\rangle{210}$ symmetric tilt grain boundary reveals solute-specific confinement effects. Cr preserves low-barrier in-plane mobility while suppressing out-of-plane transport, guiding boron into favorable midplane voids. Mo, by contrast, imposes an across-the-board reduction in boron mobility, suppressing average diffusivity by two additional orders of magnitude at 800 $^\circ$C and reducing out-of-plane transport by five orders of magnitude relative to Cr. Both elements promote segregation by producing negative segregation energies, but their roles diverge: Cr facilitates rapid redistribution and stabilization at interfacial sites, consistent with Cr-rich boride formation, while Mo creates deeper and more uniform segregation wells that strongly anchor boron. Together, these complementary behaviors explain the experimental prevalence of Cr- and Mo-rich borides at grain boundaries and carbide interfaces in Ni-based superalloys. More broadly, we establish spectral sampling as a transferable framework for interpreting diffusion in disordered alloys and for designing dopant strategies that control transport across complex interfaces.

cond-mat.mtrl-sci

Thermal Equilibrium Vacancy Concentration in an Alloy with Chemical Short-Range Order

The equilibrium vacancy concentration in multi-principal element alloys remains a controversial and nontrivial subject, primarily because of chemical complexity and chemical short-range order (CSRO). Here we derive an exact expression that is amenable to atomistic calculations, using multiple perspectives. We applied this expression to equiatomic CrCoNi alloys in the face-centered cubic structure. The derived equilibrium vacancy concentration is used in our recent work, which predicts the chemical short-range order formation timescale consistent with experimental observation. The results demonstrate the practical utility of the approach for predicting equilibrium vacancy concentrations in compositionally complex alloys.

cond-mat.mtrl-sci

Universally Converging Representations of Matter Across Scientific Foundation Models

Machine learning models of vastly different modalities and architectures are being trained to predict the behavior of molecules, materials, and proteins. However, it remains unclear whether they learn similar internal representations of matter. Understanding their latent structure is essential for building scientific foundation models that generalize reliably beyond their training domains. Although representational convergence has been observed in language and vision, its counterpart in the sciences has not been systematically explored. Here, we show that representations learned by nearly sixty scientific models, spanning string-, graph-, 3D atomistic, and protein-based modalities, are highly aligned across a wide range of chemical systems. Models trained on different datasets have highly similar representations of small molecules, and machine learning interatomic potentials converge in representation space as they improve in performance, suggesting that foundation models learn a common underlying representation of physical reality. We then show two distinct regimes of scientific models: on inputs similar to those seen during training, high-performing models align closely and weak models diverge into local sub-optima in representation space; on vastly different structures from those seen during training, nearly all models collapse onto a low-information representation, indicating that today's models remain limited by training data and inductive bias and do not yet encode truly universal structure. Our findings establish representational alignment as a quantitative benchmark for foundation-level generality in scientific models. More broadly, our work can track the emergence of universal representations of matter as models scale, and for selecting and distilling models whose learned representations transfer best across modalities, domains of matter, and scientific tasks.

cs.LG

Thermonuclear Explosions for Large-Scale Carbon Sequestration: A Call for Exploration

Climate change is a rapidly accelerating problem that requires fast and large-scale carbon sequestration to prevent catastrophe. This paper proposes a novel approach to use explosives for large-scale carbon sequestration. Combining the long-practiced method of explosive mining with newer enhanced rock weathering techniques, we propose a faster, greener, and profitable method of large-scale carbon sequestration. This method is applicable for all explosives, including thermonuclear, and can be done safely with minimal anthropological and ecological impact. We estimate a cost of $0.68/ton of CO2 sequestered.

physics.soc-ph

LightPFP: A Lightweight Route to Ab Initio Accuracy at Scale

Atomistic simulation methods have evolved through successive computational levels, each building upon more fundamental approaches: from quantum mechanics to density functional theory (DFT), and subsequently, to machine learning interatomic potentials (MLIPs). While universal MLIPs (u-MLIPs) offer broad transferability, their computational overhead limits large-scale applications. Task-specific MLIPs (ts-MLIPs) achieve superior efficiency but require prohibitively expensive DFT data generation for each material system. In this paper, we propose LightPFP, a data-efficient knowledge distillation framework. Instead of using costly DFT calculations, LightPFP generates a distilled ts-MLIP by leveraging u-MLIP to generate high-quality training data tailored for specific materials and utilizing a pre-trained light-weight MLIP to further enhance data efficiency. Across a broad spectrum of materials, including solid-state electrolytes, high-entropy alloys, and reactive ionic systems, LightPFP delivers three orders of magnitude faster model development than conventional DFT-based methods, while maintaining accuracy on par with first-principles predictions. Moreover, the distilled ts-MLIPs further sustain the computational efficiency essential for large-scale molecular dynamics, achieving 1-2 orders of magnitude faster inference than u-MLIPs. The framework further enables efficient precision transfer learning, where systematic errors from the u-MLIP can be corrected using as few as 10 high-accuracy DFT data points, as demonstrated for MgO melting point prediction. This u-MLIP-driven distillation approach enables rapid development of high-fidelity, efficient MLIPs for materials science applications.

cond-mat.mtrl-sci

Zero-field identification and control of hydrogen-related electron-nuclear spin registers in diamond

Spin defects in diamond serve as powerful building blocks for quantum technologies, especially for applications in quantum sensing and quantum networking. Electron-nuclear defects formed in the environment of optically active spins, such as the nitrogen-vacancy (NV) center, provide a resource for multi-qubit quantum registers. However, many of these defects have yet to be characterized, limiting their control and integration in quantum devices. Here, we apply two hybrid electron-nuclear spin control schemes to self-consistently characterize unknown spin defects at the single-spin level. We perform double electron-electron resonance at zero field (ZF-DEER) to extract hyperfine components and introduce a nuclear-electron-electron triple resonance (NEETR) protocol to control and identify the nuclear spin through the stronger electronic spin interaction. These results provide a guide to resolving the defect structures using ab initio calculations, leading to the identification of a new hydrogen-related defect structure as well as an accurate match to a previously identified nitrogen-related defect. We further apply our NEETR protocol to demonstrate initialization, unitary control, and long-lived coherence of the hydrogen nuclear spin qubit with $T_2 = 1.0(3)\,\mathrm{ms}$. Together, these characterization and control tools establish a framework to harness previously unknown electron-nuclear defects for quantum register applications.

quant-ph

Atomistic mechanisms of oxidation and chlorine corrosion in Ni-based superalloys: The role of boron and light interstitial segregation

Hybrid Monte Carlo and molecular dynamics simulations were used to investigate the interaction of light interstitials in multi-element Ni-based alloys. We show that light interstitials such as boron and oxygen fundamentally alter interfacial chemistry by reshaping alloy-element distribution and segregation. Oxygen adsorption drove boron migration from the grain boundary to the free surface, where it co-enriched with Cr, Fe, and Mo and formed BO3 trigonal motifs embedded within mixed-metal oxide networks. Oxygen also promoted M-O-M chain formation, including Nb2O5 clusters at the free surface. In the absence of oxygen, boron segregated to the grain boundary, altering local metal chemistry and underscoring a dynamic, environment-sensitive behavior. Following chlorine exposure, the oxidized surfaces retained strong O-mediated connectivity while forming new Cl-M associations, particularly with Nb and Cr, and exhibited further surface enrichment in Cr, Fe, and Mo. High-temperature MD simulations revealed a dynamic tug-of-war: chlorine exerted upward pull and disrupted weakly anchored sites, while Nb- and BO3-rich oxide motifs resisted deformation. A new stabilization mechanism was identified in which subsurface boron atoms anchored overlying Cr centers, suppressing their mobility and mitigating chlorine-driven displacement. These results demonstrate boron's dual role as a modifier of alloy-element segregation and a stabilizer of oxide networks, and identify Nb as a key element in reinforcing cohesion under halogen attack. More broadly, this study highlights the need to track light interstitial cross-talk and solute migration under reactive conditions, offering atomistic criteria for designing corrosion-resistant surface chemistries in Ni-based superalloys exposed to halogenated or oxidative environments.

cond-mat.mtrl-sci