arXiv ScienceSearch

arXiv subjects

Mingda Li

Publications and source records attributed to Mingda Li.

At least 19 recordsLinked to original sources

JANUS: A Multi-modal Foundation Neural Sampler for Disordered Materials

Many problems in disordered materials require sampling beyond fixed composition and volume, where coupled changes in atomic identities and structure create a prohibitively expensive discrete-continuous sampling problem. Here we introduce JANUS, a multimodal neural sampler that couples continuous and masked discrete diffusion through an equivariant graph neural network trained directly from energy evaluations, without pre-generated equilibrium data. In benchmark Ising and isobaric $\Delta\mu NPT$ alloy systems, JANUS reproduces reference Monte Carlo equilibrium observables and recovers free energies and phase behavior with more than three orders of magnitude fewer energy evaluations. In multicomponent alloys, JANUS enables conditional steering toward prescribed chemical short-range order and enhanced bulk modulus and, when coupled to a large language model evolutionary agent, performs efficient inverse design for balanced optical and mechanical properties. In semiconductors like silicon and diamond, JANUS explores vacancies and dopants spanning 15 elements in grand-canonical $\mu VT$ ensembles, recovers established defects including the silicon $E$ centre, and identifies new candidate defect pairs and triplets for quantum engineering, including S-Ti in silicon and B-O-O in diamond, with deep in-gap states validated by hybrid-functional density functional theory. By unifying discrete site identities with continuous structural and volumetric relaxation, JANUS provides a foundation for thermodynamic sampling, characterization and inverse design of chemically disordered materials.

cond-mat.mtrl-sci

MARCO: Click-Intent Decomposition for Calibrated Ads Conversion Prediction

Not all clicks are equal. Industrial ads ranking decouples conversion probability into click-through rate (CTR) and post-click conversion rate (CVR), yet treats every click as the same event. In reality, users provide a free, self-generated signal of intent through their physical UI interactions. Different click types on the same ad exhibit a 4-fold difference in actual conversion rates. By conflating these signals, the standard CVR model under-predicts high-intent clicks and over-predicts low-intent ones, which is a bias masked by near-perfect aggregate calibration. We propose MARCO (Multi-intent Ads Ranking Composition Optimization), a framework that resolves this bias by decomposing each click by intent. Using the logged click type as a free behavioral label, MARCO trains per-intent CVR heads on homogeneous populations, and at serving time composes their per-intent CVR estimates under a predicted distribution over intents. Theoretically, we prove that decomposition never raises population risk, give the exact headroom under squared loss and non-negativity under the deployed loss, and show through a routing-efficiency dial how much of it reaches serving. Because the population-optimal score is unchanged, any gain is a finite-capacity estimation and calibration effect that we validated both offline and online. For deployment at scale, we further cast multi-impression, multi-click attribution as credit assignment with a bias-variance tradeoff analogous to RL return estimation, showing last-impression, first-click attribution is the low-bias, low-variance, deterministic choice under production constraints, and derive three consistency conditions enforced end-to-end at scale. Deployed at binary intent granularity, MARCO corrects per-intent calibration to approximately 100%, lifts conversions per click by +2.80%, and drives +0.98% cumulative improvement in topline metrics.

cs.LG

LUT: Latent Utility Training for Visual Reasoning

Multimodal large language models have advanced visual understanding, yet perception-intensive reasoning remains challenging. Recent latent visual reasoning methods introduce hidden-space computation before answering, but they often rely on costly intermediate supervision, such as bounding boxes, sketches, or interleaved rationales. These strategies focus on how latent states should be shaped, but do not explicitly assess whether the latent is useful for the final answer. We propose LUT, a latent reasoning framework trained with only standard VQA pairs. LUT centers training on Latent Utility at two levels. At the trajectory level, we propose Utility-Aware Latent Distillation SFT, which explores answer-relevant latent trajectories, selects qualified trajectories by their information gain, and distills more reliable and learnable supervision through curriculum learning. At the step level, we propose Latent Attribution Policy Optimization, which uses answer-to-latent attribution to differentially optimize latent steps during reinforcement learning. Experiments on perception-intensive visual reasoning benchmarks show that LUT outperforms previous latent reasoning methods and remains competitive with latent-text interleaved methods with lower annotation cost.

cs.CV

Agentic AI for Scientific Reasoning in Autonomous Quantum Sensing Experiments

We implement an agentic AI workflow built around a large language model (LLM) agent for autonomous experiments with nitrogen-vacancy (NV) centers in diamond. NV centers are a widely used platform for quantum sensing, and the ability to control many measurements from a computer makes NV experiments a natural setting for autonomous workflows. We make two main contributions. First, we demonstrate an autonomous NV experiment workflow that combines persistent project records, quantitative calculation and data analysis tools, and deterministic experiment control. In one autonomous experiment, the agent selected a single NV center, calibrated its resonant frequency, measured \(T_2^\ast\) with Ramsey measurements, and added a Carr--Purcell--Meiboom--Gill (CPMG) measurement to check a weak feature that could be related to nearby \(^{13}\mathrm{C}\). Second, we introduce two offline benchmarks that evaluate the agent's reasoning separately from laboratory execution. We evaluated both benchmarks with GPT-5.4, GPT-5.5, and GPT-5.6 Sol. In the Ramsey checkpoint benchmark, greater reasoning effort generally improved recognition of a residual resonance calibration offset. By contrast, in the pulsed optically detected magnetic resonance (pODMR) data evaluation benchmark, pulse sequence information alone produced more false positive resonance judgments at higher reasoning effort. Requiring an expected signal calculation kept false positive rates low across all three models and reasoning settings. The results suggest a clear division of labor for autonomous experiments. The agent forms scientific hypotheses and uses quantitative tools to evaluate data, while deterministic code controls the hardware and enforces safety constraints.

quant-ph

ATLAS: A Foundation Neural Sampler for Amorphous Materials

Amorphous materials exhibit exceptional mechanical and functional properties, yet their rugged energy landscapes are notoriously difficult to sample. Below the glass-transition temperature, conventional molecular dynamics and Monte Carlo become inefficient because equilibration relies on rare barrier-crossing events, while data-driven generative models are constrained by scarce and biased reference ensembles. Here, we introduce ATLAS, an efficient sampler that learns a diffusion process to generate Boltzmann-distributed amorphous structures directly from a target energy function. Parameterized by an equivariant graph neural network, ATLAS generalizes across system size, temperature, and composition. By exploiting the time reversal of the diffusion process, it enables efficient estimation of thermodynamic quantities and steering toward target observables. In two-dimensional Kob-Andersen systems, ATLAS reproduces parallel tempering Markov chain Monte Carlo structural distributions, free energies and entropies, achieving below 0.2% free energy error in the low-temperature glass regime with over 500-fold fewer energy evaluations. In Cu-Zr and Cr-Co-Ni metallic glasses, ATLAS recovers experimentally observed short-range-order trends and steers structures toward prescribed order parameters and optimized bulk moduli. Moreover, composition-amortized pretraining outperforms composition-specific training from scratch, reduces inverse-design costs by several hundred-fold, and enables sampling with expensive universal machine learning interatomic potentials. Coupled to a large language model agent, ATLAS searches an eight-element space for high-entropy metallic glasses balancing stiffness and ductility, identifying a converged Pareto frontier within 480 oracle evaluations. Together, these results establish ATLAS as a foundation model for sampling, steering and designing amorphous materials.

cond-mat.mtrl-sci

Order of Magnitude Improved Optical Trapping of Molecules Through Transverse Cooling

We demonstrate a two-dimensional Sisyphus laser cooling method that increases the number of strontium monohydroxide (SrOH) molecules loaded into a magneto-optical trap by a factor of 12. Subsequent loading into an optical dipole trap (ODT) achieves $2.2 (3)\times10^4$ ultracold SrOH molecules with a peak density of $\sim2(1)\times10^{10}~\mathrm{cm^{-3}}$. The lifetime of molecules in the ODT is limited by two-body collisions characterized by a measured collision rate constant $\beta \sim 4\times10^{-10}~\mathrm{cm^3/s}$. The cooling method developed here is generally applicable to all known cases of direct molecular laser cooling, including symmetric and asymmetric top molecules. Increases in trapped molecule number will directly improve the search for ultralight dark matter, position polyatomic molecules as a platform for probing CP-violating new particles with masses $\gg$10 TeV, and facilitate a broad range of further research in quantum science.

physics.atom-ph

Tensor-Network Finite Elements for Analytic Operator Equations

Operator equations (OEs) underpin quantitative modeling across science and engineering. Finite-element (FE) methods discretize continuous OEs into finite-dimensional algebraic systems, whereas tensor networks (TNs) provide flexible variational representations of correlated discrete systems. Here, we develop a framework that connects FE with TN for analytic OEs. The power of this method comes from its ability to convert highly non-linear partial differential equations into linear matrix equations. In particular, we show that FE discretization induces a hierarchy of multilinear interaction tensors, through which differential, integral, nonlinear, memory, and delay equations can be expressed within a common algebraic structure. The resulting systems are reformulated as weighted-residual optimization problems over TN degrees of freedom. Matrix-product-state calculations for one-dimensional linear and nonlinear diffusion reproduce conventional solutions with controlled error while preserving continuity and Neumann boundary conditions. The framework provides a common variational language for analytic OEs and establishes a direct connection between FE numerical formalism and TN variational algorithms, offering a general foundation for TN-based and quantum-inspired approaches to solving OEs.

math.NA

XRDiff: Crystal Structure Prediction from Powder X-Ray Diffraction Data Using Diffusion Models

Determining the crystal structure of a material from its powder X-ray diffraction (PXRD) pattern is a central challenge in materials science. PXRD is an accessible and widely used characterization technique, yet recovering the atomic structure from diffraction data requires solving an underdetermined inverse problem due to the loss of phase information. Generative modeling can provide a prior over atomic structure and learn the mapping from PXRD patterns to crystal structures via simulated structure-spectrum pairs. We present XRDiff, a diffusion model that recovers crystal structures from PXRD given either the stoichiometry or, in a more challenging setting, the elemental constituents and total number of atoms in the unit cell. We evaluate on datasets where each stoichiometry has multiple polymorphs and all polymorphs of a given composition are held out together, ensuring that high performance reflects genuine use of the diffraction signal. XRDiff achieves strong structure recovery rates on simulated benchmarks, indicating that the model learns a spectrum-to-structure mapping precise enough to differentiate between polymorphs. To address generalization to experimental data, we compare a full-spectrum encoding against an encoding based on peak descriptors. The peak-based encoding generalizes substantially better, outperforming even a model trained on full spectra with augmentations fitted to the experimental noise distribution. These results demonstrate that representations robust to the noise and artifacts present in real-world PXRD offer a practical and scalable path toward closing the simulation-to-experiment gap, enabling zero-shot crystal structure solution from experimental PXRD with full or partial chemical composition input.

cond-mat.mtrl-sci

Can LLMs extract scientific consensus? A case study in high-temperature superconductivity

Scientific knowledge is increasingly dispersed across vast and heterogeneous scientific literature, where important claims are often implicit, evolving, and internally debated. While large language models (LLMs) have shown impressive performance in information extraction and summarization, their ability to recover latent scientific consensus remains unclear. Here, we investigate this problem in the context of high-temperature superconductivity (HTS), a long-standing and highly debated topic in condensed matter physics, as a challenging testbed. Using near 18,000 highly-cited publications over the past seven decades, we construct a structured knowledge graph linking competing superconducting mechanisms, material families, evidential modalities, and citation relations. We find that LLM-extracted representations recover coherent and physically interpretable structures, including family-dependent mechanism profiles, evidence-specific correlations, and citation-mediated temporal evolution of scientific beliefs. Ablation studies on LLM further show that the global structure remains robust across prompting, decoding, and model variations. Our results suggest that LLMs can indeed serve as scalable tools for deciphering scientific knowledge in domains characterized by competing interpretations and evolving knowledge.

cs.DL

Universal Magnetic Structure Prediction from Atomic Coordinates with Near-Experimental Accuracy

Magnetic order is a fundamental property of materials, governing collective behavior and enabling a broad range of functionalities. Yet magnetic structure remains difficult to determine: experiments are costly and specialized, while first-principles methods often struggle with the noncollinear and incommensurate orders found in real materials. Here we introduce magnetic structure network (MSN), an E(3) equivariant graph neural network that predicts both collinear and non-collinear magnetic structures directly from atomic crystal structures, trained directly on experimentally determined structures from MAGNDATA. By proposing the primitive modulated structure representation (PMSR), we are able to encode commensurate and incommensurate structures in a unified way without symmetry assumptions. The model achieves strong performance across all modulation components and reconstructs experimental magnetic structures with high fidelity. Our approach provides a scalable framework for rapid magnetic structure prediction and opens a route to data-driven discovery of magnetic materials.

cond-mat.mtrl-sci

Probing Non-Equilibrium Grain Boundary Dynamics with XPCS and Domain-Adaptive Machine Learning

Grain-boundary (GB) dynamics control the stability, mechanical, and functional response of nanocrystalline materials, but direct experimental access to their slow non-equilibrium motion has been limited. Here we establish X-ray photon correlation spectroscopy (XPCS), combined with domain-adaptive machine learning, as a quantitative probe of GB dynamics. Temperature- and grain-size-dependent two-time XPCS measurements in nanocrystalline silicon reveal pronounced departures from time-translation invariance, showing that GB relaxation can remain far from equilibrium over experimental timescales. However, direct extraction of quantitative physical information from these high-dimensional, noisy fluctuation maps faces a significant challenge. To overcome this barrier, we develop a semi-supervised learning framework that transfers physical parameter labels from continuum simulations to unlabeled experimental XPCS maps through domain-adaptive representation alignment. This AI-augmented approach enables the extraction of key kinetic parameters, including bulk diffusivity, GB stiffness, and effective GB concentration, directly from experimental XPCS measurements. Our results show how machine learning can transform indirect fluctuation signals into quantitative materials dynamics, providing a general route to study non-equilibrium defect motion in solids.

cond-mat.mtrl-sci

Gradients with Respect to Semantics Preserving Embeddings Tell the Uncertainty of Large Language Models

Uncertainty quantification (UQ) is an important technique for ensuring the trustworthiness of LLMs, given their tendency to hallucinate. Existing state-of-the-art UQ approaches for free-form generation rely heavily on sampling, which incurs high computational cost and variance. In this work, we propose the first gradient-based UQ method for free-form generation, SemGrad, which is sampling-free and computationally efficient. Unlike prior gradient-based methods developed for classification tasks that operates in parameter space, we propose to consider gradients in semantic space. Our method builds on the key intuition that a confident LLM should maintain stable output distributions under semantically equivalent input perturbations. We interpret the stability as the gradients in semantic space and introduce a Semantic Preservation Score (SPS) to identify embeddings that best capture semantics, with respect to which gradients are computed. We further propose HybridGrad, which combines the strengths of SemGrad and parameter gradients. Experiments demonstrate that both of our methods provide efficient and effective uncertainty estimates, achieving superior performance than state-of-the-art methods, particularly in settings with multiple valid responses.

cs.CL

Quantum Theory of Functionally Graded Materials

Functionally graded materials (FGMs) are composites whose composition or microstructure varies continuously in space, producing position-dependent mechanical and functional properties. In recent years, FGMs have gained significant attention due to advances in additive manufacturing, which enable precise spatial control of composition and orientation. However, their graded, aperiodic structure breaks the assumptions of Bloch's theorem, making first-principles electronic and electromagnetic calculations challenging. Here we develop an ab initio quantum theoretical framework for the electromagnetic properties of FGMs. Using a non-interacting electron model, we formulate a theory of modulated Bloch states, derive effective field equations, and solve them by proposing a generalized WKB (GWKB) method, an effective mass approximation, the Boltzmann equation, and numerical approaches. Our GWKB solution is not semiclassical but remains valid in the fully quantum regime. We show that effective observables such as conductivity, magnetic permeability, and electric permittivity generally do not admit a tensorial description in graded media, and that engineered orientational gradients enable precise control of Landau quantization. As a device example, we further develop a theory of graded p-n junctions with enhanced electronic tunability. This framework lays the quantum foundation for predictive design of graded composite materials, enabling AI-accelerated discovery of next-generation functional architectures.

cond-mat.mtrl-sci

Large Transverse Thermoelectric Effect in Weyl Semimetal TaIrTe$_4$ Engineered for Photodetection

Anomalous local photocurrent generation via second-order nonlinear and thermoelectric responses is a signature of many topological semimetals. The emergence of these photocurrents is inherently linked to symmetry breaking and anisotropy of their crystal lattices. Studies of type-II Weyl semimetals of group C$_{2v}$ (WTe$_2$, MoTe$_2$, TaIrTe$_4$) have reported anomalous, nonlocal photocurrents localized to crystals edges or far from electrodes, which are highly dependent on the geometry of the material sample. While originally attributed to a nonlinear charge current response, it was recently shown that these currents could instead be attributed to the anisotropic Seebeck coefficients of the materials. Here, we confirm that anomalous photocurrents observed in TaIrTe$_4$ under either visible or far-infrared far-field illumination originate from the large transverse thermoelectric effect. We engineer the mutual orientation of crystal edges and electrodes as well as the thermal environment of TaIrTe$_4$ to control and amplify its spatial photocurrent response. We show that substrate engineering can locally enhance photocurrent. This framework of thermal device engineering can enable broadband photo detection schemes by leveraging spectral and spatial dependence of photocurrents for applications like wavefront sensing, beam positioning, and edge detection.

cond-mat.mtrl-sci

Yunque DeepResearch Technical Report

Deep research has emerged as a transformative capability for autonomous agents, empowering Large Language Models to navigate complex, open-ended tasks. However, realizing its full potential is hindered by critical limitations, including escalating contextual noise in long-horizon tasks, fragility leading to cascading errors, and a lack of modular extensibility. To address these challenges, we introduce Yunque DeepResearch, a hierarchical, modular, and robust framework. The architecture is characterized by three key components: (1) a centralized Multi-Agent Orchestration System that routes subtasks to an Atomic Capability Pool of tools and specialized sub-agents; (2) a Dynamic Context Management mechanism that structures completed sub-goals into semantic summaries to mitigate information overload; and (3) a proactive Supervisor Module that ensures resilience through active anomaly detection and context pruning. Yunque DeepResearch achieves state-of-the-art performance across a range of agentic deep research benchmarks, including GAIA, BrowseComp, BrowseComp-ZH, and Humanity's Last Exam. We open-source the framework, reproducible implementations, and application cases to empower the community.

cs.CL

Frustrated Magnetism in FeGe$_3$O$_4$ with a Chiral Trillium Network

The discovery of new magnetic ground states in geometrically frustrated lattices remains a central challenge in materials science. Here, we report the synthesis, structural characterization, and frustrated magnetic properties of FeGe$_3$O$_4$, a newly identified compound that crystallizes in the noncentrosymmetric cubic space group $P2_13$. In this structure, Fe atoms form an intricate double-trillium lattice with nearest-neighbor Fe--Fe distances of $\sim$4.2~\AA{}, while Ge$^{2+}$ ions mediate magnetic interactions through Fe-Ge-Fe pathways. Field-dependent magnetization at 2~K shows a pronounced nonlinearity, reaching a maximum moment of 2.55(3)~$\mu_\mathrm{B}$/Fe$^{2+}$ at 70~kOe without evidence of saturation. Magnetic susceptibility, heat capacity, and neutron scattering collectively reveal the onset of short-range magnetic interactions near 5~K, with no long-range ordering detected down to 0.06~K. Specific heat measurements demonstrate strong frustration: only $\sim$34\% of the expected magnetic entropy is recovered at 2.4~K. Taken together, these results establish FeGe$_3$O$_4$ as a rare example of a geometrically frustrated trillium-lattice magnet, offering a promising platform for exploring exotic quantum magnetic phenomena.

cond-mat.str-el

The Loss Landscape of Powder X-Ray Diffraction-Based Structure Optimization Is Too Rough for Gradient Descent

Solving crystal structures from powder X-ray diffraction (XRD) is a central challenge in materials characterization. In this work, we study the powder XRD-to-structure mapping using gradient descent optimization, with the goal of recovering the correct structure from moderately distorted initial states based solely on XRD similarity. We show that commonly used XRD similarity metrics result in a highly non-convex landscape, complicating direct optimization. Constraining the optimization to the ground-truth crystal family significantly improves recovery, yielding higher match rates and increased mutual information and correlation scores between structural similarity and XRD similarity. Nevertheless, the landscape may remain non-convex along certain symmetry axes. These findings suggest that symmetry-aware inductive biases could play a meaningful role in helping learning models navigate the inverse mapping from diffraction to structure.

cond-mat.mtrl-sci

Chimeric states of matter: Meissner effect without superconductivity

Symmetry is central to how we classify phases of matter: solids break spatial translations, superfluids break particle-number conservation, and superconductors "break" gauge symmetry. Mixed anomalies involving higher-form symmetries, however, present a generalization of spontaneous symmetry breaking that admits a wider and more versatile set of possibilities. We introduce chimeric states of matter, in which aspects of broken and unbroken phases coexist. We find that the Meissner effect -- usually regarded as the defining hallmark of superconductivity -- can occur in media that are resistive or even insulating when probed by electric fields. We demonstrate this by constructing an effective field theory of "symmetry chimerization" and propose that Josephson junction networks could provide a laboratory realization. These results broaden the landscape of possible phases of matter, showing that physical media can mix features of symmetry-restored and symmetry-broken states in a single substrate.

cond-mat.supr-con