arXiv ScienceSearch

arXiv subjects

Xiaogang Li

Publications and source records attributed to Xiaogang Li.

At least 19 recordsLinked to original sources

BilliardPhys-Bench: Benchmarking Physical Reasoning and Visual Dynamics of Multimodal LLMs

Current multimodal models handle static image recognition well, but intuitive physical reasoning remains a weakness. Predicting how objects will move and interact from a single image is still difficult for these systems. We present BilliardPhys-Bench, a benchmark for physical reasoning in synthetic billiards environments. Its procedural engine generates randomized scenarios with friction and elastic collisions. The benchmark tests three abilities: (1) predicting ball-to-ball collisions, (2) reasoning about wall bounces, and (3) estimating final ball positions after motion stops. We evaluate recent MLLMs from the GPT, Claude, Gemini, and Qwen families. Performance drops as simulation time increases and scene geometry grows more complex. We also observe a consistent failure mode we call "stasis bias": when the correct physical outcome is harder to infer, models tend to predict no interaction. These findings show where current MLLMs break down on visual dynamics and point toward the need for better physical inductive biases in multimodal architectures.

cs.AI

CrystalXRD-Bench: Benchmarking Vision-Language Models for XRD Peak Indexing Across Diverse Crystalline Materials

Miller-index identification from powder XRD patterns requires capabilities untested by existing multimodal benchmarks: the model must read a narrow peak location from a rendered scientific curve and then connect that observation to multi-step crystallographic reasoning. We introduce CrystalXRD-Bench, a 250-sample benchmark built from 10 public crystallographic databases for a single task: recover the full set of HKLs contributing to the highest-intensity peak in an XRD pattern. Each sample pairs the rendered XRD image with the source CIF text and chemical formula, so visual extraction errors and reasoning errors can be examined side by side. We evaluate seven vision-language models. The best Jaccard score is 0.5888 (GPT-5.4) with an exact-match rate of 37.6%, yet six of seven models remain below Jaccard 0.50; the task is far from solved. Error patterns vary systematically: double-peak cases are especially brittle, recall-heavy models gain coverage by over-predicting HKLs, and access to CIF text does not close the gap in crystallographic calculation. Alongside model rankings, the benchmark identifies the conditions under which current VLMs fail on quantitative scientific figures. All data and evaluation code will be publicly available.

cs.AI

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning

Current multimodal benchmarks for scientific reasoning primarily evaluate local information extraction -- models recognize symbols and values and then perform textual inference. They do not assess whether models can reason over the global structural properties of formal diagrams, such as topology, conservation constraints, and the consistent mapping between visual patterns and algebraic expressions. We introduce FeynmanBench, a benchmark of over 2,000 tasks centered on Feynman diagrams spanning the electromagnetic, weak, and strong interactions of the Standard Model. Each instance couples a diagram image with minimal textual conventions and requires models to recover the full physical content -- vertex inventory, propagator types, topological connectivity, momentum routing, and the complete scattering amplitude. An automated generation and verification pipeline produces the diagrams, annotations, and reference answers under standardized rules. Evaluating 19 state-of-the-art multimodal LLMs, we find a consistent failure pattern: models achieve 70--95\% on local recognition (vertex and propagator identification) but collapse to 13--17\% on topological reconstruction (CP3), and near zero on full algebraic derivation (CP5). FeynmanBench offers a controlled testbed for multimodal reasoning over formal scientific diagrams and highlights fundamental limitations of current architectures in topology-sensitive scientific reasoning.

cs.AI

SPM-Bench: Benchmarking Large Language Models for Scanning Probe Microscopy

As LLMs achieved breakthroughs in general reasoning, their proficiency in specialized scientific domains reveals pronounced gaps in existing benchmarks due to data contamination, insufficient complexity, and prohibitive human labor costs. Here we present SPM-Bench, an original, PhD-level multimodal benchmark specifically designed for scanning probe microscopy (SPM). We propose a fully automated data synthesis pipeline that ensures both high authority and low-cost. By employing Anchor-Gated Sieve (AGS) technology, we efficiently extract high-value image-text pairs from arXiv and journal papers published between 2023 and 2025. Through a hybrid cloud-local architecture where VLMs return only spatial coordinates "llbox" for local high-fidelity cropping, our pipeline achieves extreme token savings while maintaining high dataset purity. To accurately and objectively evaluate the performance of the LLMs, we introduce the Strict Imperfection Penalty F1 (SIP-F1) score. This metric not only establishes a rigorous capability hierarchy but also, for the first time, quantifies model "personalities" (Conservative, Aggressive, Gambler, or Wise). By correlating these results with model-reported confidence and perceived difficulty, we expose the true reasoning boundaries of current AI in complex physical scenarios. These insights establish SPM-Bench as a generalizable paradigm for automated scientific data synthesis.

cs.AI

HLE-Verified: A Systematic Verification and Structured Revision of Humanity's Last Exam

Humanity's Last Exam (HLE) has become a widely used benchmark for evaluating frontier large language models on challenging, multi-domain questions. However, community-led analyses have raised concerns that HLE contains a non-trivial number of noisy items, which can bias evaluation results and distort cross-model comparisons. To address this challenge, we introduce HLE-Verified, a verified and revised version of HLE with a transparent verification protocol and fine-grained error taxonomy. Our construction follows a two-stage validation-and-repair workflow resulting in a certified benchmark. In Stage I, each item undergoes binary validation of the problem and final answer through domain-expert review and model-based cross-checks, yielding 668 verified items. In Stage II, flawed but fixable items are revised under strict constraints preserving the original evaluation intent, through dual independent expert repairs, model-assisted auditing, and final adjudication, resulting in 1,143 revised-and-certified items. The remaining 689 items are released as a documented uncertain set with explicit uncertainty sources and expertise tags for future refinement. We evaluate eight state-of-the-art language models on HLE and HLE-Verified, observing an average absolute accuracy gain of 7--10 percentage points on HLE-Verified. The improvement is particularly pronounced on items where the original problem statement and/or reference answer is erroneous, with gains of 30--40 percentage points. Our analyses further reveal a strong association between model confidence and the presence of errors in the problem statement or reference answer, supporting the effectiveness of our revisions. Overall, HLE-Verified improves HLE-style evaluations by reducing annotation noise and enabling more faithful measurement of model capabilities. Data is available at: https://huggingface.co/datasets/skylenage/HLE-Verified

cs.CL

Permutation groups and symmetric Hecke algebras

The endomorphism algebras of the permutation modules for transitive permutation groups, known as Hecke algebras, are fundamental objects in representation theory. While group algebras are known to be symmetric over any field, it is natural to ask whether this property extends to Hecke algebras. To study this, we introduce the new concepts of $p$-$S$-permutation groups (for a prime $p$) and $S$-permutation groups. A \emph{ $p$-$S$-permutation group} is a transitive permutation group whose associated Hecke algebra is symmetric over every field of characteristic $p$. An \emph{ $S$-permutation group} is a transitive permutation group that is a $p$-$S$-permutation group for all primes $p$. In this paper, we study Hecke algebras from a group-theoretical perspective and we show that several classes of permutation groups are $p$-$S$-permutation groups and $S$-permutation groups in our sense. This result represents a substantial extension of earlier work by Li and He. (Transform Groups, 30(4), 2025), and reframes the question of determining when the algebra \(\End_{KG}(K\Omega)\) is symmetric within a more general theoretical framework.

math.RT

A classification of regular maps with Euler characteristic $-pq$

In this paper, we give a classification of regular maps with Euler characteristic $-pq$ for distinct primes $q>p\geq 5$. This together with previous classification of regular maps with Euler characteristic $-2p,-3p$ and $-p^2$ completes the classification of regular maps with Euler characteristic $-pq$ for two primes $p$ and $q$. An interesting consequence is that, for every pair of twin primes $p$ and $q$ greater than $5$, there exist three regular maps with solvable automorphism groups and Euler characteristic $-pq$, up to duality and isomorphism.

math.GR

A classification of regular maps with Euler characteristic $-p^4$ for a prime $p\geq 5$

A map is a cellular decomposition of a closed surface. In the framework of classifying all regular maps by their supporting surface, it is an open problem to find all closed surfaces that support no regular maps. Classification of regular maps on surfaces with Euler characteristic $-p, -p^2, -p^3, -2p,$ and $-3p$ has already been done by several authors in a series of papers, which also show that surfaces with these Euler characteristic support no regular maps if the corresponding prime $p$ satisfies certain conditions. In this paper, assuming that $p\geq 5$ is a prime and $i\geq 4$, we show that the order of a Sylow $p$-subgroup of a regular map with Euler characteristic $-p^i$ is bounded by $p^{i-1}$ unless $p\in \{5, 7, 13\}$, and we show the existence of a normal $p$-subgroup for these regular maps whenever a Sylow $p$-subgroup has order at least $\sqrt{p^i}$, laying a solid foundation for using an inductive method to completely characterize regular maps of Euler characteristic $-p^i$. Based on this, we classify all regular maps with Euler characteristic $-p^4$ for a prime $p\geq 5$ in terms of reduced presentations of their automorphism groups. Consequently, a closed surface with Euler characteristic $-p^4$ supports no regular maps if and only if $p\notin \{2,3,5,7,13\}$.

math.GR

Stable equivalences and homological dimensions

As is known, every finite-dimensional algebra over a field is isomorphic to the centralizer algebra of \textbf{two} matrices. So it is fundamental to study first the centralizer algebra of a single matrix, called a centralizer matrix algebra. In this article, stable equivalences between centralizer matrix algebras over arbitrary fields are completely characterized in terms of a new type of equivalence relation on matrices. Moreover, stable equivalences of centralizer matrix algebras over any fields induce stable equivalences of Morita type, thus preserve dominant, finitistic and global dimensions. Our methods also show that the Alperin--Auslander/Auslander--Reiten conjecture holds true for stable equivalences between an arbitrary algebra and a centralizer matrix algebra over a common field.

math.RT

Derived equivalences, new matrix equivalences, and homological conjectures

Based on the fact that every finite-dimensional algebra over a field is isomorphic to the centralizer of \textbf{two} matrices, we approach the representation theory of finite-dimensional algebras over fields by centralizers of matrices. The first fundamental question is to study the centralizer of a single matrix, called a centralizer matrix algebra. By introducing three new equivalence relations on all square matrices over a field, we completely characterize Morita, derived and almost $\nu$-stable derived equivalences between centralizer matrix algebras in terms of these matrix equivalences, respectively. Further, we show that a derived equivalence between centralizer matrix algebras of permutation matrices induces both a Morita equivalence and additional derived equivalences for $p$-regular parts and for $p$-singular parts. As an application, we show that the finitistic dimension conjecture and the Nakayama conjecture are valid for centralizer matrix algebras.

math.RT

Dynamics Simulation of Arbitrary Non-Hermitian Systems Based on Quantum Monte Carlo

Non-Hermitian quantum systems exhibit unique properties and hold significant promise for diverse applications, yet their dynamical simulation poses a particular challenge due to intrinsic openness and non-unitary evolution. Here, we introduce a hybrid classical-quantum algorithm based on Quantum Monte Carlo (QMC) for simulating the dynamics of arbitrary time-dependent non-Hermitian systems. Notably, this approach constitutes a natural extension of the quantum imaginary-time evolution (QITE) algorithm. This algorithm combines the advantages of both classical and quantum computation and exhibits good applicability and adaptability, making it promising for simulating arbitrary non-Hermitian systems such as PT-symmetric systems, non-physical processes, and open quantum systems. To validate the algorithm, we applied it to the dynamic simulation of open quantum systems and achieved the desired results.

quant-ph

A Time-Symmetric Quantum Algorithm for Direct Eigenstate Determination

Time symmetry in quantum mechanics, where the current quantum state is determined jointly by both the past and the future, offers a more comprehensive description of physical phenomena. This symmetry facilitates both forward and backward time evolution, providing a computational advantage over methods that rely on a fixed time direction. In this work, we present a nonvariational and \textit{time-symmetric quantum algorithm} for addressing the eigenvalue problem of the Hamiltonian, leveraging the coherence between forward and backward time evolution. Our approach enables the simultaneous determination of both the ground state and the highest excited state, as well as the direct identification of arbitrary eigenstates of the Hamiltonian. Unlike existing methods, our algorithm eliminates the need for prior computation of lower eigenstates, allowing for the direct extraction of any eigenstate and energy bandwidth while avoiding error accumulation. Its non-variational nature ensures convergence to target states without encountering the barren plateau problem. We demonstrate the feasibility of implementing the non-unitary evolution using both the linear combination of unitaries and quantum Monte Carlo methods. Our algorithm is applied to compute the energy bandwidth and spectrum of various molecular systems, as well as to identify topological states in condensed matter systems, including the Kane-Mele model and the Su-Schrieffer-Heeger model. We anticipate that this algorithm will provide an efficient solution for eigenvalue problems, particularly in distinguishing quantum phases and calculating energy bands.

quant-ph

Quantum fluctuation energies over a spatially inhomogeneous field background in a chiral soliton model

Based on chiral soliton models, the quantum fluctuation energies of quarks over a spatially inhomogeneous meson field background have been thoroughly studied. We have used a systematic calculation scheme initiated by Schwinger, in which the loop quantum fluctuation energies are evaluated by a nontrivial level summation over the eigenvalue spectrum of the effective Hamiltonian of the system. The effective Hamiltonian can be constructed by one loop effective action of fluctuations of quarks over a static chiral soliton field background. The corresponding Dirac equation is obtained. In a static and spatially spherical case and by the hedgehog ansatz the radial part and the angular part of the grand spin of the wave function for the Dirac equation can be separated. Due to the soliton background the eigenvalue spectrum are distorted. The scattering phase shift can be determined by solve the radial equations at different momentum. The density of states in momentum space can be derived. The effective Hamiltonian has been diagonalized in a Hilbert space where the eigenfunctions are labeled by the parity, grand spin and energy. The renormalization scheme can be carried out by a Born subtraction of the phase shift and the compensating Feynman diagram renormalization. Finally the finite quantum fluctuation energies over chiral soliton background at different parities and grand spins have been numerically evaluated, compared and discussed.

nucl-th

Exponentially reduced circuit depths in Lindbladian simulation

Quantum computers can efficiently simulate Lindbladian dynamics, enabling powerful applications in open system simulation, thermal and ground-state preparation, autonomous quantum error correction, dissipative engineering, and more. Despite the abundance of well-established algorithms for closed-system dynamics, simulating open quantum systems on digital quantum computers remains challenging due to the intrinsic requirement for non-unitary operations. Existing methods face a critical trade-off: either relying on resource-intensive multi-qubit operations with experimentally challenging approaches or employing deep quantum circuits to suppress simulation errors using experimentally friendly methods. In this work, we challenge this perceived trade-off by proposing an efficient Lindbladian simulation framework that minimizes circuit depths while remaining experimentally accessible. Based on the incoherent linear combination of superoperators, our method achieves exponential reductions in circuit depth using at most two ancilla qubits and the straightforward Trotter decomposition of the process. Furthermore, our approach extends to simulate time-dependent Lindbladian dynamics, achieving logarithmic dependence on the inverse accuracy for the first time. Rigorous numerical simulations demonstrate clear advantages of our method over existing techniques. This work provides a practical and scalable solution for simulating open quantum systems on quantum devices.

quant-ph

A practical applicable quantum-classical hybrid ant colony algorithm for the NISQ era

Quantum ant colony optimization (QACO) has drew much attention since it combines the advantages of quantum computing and ant colony optimization (ACO) algorithm overcoming some limitations of the traditional ACO algorithm. However,due to the hardware resource limitations of currently available quantum computers, the practical application of the QACO is still not realized. In this paper, we developed a quantum-classical hybrid algorithm by combining the clustering algorithm with QACO algorithm.This extended QACO can handle large-scale optimization problems with currently available quantum computing resource. We have tested the effectiveness and performance of the extended QACO algorithm with the Travelling Salesman Problem (TSP) as benchmarks, and found the algorithm achieves better performance under multiple diverse datasets. In addition, we investigated the noise impact on the extended QACO and evaluated its operation possibility on current available noisy intermediate scale quantum(NISQ) devices. Our work shows that the combination of the clustering algorithm with QACO effectively improved its problem solving scale, which makes its practical application possible in current NISQ era of quantum computing.

quant-ph

A Novel Quantum Algorithm for Ant Colony Optimization

Quantum ant colony optimization (QACO) has drew much attention since it combines the advantages of quantum computing and ant colony optimization (ACO) algorithms and overcomes some limitations of the traditional ACO algorithm. However, due to the hardware resource limitations of currently available quantum computers, such as the limited number of qubits, lack of high-fidelity gating operation, and low noisy tolerance, the practical application of the QACO is quite challenging. In this paper, we introduce a hybrid quantum-classical algorithm by combining the clustering algorithm with QACO algorithm, so that this extended QACO can handle large-scale optimization problems, which makes the practical application of QACO based on available quantum computation resource possible. To verify the effectiveness and performance of the algorithm, we tested the developed QACO algorithm with the Travelling Salesman Problem (TSP) as benchmarks. The developed QACO algorithm shows better performance under multiple data set. In addition, the developed QACO algorithm also manifests the robustness to noise of calculation process, which is typically a major barrier for practical application of quantum computers. Our work shows that the combination of the clustering algorithm with QACO has effectively extended the application scenario of QACO in current NISQ era of quantum computing.

quant-ph

The endomorphism rings of permutation modules of $\frac{3}{2}$-transitive permutation groups

Recent classification of $\frac{3}{2}$-transitive permutation groups leaves us with six families of groups which are $2$-transitive, or Frobenius, or one-dimensional affine, or the affine solvable subgroups of $ \mathrm{AGL}(2, q)$, or special projective linear group $\mathrm{PSL}(2, q)$, or $\mathrm{P\Gamma L}(2, q)$, where $q=2^p $ with $p$ prime. According to a case by case analysis, we prove that the endomorphism ring of the natural permutation module for a $\frac{3}{2}$-transitive permutation group is a symmetric algebra.

math.GR

On finite totally k-closed groups

Let $G$ be a finite group acting faithfully on a finite set $\Omega$. For a positive integer $k$, $G$ acts naturally on the Catesian product $\Omega^k := \Omega \times ...\times \Omega$. In this paper, we prove that finite nilpotent group $G$ with $2\nmid |G|$ is a totally $k$-closed group if and only if $G$ is abelian with $n(G)\leq k-1$ or cyclic, where $n(G)$ is the number of invariant factors in the invariant factor decomposition of $G$.

math.GR