arXiv ScienceSearch

arXiv subjects

Zhi Lin

Publications and source records attributed to Zhi Lin.

At least 19 recordsLinked to original sources

Grand-canonical phase diagram and chiral-current suppression at $\pi$ flux in a bosonic two-leg ladder

We investigate the ground-state phase diagram of repulsively interacting bosons on a two-leg ladder threaded by a uniform artificial magnetic flux, using the cluster Gutzwiller mean-field method. In the strong-rung-coupling regime, self-consistent calculations are performed on a $2\times4$ cluster. By analyzing the superfluid order parameter, leg-resolved currents, chiral current, the current ratio on adjacent legs, and the density imbalance between the two legs, we distinguish Mott-insulating from superfluid regimes and characterize the observed states as Meissner-like, vortex-like (superfluid or Mott insulating), or biased-ladder. In regions overlapping with previous DMRG studies, our results qualitatively agree with the established phase structure, demonstrating that the cluster Gutzwiller approach balances computational efficiency and physical accuracy. We then construct the first grand-canonical $t$--$\mu$ phase diagrams for this system, revealing how the magnetic flux modifies the shape, tilt, and extent of the Mott lobes. We further explore previously inaccessible regimes, including higher fillings $\rho\gtrsim1$ and the intermediate interaction window $U/t\in[7.69,9.09]$. Special attention is paid to $\varphi=\pi$, where the effective triangular-ladder mapping becomes singular. Owing to the equivalence of $\varphi=\pi$ and $-\pi$ modulo $2\pi$, a combined symmetry forbids net chiral currents, leading to a nonchiral Mott-insulating state, in contrast to the chiral-superfluid tendency expected away from $\varphi=\pi$. Our results offer a computationally efficient route for mapping the global phase structure of bosonic flux ladders and provide guidance for future ultracold-atom experiments in artificial gauge fields.

cond-mat.quant-gas

HAVE: Head-Adaptive Gating and ValuE Calibration for Hallucination Mitigation in Large Language Models

Large Language Models (LLMs) often produce hallucinations in retrieval-augmented or long-context generation, even when relevant evidence is present. This stems from two issues: head importance is treated as input-agnostic, and raw attention weights poorly reflect each token's true contribution. We present HAVE (Head-Adaptive Gating and ValuE Calibration), a parameter-free decoding framework that directly addresses both challenges. HAVE introduces head-adaptive gating, which performs instance-level soft reweighing of attention heads, and value calibration, which augments attention with the magnitude of value vectors to approximate write-back contribution. Together, these modules construct token-level evidence aligned with model updates and fuse it with the LM distribution through a lightweight uncertainty-scaled policy. HAVE requires no finetuning and operates in a single forward pass, making it efficient and broadly applicable. Experiments across multiple QA benchmarks and LLM families demonstrate that HAVE consistently reduces hallucinations and outperforms strong baselines, including DAGCD, with modest overhead. The framework is transparent, reproducible, and readily integrates with off-the-shelf LLMs, advancing trustworthy generation in real-world settings.

cs.CL

MEUV: Achieving Fine-Grained Capability Activation in Large Language Models via Mutually Exclusive Unlock Vectors

Large language models (LLMs) enforce safety alignment to reliably refuse malicious requests, yet the same blanket safeguards also block legitimate uses in policing, defense, and other high-stakes settings. Earlier "refusal-direction" edits can bypass those layers, but they rely on a single vector that indiscriminately unlocks all hazardous topics, offering no semantic control. We introduce Mutually Exclusive Unlock Vectors (MEUV), a lightweight framework that factorizes the monolithic refusal direction into topic-aligned, nearly orthogonal vectors, each dedicated to one sensitive capability. MEUV is learned in a single epoch with a multi-task objective that blends a differential-ablation margin, cross-topic and orthogonality penalties, and several auxiliary terms. On bilingual malicious-prompt benchmarks, MEUV achieves an attack success rate of no less than 87% on Gemma-2-2B, LLaMA-3-8B, and Qwen-7B, yet cuts cross-topic leakage by up to 90% compared with the best single-direction baseline. Vectors trained in Chinese transfer almost unchanged to English (and vice versa), suggesting a language-agnostic refusal subspace. The results show that fine-grained, topic-level capability activation is achievable with minimal utility loss, paving the way for controlled LLMs deployment in security-sensitive domains.

cs.LG

A mini-batch training strategy for deep subspace clustering networks

Mini-batch training is a cornerstone of modern deep learning, offering computational efficiency and scalability for training complex architectures. However, existing deep subspace clustering (DSC) methods, which typically combine an autoencoder with a self-expressive layer, rely on full-batch processing. The bottleneck arises from the self-expressive module, which requires representations of the entire dataset to construct a self-representation coefficient matrix. In this work, we introduce a mini-batch training strategy for DSC by integrating a memory bank that preserves global feature representations. Our approach enables scalable training of deep architectures for subspace clustering with high-resolution images, overcoming previous limitations. Additionally, to efficiently fine-tune large-scale pre-trained encoders for subspace clustering, we propose a decoder-free framework that leverages contrastive learning instead of autoencoding for representation learning. This design not only eliminates the computational overhead of decoder training but also provides competitive performance. Extensive experiments demonstrate that our approach not only achieves performance comparable to full-batch methods, but outperforms other state-of-the-art subspace clustering methods on the COIL100 and ORL datasets by fine-tuning deep networks.

cs.CV

Fourth- and Higher-Order Semi-Lagrangian Finite Volume Methods for the Two-dimensional Advection Equation on Arbitrarily Complex Domains

To numerically solve the two-dimensional advection equation, we propose a family of fourth- and higher-order semi-Lagrangian finite volume (SLFV) methods that feature (1) fourth-, sixth-, and eighth-order convergence rates, (2) applicability to both regular and irregular domains with arbitrarily complex topology and geometry, (3) ease of handling both zero and nonzero source terms, and (4) the same algorithmic steps for both periodic and incoming penetration conditions. Test results confirm the analysis and demonstrate the accuracy, flexibility, robustness, and excellent conditioning of the proposed SLFV method.

math.NA

Floquet-Engineered Hybrid Topological Orders with Majorana Edge Modes in Number-Conserving Fermionic Quantum Simulators

We develop an experimental protocol based on Floquet-engineered ultracold fermions in optical lattices, enabling the emulation of pair-hopping and competing singlet/triplet pairing interactions. Through large-scale density matrix renormalization group (DMRG) simulations, we uncover three emergent topological phases: (i) A Majorana-enabled spin-density-wave (MS) phase featuring exponentially localized edge charges, non-local fermionic edge correlations, and doubly degenerate entanglement spectra; (ii) A z-axis polarized triplet superconducting (TS) phase exhibiting fractionalized edge spins (S=1/4 per edge), two-fold ground state degeneracy and a bulk single-particle gap; (iii) A hybrid x-directional triplet superconducting (XTS) phase that uniquely combines fractional spin textures and Majorana-type edge correlations, defining a new universality class of hybrid orders in number-conserving systems. These findings establish a universal framework for engineering non-Abelian topological matter, crucially bypassing the need for external pairing fields while maintaining experimental feasibility with current cold-atom techniques.

cond-mat.quant-gas

Beyond Conventional Pairing: Bosonic Quartic Superfluidity, Exotic XY Magnetism, and Phase Criticality

We report two unprecedented bosonic quartic superfluid (BQSF) phases in binary boson mixtures with synthetic pair-hopping (SPH) interaction realizable through Floquet engineering, transcending the conventional bosonic superfluidity paradigm via high-order correlations. Through extensive numerical simulations, we demonstrate: (i) The paired super-counter-fluid (PSCF) phase that exhibits bosonic 'four-particle' correlations, directly mirroring fermionic 4e superconductivity; (ii) The symbiotic super-counter-fluid which is an anomalous quantum phase featuring intrinsically intertwined, non-zero PSCF and super-counter-fluid orders. Both phases further reveal unreported pseudo-spin-ordered XY ferromagnetic states with filling-dependent magnetization textures. Adjacent to BQSF phases, we uncover both a previously unreported quantum quadruple critical point induced by interexchange asymmetry and SPH-driven state-dependent criticality-the latter exhibiting distinct behavior from conventional two-component bosonic systems. We propose momentum-resolved noise correlation spectroscopy-building on established quantum gas microscopy techniques-as a direct probe for the characteristic signatures of these BQSF states. Our results establish a new paradigm in quantum matter, opening unprecedented avenues to explore (i) quartic quantum coherence, (ii) emergent collective phenomena, and (iii) extended XY magnetism-resolving a longstanding classification challenge while discove

cond-mat.quant-gas

Delta-Learning approach combined with the cluster Gutzwiller approximation for strongly correlated bosonic systems

The cluster Gutzwiller method is widely used to study the strongly correlated bosonic systems, owing to its ability to provide a more precise description of quantum fluctuations. However, its utility is limited by the exponential increase in computational complexity as the cluster size grows. To overcome this limitation, we propose an artificial intelligence-based method known as $\Delta$-Learning. This approach constructs a predictive model by learning the discrepancies between lower-precision (small cluster sizes) and high-precision (large cluster sizes) implementations of the cluster Gutzwiller method, requiring only a small number of training samples. Using this predictive model, we can effectively forecast the outcomes of high-precision methods with high accuracy. Applied to various Bose-Hubbard models, the $\Delta$-Learning method effectively predicts phase diagrams while significantly reducing the computational resources and time. Furthermore, we have compared the predictive accuracy of $\Delta$-Learning with other direct learning methods and found that $\Delta$-Learning exhibits superior performance in scenarios with limited training data. Therefore, when combined with the cluster Gutzwiller approximation, the $\Delta$-Learning approach offers a computationally efficient and accurate method for studying phase transitions in large, complex bosonic systems.

cond-mat.quant-gas

Frustrated hopping from orbital decoration of a primitive two-dimensional lattice

Materials hosting flat electronic bands are a central focus of condensed matter physics as promising venues for novel electronic ground states. Two-dimensional (2D) geometrically frustrated lattices such as the kagome, dice, and Lieb lattices are attractive targets in this direction, anticipated to realize perfectly flat bands. Synthesizing these special structures, however, poses a formidable challenge, exemplified by the absence of solid-state materials realizing the dice and Lieb lattices. An alternative route leverages atomic orbitals to create the characteristic electron hopping of geometrically frustrated lattices. This strategy promises to expand the list of candidate materials to simpler structures, but is yet to be demonstrated experimentally. Here, we report the realization of frustrated hopping in the van der Waals (vdW) intermetallic Pd$_5$AlI$_2$, emerging from orbital decoration of a primitive square lattice. Using angle-resolved photoemission spectroscopy and quantum oscillations measurements, we demonstrate that the band structure of Pd$_5$AlI$_2$ includes linear Dirac-like bands intersected at their crossing point by a flat band, essential characteristics of frustrated hopping in the Lieb and dice lattices. Moreover, Pd$_5$AlI$_2$ is exceptionally stable, with the unusual bulk band structure and metallicity persisting in ambient conditions down to the monolayer limit. Our ability to realize an electronic structure characteristic of geometrically frustrated lattices establishes orbital decoration of primitive lattices as a new approach towards electronic structures that remain elusive to prevailing lattice-centric searches.

cond-mat.str-el

CPSDBench: A Large Language Model Evaluation Benchmark and Baseline for Chinese Public Security Domain

Large Language Models (LLMs) have demonstrated significant potential and effectiveness across multiple application domains. To assess the performance of mainstream LLMs in public security tasks, this study aims to construct a specialized evaluation benchmark tailored to the Chinese public security domain--CPSDbench. CPSDbench integrates datasets related to public security collected from real-world scenarios, supporting a comprehensive assessment of LLMs across four key dimensions: text classification, information extraction, question answering, and text generation. Furthermore, this study introduces a set of innovative evaluation metrics designed to more precisely quantify the efficacy of LLMs in executing tasks related to public security. Through the in-depth analysis and evaluation conducted in this research, we not only enhance our understanding of the performance strengths and limitations of existing models in addressing public security issues but also provide references for the future development of more accurate and customized LLM models targeted at applications in this field.

cs.AI

Optimal Stirring Strategies for Passive Scalars in a Domain with a General Shape and No-Flux Boundary Condition

Multiscale metrics such as negative Sobolev norms are effective for quantifying the degree of mixedness of a passive scalar field advected by an incompressible flow in the absence of diffusion. In this paper we introduce a mix norm that is motivated by Sobolev norm $H^{-1}$ for a general domain with a no-flux boundary. We then derive an explicit expression for the optimal flow that maximizes the instantaneous decay rate of the mix norm under fixed energy and enstrophy constraints. Numerical simulations indicate that the mix norm decays exponentially or faster for various initial conditions and geometries and the rate is closely related to the smallest non-zero eigenvalue of the Laplace operator. These results generalize previous findings restricted for a periodic domain for its analytical and numerical simplicity. Additionally, we observe that periodic boundaries tend to induce a faster decay in mix norm compared to no-flux conditions under the fixed energy constraint, while the comparison is reversed for the fixed enstrophy constraint. In the special case of even initial distributions, two types of boundary conditions yield the same optimal flow and mix norm decay.

math.OC

A Causal Framework to Unify Common Domain Generalization Approaches

Domain generalization (DG) is about learning models that generalize well to new domains that are related to, but different from, the training domain(s). It is a fundamental problem in machine learning and has attracted much attention in recent years. A large number of approaches have been proposed. Different approaches are motivated from different perspectives, making it difficult to gain an overall understanding of the area. In this paper, we propose a causal framework for domain generalization and present an understanding of common DG approaches in the framework. Our work sheds new lights on the following questions: (1) What are the key ideas behind each DG method? (2) Why is it expected to improve generalization to new domains theoretically? (3) How are different DG methods related to each other and what are relative advantages and limitations? By providing a unified perspective on DG, we hope to help researchers better understand the underlying principles and develop more effective approaches for this critical problem in machine learning.

cs.LG

Two-Stage Holistic and Contrastive Explanation of Image Classification

The need to explain the output of a deep neural network classifier is now widely recognized. While previous methods typically explain a single class in the output, we advocate explaining the whole output, which is a probability distribution over multiple classes. A whole-output explanation can help a human user gain an overall understanding of model behaviour instead of only one aspect of it. It can also provide a natural framework where one can examine the evidence used to discriminate between competing classes, and thereby obtain contrastive explanations. In this paper, we propose a contrastive whole-output explanation (CWOX) method for image classification, and evaluate it using quantitative metrics and through human subject studies. The source code of CWOX is available at https://github.com/vaynexie/CWOX.

cs.CV

Consistency Regularization for Domain Generalization with Logit Attribution Matching

Domain generalization (DG) is about training models that generalize well under domain shift. Previous research on DG has been conducted mostly in single-source or multi-source settings. In this paper, we consider a third, lesser-known setting where a training domain is endowed with a collection of pairs of examples that share the same semantic information. Such semantic sharing (SS) pairs can be created via data augmentation and then utilized for consistency regularization (CR). We present a theory showing CR is conducive to DG and propose a novel CR method called Logit Attribution Matching (LAM). We conduct experiments on five DG benchmarks and four pretrained models with SS pairs created by both generic and targeted data augmentation methods. LAM outperforms representative single/multi-source DG methods and various CR methods that leverage SS pairs. The code and data of this project are available at https://github.com/Gaohan123/LAM

cs.LG

Gain without Pain: Recycling Reflected Energy from Wireless Powered RIS-aided Communications

In this paper, we investigate and analyze energy recycling for a reconfigurable intelligent surface (RIS)-aided wireless-powered communication network. As opposed to the existing works where the energy harvested by Internet of things (IoT) devices only come from the power station, IoT devices are also allowed to recycle energy from other IoT devices. In particular, we propose group switching- and user switching-based protocols with time-division multiple access to evaluate the impact of energy recycling on system performance. Two different optimization problems are respectively formulated for maximizing the sum throughput by jointly optimizing the energy beamforming vectors, the transmit power, the transmission time, the receive beamforming vectors, the grouping factors, and the phase-shift matrices, where the constraints of the minimum throughput, the harvested energy, the maximum transmit power, the phase shift, the grouping, and the time allocation are taken into account. In light of the intractability of the above problems, we respectively develop two alternating optimization-based iterative algorithms by combining the successive convex approximation method and the penalty-based method to obtain corresponding sub-optimal solutions. Simulation results verify that the energy recycling-based mechanism can assist in enhancing the performance of IoT devices in terms of energy harvesting and information transmission. Besides, we also verify that the group switching-based algorithm can improve more sum throughput of IoT devices, and the user switching-based algorithm can harvest more energy.

cs.IT

Generalized effective-potential Landau theory for a tunable state-dependent hexagonal optical lattice

We analytically study the ground-state phase diagrams of ultracold bosons with various values of the effective magnetic quantum number $m$ in a state-dependent hexagonal optical lattice by using the generalized effective-potential Landau theory, where the site-offset energy between the two triangular sublattice A and B is tunable. Our analytical calculations of third-order corrections are in reasonably good agreement with the previous cluster Gutzwiller calculations. Furthermore, we reveal the reason why the regions of the Mott lobes $(n,n)$ in phase diagrams for $m=0.02$ are unexpectedly expanded with increasing $J/U$ in deep lattice.

cond-mat.quant-gas

A New Dataset and A Baseline Model for Breast Lesion Detection in Ultrasound Videos

Breast lesion detection in ultrasound is critical for breast cancer diagnosis. Existing methods mainly rely on individual 2D ultrasound images or combine unlabeled video and labeled 2D images to train models for breast lesion detection. In this paper, we first collect and annotate an ultrasound video dataset (188 videos) for breast lesion detection. Moreover, we propose a clip-level and video-level feature aggregated network (CVA-Net) for addressing breast lesion detection in ultrasound videos by aggregating video-level lesion classification features and clip-level temporal features. The clip-level temporal features encode local temporal information of ordered video frames and global temporal information of shuffled video frames. In our CVA-Net, an inter-video fusion module is devised to fuse local features from original video frames and global features from shuffled video frames, and an intra-video fusion module is devised to learn the temporal information among adjacent video frames. Moreover, we learn video-level features to classify the breast lesions of the original video as benign or malignant lesions to further enhance the final breast lesion detection performance in ultrasound videos. Experimental results on our annotated dataset demonstrate that our CVA-Net clearly outperforms state-of-the-art methods. The corresponding code and dataset are publicly available at \url{https://github.com/jhl-Det/CVA-Net}.

eess.IV

Example Perplexity

Some examples are easier for humans to classify than others. The same should be true for deep neural networks (DNNs). We use the term example perplexity to refer to the level of difficulty of classifying an example. In this paper, we propose a method to measure the perplexity of an example and investigate what factors contribute to high example perplexity. The related codes and resources are available at https://github.com/vaynexie/Example-Perplexity.

cs.LG