arXiv ScienceSearch

arXiv subjects

Zilong Li

Publications and source records attributed to Zilong Li.

At least 19 recordsLinked to original sources

Implicit discretization schemes for full-kinetic ion and drift-kinetic electron simulations

We present a new electromagnetic plasma simulation model with full-kinetic ions and drift-kinetic electrons. This model (termed as FIDES) solves the electric field using the implicit perpendicular Ohm's law and a novel implicit parallel Ampere's law, where the latter requires an implicit scheme for the parallel electric field in advancing the electron weights. To suppress unphysical high-frequency instabilities, ion weights are advanced using an implicit scheme for perpendicular electric fields. Simulations of perpendicular and parallel waves validate the model's capability in handling high-frequency physics. Low-frequency wave simulations demonstrate that the implicit parallel Ampere's law can mitigate the cancellation problem more effectively than the conventional schemes using the parallel Ohm's law. To reduce the numerical damping from implicit time-stepping, we develop a second-order scheme for particle pushing. Meanwhile, an integrated strategy combining the first- and second-order schemes is employed to suppress odd-even decoupling while maintaining the accuracy of the second-order formulation.

physics.plasm-ph

Andreev Reflection to Probe Momentum-Dependent Spin Polarization in Altermagnet CrSb

Altermagnetic materials have recently emerged as promising candidates for next-generation spintronic applications, characterized by the k-dependent spin-splitted band structure and a simultaneous zero-net-magnetization. Among them, altermagnetic candidate CrSb has attracted considerable attention, owing to its g-wave spin splitting and high N\'eel temperature. In this article, we employed mechanical point-contact spectroscopy (MPCS) with superconducting Nb tips to probe the Andreev reflection on CrSb single crystals along three principal crystallographic orientations. The extracted momentum-dependent spin polarizations are approximately 73.4% for the (0001) plane, 67.9% for the (-1-120) plane, and 61.9% for the (10-10) plane, respectively, distinct from conventional antiferromagnets. Furthermore, conductance spectra from spatial line-scans on the sample surface support the existence of altermagnetic domains with a characteristic size of 250-500 nm separated by domain-walls with width about 250 nm. These results strongly support the momentum-dependent spin polarization in altermagnetic CrSb and establish Andreev reflection as a new paradigm to probe k-dependent spin textures.

cond-mat.supr-con

GraphMAR: Geometry-Aware Graph Learning Framework for Spatially Adaptive CT Metal Artifact Reduction

Computed tomography (CT) metal artifact reduction (MAR) aims to reduce the severe streaking artifacts induced by metallic implants and other high-density objects. Effective MAR generally requires both accurate artifact localization and artifact removal. Sinogram-domain methods can exploit explicit geometric cues, such as metal traces, to identify metal-corrupted measurements, while requiring raw projection data, which is often unavailable in clinical and practical scenarios. Image-domain methods are more flexible and widely applicable, yet they usually lack comparable geometric guidance, limiting their ability to localize artifacts and leading to suboptimal results. To address this limitation, we propose GraphMAR, a geometry-aware learning framework for explicit artifact identification and spatially adaptive MAR in the image domain. The key idea is to introduce graph-based geometric modeling as an image-domain analogue of sinogram metal traces. Specifically, we first construct a geometric graph from the metal mask and derive a geometric density graph that coarsely localizes artifact-prone regions according to inter-implant geometry. We then design GraphMoE, a graph-routed mixture-of-experts module that builds a polar-coordinate artifact graph in feature space and adaptively routes different experts to different spatial regions for MAR. By aligning the learned routing maps with the geometric density graph, GraphMAR provides explicit and interpretable artifact localization while enabling region-adaptive artifact reduction. Experiments on both simulated and real-world datasets demonstrate that GraphMAR achieves superior MAR performance compared with existing methods. To the best of our knowledge, this is the first work to introduce graph-based modeling for CT MAR and to enable explicit artifact identification in the image domain, improving both restoration quality and interpretability.

cs.CV

Antiferromagnetic Dimers in the Parent Phase of a Correlated Kagome Superconductor

Kagome metals are prone to charge-density wave (CDW), magnetic, and superconducting phases, with their flat electronic band conducive for correlated physics. In contrast to the weakly correlated $A$V$_3$Sb$_5$ ($A$ = K, Rb, Cs) kagome metals with a $2\times2$ CDW, CsCr$_3$Sb$_5$ is a correlated metal with a flat band close to the Fermi level, and exhibits a $4\times1$ CDW intertwined with magnetic order. Under pressure, the intertwined orders are suppressed and give way to a dome of superconductivity that emerges from a non-Fermi liquid normal state. Here, we solve the crystal structure of the $4\times 1$ CDW state in CsCr$_3$Sb$_5$, and show it consists of Cr dimers separated by Cr chains. First-principles calculations show the dominant exchange interaction is antiferromagnetic within the dimers, while the intra-chain and dimer-chain couplings are much weaker. The CDW transition of CsCr$_3$Sb$_5$ is found to be more strongly first-order than those in $A$V$_3$Sb$_5$, without significant soft phonons or diffuse scattering above the CDW transition temperature. These findings suggest that fluctuating antiferromagnetic dimers may play a major role in the electron pairing of superconducting CsCr$_3$Sb$_5$.

cond-mat.str-el

PHASOR: Anatomy- and Phase-Consistent Volumetric Diffusion for CT Virtual Contrast Enhancement

Contrast-enhanced computed tomography (CECT) is pivotal for highlighting tissue perfusion and vascularity, yet its clinical ubiquity is impeded by the invasive nature of contrast agents and radiation risks. While virtual contrast enhancement (VCE) offers an alternative to synthesizing CECT from non-contrast CT (NCCT), existing methods struggle with anatomical heterogeneity and spatial misalignment, leading to inconsistent enhancement patterns and incorrect details. This paper introduces PHASOR, a volumetric diffusion framework for high-fidelity CT VCE. By treating CT volumes as coherent sequences, we leverage a video diffusion model to enhance structural coherence and volumetric accuracy. To ensure anatomy-phase consistent synthesis, we introduce two complementary modules. First, anatomy-routed mixture-of-experts (AR-MoE) anchors distinct enhancement patterns to anatomical semantics, with organ-specific memory to capture salient details. Second, intensity-phase aware representation alignment (IP-REPA) highlights intricate contrast signals while mitigating the impact of imperfect spatial alignment. Extensive experiments across three datasets demonstrate that PHASOR significantly outperforms state-of-the-art methods in both synthesis quality and enhancement accuracy.

cs.CV

Hardware-Impaired Over-the-Air Computation with Fluid Antenna Array

This paper investigates a fluid antenna (FA) array-enhanced over-the-air computation (AirComp) system in the presence of hardware impairments (HWIs), exploiting the new degrees of freedom offered by reconfigurable antenna positioning. To minimize the mean squared error (MSE) of the aggregated signal, we jointly optimize transmit power control, receive beamforming, and the antenna position vector (APV), subject to practical constraints such as HWI-induced distortion noise, FA movement energy consumption, and total power budgets. The resulting optimization problem is non-convex and highly coupled. To address it efficiently, we adopt a block coordinate descent (BCD) framework, decomposing it into three manageable subproblems. For each subproblem, closed-form solutions or efficient numerical algorithms are derived. Simulation results demonstrate that the proposed joint transceiver and APV design significantly reduces the MSE compared to conventional fixed-position antenna (FPA) arrays and exhibits enhanced robustness against hardware impairments. The effectiveness and convergence of the proposed algorithm are further validated under various system configurations.

eess.SP

EduEval: A Hierarchical Cognitive Benchmark for Evaluating Large Language Models in Chinese Education

Large language models (LLMs) demonstrate significant potential for educational applications. However, their unscrutinized deployment poses risks to educational standards, underscoring the need for rigorous evaluation. We introduce EduEval, a comprehensive hierarchical benchmark for evaluating LLMs in Chinese K-12 education. This benchmark makes three key contributions: (1) Cognitive Framework: We propose the EduAbility Taxonomy, which unifies Bloom's Taxonomy and Webb's Depth of Knowledge to organize tasks across six cognitive dimensions including Memorization, Understanding, Application, Reasoning, Creativity, and Ethics. (2) Authenticity: Our benchmark integrates real exam questions, classroom conversation, student essays, and expert-designed prompts to reflect genuine educational challenges; (3) Scale: EduEval comprises 24 distinct task types with over 11,000 questions spanning primary to high school levels. We evaluate 14 leading LLMs under both zero-shot and few-shot settings, revealing that while models perform well on factual tasks, they struggle with classroom dialogue classification and exhibit inconsistent results in creative content generation. Interestingly, several open source models outperform proprietary systems on complex educational reasoning. Few-shot prompting shows varying effectiveness across cognitive dimensions, suggesting that different educational objectives require tailored approaches. These findings provide targeted benchmarking metrics for developing LLMs specifically optimized for diverse Chinese educational tasks.

cs.CL

Field Inversion Machine Learning for Time-Resolved Unsteady Flows in Airfoil Dynamic Stall

While many existing machine learning studies have focused on augmenting Reynolds averaged Navier Stokes (RANS) turbulence models for steady or time averaged unsteady flows, this paper takes a first step toward extending such augmentation to time resolved unsteady flows. An unsteady field inversion and machine learning (FIML) method is developed, in which a temporally evolving correction field (beta) is incorporated into the production term of a RANS turbulence model. The inverse problem is solved by optimizing the spatial temporal distribution of beta to minimize the regularized prediction errors. The resulting optimized beta field is then used to train a multi layer neural network that learns the time dependent relationship between local flow features and beta. The approach is demonstrated using the unsteady flow over a NACA0012 airfoil undergoing dynamic stall. Results show that the unsteady FIML model, trained using only the time series of drag data at a given pitch rate, can accurately reproduce the spatial temporal evolution of reference drag, lift, pitching moment, surface pressure, and velocity fields at both identical and different pitch rates. The unsteady FIML is integrated into the open source DAFoam framework, enabling a pathway toward developing accurate and generalizable RANS turbulence models for time resolved unsteady flows.

physics.flu-dyn

Translation via Annotation: A Computational Study of Translating Classical Chinese into Japanese

Ancient people translated classical Chinese into Japanese using a system of annotations placed around characters. We abstract this process as sequence tagging tasks and fit them into modern language technologies. The research on this annotation and translation system faces a low resource problem. We alleviate this problem by introducing an LLM-based annotation pipeline and constructing a new dataset from digitized open-source translation data. We show that in the low-resource setting, introducing auxiliary Chinese NLP tasks enhances the training of sequence tagging tasks. We also evaluate the performance of Large Language Models (LLMs) on this task. While they achieve high scores on direct machine translation, our method could serve as a supplement to LLMs to improve the quality of character's annotation.

cs.CL

Spin triplet pairing by suppressing altermagnetism

The interplay between unconventional superconductivity and altermagnetic order has attracted much attention. In particular, whether spin-triplet superconductivity can be achieved by suppressing altermagnetism remains an open issue. We investigate this issue using a minimal single-orbital Hubbard model on a square lattice with a vacancy superstructure, in which both conventional antiferromagnetic and altermagnetic orders can emerge on an equal footing. We illustrate the existence of a metallic normal phase with altermagnetism even at half-filling due to geometric frustration and Coulomb interaction. Suppressing the altermagnetic long-range order using charge doping can lead to both conventional antiferromagnetic and altermagnetic spin fluctuations. Spin-singlet pairing is always favored when conventional antiferromagnetic spin fluctuations dominate. However, when altermagnetic spin fluctuations dominate, spin-triplet pairing will be induced. Implications of our results for possible material candidates are also briefly discussed.

cond-mat.supr-con

FoundDiff: Foundational Diffusion Model for Generalizable Low-Dose CT Denoising

Low-dose computed tomography (CT) denoising is crucial for reduced radiation exposure while ensuring diagnostically acceptable image quality. Despite significant advancements driven by deep learning (DL) in recent years, existing DL-based methods, typically trained on a specific dose level and anatomical region, struggle to handle diverse noise characteristics and anatomical heterogeneity during varied scanning conditions, limiting their generalizability and robustness in clinical scenarios. In this paper, we propose FoundDiff, a foundational diffusion model for unified and generalizable LDCT denoising across various dose levels and anatomical regions. FoundDiff employs a two-stage strategy: (i) dose-anatomy perception and (ii) adaptive denoising. First, we develop a dose- and anatomy-aware contrastive language-image pre-training model (DA-CLIP) to achieve robust dose and anatomy perception by leveraging specialized contrastive learning strategies to learn continuous representations that quantify ordinal dose variations and identify salient anatomical regions. Second, we design a dose- and anatomy-aware diffusion model (DA-Diff) to perform adaptive and generalizable denoising by synergistically integrating the learned dose and anatomy embeddings from DA-CLIP into diffusion process via a novel dose and anatomy conditional block (DACB) based on Mamba. Extensive experiments on a large simulated multi-dose CT dataset spanning three anatomical regions, together with cross-dataset evaluations on Mayo-2016, CQ500, and piglet datasets, demonstrate superior denoising performance and strong generalization to unseen dose levels and anatomical regions. The codes and models are available at https: //github.com/hao1635/FoundDiff.

cs.CV

Unified Medical Image Tokenizer for Autoregressive Synthesis and Understanding

Autoregressive modeling has driven major advances in multimodal AI, yet its application to medical imaging remains constrained by the absence of a unified image tokenizer that simultaneously preserves fine-grained anatomical structures and rich clinical semantics across heterogeneous modalities. Existing approaches jointly optimize image reconstruction and textual semantic objectives, relying on large-scale image-caption pairs and are prone to gradient interference. This is ill-suited for the medical domain where paired data are scarce and abundant unpaired images remain unexploited. This work identifies these issues in building unified medical image tokenizers, and introduces a principled two-stage training framework using visual representation as a bridge to address them. The propose visual representation alignment stage enables the utilization of large-scale unpaired medical images to ensure reconstruction fidelity and establish foundational semantics, alleviating the interference and better preparing for the second stage where fine-grained textual semantics are injected using image-text pairs. The resulting tokenizer, MedITok, is trained on over 33 million medical images spanning 9 modalities and 2 million image-text pairs. MedITok achieves state-of-the-art performance on 30+ benchmarks spanning 9 imaging modalities and 4 task families. It further enables autoregressive modeling for diagnostic and generative applications, serving as a scalable component for future multimodal models with unified synthesis and understanding capabilities in the medical domain. Project page: https://github.com/Masaaki-75/meditok

eess.IV

A kinetic CMA diagram

We present a kinetic Clemmow-Mullaly-Allis (CMA) diagram by systematically analysing the kinetic effects on the wave propagation in a homogeneous thermal plasma. The differences between the cold and kinetic CMA diagrams are outlined. It is found that new boundaries for weakly damped left- and right-handed circularly polarized waves are located above the ion and electron cyclotron frequency lines in the kinetic CMA diagram. Additionally, Langmuir waves in the kinetic CMA diagram occupy a specific region between the new Langmuir wave boundary and the plasma frequency line, while in the cold CMA diagram, they exist on the plasma frequency line. The extraordinary-Bernstein mode transformation frequency lines in the kinetic CMA diagram replace the hybrid resonant frequency lines of the cold CMA diagram, with discontinuities between different cyclotron harmonics. These new boundaries partition the parameter space in the kinetic CMA diagram differently, leading to new inverse wave normal surfaces in the regions bounded by new boundaries. The kinetic CMA diagram not only contributes to a basic understanding of wave properties in thermal plasmas, but also can provide a powerful tool to explore new possible propagation paths.

physics.plasm-ph

Pressure Induced Altermagnetism in Layered Ternary Iron-Selenides

Employing first-principles based calculations, we reexamined the high-pressure phases of the vacancy-ordered iron-selenides, i.e. A2Fe4Se5 phase. A magnetic transition from the block-spin antiferromagnetic phase to Neel-AM phase is observed under high pressure when the iron-vacancy order is preserved. The transition is first-order, driven by the collapse of c-axis and accompanied by an insulator-metal transition. In addition, the Neel-AM phase is a rare example of intrinsic altermagnetism that does not require ligand atoms nor supercell construction to break the spatial inversion or translational symmetry between the two spin sublattices, with a spin-splitting of the band structure as large as 300 meV. If the re-entrant superconducting phase of KyFe2-xSe2 under high pressure emerges from the Neel-AM normal state, an equal-spin triplet pairing would be naturally favored.

cond-mat.mtrl-sci

Radiologist-in-the-Loop Self-Training for Generalizable CT Metal Artifact Reduction

Metal artifacts in computed tomography (CT) images can significantly degrade image quality and impede accurate diagnosis. Supervised metal artifact reduction (MAR) methods, trained using simulated datasets, often struggle to perform well on real clinical CT images due to a substantial domain gap. Although state-of-the-art semi-supervised methods use pseudo ground-truths generated by a prior network to mitigate this issue, their reliance on a fixed prior limits both the quality and quantity of these pseudo ground-truths, introducing confirmation bias and reducing clinical applicability. To address these limitations, we propose a novel Radiologist-In-the-loop SElf-training framework for MAR, termed RISE-MAR, which can integrate radiologists' feedback into the semi-supervised learning process, progressively improving the quality and quantity of pseudo ground-truths for enhanced generalization on real clinical CT images. For quality assurance, we introduce a clinical quality assessor model that emulates radiologist evaluations, effectively selecting high-quality pseudo ground-truths for semi-supervised training. For quantity assurance, our self-training framework iteratively generates additional high-quality pseudo ground-truths, expanding the clinical dataset and further improving model generalization. Extensive experimental results on multiple clinical datasets demonstrate the superior generalization performance of our RISE-MAR over state-of-the-art methods, advancing the development of MAR models for practical application. Code is available at https://github.com/Masaaki-75/rise-mar.

eess.IV

Team Ryu's Submission to SIGMORPHON 2024 Shared Task on Subword Tokenization

This papers presents the submission of team Ryu to the canceled SIGMORPHON 2024 shared task on subword tokenization. My submission explores whether morphological segmentation methods can be used as a part of subword tokenizers. I adopt two approaches: the statistical segmentation method Morfessor and a transformer based sequence-to-sequence (seq2seq) segmentation model in tokenizers. The prediction results show that morphological segmentation could be as effective as commonly used subword tokenizers. Additionally, I investigate how a tokenizer's vocabulary influences the performance of language models. A tokenizer with a balanced token frequency distribution tends to work better. A balanced token vocabulary can be achieved by keeping frequent words as unique tokens.

cs.CL

The stability and instability of the language control network: a longitudinal resting-state functional magnetic resonance imaging study

The language control network is vital among language-related networks responsible for solving the problem of multiple language switching. Researchers have expressed concerns about the instability of the language control network when exposed to external influences (e.g., Long-term second language learning). However, some studies have suggested that the language control network is stable. Therefore, whether the language control network is stable or not remains unclear. In the present study, we directly evaluated the stability and instability of the language control network using resting-state functional magnetic resonance imaging (rs-fMRI). We employed cohorts of Chinese first-year college students majoring in English who underwent second language (L2) acquisition courses at a university and those who did not. Two resting-state fMRI scans were acquired approximately 1 year apart. We found that the language control network was both moderately stable and unstable. We further investigated the morphological coexistence patterns of stability and instability within the language control network. First, we extracted connections representing stability and plasticity from the entire network. We then evaluated whether the coexistence patterns were modular (stability and instability involve different brain regions) or non-modular (stability and plasticity involve the same brain regions but have unique connectivity patterns). We found that both stability and instability coexisted in a non-modular pattern. Compared with the non-English major group, the English major group has a more non-modular coexistence pattern.. These findings provide preliminary evidence of the coexistence of stability and instability in the language control network.

q-bio.NC

Prompted Contextual Transformer for Incomplete-View CT Reconstruction

Incomplete-view computed tomography (CT) can shorten the data acquisition time and allow scanning of large objects, including sparse-view and limited-angle scenarios, each with various settings, such as different view numbers or angular ranges. However, the reconstructed images present severe, varying artifacts due to different missing projection data patterns. Existing methods tackle these scenarios/settings separately and individually, which are cumbersome and lack the flexibility to adapt to new settings. To enjoy the multi-setting synergy in a single model, we propose a novel Prompted Contextual Transformer (ProCT) for incomplete-view CT reconstruction. The novelties of ProCT lie in two folds. First, we devise a projection view-aware prompting to provide setting-discriminative information, enabling a single ProCT to handle diverse incomplete-view CT settings. Second, we propose artifact-aware contextual learning to sense artifact pattern knowledge from in-context image pairs, making ProCT capable of accurately removing the complex, unseen artifacts. Extensive experimental results on two publicly available clinical CT datasets demonstrate the superior performance of ProCT over state-of-the-art methods -- including single-setting models -- on a wide range of incomplete-view CT settings, strong transferability to unseen datasets and scenarios, and improved performance when sinogram data is available. The code is available at: https://github.com/Masaaki-75/proct

eess.IV