arXiv ScienceSearch

arXiv subjects

Chuan Chen

Publications and source records attributed to Chuan Chen.

At least 19 recordsLinked to original sources

Structural Entropy-Driven Graph Diffusion Generation for One-Shot Federated Graph Learning

One-shot federated graph learning (FGL) requires the server to estimate client contributions from highly compressed information, yet conventional volume-based weighting captures the amount of client data while overlooking how its connectivity is organized. In this paper, we propose SPIRE, a Structural Entropy-Driven Graph Diffusion Generation method that introduces topology-aware client differentiation into one-shot FGL. Specifically, we employ first-order degree-distribution structural entropy as a compact descriptor of degree-mass dispersion and use it to derive structural client weights, providing an inductive bias that accounts for differences in graph topology beyond data volume. On the generation side, a graph diffusion model on the server synthesizes pseudographs conditioned on the weighted client prototypes, capturing both semantic and structural information without requiring additional client-side training. The generated pseudographs are then assembled via disjoint union fusion to train a global graph neural network. Extensive experiments on seven real-world graph datasets demonstrate that SPIRE consistently outperforms conventional and one-shot FGL methods, with particularly strong gains under highly heterogeneous (non-IID) and graph-perturbed settings.

cs.LG

Dynamical Response of the Kitaev Spin Liquid under Third-Nearest-Neighbor Heisenberg Interaction

Motivated by growing evidence for the significance of the third-nearest-neighbor Heisenberg ($J_3$) interaction in candidate Kitaev materials, we investigate the dynamical properties of the Kitaev spin liquid (KSL) under a $J_3$ perturbation, focusing on its spin dynamical structure factor (DSF) and Raman scattering. Within a self-consistent parton mean-field plus random-phase approximation framework, we find that $J_3$ induces coherent, paramagnon-like collective modes that coexist with a high-energy Majorana continuum in the spin DSF. The softening of these modes with increasing $|J_3|$ signals a quantum phase transition to magnetic order. Remarkably, magnetic ordering sets in at a common critical $J_3$ for both ferromagnetic ($K<0$) and antiferromagnetic ($K>0$) Kitaev models, with the resulting ordered states forming exact dual pairs under a four-sublattice duality transformation that maps $(K,J_3) \rightarrow (-K,J_3)$. An external magnetic field further softens the preexisting paramagnon modes, thereby enhancing magnetic order. Perturbative Raman calculations show that while the Kitaev-like Raman vertex probes only itinerant matter Majorana fermions, the response from the $J_3$-like vertex features both matter Majoranas and visons. Four-vison excitations produce a sharp peak accompanied by a two-fermion continuum, whereas two-vison excitations yield a continuum closely resembling the single-matter-fermion density of states. These results provide a unified perspective on the dynamical signatures of $J_3$-perturbed KSL and are helpful for interpreting experimental spectra in candidate Kitaev materials with sizable $J_3$ interactions.

cond-mat.str-el

Automatic Echocardiography Segmentation via Transition Probability Correlation for Stable Semantic Extraction

While echocardiography is essential for cardiovascular diagnosis, inherent speckle noise and low signal-to-noise ratio often lead to ambiguous semantic features and fragmented boundaries. These limitations significantly hinder the segmentation accuracy of deep learning models in complex clinical cases. Moreover, temporal motion of the heart plays a critical role in recognizing anatomical structures. To address these challenges, we designed a STLSF module which comprises a window-matching-based semantic correction component and a semantics-guided texture enhancement component. By leveraging local transition probability correlations to correct semantics and employing semantics-guided texture enhancement, the STLSF module effectively mitigates texture instability and ambiguous semantic interpretations caused by disadvantaged echocardiography quality. Additionally, to facilitate the encoder's adaptation to the intrinsic priors of ultrasound-specific imaging patterns, we propose a frequency-aware denoising pre-training method. The entire work builds a convolution-based network with locality inductive bias and long-range dependencies. Extensive experiments confirm our SOTA performance, achieving 93.87\% Dice on CAMUS and 92.62\% on EchoNet-Dynamic, with respective HD95 values of 3.29mm and 2.73mm.

cs.CV

Evidence for Deconfined Magnetic Order in the Kitaev-$J_3$ Model

We investigate the Kitaev-$J_3$ honeycomb model using variational Monte Carlo calculations combined with a vison-quasiparticle analysis of the parent Kitaev spin liquid (KSL). We provide evidence for deconfined magnetic phases in which zigzag or antiferromagnetic order coexists with remnant $\mathbb{Z}_2$ topological structure inherited from the KSL. The optimized variational wave functions retain multiple linearly independent topological sectors on a torus, whereas those of conventional ordered phases collapse to a single sector. The vison-quasiparticle analysis shows that magnetic order naturally arises from vison-pair condensation while single visons remain gapped, yielding a microscopic mechanism for magnetic ordering without immediate confinement. The resulting phases further host gapless spinons with multiple Majorana cones, offering a possible microscopic scenario for the anomalous low-temperature longitudinal thermal transport reported in magnetically ordered Kitaev materials such as Na$_2$Co$_2$TeO$_6$. Our results reveal a microscopic route to fractionalized magnetism beyond the conventional dichotomy between magnetic order and spin-liquid behavior.

cond-mat.str-el

LayersReg: A Layer-by-Layer Progressive Regressor for Reliable Intraoperative 3D/2D Registration

3D/2D registration serves as a cornerstone technique in surgical navigation. Traditional iterative optimization algorithms suffer from low efficiency and high failure rates in intraoperative settings. Deep learning-based methods reformulate registration from iterative optimization to a regression problem that maps image appearance features to spatial pose, typically achieving improved real-time performance and accuracy. However, such learnable methods are confined to memory-driven retrieval of specific pose features rather than understanding the task of image alignment itself, which limits their generalization in complex scenarios. We propose LayersReg, a pioneering regression paradigm that endows the model with 3D anatomical awareness and searches for the correct pose in a progressive, layer-by-layer manner. Inspired by the iterative pose-searching optimization criterion of classical registration, LayersReg searches for correlations between the moving and fixed images in feature space, capturing the trend of pixel flow and thereby converging iteratively toward the correct spatial pose transformation. We further design a coupling of node-wise regression with the progressive registration framework to enhance the model's perception of spatial pose changes. Experimental results demonstrate that under large offsets and multimodality conditions, LayersReg achieves high accuracy on both X-ray/CT registration (0.68°, 1.41 mm) and slice localization (0.73°, 1.55 mm) tasks, outperforming existing state-of-the-art methods while meeting the intraoperative demands for precision and real-time capability.

cs.CV

Getting There and Getting In: How Mobility and Sorting Keep Women out of Top Startup Accelerators

Startup accelerators are a leading gateway to venture capital, but top programs often require founders to relocate to a venture hub. From a hand-collected census of U.S. accelerator startups (2008-2011) followed for five years, we estimate a two-sided matching model that separates two channels behind the gender funding gap, geographic mobility and sorting across accelerator tiers. Women raise about 60% less than men over five years; the gap concentrates among non-relocating women, is largest at active-childrearing ages, and vanishes for relocators, while the mobility cost is near zero for men. Removing mobility frictions raises women's match quality but not their tier; reaching the high-funding top tier also requires removing the sorting disadvantage that women face. The 2012 JOBS Act eased the legal barrier and capacity grew tenfold, yet the U.S. VC dollar gap still tripled (2011-2020): closing it needs mobility, sorting, and capacity together.

econ.GN

DataClawBench: An Agent Benchmark for Exploratory Real-World Financial Data Analysis

Autonomous data analysis agents are increasingly expected to conduct exploratory analysis with limited human guidance about data. However, existing benchmarks typically evaluate such agents in prior-guided settings, providing selected data sources, explicit data schemas, or cleaned data, thereby understating the exploratory burden. To evaluate this realistic exploratory data analysis task, we introduce DataClawBench, a benchmark built from financial think-tank consulting scenarios where agents must independently explore unfamiliar, noisy, cross-domain data and produce verifiable conclusions. DataClawBench provides a unified real-world data environment with approximately 2.06 million records across enterprise, industry, and policy domains, with native data noise preserved. On top of this data environment, it defines 492 multi-step cross-domain tasks, each annotated with intermediate milestones that diagnose exploration and reasoning failures beyond outcome accuracy. A systematic evaluation of eight advanced LLMs under the OpenClaw agent reveals that exploratory data analysis breaks agent reliability: more exploration does not reliably translate into task-relevant progress or correct final answers.

cs.AI

NTIRE 2026 The Second Challenge on Day and Night Raindrop Removal for Dual-Focused Images: Methods and Results

This paper presents an overview of the NTIRE 2026 Second Challenge on Day and Night Raindrop Removal for Dual-Focused Images. Building upon the success of the first edition, this challenge attracted a wide range of impressive solutions, all developed and evaluated on our real-world Raindrop Clarity dataset~\cite{jin2024raindrop}. For this edition, we adjust the dataset with 14,139 images for training, 407 images for validation, and 593 images for testing. The primary goal of this challenge is to establish a strong and practical benchmark for the removal of raindrops under various illumination and focus conditions. In total, 168 teams have registered for the competition, and 17 teams submitted valid final solutions and fact sheets for the testing phase. The submitted methods achieved strong performance on the Raindrop Clarity dataset, demonstrating the growing progress in this challenging task.

cs.CV

SWE-CI: Evaluating Agent Capabilities in Maintaining Codebases via Continuous Integration

Large language model (LLM)-powered agents have demonstrated strong capabilities in automating software engineering tasks such as static bug fixing. However, in the real world, the development of mature software is typically predicated on complex requirement changes and long-term feature iterations -- a process that static, one-shot repair paradigms fail to capture. To bridge this gap, we propose SWE-CI, the first repository-level benchmark built upon the Continuous Integration loop, aiming to shift the evaluation paradigm for code generation from static, short-term functional correctness toward dynamic, long-term maintainability. The key insight is simple: Maintainability can be revealed by tracking how functional correctness changes over time. The benchmark comprises 100 tasks, each deriving from a real-world code repository with a development history spanning an average of 233 days and 71 consecutive commits. SWE-CI requires agents to systematically resolve these tasks through dozens of rounds of analysis and coding iterations. SWE-CI provides valuable insights into how well agents can sustain code quality throughout long-term evolution.

cs.SE

AIConfigurator: Lightning-Fast Configuration Optimization for Multi-Framework LLM Serving

Optimizing Large Language Model (LLM) inference in production systems is increasingly difficult due to dynamic workloads, stringent latency/throughput targets, and a rapidly expanding configuration space. This complexity spans not only distributed parallelism strategies (tensor/pipeline/expert) but also intricate framework-specific runtime parameters such as those concerning the enablement of CUDA graphs, available KV-cache memory fractions, and maximum token capacity, which drastically impact performance. The diversity of modern inference frameworks (e.g., TRT-LLM, vLLM, SGLang), each employing distinct kernels and execution policies, makes manual tuning both framework-specific and computationally prohibitive. We present AIConfigurator, a unified performance-modeling system that enables rapid, framework-agnostic inference configuration search without requiring GPU-based profiling. AIConfigurator combines (1) a methodology that decomposes inference into analytically modelable primitives - GEMM, attention, communication, and memory operations while capturing framework-specific scheduling dynamics; (2) a calibrated kernel-level performance database for these primitives across a wide range of hardware platforms and popular open-weights models (GPT-OSS, Qwen, DeepSeek, LLama, Mistral); and (3) an abstraction layer that automatically resolves optimal launch parameters for the target backend, seamlessly integrating into production-grade orchestration systems. Evaluation on production LLM serving workloads demonstrates that AIConfigurator identifies superior serving configurations that improve performance by up to 40% for dense models (e.g., Qwen3-32B) and 50% for MoE architectures (e.g., DeepSeek-V3), while completing searches within 30 seconds on average. Enabling the rapid exploration of vast design spaces - from cluster topology down to engine specific flags.

cs.LG

MemHunter: Automated and Verifiable Memorization Detection at Dataset-scale in LLMs

Large language models (LLMs) have been shown to memorize and reproduce content from their training data, raising significant privacy concerns, especially with web-scale datasets. Existing methods for detecting memorization are primarily sample-specific, relying on manually crafted or discretely optimized memory-inducing prompts generated on a per-sample basis, which become impractical for dataset-level detection due to the prohibitive computational cost of iterating through all samples. In real-world scenarios, data owners may need to verify whether a susceptible LLM has memorized their dataset, particularly if the LLM may have collected the data from the web without authorization. To address this, we introduce MemHunter, which trains a memory-inducing LLM and employs hypothesis testing to efficiently detect memorization at the dataset level, without requiring sample-specific memory inducing. Experiments on models like Pythia and Llama demonstrate that MemHunter can extract up to 40% more training data than existing methods under constrained time resources and reduce search time by up to 80% when integrated as a plug-in. Crucially, MemHunter is the first method capable of dataset-level memorization detection, providing a critical tool for assessing privacy risks in LLMs powered by large-scale datasets.

cs.CR

GI-NAS: Boosting Gradient Inversion Attacks Through Adaptive Neural Architecture Search

Gradient Inversion Attacks invert the transmitted gradients in Federated Learning (FL) systems to reconstruct the sensitive data of local clients and have raised considerable privacy concerns. A majority of gradient inversion methods rely heavily on explicit prior knowledge (e.g., a well pre-trained generative model), which is often unavailable in realistic scenarios. This is because real-world client data distributions are often highly heterogeneous, domain-specific, and unavailable to attackers, making it impractical for attackers to obtain perfectly matched pre-trained models, which inevitably suffer from fundamental distribution shifts relative to target private data. To alleviate this issue, researchers have proposed to leverage the implicit prior knowledge of an over-parameterized network. However, they only utilize a fixed neural architecture for all the attack settings. This would hinder the adaptive use of implicit architectural priors and consequently limit the generalizability. In this paper, we further exploit such implicit prior knowledge by proposing Gradient Inversion via Neural Architecture Search (GI-NAS), which adaptively searches the network and captures the implicit priors behind neural architectures. Extensive experiments verify that our proposed GI-NAS can achieve superior attack performance compared to state-of-the-art gradient inversion methods, even under more practical settings with high-resolution images, large-sized batches, and advanced defense strategies. To the best of our knowledge, we are the first to successfully introduce NAS to the gradient inversion community. We believe that this work exposes critical vulnerabilities in real-world federated learning by demonstrating high-fidelity reconstruction of sensitive data without requiring domain-specific priors, forcing urgent reassessment of FL privacy safeguards.

cs.AI

Composable Score-based Graph Diffusion Model for Multi-Conditional Molecular Generation

Controllable molecular graph generation is essential for material and drug discovery, where generated molecules must satisfy diverse property constraints. While recent advances in graph diffusion models have improved generation quality, their effectiveness in multi-conditional settings remains limited due to reliance on joint conditioning or continuous relaxations that compromise fidelity. To address these limitations, we propose Composable Score-based Graph Diffusion model (CSGD), the first model that extends score matching to discrete graphs via concrete scores, enabling flexible and principled manipulation of conditional guidance. Building on this foundation, we introduce two score-based techniques: Composable Guidance (CoG), which allows fine-grained control over arbitrary subsets of conditions during sampling, and Probability Calibration (PC), which adjusts estimated transition probabilities to mitigate train-test mismatches. Empirical results on four molecular datasets show that CSGD achieves state-of-the-art performance, with a 15.3% average improvement in controllability over prior methods, while maintaining high validity and distributional fidelity. Our findings highlight the practical advantages of score-based modeling for discrete graph generation and its capacity for flexible, multi-property molecular design.

cs.LG

CueGCL: Cluster-aware Personalized Self-Training for Unsupervised Graph Contrastive Learning

Recently, graph contrastive learning (GCL) has emerged as one of the optimal solutions for node-level and supervised tasks. However, for structure-related and unsupervised tasks such as graph clustering, current GCL algorithms face difficulties acquiring the necessary cluster-level information, resulting in poor performance. In addition, general unsupervised GCL improves the performance of downstream tasks by increasing the number of negative samples, which leads to severe class collision and unfairness of graph clustering. To address the above issues, we propose a Cluster-aware Graph Contrastive Learning Framework (CueGCL) to jointly learn clustering results and node representations. Specifically, we design a personalized self-training (PeST) strategy for unsupervised scenarios, which enables our model to capture precise cluster-level personalized information. With the benefit of the PeST, we alleviate class collision and unfairness without sacrificing the overall model performance. Furthermore, aligned graph clustering (AGC) is employed to obtain the cluster partition, where we align the clustering space of our downstream task with that in PeST to achieve more consistent node embeddings. Finally, we theoretically demonstrate the effectiveness of our model, showing it yields an embedding space with a significantly discernible cluster structure. Extensive experimental results also show our CueGCL exhibits state-of-the-art performance on five benchmark datasets with different scales.

cs.SI

ParaAegis: Parallel Protection for Flexible Privacy-preserved Federated Learning

Federated learning (FL) faces a critical dilemma: existing protection mechanisms like differential privacy (DP) and homomorphic encryption (HE) enforce a rigid trade-off, forcing a choice between model utility and computational efficiency. This lack of flexibility hinders the practical implementation. To address this, we introduce ParaAegis, a parallel protection framework designed to give practitioners flexible control over the privacy-utility-efficiency balance. Our core innovation is a strategic model partitioning scheme. By applying lightweight DP to the less critical, low norm portion of the model while protecting the remainder with HE, we create a tunable system. A distributed voting mechanism ensures consensus on this partitioning. Theoretical analysis confirms the adjustments between efficiency and utility with the same privacy. Crucially, the experimental results demonstrate that by adjusting the hyperparameters, our method enables flexible prioritization between model accuracy and training time.

cs.LG

Anyon polarons as a window into the competing phases of the Kitaev-Gamma-Gamma' model

We investigate the dispersions of anyon quasi-particles in the Kitaev honeycomb spin-liquid perturbed by $Γ$ and $Γ'$ couplings in order to understand phase transitions into competing states through anyon gap-closing instabilities. We demonstrate how anyon gap closings allow to understand phase transitions into a plethora of previously identified competing phases -- including zigzag, stripy, $120^\circ$, and incommensurate spiral phases -- and are in agreement with numerical studies not only on the nature of the phases, but also on the specific critical values of $Γ$ and $Γ'$ couplings. Remarkably, when the anti-ferromagnetic Kitaev model is perturbed by an ferromagnetic $Γ$ interaction, we find that the single-vison and fermion gaps remain open while the gap of a magnon-like local boson vanishes, implying that the resulting state has coexistence of a spontaneous broken symmetry and the fractionalization pattern of the Kitaev spin liquid. The magnetic long-range order could be either a stripy antiferromagnet or an incommensurate spiral, depending on the sign of $Γ'$.

cond-mat.str-el

Universal non-Hermitian transport in disordered systems

In disordered Hermitian systems, localization of energy eigenstates prohibits wave propagation. In non-Hermitian systems, however, wave propagation is possible even when the eigenstates of Hamiltonian are exponentially localized by disorders. We find in this regime that non-Hermitian wave propagation exhibits novel universal scaling behaviors without Hermitian counterpart. Furthermore, our theory demonstrates how the tail of imaginary-part density of states dictates wave propagation in the long-time limit. Specifically, for the three typical classes, namely the Gaussian, the uniform, and the linear imaginary-part density of states, we obtain logarithmically suppressed sub-ballistic transport, and two types of subdiffusion with exponents that depend only on spatial dimensions, respectively. Our work highlights the fundamental differences between Hermitian and non-Hermitian Anderson localization, and uncovers unique universality in non-Hermitian wave propagation.

quant-ph

Anyon polarons as a window into competing phases of the Kitaev honeycomb model under a Zeeman field

We compute the spectra of anyon quasiparticles in all three super-selection sectors of the Kitaev model (i.e., visons, fermions and bosons), perturbed by a Zeeman field away from its exactly solvable limit, to gain insights on the competition of its non-abelian spin-liquid with other nearby phases, such as the mysterious intermediate state observed in the antiferromagnetic model. Both for the ferro- and antiferro-magnetic models we find that the fermions and visons become gapless at nearly identical critical Zeeman couplings. In the ferromagnetic model this is consistent with a direct transition into a polarized state. In the anti-ferromagnetic model this implies that previous theories of the intermediate phase viewed as a spin liquid with a different fermion Chern number are inadequate, as they presume that the vison gap does not close. In the antiferromagnetic model we also find that a bosonic quasiparticle becomes gapless at nearly the same critical field as the fermions and visons. This boson carries the quantum numbers of an anti-ferromagnetic order parameter, suggesting that the intermediate phase has spontaneously broken symmetry with this order.

cond-mat.str-el