arXiv ScienceSearch

arXiv subjects

Chuang Chen

Publications and source records attributed to Chuang Chen.

At least 19 recordsLinked to original sources

Higher-Order Topological States with Cleavage-Dependent Dirac Mass

Topological quantum chemistry based on local charge profiles lacks predictive power for the crystalline cleavage of higher-order topological insulators (HOTIs). By cleaving an obstructed atomic insulator, we discover a topological phase characterized by e/2-fractional charges localized at precisely half of the corners, while the remaining empty corners host complementary vacancies of interstice charge. These zero-energy charge vacancies and topological corners form a spatially balanced geometry, separately localized at four corners. Crucially, we demonstrate that the emergence of corner zero modes dictates that specific dangling bonds acting as the mass of a Dirac fermion-must explicitly expose in, and subtly slope toward, the corner regions. This strict directionality is verified by the anisotropic evolution of the mass term within a (2+1) dimensional parameter space. Moreover, we find that the topological corners acquire lower entanglement entropy compared to the bulk, a behavior opposite to that of the real-space energy distribution which forms an energy-entropy compensation, essentially derived from the topological charge compensation. Our work paves the way for the local chemical environment at topological boundaries, and demonstrates the higher order quantum transport counterparts for high energy Dirac physics.

cond-mat.mes-hall

Deconfined criticality between an antiferromagnetic insulator and a nodal d-wave superconductor: a quantum Monte Carlo study

We present a quantum Monte Carlo study of the transition between the insulating N\'eel state and the nodal $d$-wave superconductor on the square lattice at half-filling. We access a regime of frustrated magnetic order without a sign problem using a parton representation of the electron in terms of fermionic spinons and bosonic chargons. Both partons move in a background $\pi$-flux (so the electron experiences no net flux) and are coupled to a quantum fluctuating SU(2) lattice gauge field. In contrast to earlier studies directly on the electronic degrees of freedom, we find evidence for a second-order deconfined quantum phase transition at which both the N\'eel and $d$-wave superconductivity orders vanish continuously. We compute correlators of the spinon-chargon composite with the same quantum numbers as the electron: we find a gapless Dirac dispersion inside the $d$-wave superconductor, turning into a gapped dispersion in the antiferromagnet.

cond-mat.str-el

Deep Joint Source-Channel Coding for Wireless Video Transmission with Asymmetric Context

In this paper, we propose a high-efficiency deep joint source-channel coding (JSCC) method for video transmission based on conditional coding with asymmetric context. The conditional coding-based neural video compression requires to predict the encoding and decoding conditions from the same context which includes the same reconstructed frames. However in JSCC schemes which fall into pseudo-analog transmission, the encoder cannot infer the same reconstructed frames as the decoder even a pipeline of the simulated transmission is constructed at the encoder. In the proposed method, without such a pipeline, we guide and design neural networks to learn encoding and decoding conditions from asymmetric contexts. Additionally, we introduce feature propagation, which allows intermediate features to be independently propagated at the encoder and decoder and help to generate conditions, enabling the framework to greatly leverage temporal correlation while mitigating the problem of error accumulation. To further exploit the performance of the proposed transmission framework, we implement content-adaptive coding which achieves variable bandwidth transmission using entropy models and masking mechanisms. Experimental results demonstrate that our method outperforms existing deep video transmission frameworks in terms of performance and effectively mitigates the error accumulation. By mitigating the error accumulation, our schemes can reduce the frequency of inserting intra-frame coding modes, further enhancing performance.

eess.IV

DAGLFNet: Deep Feature Attention Guided Global and Local Feature Fusion for Pseudo-Image Point Cloud Segmentation

Environmental perception systems are crucial for high-precision mapping and autonomous navigation, with LiDAR serving as a core sensor providing accurate 3D point cloud data. Efficiently processing unstructured point clouds while extracting structured semantic information remains a significant challenge. In recent years, numerous pseudo-image-based representation methods have emerged to balance efficiency and performance by fusing 3D point clouds with 2D grids. However, the fundamental inconsistency between the pseudo-image representation and the original 3D information critically undermines 2D-3D feature fusion, posing a primary obstacle for coherent information fusion and leading to poor feature discriminability. This work proposes DAGLFNet, a pseudo-image-based semantic segmentation framework designed to extract discriminative features. It incorporates three key components: first, a Global-Local Feature Fusion Encoding (GL-FFE) module to enhance intra-set local feature correlation and capture global contextual information; second, a Multi-Branch Feature Extraction (MB-FE) network to capture richer neighborhood information and improve the discriminability of contour features; and third, a Feature Fusion via Deep Feature-guided Attention (FFDFA) mechanism to refine cross-channel feature fusion precision. Experimental evaluations demonstrate that DAGLFNet achieves mean Intersection-over-Union (mIoU) scores of 69.9% and 78.7% on the validation sets of SemanticKITTI and nuScenes, respectively. The method achieves an excellent balance between accuracy and efficiency.

cs.CV

Scalable hybrid quantum Monte Carlo simulation of U(1) gauge field coupled to fermions on GPU

We develop a GPU-accelerated hybrid quantum Monte Carlo (QMC) algorithm to solve the fundamental yet difficult problem of $U(1)$ gauge field coupled to fermions, which gives rise to a $U(1)$ Dirac spin liquid state under the description of (2+1)d quantum electrodynamics QED$_3$. The algorithm renders a good acceptance rate and, more importantly, nearly linear space-time volume scaling in computational complexity $O(N_{\tau} V_s)$, where $N_\tau$ is the imaginary time dimension and $V_s$ is spatial volume, which is much more efficient than determinant QMC with scaling behavior of $O(N_\tau V_s^3)$. Such acceleration is achieved via a collection of technical improvements, including (i) the design of the efficient problem-specific preconditioner, (ii) customized CUDA kernel for matrix-vector multiplication, and (iii) CUDA Graph implementation on the GPU. These advances allow us to simulate the $U(1)$ Dirac spin liquid state with unprecedentedly large system sizes, which is up to $N_\tau\times L\times L = 660\times66\times66$, and reveal its novel properties. With these technical improvements, we see the asymptotic convergence in the scaling dimensions of various fermion bilinear operators and the conserved current operator when approaching the thermodynamic limit. The scaling dimensions find good agreement with field-theoretical expectation, which provides supporting evidence for the conformal nature of the $U(1)$ Dirac spin liquid state in the QED$_3$. Our technical advancements open an avenue to study the Dirac spin liquid state and its transition towards symmetry-breaking phases at larger system sizes and with less computational burden.

cond-mat.str-el

Emergent gauge flux in mixed QED$_3$ with flavor chemical potential: application to magnetized U(1) Dirac spin liquids

We design a lattice model of a "mixed" U(1) gauge field coupled to fermions with a flavor chemical potential and solve it with large-scale determinant quantum Monte Carlo simulations, For zero flavor chemical potential, the model realizes three-dimensional quantum electrodynamics (QED$_3$) which has been argued to describe the ground state and low-energy excitations of the Dirac spin liquid phase of quantum antiferromagnets. At finite flavor chemical potential, corresponding to a Zeeman field perturbing the Dirac spin liquid, we find a "chiral flux" phase which is characterized by the generation of a finite mean emergent gauge flux and, accordingly, the formation of relativistic Landau levels for the Dirac fermions. In this state, the U(1)$_m$ magnetic symmetry is spontaneously broken, leading to a gapless free photon mode which, due to spin-flux-attachment, is observable in the longitudinal spin structure factor. We numerically compute longitudinal and transverse spin structure factors which match our continuum and lattice mean-field theory predictions. In a different region of the phase diagram, strong fluctuations of the emergent gauge field give rise to an antiferromagnetically ordered state with gapped Dirac fermions coexisting with a deconfined gauge field. We also find an interesting intermediate phase where the chiral flux phase and the antiferromagnetic phase coexist. We argue that our results pave the way to testable predictions for magnetized Dirac spin liquids in frustrated quantum antiferromagnets.

cond-mat.str-el

MoGaFace: Momentum-Guided and Texture-Aware Gaussian Avatars for Consistent Facial Geometry

Existing 3D head avatar reconstruction methods adopt a two-stage process, relying on tracked FLAME meshes derived from facial landmarks, followed by Gaussian-based rendering. However, misalignment between the estimated mesh and target images often leads to suboptimal rendering quality and loss of fine visual details. In this paper, we present MoGaFace, a novel 3D head avatar modeling framework that continuously refines facial geometry and texture attributes throughout the Gaussian rendering process. To address the misalignment between estimated FLAME meshes and target images, we introduce the Momentum-Guided Consistent Geometry module, which incorporates a momentum-updated expression bank and an expression-aware correction mechanism to ensure temporal and multi-view consistency. Additionally, we propose Latent Texture Attention, which encodes compact multi-view features into head-aware representations, enabling geometry-aware texture refinement via integration into Gaussians. Extensive experiments show that MoGaFace achieves high-fidelity head avatar reconstruction and significantly improves novel-view synthesis quality, even under inaccurate mesh initialization and unconstrained real-world settings.

cs.CV

SRMambaV2: Biomimetic Attention for Sparse Point Cloud Upsampling in Autonomous Driving

Upsampling LiDAR point clouds in autonomous driving scenarios remains a significant challenge due to the inherent sparsity and complex 3D structures of the data. Recent studies have attempted to address this problem by converting the complex 3D spatial scenes into 2D image super-resolution tasks. However, due to the sparse and blurry feature representation of range images, accurately reconstructing detailed and complex spatial topologies remains a major difficulty. To tackle this, we propose a novel sparse point cloud upsampling method named SRMambaV2, which enhances the upsampling accuracy in long-range sparse regions while preserving the overall geometric reconstruction quality. Specifically, inspired by human driver visual perception, we design a biomimetic 2D selective scanning self-attention (2DSSA) mechanism to model the feature distribution in distant sparse areas. Meanwhile, we introduce a dual-branch network architecture to enhance the representation of sparse features. In addition, we introduce a progressive adaptive loss (PAL) function to further refine the reconstruction of fine-grained details during the upsampling process. Experimental results demonstrate that SRMambaV2 achieves superior performance in both qualitative and quantitative evaluations, highlighting its effectiveness and practical value in automotive sparse point cloud upsampling tasks.

cs.CV

Self-learning Monte Carlo Method: A Review

The Self-Learning Monte Carlo (SLMC) method is a Monte Carlo approach that has emerged in recent years by integrating concepts from machine learning with conventional Monte Carlo techniques. Designed to accelerate the numerical study of interacting many-body systems, SLMC significantly improves sampling efficiency by constructing an effective model -- via machine learning methods -- based on configurations generated by conventional Monte Carlo methods and then proposes global updates based on the effective model. This enhancement leads to a substantial reduction in autocorrelation time, especially near the critical region, where traditional methods typically suffer from critical slowing down and increased computation complexity. Moreover, SLMC maintains statistical accuracy by implementing a cumulative update scheme that rigorously satisfies the detailed balance condition. And more recent applications have extended the SLMC to convolutional neural networks with applications not only in condensed matter physics but also high-energy physics, quantum chemistry, and quantum simulations. The generic applicability and high computational efficiency make SLMC a powerful and scalable framework for quantum Monte Carlo simulations of strongly correlated electron systems, extending the reach of numerical investigations beyond the limitations of conventional techniques.

cond-mat.str-el

Arcturus: A Cloud Overlay Network for Global Accelerator with Enhanced Performance and Stability

Global Accelerator (GA) services play a vital role in ensuring low-latency, high-reliability communication for real-time interactive applications. However, existing GA offerings are tightly bound to specific cloud providers, resulting in high costs, rigid deployment, and limited flexibility, especially for large-scale or budget-sensitive deployments. Arcturus is a cloud-native GA framework that revisits the design of GA systems by leveraging low-cost, heterogeneous cloud resources across multiple providers. Rather than relying on fixed, high-end infrastructure, Arcturus dynamically constructs its acceleration network and balances performance, stability, and resource efficiency. To achieve this, Arcturus introduces a two-plane design: a forwarding plane that builds a proxy network with adaptive control, and a scheduling plane that coordinates load and routing through lightweight, quantitative optimization. Evaluations under millions of RPS show that Arcturus outperforms commercial GA services by up to 1.7X in acceleration performance, reduces cost by 71%, and maintains over 80% resource efficiency--demonstrating efficient use of cloud resources at scale.

cs.NI

Interactive Drawing Guidance for Anime Illustrations with Diffusion Model

Creating high-quality anime illustrations presents notable challenges, particularly for beginners, due to the intricate styles and fine details inherent in anime art. We present an interactive drawing guidance system specifically designed for anime illustrations to address this issue. It offers real-time guidance to help users refine their work and streamline the creative process. Our system is built upon the StreamDiffusion pipeline to deliver real-time drawing assistance. We fine-tune Stable Diffusion with LoRA to synthesize anime style RGB images from user-provided hand-drawn sketches and prompts. Leveraging the Informative Drawings model, we transform these RGB images into rough sketches, which are further refined into structured guidance sketches using a custom-designed optimizer. The proposed system offers precise, real-time guidance aligned with the creative intent of the user, significantly enhancing both the efficiency and accuracy of the drawing process. To assess the effectiveness of our approach, we conducted a user study, gathering empirical feedback on both system performance and interface usability.

cs.GR

SRMamba: Mamba for Super-Resolution of LiDAR Point Clouds

In recent years, range-view-based LiDAR point cloud super-resolution techniques attract significant attention as a low-cost method for generating higher-resolution point cloud data. However, due to the sparsity and irregular structure of LiDAR point clouds, the point cloud super-resolution problem remains a challenging topic, especially for point cloud upsampling under novel views. In this paper, we propose SRMamba, a novel method for super-resolution of LiDAR point clouds in sparse scenes, addressing the key challenge of recovering the 3D spatial structure of point clouds from novel views. Specifically, we implement projection technique based on Hough Voting and Hole Compensation strategy to eliminate horizontally linear holes in range image. To improve the establishment of long-distance dependencies and to focus on potential geometric features in vertical 3D space, we employ Visual State Space model and Multi-Directional Scanning mechanism to mitigate the loss of 3D spatial structural information due to the range image. Additionally, an asymmetric U-Net network adapts to the input characteristics of LiDARs with different beam counts, enabling super-resolution reconstruction for multi-beam point clouds. We conduct a series of experiments on multiple challenging public LiDAR datasets (SemanticKITTI and nuScenes), and SRMamba demonstrates significant superiority over other algorithms in both qualitative and quantitative evaluations.

cs.CV

A Diff-Attention Aware State Space Fusion Model for Remote Sensing Classification

Multispectral (MS) and panchromatic (PAN) images describe the same land surface, so these images not only have their own advantages, but also have a lot of similar information. In order to separate these similar information and their respective advantages, reduce the feature redundancy in the fusion stage. This paper introduces a diff-attention aware state space fusion model (DAS2F-Model) for multimodal remote sensing image classification. Based on the selective state space model, a cross-modal diff-attention module (CMDA-Module) is designed to extract and separate the common features and their respective dominant features of MS and PAN images. Among this, space preserving visual mamba (SPVM) retains image spatial features and captures local features by optimizing visual mamba's input reasonably. Considering that features in the fusion stage will have large semantic differences after feature separation and simple fusion operations struggle to effectively integrate these significantly different features, an attention-aware linear fusion module (AALF-Module) is proposed. It performs pixel-wise linear fusion by calculating influence coefficients. This mechanism can fuse features with large semantic differences while keeping the feature size unchanged. Empirical evaluations indicate that the presented method achieves better results than alternative approaches. The relevant code can be found at:https://github.com/AVKSKVL/DAS-F-Model

cs.CV

UniEmoX: Cross-modal Semantic-Guided Large-Scale Pretraining for Universal Scene Emotion Perception

Visual emotion analysis holds significant research value in both computer vision and psychology. However, existing methods for visual emotion analysis suffer from limited generalizability due to the ambiguity of emotion perception and the diversity of data scenarios. To tackle this issue, we introduce UniEmoX, a cross-modal semantic-guided large-scale pretraining framework. Inspired by psychological research emphasizing the inseparability of the emotional exploration process from the interaction between individuals and their environment, UniEmoX integrates scene-centric and person-centric low-level image spatial structural information, aiming to derive more nuanced and discriminative emotional representations. By exploiting the similarity between paired and unpaired image-text samples, UniEmoX distills rich semantic knowledge from the CLIP model to enhance emotional embedding representations more effectively. To the best of our knowledge, this is the first large-scale pretraining framework that integrates psychological theories with contemporary contrastive learning and masked image modeling techniques for emotion analysis across diverse scenarios. Additionally, we develop a visual emotional dataset titled Emo8. Emo8 samples cover a range of domains, including cartoon, natural, realistic, science fiction and advertising cover styles, covering nearly all common emotional scenes. Comprehensive experiments conducted on six benchmark datasets across two downstream tasks validate the effectiveness of UniEmoX. The source code is available at https://github.com/chincharles/u-emo.

cs.AI

Universal collective Larmor-Silin mode emerging in magnetized correlated Dirac fermions

Employing large-scale quantum Monte Carlo simulations, we find that in the magnetized interacting Dirac fermion model there emerges a universal collective Larmor-Silin spin wave mode in the transverse dynamical spin susceptibility. Such mode purely originates from the interaction among Dirac fermions and distinguishes itself from the usual particle-hole continuum with finite lifetime and clear dispersion, both at small and large momenta in a large portion of the Brillouin zone. Our unbiased numerical results offer the dynamic signature of this collective excitation in interacting Dirac fermion systems, and provide experimental guidance for inelastic neutron scattering, electron spin resonance, and other spectroscopic approaches in the investigation of such universal collective modes in quantum Moire materials, topological insulators, and quantum spin liquid materials under magnetic field, with quintessential interaction nature beyond the commonly assumed noninteracting Dirac fermion or spinon approximations.

cond-mat.str-el

Mutilmodal Feature Extraction and Attention-based Fusion for Emotion Estimation in Videos

The continuous improvement of human-computer interaction technology makes it possible to compute emotions. In this paper, we introduce our submission to the CVPR 2023 Competition on Affective Behavior Analysis in-the-wild (ABAW). Sentiment analysis in human-computer interaction should, as far as possible Start with multiple dimensions, fill in the single imperfect emotion channel, and finally determine the emotion tendency by fitting multiple results. Therefore, We exploited multimodal features extracted from video of different lengths from the competition dataset, including audio, pose and images. Well-informed emotion representations drive us to propose a Attention-based multimodal framework for emotion estimation. Our system achieves the performance of 0.361 on the validation dataset. The code is available at [https://github.com/xkwangcn/ABAW-5th-RT-IAI].

cs.CV

Quantum Many-Body Simulations of the 2D Fermi-Hubbard Model in Ultracold Optical Lattices

Understanding quantum many-body states of correlated electrons is one main theme in modern condensed matter physics. Given that the Fermi-Hubbard model, the prototype of correlated electrons, has been recently realized in ultracold optical lattices, it is highly desirable to have controlled numerical methodology to provide precise finite-temperature results upon doping, to directly compare with experiments. Here, we demonstrate the exponential tensor renormalization group (XTRG) algorithm [Phys. Rev. X 8, 031082 (2018)], complemented with independent determinant quantum Monte Carlo (DQMC) offer a powerful combination of tools for this purpose. XTRG provides full and accurate access to the density matrix and thus various spin and charge correlations, down to unprecedented low temperature of few percents of the fermion tunneling energy scale. We observe excellent agreement with ultracold fermion measurements at both half-filling and finite-doping, including the sign-reversal behavior in spin correlations due to formation of magnetic polarons, and the attractive hole-doublon and repulsive hole-hole pairs that are responsible for the peculiar bunching and antibunching behavior of the antimoments.

cond-mat.str-el

Fermi arcs and pseudogap in a lattice model of a doped orthogonal metal

Since the discovery of the pseudogap and Fermi arc states in underdoped cuprates, the understanding of such non-Fermi-liquid states and the associated violation of Luttinger's theorem have been the central theme in correlated electron systems. However, still lacking is a well-accepted theoretical framework to unambiguously explain these metallic states that are clearly beyond Landau's Fermi liquid and Luttinger's theorem of a Fermi surface and electron filling. Here, we design a lattice model of orthogonal metals with fermion and Ising matter fields coupled to topological order and, by solving the model via unbiased quantum Monte Carlo simulation at generic electron fillings, find that the system gives birth to phenomena of the Fermi arc and pseudogap in the single-particle spectrum that go beyond the Luttinger sum rule with broken Fermi surface but no symmetry breaking. The pseudogap and Fermi arcs coexist with a background of a deconfined Z2 gauge field, and we further find that the confinement transition of the gauge field triggers a superconductivity instability and that the hopping of the gauge-neutral fermions brings the "large" Fermi surface back from the Fermi arc state. Our unbiased numerical results provide a concrete model realization and theoretical framework for the coupling between gauge field and fermions and, in the process, generate the rich phenomena of the pseudogap, the Fermi arc, and superconductivity in generic correlated electron systems.

cond-mat.str-el