arXiv ScienceSearch

arXiv subjects

Hong Zhao

Publications and source records attributed to Hong Zhao.

At least 19 recordsLinked to original sources

Gaussian Linear Functional Manifold Method for Massive Point Cloud Data

Reconstructing continuous terrain manifolds from massive, unstructured airborne LiDAR point clouds remains challenging in complex Wildland-Urban Interface (WUI) environments, where deep neural networks require costly point-wise annotations and nonparametric surface reconstruction methods often lack structural interpretability. This paper introduces the Gaussian Linear Functional Manifold (GLFM), a physics-informed statistical framework that represents continuous surface topography using deterministic linear functional bases while modeling microscale diffuse laser backscatter as an isotropic Gaussian process. To avoid the quadratic computational cost of exact constrained maximum likelihood estimation, we develop an algebraic singular value decomposition (SVD) rank-reduction algorithm that enables linear-time parameter estimation and closed-form quadric classification. Evaluated on 35.2 km^2 of real-world aerial LiDAR data, GLFM automatically filters ground points and extracts morphological features, achieving an adjusted Rand index (ARI) of 0.9933 against field-verified ground truth and outperforming four leading baselines while maintaining an out-of-core memory footprint. The framework provides a rigorous, interpretable, and scalable foundation for large-scale point cloud analytics.

cs.CV

Equilibrium Distributions for Strongly Nonlinear Many-Body Systems

Obtaining equilibrium distributions of nonlinear systems is essential for accurately computing macroscopic observables. Conventional theoretical corrections are typically limited to weak nonlinearities, where interaction terms can be treated as effectively uncorrelated perturbations and the random phase approximation applies. In this Letter, we develop a framework to determine equilibrium distributions based on the generalized energy equipartition principle. Our approach recovers existing corrections in the weakly nonlinear regime and, crucially, remains valid for strong nonlinearities, where perturbative contributions become correlated and conventional approaches break down. Numerical simulations of the nonlinear Schr\"odinger equation, the Majda-McLaughlin-Tabak model, and the Fermi-Pasta-Ulam-Tsingou model demonstrate accurate corrections for nonlinearities more than an order of magnitude stronger than those accessible to conventional theories.

cond-mat.stat-mech

Exact Resonances Are Not Sufficient for Phonon Energy Diffusion

Multi-phonon resonance conditions underpin kinetic theories of phonon transport and lattice thermalization. We show that exact resonance matching, nonzero interaction coefficients, and network connectivity do not guarantee persistent energy diffusion. Symmetry-enforced balance relations drive exact-resonant collision currents to nonthermal zero-flux states, producing kinetic arrest from individual resonant sets to connected networks. Complete energy spreading is sustained by quasi-resonances. The thermodynamic and weak-nonlinearity limits do not commute: the leading kinetic behavior is recovered in the former, whereas at fixed finite size the thermalization time diverges through higher-order crossovers as the nonlinearity vanishes. Exact-resonance existence and connectivity are therefore kinematic, not sufficient dynamical, criteria for phonon energy diffusion.

cond-mat.stat-mech

Beyond Backpropagation: Monte Carlo Method Can Train Deep Neural Networks

Backpropagation (BP) dominates deep learning training, but its reliance on gradients brings inherent troubles -- vanishing and exploding gradients. The pursuit of gradient-free methods has long been a goal in the field of artificial intelligence. This paper shows that indeed the simplest Monte Carlo algorithm implemented on a single GPU -- randomly mutate a parameter, keep it if the loss decreases, otherwise retry -- can practically train deep networks. This gradient-free method does not even need common techniques such as batch normalization or residual connections to directly train sufficiently deep networks. More remarkably, its flexibility extends to several nontrivial scenarios: it enables pure pruning training, supports discrete weights, accommodates unconventional transfer functions such as Gaussian, and reveals the substantial redundancy of deep networks. We have demonstrated its feasibility on deep networks with more than 20 layers, single-hidden-layer wide networks with up to 16,384 hidden neurons, and even a simple Transformer architecture trained on both image classification (MNIST) and character-level language modeling (Tiny Shakespeare). This simple gradient-free method may offer a complementary perspective for understanding the self-organization and learning mechanisms of neural networks, and also provides an alternative route for building physically inspired deep learning systems.

cs.LG

CodeCytos: AI-assisted spatial molecular imaging analysis via code-augmented agent action space

Conventional tissue image analysis software provides foundational capabilities for cellular analysis, including segmentation, basic morphological feature extraction, and spatial organization analysis. However, these tools often require manual intervention and are not well integrated with code-driven automation, limiting efficiency and scalability for complex spatial tissue studies. In addition, they offer limited flexibility for custom analyses, as they typically support only a fixed set of pre-implemented spatial cellular features. To address these limitations, we propose CodeCytos, a coding-based reasoning agent framework that enables dynamic, programmable interaction with spatial molecular imaging data to improve automation and customization. CodeCytos is designed to streamline the exploration of custom spatial cellular features and adapt to diverse research needs. We demonstrate its utility through case studies on four expert-curated datasets from distinct tissue types: frontal cortex, non-small-cell lung cancer, pancreas, and tonsil. We evaluate CodeCytos under a realistic minimal prompt setting, where bioscientists pose simple questions without task-specific instructions or contextual information about spatial cellular analysis, and benchmark multiple LLM backbones with strong coding capabilities. We further show that incorporating tailored, domain-agnostic few-shot in-context coding-reasoning examples (randomly sampled demonstrations outside the spatial analysis domain) can substantially improve performance without requiring costly, expert-crafted in-domain demonstrations. Overall, CodeCytos outperforms baseline approaches, highlighting the potential of code-action agents to assist with custom feature exploration in spatial molecular imaging and to accelerate biomarker discovery.

cs.CV

Boltzmann Distribution from Invariance of Coarse-Graining-Scale and Energy-Shift

We present a concise derivation of the Boltzmann form for single-particle energy distributions in classical many-body Hamiltonian systems. The derivation relies on two physical facts: coarse-graining-scale invariance of the empirical distribution and invariance under a uniform shift of the energy zero. These conditions uniquely yield the Boltzmann factor, whose parameter is fixed by the mean energy per particle. For separable Hamiltonians, the equilibrium weight factorizes into kinetic and configurational contributions sharing the same parameter, identified from the kinetic part as the inverse kinetic temperature. The principle extends to any physical quantity with a stationary distribution and translational invariance. It is illustrated in a one-dimensional diatomic hard-core gas and a nonlinear lattice chain, where it predicts velocity, energy, spacing, collision-time, and pressure-dependent displacement distributions in agreement with simulations. The lattice model further shows how harmonic elasticity, anharmonic corrections, internal pressure, and thermal expansion emerge from the same exponential equilibrium weights. Finally, the relationships among different ensembles are briefly discussed.

cond-mat.stat-mech

The Fermi-Pasta-Ulam-Tsingou problem after 70 years: toward universal laws of thermalization in lattice systems in the thermodynamic limit

The Fermi--Pasta--Ulam--Tsingou (FPUT) problem provides a paradigmatic framework for understanding thermalization in weakly nonlinear many-body Hamiltonian systems. This focused review summarizes major developments in near-integrable dynamics, wave resonances, and phonon kinetic theory, emphasizing recent results on thermalization-time scaling in nonlinear lattices. A coherent picture emerges in the thermodynamic limit based on the eigenmode properties of an appropriate integrable reference system. If the reference system has extended eigenmodes and the leading resonant or quasi-resonant processes form a sufficiently connected network, the thermalization time follows $T_{\mathrm{eq}}\propto g^{-2}$, where $g$ measures the effective deviation from integrability. This scaling is found broadly in ordered and weakly disordered lattices and is robust to dimensionality, interaction potential, integrability-breaking mechanism, and multimode initial conditions. Identifying the correct integrable reference is essential; in some one-dimensional FPUT-type lattices with cubic interactions, the nearby Toda lattice, rather than the harmonic chain, provides the proper reference. If all reference eigenmodes are localized, spatial-overlap constraints progressively fragment low-order resonance networks as $g$ decreases, and thermalization becomes controlled by higher-order processes. Numerical studies reveal successive regimes $T_{\mathrm{eq}}\propto g^{-\gamma}$ with $\gamma=2,4,6$, together with weak system-size dependence. Whether this hierarchy persists asymptotically and whether a finite thermalization threshold exists remain open questions. We also discuss finite-size effects, strongly nonintegrable dynamics, heat transport, localization, and higher-dimensional lattices.

cond-mat.stat-mech

From Near-Integrable to Far-from-Integrable: A Unified Picture of Thermalization and Heat Transport

Whether and how a system approaches equilibrium is central in nonequilibrium statistical physics, crucial to understanding thermalization and transport. Bogoliubov's three-stage (initial, kinetic, and hydrodynamic) evolution hypothesis offers a qualitative framework, but quantitative progress has focused on near-integrable systems like dilute gases. In this work, we investigate the relaxation dynamics of a one-dimensional diatomic hard-point (DHP) gas, presenting a phase diagram that characterizes relaxation behavior across the full parameter space, from near-integrable to far-from-integrable regimes. We analyze thermalization (local energy relaxation in nonequilibrium states) and identify three universal dynamical regimes: (i) In the near-integrable regime, kinetic processes dominate, local energy relaxation decays exponentially, and the thermalization time $\tau$ scales as $\tau \propto \delta^{-2}$. (ii) In the far-from-integrable regime, hydrodynamic effects dominate, energy relaxation decays power-law, and thermalization time scales linearly with system size $N$. (iii) In the intermediate regime, the Bogoliubov phase emerges, characterized by the transition from kinetic to hydrodynamic relaxation. The phase diagram also shows that hydrodynamic behavior can emerge in small systems when sufficiently far from the integrable regime, challenging the view that such effects occur only in large systems. In the thermodynamic limit, the system's relaxation depends on the order in which the limits ($N \to \infty$ or $\delta \to 0$) are taken. We then analyze heat transport (decay of heat-current fluctuations in equilibrium), demonstrating its consistency with thermalization, leading to a unified theoretical description of thermalization and transport. Our approach provides a pathway for studying relaxation dynamics in many-body systems, including quantum systems.

cond-mat.stat-mech

Dual-Head Physics-Informed Graph Decision Transformer for Distribution System Restoration

Driven by recent advances in sensing and computing, deep reinforcement learning (DRL) technologies have shown great potential for addressing distribution system restoration (DSR) under uncertainty. However, their data-intensive nature and reliance on the Markov Decision Process (MDP) assumption limit their ability to handle scenarios that require long-term temporal dependencies or few-shot and zero-shot decision making. Emerging Decision Transformers (DTs), which leverage causal transformers for sequence modeling in DRL tasks, offer a promising alternative. However, their reliance on return-to-go (RTG) cloning and limited generalization capacity restricts their effectiveness in dynamic power system environments. To address these challenges, we introduce an innovative Dual-Head Physics-informed Graph Decision Transformer (DH-PGDT) that integrates physical modeling, structural reasoning, and subgoal-based guidance to enable scalable and robust DSR even in zero-shot or few-shot scenarios. DH-PGDT features a dual-head physics-informed causal transformer architecture comprising Guidance Head, which generates subgoal representations, and Action Head, which uses these subgoals to generate actions independently of RTG. It also incorporates an operational constraint-aware graph reasoning module that encodes power system topology and operational constraints to generate a confidence-weighted action vector for refining DT trajectories. This design effectively improves generalization and enables robust adaptation to unseen scenarios. While this work focuses on DSR, the underlying computing model of the proposed PGDT is broadly applicable to sequential decision making across various power system operations and other complex engineering domains.

eess.SY

The equilibrium distribution function for strongly nonlinear systems

The equilibrium distribution function determines macroscopic observables in statistical physics. While conventional methods correct equilibrium distributions in weakly nonlinear or near-integrable systems, they fail in strongly nonlinear regimes. We develop a framework to get the equilibrium distributions and dispersion relations in strongly nonlinear many-body systems, incorporating corrections beyond the random phase approximation and capturing intrinsic nonlinear effects. The theory is verified on the nonlinear Schrodinger equation, the Majda-McLaughlin-Tabak model, and the FPUT-beta model, demonstrating its accuracy across distinct types of nonlinear systems. Numerical results show substantial improvements over existing approaches, even in strong nonlinear regimes. This work establishes a theoretical foundation for equilibrium statistical properties in strongly nonlinear systems.

cond-mat.stat-mech

Revisiting Multi-Wave Resonance in Classical Lattices: Quasi-Resonances, Not Exact Resonance, Govern Energy Redistribution

The multi-wave exact resonance condition is a fundamental principle for understanding energy transfer in condensed matter systems, yet the dynamical evolution of waves satisfying this condition remains unexplored. Here, we reveal that the multi-wave resonant kinetic equations possess distinctive symmetry properties that preferentially induce energy equalization between counter-propagating waves of identical frequency. This initial equalization disrupts the exact resonance condition, rendering it dynamically invalid. We further demonstrate that nonlinearity-mediated multi-wave quasi-resonances--not exact resonances--overn energy transfer and drive the system toward thermalization. Crucially, the strength of exact resonances decays with increasing system size, while quasi-resonance strength grows. Moreover, exact resonance strength remains independent of nonlinearity, whereas quasi-resonance strength diminishes with reduced nonlinearity. These observations provide additional evidence supporting the aforementioned conclusion while elucidating the size-dependent thermalization characteristics in lattice systems.

cond-mat.stat-mech

Multi-Type Instability Processes of Periodic Orbits in Nonlinear Chains

Nonlinear normal modes are periodic orbits that survive in nonlinear many-body Hamiltonian systems, and their instability is crucial for relaxation dynamics. Here, we study the instability process of the $\pi/3$-mode in the Fermi-Pasta-Ulam-Tsingou-$\alpha$ chain with fixed boundary conditions. We find that three types of bifurcations -- period-doubling, tangent, and Hopf -- coexist in this system, each driving instability at specific reduced wave-number $\tilde{k}$. Our analysis reveals a universal scaling law for the instability time $\mathcal{T} \propto (\lambda - \lambda_{\rm c})^{-1/2}$, independent of bifurcation types and models, where the critical perturbation strength $\lambda_{\rm c}$ scales as $\lambda_{\rm c} \propto (\tilde{k} - \tilde{k}_{\rm c})$, with $\tilde{k}_{\rm c}$ varying across bifurcations. We also observe a double instability phenomenon for certain system sizes, meaning that larger perturbations do not always lead to faster thermalization. These results provide new insights into the relaxation and thermalization dynamics in many-body systems.

cond-mat.stat-mech

Network Dynamics-Based Framework for Understanding Deep Neural Networks

Advancements in artificial intelligence call for a deeper understanding of the fundamental mechanisms underlying deep learning. In this work, we propose a theoretical framework to analyze learning dynamics through the lens of dynamical systems theory. We redefine the notions of linearity and nonlinearity in neural networks by introducing two fundamental transformation units at the neuron level: order-preserving transformations and non-order-preserving transformations. Different transformation modes lead to distinct collective behaviors in weight vector organization, different modes of information extraction, and the emergence of qualitatively different learning phases. Transitions between these phases may occur during training, accounting for key phenomena such as grokking. To further characterize generalization and structural stability, we introduce the concept of attraction basins in both sample and weight spaces. The distribution of neurons with different transformation modes across layers, along with the structural characteristics of the two types of attraction basins, forms a set of core metrics for analyzing the performance of learning models. Hyperparameters such as depth, width, learning rate, and batch size act as control variables for fine-tuning these metrics. Our framework not only sheds light on the intrinsic advantages of deep learning, but also provides a novel perspective for optimizing network architectures and training strategies.

cs.LG

Study on the efficiency droop in high-quality GaN material under high photoexcitation intensity

III-V nitride semiconductors, represented by GaN, have attracted significant research attention. Driven by the growing interest in smart micro-displays, there is a strong desire to achieve enhanced light output from even smaller light-emitting diode (LED) chips. However, the most perplexing phenomenon and the most significant challenge in the study of emission properties under high-injection conditions in GaN has always been efficiency droop for decades, where LEDs exhibit a substantial loss in efficiency at high driving currents. In this paper, we present our study on the intrinsic emission properties of high-quality GaN material based on the density of states and the principles of momentum conservation. Our theoretical calculations reveal a momentum distribution mismatch between the non-equilibrium excess electrons and holes, which becomes more significant as the carrier concentration increases. Our excitation-dependent photoluminescence measurements conducted at 6 K exhibited a clear droop for all exciton recombinations, but droop-free for phonon-assisted recombination due to phonons compensating for the momentum mismatch. These findings indicate that the momentum distribution mismatch between the non-equilibrium excess electrons and holes is one of the intrinsic causes of the efficiency droop, which originates from the intrinsic band properties of GaN. These results suggest that proper active region design aimed at reducing this mismatch will contribute to the development of ultra-highly efficient lighting devices in the future.

physics.optics

Cavity Plasmon: Enhanced Luminescence Effect on InGaN Light Emitting Diodes

We fabricated polygonal nanoholes in the top p-GaN layer of the InGaN/GaN light-emitting diode, followed by the deposition of Au/Al metal thin film within the nanoholes to create metal microcavities, thereby constructing the surface plasmon structure. The findings indicate that with increased current injection, the light output of the LEDs rose by 46%, accompanied by a shift of the gain peak position towards the plasmon resonance energy. The maximum enhancement factor increases to 2.38 as the coupling distance decreases from 60 nm to 30 nm. Interestingly, time-resolved photoluminescence data showed that the spontaneous emission decay time lengthened due to the plasmon coupling, suggesting the presence of a new plasmon coupling mechanism. Finite-Difference Time-Domain simulation results show that the electric field is localized at certain locations around the metal microcavity, generating a new type of shape-sensitive plasmon, named Cavity Plasmon here. This intense localization leads to a longer lifetime and enhances the recombination efficiency of excitons. We discuss several unique properties of the cavity plasmon generated by the polygonal metal microcavity with several specific angular shapes. The results demonstrate that the cavity plasmon generated by the polygonal metal microcavity is a highly promising technique for enhancing the light emission performance of of relevant semiconductor optoelectronic devices.

physics.optics

Unsupervised Machine Learning for Detecting and Locating Human-Made Objects in 3D Point Cloud

A 3D point cloud is an unstructured, sparse, and irregular dataset, typically collected by airborne LiDAR systems over a geological region. Laser pulses emitted from these systems reflect off objects both on and above the ground, resulting in a dataset containing the longitude, latitude, and elevation of each point, as well as information about the corresponding laser pulse strengths. A widely studied research problem, addressed in many previous works, is ground filtering, which involves partitioning the points into ground and non-ground subsets. This research introduces a novel task: detecting and identifying human-made objects amidst natural tree structures. This task is performed on the subset of non-ground points derived from the ground filtering stage. Marked Point Fields (MPFs) are used as models well-suited to these tasks. The proposed methodology consists of three stages: ground filtering, local information extraction (LIE), and clustering. In the ground filtering stage, a statistical method called One-Sided Regression (OSR) is introduced, addressing the limitations of prior ground filtering methods on uneven terrains. In the LIE stage, unsupervised learning methods are lacking. To mitigate this, a kernel-based method for the Hessian matrix of the MPF is developed. In the clustering stage, the Gaussian Mixture Model (GMM) is applied to the results of the LIE stage to partition the non-ground points into trees and human-made objects. The underlying assumption is that LiDAR points from trees exhibit a three-dimensional distribution, while those from human-made objects follow a two-dimensional distribution. The Hessian matrix of the MPF effectively captures this distinction. Experimental results demonstrate that the proposed ground filtering method outperforms previous techniques, and the LIE method successfully distinguishes between points representing trees and human-made objects.

cs.CV

Exploring a Physics-Informed Decision Transformer for Distribution System Restoration: Methodology and Performance Analysis

Driven by advancements in sensing and computing, deep reinforcement learning (DRL)-based methods have demonstrated significant potential in effectively tackling distribution system restoration (DSR) challenges under uncertain operational scenarios. However, the data-intensive nature of DRL poses obstacles in achieving satisfactory DSR solutions for large-scale, complex distribution systems. Inspired by the transformative impact of emerging foundation models, including large language models (LLMs), across various domains, this paper explores an innovative approach harnessing LLMs' powerful computing capabilities to address scalability challenges inherent in conventional DRL methods for solving DSR. To our knowledge, this study represents the first exploration of foundation models, including LLMs, in revolutionizing conventional DRL applications in power system operations. Our contributions are twofold: 1) introducing a novel LLM-powered Physics-Informed Decision Transformer (PIDT) framework that leverages LLMs to transform conventional DRL methods for DSR operations, and 2) conducting comparative studies to assess the performance of the proposed LLM-powered PIDT framework at its initial development stage for solving DSR problems. While our primary focus in this paper is on DSR operations, the proposed PIDT framework can be generalized to optimize sequential decision-making across various power system operations.

eess.SY

Methodology and Real-World Applications of Dynamic Uncertain Causality Graph for Clinical Diagnosis with Explainability and Invariance

AI-aided clinical diagnosis is desired in medical care. Existing deep learning models lack explainability and mainly focus on image analysis. The recently developed Dynamic Uncertain Causality Graph (DUCG) approach is causality-driven, explainable, and invariant across different application scenarios, without problems of data collection, labeling, fitting, privacy, bias, generalization, high cost and high energy consumption. Through close collaboration between clinical experts and DUCG technicians, 46 DUCG models covering 54 chief complaints were constructed. Over 1,000 diseases can be diagnosed without triage. Before being applied in real-world, the 46 DUCG models were retrospectively verified by third-party hospitals. The verified diagnostic precisions were no less than 95%, in which the diagnostic precision for every disease including uncommon ones was no less than 80%. After verifications, the 46 DUCG models were applied in the real-world in China. Over one million real diagnosis cases have been performed, with only 17 incorrect diagnoses identified. Due to DUCG's transparency, the mistakes causing the incorrect diagnoses were found and corrected. The diagnostic abilities of the clinicians who applied DUCG frequently were improved significantly. Following the introduction to the earlier presented DUCG methodology, the recommendation algorithm for potential medical checks is presented and the key idea of DUCG is extracted.

cs.AI