arXiv ScienceSearch

arXiv subjects

Lin Cheng

Publications and source records attributed to Lin Cheng.

At least 19 recordsLinked to original sources

AirAnchor: Bridging Local and Global Spatial Information for Zero-Shot Aerial Vision-and-Language Navigation

Aerial Vision-and-Language Navigation requires drones to follow natural-language instructions and navigate through complex urban environments. Accurate navigation relies on both local and global spatial information, which support immediate action grounding and long-horizon path planning, respectively. However, existing zero-shot methods typically operate at a single spatial scale, relying either on local representations constructed online from current observations or on global memories built offline from historical experience. To address this limitation, we propose AirAnchor, a new paradigm that bridges local and global spatial information through spatial anchors and integrates both into a shared navigation framework, enabling comprehensive spatial grounding for decision-making. AirAnchor consists of three core components: (1) Query-Driven Spatial Anchor Grounding, which identifies decision-relevant anchors from visual observations and organizes them into local spatial representations; (2) Persistent Object Spatial Memory, which incrementally maintains an object knowledge base as persistent global spatial memory and retrieves landmark-related spatial priors; and (3) a Spatially-Informed Navigation Agent, which explicitly integrates both local and global spatial information into an agentic framework for decision-making. Extensive experiments on AerialVLN demonstrate that AirAnchor substantially outperforms existing zero-shot baselines, validating the effectiveness and efficiency of the proposed paradigm.

cs.CV

Diffractive Sail H-Reversal Trajectory: Theoretical Feasibility, Design Strategies, and Applications

With the growing threat of near-Earth asteroid, planetary defense serves as a vital shield against catastrophic disasters. Kinetic impact utilizing an angular momentum reversal (H-reversal) trajectory of a solar sail is a highly advantageous defense approach. However, traditional reflective sails (RS) are constrained during these maneuvers by attitude-thrust coupling and a degradation of solar radiation pressure utilization at the high cone angles required for transverse acceleration. To enhance impact performance and simplify control, this paper proposes an H-reversal impact scheme utilizing a Sun-facing diffractive sail (SFDS) under a one-stage diffraction angle {\theta}d strategy and a two-stage {\theta}d strategy. The feasible parameter spaces for both strategies were mapped using the hodograph method. Tailored trajectory design methods were established for both strategies based on the feasibility analysis. Apophis impact scenario was considered, and the corresponding trajectories were constructed. Simulations demonstrate that the one-stage {\theta}d SFDS outperforms RS through a 21% increase in the impact velocity and a 35% decrease in the mission duration. Furthermore, the proposed two-stage {\theta}d strategy yields an additional 9km/s gain in impact velocity. By utilizing SFDS H-reversal trajectories, this research establishes an emergency planetary defense framework characterized by rapid response and high kinetic energy.

astro-ph.EP

Process-Knowledge-Embedded Safe DRL for Real-Time Dispatch of Process Loads in Industrial Microgrids

Steelmaking process loads (SPLs) are flexible resources that enhance local renewable-energy utilization and reduce electricity procurement costs in industrial microgrids. However, strong multistage coupling makes current decisions affect subsequent feasibility, challenging conventional deep reinforcement learning to reduce costs while maintaining process feasibility throughout production. This paper proposes a process-knowledge-embedded safe deep reinforcement learning framework for the real-time dispatch of SPLs in industrial microgrids. Specifically, a lossless active-frontier action space is constructed, and a process-distance-guided action-processing mechanism reallocates excluded-action probabilities according to process distance and the actor's safe-action preference. Recursive process feasibility is established to guarantee admissible execution and feasible continuation. Furthermore, the expected process-correction distance is incorporated into PPO through a correction budget and a primal-dual update to internalize process knowledge into the raw policy, while a derived bound quantifies the raw policy's dependence on safety processing. Case studies using real-world data demonstrate zero process losses, electricity-cost reductions of 49.2% and 25.9% relative to rule-based scheduling and rolling MILP, respectively, within an acceptable computation time.

eess.SY

Coordinated Primary Frequency Regulation and Grid-Forming Control for Wind Turbine Generators

Conventional grid-forming (GFM) control strategies often treat the DC source as an unconstrained link, creating mismatches when applied to the wind turbine generators (WTGs). Focusing on primary frequency regulation, this paper systematically investigates the mismatch between the GFM-WTGs behavior and the droop-based primary frequency regulation. To address this issue, a novel coordination strategy between WTG primary frequency regulation and GFM control is proposed. By establishing well-designed relationships among the power-tracking coefficient, power set-point, and frequency deviation, the proposed strategy enables GFM-WTGs to participate consistently in primary frequency regulation within predefined frequency limits while maintaining appropriate power points and effectively utilizing the allowable power reserve. Furthermore, the proposed method preserves the control structure and dynamic performance of conventional GFM control and inherently adapts to varying wind-speed conditions. Comparative case studies under different operating scenarios demonstrate the effectiveness and superiority of the proposed strategy.

eess.SY

Optimal sizing of a hydrogen-based direct reduced iron-electric arc furnace system integrated with methanol synthesis toward zero-carbon steel production

A hydrogen-based direct reduced iron-electric arc furnace system integrated with methanol synthesis (H2-DRI-EAF-MeOH) provides a feasible pathway toward zero-carbon steel production. However, the volatility and intermittency of renewable energy sources (RES) may lead to capacity oversizing, while conventional annualized-cost and levelized-cost metrics are insufficient to evaluate investment return. To address these issues, this paper develops a fractional programming-based optimal sizing model for the H2-DRI-EAF-MeOH system. Hourly production rate limits are decoupled from annual production capacities and formulated as independent decision variables. Theoretical analysis demonstrates that increasing hourly production rate limits expands the multi-period operating feasible region, increases the theoretical upper bound on RES matching, and provides an investment-efficiency condition for increasing production rate limits. The annualized net return on investment (ANROI) is then adopted to represent investment return, and the resulting mixed-integer linear fractional programming model is solved using the Dinkelbach method. Case studies show that compared with the annualized net return (ANR) benchmark, the ANROI-based sizing yields higher internal rates of return (IRR), reaching 20.67% and 17.08% under the on-grid and off-grid scenarios, respectively. As the equivalent annual operating hours decrease from 8000 to 4000 h, the IRR increases from 19.68% to 20.67% under the on-grid scenario and from 12.64% to 17.08% under the off-grid scenario, while the corresponding battery storage capacities decrease by 113.56 and 796.31 MWh. The off-grid zero-carbon constraint enables CO2-to-MeOH conversion but reduces the system IRR to 14.99%, indicating that the resulting revenues are insufficient to offset the additional costs.

eess.SY

Geometry-Resolved Projection of RF Imbalance to Ion Micromotion in a Same-Phase Dual-RF Blade Trap

Common-mode metrics of a high-$Q$ helical resonator do not determine the residual ion-side field in a dual-electrode drive. We combine a two-node differential RF model with single-electrode finite-element bases to obtain a computation-only, geometry-resolved projection for a same-phase blade trap. For a 3.5 pF external load per branch, the model gives a total effective branch capacitance of 7.640 pF and an HWHM-equivalent full branch-difference scale of 12.7 fF at $Q_{\mathrm{loaded}}=600$. The seven-segment geometry gives center and axial-RMS differential field coefficients of 640 V m$^{-1}$ and 635 V m$^{-1}$ per differential peak volt. A representative 10 fF mismatch with an effective 0.1 pF balance scale projects to 44.5/44.1 nm center/RMS $^{171}\mathrm{Yb}^{+}$ micromotion at 100 V common peak voltage. Supplementary thermal, bypass-admittance, and tested numerical cases characterize model sensitivity. All reported displacements are projections; no RF-bench or ion-side validation is claimed.

quant-ph

Boundary-Phase Control of Sequentially Addressed Trapped-Ion ZZ Interactions

Motion-mediated trapped-ion interactions commonly coordinate state-dependent forces on both target ions. Sequential optical access reduces the number of concurrent target channels but makes the relative phase between disjoint force windows a control variable. We derive a complex near-resonant description in which each window generates a displacement vector and ordered symplectic products between vectors on different ions produce the ZZ phase. Only relative boundary phases affect this area; a common phase shift is a gauge transformation. Building on the experimental precedent for alternating single-ion addressing, we develop matched-envelope phase and contrast controls that isolate this boundary-phase dependence without target-window overlap or hidden force in the dark gaps. The analysis separates phase generation from differential closure, projector-common motion, deterministic local-Z phases, spectator coupling, and control-parameter transfer. A conditional-Ramsey sequence gives continuous and reset contrasts of 0.998 and 0.996, with a reset-induced phase separation of 0.581 rad modulo $\pi/2$. In the representative comparison, sequential control uses fewer concurrent target channels but greater normalized force action than independently calibrated simultaneous control. All results are model-level estimates within the stated Lamb-Dicke, rotating-wave, and apparatus-input limits.

quant-ph

Automated Synthesis of Facial Mechanisms for Conversational Animatronic Robots

Animatronic faces are a central component of socially interactive robots, enabling rich nonverbal communication through facial articulation. However, state-of-the-art animatronic faces are typically tailored systems: each new facial geometry requires extensive manual mechanical redesign, making large-scale personalization prohibitively slow and costly. In this work, we pursue automated and scalable mechanical face synthesis, aiming to rapidly generate a physically realizable facial mechanism for a wide range of facial geometries. We introduce a parametric, linkage-driven mechanical face template whose topology and actuator layout are explicitly parameterized to support systematic scaling and retargeting across diverse facial morphologies. Building on this template, we propose a hierarchical automatic design algorithm that takes a single 2D portrait as input, reconstructs a target 3D face, and synthesizes a collision-free, manufacturable internal mechanism. The algorithm combines anatomy-guided feasible motion volumes, Action Unit (AU)-derived trajectory-based expressiveness objectives, and a collision-driven outer-loop refinement strategy. Beyond hardware synthesis, we argue that future mechanical faces deployed at scale must engage in bidirectional, multi-turn conversation rather than functioning solely as speaking or listening heads. To this end, we develop a dual-identity conversational facial motion synthesis framework that jointly models speaking and listening behaviors from audio, producing temporally coherent 3D facial motion suitable for physical execution. We validate our system through extensive experiments, including (i) quantitative evaluation of automatic mechanism synthesis across diverse facial geometries, (ii) comparisons against manual mechanical design, (iii) benchmarks on conversational facial motion synthesis and real-time deployment, and (iv) perceptual user studies.

cs.RO

Digital Quantum Simulation of Nonequilibrium Dynamics in the Schwinger Model under a Strong External Electric Field

We use the (1+1)-dimensional Schwinger model to investigate the nonequilibrium dynamics of a finite lattice system under a constant external electric field. The lattice Hamiltonian is constructed under open boundary conditions. The vacuum state is prepared using the variational quantum eigensolver (VQE). Scans over the external field strength show the flip of the vacuum state at several field strengths. The critical field strengths agree with theoretical predictions. We further investigate the real-time evolution of the zero-field vacuum under an external electric field using a second-order Trotter-Suzuki decomposition. By comparison with exact diagonalization (ED), we verify that the quantum-simulation protocol reproduces the main features of field-induced boundary charge separation, decay of the vacuum-state fidelity, and quasiperiodic energy redistribution between the electric-field energy term and the fermionic sector. Our results indicate that combining VQE-based state preparation with digital real-time evolution provides a useful approach for studying nonequilibrium dynamics in strong-field lattice gauge theories.

hep-lat

Reference-Free Heterogeneous Multi-Agent Reinforcement Learning for Grid-Friendly Tie-Line Power Shaping in Industrial Microgrids

Tie-line power (TLP) shaping is a key requirement for the grid-friendly operation of industrial microgrids (IMGs). This paper studies the coordination of multi-timescale heterogeneous adjustable resources in a steel IMG to shape a grid-friendly TLP trajectory considering multiple objectives. A sequential heterogeneous-agent coordination (SHAC) framework is proposed, where process loads, hydrogen storage, and battery storage are modeled as functionally heterogeneous agents with cross-role observations, asynchronous decision intervals, role-specific rewards and critics. This design captures the heterogeneous temporal effects of different resources on the TLP trajectory and alleviates ambiguous credit assignment and weak inter-agent coordination. To ensure feasible real-time execution, process-knowledge-based action masking and feasibility projection are embedded into policy execution, and a role-aware multi-timescale actor--critic training scheme is developed for agents with different action structures and decision intervals. Numerical studies using real renewable generation and electricity market data show that SHAC effectively eliminates the dependence on predefined reference trajectories and enables adaptive 1-min online decision-making, achieving zero production failures with an average computational time of only 0.4 ms per step. Compared with the original operation, SHAC reduces the total grid purchase cost, contract-demand exceedance time, and cumulative ramp excess by 91.27\%, 98.64\%, and 96.91\%, respectively. These results demonstrate that the proposed framework improves the economic efficiency and grid friendliness of industrial microgrid operation while satisfying strict process-safety constraints and real-time computational requirements.

eess.SY

A Process-Aware Demand Response Evaluation Framework for Hydrogen-Integrated Zero-Carbon Steel Plants Coupled with Methanol Production

High penetration of renewables (RES) and the retirement of thermal units aggravate flexibility scarcity in power systems. Hydrogen-based low-carbon steel production systems possess substantial demand response (DR) potential. This paper proposes a process-aware DR evaluation framework for hydrogen-integrated zero-carbon steel plants coupled with methanol production (H2-DRI-EAF-MeOH). First, a novel H2-DRI-EAF-MeOH architecture is introduced to eliminate residual emissions via methanol synthesis. Integrated energy-material flows are formulated to reflect coupling interactions governing DR potential. Second, to capture electric arc furnace (EAF) operational constraints while preserving tractability, an operating feasible region model is developed and validated using field data from a pure hydrogen direct reduced iron and EAF plant, yielding a 4.1% average relative error. Third, a process-aware DR potential evaluation model is formulated, incorporating a nonlinear asymmetric penalty and an adaptive rolling mechanism to reflect operators' aversion to process deviations and avoid myopic scheduling. Finally, dual-side evaluation metrics are established to quantify grid-side delivered DR capacity and ramping risks, making load-side unit-level regulation behaviors observable. Case studies show the proposed framework achieves an average effective delivered DR capacity of 178.3 MW, improves RES-load matching from 0.257 to 0.587, and reduces costs by 15.68% compared to the baseline. Furthermore, the exponential asymmetric penalty mitigates extreme tail risks of process deviations. Ultimately, this work provides a theoretical foundation for leveraging RES-steel-chemical synergies to mitigate flexibility scarcity.

eess.SY

Demand response potential evaluation of a zero carbon hydrogen metallurgy system considering shaft furnace's flexibility

The increasing penetration of intermittent renewable energy sources and the retirement of thermal units have widened the power system flexibility gap. Industrial demand response (DR) driven by real-time pricing is widely regarded as a viable solution. In this paper, we propose a framework to quantify the DR potential of a zero-carbon hydrogen metallurgy system (ZCHMS) considering shaft furnace's flexibility. First, we model the shaft furnace as a constrained flexible load and validate the model via simulation, achieving a root mean square error of 4.48\% of the rated load. Second, we formulate a DR potential evaluation method that determines baseline and DR-based production scheduling schemes by minimizing operating cost subject to production orders. Finally, the numerical results show that compared with the baseline, DR-based ZCHMS reduces operating cost by 6.6\%, incentivizing demand-side management in ironmaking and strengthening power-ironmaking synergies.

eess.SY

Full Timescale Hierarchical MPC-MTIP Framework for Hybrid Energy Storage Management in Low-Carbon Industrial Microgrid

Uncertainties in balancing generation and load in low-carbon industrial microgrids (IMGs) make hybrid energy storage systems (HESS) crucial for their stable and economic operation. Existing model predictive control (MPC) techniques typically enforce periodic state of charge (SOC) constraints to maintain long term stability. However, these hard constraints compromise dispatch flexibility near the end of the prediction horizon, preventing sufficient energy release during critical peaks and leading to optimization infeasibility. This paper eliminates the periodic SOC constraints of individual storage units and proposes a novel full-timescale hierarchical MPC scheduling framework. Specifically, comprehensive physical and cost models are established for the HESS composed of flywheel, battery, compressed-air, and hydrogen-methanol energy storage. The control problem is decoupled into a hierarchical MPC architecture. Furthermore, a novel adaptive feedback mechanism based on micro trajectory inverse projection (MTIP) is embedded into the scheduling process, accurately mapping the high frequency dynamic buffering capabilities of lower tier storages into the upper decision space to generate dynamic boundaries. Experiments using 14 consecutive months of second-level data from a real-world IMG validate the effectiveness of the proposed method, demonstrating its significant superiority over existing approaches. By effectively preventing limit violations and deadlocks in lower-tier storages under extreme fluctuations, it achieves a 97.4\% net load smoothing rate and a 62.2\% comprehensive cycle efficiency.

eess.SY

Parameter Optimization in Trajectory Planning via Differentiable Convex Programming

Sequential convex programming has been established as an effective framework for solving nonconvex trajectory planning problems. However, its performance is highly sensitive to problem parameters, including trajectory variables, algorithmic hyperparameters, and physical vehicle parameters. This paper introduces a differentiable sequential convex programming framework that integrates differentiable convex optimization with sequential convex programming to enable end-to-end parameter optimization. By deriving first-order sensitivity relations of second-order cone programming solutions with respect to problem data, exact gradients of trajectory performance metrics with respect to arbitrary parameters are obtained and propagated through iterations. The effectiveness of the proposed framework is validated through three representative applications: optimal terminal-time prediction for powered landing, trust-region penalty optimization in subproblems, and surface-to-mass ratio optimization for hypersonic gliding vehicles. Simulation results show that the proposed framework enables reliable gradient-based parameter learning and significantly improves numerical performance, convergence behavior, and design efficiency. These results indicate that the differentiable sequential convex programming framework provides a powerful and general tool for vehicle design, mission optimization, and hyperparameter selection in aerospace trajectory planning.

math.OC

Subgoal Graph-Augmented Planning for LLM-Guided Open-World Reinforcement Learning

Large language models (LLMs) offer strong high-level planning capabilities for reinforcement learning (RL) by decomposing tasks into subgoals. However, their practical utility is limited by poor planning-execution alignment, which reflects a critical gap between abstract plans and actionable, environment-compatible behaviors. This misalignment arises from two interrelated limitations: (1) LLMs often produce subgoals that are semantically plausible but infeasible or irrelevant in the target environment due to insufficient grounding in environment-specific knowledge, and (2) single-LLM planning conflates generation with self-verification, resulting in overconfident yet unreliable subgoals that frequently fail during execution. To address these challenges, we propose Subgoal Graph-Augmented Actor-Critic-Refiner (SGA-ACR), a framework that integrates an environment-specific subgoal graph and structured entity knowledge with a multi-LLM planning pipeline that explicitly separates generation, critique, and refinement to produce executable and verifiable subgoals. A subgoal tracker further monitors execution progress, provides auxiliary rewards, and adaptively updates the subgoal graph to maintain alignment between plans and actions. Experimental results on 22 diverse tasks in the open-world game "Crafter" demonstrate the effectiveness of our proposed method.

cs.LG

Sparse Kalman Identification for Partially Observable Systems via Adaptive Bayesian Learning

Sparse dynamics identification is an essential tool for discovering interpretable physical models and enabling efficient control in engineering systems. However, existing methods rely on batch learning with full historical data, limiting their applicability to real-time scenarios involving sequential and partially observable data. To overcome this limitation, this paper proposes an online Sparse Kalman Identification (SKI) method by integrating the Augmented Kalman Filter (AKF) and Automatic Relevance Determination (ARD). The main contributions are: (1) a theoretically grounded Bayesian sparsification scheme that is seamlessly integrated into the AKF framework and adapted to sequentially collected data in online scenarios; (2) an update mechanism that adapts the Kalman posterior to reflect the updated selection of the basis functions that define the model structure; (3) an explicit gradient-descent formulation that enhances computational efficiency. Consequently, the SKI method achieves accurate model structure selection with millisecond-level efficiency and higher identification accuracy, as demonstrated by extensive simulations and real-world experiments (showing an 84.21\% improvement in accuracy over the baseline AKF).

eess.SY

An Iterative Problem-Driven Scenario Reduction Framework for Stochastic Optimization with Conditional Value-at-Risk

Scenario reduction (SR) alleviates the computational complexity of scenario-based stochastic optimization with conditional value-at-risk (SBSO-CVaR) by identifying representative scenarios to depict the underlying uncertainty and tail risks. Existing distribution-driven SR methods emphasize statistical similarity but often exclude extreme scenarios, leading to weak tail-risk awareness and insufficient problem-specific representativeness. Instead, this paper proposes an iterative problem-driven scenario reduction framework. Specifically, we integrate the SBSO-CVaR problem structure into SR process and project the original scenario set from the distribution space onto the problem space. Subsequently, to minimize the SR optimality gap with acceptable computation complexity, we propose a tractable iterative problem-driven scenario reduction (IPDSR) method that selects representative scenarios that best approximate the optimality distribution of the original scenario set while preserving tail risks. Furthermore, the iteration process is rendered as a mixed-integer program to enable scenario partitioning and representative scenarios selection. And ex-post problem-driven evaluation indices are proposed to evaluate the SR performance. Numerical experiments show IPDSR significantly outperforms existing SR methods by achieving an optimality gap of less than 1% within an acceptable computation time.

math.OC

Recursive Inference for Heterogeneous Multi-Output GP State-Space Models with Arbitrary Moment Matching

Accurate learning of system dynamics is becoming increasingly crucial for advanced control and decision-making in engineering. However, real-world systems often exhibit multiple channels and highly nonlinear transition dynamics, challenging traditional modeling methods. To enable online learning for these systems, this paper formulates the system as Gaussian process state-space models (GPSSMs) and develops a recursive learning method. The main contributions are threefold. First, a heterogeneous multi-output kernel is designed, allowing each output dimension to adopt distinct kernel types, hyperparameters, and input variables, improving expressiveness in multi-dimensional dynamics learning. Second, an inducing-point management algorithm enhances computational efficiency through independent selection and pruning for each output dimension. Third, a unified recursive inference framework for GPSSMs is derived, supporting general moment matching approaches, including the extended Kalman filter (EKF), unscented Kalman filter (UKF), and assumed density filtering (ADF), enabling accurate learning under strong nonlinearity and significant noise. Experiments on synthetic and real-world datasets show that the proposed method matches the accuracy of SOTA offline GPSSMs with only 1/100 of the runtime, and surpasses SOTA online GPSSMs by around 70% in accuracy under heavy noise while using only 1/20 of the runtime.

stat.ML