arXiv ScienceSearch

arXiv subjects

Tan Xin

Publications and source records attributed to Tan Xin.

3 recordsLinked to original sources

Graph-Based Safe Reinforcement Learning for Multi-Agent Systems with Time-Varying Topology

This paper presents a graph-based safe multi-agent reinforcement learning (MARL) framework for cooperative navigation with time-varying topology. To address the critical challenge of ensuring safety in environments with sensing constraints, a safety-decoupled mechanism is introduced through a Control Barrier-Like Function (CBLF) action screening layer. This mechanism bridges the gap between discrete LiDAR perception and continuous safety constraints, ensuring that physical safety constraints are strictly satisfied regardless of the learning progress. Building upon this safety foundation, a unified structural architecture is proposed, integrating a attention-based actor and a Graph Attention Network (GAT) centralized critic. The actor utilizes a value vector reconstruction mechanism that explicitly encodes relative geometric relations through a collaborative tracking error matrix, enabling scale-insensitive policy learning under time-varying communication topologies. Meanwhile, the GAT-based critic models evolving interaction structures for accurate global value estimation. The proposed framework is validated on real differential-drive robot platforms, and experimental results demonstrate superior stability and safety in dynamic scenarios with limited fields-of-view.

cs.RO

Human Motion Synthesis in 3D Scenes via Unified Scene Semantic Occupancy

Human motion synthesis in 3D scenes relies heavily on scene comprehension, while current methods focus mainly on scene structure but ignore the semantic understanding. In this paper, we propose a human motion synthesis framework that take an unified Scene Semantic Occupancy (SSO) for scene representation, termed SSOMotion. We design a bi-directional tri-plane decomposition to derive a compact version of the SSO, and scene semantics are mapped to an unified feature space via CLIP encoding and shared linear dimensionality reduction. Such strategy can derive the fine-grained scene semantic structures while significantly reduce redundant computations. We further take these scene hints and movement direction derived from instructions for motion control via frame-wise scene query. Extensive experiments and ablation studies conducted on cluttered scenes using ShapeNet furniture, as well as scanned scenes from PROX and Replica datasets, demonstrate its cutting-edge performance while validating its effectiveness and generalization ability. Code will be publicly available at https://github.com/jingyugong/SSOMotion.

cs.CV

Super-Resolution Coherent Diffractive Imaging via Titled-Incidence Multi-Rotation-Angle Fusion Ptychography

Coherent diffractive imaging (CDI) enables lensless imaging with experimental simplicity and a flexible field of view, yet its resolution is fundamentally constrained by the Abbe diffraction limit. To overcome this limitation, we introduce a novel Tilted-Incidence Multi-Rotation-Angle Fusion Ptychography technique. This approach leverages a tilted-incidence geometry to extend the collection angle beyond the Abbe limit, achieving up to a -fold resolution enhancement. By acquiring diffraction patterns at multiple sample rotation angles, we capture complementary spatial frequency information. A tilted-incidence multi-rotation-angle fusion ptychographic iterative engine (tmf-PIE) algorithm is then employed to integrate these datasets, enabling super-resolution image reconstruction. Additionally, this method mitigates the anisotropic resolution artifacts inherent to tilted CDI geometries. Our technique represents a novel advancement in super-resolution imaging, providing a novel alternative alongside established methods such as STED, SIM, and SMLM.

physics.optics