arXiv ScienceSearch

arXiv subjects

Cunhua Pan

Publications and source records attributed to Cunhua Pan.

At least 19 recordsLinked to original sources

WiFi based Multi user Activity Recognition via User Conditioned Spatial Attention

WiFi-based human activity recognition has achieved high accuracy under the single-user scene. Recognizing activities performed by multiple concurrent users remains challenging, because their body-reflected propagation paths superimpose in the channel state information (CSI) measurement. Existing multi-user methods either decompose the signal before classification, suffering from error propagation, or predict activities jointly without reliably associating each prediction with the correct user. Moreover, attention mechanisms that have proven effective for single-user sensing compute a single user-agnostic attention map, blending patterns from different users and providing no mechanism to associate features with specific user slots. This paper proposes a user-conditioned spatial attention (UCSA) module that generates per-user spatial attention maps by conditioning multi-scale spatial attention on learnable user-ID embeddings via feature-wise linear modulation (FiLM). UCSA creates a deep coupling between spatial feature enhancement and user differentiation: the attention module itself becomes user-aware, producing distinct attention maps for each user slot rather than relying on a late-fusion embedding addition. Combined with a shareable multi-semantic spatial attention (SMSA) module for multi-scale spatial enhancement, the proposed framework addresses both feature quality and user differentiation in a unified architecture. Experiments on the WiMANS benchmark across three indoor environments and up to five concurrent users show that the proposed method outperforms nine baselines in all settings, achieving an average accuracy of over 93\% across three environments even with five users. Ablation studies confirm that both SMSA and UCSA contribute substantially to recognition accuracy, with UCSA providing the critical link between shared features and per-user predictions.

eess.SP

Adaptive Antenna Element Activation for Multiuser Uniform Spherical Arrays

Uniform spherical arrays provide full-space coverage. However, owing to directional element patterns, full-array transmission activates many elements that contribute little to a given user, resulting in increased system overhead and inefficient power allocation. To address this issue, this letter proposes an adaptive element activation strategy for multiuser uniform spherical arrays that reduces the number of physically active elements while satisfying individual user rate requirements. The proposed strategy constructs a user-specific candidate sequence according to the angle between each element boresight and the user direction. At each iteration, it selects, among the unsatisfied users, the user-element connection that provides the largest improvement in the aggregate rate deficit. Simulation results show that fewer than half of the elements are sufficient for a single user to achieve 80% of its full-array rate, while fewer than 70% are sufficient to achieve the full-array rates of all users in multiuser scenarios. Moreover, the proposed strategy consistently requires a lower active-element ratio than the baseline methods across all tested configurations.

eess.SP

Secure Relay Low-Altitude Networks via Hybrid Fixed-Position and Rotatable Antenna Arrays

In this paper, a relay network with hybrid fixed-position and rotatable antenna arrays is proposed. The deployment of rotatable arrays in conventional relay networks is considered to provide more secure communications for low-altitude economy applications. Specifically, both the base station and the relay station are equipped with fixed-position antenna arrays and rotatable arrays to serve ground users and aerial users, respectively. To address the challenge of multi-user interference, a low-cost reconfigurable intelligent surface is exploited as a candidate path. Accordingly, under constraints on transmit power, user quality of service, rotatable range, and path selection, the objective is to maximize the worst-case secrecy rate (SR) through joint beamforming, power allocation, and rotatable antenna orientation design. First, the SR performance in the \textcolor{black}{single ground-aerial user pair scenario} is investigated, and a step-by-step leakage-based scheme is proposed. Then, the general multi-user scenario is studied, and a Distributional Soft Actor-Critic with Three refinements (DSAC-T)-based learning scheme, which supports hybrid discrete and continuous actions, is proposed to maximize the worst-case SR. Simulation results validate the effectiveness of the proposed schemes. The proposed schemes achieve approximately a twofold improvement in SR performance compared to isotropic antennas. \textcolor{black}{The proposed system can achieve approximately 99.7\% power saving, 55\% antenna saving, and can serve more users.}

eess.SP

An Optical Pathway to Movable Rydberg Atomic Quantum Receivers

This paper develops an optically movable Rydberg atomic quantum receiver (RAQR), in which the probe and coupling beams are steered within each vapor cell to dynamically reconfigure the effective radio-frequency (RF) sensing position without mechanical actuation. A closed-form equivalent baseband model is derived by separating the atomic transduction coefficient, optical steering phase, and cell-center array response into distinct factors and the accuracy of the resulting model is validated against numerical solutions of the Lindblad master equation. Based on the derived model, we reveal two complementary channel-shaping mechanisms, including intrinsic beam-pattern shaping through RF-to-optical transduction and per-cell phase control enabled by optical displacement. To further exploit these capabilities, a non-convex sum-rate maximization problem is formulated over the optical positions and local oscillator design and solved via an alternating optimization framework with analytical gradients. Simulation results validate the derived model and demonstrate substantial performance gains enabled by optical movability, highlighting its potential as a programmable receiver architecture for future wireless networks.

eess.AS

Tensor Train Decomposition-based Channel Estimation for MIMO-AFDM Systems with Fractional Delay and Doppler

Affine Frequency Division Multiplexing (AFDM) has emerged as a promising chirp-based multicarrier technology for high-speed communication systems. To fully exploit the diversity gain offered by AFDM, accurate channel estimation is essential. However, existing studies have mainly focused on the integer-delay-tap scenario and single-symbol pilot-based estimation. Since delay taps in practice are generally fractional, approximating them as integers not only degrades delay estimation accuracy but also severely affects Doppler frequency estimation. To address this problem, in this paper, we investigate channel estimation for multiple-input multiple-output (MIMO)-AFDM systems. A time-affine frequency (T-AF) domain pilot structure is proposed to exploit time-domain phase variations. By leveraging the rotational invariance property in the spatial and temporal domains, a channel estimation algorithm based on Vandermonde-structured tensor-train (TT) decomposition is developed. The proposed algorithm demonstrates superior computational efficiency compared with state-of-the-art parameter estimation methods. Moreover, diverging from current studies, we derive the global Ziv-Zakai bound (ZZB) as an alternative parameter estimation error lower bound to the Cramér-Rao bound (CRB). Numerical results show that the derived ZZB provides tighter global performance characterization and successfully captures the threshold phenomenon in mean square error (MSE) performance in the low-SNR regime. Furthermore, the proposed algorithm achieves superior communication performance relative to the existing schemes, while offering a computational speedup, reducing the execution time by an order of magnitude compared to the state-of-the-art iterative algorithms.

cs.IT

U-Net-Based Generative Joint Source-Channel Coding for Wireless Image Transmission

Deep learning (DL)-based joint source-channel coding (JSCC) methods have achieved remarkable success in wireless image transmission. However, these methods either focus on conventional distortion metrics that do not necessarily yield high perceptual quality or incur high computational complexity. In this paper, we propose two DL-based JSCC (DeepJSCC) methods that leverage deep generative architectures for wireless image transmission. Specifically, we propose G-UNet-JSCC, a scheme comprising an encoder and a U-Net-based generator serving as the decoder. Its skip connections enable multi-scale feature fusion to improve both pixel-level fidelity and perceptual quality of reconstructed images by integrating low- and high-level features. To further enhance pixel-level fidelity, the encoder and the U-Net-based decoder are jointly optimized using a weighted sum of structural similarity and mean-squared error (MSE) losses. Building upon G-UNet-JSCC, we further develop a DeepJSCC method called cGAN-JSCC, where the decoder is enhanced through adversarial training. In this scheme, we retain the encoder of G-UNet-JSCC and adversarially train the decoder's generator against a patch-based discriminator. cGAN-JSCC employs a two-stage training procedure. The outer stage trains the encoder and the decoder end-to-end using an MSE loss, while the inner stage adversarially trains the decoder's generator and the discriminator by minimizing a joint loss combining adversarial and distortion losses. Simulation results demonstrate that the proposed methods achieve superior pixel-level fidelity and perceptual quality on both high- and low-resolution images. For low-resolution images, cGAN-JSCC achieves better reconstruction performance and greater robustness to channel variations than G-UNet-JSCC.

eess.IV

AFDM-ISAC With Fractional Delay-Doppler Coupling

Affine frequency division multiplexing (AFDM) is a promising chirp-based multicarrier waveform for high-mobility integrated sensing and communication (ISAC). Accurate angle, delay, and Doppler estimation is essential for AFDM sensing. Since target delays and Doppler shifts are generally continuous-valued, representing them on a discrete delay--Doppler grid causes energy leakage and peak displacement in the discrete affine Fourier transform (DAFT) domain. The AFDM chirp also induces delay--Doppler coupling in the DAFT-domain response. The resulting DAFT-domain matching-score surface exhibits a local ridge that is not aligned with the normalized-delay and normalized-Doppler axes. To address these issues, this paper investigates joint estimation of angle and continuous-valued delay--Doppler parameters for a colocated AFDM-ISAC sensing architecture. A transform-domain sparse sensing model is formulated from the fractional DAFT-domain response. Based on this model, a coupled-coordinate Newtonized orthogonal matching pursuit (CC-NOMP) estimator is developed. CC-NOMP uses the AFDM-induced coupling coordinate to parameterize the dominant local ridge. It combines coupled-coordinate Newton refinement with safeguarded updates, coupling-aligned delay refinement, and cyclic multi-target refinement to estimate angle, continuous normalized delay, and normalized Doppler. A deterministic Cramér--Rao bound and a dominant-order complexity analysis are also derived. Simulation results with continuous-valued off-grid target parameters show that CC-NOMP achieves lower delay and Doppler error floors than the considered baselines while maintaining comparable angle-estimation accuracy.

cs.IT

Joint Optimization for RIS-Assisted Wireless Communications: From Physical and Electromagnetic Perspectives

Reconfigurable intelligent surfaces (RISs) are envisioned to be a disruptive wireless communication technique that is capable of reconfiguring the wireless propagation environment. In this paper, we study a free-space RIS-assisted multiple-input single-output (MISO) communication system in far-field operation. To maximize the received power from the physical and electromagnetic nature point of view, a comprehensive optimization, including beamforming of the transmitter, phase shifts of the RIS, orientation and position of the RIS is formulated and addressed. After exploiting the property of line-of-sight (LoS) links, we derive closed-form solutions of beamforming and phase shifts. For the non-trivial RIS position optimization problem in arbitrary three-dimensional space, a dimensional-reducing theory is proved. The simulation results show that the proposed closed-form beamforming and phase shifts approach the upper bound of the received power. The robustness of our proposed solutions in terms of the perturbation is also verified. Moreover, the RIS significantly enhances the performance of the mmWave/THz communication system.

cs.IT

Cross-Field Channel Parameter Estimation and Channel Characterization at THz Bands in Indoor Scenarios

The terahertz (THz) frequency band offers the potential for ultra-high data rate transmission in future wireless communication systems. To extend the transmission distance and enhance spectral efficiency, the deployment of large-scale antenna arrays emerges as a promising solution in the THz band. This paper targets the critical challenge of cross-field (hybrid near-field/far-field) channel parameter estimation and channel characterization in such configurations. We first establish a 260-380 GHz virtual uniform linear array (ULA) measurement framework in an indoor scenario, capturing high-resolution channel transfer functions (CTFs) that reveal spatial non-stationarity and cross-field wavefront characteristics. Building upon these empirical observations, we propose a cross-field space-alternating generalized expectation-maximization (SAGE) algorithm that discriminatively estimates near-field and far-field multipath components (MPCs) via Bayesian phase-curvature classification, while explicitly tracking spatial birth-death phenomena through visibility region estimation. Analysis of the measurement data validates the algorithm's effectiveness in resolving cross-field MPCs and quantifies that near-field MPCs account for over 90% of total MPCs at 2 m transmission distance (380 GHz). We observe that spatial non-stationarity intensifies as the carrier frequency increases and the transmission distance decreases. These findings offer quantitative guidelines for channel modeling and system design in wireless THz communication systems.

eess.SP

Selective Depthwise Separable Convolution for Lightweight Joint Source-Channel Coding in Wireless Image Transmission

Depthwise separable convolutional (DSConv) layers have been successfully applied to deep learning (DL)-based joint source-channel coding (JSCC) schemes to reduce computational complexity. However, a systematic investigation of the layerwise and ratio-wise replacement of standard convolutional (Conv) layers with DSConv layers in JSCC systems for wireless image transmission remains largely unexplored. In this letter, we propose a configurable lightweight JSCC framework that incorporates a selective replacement strategy, enabling flexible Conv-to-DSConv replacement at different replacement ratios and positions. By varying the replacement ratio, we obtain models with different computational complexities and analyze their impact on reconstruction performance. Furthermore, we investigate how replacements at different encoder and decoder depths influence reconstruction quality under a fixed replacement ratio. Our results show that Conv-to-DSConv replacement at the intermediate layers of the encoder and decoder achieves a favorable complexity-performance trade-off, revealing layer-wise redundancy in DL-based JSCC systems. Extensive experiments further demonstrate that the proposed framework achieves substantial parameter reduction with only slight performance degradation, enabling flexible complexity-performance trade-offs for resource-constrained edge devices.

eess.IV

Cell-Level Channel Shaping for Rydberg Atomic Quantum Receivers in Satellite Uplinks With Doppler-Enabled Superheterodyne Reception

In this paper, we propose a self-superheterodyne Rydberg uniform array receiver for satellite uplink communications, in which the Doppler shift naturally induced by satellite motion is exploited to generate the intermediate-frequency signal. We first develop a near-field local oscillator (LO) synthesis model and characterize the spatially varying LO electric field across the Rydberg vapor cells. Based on a vapor-cell-center approximation, a closed-form radio frequency (RF)-to-optical conversion is derived, establishing an explicit bridge between the incident satellite signal and the LO-induced cell-level response. The derived model reveals that the programmable LO serves as an analog-domain channel-shaping mechanism by controlling the cell-level transduction gain, phase response, and phase-matching behavior. Building upon this equivalent channel model, we formulate an LO design problem that maximizes the Shannon capacity of the effective channel, and develop an efficient optimization algorithm for the LO amplitudes and phases. Simulation results demonstrate that the vapor-cell transduction can reshape the effective channel, adjust the beam-pattern alignment, and moderately reduce the inter-user correlation under suitable LO configurations. Furthermore, the proposed LO design significantly improves the achievable capacity over benchmark schemes, offering a promising self-superheterodyne Rydberg architecture for future satellite communication systems.

eess.SP

RIS-Position and Orientation Estimation in MIMO-OFDM Systems with Practical Scatterers

In this paper, we investigate the problem of estimating the position and the angle of rotation of a mobile station (MS) in a millimeter wave (mmWave) multiple-input-multiple-output (MIMO) system aided by a reconfigurable intelligent surface (RIS). The virtual line-of-sight (VLoS) link created by the RIS and the non-line-of-sight (NLoS) links that originate from scatterers in the considered environment are utilized to facilitate the estimation. A two-step positioning scheme is exploited, where the channel parameters are first acquired, and the position-related parameters are then estimated. The channel parameters are obtained through a coarser and a subsequent finer estimation processes. As for the coarse estimation, the distributed compressed sensing orthogonal simultaneous matching pursuit (DCS-SOMP) algorithm, the maximum likelihood (ML) algorithm, and the discrete Fourier transform (DFT) are utilized to separately estimate the channel parameters. The obtained channel parameters are then jointly refined by using the space-alternating generalized expectation maximization (SAGE) algorithm, which circumvents the high-dimensional optimization issue of ML estimation. Departing from the estimated channel parameters, the positioning-related parameters are estimated. The performance of estimating the channel-related and position-related parameters is theoretically quantified by using the Cramer-Rao lower bound (CRLB). Simulation results demonstrate the superior performance of the proposed positioning algorithms.

eess.SP

Explainable Task-Oriented Token Communication for AI-Native 6G Networks

The integration of Foundation Models (FMs) and wireless communications is driving the evolution of image communication from bit-accurate transmission toward task-oriented transmission. However, existing task-oriented image communication methods still face three major challenges: insufficient task-oriented Token representation, inadequate collaboration between Visual Tokens and Task Tokens, and limited interpretability of task decisions. To address these challenges, we propose an Explainable Task-Oriented Token Communication (ET-TokenCom) framework. By treating Tokens as unified units for information representation and transmission, the proposed framework constructs an end-to-end communication link that spans visual perception, wireless transmission, and task reasoning. At the transmitter, the ET-TokenCom framework extracts Visual Tokens from images to preserve low-level visual information. Meanwhile, Task Tokens generated by the FM are introduced to represent the target information and decision intent required by the current task. A Cross-Modal Attention (CMA) fusion mechanism is further designed, enabling Task Tokens to explicitly guide the selection, weighting, and transmission of Visual Tokens. At the receiver, the framework integrates Token decoding with an explainable output mechanism, where attention heatmaps are generated to highlight critical perceptual regions under different task objectives and reveal the influence of Task Tokens on the outputs. Finally, simulation results validate the effectiveness and robustness of the proposed ET-TokenCom framework.

eess.IV

Vision-Language-Action Models Meet World Models: Embodied Agentic AI for Low-Altitude Wireless Networks

Low-Altitude Wireless Networks (LAWNs), composed of Unmanned Aerial Vehicles (UAVs) and other aerial platforms, provide integrated perception, communication, and computation services in low-altitude airspace. However, deploying large generative models in this domain faces three major challenges: 1) Limited embodied action mapping; 2) Inadequate physical environment modeling; 3) Insufficient closed-loop optimization. To address these challenges, this study proposes an Embodied Agentic UAV framework. Centered on a Vision-Language-Action (VLA) model as the execution core, the framework establishes an end-to-end embodied decision-making pipeline from multimodal environmental perception to continuous control generation. In addition, a World Model (WM) is introduced to capture the coupling between UAV actions and environmental state evolution, thereby supporting environment prediction, policy verification, and dynamic optimization. Furthermore, memory and reflection mechanisms are incorporated to form an adaptive closed-loop optimization paradigm of decision, execution, evaluation, and update, thereby enhancing the system's autonomous decision-making capability and continual evolution ability in complex dynamic environments. Experimental results validate its effectiveness in enabling robust, predictive, and sustainable autonomous control in LAWNs.

cs.IT

A Hybrid Near-field Indoor Channel Model for THz Bands Based on Surface Scattering Characteristics

Terahertz (THz) communication and extremely large-scale MIMO (XL-MIMO) are essential for achieving ultra-high data rates in future 6G systems. However, at sub-millimeter wavelengths, typical indoor materials exhibit significant roughness that invalidates conventional ideal smooth surface assumptions, while massive array apertures introduce pronounced near-field effects and spatial non-stationarity. To address these challenges, this paper proposes a hybrid near-field channel model utilizing surface scattering characteristics based on distinct measurement campaigns. First, based on typical indoor materials scattering measurements across the 260-400 GHz band, an improved Beckmann-Kirchhoff (B-K) model is developed to accurately characterize surface roughness and diffuse scattering behavior. The model independently analyzes single-bounce (SB) and multi-bounce (MB) clusters by applying deterministic rough surface scattering theory and geometry-statistical approach, respectively. Then, using near-field spatial non-stationarity measurements from a 630-element virtual array in the 330-360 GHz band, a Dual-Gaussian Mixture Model (DMM) and a Negative Binomial (NB) distribution are adopted to describe the lengths and the number of spatial visibility regions (VRs), respectively. Additionally, a Weibull distribution is employed to model the intra-region power fluctuations. Finally, comprehensive XL-MIMO channel evaluations within the same band demonstrate that the proposed model aligns closely with measured results in terms of the spatial cross-correlation function (SCCF), frequency cross-correlation function (FCF), and channel capacity. By reproducing the spatial sparsity of THz band, the proposed model overcomes the limitation of conventional standard models, such as 3GPP 38.901 and WINNER II, in significantly overestimating channel capacity.

eess.SP

RFDT-Channel: RGB-LiDAR-Based RF Digital Twin Scene Construction for 28 GHz Indoor Ray-Tracing Channel Simulation

Real-scene indoor millimeter-wave simulation requires efficient modeling of radio frequency (RF)-computable geometry and electromagnetic material properties. To address the low efficiency of manual scene modeling, the limited RF adaptability of visually reconstructed meshes, and the lack of material binding in 28 GHz ray-tracing simulation, RFDT-Channel is developed as an RF digital twin scene construction workflow based on red-green-blue (RGB) images and light detection and ranging (LiDAR) point clouds. Indoor videos and point clouds are collected by a Jetson Orin platform with LiDAR and GMSL cameras. An initial triangular mesh is generated through COLMAP, 3D Gaussian Splatting, and SuGaR. The LiDAR point cloud then provides geometric and scale references for RF-oriented regularization in Blender, including alignment, wall solidification, door/window opening construction, and topology repair. OpenScene semantic segmentation maps major indoor structures to concrete, glass, wood, and metal materials, and Sionna RT performs 28 GHz ray tracing. Under a fixed transmitter-receiver deployment, the generated channel impulse response (CIR), channel frequency response (CFR), and Radio Map results show that material binding mainly changes weak reflection, transmission, and scattering paths, reducing the number of effective paths from about 742 to about 52 while keeping the dominant path amplitude nearly unchanged.

eess.IV

Beyond the RF Paradigm: Rydberg Atomic Receivers for Next-Generation IoT

Next-generation Internet-of-Things (IoT) is evolving toward a ubiquitous, ultra-low-power, and multi-band heterogeneous networking paradigm that seamlessly integrates terrestrial, non-terrestrial, and ambient devices. This vision places unprecedented demands on conventional radio frequency (RF) receivers, whose fundamental bottlenecks in sensitivity, power consumption, coverage, and multi-band operation are rooted in the RF antenna. To tackle these issues, we show that the quantum properties of Rydberg atomic quantum receivers (RAQRs), including ultra-high sensitivity, broad frequency agility, and diverse reception modalities, provide a physically distinct receiver-side path that replaces the conventional antenna-and-low-noise-amplifier chain. Using LoRa, narrowband IoT, and ambient IoT as case studies, this article shows that RAQRs deliver significant gains in weak-uplink, low-power, and battery-free regimes. A stochastic-geometry analysis in cellular and cell-free architectures then maps these device-level gains onto network coverage, where the RAQR retains roughly a 4 dB half-coverage advantage over the RF receiver in sparse deployments at \(λ\sim 10^{-5}~{\mathrm m}^{-2}\), with the gain eroded as device density grows. The open challenges are presented to stand between current RAQR prototypes and deployable IoT infrastructure.

eess.SP

Channel Measurements and Characterization with Phase Drift Compensation for Outdoor 330-360 GHz MIMO Communications

In this paper, an outdoor channel measurement campaign at 330-360 GHz employing a 128 * 4 virtual antenna array (VAA)-based multiple-input multiple-output (MIMO) configuration is conducted. The transmitter (Tx) and receiver (Rx) location pairs are classified into line-of-sight (LoS) and obstructed-LoS (OLoS) scenarios to enable a detailed investigation of outdoor terahertz (THz) band channel characteristics. During the measurement process, the stationarity of the outdoor environment is carefully verified, and a linear phase drift (PD) effect is identified. Then, we propose a PD-aware Space-Alternating Generalized Expectation-Maximization (SAGE) algorithm, which significantly improves both delay resolution and channel parameter estimation accuracy. Based on the processed measurement data, we characterize key channel properties, including the power delay profile, path loss, shadow fading, delay spread, angular spread, Rician K-factor, as well as their cumulative distribution functions and correlation characteristics. In addition, near-field effects and MIMO-specific properties, including the spatial non-stationarity and the cluster birth-death property, are analyzed.

eess.SP