arXiv ScienceSearch

arXiv subjects

Linglong Dai

Publications and source records attributed to Linglong Dai.

At least 19 recordsLinked to original sources

From Source Reconstruction to Predictive State Preservation: An Information-Theoretic Framework for AI-Native Communication

AI-native communication increasingly aims to support prediction rather than reproduce every detail of the source. This shift raises a basic question left implicit by conventional source coding: what should be preserved when the terminal goal is prediction? We take the source-induced predictive state as the fidelity object. It is the distribution of the specified future conditioned on the source observation and shared context. We show that this state is sufficient and minimal for exact predictive preservation. The terminal prediction loss then defines communication distortion as lost predictive performance rather than source reconstruction error. Under logarithmic loss, this distortion equals the conditional mutual information lost through communication. Using the receiver-side predictive state as a Bayes reference, we separate AI-receiver error into predictive value lost in communication, receiver-available value unusable by the model family, and family capability not realized by the deployed model. The same predictive state also suffices for matched compression. For finite-alphabet memoryless sources, access to the raw source gives no rate-distortion advantage over coding the state directly. Applying the same target-conditioned construction to sequential prediction reveals a dynamic boundary. A state induced by a fixed horizon is minimal for that horizon but may not support recursive updating as the target window shifts. Taking the entire future as the target yields a minimal full-future state that updates recursively and admits a Markov representation. Together, these results shift AI communication from source reconstruction to predictive-state preservation.

cs.IT

Why Is Cubic-Phase Airy Beamforming Sufficient for Blockage Recovery?

Blockage is a critical challenge for near-field communications, where reliable transmission depends heavily on the line-of-sight (LoS) path and can suffer severe power degradation when that path is obstructed. Near-field Airy beams offer a promising solution for blockage mitigation by forming curved trajectories that guide energy around obstacles, and can be practically generated with phased arrays by imposing a cubic source phase. However, trajectory-based interpretations explain how Airy beams propagate, but not why cubic-phase Airy beamforming is sufficient for blockage recovery or how much received-power gain the cubic term itself contributes. To answer these questions, we identify the blockage-induced phase mismatch relative to conventional near-field focusing and quantify how successive phase orders compensate it. The resulting analysis reveals that the linear and quadratic degrees of freedom, originally used to compensate the free-space geometric phase, can be reoptimized under blockage to provide resteering and refocusing, respectively. The quadratic term can compensate the dominant quadratic component of the additional mismatch, while the Airy cubic provides the first independent correction to the remaining non-quadratic mismatch. Simulations show that linear and quadratic compensation recover most of the available gain. The Airy cubic adds only $0.1433$ dB on average, yet enables the cubic-phase family to attain $99.77\%$ of the phase-only upper bound. Residual-phase analysis further determines when the remaining higher-order components are negligible within a prescribed received-power tolerance. These results explain why cubic-phase Airy beamforming is sufficient: lower-order phase terms provide most of the recovery, while the cubic term closes nearly all of the remaining gap.

eess.SP

Physics-Guided Neural Airy Beamforming for Near-Field Blockage Mitigation

High-frequency communication systems heavily rely on line-of-sight(LoS) paths, so blockage of the LoS path can cause severe performance loss. Near-field Airy beams with curved trajectories can steer energy around obstacles, offering a promising solution for blockage mitigation. However, existing methods for selecting a near-optimal Airy beam trajectory either rely on high-overhead beam training, or employ data-driven learning without a clear, physically interpretable rule. To address this problem, we propose a physics-guided neural Airy beamforming framework that selects a near-optimal trajectory in one shot with clear physical interpretability. Specifically, we first formulate a single-edge representation of the blocker model in 3GPP TR 38.901 and reveal the trajectory--edge coupling mechanism. This analysis yields a trajectory-selection optimality condition that defines the candidate trajectories. Although these trajectories generally cannot be expressed in closed form, we show that they form a continuous structure. This continuous structure is then exploited to construct a compact physics-defined region that captures near-optimal trajectories. Guided by this region, a lightweight neural predictor is finally designed to directly select a near-optimal trajectory without beam training. Simulations show that the compact physics-defined region effectively captures near-optimal trajectories, while occupying only about 6% of the candidate-space area on average. The proposed framework retains 99.7% of the reference rate obtained through numerical optimization, and nearly matches the rate of the data-driven method despite using approximately 112\times fewer neural-network parameters.

eess.SP

Near-Field Communications with Grating Lobes for Quasi-Distributed Arrays: From ULA to MRA

Extremely large-scale antenna array (ELAA) has emerged as a common feature of many key candidate technologies for 6G, where the near-field characteristics become dominant. The quasi-distributed array can further extend the near-field range and utilize the near-field benefits to improve the system performance. However, its typical implementation with modular arrays suffers from severe grating lobes that cause non-negligible inter-user interferences. To solve this problem, we propose the modular minimum-redundancy array (M-MRA) to suppress near-field grating lobes by redesigning the subarray configuration. Specifically, we first characterize the beam pattern of the conventional modular uniform linear array (M-ULA). Contrary to the common belief that grating lobes only exist in the angle domain, we reveal that near-field grating lobes may also occur in the distance domain. We further analyze how to suppress near-field grating lobes for the M-ULA. The results demonstrate that increasing the number of antennas per module can suppress grating lobes. In particular, the number of antennas required grows linearly with the inter-module spacing, thus the grating lobe interferences are severe under a limited number of antennas. This limitation inspires us to propose the M-MRA by redesigning the subarray structure. For M-MRA, the nonuniform antenna spacing within each subarray provides a narrower spatial envelope, allowing it to suppress near-field grating lobes in the angle and distance domains simultaneously. Simulation results verify that the proposed M-MRA can significantly improve the spectrum efficiency of multi-user near-field communications under the same number of antennas.

eess.SP

FM-Receiver: A Foundation Model Enabled Unified Inner and Outer Neural Receiver Towards AI-Native Wireless Communications

With the development of artificial intelligence (AI) techniques, neural receivers, which apply AI to improve wireless receivers have been developed. However, most existing neural receivers apply deep learning only to the outer receiver while retaining conventional channel decoding for the inner receiver, which prevents joint optimization and makes it difficult to build efficient and unified AI-native receivers. To address this issue, we propose a foundation model (FM)-enabled unified neural receiver, FM-Receiver, that integrates the outer and inner receivers into a single AI-native framework, by leveraging the strong representation capability of FMs. Specifically, we introduce a grouped error correction code Transformer that performs symbol-level channel decoding, enabling seamless integration of the inner and outer receiver. Building on this, we illustrate the proposed FM-Receiver, that directly takes the received signals as input of FM and outputs the recovered transmitted bits. In addition, a three-stage configuration-adaptive pre-training strategy is designed to improve the generalization ability to diverse system configurations and scenarios. Extensive simulations show that the proposed FM-Receiver achieves better performance than baselines across different system configurations. It also demonstrates strong zero-shot generalization to unseen frequency bands and scenarios.

cs.IT

Blockage-Robust Beamforming for Near-Field Communications: From Single-Airy to Multi-Airy

High-frequency communications strongly depend on the line-of-sight (LoS) path, and obstacle blockage can severely degrade the received signal power and achievable rate. Near-field Airy beams with curved trajectories can circumvent obstacles, offering a promising way to alleviate blockage. However, since an Airy beam carries most useful energy along a single curved trajectory, existing Airy beamforming methods are highly sensitive to estimation errors of transmitter-obstacle-receiver geometry. That is to say, even a small error in the estimated geometry may cause the mismatched Airy trajectory, leading to severe performance loss. To address this problem, we propose a multi-Airy beamforming scheme for blockage-robust near-field communications. Specifically, we first reveal and analyze the sensitivity mechanism of single-Airy beamforming. This mechanism motivates us to extend the single-Airy generation method to a coordinated multi-Airy generation method by deriving the phase offsets required to coherently combine multiple Airy beams at the target user. Based on this coordinated generation method, we partition the transmit array into multiple sub-arrays and configure a tailored Airy beam for each sub-array, so that the resulting Airy beams formed by multiple curved trajectories can be coherently combined at the target user. Simulation results verify the sensitivity of single-Airy beamforming and the robustness of multi-Airy beamforming under estimation errors of transmitter-obstacle-receiver geometry. Moreover, the proposed scheme achieves higher achievable rates than single-Airy beamforming in blocked scenarios without geometry estimation errors.

eess.SP

Near-field Beam Training under Multi-path Channels: A Hybrid Learning-and-Optimization Approach

For extremely large-scale arrays (XL-arrays), the discrete Fourier transform (DFT) codebook, conventionally used in the far-field, has recently been employed for near-field beam training. However, most existing methods rely on the line-of-sight (LoS) dominant channel assumption, which may suffer degraded communication performance when applied to the general multi-path scenario due to the more complex received signal power pattern at the user. To address this issue, we propose in this paper a new hybrid learning-and-optimization-based beam training method that first leverages deep learning (DL) to obtain coarse channel parameter estimates, and then refines them via a model-based optimization algorithm, hence achieving high-accuracy estimation with low computational complexity. Specifically, in the first stage, a tailored U-Net architecture is developed to learn the non-linear mapping from the received power pattern to coarse estimates of the angles and ranges of multi-path components. In particular, the inherent permutation ambiguity in multi-path parameter matching is effectively resolved by a permutation invariant training (PIT) strategy, while the unknown number of paths is estimated based on defined path existence logits. In the second stage, we further propose an efficient particle swarm optimization method to refine the angular and range parameters within a confined search region; in the meanwhile, a Gerchberg-Saxton algorithm is used to retrieve multi-path channel gains from the received power pattern. Last, numerical results demonstrate that the proposed hybrid design significantly outperforms various benchmarks in terms of parameter estimation accuracy and achievable rate, yet with low computational complexity.

eess.SP

Wavenumber-domain signal processing for holographic MIMO: Foundations, methods, and future directions

Holographic multiple-input multiple-output (H-MIMO) systems represent a paradigm shift in wireless communications by enabling quasi-continuous apertures. Unlike conventional MIMO systems, H-MIMO with subwavelength antenna spacing operates in both far-field and near-field regimes, where classical discrete Fourier transform (DFT) representations fail to sufficiently capture the channel characteristics. To address this challenge, this article provides an overview of the emerging wavenumber-domain signal processing framework. Specifically, by leveraging spatial Fourier plane-wave decomposition to model H-MIMO channels, the wavenumber domain offers a unified and physically consistent basis for characterizing subwavelength-level spatial correlation and spherical wave propagation. This article first introduces the concept of H-MIMO and the wavenumber representation of H-MIMO channels. Next, it elaborates on wavenumber-domain signal processing technologies reported in the literature, including multiplexing, channel estimation, and waveform designs. Finally, it highlights open challenges and outlines future research directions in wavenumber-domain signal processing for next-generation wireless systems.

eess.SP

RAQ-MIMO: MIMO for Multi-Band Rydberg Atomic Quantum Receiver

Rydberg atomic quantum receivers (RAQRs) are capable of receiving multi-band radio-frequency (RF) signals simultaneously, which are expected to break Chu's limit for classical electronic antennas. However, signals from different users will interfere with each other in the optical intermediate frequency (IF) domain of the multi-band quantum receiver, which is termed the IF interference (IFI) problem. To address this problem, in this paper, we propose a multi-input multi-output (MIMO) architecture for Rydberg atomic quantum receiver (RAQ-MIMO) by exploiting the additional spatial diversity of MIMO receivers. Specifically, by applying the dynamic signal model of RAQRs, we clarify the physical relationship between the quantum local oscillator (LO) configurations and the multi-band gains with the concept of quantum transconductance. Then, with the quantum transconductance-based signal model, we formulate the spectral efficiency (SE) maximization problem and further propose the quantum weighted minimum mean square error (qWMMSE) algorithm, which jointly optimizes the quantum LO configurations and the classical precoder/combiner matrices. Furthermore, we test the qWMMSE algorithm within the standard space division multiple access (SDMA) scheme and the frequency division multiple access (FDMA) scheme. Simulation results demonstrate that the qWMMSE optimization framework can significantly improve the SE of RAQ-MIMO systems for both multiple access schemes, and that RAQ-MIMO systems can outperform classical electronic receiver-based multi-user MIMO systems by eliminating the mutual coupling effect between classical antennas.

cs.IT

MUSE-FM: Multi-task Environment-aware Foundation Model for Wireless Communications

Recent advancements in foundation models (FMs) have attracted increasing attention in the wireless communication domain. Leveraging the powerful multi-task learning capability, FMs hold the promise of unifying multiple tasks of wireless communication with a single framework. Nevertheless, existing wireless FMs face limitations in the uniformity to address multiple tasks with diverse inputs/outputs across different communication scenarios. In this paper, we propose a MUlti-taSk Environment-aware FM (MUSE-FM) with a unified architecture to handle multiple tasks in wireless communications, while effectively incorporating scenario information. Specifically, to achieve task uniformity, we propose a unified prompt-guided data encoder-decoder pair to handle data with heterogeneous formats and distributions across different tasks. Besides, we integrate the environmental context as a multi-modal input, which serves as prior knowledge of environment and channel distributions and facilitates cross-scenario feature extraction. Simulation results illustrate that the proposed MUSE-FM outperforms existing methods for various tasks, and its prompt-guided encoder-decoder pair facilitates few-shot adaptation to new task configurations. Moreover, the incorporation of environment information improves the ability to adapt to different scenarios.

cs.IT

Near-Field Challenges in Ultra-Wideband ISAC: Beamforming Strategies and System Insights

The shift toward sixth-generation (6G) wireless networks places integrated sensing and communications (ISAC) at the core of future applications such as autonomous driving, extended reality, and smart manufacturing. However, the combination of large antenna arrays and ultra-wide bandwidths brings near-field propagation effects and beam squint to the forefront, fundamentally challenging traditional far-field designs. True time delay units (TTDs) offer a potential solution, but their cost and hardware complexity limit scalability. In this article, we present practical beamforming strategies for near-field ultra-wideband ISAC systems. We explore codebook designs across analog and digital domains that mitigate beam squint, ensure reliable user coverage, and enhance sensing accuracy. We further validate these approaches through large-scale system-level simulations, including 3D map-based evaluations that reflect real-world urban environments. Our results demonstrate how carefully designed beamforming can balance communication throughput with sensing performance, achieving reliable coverage and efficient resource use even under severe near-field conditions. We conclude by highlighting open challenges in hardware, algorithms, and system integration, pointing toward research directions that will shape the deployment of 6G-ready ISAC networks.

eess.SP

General Signal Model and Capacity Limit for Rydberg Quantum Information System

Rydberg atomic receivers represent a transformative approach to achieving high-sensitivity, broadband, and miniaturized radio frequency (RF) reception. However, existing static signal models for Rydberg atomic receivers rely on the steady-state assumption of atomic quantum states, which cannot fully describe the signal reception process of dynamic signals. To fill in this gap, in this paper, we present a general model to compute the dynamic signal response of Rydberg atomic receivers in closed form. Specifically, by applying small-signal perturbation techniques to the quantum master equation, we derive closed-form Laplace domain transfer functions that characterize the receiver's dynamic responses to time-varying signal fields. To gain more insights into the quantum-based RF-photocurrent conversion process, we further introduce the concept of quantum transconductance that describes the quantum system as an equivalent classical system. By applying quantum transconductance, we quantify the influence of in-band blackbody radiation (BBR) noise on the atomic receiver sensitivity. Extensive simulations for Rydberg atomic receivers validate the proposed signal model, and demonstrate the possibility of quantum receivers to outperform classical electronic receivers through the improvement of quantum transconductance.

eess.SP

Empowering Near-Field Communications in Low-Altitude Economy with LLM: Fundamentals, Potentials, Solutions, and Future Directions

The low-altitude economy (LAE) is gaining significant attention from academia and industry. Fortunately, LAE naturally aligns with near-field communications in extremely large-scale MIMO (XL-MIMO) systems. By leveraging near-field beamfocusing, LAE can precisely direct beam energy to unmanned aerial vehicles, while the additional distance dimension boosts overall spectrum efficiency. However, near-field communications in LAE still face several challenges, such as the increase in signal processing complexity and the necessity of distinguishing between far and near-field users. Inspired by the large language models (LLM) with powerful ability to handle complex problems, we apply LLM to solve challenges of near-field communications in LAE. The objective of this article is to provide a comprehensive analysis and discussion on LLM-empowered near-field communications in LAE. Specifically, we first introduce fundamentals of LLM and near-field communications, including the key advantages of LLM and key characteristics of near-field communications. Then, we reveal the opportunities and challenges of near-field communications in LAE. To address these challenges, we present a LLM-based scheme for near-field communications in LAE, and provide a case study which jointly distinguishes far and near-field users and designs multi-user precoding matrix. Finally, we outline and highlight several future research directions and open issues.

eess.SP

A General DoF and Pattern Analyzing Scheme for Electromagnetic Information Theory

Electromagnetic information theory (EIT) is one of the emerging topics for 6G communication due to its potential to reveal the performance limit of wireless communication systems. For EIT, one of the most important research directions is degree of freedom (DoF) analysis. Existing research works on DoF analysis for EIT focus on asymptotic conclusions of DoF, which do not well fit the practical wireless communication systems with finite spatial regions and finite frequency bandwidth. In this paper, we use the theoretical analyzing tools from Slepian concentration problem and extend them to three-dimensional space domain and four-dimensional space-time domain under electromagnetic constraints. Then we provide asymptotic DoF conclusions and non-asymptotic DoF analyzing scheme, which suits practical scenarios better, under different scenarios like three-dimensional antenna array. Moreover, we theoretically prove that the channel DoF is upper bounded by the proposed DoF of electromagnetic fields. Finally, we use numerical analysis to provide some insights about the optimal spatial sampling interval of the antenna array, the DoF of three-dimensional antenna array, the impact of unequal antenna spacing, the orthogonal space-time patterns, etc.

cs.IT

Decoding for Punctured Convolutional and Turbo Codes: A Deep Learning Solution for Protocols Compliance

Neural network-based decoding methods show promise in enhancing error correction performance but face challenges with punctured codes. In particular, existing methods struggle to adapt to variable code rates or meet protocol compatibility requirements. This paper proposes a unified long short-term memory (LSTM)-based neural decoder for punctured convolutional and Turbo codes to address these challenges. The key component of the proposed LSTM-based neural decoder is puncturing-aware embedding, which integrates puncturing patterns directly into the neural network to enable seamless adaptation to different code rates. Moreover, a balanced bit error rate training strategy is designed to ensure the decoder's robustness across various code lengths, rates, and channels. In this way, the protocol compatibility requirement can be realized. Extensive simulations in both additive white Gaussian noise (AWGN) and Rayleigh fading channels demonstrate that the proposed neural decoder outperforms conventional decoding techniques, offering significant improvements in decoding accuracy and robustness.

cs.LG

Large Language Model Enabled Multi-Task Physical Layer Network

The advance of Artificial Intelligence (AI) is continuously reshaping the future 6G wireless communications. Particularly, the development of Large Language Models (LLMs) offers a promising approach to effectively improve the performance and generalization of AI in different physical-layer (PHY) tasks. However, most existing works finetune dedicated LLM networks for a single wireless communication task separately. Thus performing diverse PHY tasks requires extremely high training resources, memory usage, and deployment costs. To solve the problem, we propose a LLM-enabled multi-task PHY network to unify multiple tasks with a single LLM, by exploiting the excellent semantic understanding and generation capabilities of LLMs. Specifically, we first propose a multi-task LLM framework, which finetunes LLM to perform multi-user precoding, signal detection and channel prediction simultaneously. Besides, multi-task instruction module, input encoders, as well as output decoders, are elaborately designed to distinguish different tasks. The proposed design allows different wireless data types to be well aligned with the LLM input format. Moreover, low-rank adaptation (LoRA) is utilized for LLM fine-tuning. To reduce the memory requirement during LLM fine-tuning, a LoRA fine-tuning-aware quantization method is introduced. Extensive numerical simulations are also displayed to verify the effectiveness of the proposed method.

cs.IT

Spatio-Temporal Electromagnetic Kernel Learning for Channel Prediction

Accurate channel prediction is essential for addressing channel aging caused by user mobility. However, the actual channel variations over time are highly complex in high-mobility scenarios, which makes it difficult for existing predictors to obtain future channels accurately. The low accuracy of channel predictors leads to difficulties in supporting reliable communication. To overcome this challenge, we propose a channel predictor based on spatio-temporal electromagnetic (EM) kernel learning (STEM-KL). Specifically, inspired by recent advancements in EM information theory (EIT), the STEM kernel function is derived. The velocity and the concentration kernel parameters are designed to reflect the time-varying propagation of the wireless signal. We obtain the parameters through kernel learning. Then, the future channels are predicted by computing their Bayesian posterior, with the STEM kernel acting as the prior. To further improve the stability and model expressibility, we propose a grid-based EM mixed kernel learning (GEM-KL) scheme. We design the mixed kernel to be a convex combination of multiple sub-kernels, where each of the sub-kernel corresponds to a grid point in the set of pre-selected parameters. This approach transforms non-convex STEM kernel learning problem into a convex grid-based problem that can be easily solved by weight optimization. Finally, simulation results verify that the proposed STEM-KL and GEM-KL schemes can achieve more accurate channel prediction. This indicates that EIT can improve the performance of wireless system efficiently.

eess.SP

AI and Deep Learning for Terahertz Ultra-Massive MIMO: From Model-Driven Approaches to Foundation Models

This study explored the transformative potential of artificial intelligence (AI) in addressing the challenges posed by terahertz ultra-massive multiple-input multiple-output (UM-MIMO) systems. It begins by outlining the characteristics of terahertz UM-MIMO systems and identifies three primary challenges for transceiver design: computational complexity, modeling difficulty, and measurement limitations. The study posits that AI provides a promising solution to these challenges. Three systematic research roadmaps are proposed for developing AI algorithms tailored to terahertz UM-MIMO systems. The first roadmap, model-driven deep learning (DL), emphasizes the importance of leveraging available domain knowledge and advocates the adoption of AI only to enhance bottleneck modules within an established signal processing or optimization framework. Four essential steps are discussed: algorithmic frameworks, basis algorithms, loss-function design, and neural architecture design. The second roadmap presents channel state information (CSI) foundation models, aimed at unifying the design of different transceiver modules by focusing on their shared foundation, that is, the wireless channel. The training of a single compact foundation model is proposed to estimate the score function of wireless channels, which serve as a versatile prior for designing a wide variety of transceiver modules. Four essential steps are outlined: general frameworks, conditioning, site-specific adaptation, joint design of CSI foundation models, and model-driven DL. The third roadmap aims to explore potential directions for applying pretrained large language models (LLMs) to terahertz UM-MIMO systems. Several application scenarios are envisioned, including LLM-based estimation, optimization, search, network management, and protocol understanding. Finally, the study highlights open problems and future research directions.

eess.SP