arXiv Science⌕ Search

arXiv · 2610.06407

Reinforcement Learning-Based 3D Beam Adaptation for Underwater Wireless Optical Communication with AUVs

Abstract

Underwater Wireless Optical Communications (UWOC) provide essential high data rates for Autonomous Underwater Vehicles (AUVs), but reliable connectivity is critically affected by transmitter-receiver misalignment. This work addresses the beam pointing problem for a moving AUV subject to unknown ocean currents through a Deep Reinforcement Learning (DRL) framework. We develop a comprehensive 3D UWOC channel model incorporating depth-dependent attenuation, turbulence, and a discrete-ray method to accurately quantify geometric and misalignment losses. The resulting agent jointly optimizes beam steering and divergence angles, learning a policy that prioritizes continuous link maintenance. Evaluated using real oceanographic data, the framework's performance is assessed via the excess outage metric, which isolates outages occurring exclusively due to pointing errors. Relative to perfect alignment between nodes, the proposed method bounds this metric to 8.5% under ocean current-free tracking conditions and limits it to 16.2% when subjected to current-induced drift. The findings obtained in this work confirms that DRL-enabled beam control is capable of adaptive tracking and mitigating link outages in dynamic underwater environments.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Viviana Centritto~Arrojo, Ama Bandara, Sergi Abadal, Evgenii Vinogradov. 2026-10-05. Reinforcement Learning-Based 3D Beam Adaptation for Underwater Wireless Optical Communication with AUVs. https://arxiv.org/abs/2610.06407

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Resource-Efficient Wi-Fi CSI-Based Sensing via Exploiting the Age of Samples

Wi-Fi channel state information (CSI)-based sensing must coexist with data communications, limiting the availability of temporally-dense CSI measurements. We formulate CSI-based human activity and identity recognition under an average sensing budget that limits the fraction of CSI measurement and reporting opportunities within a sensing session. The budget captures sensing-communication resource sharing, packet loss, and traffic-induced irregularity, which we model using deterministic (accumulated) and stochastic (Bernoulli) sampling policies. We propose a low-cost, age-aware WiFi sensing framework that encodes the age of each retained CSI sample and multiplicatively fuses it with the CSI embedding. On the NTU-Fi human activity recognition and person identification datasets, the proposed model outperforms both a CSI-only baseline and the time-aware attention model of the UniFi benchmark across most operating regimes. For person identification, it improves over UniFi by more than 10 percentage points, with the largest gains under strict sensing budgets.

eess.SP↗

Graph Learning for Cross-Subject, Cross-Population EEG Emotion Decoding and Model-Derived Spatial-Spectral Neural Signatures

Cross subject emotion decoding from electroencephalography EEG requires representations that accommodate individual variability while preserving spatial spectral structure for interpretation. This study introduces EmoDiPyraTrans, a differential graph Transformer that integrates adaptive graph recurrence, differential attention, pyramid fusion and distribution regularization over sequential relative power spectral density graphs. Across SEED, FACED, MAHNOB HCI, DEAP and DREAMER, the model achieved the highest participant mean accuracy and positive class F1 among the evaluated methods, with accuracy and F1 both reaching 0.928 on SEED. On DEP EEG, positive versus neutral accuracy reached 0.802 within healthy controls and 0.704 within participants with depression, compared with 0.591 under healthy to depression transfer and 0.581 with mixed population development. Complementary SEED analyses identified distributed spatial weighting and an alpha centred spectral preference, while configurations averaging six channels retained near full performance. These findings link generalization assessment with model derived candidate signatures to support interpretable EEG emotion decoding, with code available at https://github.com/hdy6438/EmoDiPyraTrans.

eess.SP↗