arXiv Science⌕ Search

arXiv · 2610.05485

RF-Agent: Hierarchical Language-Agent Control for Instruction-Conditioned Active Spectrum Sensing

Abstract

Active radio-frequency (RF) sensing must acquire evidence under receiver limits while following confirmation, recovery, and stopping instructions. \newhl{We propose RF-Agent, a closed-loop architecture combining language supervision, episode memory, deterministic RF execution, perception feedback, and evidence-based reporting.} An independent auditor checks trajectory compliance. We derive acquisition and supervisory-request bounds, characterize full-history token complexity, and bound joint success by output validity and evidence capacity. \newhl{ActiveRF-AD tests instruction changes at fixed states and evaluates joint success, requiring compliance and RF correctness in the same episode.} \newhl{Training on different instructions at shared sensing states raises counterfactual-pair accuracy from 24.03\% to 96.22\% and Core joint success from 65.13\% to 72.12\%. Of the 6.99-point joint-success gain, 6.73 points reflect a decrease in the fraction of all episodes that are RF-correct but noncompliant.} \newhl{At Core budget three, paired-training hierarchical control achieves 66.60\% joint success versus 9.17\% for direct-action control.} \newhl{In a separate comparison at the same budget, RF-Agent improves joint success by 2.31 points and reduces RF acquisitions by 12.3\% relative to an initial-state planner, at increased inference cost.} Cross-backbone diagnostics cover three model families. These results show how instruction-responsive control improves task completion beyond what RF accuracy alone reveals.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Hao Zhang, Dongfang Xu, Hang Zou, Anis Bara, Brahim Mefgouda, Lina Bariah, Yuzhi Yang, Fuhui Zhou, Qihui Wu, Merouane Debbah. 2026-10-04. RF-Agent: Hierarchical Language-Agent Control for Instruction-Conditioned Active Spectrum Sensing. https://arxiv.org/abs/2610.05485

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

PyDPF: A Python Package for Differentiable Particle Filtering

State-space models (SSMs) are a widely used tool in time series analysis. In the complex systems that arise from real-world data, it is common to employ particle filtering (PF), an efficient Monte Carlo method for estimating the hidden state corresponding to a sequence of observations. Applying particle filtering requires specifying both the parametric form and the parameters of the system, which are often unknown and must be estimated. Gradient-based optimisation techniques cannot be applied directly to standard particle filters, as the filters themselves are not differentiable. However, several recently proposed methods modify the resampling step to make particle filtering differentiable. In this paper, we present an implementation of several such differentiable particle filters (DPFs) with a unified API built on the popular PyTorch framework. Our implementation makes these algorithms easily accessible to a broader research community and facilitates straightforward comparison between them. We validate our framework by reproducing experiments from several existing studies and demonstrate how DPFs can be applied to address several common challenges with state space modelling.

eess.SP↗

Rigid Body Localization via Gaussian Belief Propagation with Quadratic Angle Approximation

Gaussian belief propagation (GaBP) is a technique that relies on linearized error and input output models to yield low-complexity solutions to complex estimation problems, which has been recently shown to be effective in the design of range-based GaBP schemes for stationary and moving rigid body localization (RBL) in three-dimensional (3D) space, as long as the relative rotation between the prior position and the target rigid body is sufficiently small. In this article we present a novel range-based RBL scheme via GaBP that relaxes the latter limitation significantly. To this end, the proposed method incorporates a quadratic angle approximation to linearize the relative orientation between the prior and the target rigid body, enabling high precision estimates of corresponding rotation angles even for large deviations. Leveraging the resulting linearized model, we derive the corresponding message-passing (MP) rules to obtain estimates of the translation vector and rotation matrix of the target rigid body, relative to a prior reference frame. Numerical results corroborate the good performance of the proposed angle approximation itself, as well as the consequent RBL performance in terms of root mean square errors (RMSEs) in comparison to the state-of-the-art (SotA), while maintaining a low computational complexity.

eess.SP↗

Radar Intelligent Detection with Coarse-Grained Labels

Many deep learning based radar target detectors rely on range-cell level labels for training, which are expensive to obtain. To reduce the labeling burden, this paper presents a training strategy that uses only range-window level labels. Specifically, two sub-echoes are randomly cropped from the same window echo, resulting in a known relative shift between them. A siamese Transformer is then used to identify target-salient responses from one sub-echo and form pseudo labels, which are mapped to the paired sub-echo according to the known shift. In addition, a shift-equivariant consistency (SEC) constraint is imposed to align the predicted responses of the two sub-echoes after shift compensation. By exploiting shift priors within each labeled window, the proposed approach learns reliable detection without requiring cell-level labels. Experimental results show that it consistently surpasses classical constant false alarm rate (CFAR) detectors and achieves performance close to range-cell level labels training, while substantially reducing labeling cost.

eess.SP↗