arXiv Science⌕ Search

arXiv · 2610.04201

Quantized Transformers for Massive MIMO Precoding with Automatic Resolution Tuning

Abstract

Deep learning precoders, particularly those based on Transformer architectures, offer superior spectral efficiency for Massive MIMO systems but may incur high computational cost. While existing mixed-precision quantization studies on convolutional neural networks establish important energy baselines, traditional discrete search methods fail to scale to the large parameter spaces of modern Transformers. To bridge this gap, we apply a differentiable precision learning technique to massive MIMO precoding. This approach autonomously and jointly optimizes weights, activations, quantization step sizes, and layer-wise bit-widths in a single training loop, directly optimizing for energy efficiency. We further expand this approach by incorporating Neural Architecture Search across various Transformer model sizes and evaluating the impact of random versus floating-point trained initializations on resolution tuning. Ultimately, we demonstrate that the proposed training method enables the deployment of highly compact Transformer precoders, improving energy efficiency by up to 288\(\times\) compared to the classical Weighted Minimum Mean Square Error algorithm at equal sum-rate performance in a dense downtown environment.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Ghazal Kasalaee, Glodi Sala Mangituka, Ali Hasanzadeh Karkan, Jean-François Frigon, François Leduc-Primeau. 2026-10-03. Quantized Transformers for Massive MIMO Precoding with Automatic Resolution Tuning. https://arxiv.org/abs/2610.04201

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Resource-Efficient Wi-Fi CSI-Based Sensing via Exploiting the Age of Samples

Wi-Fi channel state information (CSI)-based sensing must coexist with data communications, limiting the availability of temporally-dense CSI measurements. We formulate CSI-based human activity and identity recognition under an average sensing budget that limits the fraction of CSI measurement and reporting opportunities within a sensing session. The budget captures sensing-communication resource sharing, packet loss, and traffic-induced irregularity, which we model using deterministic (accumulated) and stochastic (Bernoulli) sampling policies. We propose a low-cost, age-aware WiFi sensing framework that encodes the age of each retained CSI sample and multiplicatively fuses it with the CSI embedding. On the NTU-Fi human activity recognition and person identification datasets, the proposed model outperforms both a CSI-only baseline and the time-aware attention model of the UniFi benchmark across most operating regimes. For person identification, it improves over UniFi by more than 10 percentage points, with the largest gains under strict sensing budgets.

eess.SP↗

Graph Learning for Cross-Subject, Cross-Population EEG Emotion Decoding and Model-Derived Spatial-Spectral Neural Signatures

Cross subject emotion decoding from electroencephalography EEG requires representations that accommodate individual variability while preserving spatial spectral structure for interpretation. This study introduces EmoDiPyraTrans, a differential graph Transformer that integrates adaptive graph recurrence, differential attention, pyramid fusion and distribution regularization over sequential relative power spectral density graphs. Across SEED, FACED, MAHNOB HCI, DEAP and DREAMER, the model achieved the highest participant mean accuracy and positive class F1 among the evaluated methods, with accuracy and F1 both reaching 0.928 on SEED. On DEP EEG, positive versus neutral accuracy reached 0.802 within healthy controls and 0.704 within participants with depression, compared with 0.591 under healthy to depression transfer and 0.581 with mixed population development. Complementary SEED analyses identified distributed spatial weighting and an alpha centred spectral preference, while configurations averaging six channels retained near full performance. These findings link generalization assessment with model derived candidate signatures to support interpretable EEG emotion decoding, with code available at https://github.com/hdy6438/EmoDiPyraTrans.

eess.SP↗