arXiv ScienceSearch

arXiv subjects

Tonmoy Rajkhowa

Publications and source records attributed to Tonmoy Rajkhowa.

2 recordsLinked to original sources

Performance Analysis of AFDM Transceiver Under Practical IQ Imbalance

This paper investigates the impact of in-phase/quadrature (IQ) imbalance on signal detection in affine frequency-division multiplexing (AFDM) systems over doubly dispersive Rayleigh fading channels. The bit error rate (BER) performance of AFDM under various IQ mismatch conditions is evaluated and compared with that of an ideal transceiver. Simulation results show that, under ideal conditions, BER decreases rapidly with increasing signal-to-noise ratio (SNR). However, IQ imbalance introduces image interference that spreads across the affine domain, causing significant performance degradation. The results further highlight the impact of practical IQ imbalance on AFDM systems across different quadrature amplitude modulation (QAM) orders, channel scenarios, and comparisons with orthogonal frequency-division multiplexing (OFDM) and orthogonal time frequency space (OTFS) systems. In particular, higher-order QAM performance degrades significantly, indicating the need for IQ imbalance compensation techniques or impairment-aware detection algorithms to ensure robust operation in practical AFDM systems.

eess.SP

TM-PATHVQA:90000+ Textless Multilingual Questions for Medical Visual Question Answering

In healthcare and medical diagnostics, Visual Question Answering (VQA) mayemergeasapivotal tool in scenarios where analysis of intricate medical images becomes critical for accurate diagnoses. Current text-based VQA systems limit their utility in scenarios where hands-free interaction and accessibility are crucial while performing tasks. A speech-based VQA system may provide a better means of interaction where information can be accessed while performing tasks simultaneously. To this end, this work implements a speech-based VQA system by introducing a Textless Multilingual Pathological VQA (TMPathVQA) dataset, an expansion of the PathVQA dataset, containing spoken questions in English, German & French. This dataset comprises 98,397 multilingual spoken questions and answers based on 5,004 pathological images along with 70 hours of audio. Finally, this work benchmarks and compares TMPathVQA systems implemented using various combinations of acoustic and visual features.

cs.CV