arXiv ScienceSearch

arXiv subjects

Ruchi Pandey

Publications and source records attributed to Ruchi Pandey.

At least 19 recordsLinked to original sources

Star formation, stellar evolution, and planets in high-resolution X-ray imaging

Stars set the conditions for planet formation, planet evolution, and habitability -- all major topics in astronomy today. Stars are also important in their own right as the most visible component of galaxies. In cool stars, X-ray emission is powered by magnetic fields, and so far our Sun is the only system in which those fields are spatially resolved. High-resolution X-ray (HiReX) imaging can track the origin and evolution of those fields, see how they connect young stars to their disks and outflows, and measure the energy, mass, and momentum that radiation and coronal mass ejections (CMEs) carry into the circumstellar environment. This is crucial for understanding whether planets can form and survive in young stellar systems, whether life can develop on those planets, and how the star evolves over time. Intermediate-mass and high-mass stars blow winds and eventually evolve into degenerate objects such as white dwarfs, neutron stars, and black holes. Their evolution and death drive the chemical evolution of galaxies. High-resolution X-ray (HiReX) imaging can study the hottest components in those systems, such as the colliding winds of massive stars, accretion and nova explosions in CVs, and the shocks in outflows that form planetary nebulae. All these cases have in common that the X-ray emission is tracing the hottest, fastest, and most energetic components of the shocks. HiReX observations can reveal the temperature, spatial structures, and elemental abundances of different system components that no other wavelength can. While stars are physically small compared to more powerful objects such as accreting black holes and AGN, they are also much closer to us, allowing a HiReX mission to resolve a variety of physical phenomena fundamental to our understanding of how stellar systems form, evolve, and interact with their environment.

astro-ph.IM

How Bilingual Are SSL Speech Models? Cross-Lingual Probing of Articulatory Encoding with Finnish and Russian EMA

SSL speech models capture rich phonetic, prosodic, and acoustic patterns from raw audio, yet how they encode articulatory information across diverse languages remains unclear. Using EMA data from bilingual Finnish-Russian speakers, we evaluate cross-lingual correlations between SSL latent representations and articulatory movements. Models achieve strong prediction performance (Pearson r up to 0.68) even with approximately 5 minutes of training data, with multilingual models outperforming monolingual ones. Intermediate layers encode articulatory features most effectively, and tongue movements are more predictable than lip movements. We also assess the impact of task type (read versus spontaneous speech) and language proficiency, finding higher accuracy for structured tasks and strong generalization across proficiency levels. These results enhance the interpretability of SSL models and show their potential for speech-technology applications.

eess.AS

Beyond Speaker Independence: Evaluating Cross-Lingual Acoustic-to-Articulatory Inversion Across Finnish and Russian

Acoustic-to-articulatory inversion (AAI) remains challenging under domain shifts where changes in speaker attributes and cross-language conditions often degrade performance. We conduct a systematic evaluation under such shifts and establish baseline benchmarks on FROST-EMA, a Finnish-Russian bilingual EMA corpus. FROST-EMA addresses the English bias and limited speaker diversity of existing resources. We benchmark (i) articulatory targets (raw EMA coordinates vs tract variables), (ii) acoustic front-ends (MFCC vs SSL features), and (iii) inversion back-ends (BiLSTM vs a lightweight attention-based sequence model). We further define evaluation protocols for cross-gender transfer (within language) and cross-language transfer (within gender). The results indicate that cross-gender mismatch introduces moderate Pearson correlation declines (approximately 0.05 to 0.10) relative to the in-domain baseline, whereas cross-language mismatch causes larger drops (approximately 0.10 to 0.20).

eess.AS

Learning Input-Channel Permutation Equivariance for Multi-Channel Source Separation: Reducing Bleeding in Small Music Ensembles

Microphone bleed is a persistent challenge in small ensembles and orchestral recordings, where close microphones intended for individual instruments also capture leakage from nearby sources. This overlap degrades track isolation and complicates mixing. This paper addresses the bleeding problem by making channel-permutation-equivariance a core learning principle. During training, we apply the same random permutation to the input microphone channels and their corresponding reference targets. This discourages reliance on fixed channel-instrument associations and improves robustness to changes in the recording setup and even in the recorded instruments. The proposed model is trained on synthetic ensembles with diverse simulated room acoustics and microphone placements, and evaluated on unseen simulated conditions and real URMP recordings. The results show that permutation-aware training consistently improves SDR and reduces bleeding under unseen conditions compared with non-permutation baselines. The findings highlight permutation-equivariance as a simple, data-centric strategy for robust debleeding and practical multi-channel source separation in music production workflows.

eess.AS

Development of ProtoPol: a medium resolution echelle spectro-polarimeter for PRL telescopes, Mt Abu, India -- Part II : the data-reduction pipeline, on-sky characterization $\&$ performance verification and first science results

We present the development of ProtoPol - a medium resolution echelle spectro-polarimeter for the PRL 1.2m and 2.5m telescopes at Mt Abu observatory, India. In this second and final part of the paper series, we report on the development of a dedicated data reduction pipeline of ProtoPol along with several characterization, performance evaluation, and scientific observations to quantify the performance of the instrument. ProtoPol provides a spectral resolution in the range of $\sim$0.4 - 0.75 angstrom across various orders in the visible wavelength range of 4000-9600 angstrom. On PRL 2.5m telescope, an SNR of 10 is achieved for $m_V\sim13.2$ source in 1 hour of integration time, and its throughput is estimated to be $\sim$6\% including all the contributing factors such as atmospheric transmission, telescope reflectivity, instrument's optics, CCD efficiency etc. ProtoPol achieved a linear polarization accuracy $\delta P \approx 0.1-0.2\%$ in 2 hours of integration time for a source with $m_V\approx8$. The instrumental polarization is determined to be around $0.1\%$. We also present the first science results with ProtoPol to demonstrate the capabilities of the instrument. A sample of Herbig Ae/Be stars, classical Herbig stars, Symbiotic stars, and AGB/post-AGB stars were observed over the period of one and half years for their spectro-polarimetry measurements covering various physical mechanisms such as intrinsic line polarization in Herbig and classical Be stars, Raman scattered features in Symbiotic stars, as well as continuum polarization in AGB/post-AGB stars to verify the polarization performance of the instrument.

astro-ph.IM

WST-X Series: Wavelet Scattering Transform for Interpretable Speech Deepfake Detection

In this work, we focus on front-end design for speech deepfake detectors, the component that determines the discriminative acoustic cues provided to the classifier. Existing approaches are primarily categorized into two types. Hand-crafted filterbank features are transparent but limited in capturing higher-level information. SSL features, in turn, lack interpretability and may overlook fine-grained spectral anomalies. We propose the WST-X series, a novel family of feature extractors that combines the best of both worlds via the wavelet scattering transform (WST), which cascades wavelet convolutions with modulus nonlinearities to produce deformation-stable, multi-scale features. Experiments on the recent Deepfake-Eval-2024 benchmark, together with cross-dataset evaluations on the SpoofCeleb and In-the-Wild, show that WST-X outperforms existing front-ends by a wide margin. Our analysis reveals that a small averaging scale ($J$), combined with high-frequency and directional resolutions ($Q$, $L$), is critical for capturing subtle artifacts. This underscores the value of stable and translation-invariant features for speech deepfake detection. The code is available at https://github.com/xxuan-acoustics/WST-X-Series.

eess.AS

The Need for Ultra High Resolution X-ray Imaging

This paper discusses the broad science case for obtaining milliarcsecond to microarcsecond astronomical imaging resolution in the soft to medium-energy X-ray band (~0.5 to ~8 keV). Astronomy across much of the electromagnetic spectrum has been fundamentally transformed with a rapid increase in ground-based and space-based capabilities to examine celestial objects on small scales that relate directly to their relevant physical processes. X-ray imaging capabilities, however, have fallen far behind observations at longer wavelengths. As such, without decisive advances in X-ray imaging, we will be unable to uncover key phenomena on the smallest astrophysical scales, leaving entire classes of high-energy discoveries beyond our reach. Here we describe several science goals for which high quality X-ray imaging is crucial and the status of some current technologies or mission concepts that would be required for these advances. In particular, we discuss the Accretion Explorer, a mission architecture under current study for a dispersed aperture X-ray interferometer.

astro-ph.HE

Stereo Sound Event Localization and Detection with Onscreen/offscreen Classification

This paper presents the objective, dataset, baseline, and metrics of Task 3 of the DCASE2025 Challenge on sound event localization and detection (SELD). In previous editions, the challenge used four-channel audio formats of first-order Ambisonics (FOA) and microphone array. In contrast, this year's challenge investigates SELD with stereo audio data (termed stereo SELD). This change shifts the focus from more specialized 360{\deg} audio and audiovisual scene analysis to more commonplace audio and media scenarios with limited field-of-view (FOV). Due to inherent angular ambiguities in stereo audio data, the task focuses on direction-of-arrival (DOA) estimation in the azimuth plane (left-right axis) along with distance estimation. The challenge remains divided into two tracks: audio-only and audiovisual, with the audiovisual track introducing a new sub-task of onscreen/offscreen event classification necessitated by the limited FOV. This challenge introduces the DCASE2025 Task3 Stereo SELD Dataset, whose stereo audio and perspective video clips are sampled and converted from the STARSS23 recordings. The baseline system is designed to process stereo audio and corresponding video frames as inputs. In addition to the typical SELD event classification and localization, it integrates onscreen/offscreen classification for the audiovisual track. The evaluation metrics have been modified to introduce an onscreen/offscreen accuracy metric, which assesses the models' ability to identify which sound sources are onscreen. In the experimental evaluation, the baseline system performs reasonably well with the stereo audio data.

cs.SD

Class-Incremental Learning for Sound Event Localization and Detection

This paper investigates the feasibility of class-incremental learning (CIL) for Sound Event Localization and Detection (SELD) tasks. The method features an incremental learner that can learn new sound classes independently while preserving knowledge of old classes. The continual learning is achieved through a mean square error-based distillation loss to minimize output discrepancies between subsequent learners. The experiments are conducted on the TAU-NIGENS Spatial Sound Events 2021 dataset, which includes 12 different sound classes and demonstrate the efficacy of proposed method. We begin by learning 8 classes and introduce the 4 new classes at next stage. After the incremental phase, the system is evaluated on the full set of learned classes. Results show that, for this realistic dataset, our proposed method successfully maintains baseline performance across all metrics.

eess.AS

Fate and detectability of rare gas hydride ions in nova ejecta: A case study with nova templates

HeH$^+$ was the first heteronuclear molecule to form in the metal-free Universe after the Big Bang. The molecule gained significant attention following its first circumstellar detection in the young and dense planetary nebula NGC 7027. We target some hydride ions associated with the noble gases (HeH$^+$, ArH$^+$, and NeH$^+$) to investigate their formation in harsh environments like the nova outburst region. We use a photoionization modeling (based on previously published best-fit physical parameters) of the moderately fast ONe type nova, QU Vulpeculae 1984, and the CO type novae, RS Ophiuchi and V1716 Scorpii. Our steady-state modeling reveals a convincing amount of HeH$^+$, especially in the dense clump of RS Ophiuchi and V1716 Scorpii. The calculated upper limit on the surface brightness of HeH$^+$ transitions suggests that the James Webb Space Telescope (JWST) could detect some of them, particularly in sources like RS Ophiuchi and V1716 Scorpii, which have similar physical and chemical conditions and evolution. It must be clearly noted that the sources studied are used as templates, and not as targets for observations. The detection of these lines could be useful for determining the physical conditions in similar types of systems and for validating our predictions based on new electron-impact ro-vibrational collisional data at temperatures of up to 20,000 K.

astro-ph.GA

A phenomenological study of the evolution of shock-induced O I emission lines in the spectrum of nova V2891 Cygni

The eruption of Nova V2891 Cygni in 2019 offers a rare opportunity to explore the shock-induced processes in novae ejecta. The spectral evolution shows noticeable differences in the evolution of various oxygen emission lines such as O I 7773 {\AA}, O I 8446 {\AA}, O I 1.1286 $\mu$m, O I 1.3164 $\mu$m, etc. Here, we use spectral synthesis code CLOUDY to study the temporal evolution of these oxygen emission lines. Our photoionization model requires the introduction of a component with a very high density ($n ~ 10^{11}$ cm$^{-3}$) and an enhanced oxygen abundance (O/H $\sim$ 28) to produce the O I 7773 {\AA} emission line, suggesting a stratification of material with high oxygen abundance within the ejecta. An important outcome is the behaviour of the O I 1.3164 $\mu$m line, which could only be generated by invoking the collisional ionization models in CLOUDY. Our phenomenological analysis suggests that O I 1.3164 $\mu$m emission originates from a thin, dense shell characterized by a high density of about $10^{12.5} - 10^{12.8}$ cm$^{-3}$, which is most likely formed due to the strong internal collisions. If such is the case, the O I 1.3164 $\mu$m emission presents itself as a tracer of shock-induced dust formation in V2891 Cyg. The collisional ionization models have also been successful in creating the high-temperature conditions ($~ 7.07 - 7.49 \times 10^5$ K) required to reproduce the observed high ionization potential coronal lines, which coincide with the epoch of dust formation and evolution of the O I 1.3164 $\mu$m emission line.

astro-ph.SR

IR-UWB Radar-based Situational Awareness System for Smartphone-Distracted Pedestrians

With the widespread adoption of smartphones, ensuring pedestrian safety on roads has become a critical concern due to smartphone distraction. This paper proposes a novel and real-time assistance system called UWB-assisted Safe Walk (UASW) for obstacle detection and warns users about real-time situations. The proposed method leverages Impulse Radio Ultra-Wideband (IR-UWB) radar embedded in the smartphone, which provides excellent range resolution and high noise resilience using short pulses. We implemented UASW specifically for Android smartphones with IR-UWB connectivity. The framework uses complex Channel Impulse Response (CIR) data to integrate rule-based obstacle detection with artificial neural network (ANN) based obstacle classification. The performance of the proposed UASW system is analyzed using real-time collected data. The results show that the proposed system achieves an obstacle detection accuracy of up to 97% and obstacle classification accuracy of up to 95% with an inference delay of 26.8 ms. The results highlight the effectiveness of UASW in assisting smartphone-distracted pedestrians and improving their situational awareness.

cs.CV

Study of the fastest classical nova, V1674 Her: Photoionization and Morpho-kinemetic model analysis

We present the results of the investigation of the nova V1674 Her (2021), recognised as the swiftest classical nova, with $t_2 \sim 0.90$ days. The distance to the nova is estimated to be 4.97 kpc. The mass and radius of the WD are calculated to be $\sim~1.36~M_\odot$ and $\sim 0.15~R_\oplus$, respectively. Over the course of one month following the outburst, V1674 Her traversed distinct phases -- pre-maxima, early decline, nebular, and coronal -- displaying a remarkably swift transformation. The nebular lines emerged on day 10.00, making it the classical nova with the earliest observed commencement to date. We modelled the observed optical spectrum using the photoionization code \textsc{cloudy}. From the best-fitting model we deduced different physical and chemical parameters associated withe the system. The temperature and luminosity of the central ionizing sources are found in the range of $1.99 - 2.34~\times 10^5$ K and $1.26 - 3.16~ \times 10^{38}$ \ergs, respectively. Elements such as He, O, N, and Ne are found to be overabundant compared to solar abundance in both the nebular and coronal phases. According to the model, Fe II abundance diminishes while Ne abundance increases, potentially elucidating the rare hybrid transition between Fe and He/N nova classes. The ejected mass across all epochs spanned from $3.42 - 7.04~ \times 10^{-5}~M_\odot$. Morpho-kinematic modelling utilising \textsc{shape} revealed that the nova V1674 Her possesses a bipolar structure with an equatorial ring at the centre and an inclination angle of i = 67$\pm$ 1.5$^{\circ}$.

astro-ph.SR

Localization of DOA trajectories -- Beyond the grid

The direction of arrival (DOA) estimation algorithms are crucial in localizing acoustic sources. Traditional localization methods rely on block-level processing to extract the directional information from multiple measurements processed together. However, these methods assume that DOA remains constant throughout the block, which may not be true in practical scenarios. Also, the performance of localization methods is limited when the true parameters do not lie on the parameter search grid. In this paper we propose two trajectory models, namely the polynomial and bandlimited trajectory models, to capture the DOA dynamics. To estimate trajectory parameters, we adopt two gridless algorithms: i) Sliding Frank-Wolfe (SFW), which solves the Beurling LASSO problem and ii) Newtonized Orthogonal Matching Pursuit (NOMP), which improves over OMP using cyclic refinement. Furthermore, we extend our analysis to include wideband processing. The simulation results indicate that the proposed trajectory localization algorithms exhibit improved performance compared to grid-based methods in terms of resolution, robustness to noise, and computational efficiency.

eess.AS

Improving trajectory localization accuracy via direction-of-arrival derivative estimation

Sound source localization is crucial in acoustic sensing and monitoring-related applications. In this paper, we do a comprehensive analysis of improvement in sound source localization by combining the direction of arrivals (DOAs) with their derivatives which quantify the changes in the positions of sources over time. This study uses the SALSA-Lite feature with a convolutional recurrent neural network (CRNN) model for predicting DOAs and their first-order derivatives. An update rule is introduced to combine the predicted DOAs with the estimated derivatives to obtain the final DOAs. The experimental validation is done using TAU-NIGENS Spatial Sound Events (TNSSE) 2021 dataset. We compare the performance of the networks predicting DOAs with derivative vs. the one predicting only the DOAs at low SNR levels. The results show that combining the derivatives with the DOAs improves the localization accuracy of moving sources.

eess.AS

Study of 2021 outburst of the recurrent nova RS Ophiuchi: Photoionization and morpho-kinematic modelling

We present the evolution of the optical spectra of the 2021 outburst of RS Ophiuchi (RS Oph) over about a month after the outburst. The spectral evolution is similar to the previous outbursts. Early spectra show prominent P Cygni profiles of hydrogen Balmer, \ion{Fe}{ii}, and \ion{He}{i} lines. The emission lines were very broad during the initial days, which later became narrower and sharper as the nova evolved. This is interpreted as the expanding shocked material into the winds of the red giant companion. We find that the nova ejecta expanded freely for $\sim 4$ days, and afterward, the shock velocity decreased monotonically with time as $v\propto t^{-0.6}$. The physical and chemical parameters associated with the system are derived using the photoionization code \textsc{cloudy}. The best-fit \textsc{cloudy} model shows the presence of a hot central white dwarf source with a roughly constant luminosity of $\sim$1.00 $\times$ 10$^{37}$ erg s$^{-1}$. The best-fit photoionization models yield absolute abundance values by number, relative to solar of He/H $\sim 1.4 - 1.9$, N/H = $70 - 95$, O/H = $0.60 - 2.60$, and Fe/H $\sim 1.0 - 1.9$ for the ejecta during the first month after the outburst. Nitrogen is found to be heavily overabundant in the ejecta. The ejected hydrogen shell mass of the system is estimated to be in the range of $3.54 - 3.83 \times 10^{-6} M_{\odot}$. The 3D morpho-kinematic modelling shows a bipolar morphology and an inclination angle of $i=30^{\circ}$ for the RS Oph binary system.

astro-ph.SR

Parametric Models for DOA Trajectory Localization

Directions of arrival (DOA) estimation or localization of sources is an important problem in many applications for which numerous algorithms have been proposed. Most localization methods use block-level processing that combines multiple data snapshots to estimate DOA within a block. The DOAs are assumed to be constant within the block duration. However, these assumptions are often violated due to source motion. In this paper, we propose a signal model that captures the linear variations in DOA within a block. We applied conventional beamforming (CBF) algorithm to this model to estimate linear DOA trajectories. Further, we formulate the proposed signal model as a block sparse model and subsequently derive sparse Bayesian learning (SBL) algorithm. Our simulation results show that this linear parametric DOA model and corresponding algorithms capture the DOA trajectories for moving sources more accurately than traditional signal models and methods.

eess.SP

Photoionization modeling of the dusty nova V1280 Scorpii

We perform photoionization modeling of the dusty nova V1280 Scorpii (V1280 Sco) with an aim to study the changes in the physical and chemical parameters. We model pre and post dust phase, optical and near-Infrared (NIR), spectra using the photoionization code \textsc{cloudy}, v.17.02, considering a two-component (low density and high density region) model. From the best-fit model, we find that the temperature and luminosity of the central ionizing source in the pre-dust phase are in the range 1.32 - 1.50 $\times 10^4$ K and 2.95 - 3.16 $\times 10^{36}$ ergs$^{-1}$, respectively, which increase to 1.58 - 1.62 $\times 10^4$ K and 3.23 - 3.31 $\times 10^{36}$ ergs$^{-1}$, respectively, in the post-dust phase. It is found that a very high hydrogen density ($\sim 10^{13} - 10^{14}$ cm$^{-3}$) is required for the generation of spectra properly. Dust condensation conditions are achieved at high ejecta density ($\sim 3.16 \times 10^{8}$cm$^{-3}$) and low temperature ($\sim$2000 K) in the outer region of the ejecta. It is found that a mixture of small (0.005 - 0.25$\mu$m) amorphous carbon dust grains and large (0.03 - 3.0$\mu$m) astrophysical silicate dust grains iis present n the ejecta in the post-dust phase. Our model yields very high elemental abundance values as C/H = 13.5 - 20, N/H = 250, O/H = 27 - 35, by number, relative to solar in the ejecta, during the pre-dust phase, which decrease in the post-dust phase.

astro-ph.SR