arXiv ScienceSearch

arXiv subjects

Devansh Agarwal

Publications and source records attributed to Devansh Agarwal.

At least 19 recordsLinked to original sources

Neural-Spectral Discovery of Rotating Black Holes Beyond General Relativity

Finding rotating black hole solutions in higher-curvature theories of gravity is a problem of fundamental importance. Virtually every approach to reconcile gravity with quantum mechanics predicts corrections to the Einstein-Hilbert action, yet no systematic solution-generating method exists for the stationary sector. We close this gap with {\sc Akribeia}, a novel hybrid framework that pairs physics-informed neural networks with a pseudo-spectral refinement step, yielding certified neural-field rotating black hole solutions -- continuous, globally defined functions, parametric in the coupling constants -- whose residuals against the field equations are verified to extreme precision. We apply the method to theories quadratic and cubic in the curvature and construct, for the first time, families of rotating black holes featuring multiple non-vanishing angular momenta, parametric in the new coupling constants. After validating against previously known five-dimensional spacetimes, we present new solutions in scenarios leading to a highly non-linear/non-perturbative coupled system of ordinary differential equations. Our method can be systematically adapted to other setups involving partial differential equations as well.

gr-qc

FinBalance: A Multi-Document Accounting Reconciliation Benchmark

Existing financial-NLP benchmarks mostly evaluate prepared artifacts such as filings, tables, or extracted values. Real accounting begins earlier: source documents must be reconciled into cited journal entries, aggregated into a balance sheet, and checked for contradictions. We introduce FinBalance, a multi-document accounting reconciliation benchmark built from source-document bundles across eight industries, three period types, and five difficulty levels. Human-authored business scenarios, accounting policies, tax/FX treatments, document schemas, distractors, and inconsistency templates are composed by a deterministic generator whose ledger produces journal entries,balance sheets, and 23 inconsistency-code labels. On a 710-record evaluation split, six contemporary LLMs reach at most 46% exact final-balance-sheet accuracy. Four models show a 26-41 pp gap between BS_exact, the model's reported balance sheet, and BS_recon, the balance sheet obtained by replaying its entries through our ledger. Models often recover numerically plausible entries but fail to bind them to supporting documents and aggregate them consistently. Citation-pressure prompting barely changes document-linking errors, while ledger-feedback ablations substantially improve reported balance sheets and expose inconsistency-detection trade-offs. Expert finance reviewers validate the benchmark design and labels.

cs.CL

The RRATalog: a Galactic census of rotating radio transients

Rotating radio transients (RRATs) represent a significant but poorly understood component of the Galactic neutron star population, characterized by sporadic emission first detectable only through single-pulse searches. We present the RRATalog, an up-to-date catalogue of 335 RRATs, and utilize a uniform sample of RRATs discovered in four Parkes telescope surveys to model their Galactic population. Accounting in detail for observational selection effects, we find a radial density profile similar to pulsars, but identify a significantly steeper luminosity function (power-law index $\alpha \simeq -1.3$) than previously assumed. For sources beaming towards Earth, we estimate $34000 \pm 1600$ potentially observable RRATs above a peak luminosity of 30 mJy kpc$^2$. At these high luminosities, the RRAT population is comparable in size to that of canonical pulsars. Consistent with the observed distribution, the underlying period distribution is significantly shifted toward longer periods compared to canonical pulsars, suggesting RRATs represent a more evolved population. We find evidence for a turnover in the luminosity function below 30 mJy kpc$^2$, and predict that the total number of potentially observable RRATs is $\lesssim 70,000$. Applying the Tauris \& Manchester beaming model, we estimate the total Galactic RRAT population to be $\lesssim 400,000$. The implied birth rate of $\lesssim 1.4$ RRATs per century is consistent with the Galactic core-collapse supernova rate, suggesting RRATs can be reconciled with known progenitor rates without requiring a separate evolutionary origin. We provide predictions for RRAT discoveries in ongoing and future surveys.

astro-ph.HE

LogSyn: A Few-Shot LLM Framework for Structured Insight Extraction from Unstructured General Aviation Maintenance Logs

Aircraft maintenance logs hold valuable safety data but remain underused due to their unstructured text format. This paper introduces LogSyn, a framework that uses Large Language Models (LLMs) to convert these logs into structured, machine-readable data. Using few-shot in-context learning on 6,169 records, LogSyn performs Controlled Abstraction Generation (CAG) to summarize problem-resolution narratives and classify events within a detailed hierarchical ontology. The framework identifies key failure patterns, offering a scalable method for semantic structuring and actionable insight extraction from maintenance logs. This work provides a practical path to improve maintenance workflows and predictive analytics in aviation and related industries.

cs.LG

Reinforcement Learning for Self-Healing Material Systems

The transition to autonomous material systems necessitates adaptive control methodologies to maximize structural longevity. This study frames the self-healing process as a Reinforcement Learning (RL) problem within a Markov Decision Process (MDP), enabling agents to autonomously derive optimal policies that efficiently balance structural integrity maintenance against finite resource consumption. A comparative evaluation of discrete-action (Q-learning, DQN) and continuous-action (TD3) agents in a stochastic simulation environment revealed that RL controllers significantly outperform heuristic baselines, achieving near-complete material recovery. Crucially, the TD3 agent utilizing continuous dosage control demonstrated superior convergence speed and stability, underscoring the necessity of fine-grained, proportional actuation in dynamic self-healing applications.

cs.LG

Semantic Anchoring in Agentic Memory: Leveraging Linguistic Structures for Persistent Conversational Context

Large Language Models (LLMs) have demonstrated impressive fluency and task competence in conversational settings. However, their effectiveness in multi-session and long-term interactions is hindered by limited memory persistence. Typical retrieval-augmented generation (RAG) systems store dialogue history as dense vectors, which capture semantic similarity but neglect finer linguistic structures such as syntactic dependencies, discourse relations, and coreference links. We propose Semantic Anchoring, a hybrid agentic memory architecture that enriches vector-based storage with explicit linguistic cues to improve recall of nuanced, context-rich exchanges. Our approach combines dependency parsing, discourse relation tagging, and coreference resolution to create structured memory entries. Experiments on adapted long-term dialogue datasets show that semantic anchoring improves factual recall and discourse coherence by up to 18% over strong RAG baselines. We further conduct ablation studies, human evaluations, and error analysis to assess robustness and interpretability.

cs.CL

EchoGuide: Active Acoustic Guidance for LLM-Based Eating Event Analysis from Egocentric Videos

Self-recording eating behaviors is a step towards a healthy lifestyle recommended by many health professionals. However, the current practice of manually recording eating activities using paper records or smartphone apps is often unsustainable and inaccurate. Smart glasses have emerged as a promising wearable form factor for tracking eating behaviors, but existing systems primarily identify when eating occurs without capturing details of the eating activities (E.g., what is being eaten). In this paper, we present EchoGuide, an application and system pipeline that leverages low-power active acoustic sensing to guide head-mounted cameras to capture egocentric videos, enabling efficient and detailed analysis of eating activities. By combining active acoustic sensing for eating detection with video captioning models and large-scale language models for retrieval augmentation, EchoGuide intelligently clips and analyzes videos to create concise, relevant activity records on eating. We evaluated EchoGuide with 9 participants in naturalistic settings involving eating activities, demonstrating high-quality summarization and significant reductions in video data needed, paving the way for practical, scalable eating activity tracking.

cs.HC

SonicID: User Identification on Smart Glasses with Acoustic Sensing

Smart glasses have become more prevalent as they provide an increasing number of applications for users. They store various types of private information or can access it via connections established with other devices. Therefore, there is a growing need for user identification on smart glasses. In this paper, we introduce a low-power and minimally-obtrusive system called SonicID, designed to authenticate users on glasses. SonicID extracts unique biometric information from users by scanning their faces with ultrasonic waves and utilizes this information to distinguish between different users, powered by a customized binary classifier with the ResNet-18 architecture. SonicID can authenticate users by scanning their face for 0.06 seconds. A user study involving 40 participants confirms that SonicID achieves a true positive rate of 97.4%, a false positive rate of 4.3%, and a balanced accuracy of 96.6% using just 1 minute of training data collected for each new user. This performance is relatively consistent across different remounting sessions and days. Given this promising performance, we further discuss the potential applications of SonicID and methods to improve its performance in the future.

cs.HC

MunchSonic: Tracking Fine-grained Dietary Actions through Active Acoustic Sensing on Eyeglasses

We introduce MunchSonic, an AI-powered active acoustic sensing system integrated into eyeglasses to track fine-grained dietary actions. MunchSonic emits inaudible ultrasonic waves from the eyeglass frame, with the reflected signals capturing detailed positions and movements of body parts, including the mouth, jaw, arms, and hands involved in eating. These signals are processed by a deep learning pipeline to classify six actions: hand-to-mouth movements for food intake, chewing, drinking, talking, face-hand touching, and other activities (null). In an unconstrained study with 12 participants, MunchSonic achieved a 93.5% macro F1-score in a user-independent evaluation with a 2-second resolution in tracking these actions, also demonstrating its effectiveness in tracking eating episodes and food intake frequency within those episodes.

cs.HC

ActSonic: Recognizing Everyday Activities from Inaudible Acoustic Wave Around the Body

We present ActSonic, an intelligent, low-power active acoustic sensing system integrated into eyeglasses that can recognize 27 different everyday activities (e.g., eating, drinking, toothbrushing) from inaudible acoustic waves around the body. It requires only a pair of miniature speakers and microphones mounted on each hinge of the eyeglasses to emit ultrasonic waves, creating an acoustic aura around the body. The acoustic signals are reflected based on the position and motion of various body parts, captured by the microphones, and analyzed by a customized self-supervised deep learning framework to infer the performed activities on a remote device such as a mobile phone or cloud server. ActSonic was evaluated in user studies with 19 participants across 19 households to track its efficacy in everyday activity recognition. Without requiring any training data from new users (leave-one-participant-out evaluation), ActSonic detected 27 activities, achieving an average F1-score of 86.6% in fully unconstrained scenarios and 93.4% in prompted settings at participants' homes.

cs.HC

Ring-a-Pose: A Ring for Continuous Hand Pose Tracking

We present Ring-a-Pose, a single untethered ring that tracks continuous 3D hand poses. Located in the center of the hand, the ring emits an inaudible acoustic signal that each hand pose reflects differently. Ring-a-Pose imposes minimal obtrusions on the hand, unlike multi-ring or glove systems. It is not affected by the choice of clothing that may cover wrist-worn systems. In a series of three user studies with a total of 30 participants, we evaluate Ring-a-Pose's performance on pose tracking and micro-finger gesture recognition. Without collecting any training data from a user, Ring-a-Pose tracks continuous hand poses with a joint error of 14.1mm. The joint error decreases to 10.3mm for fine-tuned user-dependent models. Ring-a-Pose recognizes 7-class micro-gestures with a 90.60% and 99.27% accuracy for user-independent and user-dependent models, respectively. Furthermore, the ring exhibits promising performance when worn on any finger. Ring-a-Pose enables the future of smart rings to track and recognize hand poses using relatively low-power acoustic sensing.

cs.HC

EchoWrist: Continuous Hand Pose Tracking and Hand-Object Interaction Recognition Using Low-Power Active Acoustic Sensing On a Wristband

Our hands serve as a fundamental means of interaction with the world around us. Therefore, understanding hand poses and interaction context is critical for human-computer interaction. We present EchoWrist, a low-power wristband that continuously estimates 3D hand pose and recognizes hand-object interactions using active acoustic sensing. EchoWrist is equipped with two speakers emitting inaudible sound waves toward the hand. These sound waves interact with the hand and its surroundings through reflections and diffractions, carrying rich information about the hand's shape and the objects it interacts with. The information captured by the two microphones goes through a deep learning inference system that recovers hand poses and identifies various everyday hand activities. Results from the two 12-participant user studies show that EchoWrist is effective and efficient at tracking 3D hand poses and recognizing hand-object interactions. Operating at 57.9mW, EchoWrist is able to continuously reconstruct 20 3D hand joints with MJEDE of 4.81mm and recognize 12 naturalistic hand-object interactions with 97.6% accuracy.

cs.HC

The Petabyte Project

Transient radio sources, such as fast radio bursts, intermittent pulsars, and rotating radio transients, can offer a wealth of information regarding extreme emission physics as well as the intervening interstellar and/or intergalactic medium. Vital steps towards understanding these objects include characterizing their source populations and estimating their event rates across observing frequencies. However, previous efforts have been undertaken mostly by individual survey teams at disparate observing frequencies and telescopes, and with non-uniform algorithms for searching and characterization. The Petabyte Project (TPP) aims to address these issues by uniformly reprocessing data from several petabytes of radio transient surveys covering two decades of observing frequency (300 MHz-20 GHz). The TPP will provide robust event rate analyses, in-depth assessment of survey and pipeline completeness, as well as revealing discoveries from archival and ongoing radio surveys. We present an overview of TPP's processing pipeline, scope, and our potential to make new discoveries.

astro-ph.IM

Masked Image Modeling Advances 3D Medical Image Analysis

Recently, masked image modeling (MIM) has gained considerable attention due to its capacity to learn from vast amounts of unlabeled data and has been demonstrated to be effective on a wide variety of vision tasks involving natural images. Meanwhile, the potential of self-supervised learning in modeling 3D medical images is anticipated to be immense due to the high quantities of unlabeled images, and the expense and difficulty of quality labels. However, MIM's applicability to medical images remains uncertain. In this paper, we demonstrate that masked image modeling approaches can also advance 3D medical images analysis in addition to natural images. We study how masked image modeling strategies leverage performance from the viewpoints of 3D medical image segmentation as a representative downstream task: i) when compared to naive contrastive learning, masked image modeling approaches accelerate the convergence of supervised training even faster (1.40$\times$) and ultimately produce a higher dice score; ii) predicting raw voxel values with a high masking ratio and a relatively smaller patch size is non-trivial self-supervised pretext-task for medical images modeling; iii) a lightweight decoder or projection head design for reconstruction is powerful for masked image modeling on 3D medical images which speeds up training and reduce cost; iv) finally, we also investigate the effectiveness of MIM methods under different practical scenarios where different image resolutions and labeled data ratios are applied.

cs.CV

Comprehensive analysis of a dense sample of FRB 121102 bursts

We present an analysis of a densely repeating sample of bursts from the first repeating fast radio burst, FRB 121102. We reanalysed the data used by Gourdji et al. (2019) and detected 93 additional bursts using our single-pulse search pipeline. In total, we detected 133 bursts in three hours of data at a center frequency of 1.4 GHz using the Arecibo telescope, and develop robust modeling strategies to constrain the spectro-temporal properties of all the bursts in the sample. Most of the burst profiles show a scattering tail, and burst spectra are well modeled by a Gaussian with a median width of 230 MHz. We find a lack of emission below 1300 MHz, consistent with previous studies of FRB 121102. We also find that the peak of the log-normal distribution of wait times decreases from 207 s to 75 s using our larger sample of bursts, as compared to that of Gourdji et al. (2019). Our observations do not favor either Poissonian or Weibull distributions for the burst rate distribution. We searched for periodicity in the bursts using multiple techniques but did not detect any significant period. The cumulative burst energy distribution exhibits a broken power-law shape, with the lower and higher-energy slopes of $-0.4\pm0.1$ and $-1.8\pm0.2$, with the break at $(2.3\pm0.2)\times 10^{37}$ ergs. We provide our burst fitting routines as a python package BURSTFIT that can be used to model the spectrogram of any complex FRB or pulsar pulse using robust fitting techniques. All the other analysis scripts and results are publicly available.

astro-ph.HE

Your: Your Unified Reader

The advancement in signal processing and GPU based systems has enabled new transient detectors at various telescopes to perform much more sensitive searches than their predecessors. Typically the data output from the telescopes is in one of the two commonly used formats: psrfits and Sigproc filterbank. Software developed for transient searches often only works with one of these two formats, limiting their general applicability. Therefore, researchers have to write custom scripts to read/write the data in their format of choice before they can begin any data analysis relevant for their research. \textsc{Your} (Your Unified Reader) is a python-based library that unifies the data processing across multiple commonly used formats. \textsc{Your} implements a user-friendly interface to read and write in the data format of choice. It also generates unified metadata corresponding to the input data file for a quick understanding of observation parameters and provides utilities to perform common data analysis operations. \textsc{Your} also provides several state-of-the-art radio frequency interference mitigation (RFI) algorithms, which can now be used during any stage of data processing (reading, writing, etc.) to filter out artificial signals.

astro-ph.IM

A targeted search for repeating fast radio bursts associated with gamma-ray bursts

The origin of fast radio bursts (FRBs) still remains a mystery, even with the increased number of discoveries in the last three years. Growing evidence suggests that some FRBs may originate from magnetars. Large, single-dish telescopes such as Arecibo Observatory (AO) and Green Bank Telescope (GBT) have the sensitivity to detect FRB~121102-like bursts at gigaparsec distances. Here we present searches using AO and GBT that aimed to find potential radio bursts at 11 sites of past $\gamma$--ray bursts that show evidence for the birth of a magnetar. We also performed a search towards GW170817, which has a merger remnant whose nature remains uncertain. We place $10\,\sigma$ fluence upper limits of $\approx 0.036$ Jy ms at 1.4 GHz and $\approx 0.063$ Jy ms at 4.5 GHz for AO data and fluence upper limits of $\approx 0.085$ Jy ms at 1.4 GHz and $\approx 0.098$ Jy ms at 1.9 GHz for GBT data, for a maximum pulse width of $\approx 42$ ms. The AO observations had sufficient sensitivity to detect any FRB of similar luminosity to the one recently detected from the Galactic magnetar SGR 1935+2154. Assuming a Schechter function for the luminosity function of FRBs, we find that our non-detections favor a steep power--law index ($\alpha\lesssim-1.0$) and a large cut--off luminosity ($L_0 \gtrsim 10^{42}$ erg/s).

astro-ph.HE

On Adversarial Robustness: A Neural Architecture Search perspective

Adversarial robustness of deep learning models has gained much traction in the last few years. Various attacks and defenses are proposed to improve the adversarial robustness of modern-day deep learning architectures. While all these approaches help improve the robustness, one promising direction for improving adversarial robustness is unexplored, i.e., the complex topology of the neural network architecture. In this work, we address the following question: Can the complex topology of a neural network give adversarial robustness without any form of adversarial training?. We answer this empirically by experimenting with different hand-crafted and NAS-based architectures. Our findings show that, for small-scale attacks, NAS-based architectures are more robust for small-scale datasets and simple tasks than hand-crafted architectures. However, as the size of the dataset or the complexity of task increases, hand-crafted architectures are more robust than NAS-based architectures. Our work is the first large-scale study to understand adversarial robustness purely from an architectural perspective. Our study shows that random sampling in the search space of DARTS (a popular NAS method) with simple ensembling can improve the robustness to PGD attack by nearly~12\%. We show that NAS, which is popular for achieving SoTA accuracy, can provide adversarial accuracy as a free add-on without any form of adversarial training. Our results show that leveraging the search space of NAS methods with methods like ensembles can be an excellent way to achieve adversarial robustness without any form of adversarial training. We also introduce a metric that can be used to calculate the trade-off between clean accuracy and adversarial robustness. Code and pre-trained models will be made available at \url{https://github.com/tdchaitanya/nas-robustness}

cs.LG