arXiv ScienceSearch

arXiv subjects

Aryan Sharma

Publications and source records attributed to Aryan Sharma.

8 recordsLinked to original sources

Characterizing Paraphrase-Induced Failures in Lean 4 Autoformalization

Lean 4 autoformalization has become increasingly popular in recent years, with frontier language models and open-weight autoformalizers now producing valid formalizations of mathematical theorems. However, these evaluations often rely on single canonical phrasings of theorems and rarely probe whether outputs are robust to natural variation in inputs, while prior work has shown that semantically equivalent paraphrases often induce divergent formal outputs. We study the structure of these divergences in Lean 4 by applying deterministic paraphrase rules to datasets of undergraduate and Olympiad-level math problems. Across four frontier models and three open-weight autoformalizers, we find that paraphrase sensitivity is dominated by failures at the code-generation layer, and that these failures are typed differently by dataset. Furthermore, these patterns generalize to open-weight models, showing that state-of-the-art autoformalizers still struggle to generate valid Lean code. Our results provide a failure-mode taxonomy for autoformalization and motivate training-time interventions targeted at specific compilation failures.

cs.LG

Dissociating Decodability and Causal Use in Bracket-Sequence Transformers

When trained on tasks requiring an understanding of hierarchical structure, transformers have been found to represent this hierarchy in distinct ways: in the geometry of the residual stream, and in stack-like attention patterns maintaining a last-in, first-out ordering. However, it remains unclear whether these representations are causally used or merely decodable. We examine this gap in transformers trained on the Dyck language (a formal language of balanced bracket sequences), where the hierarchical ground truth is explicit. By probing and intervening on the residual stream and attention patterns, we find that depth, distance, and top-of-stack signals are all decodable, yet their causal roles diverge. Specifically, masking attention to the true top-of-stack position causes a sharp drop in long-distance accuracy, while ablating low-dimensional residual stream subspaces has comparatively little effect. These results, which extend to a templated natural language setting, suggest that even in a controlled setting where the relevant hierarchical variables are known, decodability alone does not imply causal use.

cs.CL

H-Probes: Extracting Hierarchical Structures From Latent Representations of Language Models

Representing and navigating hierarchy is a fundamental primitive of reasoning. Large language models have demonstrated proficiency in a wide variety of tasks requiring hierarchical reasoning, but there exists limited analysis on how the models geometrically represent the necessary latent constructions for such thinking. To this end, we develop H-probes, a collection of linear probes that extract hierarchical structure, specifically depth and pairwise distance, from latent representations. In synthetic tree traversal tasks, the H-probes robustly find the subspaces containing hierarchical structure necessary to complete the tasks; furthermore, in comprehensive ablation experiments, we show that these hierarchy-containing subspaces are low-dimensional, causally important for high task performance, and generalize within- and out-of-domain. Furthermore, we find analogous, though weaker, hierarchical structure in real-world hierarchical contexts such as mathematical reasoning traces. These results demonstrate that models represent hierarchy not only at the level of syntax and concepts, but at deeper levels of abstraction -- including the reasoning process itself.

cs.CL

Quantifying Global Networks of Exchange through the Louvain Method

Congressional Research Service (CRS) reports provide detailed analyses of major policy issues to members of the US Congress. We extract and analyze data from 2,010 CRS reports written between 1996 and 2024 to quantify inter-country relationships, representing 172 countries as nodes and 4,137 shared interests as edges within a weighted, bidirectional network. Through the Louvain method, we extract non-overlapping communities from our network and identify clusters with shared interests. We then compute the eigenvector centrality of countries to highlight their network influence. The results of this work could enable improvements in sourcing evidence for analytic products and understanding the connectivity of our world.

cs.SI

Non-Linear behavior of the Electron Cyclotron Drift Instability and the Suppression of Anomalous Current

We present results of one-dimensional collisionless simulations of plasma turbulence and related anomalous electron current of the Electron Cyclotron Drift Instability (ECDI). Our highly resolved, long-term simulations of xenon plasma in the magnetic field performed with the WarpX particle-in-cell (PIC) code show several intermediate non-linear stages before the system enters a stationary state with significantly increased electron temperature and a finite level of energy in the electrostatic fluctuations. In early and intermediate non-linear stages, the fluctuations are driven by the electron cyclotron resonances gradually shifting from higher ($m>1$) modes to the fundamental $m=1$ resonance. Enhanced resonant growth is observed from the point when the cyclotron $m=1$ mode coincides with the most unstable ion-acoustic mode. In the final stage, the anomalous electron current existing in intermediate stages is quenched to zero. Following this quenching, our simulations reveal a transition from ECDI-driven dynamics to saturated ion-acoustic turbulence. The modification of the electron and ion distribution functions and their roles in the non-linear developments and saturation of the instability are analyzed at different non-linear stages. The non-linear development of ECDI driven by the $\mathbf{E} \times \mathbf{B}$ electron drift from the applied current and the ECDI driven by the ion beam perpendicular to the magnetic field are compared and characterized as two perspectives of the instability, observed through different Doppler-shifted frames. An extension of this work incorporating full the dynamics of magnetized ions for ECDI driven by a hydrogen ion beam is shown to develop full beam inversion, with the periodic bursts of growth-saturation cycles of ECDI.

physics.plasm-ph

Deep Learning for Wildfire Risk Prediction: Integrating Remote Sensing and Environmental Data

Wildfires pose a significant threat to ecosystems, wildlife, and human communities, leading to habitat destruction, pollutant emissions, and biodiversity loss. Accurate wildfire risk prediction is crucial for mitigating these impacts and safeguarding both environmental and human health. This paper provides a comprehensive review of wildfire risk prediction methodologies, with a particular focus on deep learning approaches combined with remote sensing. We begin by defining wildfire risk and summarizing the geographical distribution of related studies. In terms of data, we analyze key predictive features, including fuel characteristics, meteorological and climatic conditions, socioeconomic factors, topography, and hydrology, while also reviewing publicly available wildfire prediction datasets derived from remote sensing. Additionally, we emphasize the importance of feature collinearity assessment and model interpretability to improve the understanding of prediction outcomes. Regarding methodology, we classify deep learning models into three primary categories: time-series forecasting, image segmentation, and spatiotemporal prediction, and further discuss methods for converting model outputs into risk classifications or probability-adjusted predictions. Finally, we identify the key challenges and limitations of current wildfire-risk prediction models and outline several research opportunities. These include integrating diverse remote sensing data, developing multimodal models, designing more computationally efficient architectures, and incorporating cross-disciplinary methods--such as coupling with numerical weather-prediction models--to enhance the accuracy and robustness of wildfire-risk assessments.

cs.LG

Equilibrium states of Burgers and KdV equations

We simulate KdV and dissipation-less Burgers equations using delta-correlated random noise as initial condition. We observe that the energy fluxes of the two equations remain zero throughout, thus indicating their equilibrium nature. We characterize the equilibrium states using Gaussian probability distribution for the real space field, and using Boltzmann distribution for the modal energy. We show that the single soliton of the KdV equation too exhibits zero energy flux, hence it is in equilibrium. We argue that the energy flux is a good measure for ascertaining whether a system is in equilibrium or not.

cond-mat.stat-mech

Lightweight Multi-Drone Detection and 3D-Localization via YOLO

In this work, we present and evaluate a method to perform real-time multiple drone detection and three-dimensional localization using state-of-the-art tiny-YOLOv4 object detection algorithm and stereo triangulation. Our computer vision approach eliminates the need for computationally expensive stereo matching algorithms, thereby significantly reducing the memory footprint and making it deployable on embedded systems. Our drone detection system is highly modular (with support for various detection algorithms) and capable of identifying multiple drones in a system, with real-time detection accuracy of up to 77\% with an average FPS of 332 (on Nvidia Titan Xp). We also test the complete pipeline in AirSim environment, detecting drones at a maximum distance of 8 meters, with a mean error of $23\%$ of the distance. We also release the source code for the project, with pre-trained models and the curated synthetic stereo dataset.

cs.CV