arXiv Science⌕ Search

arXiv subjects

Abhishek Tiwari

Publications and source records attributed to Abhishek Tiwari.

At least 19 recordsLinked to original sources

The Psychological Costs of Artificial Intelligence Adoption in Software Engineering

Artificial intelligence (AI) is increasingly used to augment software engineering (SE) workflows. While code generation remains the main use case, organizations are actively seeking AI integration in other practices such as test cases generation and code reviews. Organizational AI adoption strategies seem to focus on tangible outcomes such as productivity. However, AI is a disruptive force, introduced into settings where role identity, team norms, and the sources of job satisfaction were well established before the recent advances in generative AI. Historically, technological disruptions have caused psychological and social strains in workplaces, ranging from anxiety and eroded meaning to deskilling and disrupted professional identities. The assumption that AI for SE is cost-free may not be accurate. Therefore, in this study we sought to understand the psychological costs software professionals experience during organizational AI adoption. We carried out a case study in a large software development services company, one year after the company launched its AI adoption. We collected qualitative data through meetings and semi-structured interviews (N = 21). We found that software professionals experience accountability anxiety, craft identity disruption, meaning and satisfaction erosion, cognitive and workload intensification, and uncertainty distress. Practitioners manage these costs through practices that restore control, mitigate them through protective and identity-preserving adaptations, or absorb them, carrying what neither can resolve. We contribute to AI-human collaboration in SE by repositioning AI adoption as a human transition, not only a technological and organizational one.

cs.SE↗

A Measurement Study on the Adoption of Pledges and Unveils in the OpenBSD Operating System

The paper presents a longitudinal measurement study on the adoption of the pledge and unveil system calls in OpenBSD. These system calls are used to sandbox programs and libraries. Given a dataset covering 19 releases, many programs and libraries were modified to use the system calls already before their introductions in official releases. The adoption rates have also steadily grown; a linear trend provides a coarse but sensible heuristic. Although particularly programs residing in /usr/bin and /usr/sbin have been modified to use the system calls, the sizes of programs and libraries do not correlate well with the amounts of pledge and unveil system calls invoked. Regarding the pledges made, standard input and output operations have frequently been requested, although the full fine-grained arsenal offered by pledge has generally been utilized in OpenBSD. The same observation is seen in that particularly read operations to given paths have frequently been unveiled. All in all, the measurement results indicate that the adoption of system call minimization and sandboxing techniques is not necessarily as troublesome as has often been discussed in the literature.

cs.SE↗

Empirical Derivations from an Evolving Test Suite

The paper presents a longitudinal empirical analysis of the automated, continuous, and virtualization-based software test suite of the NetBSD operating system. The longitudinal period observed spans from the initial roll out of the test suite in the early 2010s to late 2025. According to the results, the test suite has grown continuously, currently covering over ten thousand individual test cases. Failed test cases exhibit overall stability, although there have been shorter periods marked with more frequent failures. A similar observation applies to build failures, failures of the test suite to complete, and installation failures, all of which are also captured by the NetBSD's testing framework. Finally, code churn and kernel modifications do not provide longitudinally consistent statistical explanations for the failures. Although some periods exhibit larger effects, including particularly with respect to the kernel modifications, the effects are small on average. Even though only in an exploratory manner, these empirical observations contribute to efforts to draw conclusions from large-scale and evolving software test suites.

cs.SE↗

Ranking Plausible Patches by Historic Feature Frequencies

Automated program repair (APR) techniques have achieved conspicuous progress, and are now capable of producing genuinely correct fixes in scenarios that were well beyond their capabilities only a few years ago. Nevertheless, even when an APR technique can find a correct fix for a bug, it still runs the risk of ranking the fix lower than other patches that are plausible (they pass all available tests) but incorrect. This can seriously hurt the technique's practical effectiveness, as the user will have to peruse a larger number of patches before finding the correct one. This paper presents PrevaRank, a technique that ranks plausible patches produced by any APR technique according to their feature similarity with historic programmer-written fixes for similar bugs. PrevaRank implements simple heuristics, which help make it scalable and applicable to any APR tool that produces plausible patches. In our experimental evaluation, after training PrevaRank on the fix history of 81 open-source Java projects, we used it to rank patches produced by 8 Java APR tools on 168 Defects4J bugs. PrevaRank consistently improved the ranking of correct fixes: for example, it ranked a correct fix within the top-3 positions in 27% more cases than the original tools did. Other experimental results indicate that PrevaRank works robustly with a variety of APR tools and bugs, with negligible overhead.

cs.SE↗

HBAT 2: A Python Package to Analyse Hydrogen Bonds and Other Non-covalent Interactions in Macromolecular Structures

Hydrogen bonds and other non-covalent interactions play a crucial role in maintaining the structural integrity and functionality of biological macromolecules such as proteins and nucleic acids. Accurate identification and analysis of these interactions are essential for understanding molecular recognition, protein folding, and drug design. HBAT (Hydrogen Bond Analysis Tool) is software for analysing hydrogen bonds and other weak interactions in macromolecular structures. This paper presents HBAT 2, an updated Python reimplementation of the original HBAT tool published in 2007. HBAT 2 is a Python package for automated analysis of hydrogen bonds and other non-covalent interactions in macromolecular structures available in Protein Data Bank (PDB) file format. The software identifies and analyses traditional hydrogen bonds, weak hydrogen bonds, halogen bonds, X-H$\cdots$$π$, $π$-$π$ stacking, and n$\rightarrow$$π$* interactions using geometric criteria. It also detects cooperativity and anticooperativity chains and renders them as 2D visualisations. The latest version offers improved cross-platform tkinter-based graphical user interface (GUI), a web-based interface, a simple command-line interface (CLI), and a developer-friendly API, making it accessible to users with different computational backgrounds.

physics.chem-ph↗

BioME: A Resource-Efficient Bioacoustic Foundational Model for IoT Applications

Passive acoustic monitoring has become a key strategy in biodiversity assessment, conservation, and behavioral ecology, especially as Internet-of-Things (IoT) devices enable continuous in situ audio collection at scale. While recent self-supervised learning (SSL)-based audio encoders, such as BEATs and AVES, have shown strong performance in bioacoustic tasks, their computational cost and limited robustness to unseen environments hinder deployment on resource-constrained platforms. In this work, we introduce BioME, a resource-efficient audio encoder designed for bioacoustic applications. BioME is trained via layer-to-layer distillation from a high-capacity teacher model, enabling strong representational transfer while reducing the parameter count by 75%. To further improve ecological generalization, the model is pretrained on multi-domain data spanning speech, environmental sounds, and animal vocalizations. A key contribution is the integration of modulation-aware acoustic features via FiLM conditioning, injecting a DSP-inspired inductive bias that enhances feature disentanglement in low-capacity regimes. Across multiple bioacoustic tasks, BioME matches or surpasses the performance of larger models, including its teacher, while being suitable for resource-constrained IoT deployments. For reproducibility, code and pretrained checkpoints are publicly available.

eess.AS↗

Towards Analyzing N-language Polyglot Programs

Polyglot programming is gaining popularity as developers integrate multiple programming languages to harness their individual strengths. With the recent popularity of platforms like GraalVM and other multi-language runtimes, creating and managing these systems has become much more feasible. However, current research on analyzing multilingual programs mainly focuses on two languages, leaving out the increasing complexity of systems that use three or more. For example, modern web systems often link JavaScript, WebAssembly, and Rust within the same execution chain. This paper envisions the landscape of software systems with three-language polyglot communication. We identify fundamental challenges in analyzing them and propose a conceptual roadmap to advance static analysis techniques to address them. Our vision aims to stimulate discussion and inspire new research directions toward scalable, language-agnostic analysis frameworks for next-generation polyglot systems.

cs.SE↗

Generalized relativistic second order magnetohydrodynamics: A correlation function approach using Zubarev's nonequilibrium statistical operator

We use total energy-momentum conservation and the Bianchi identity (magnetic-flux conservation) to construct second-order relativistic magnetohydrodynamics in a Zubarev's non-equilibrium statistical operator (NESO) framework. We obtain all dissipative tensors in the medium by focusing on a relativistic magnetized plasma that preserves parity and is symmetric to charge-conjugation. We also provide Kubo formulas for all transport coefficients that arise at second order. Moreover, we extend the NESO formalism to systematically take into account for nonlocal contributions.

physics.flu-dyn↗

$T_i/T_e$ Dependence of Core Turbulence and Transport in DIII-D QH-Mode Plasmas

This study investigates the effect of the ion-to-electron temperature ratio ($T_i/T_e$) on microturbulence driven transport in Quiescent H-mode (QH-mode) plasmas in the DIII-D tokamak. Utilizing the Gyrokinetic Toroidal Code (GTC) and the QH-mode equilibrium, we perform linear and nonlinear simulations to analyze transport properties and instability dynamics under variations of $T_i$ and $T_e$. Our results demonstrate that decreasing $T_i/T_e$ leads to a relative destabilization of trapped electron modes (TEM) over ion temperature gradient (ITG) modes, with the transition between these regimes dictated by $T_i/T_e$. When the electron temperature is increased at fixed ion temperature, we observe an increase in transport saturation levels. In contrast, decreasing the ion temperature at fixed electron temperature results in more modest transport enhancement. The radial correlation length, which characterizes eddy size, increases with rising $T_e$ and decreases with falling $T_i$, consistent with the observed trends in turbulent transport. Additionally, we examine the impact of impurity addition on turbulence and growth rates, finding that impurity presence does not significantly alter transport quantities compared to the impurity-free case. Finally, investigating helium as an alternative main ion species, we find that helium plasmas exhibit higher linear growth rates but result in lower transport saturation levels than deuterium plasmas, suggesting potential confinement benefits. These findings provide quantitative insights into the temperature ratio dependence in QH-mode plasmas and highlight the role of temperature profiles and zonal flows in influencing plasma confinement.

physics.plasm-ph↗

Towards Systematic Specification and Verification of Fairness Requirements: A Position Paper

Decisions suggested by improperly designed software systems might be prone to discriminate against people based on protected characteristics, such as gender and ethnicity. Previous studies attribute such undesired behavior to flaws in algorithmic design or biased data. However, these studies ignore that discrimination is often the result of a lack of well-specified fairness requirements and their verification. The fact that experts' knowledge about fairness is often implicit makes the task of specifying precise and verifiable fairness requirements difficult. In related domains, such as security engineering, knowledge graphs have been proven to be effective in formalizing knowledge to assist requirements specification and verification. To address the lack of formal mechanisms for specifying and verifying fairness requirements, we propose the development of a knowledge graph-based framework for fairness. In this paper, we discuss the challenges, research questions, and a road map towards addressing the research questions.

cs.SE↗

Vulnerability Patching Across Software Products and Software Components: A Case Study of Red Hat's Product Portfolio

Motivated by software maintenance and the more recent concept of security debt, the paper presents a time series analysis of vulnerability patching of Red Hat's products and components between 1999 and 2024. According to the results based on segmented regression analysis, the amounts of vulnerable products and components have not been stable; a linear trend describes many of the series well. Nor do the amounts align well with trends characterizing vulnerabilities in general. There are also visible breakpoints indicating that the linear trend is not universally applicable and that the growing security debt may be stabilizing.

cs.SE↗

Quantum-Assisted Machine Learning Models for Enhanced Weather Prediction

Quantum Machine Learning (QML) presents as a revolutionary approach to weather forecasting by using quantum computing to improve predictive modeling capabilities. In this study, we apply QML models, including Quantum Gated Recurrent Units (QGRUs), Quantum Neural Networks (QNNs), Quantum Long Short-Term Memory(QLSTM), Variational Quantum Circuits(VQCs), and Quantum Support Vector Machines(QSVMs), to analyze meteorological time-series data from the ERA5 dataset. Our methodology includes preprocessing meteorological features, implementing QML architectures for both classification and regression tasks. The results demonstrate that QML models can achieve reasonable accuracy in both prediction and classification tasks, particularly in binary classification. However, challenges such as quantum hardware limitations and noise affect scalability and generalization. This research provides insights into the feasibility of QML for weather prediction, paving the way for further exploration of hybrid quantum-classical frameworks to enhance meteorological forecasting.

quant-ph↗

Second-order spin hydrodynamics from Zubarev's nonequilibrium statistical operator formalism

Using the Zubarev's nonequilibrium statistical operator formalism, we derive the second-order expression for the dissipative tensors in relativistic spin hydrodynamics, {\em viz.} rotational stress tensor ($τ_{μν}$), boost heat vector ($q_μ$), shear stress tensor ($π_{μν}$), and bulk viscous pressure ($Π$). The first two ($τ_{μν}$ and $q_μ$) emerge due to the inclusion of the antisymmetric part in the energy-momentum tensor, which, in turn, governs the conservation of spin angular momentum ($Σ^{αμν}$). As a result, new thermodynamic forces, generated due to the antisymmetric part of $T_{μν}$, contain the spin chemical potential. In this work, we have also taken the spin density ($S^{μν}$) as an independent thermodynamic variable, in addition to the energy density and particle density, thereby resulting in two novel transport coefficients given by the correlation between spin density tensor and rotational stress tensor and vice versa. Additionally, the newly found terms in $π_{μν}$ and $Π$ are the artifacts of the new thermodynamic forces that arise due to the antisymmetric part of $T^{μν}$. Finally, we have derived the evolution equations for the aforesaid tensors: $τ_{μν}$, $q_μ$, $π_{μν}$, and $Π$.

hep-th↗

Zonal flow suppression of turbulent transport in the optimized stellarators W7-X and QSTK

We present a comparative study of transport in two optimized stellarator configurations: Wendelstein 7-X (W7-X) and a recent design called Quasi-Symmetric Turbulence Konzept (QSTK). Using global Gyrokinetic Toroidal Code (GTC), we explore the role of zonal flows (ZFs) in suppressing electrostatic Ion Temperature Gradient (ITG) driven turbulence in both configurations. The simulations reveal that ZFs significantly reduce ion heat transport in both W7-X and QSTK, with a lower value of heat flux on the latter configuration, as suggested by the apparently higher linear threshold (''critical'') gradients for ITG modes. The study also highlights that both stellarators exhibit similar mode structures. The results support the notion that linear stability measures, in combination with nonlinear stabilization by zonal flows, can play an important role in the suppression of nonlinear heat fluxes.

physics.plasm-ph↗

Taming the Tail: Leveraging Asymmetric Loss and Pade Approximation to Overcome Medical Image Long-Tailed Class Imbalance

Long-tailed problems in healthcare emerge from data imbalance due to variability in the prevalence and representation of different medical conditions, warranting the requirement of precise and dependable classification methods. Traditional loss functions such as cross-entropy and binary cross-entropy are often inadequate due to their inability to address the imbalances between the classes with high representation and the classes with low representation found in medical image datasets. We introduce a novel polynomial loss function based on Pade approximation, designed specifically to overcome the challenges associated with long-tailed classification. This approach incorporates asymmetric sampling techniques to better classify under-represented classes. We conducted extensive evaluations on three publicly available medical datasets and a proprietary medical dataset. Our implementation of the proposed loss function is open-sourced in the public repository:https://github.com/ipankhi/ALPA.

cs.CV↗

Challenges of Multilingual Program Specification and Analysis

Multilingual programs, whose implementations are made of different languages, are gaining traction especially in domains, such as web programming, that particularly benefit from the additional flexibility brought by using multiple languages. In this paper, we discuss the impact that the features commonly used in multilingual programming have on our capability of specifying and analyzing them. To this end, we first outline a few broad categories of multilingual programming, according to the mechanisms that are used for inter-language communication. Based on these categories, we describe several instances of multilingual programs, as well as the intricacies that formally reasoning about their behavior would entail. We also summarize the state of the art in multilingual program analysis, including the challenges that remain open. These contributions can help understand the lay of the land in multilingual program specification and analysis, and motivate further work in this area.

cs.PL↗

Unifying Pointer Analyses for Polyglot Inter-operations through Summary Specialization

Modular analysis of polyglot applications is challenging because heap object flows across language boundaries must be resolved. The state-of-the-art analyses for polyglot applications have two fundamental limitations. First, they assume explicit boundaries between the host and the guest language to determine inter-language dataflows. Second, they rely on specific analyses of the host and guest languages. The former assumption is impractical concerning recent advancements in polyglot programming techniques, while the latter disregards advances in pointer analysis of the underlying languages. In this work, we propose to extend existing pointer analyses with a novel summary specialization technique so that points-to set across language boundaries can be unified. Our novel technique leverages various combinations of host and guest analyses with minor modifications. We demonstrate the efficacy and generalizability of our approach by evaluating it with two polyglot language models: Java-C communication via Android's NDK and Java-Python communication in GraalVM.

cs.SE↗

Our fingerprints don't fade from the Apps we touch: Fingerprinting the Android WebView

Numerous studies demonstrated that browser fingerprinting is detrimental to users' security and privacy. However, little is known about the effects of browser fingerprinting on Android hybrid apps -- where a stripped-down Chromium browser is integrated into an app. These apps expand the attack surface by employing two-way communication between native apps and the web. This paper studies the impact of browser fingerprinting on these embedded browsers. To this end, we instrument the Android framework to record and extract information leveraged for fingerprinting. We study over 20,000 apps, including the most popular apps from the Google play store. We exemplify security flaws and severe information leaks in popular apps like Instagram. Our study reveals that fingerprints in hybrid apps potentially contain account-specific and device-specific information that identifies users across multiple devices uniquely. Besides, our results show that the hybrid app browser does not always adhere to standard browser-specific privacy policies.

cs.CR↗