arXiv ScienceSearch

arXiv subjects

Preet Baxi

Publications and source records attributed to Preet Baxi.

6 recordsLinked to original sources

Automated Design of Inventory Policy with Large Language Models: An Exploratory Study

Firms making inventory decisions have access to operational data, optimization tools, and large language models (LLMs). Typically, data characterize the operating environment, optimization selects parameters within a prespecified inventory policy class, and LLMs support coding and decision analysis. We develop an integrated framework that combines these resources to automate inventory policy design. Given demand data, the framework iteratively uses an LLM to generate parameterized policy classes and an external solver to optimize its parameters within each class. Across 30 lost-sales inventory instances, the mean cost reduction relative to optimized base-stock benchmarks increases from 17.5% after one generation to 30.0% after ten generations. Parameter optimization is central to this performance: an LLM-only variant performs substantially worse, whereas optimization-guided feedback improves policy quality, accelerates search, and directs the LLM toward better policy classes rather than merely better parameter values within a fixed class. The strongest discovered policies are also interpretable: they combine recognizable inventory-control motifs, including capped orders, discounted or weighted pipeline inventory, and threshold-based replenishment logic. The search thereby produces new policy-class functional forms that, to our knowledge, have not previously been studied in the lost-sales inventory literature. These functional forms are not specified ex ante but emerge from the search process. Moreover, after their parameters are re-optimized, three discovered policy classes achieve average cost reductions of 21.75% to 22.60% across 10,064 new inventory instances. Overall, the results show that data-driven parameter optimization can guide LLM-based search over a broad space of inventory policy classes and identify high-performing, interpretable, and transferable decision rules.

cs.AI

Prompt Injection in Automated R\'esum\'e Screening with Large Language Models: Single and Multi-Injection Settings

Large language models (LLMs) are increasingly used to screen and rank job applicants, creating incentives for candidates to strategically manipulate algorithmic hiring systems. We study prompt injection in automated r\'esum\'e screening, defined as subtle self-promotional text that introduces no new qualifications but is designed to influence LLM evaluations. Using controlled experiments, we show that prompt injection reliably improves applicant rankings when r\'esum\'e quality is homogeneous and few candidates inject. However, its effectiveness rapidly diminishes as more candidates inject, collapsing when manipulation becomes widespread. When candidate quality is heterogeneous, prompt injection is less effective on average, but can occasionally allow lower-quality candidates to outrank higher-quality ones, raising fairness concerns. Overall, LLM-based screening is most vulnerable when manipulation is rare and candidate quality differences are small. Code and resources are publicly available at: https://github.com/preetb1199/Prompt_Injection_ACL26

cs.AI

Asymptotically Optimal Sequential Testing with Heterogeneous LLMs

We study a Bayesian binary sequential hypothesis testing problem with multiple large language models (LLMs). Each LLM $j$ has per-query cost $c_j>0$, random waiting time with mean $\mu_j>0$ and sub-Gaussian tails, and \emph{asymmetric} accuracies: the probability of returning the correct label depends on the true hypothesis $\theta\in\{A,B\}$ and needs not be the same under $A$ and $B$. This asymmetry induces two distinct information rates $(I_{j,A}, I_{j,B})$ per LLM, one under each hypothesis. The decision-maker chooses LLMs sequentially, observes their noisy binary answers, and stops when the posterior probability of one hypothesis exceeds $1-\alpha$. The objective is to minimize the sum of expected query cost and expected waiting cost, $\mathbb{E}[C_\pi] + \mathbb{E}[g(W_\pi)]$, where $C_\pi$ is the total query cost, $W_\pi$ is the total waiting time and $g$ is a polynomial function (e.g., $g(x)=x^\rho$ with $\rho\ge 1$). We prove that as the error tolerance $\alpha\to0$, the optimal policy is asymptotically equivalent to one that uses at most two LLMs. In this case, a single-LLM policy is \emph{not} generically optimal: optimality now requires exploiting a two-dimensional tradeoff between information under $A$ and information under $B$. Any admissible policy induces an expected information-allocation vector in $\mathbb{R}_+^2$, and we show that the optimal allocation lies at an extreme point of the associated convex set when $\alpha$ is relatively small, and hence uses at most two LLMs. We construct belief-dependent policies that first mix between two LLMs when the posterior is ambiguous, and then switch to a single "specialist" LLM when the posterior is sufficiently close to one of the hypotheses. These policies match the universal lower bound up to a $(1+o(1))$ factor as $\alpha\rightarrow 0$.

cs.DS

Search for continuous gravitational waves from neutron stars in five globular clusters in the first part of the fourth LIGO-Virgo-KAGRA observing run

We present the results of directed searches for continuous gravitational waves from unknown neutron stars in five Milky Way globular clusters. We carry out these searches in the LIGO data from the first eight months of the fourth LIGO-Virgo-KAGRA observing run using the WEAVE semicoherent program, which sums matched-filter detection-statistic values over many time segments spanning the observation period. No gravitational wave signal is detected in the search band of 20-475 Hz. Injections of simulated continuous wave signals in the data indicate that we achieve the most sensitive results to date across most of the explored parameter space volume, obtaining median 95% confidence level upper limits as low as $\sim 4.2 \times 10^{-26}$ near 282 Hz for NGC 6397. We also derive upper limits on neutron star ellipticity and $r$-mode amplitudes, reaching $\lesssim 10^{-5}$ and $\lesssim 10^{-3}$, respectively, at frequencies above 200 Hz.

gr-qc

Monitoring of Continuous-Wave Hardware Injections in LIGO Interferometers during the O4 Observing Run

Although there have now been hundreds of transient gravitational-wave detections of merging compact stars by the LIGO-Virgo-KAGRA (LVK) detector network, no continuous-wave (CW) signals have yet been discovered. To ensure that such signals, expected to be exceedingly weak, can be detected in the ongoing O4 observing run by coherent integration over years, simulated waveforms ('hardware injections') are injected directly into the LIGO data by continuously modulating the positions of the interferometer mirrors so as to mimic nearly sinusoidal signals from fast-spinning galactic neutron stars. A set of 18 such simulated CW sources are injected with signal frequencies spanning much of the LIGO detection band and with varying sky locations. By verifying the successful recovery of the simulated signals, including preservation of absolute phase over as many as 10^{11} signal cycles, we validate our understanding of detector response and end-to-end search pipelines, including data cleaning. Daily and weekly monitoring of the signal reconstruction is meant to catch any unexpected sudden changes in interferometer response, to verify that signal-to-noise ratio increases as expected and to verify that simulated source parameters are recovered correctly. We describe three methods of monitoring: 1) a highly templated matched filter to extract signal amplitude and phase precisely; 2) a frequentist Fstatistic evaluation that marginalizes over amplitude, phase and orientation of the star; and 3) a Bayesian reconstruction of the source parameters together with noise characterization. Results from each method are shown, with emphasis on the new templated method, which yields precise measurement of the critical phase offset parameter and therefore validates understanding of absolute timing delays in the detector response and data stream.

astro-ph.IM

Detectability of gravitational higher order modes in the third-generation era

Detection of higher order modes of gravitational waves in third-generation (3G) ground-based detectors such as Cosmic Explorer and Einstein Telescope is explored. Using the astrophysical population of binary black holes based on events reported in the second gravitational wave catalog by Laser Interferometer Gravitational Wave Observatory (LIGO) and Virgo (GWTC-2), in conjunction with the Madau-Dickinson model for redshift evolution of the binary black hole mergers, we assess the detectability of these higher order modes using a network consisting of three third-generation detectors. We find that the two subleading modes [(3,3) and (4,4)] can be detected in approximately 30% of the population with a network signal-to-noise ratio of 3 or more, and for nearly 10% of the sources, the five leading modes will be detectable. Besides, a study concerning the effect of binary's mass ratio and its orbital inclination with the observer's line-of-sight in detecting various modes is presented. For a few selected events of the LIGO-Virgo catalog, we identify the modes that would have been detected if a third-generation detector was operational when these events were recorded. We also compute the detectability of higher modes by Voyager and find that only $\sim$ 6 and 2% of the detectable population will have an associated detection of (3,3) and (4,4) modes, respectively. Observing these higher order modes in the 3G era would have a huge impact on the science possible with these detectors ranging from astrophysics and cosmology to testing strong-field gravity.

gr-qc