arXiv ScienceSearch

arXiv subjects

Roman Levin

Publications and source records attributed to Roman Levin.

8 recordsLinked to original sources

Has My System Prompt Been Used? Large Language Model Prompt Membership Inference

Prompt engineering has emerged as a powerful technique for optimizing large language models (LLMs) for specific applications, enabling faster prototyping and improved performance, and giving rise to the interest of the community in protecting proprietary system prompts. In this work, we explore a novel perspective on prompt privacy through the lens of membership inference. We develop Prompt Detective, a statistical method to reliably determine whether a given system prompt was used by a third-party language model. Our approach relies on a statistical test comparing the distributions of two groups of model outputs corresponding to different system prompts. Through extensive experiments with a variety of language models, we demonstrate the effectiveness of Prompt Detective for prompt membership inference. Our work reveals that even minor changes in system prompts manifest in distinct response distributions, enabling us to verify prompt usage with statistical significance.

cs.AI

A Performance-Driven Benchmark for Feature Selection in Tabular Deep Learning

Academic tabular benchmarks often contain small sets of curated features. In contrast, data scientists typically collect as many features as possible into their datasets, and even engineer new features from existing ones. To prevent overfitting in subsequent downstream modeling, practitioners commonly use automated feature selection methods that identify a reduced subset of informative features. Existing benchmarks for tabular feature selection consider classical downstream models, toy synthetic datasets, or do not evaluate feature selectors on the basis of downstream performance. Motivated by the increasing popularity of tabular deep learning, we construct a challenging feature selection benchmark evaluated on downstream neural networks including transformers, using real datasets and multiple methods for generating extraneous features. We also propose an input-gradient-based analogue of Lasso for neural networks that outperforms classical feature selection methods on challenging problems such as selecting from corrupted or second-order features.

cs.LG

Transfer Learning with Deep Tabular Models

Recent work on deep learning for tabular data demonstrates the strong performance of deep tabular models, often bridging the gap between gradient boosted decision trees and neural networks. Accuracy aside, a major advantage of neural models is that they learn reusable features and are easily fine-tuned in new domains. This property is often exploited in computer vision and natural language applications, where transfer learning is indispensable when task-specific training data is scarce. In this work, we demonstrate that upstream data gives tabular neural networks a decisive advantage over widely used GBDT models. We propose a realistic medical diagnosis benchmark for tabular transfer learning, and we present a how-to guide for using upstream data to boost performance with a variety of tabular neural network architectures. Finally, we propose a pseudo-feature method for cases where the upstream and downstream feature sets differ, a tabular-specific problem widespread in real-world applications. Our code is available at https://github.com/LevinRoman/tabular-transfer-learning .

cs.LG

Where do Models go Wrong? Parameter-Space Saliency Maps for Explainability

Conventional saliency maps highlight input features to which neural network predictions are highly sensitive. We take a different approach to saliency, in which we identify and analyze the network parameters, rather than inputs, which are responsible for erroneous decisions. We find that samples which cause similar parameters to malfunction are semantically similar. We also show that pruning the most salient parameters for a wrongly classified sample often improves model behavior. Furthermore, fine-tuning a small number of the most salient parameters on a single sample results in error correction on other samples that are misclassified for similar reasons. Based on our parameter saliency method, we also introduce an input-space saliency technique that reveals how image features cause specific network components to malfunction. Further, we rigorously validate the meaningfulness of our saliency maps on both the dataset and case-study levels.

cs.CV

Echo Chambers in Collaborative Filtering Based Recommendation Systems

Recommendation systems underpin the serving of nearly all online content in the modern age. From Youtube and Netflix recommendations, to Facebook feeds and Google searches, these systems are designed to filter content to the predicted preferences of users. Recently, these systems have faced growing criticism with respect to their impact on content diversity, social polarization, and the health of public discourse. In this work we simulate the recommendations given by collaborative filtering algorithms on users in the MovieLens data set. We find that prolonged exposure to system-generated recommendations substantially decreases content diversity, moving individual users into "echo-chambers" characterized by a narrow range of content. Furthermore, our work suggests that once these echo-chambers have been established, it is difficult for an individual user to break out by manipulating solely their own rating vector.

cs.IR

A Proof of Principle: Multi-Modality Radiotherapy Optimization

Radiotherapy is used to treat cancer patients by damaging DNA of tumor cells using ionizing radiation. Photons are the most widely used radiation type for therapy, having been put into use soon after the first discovery of X-rays in 1895. However, there are emerging interests and developments of other radiation modalities such as protons and carbon ions, owing to their unique biological and physical characteristics that distinguish these modalities from photons. Current attempts to determine an optimal radiation modality or an optimal combination of multiple modalities are empirical and in the early stage of development. In this paper, we propose a mathematical framework to optimize full radiation dose distributions and fractionation schedules of multiple radiation modalities, aiming to maximize the damage to the tumor while limiting the damage to the normal tissue to the corresponding tolerance level. This formulation gives rise to a non-convex, mixed integer program and we propose a bilevel optimization algorithm, to efficiently solve it. The upper level problem is to optimize the fractionation schedule using the dose distribution optimized in the lower level. We demonstrate the feasibility of our novel framework and algorithms in a simple 2-dimensional phantom with two different radiation modalities, where clinical intuition can be easily drawn. The results of our numerical simulations agree with the clinical intuition, validating our approach and showing the promise of the framework for further clinical investigation.

math.OC

Study of plasma heating in ohmically and auxiliary heated regimes in spherical tokamak Globus-M

The ion temperature behavior in the plasma core of the spherical tokamak Globus-M (major radius 0.36 m, minor radius 0.24 m, torus aspect ratio 1.5, toroidal magnetic field near the plasma axis 0.4 T, plasma current up to 0.3 MA) was studied by means of 12-channels neutral particle analyzer (NPA) ACORD-12. The experiments were performed in ohmic regimes as well as in regimes with the ion cyclotron resonance heating (ICRH) in the vicinity of fundamental harmonic for hydrogen minority in deuterium bulk plasma and the neutral beam injection (NBI). The total auxiliary power exceeded the magnitude of 0.8-1 MW. The NPA provided the simultaneous measurements of deuterium and hydrogen energy spectra and the percentage of both isotopes. The ion temperature was studied in a wide range of the plasma current 0.08-0.3 MA and the plasma average density (1-7)x1019 m-3 at various values of the plasma vertical elongation and the triangularity. The experimental data were compared with the results of numerical simulation. We employed a simple 1D model describing the ion energy balance by using the neoclassical transport coefficients. The charge-exchange losses were also taken into account. The experimentally measured ion temperature dependence on the plasma current and plasma density differed from the Artsimovich scaling law predictions well describing the ion heating in the case of the neoclassical plateau regime in the conventional tokamak even at a relatively low temperature in ohmic plasma. In particular strong, almost linear plasma current temperature dependence was revealed. It indicates a dominating role of trapped particles in the ion energy balance at a low aspect ratio. The estimates of energy confinement time values derived from the magnetic measurements in ohmic and auxiliary heating regimes are also presented.

physics.plasm-ph