arXiv ScienceSearch

arXiv subjects

David Sullivan

Publications and source records attributed to David Sullivan.

7 recordsLinked to original sources

Defense against Prompt Injection Attacks via Mixture of Encodings

Large Language Models (LLMs) have emerged as a dominant approach for a wide range of NLP tasks, with their access to external information further enhancing their capabilities. However, this introduces new vulnerabilities, known as prompt injection attacks, where external content embeds malicious instructions that manipulate the LLM's output. Recently, the Base64 defense has been recognized as one of the most effective methods for reducing success rate of prompt injection attacks. Despite its efficacy, this method can degrade LLM performance on certain NLP tasks. To address this challenge, we propose a novel defense mechanism: mixture of encodings, which utilizes multiple character encodings, including Base64. Extensive experimental results show that our method achieves one of the lowest attack success rates under prompt injection attacks, while maintaining high performance across all NLP tasks, outperforming existing character encoding-based defense methods. This underscores the effectiveness of our mixture of encodings strategy for both safety and task performance metrics.

cs.CL

A Framework for Automated Measurement of Responsible AI Harms in Generative AI Applications

We present a framework for the automated measurement of responsible AI (RAI) metrics for large language models (LLMs) and associated products and services. Our framework for automatically measuring harms from LLMs builds on existing technical and sociotechnical expertise and leverages the capabilities of state-of-the-art LLMs, such as GPT-4. We use this framework to run through several case studies investigating how different LLMs may violate a range of RAI-related principles. The framework may be employed alongside domain-specific sociotechnical expertise to create measurements for new harm areas in the future. By implementing this framework, we aim to enable more advanced harm measurement efforts and further the responsible use of LLMs.

cs.CL

Relative baryon-dark matter velocities in cosmological zoom simulations

Supersonic relative motion between baryons and dark matter due to the decoupling of baryons from the primordial plasma after recombination affects the growth of the first small-scale structures. Large box sizes (greater than a few hundred Mpc) are required to sample the full range of scales pertinent to the relative velocity, while the effect of the relative velocity is strongest on small scales (less than a few hundred kpc). This separation of scales naturally lends itself to the use of `zoom' simulations, and here we present our methodology to self-consistently incorporate the relative velocity in zoom simulations, including its cumulative effect from recombination through to the start time of the simulation. We apply our methodology to a large-scale cosmological zoom simulation, finding that the inclusion of relative velocities suppresses the halo baryon fraction by $46$--$23$ per cent between $z=13.6$ and $11.2$, in qualitative agreement with previous works. In addition, we find that including the relative velocity delays the formation of star particles by $\sim 20 {~\rm Myr}$ Myr on average (of the order of the lifetime of a $\sim 9~{\rm M}_\odot$ Population III star) and suppresses the final stellar mass by as much as $79$ per cent at $z=11.2$.

astro-ph.CO

Semantic Classification of Tabular Datasets via Character-Level Convolutional Neural Networks

A character-level convolutional neural network (CNN) motivated by applications in "automated machine learning" (AutoML) is proposed to semantically classify columns in tabular data. Simulated data containing a set of base classes is first used to learn an initial set of weights. Hand-labeled data from the CKAN repository is then used in a transfer-learning paradigm to adapt the initial weights to a more sophisticated representation of the problem (e.g., including more classes). In doing so, realistic data imperfections are learned and the set of classes handled can be expanded from the base set with reduced labeled data and computing power requirements. Results show the effectiveness and flexibility of this approach in three diverse domains: semantic classification of tabular data, age prediction from social media posts, and email spam classification. In addition to providing further evidence of the effectiveness of transfer learning in natural language processing (NLP), our experiments suggest that analyzing the semantic structure of language at the character level without additional metadata---i.e., network structure, headers, etc.---can produce competitive accuracy for type classification, spam classification, and social media age prediction. We present our open-source toolkit SIMON, an acronym for Semantic Inference for the Modeling of ONtologies, which implements this approach in a user-friendly and scalable/parallelizable fashion.

cs.CL

Abstractive Tabular Dataset Summarization via Knowledge Base Semantic Embeddings

This paper describes an abstractive summarization method for tabular data which employs a knowledge base semantic embedding to generate the summary. Assuming the dataset contains descriptive text in headers, columns and/or some augmenting metadata, the system employs the embedding to recommend a subject/type for each text segment. Recommendations are aggregated into a small collection of super types considered to be descriptive of the dataset by exploiting the hierarchy of types in a pre-specified ontology. Using February 2015 Wikipedia as the knowledge base, and a corresponding DBpedia ontology as types, we present experimental results on open data taken from several sources--OpenML, CKAN and data.world--to illustrate the effectiveness of the approach.

cs.AI

Using Artificial Neural Networks to Constrain the Halo Baryon Fraction during Reionization

Radiative feedback from stars and galaxies has been proposed as a potential solution to many of the tensions with simplistic galaxy formation models based on $\Lambda$CDM, such as the faint end of the UV luminosity function. The total energy budget of radiation could exceed that of galactic winds and supernovae combined, which has driven the development of sophisticated algorithms that evolve both the radiation field and the hydrodynamical response of gas simultaneously, in a cosmological context. We probe self-feedback on galactic scales using the adaptive mesh refinement, radiative transfer, hydrodynamics, and $N$-body code. Unlike previous studies which assume a homogeneous UV background, we self-consistently evolve both the radiation field and gas to constrain the halo baryon fraction during cosmic reionization. We demonstrate that the characteristic halo mass with mean baryon fraction half the cosmic mean, $M_{\mathrm{c}}(z)$, shows very little variation as a function of mass-weighted ionization fraction. Furthermore, we find that the inclusion of metal cooling and the ability to resolve scales small enough for self-shielding to become efficient leads to a significant drop in $M_{\mathrm{c}}$ when compared to recent studies. Finally, we develop an Artificial Neural Network that is capable of predicting the baryon fraction of haloes based on recent tidal interactions, gas temperature, and mass-weighted ionization fraction. Such a model can be applied to any reionization history, and trivially incorporated into semi-analytical models of galaxy formation.

astro-ph.GA

Cosmic Dawn (CoDa): the First Radiation-Hydrodynamics Simulation of Reionization and Galaxy Formation in the Local Universe

Cosmic reionization by starlight from early galaxies affected their evolution, thereby impacting reionization, itself. Star formation suppression, for example, may explain the observed underabundance of Local Group dwarfs relative to N-body predictions for Cold Dark Matter. Reionization modelling requires simulating volumes large enough [~(100Mpc)^3] to sample reionization "patchiness", while resolving millions of galaxy sources above ~10^8 Msun , combining gravitational and gas dynamics with radiative transfer. Modelling the Local Group requires initial cosmological density fluctuations pre-selected to form the well-known structures of the local universe today. Cosmic Dawn ("CoDa") is the first such fully-coupled, radiation-hydrodynamics simulation of reionization of the local universe. Our new hybrid CPU-GPU code, RAMSES-CUDATON, performs hundreds of radiative transfer and ionization rate-solver timesteps on the GPUs for each hydro-gravity timestep on the CPUs. CoDa simulated (91Mpc)^3 with 4096^3 particles and cells, to redshift 4.23, on ORNL supercomputer Titan, utilizing 8192 cores and 8192 GPUs. Global reionization ended slightly later than observed. However, a simple temporal rescaling which brings the evolution of ionized fraction into agreement with observations also reconciles ionizing flux density, cosmic star formation history, CMB electron scattering optical depth and galaxy UV luminosity function with their observed values. Photoionization heating suppressed the star formation of haloes below ~2 x 10^9 Msun , decreasing the abun- dance of faint galaxies around MAB_1600 = [-10,-12]. For most of reionization, star formation was dominated by haloes between 10^10 - 10^11 Msun , so low-mass halo suppression was not reflected by a distinct feature in the global star formation history. (Abridged)

astro-ph.GA