arXiv ScienceSearch

arXiv subjects

David Johnston

Publications and source records attributed to David Johnston.

At least 19 recordsLinked to original sources

Bergson: An Open Source Library for Data Attribution

Data attribution is a promising field in interpretability that aims to explain model behavior through the influence of its training data, with applications including debugging undesirable model behavior and training dataset curation. However, significant engineering effort is required to perform it at scale, and many cutting edge techniques lack open-source tooling and support. Bergson is an open source library that aims to enable faster progress in the field by providing a host of techniques that scale to very large language models and pre-training datasets. The library natively supports on-disk gradient stores and multi-node distributed training, and provides quality of life tools for researchers. Finally, we introduce the first open-source implementations of three leading data attribution methods: MAGIC, SOURCE, and TrackStar. The library is available at https://github.com/EleutherAI/bergson .

cs.LG

Examining Two Hop Reasoning Through Information Content Scaling

Prior work has found that transformers have an inconsistent ability to learn to answer latent two-hop questions -- questions of the form "Who is Bob's mother's boss?" We study why this is the case by examining how transformers' capacity to learn datasets of two-hop questions and answers (two-hop QA) scales with their size, motivated by prior work on transformer knowledge capacity for simple factual memorization. We find that capacity scaling and generalization both support the hypothesis that latent two-hop QA requires transformers to learn each fact twice, while two-hop QA with chain of thought does not. We also show that with appropriate dataset parameters, it is possible to "trap" very small models in a regime where they memorize answers to two-hop questions independently, even though they would perform better if they could learn to answer them with function composition. Our findings show that measurement of capacity scaling can complement existing interpretability methods, though there are challenges in using it for this purpose.

cs.AI

Impacts of Extreme Heat on Labor Force Dynamics

We use daily longitudinal data and a within-worker identification approach to examine the impacts of heat on labor force dynamics in Australia. High temperatures during 2001-2019 significantly reduced work attendance and hours worked, which were not compensated for in subsequent days and weeks. The largest reductions occurred in cooler regions and recent years, and were not solely concentrated amongst outdoor-based workers. Financial and Insurance Services was the most strongly affected industry, with temperatures above 38{\deg}C (100{\deg}F) increasing absenteeism by 15 percent. Adverse heat effects during the work commute and during outdoor work hours are shown to be key mechanisms.

econ.GN

Heat and Worker Health

Extreme heat negatively impacts cognition, learning, and task performance. With increasing global temperatures, workers may therefore be at increased risk of work-related injuries and illness. This study estimates the effects of temperature on worker health using records spanning 1985-2020 from an Australian mandatory insurance scheme. High temperatures are found to cause significantly more claims, particularly among manual workers in outdoor-based industries. These adverse effects have not diminished across time, with the largest effect observed for the 2015-2020 period, indicating increasing vulnerability to heat. Within occupations, the workers most adversely affected by heat are female, older-aged and higher-earning. Finally, results from firm-level panel analyses show that the percentage increase in claims on hot days is largest at "safer" firms.

econ.GN

Flood Disasters and Health Among the Urban Poor

Billions of people live in urban poverty, with many forced to reside in disaster-prone areas. Research suggests that such disasters harm child nutrition and increase adult morbidity. However, little is known about impacts on mental health, particularly of people living in slums. In this paper we estimate the effects of flood disasters on the mental and physical health of poor adults and children in urban Indonesia. Our data come from the Indonesia Family Life Survey and new surveys of informal settlement residents. We find that urban poor populations experience increases in acute morbidities and depressive symptoms following floods, that the negative mental health effects last longer, and that the urban wealthy show no health effects from flood exposure. Further analysis suggests that worse economic outcomes may be partly responsible. Overall, the results provide a more nuanced understanding of the morbidities experienced by populations most vulnerable to increased disaster occurrence.

econ.GN

Heat and Economic Preferences

The empirical evidence suggests that key accumulation decisions and risky choices associated with economic development depend, at least in part, on economic preferences such as willingness to take risk and patience. This paper studies whether temperature could be one of the potential channels that influences such economic preferences. Using data from the Indonesia Family Life Survey and NASAs Modern Era Retrospective Analysis for Research and Applications data we exploit quasi exogenous variations in outdoor temperatures caused by the random allocation of survey dates. This approach allows us to estimate the effects of temperature on elicited measures of risk aversion, rational choice violations, and impatience. We then explore three possible mechanisms behind this relationship, cognition, sleep, and mood. Our findings show that higher temperatures lead to significantly increased rational choice violations and impatience, but do not significantly increase risk aversion. These effects are mainly driven by night time temperatures on the day prior to the survey and less so by temperatures on the day of the survey. This impact is quasi linear and increasing when midnight outdoor temperatures are above 22C. The evidence shows that night time temperatures significantly deplete cognitive functioning, mathematical skills in particular. Based on these findings we posit that heat induced night time disturbances cause stress on critical parts of the brain, which then manifest in significantly lower cognitive functions that are critical for individuals to perform economically rational decision making.

econ.GN

The SDSS Coadd: A Galaxy Photometric Redshift Catalog

We present and describe a catalog of galaxy photometric redshifts (photo-z's) for the Sloan Digital Sky Survey (SDSS) Coadd Data. We use the Artificial Neural Network (ANN) technique to calculate photo-z's and the Nearest Neighbor Error (NNE) method to estimate photo-z errors for $\sim$ 13 million objects classified as galaxies in the coadd with $r < 24.5$. The photo-z and photo-z error estimators are trained and validated on a sample of $\sim 83,000$ galaxies that have SDSS photometry and spectroscopic redshifts measured by the SDSS Data Release 7 (DR7), the Canadian Network for Observational Cosmology Field Galaxy Survey (CNOC2), the Deep Extragalactic Evolutionary Probe Data Release 3(DEEP2 DR3), the VIsible imaging Multi-Object Spectrograph - Very Large Telescope Deep Survey (VVDS) and the WiggleZ Dark Energy Survey. For the best ANN methods we have tried, we find that 68% of the galaxies in the validation set have a photo-z error smaller than $\sigma_{68} =0.031$. After presenting our results and quality tests, we provide a short guide for users accessing the public data.

astro-ph.CO

The SDSS Coadd: Cross-Correlation Weak Lensing and Tomography of Galaxy Clusters

The shapes of distant galaxies are sheared by intervening galaxy clusters. We examine this effect in Stripe 82, a 275 square degree region observed multiple times in the Sloan Digital Sky Survey and coadded to achieve greater depth. We obtain a mass-richness calibration that is similar to other SDSS analyses, demonstrating that the coaddition process did not adversely affect the lensing signal. We also propose a new parameterization of the effect of tomography on the cluster lensing signal which does not require binning in redshift, and we show that using this parameterization we can detect tomography for stacked clusters at varying redshifts. Finally, due to the sensitivity of the tomographic detection to accurately marginalizing over the effect of the cluster mass, we show that tomography at low redshift (where dependence on exact cosmological models is weak) can be used to constrain mass profiles in clusters.

astro-ph.CO

The SDSS Coadd: Cosmic Shear Measurement

Stripe 82 in the Sloan Digital Sky Survey was observed multiple times, allowing deeper images to be constructed by coadding the data. Here we analyze the ellipticities of background galaxies in this 275 square degree region, searching for evidence of distortions due to cosmic shear. The E-mode is detected in both real and Fourier space with $>5$-$\sigma$ significance on degree scales, while the B-mode is consistent with zero as expected. The amplitude of the signal constrains the combination of the matter density $\Omega_m$ and fluctuation amplitude $\sigma_8$ to be $\Omega_m^{0.7}\sigma_8 = 0.252^{+0.032}_{-0.052}$.

astro-ph.CO

Lossy compression of weak lensing data

Future orbiting observatories will survey large areas of sky in order to constrain the physics of dark matter and dark energy using weak gravitational lensing and other methods. Lossy compression of the resultant data will improve the cost and feasibility of transmitting the images through the space communication network. We evaluate the consequences of the lossy compression algorithm of Bernstein et al. (2010) for the high-precision measurement of weak-lensing galaxy ellipticities. This square-root algorithm compresses each pixel independently, and the information discarded is by construction less than the Poisson error from photon shot noise. For simulated space-based images (without cosmic rays) digitized to the typical 16 bits per pixel, application of the lossy compression followed by image-wise lossless compression yields images with only 2.4 bits per pixel, a factor of 6.7 compression. We demonstrate that this compression introduces no bias in the sky background. The compression introduces a small amount of additional digitization noise to the images, and we demonstrate a corresponding small increase in ellipticity measurement noise. The ellipticity measurement method is biased by the addition of noise, so the additional digitization noise is expected to induce a multiplicative bias on the galaxies' measured ellipticities. After correcting for this known noise-induced bias, we find a residual multiplicative ellipticity bias of m ~ -4x10^{-4}. This bias is small when compared to the many other issues that precision weak lensing surveys must confront, and furthermore we expect it to be reduced further with better calibration of ellipticity measurement methods.

astro-ph.IM

An Improved Cluster Richness Estimator

Minimizing the scatter between cluster mass and accessible observables is an important goal for cluster cosmology. In this work, we introduce a new matched filter richness estimator, and test its performance using the maxBCG cluster catalog. Our new estimator significantly reduces the variance in the L_X-richness relation, from σ_{\ln L_X}^2=(0.86\pm0.02)^2 to σ_{\ln L_X}^2=(0.69\pm0.02)^2. Relative to the maxBCG richness estimate, it also removes the strong redshift dependence of the richness scaling relations, and is significantly more robust to photometric and redshift errors. These improvements are largely due to our more sophisticated treatment of galaxy color data. We also demonstrate the scatter in the L_X-richness relation depends on the aperture used to estimate cluster richness, and introduce a novel approach for optimizing said aperture which can be easily generalized to other mass tracers.

astro-ph

Constraining the Scatter in the Mass-Richness Relation of maxBCG Clusters With Weak Lensing and X-ray Data

We measure the logarithmic scatter in mass at fixed richness for clusters in the maxBCG cluster catalog, an optically selected cluster sample drawn from SDSS imaging data. Our measurement is achieved by demanding consistency between available weak lensing and X-ray measurements of the maxBCG clusters, and the X-ray luminosity--mass relation inferred from the 400d X-ray cluster survey, a flux limited X-ray cluster survey. We find σ_{\ln M|N_{200}}=0.45^{+0.20}_{-0.18} (95% CL) at N_{200} ~ 40, where N_{200} is the number of red sequence galaxies in a cluster. As a byproduct of our analysis, we also obtain a constraint on the correlation coefficient between \ln Lx and \ln M at fixed richness, which is best expressed as a lower limit, r_{L,M|N} >= 0.85 (95% CL). This is the first observational constraint placed on a correlation coefficient involving two different cluster mass tracers. We use our results to produce a state of the art estimate of the halo mass function at z=0.23 -- the median redshift of the maxBCG cluster sample -- and find that it is consistent with the WMAP5 cosmology. Both the mass function data and its covariance matrix are presented.

astro-ph

Cosmological Constraints from SDSS maxBCG Cluster Abundances

We perform a maximum likelihood analysis of the cluster abundance measured in the SDSS using the maxBCG cluster finding algorithm. Our analysis is aimed at constraining the power spectrum normalization $σ_8$, and assumes flat cosmologies with a scale invariant spectrum, massless neutrinos, and CMB and supernova priors Omega_m*h^2=0.128+/-0.01 and h=0.72+/-0.05 respectively. Following the method described in the companion paper Rozo et al. 2007, we derive σ_8=0.92+/-0.10$ (1-sigma) after marginalizing over all major systematic uncertainties. We place strong lower limits on the normalization, sigma_8>0.76 (95% CL) (>0.68 at 99% CL). We also find that our analysis favors relatively low values for the slope of the Halo Occupation Distribution (HOD), alpha=0.83+/-0.06. The uncertainties of these determinations will substantially improve upon completion of an ongoing campaign to estimate dynamical, weak lensing, and X-ray cluster masses in the SDSS maxBCG cluster sample.

astro-ph

COSMOS: 3D weak lensing and the growth of structure

We present a three dimensional cosmic shear analysis of the Hubble Space Telescope COSMOS survey, the largest ever optical imaging program performed in space. We have measured the shapes of galaxies for the tell-tale distortions caused by weak gravitational lensing, and traced the growth of that signal as a function of redshift. Using both 2D and 3D analyses, we measure cosmological parameters Omega_m, the density of matter in the universe, and sigma_8, the normalization of the matter power spectrum. The introduction of redshift information tightens the constraints by a factor of three, and also reduces the relative sampling (or "cosmic") variance compared to recent surveys that may be larger but are only two dimensional. From the 3D analysis, we find sigma_8*(Omega_m/0.3)^0.44=0.866+^0.085_-0.068 at 68% confidence limits, including both statistical and potential systematic sources of error in the total budget. Indeed, the absolute calibration of shear measurement methods is now the dominant source of uncertainty. Assuming instead a baseline cosmology to fix the geometry of the universe, we have measured the growth of structure on both linear and non-linear physical scales. Our results thus demonstrate a proof of concept for tomographic analysis techniques that have been proposed for future weak lensing surveys by a dedicated wide-field telescope in space.

astro-ph

MaxBCG: A Red Sequence Galaxy Cluster Finder

Measurements of galaxy cluster abundances, clustering properties, and mass to- light ratios in current and future surveys can provide important cosmological constraints. Digital wide-field imaging surveys, the recently-demonstrated fidelity of red-sequence cluster detection techniques, and a new generation of realistic mock galaxy surveys provide the means for construction of large, cosmologicallyinteresting cluster samples, whose selection and properties can be understood in unprecedented depth. We present the details of the "maxBCG" algorithm, a cluster-detection technique tailored to multi-band CCD-imaging data. MaxBCG primarily relies on an observational cornerstone of massive galaxy clusters: they are marked by an overdensity of bright, uniformly red galaxies. This detection scheme also exploits classical brightest cluster galaxies (BCGs), which are often found at the center of these same massive clusters. (ABRIDGED)

astro-ph

The Shear TEsting Programme 2: Factors affecting high precision weak lensing analyses

The Shear TEsting Programme (STEP) is a collaborative project to improve the accuracy and reliability of weak lensing measurement, in preparation for the next generation of wide-field surveys. We review sixteen current and emerging shear measurement methods in a common language, and assess their performance by running them (blindly) on simulated images that contain a known shear signal. We determine the common features of algorithms that most successfully recover the input parameters. We achieve previously unattained discriminatory precision in our analysis, via a combination of more extensive simulations, and pairs of galaxy images that have been rotated with respect to each other, thus removing noise from their intrinsic ellipticities. The robustness of our simulation approach is also confirmed by testing the relative calibration of methods on real data. Weak lensing measurement has improved since the first STEP paper. Several methods now consistently achieve better than 2% precision, and are still being developed. However, the simulations can now distinguish all methods from perfect performance. Our main concern continues to be the potential for a multiplicative shear calibration bias: not least because this can not be internally calibrated with real data. We determine which galaxy populations are responsible and, by adjusting the simulated observing conditions, we also investigate the effects of instrumental and atmospheric parameters. We have isolated several previously unrecognised aspects of galaxy shape measurement, in which focussed development could provide further progress towards the sub-percent level of precision desired for future surveys. [ABRIDGED]

astro-ph

Broadband Optical Properties of Massive Galaxies: the Dispersion Around the Field Galaxy Color-Magnitude Relation Out to z~0.4

Using a sample of nearly 20,000 massive early-type galaxies selected from the Sloan Digital Sky Survey, we study the color-magnitude relation for the most luminous (L > 2.2 L^{*}) field galaxies in the redshift range 0.1<z<0.4 in several colors. The intrinsic dispersion in galaxy colors is quite small in all colors studied, but the 40 milli-mag scatter in the bluest colors is a factor of two larger than the 20 milli-mag measured in the reddest bands. While each of three simple models constructed for the star formation history in these systems can satisfy the constraints placed by our measurements, none of them produce color distributions matching those observed. Subdividing by environment, we find the dispersion for galaxies in clusters to be about 11% smaller than that of more isolated systems. Finally, having resolved the red sequence, we study the color dependence of the composite spectra. Bluer galaxies on the red sequence are found to have more young stars than red galaxies; the extent of this spectral difference is marginally better described by passive evolution of an old stellar population than by a model consisting of a recent trace injection of young stars.

astro-ph

Photometric Covariance in Multi-Band Surveys: Understanding the Photometric Error in the SDSS

In this paper we describe a detailed analysis of the photometric uncertainties present within the Sloan Digital Sky Survey (SDSS) imaging survey based on repeat observations of approximately 200 square degrees of the sky. We show that, for the standard SDSS aperture systems (petrocounts, counts_model, psfcounts and cmodel_counts), the errors generated by the SDSS photometric pipeline under-estimate the observed scatter in the individual bands. The degree of disagreement is a strong function of aperture and magnitude (ranging from 20% to more than a factor of 2). We also find that the photometry in the five optical bands can be highly correlated for both point sources and galaxies, although the correlation for point sources is almost entirely due to variable objects. Without correcting for this covariance the SDSS color errors could be in over-estimated by a factor of two to three. Combining these opposing effects, the SDSS errors on the colors differ from the observed color variation by approximately 10-20% for most apertures and magnitudes. We provide a prescription to correct the errors derived from the SDSS photometric pipeline as a function of magnitude and a semi-analytic method for generating the appropriate covariance between the different photometric passbands. Given the intrinsic nature of these correlations, we expect that all current and future multi-band surveys will also observe strongly covariant magnitudes. The ability of these surveys to complete their science goals is largely dependent on color-based target selection and photometric redshifts; these results show the importance of spending a significant fraction of early survey operations on re-imaging to empirically determine the photometric covariance of any observing/reduction pipeline.

astro-ph