arXiv ScienceSearch

arXiv subjects

Edward Berman

Publications and source records attributed to Edward Berman.

13 recordsLinked to original sources

Smoothness Errors in Dynamics Models and How to Avoid Them

Modern neural networks have shown promise for solving partial differential equations over surfaces, often by discretizing the surface as a mesh and learning with a mesh-aware graph neural network. However, graph neural networks suffer from oversmoothing, where a node's features become increasingly similar to those of its neighbors. Unitary graph convolutions, which are mathematically constrained to preserve smoothness, have been proposed to address this issue. Despite this, in many physical systems, such as diffusion processes, smoothness naturally increases and unitarity may be overconstraining. In this paper, we systematically study the smoothing effects of different GNNs for dynamics modeling and prove that unitary convolutions hurt performance for such tasks. We propose relaxed unitary convolutions that balance smoothness preservation with the natural smoothing required for physical systems. We also generalize unitary and relaxed unitary convolutions from graphs to meshes. In experiments on PDEs such as the heat and wave equations over complex meshes and on weather forecasting, we find that our method outperforms several strong baselines, including mesh-aware transformers and equivariant neural networks.

cs.LG

An ultra-high-resolution map of (dark) matter

Ordinary matter-including particles such as protons and neutrons-accounts for only about one sixth of all matter in the Universe. The rest is dark matter, which does not emit or absorb light but plays a fundamental role in galaxy and structure evolution. Because it interacts only through gravity, one of the most direct probes is weak gravitational lensing: the deflection of light from distant galaxies by intervening mass. Here we present an extremely detailed, wide-area weak-lensing mass map, covering 0.77 deg x 0.70 deg, using high-resolution imaging from the James Webb Space Telescope (JWST) as part of the COSMOS-Web survey. By measuring the shapes of 129 galaxies per square arcminute-many independently in the F115W and F150W bands-we achieve an angular resolution of 1.00 +/- 0.01 arcmin. Our map has more than twice the resolution of earlier Hubble Space Telescope maps, revealing how dark and luminous matter co-evolve across filaments, clusters, and under-densities. It traces mass features out to z ~ 2, including the most distant structure at z ~ 1.1. The sensitivity to high-redshift lensing constrains galaxy environments at the peak of cosmic star formation and sets a high-resolution benchmark for testing theories about the nature of dark matter and the formation of large-scale cosmic structure

astro-ph.CO

On Uncertainty Calibration for Equivariant Functions

Data-sparse settings such as robotic manipulation, molecular physics, and galaxy morphology classification are some of the hardest domains for deep learning. For these problems, equivariant networks can help improve modeling across undersampled parts of the input space, and uncertainty estimation can guard against overconfidence. However, until now, the relationships between equivariance and model confidence, and more generally equivariance and model calibration, has yet to be studied. Since traditional classification and regression error terms show up in the definitions of calibration error, it is natural to suspect that previous work can be used to help understand the relationship between equivariance and calibration error. In this work, we present a theory relating equivariance to uncertainty estimation. By proving lower and upper bounds on uncertainty calibration errors (ECE and ENCE) under various equivariance conditions, we elucidate the generalization limits of equivariant models and illustrate how symmetry mismatch can result in miscalibration in both classification and regression. We complement our theoretical framework with numerical experiments that clarify the relationship between equivariance and uncertainty using a variety of real and simulated datasets, and we comment on trends with symmetry mismatch, group size, and aleatoric and epistemic uncertainties.

cs.LG

On Soft Clustering For Correlation Estimators

Properly estimating correlations between objects at different spatial scales necessitates $\mathcal{O}(n^2)$ distance calculations. For this reason, most widely adopted packages for estimating correlations use clustering algorithms to approximate local trends. However, methods for quantifying the error introduced by this clustering have been understudied. In response, we present an algorithm for estimating correlations that is probabilistic in the way that it clusters objects, enabling us to quantify the uncertainty caused by clustering simply through model inference. These soft clustering assignments enable correlation estimators that are theoretically differentiable with respect to their input catalogs. Thus, we also build a theoretical framework for differentiable correlation functions and describe their utility in comparison to existing surrogate models. Notably, we find that repeated normalization and distance function calls slow gradient calculations and that sparse Jacobians destabilize precision, pointing towards either approximate or surrogate methods as a necessary solution to exact gradients from correlation functions. To that end, we close with a discussion of surrogate models as proxies for correlation functions. We provide an example that demonstrates the efficacy of surrogate models to enable gradient-based optimization of astrophysical model parameters, successfully minimizing a correlation function output. Our numerical experiments cover science cases across cosmology, from point spread function (PSF) modeling efforts to gravitational simulations to galaxy intrinsic alignment (IA).

astro-ph.IM

The COSMOS-Web Lens Survey (COWLS) I: Discovery of >100 high redshift strong lenses in contiguous JWST imaging

We present the COSMOS-Web Lens Survey (COWLS), a sample of over 100 strong lens candidates from the $0.54$\,deg$^2$ COSMOS-Web survey, discovered using exquisite James Webb Space Telescope (JWST) imaging across four wavebands. Following two rounds of visual inspection, over 100 candidates were ranked as `high confidence' or `likely' by at least $50\%$ of inspectors. The COWLS sample has several notable properties: (i) magnified source galaxies spanning redshifts $z \sim 0.1$ to $z \sim 9$, which therefore extend into the epoch of reionisation; (ii) the highest-redshift lens galaxies known, pushing galaxy density profile evolution studies beyond $z \sim 2$; (iii) all lenses are distributed within a contiguous $0.54$\,deg$^2$ region, allowing for joint strong and weak lensing analyses; and (iv) a subset exhibits lensed source emission ray-traced near the lens galaxy centers, enabling studies of supermassive black holes and dust absorption. A key innovation of our approach is the use of lens modelling to aid in identifying lenses that may otherwise be missed. This paper is accompanied by the first COWLS public release, providing JWST NIRCam imaging in four bands, lens models, pixelized source reconstructions and lens redshift estimates : https://github.com/Jammy2211/COWLS_COSMOS_Web_Lens_Survey

astro-ph.GA

The COSMOS-Web Lens Survey (COWLS) III: forecasts versus data

We compare forecasts for the abundance and properties of strong gravitational lenses in the COSMOS-Web survey, a $0.54$ deg$^2$ survey of the COSMOS field using the NIRCam and MIRI instruments aboard JWST, with the first catalogue of strong lens candidates identified in the observed NIRCam data, COWLS. We modify the lenspop package to produce a forecast for strong lensing in COSMOS-Web. We add a new mock galaxy catalogue to use as the source population, as well as the COSMOS-Web survey specifications, including the transmission data for the four NIRCam filters used. We forecast 107 strong lenses can be detected in COSMOS-Web across all bands, assuming complete subtraction of the lens galaxy light. The majority of the lenses are forecast to have small Einstein radii ($\theta_{\rm E} < 1$ arcsecond) and lie at redshifts between $0 < z <2$, whilst the source redshift distribution peaks at $z\sim 3$ and has a long tail extending up to $z \sim 11$, unambiguously showing that strong lensing in JWST can probe the entirety of the epoch of reionisation. We compare our forecast with the distributions of Einstein radii, lens photometric redshifts, and lens and source magnitudes in the observed lenses, finding that whilst the forecast and observed Einstein radii distributions match, the redshifts and magnitudes do not. The observed lens redshift distribution peaks at a slightly lower redshift than the forecast one, whilst the lens magnitudes are systematically brighter in the observed data than in the forecast.

astro-ph.GA

The COSMOS-Web Lens Survey (COWLS) II: depth, resolution, and NIR coverage from JWST reveal 17 spectacular lenses

The COSMOS-Web Lens Survey (COWLS) presents the first systematic search for strong gravitational lenses in the COSMOS-Web field using data from the \textit{James Webb} Space Telescope (\textit{JWST}). Using high-resolution NIRCam imaging, we visually inspected over 42\,660 galaxies and identified over 400 lensing candidates. From this sample and based on \textit{JWST}/NIRCam imaging only, we report here the 17 most obvious and spectacular strong lensing systems. These lenses, characterised by large Einstein rings and arcs and their distinct lens and source colours, were found through only the visual inspection of the lens-light-subtracted image data and were immediately visible due to their spectacular appearance. We showcase how spectacular strong lenses are at the extremes of lens parameter space. Their exceptionally high signal-to-noise, multi-wavelength imaging enables unprecedented lensing analysis, including `\textit{HST}-dark' source galaxies that are also invisible in the deeper bluer \textit{JWST} wavebands, enabling clean deblending between the lens and the source. Sources may exhibit dramatic morphological changes across wavelengths, and dust absorption within lenses may be detectable by eye. No other instrument, including the \textit{Hubble} Space Telescope, can discover or image such lenses with comparable detail. We estimate that \textit{JWST} uncovers a new spectacular lens approximately every 10 to 12 NIRCam pointings, suggesting that over 40 such lenses remain undetected within its first three years of observations. All COWLS data is publicly available on GitHub.

astro-ph.GA

The State of Julia for Scientific Machine Learning

Julia has been heralded as a potential successor to Python for scientific machine learning and numerical computing, boasting ergonomic and performance improvements. Since Julia's inception in 2012 and declaration of language goals in 2017, its ecosystem and language-level features have grown tremendously. In this paper, we take a modern look at Julia's features and ecosystem, assess the current state of the language, and discuss its viability and pitfalls as a replacement for Python as the de-facto scientific machine learning language. We call for the community to address Julia's language-level issues that are preventing further adoption.

cs.LG

Not-so-little Red Dots: Two massive and dusty starbursts at z~5-7 pushing the limits of star formation discovered by JWST in the COSMOS-Web survey

We present the properties of two candidate massive ($M_\star\sim10^{11}M_\odot$) and dusty ($A_{\rm v}>2.5$ mag) galaxies at $z=5-7$ in the first 0.28 deg$^2$ of the COSMOS-Web survey. One object is spectroscopically confirmed at $z_{\rm spec}=5.051$, while the other has a robust $z_{\rm phot}=6.7\pm0.3$. Thanks to their extremely red colors ($F277W-F444W\sim1.7$ mag), these galaxies satisfy the nominal color-selection for the widely-studied ``little red dot" (LRD) population with the exception of their spatially-resolved morphologies. The morphology of our targets allows us to conclude that their red continuum is dominated by highly obscured stellar emission and not by reddened nuclear activity. Using a variety of SED-fitting tools and star formation histories, we estimate the stellar masses to be $\log(M_\star)=11.32^{+0.07}_{-0.15}$ $M_\odot$ and $\log(M_\star)=11.2^{+0.1}_{-0.2}$ $M_\odot$, respectively, with a red continuum emission dominated by a recent episode of star formation. We then compare their number density to the halo mass function to infer stellar baryon fractions of $\epsilon_\star\sim0.25$ and $\epsilon_\star\sim0.5$. Both are significantly higher than what is commonly observed in lower-z galaxies or more dust-obscured galaxies at similar redshifts. With very bright ultra-high-z Lyman-Break Galaxies and some non-AGN dominated LRDs, such ``extended" LRDs represent another population that may require very efficient star formation at early times.

astro-ph.GA

Efficient PSF Modeling with ShOpt.jl: A PSF Benchmarking Study with JWST NIRCam Imaging

With their high angular resolutions of 30--100 mas, large fields of view, and complex optical systems, imagers on next-generation optical/near-infrared space observatories, such as the Near-Infrared Camera (NIRCam) on the James Webb Space Telescope (JWST), present both new opportunities for science and also new challenges for empirical point spread function (PSF) characterization. In this context, we introduce ShOpt, a new PSF fitting tool developed in Julia and designed to bridge the advanced features of PIFF (PSFs in the Full Field of View) with the computational efficiency of PSFEx (PSF Extractor). Along with ShOpt, we propose a suite of non-parametric statistics suitable for evaluating PSF fit quality in space-based imaging. Our study benchmarks ShOpt against the established PSF fitters PSFEx and PIFF using real and simulated COSMOS-Web Survey imaging. We assess their respective PSF model fidelity with our proposed diagnostic statistics and investigate their computational efficiencies, focusing on their processing speed relative to the complexity and size of the PSF models. We find that ShOpt can already achieve PSF model fidelity comparable to PSFEx and PIFF while maintaining competitive processing speeds, constructing PSF models for large NIRCam mosaics within minutes.

astro-ph.IM

Can It Edit? Evaluating the Ability of Large Language Models to Follow Code Editing Instructions

A significant amount of research is focused on developing and evaluating large language models for a variety of code synthesis tasks. These include synthesizing code from natural language, synthesizing tests from code, and synthesizing explanations of code. In contrast, the behavior of instructional code editing with LLMs is understudied. These are tasks in which the model is provided a block of code and an instruction to modify the code. The editing instruction may ask for a feature to be added or removed, describe a bug and ask for a fix, or ask for a different kind of solution. We introduce a carefully crafted benchmark of code editing tasks and use it to evaluate several cutting edge LLMs. Our evaluation exposes a significant gap between the capabilities of state-of-the-art open and closed models. For example, even GPT-3.5-Turbo is better than the best open model at code editing tasks. We also introduce a new, carefully curated, permissively licensed training dataset of code editing tasks coupled with natural language instructions. Using this training dataset, we show that we can fine-tune open Code LLMs to significantly improve their code editing capabilities, closing the gap between open and closed models. All code, data, and models are available at https://github.com/nuprl/CanItEdit.

cs.SE

ShOpt.jl: A Julia Package for Empirical Point Spread Function Characterization of JWST NIRCam Data

As astronomical data grows in volume and complexity, the scalability of analysis software becomes increasingly important. At the same time, astrophysics analysis software relies heavily on open-source contributions, so languages and tools that prioritize both performance and readability are especially valuable. Julia, with its just-in-time compiler and high level syntax, offers a compelling alternative to traditional languages like Python or C. In this paper, we outline ShOpt.jl, a new software package for point spread function (PSF) characterization written in Julia. ShOpt.jl features a number of performance optimizations, such as multithreading, the use of preconditioners, and the implementation of the memory-limited Broyden-Fletcher-Goldfarb-Shanno algorithm, as well as the flexibility to choose between principal component analysis, an autoencoder, and analytic profiles for PSF characterization. As observatories like the James Webb Space Telescope bring astrophysics into a new era of wide-field, high-resolution imaging, the challenges of PSF modeling become more pronounced. Tools like ShOpt.jl provide the community with a scalable, efficient, and accurate solution to these challenges, while also demonstrating the potential of Julia as a language that meets the demands of modern astrophysical research.

astro-ph.IM

Reflexion: Language Agents with Verbal Reinforcement Learning

Large language models (LLMs) have been increasingly used to interact with external environments (e.g., games, compilers, APIs) as goal-driven agents. However, it remains challenging for these language agents to quickly and efficiently learn from trial-and-error as traditional reinforcement learning methods require extensive training samples and expensive model fine-tuning. We propose Reflexion, a novel framework to reinforce language agents not by updating weights, but instead through linguistic feedback. Concretely, Reflexion agents verbally reflect on task feedback signals, then maintain their own reflective text in an episodic memory buffer to induce better decision-making in subsequent trials. Reflexion is flexible enough to incorporate various types (scalar values or free-form language) and sources (external or internally simulated) of feedback signals, and obtains significant improvements over a baseline agent across diverse tasks (sequential decision-making, coding, language reasoning). For example, Reflexion achieves a 91% pass@1 accuracy on the HumanEval coding benchmark, surpassing the previous state-of-the-art GPT-4 that achieves 80%. We also conduct ablation and analysis studies using different feedback signals, feedback incorporation methods, and agent types, and provide insights into how they affect performance.

cs.AI