arXiv ScienceSearch

arXiv subjects

Ben Thompson

Publications and source records attributed to Ben Thompson.

8 recordsLinked to original sources

Verbalizable Representations Form a Global Workspace in Language Models

Out of everything the human brain processes, only a small fraction is consciously accessible, in the sense of being available for verbal report, deliberate control, and flexible reasoning. In this paper, we present evidence that an analogous functional distinction has emerged in large language models. Using a new interpretability technique, the Jacobian lens, we identify the representations a model is poised to verbalize at any point in its processing. These representations, which we collectively call the J-space, exhibit the functional properties characteristic of a global workspace: their contents can be reported, deliberately summoned and held, used to carry the intermediate steps of silent reasoning, and passed as arguments to arbitrary downstream computations, while automatic processing such as text parsing and routine inference proceeds without them. The J-space also has structural signatures that global workspace theory associates with conscious access: it carries coherent content only in an intermediate band of layers, holds on the order of tens of concepts at a time, and is broadcast by the model's weights more widely than other representations. These properties make it a practical window into a model's unspoken thinking. In alignment audits, it reveals strategic deliberation, evaluation awareness, and trained-in misaligned dispositions that never appear in the model's outputs. We find that post-training installs the Assistant's point of view in the workspace, and we introduce counterfactual reflection training, which improves behavior by training only what a model would say if interrupted and asked to reflect. These results indicate that language models maintain a small, privileged set of representations bearing some of the functional hallmarks of conscious access, and that decoding these representations sheds light on ongoing cognitive processes.

cs.CL

Turn-Averaged SAEs for Feature Discovery and Long-Context Attribution

Sparse autoencoders (SAEs) have become a useful tool for extracting interpretable features in language models. However, standard SAE architectures operate on individual token activations, meaning that the number of active features scales linearly with context length, and studying long model transcripts becomes difficult. We introduce turn-averaged SAEs, which represent a single Human or Assistant turn with a fixed number of features by learning to reconstruct the average model activation across the turn. We find that turn-averaged features describe a single turn's high-level characteristics more completely than per-token features when judged by an LLM. We also demonstrate that turn-averaged SAEs greatly simplify common downstream uses of SAEs like attribution graphs. Broadly, turn-averaged SAEs make interpretability techniques practical at long context lengths.

cs.CL

Introduction to Quantum Ophthalmology

Quantum technologies are rapidly advancing across multiple research domains, with a growing impact on biomedical imaging and sensing. We examine their emerging role in ophthalmology through four complementary directions: photon-limited retinal imaging, correlation based imaging, nanoscale optical probes, and quantum-limited visual perception. Advances in optical coherence tomography and single-photon detection enable imaging under strict photon budget constraints, reducing phototoxicity while preserving image quality. Correlation-based approaches, including ghost imaging, offer alternative strategies for image formation in low-light and scattering environments, although practical implementation remains limited by detection efficiency and acquisition time. In parallel, nanoscale optical platforms such as quantum dots provide tunable and photostable probes for enhanced contrast and targeted delivery, with ongoing challenges related to biocompatibility and clinical translation. Finally, experiments at the single-photon level and with structured light fields demonstrate how the visual system itself operates near physical detection limits and can be probed using controlled optical states. While many of these approaches remain at an early stage, they collectively illustrate how quantum and quantum-inspired methods may augment current ophthalmic imaging and diagnostic technologies while providing new tools for studying visual function under well-defined physical constraints.

physics.med-ph

Eye dominance and testing order effects in the circularly-oriented macular pigment optical density measurements that rely on the perception of structured light-based stimuli

Psychophysical discrimination of structured light (SL) stimuli may be useful in screening for various macular disorders, including degenerative macular diseases. The circularly-oriented macular pigment optical density (coMPOD), calculated from the discrimination performance of SL-induced entoptic phenomena, may reveal a novel functional biomarker of macular health. In this study, we investigated the potential influence of eye dominance and testing order effects on SL-based stimulus perception, factors that potentially influence the sensitivity of screening tests based on SL technology. A total of 28 participants (aged 18-38 years) were selected for the study after undergoing a comprehensive eye examination. A psychophysical task was performed where various SL-based entoptic images with multiple azimuthal fringes rotating with a specific temporal frequency were projected onto the participants' retinas. By occluding the central areas of entoptic images, we measured the retinal eccentricity ($R$) of the perceivable area of the stimuli. The slope of the coMPOD profile ($a$-value) was calculated for each participant using a spatiotemporal sensitivity model that takes into account the perceptual threshold measurements of structured light stimuli with varying spatial densities and temporal frequencies. The Pearson correlation coefficient between eye dominance and testing order effects was $r=0.8$ ($p<0.01$). The Bland-Altman plots for both factors indicated zero bias. The results indicate repeatable measurements for both eyes, implying minimal impact from eye dominance and testing order on SL-based stimulus perception. The results provide a foundation for future studies exploring the clinical utility of SL tools in eye health.

q-bio.NC

Measuring the visual angle of polarization-related entoptic phenomena using structured light

The ability to perceive polarization-related entoptic phenomena arises from the dichroism of macular pigments held in Henle's fiber layer of the retina and can be inhibited by retinal diseases such as age-related macular degeneration, which alter the structure of the macula. Structured light tools enable the direct probing of macular pigment density through the perception of polarization-dependent entoptic patterns. Here, we directly measure the visual angle of an entoptic pattern created through the illumination of the retina with a structured state of light and a perception task that is insensitive to corneal birefringence. The central region of the structured light stimuli was obstructed, with the size of the obstruction varying according to a psychophysical staircase. The perceived size of the entoptic pattern was observed to vary between participants, with an average visual angle threshold radius of $9.5^\circ \pm 0.9^\circ$, 95% C.I. = [$5.8^\circ$, $13^\circ$], in a sample of healthy participants. These results (with eleven azimuthal fringes) differ markedly from previous estimates of the Haidinger's brush phenomenon's extant (two azimuthal fringes), of $3.75^\circ$, suggesting that higher azimuthal fringe density increases pattern visibility. The increase in apparent size and clarity of entoptic phenomenon produced by the presented structured light stimuli may possess greater potential to detect the early signs of macular disease over perception tasks using uniform polarization stimuli.

physics.med-ph

Two-Dimensional Quantum Material Identification via Self-Attention and Soft-labeling in Deep Learning

In quantum machine field, detecting two-dimensional (2D) materials in Silicon chips is one of the most critical problems. Instance segmentation can be considered as a potential approach to solve this problem. However, similar to other deep learning methods, the instance segmentation requires a large scale training dataset and high quality annotation in order to achieve a considerable performance. In practice, preparing the training dataset is a challenge since annotators have to deal with a large image, e.g 2K resolution, and extremely dense objects in this problem. In this work, we present a novel method to tackle the problem of missing annotation in instance segmentation in 2D quantum material identification. We propose a new mechanism for automatically detecting false negative objects and an attention based loss strategy to reduce the negative impact of these objects contributing to the overall loss function. We experiment on the 2D material detection datasets, and the experiments show our method outperforms previous works.

cs.CV

Human psychophysical discrimination of spatially dependant Pancharatnam-Berry phases in optical spin-orbit states

We tested the ability of human observers to discriminate distinct profiles of spatially dependant geometric phases when directly viewing stationary structured light beams. Participants viewed polarization coupled orbital angular momentum (OAM) states, or ``spin-orbit'' states, in which the OAM was induced through Pancharatnam-Berry phases. The coupling between polarization and OAM in these beams manifests as spatially dependant polarization. Regions of uniform polarization are perceived as specifically oriented Haidinger's brushes, and study participants discriminated between two spin-orbit states based on the rotational symmetry in the spatial orientations of these brushes. Participants used self-generated eye movements to prevent adaptation to the visual stimuli. After initial training, the participants were able to correctly discriminate between two spin-orbit states, differentiated by OAM $=\pm1$, with an average success probability of $69\%$ ($S.D. = 22\%$, $p = 0.013$). These results support our previous observation that human observers can directly perceive spin-orbit states, and extend this finding to non-rotating beams, OAM modes induced via Pancharatnam-Berry phases, and the discrimination of states that are differentiated by OAM.

physics.optics

WIYN Open Cluster Study LXII: Comparison of Isochrone Systems using Deep Multi-Band Photometry of M35

Current generation stellar isochrone models exhibit non-negligible discrepancies due to variations in the input physics. The success of each model is determined by how well it fits the observations, and this paper aims to disentangle contributions from the various physical inputs. New deep, wide-field optical and near-infrared photometry ($UBVRIJHK_S$) of the cluster M35 is presented, against which several isochrone systems are compared: Padova, PARSEC, Dartmouth and Y$^2$. Two different atmosphere models are applied to each isochrone: ATLAS9 and BT-Settl. For any isochrone set and atmosphere model, observed data are accurately reproduced for all stars more massive then $0.7$ M$_\odot$. For stars less massive than 0.7 M$_\odot$, Padova and PARSEC isochrones consistently produce higher temperatures than observed. Dartmouth and Y$^2$ isochrones with BT-Settl atmospheres reproduce optical data accurately, however they appear too blue in IR colors. It is speculated that molecular contributions to stellar spectra in the near-infrared may not be fully explored, and that future study may reconcile these differences.

astro-ph.SR