arXiv ScienceSearch

arXiv subjects

Christopher Thomas

Publications and source records attributed to Christopher Thomas.

At least 19 recordsLinked to original sources

Superconducting NbN Resonator Parametric Amplifiers for Millimetre Wavelengths

We report the development of a reactive sputtering process for high $T_\mathrm{c}$ NbN films with high normal-state resistivity, tailored for kinetic inductance parametric amplifiers. The process includes precise control to ensure full nitridation of the target prior to deposition. Under optimized conditions, the resulting NbN thin films exhibit a critical temperature of $10.5\,\mathrm{K}$ and a resistivity of $\sim1000\,\mathrm{\mu\Omega\,cm}$. The high $T_\mathrm{c}$ of the NbN thin-films suggests strong potential for application over the entire millimetre-wave frequency range from $24\,\mathrm{GHz}$ to $300\,\mathrm{GHz}$, whereas the high resistivity suggests a reduced power requirement for the pump tone to achieve high gain. Resonator parametric amplifiers have been fabricated from these films using coplanar waveguide geometry. The devices were able to produce high gain exceeding $20\,\mathrm{dB}$ at $25\,\mathrm{GHz}$, with artefact-free, reproducible amplification profiles in good agreement with theoretical models.

cond-mat.supr-con

Superconducting Ring Resonators: Modelling, Simulation, and Experimental Characterisation

We present a theoretical and experimental study of superconducting ring resonators as an initial step toward their implementation in superconducting electronics and quantum technologies, with promising applications including superconducting parametric amplifiers with pump-signal isolation, flux-controlled quantum circuits, ultra-sensitive measurements in quantum sensing, and THz instrumentations. These devices have the potentially valuable property of supporting two orthogonal electromagnetic modes that couple to a common Cooper pair, quasiparticle, and phonon system. We present here a comprehensive theoretical and experimental analysis of the superconducting ring resonator system. We have developed superconducting ring resonator models that describe the key features of microwave behaviour to first order, providing insights into how transmission line inhomogeneities give rise to frequency splitting and mode rotation. Furthermore, we constructed signal flow graphs for a four-port ring resonator to numerically validate the behaviour predicted by our theoretical analysis. Superconducting ring resonators were fabricated in both coplanar waveguide and microstrip geometries using Al and Nb thin films. Microwave characterisation of these devices demonstrates close agreement with theoretical predictions. Our study reveals that frequency splitting and mode rotation are prevalent in ring systems with coupled degenerate modes, and these phenomena become distinctly resolved in high quality factor superconducting ring resonators.

cond-mat.supr-con

Non-degenerate pumping of superconducting resonator parametric amplifier with evidence of phase-sensitive amplification

Superconducting resonator parametric amplifiers are potentially important components for a wide variety of fundamental physics experiments and utilitarian applications. We propose and realise an operating scheme that achieves amplification through the use of non-degenerate pumps, which addresses two key challenges in the design of parametric amplifiers: non-continuous gain across the amplification band and pump tone removal. We have experimentally demonstrated the non-degenerate pumping scheme using a half-wave resonator amplifier based on NbN thin-film, and measured a peak gain of 26 dB and 3-dB bandwidth of 0.5 MHz. The two non-degenerate pump tones were positioned ~10 bandwidths above and below the frequency at which peak gain occurs. We have found the non-degenerate pumping scheme to be more stable compared to the usual degenerate pumping scheme in terms of gain drift over time, by a factor of 4. This scheme also retains the usual flexibility of NbN resonator parametric amplifiers in terms of reliable amplification in a ~4 K environment, and is suitable for cross-harmonic amplification. The use of pump tones at different frequencies allows phase-sensitive amplification when the signal tone is degenerate with the idler tone. A gain of 23 dB and squeezing ratio of 6 dB were measured.

quant-ph

Semantic Shield: Defending Vision-Language Models Against Backdooring and Poisoning via Fine-grained Knowledge Alignment

In recent years there has been enormous interest in vision-language models trained using self-supervised objectives. However, the use of large-scale datasets scraped from the web for training also makes these models vulnerable to potential security threats, such as backdooring and poisoning attacks. In this paper, we propose a method for mitigating such attacks on contrastively trained vision-language models. Our approach leverages external knowledge extracted from a language model to prevent models from learning correlations between image regions which lack strong alignment with external knowledge. We do this by imposing constraints to enforce that attention paid by the model to visual regions is proportional to the alignment of those regions with external knowledge. We conduct extensive experiments using a variety of recent backdooring and poisoning attacks on multiple datasets and architectures. Our results clearly demonstrate that our proposed approach is highly effective at defending against such attacks across multiple settings, while maintaining model utility and without requiring any changes at inference time

cs.CV

Quantum Noise Limited Phased Arrays for Single-Electron Cyclotron Radiation Emission Spectroscopy

Neutrino oscillation experiments show that neutrinos have mass; however, the absolute mass scale is exceedingly difficult to measure and is currently unknown. A promising approach is to measure the energies of the electrons released during the radioactive decay of tritium. The energies of interest are within a few eV of the 18.6 keV end point, and so are mildly relativistic. By capturing the electrons in a static magnetic field and measuring the frequency of the cyclotron radiation emitted the initial energy can be determined, but end-point events are infrequent, the observing times short, and the signal to noise ratios low. To achieve a resolution of $<$ 10 meV, single-electron emission spectra need to be recorded over large fields of view with highly sensitive receivers. The principles of Cylotron Radiation Emission Spectroscopy (CRES) have already been demonstrated by Project 8, and now there is considerable interest in increasing the FoV to $>$ 0.1 m$^3$. We consider a range of issues relating to the design and optimisation of inward-looking quantum-noise-limited microwave receivers for single-electron CRES, and present a single framework for understanding signal, noise and system-level behaviour. Whilst there is a great deal of literature relating to the design of outward-looking phased arrays for applications such as radar and telecommunications, there is very little coverage of the new issues that come into play when designing ultra-sensitive inward-looking phased arrays for volumetric spectroscopy and imaging.

physics.ins-det

Enhanced Chart Understanding in Vision and Language Task via Cross-modal Pre-training on Plot Table Pairs

Building cross-model intelligence that can understand charts and communicate the salient information hidden behind them is an appealing challenge in the vision and language(V+L) community. The capability to uncover the underlined table data of chart figures is a critical key to automatic chart understanding. We introduce ChartT5, a V+L model that learns how to interpret table information from chart images via cross-modal pre-training on plot table pairs. Specifically, we propose two novel pre-training objectives: Masked Header Prediction (MHP) and Masked Value Prediction (MVP) to facilitate the model with different skills to interpret the table information. We have conducted extensive experiments on chart question answering and chart summarization to verify the effectiveness of the proposed pre-training strategies. In particular, on the ChartQA benchmark, our ChartT5 outperforms the state-of-the-art non-pretraining methods by over 8% performance gains.

cs.CL

Metamagnetic transition in the two $f$ orbitals Kondo lattice model

In this work, we study the effects of a transverse magnetic field in a Kondo lattice model with two $f$ orbitals interacting with the conduction electrons. The $f$ electrons that are present on the same site interact through Hund's coupling, while on neighboring sites they interact through intersite exchange. We consider here that part of $f$ electrons are localized (orbital 1) while another part (orbital 2) are delocalized, as it is frequent in uranium systems. Then, only electrons in the localized orbital 1 interact through exchange interaction with the neighboring ones, while electrons in orbital 2 are coupled with conduction electrons through a Kondo interaction. We obtain a solution where ferromagnetism and Kondo effect coexist for small values of an applied transverse magnetic field for $T\rightarrow0$. Increasing the transverse field, two situations can be obtained when Kondo coupling vanishes: first, a metamagnetic transition occurs just before or at the same time of the fully polarized state, and second, a metamagnetic transition occurs when the spins are already pointing out along the magnetic field.

cond-mat.str-el

Weakly-Supervised Temporal Article Grounding

Given a long untrimmed video and natural language queries, video grounding (VG) aims to temporally localize the semantically-aligned video segments. Almost all existing VG work holds two simple but unrealistic assumptions: 1) All query sentences can be grounded in the corresponding video. 2) All query sentences for the same video are always at the same semantic scale. Unfortunately, both assumptions make today's VG models fail to work in practice. For example, in real-world multimodal assets (eg, news articles), most of the sentences in the article can not be grounded in their affiliated videos, and they typically have rich hierarchical relations (ie, at different semantic scales). To this end, we propose a new challenging grounding task: Weakly-Supervised temporal Article Grounding (WSAG). Specifically, given an article and a relevant video, WSAG aims to localize all ``groundable'' sentences to the video, and these sentences are possibly at different semantic scales. Accordingly, we collect the first WSAG dataset to facilitate this task: YouwikiHow, which borrows the inherent multi-scale descriptions in wikiHow articles and plentiful YouTube videos. In addition, we propose a simple but effective method DualMIL for WSAG, which consists of a two-level MIL loss and a single-/cross- sentence constraint loss. These training objectives are carefully designed for these relaxed assumptions. Extensive ablations have verified the effectiveness of DualMIL.

cs.CV

Incommensurate charge density wave on multiband intermetallic systems exhibiting competing orders

The appearance of an incommensurate charge density wave vector $\textbf{Q} = (Q_x,Q_y)$ on multiband intermetallic systems presenting commensurate charge density wave (CDW) and superconductivity (SC) orders is investigated. We consider a two-band model in a square lattice, where the bands have distinct effective masses. The incommensurate CDW (inCDW) and CDW phases arise from an interband Coulomb repulsive interaction, while the SC emerges due to a local intraband attractive interaction. For simplicity, all the interactions, the order parameters and hybridization between bands are considered $\textbf{k}$-independent. The multiband systems that we are interested are intermetallic systems with a $d$-band coexisting with a large $c$-band, for which a mean-field approach has proved suitable. We obtain the eigenvalues and eigenvectors of the Hamiltonian numerically and minimize the free energy density with respect to the diverse parameters of the model by means of the Hellmann-Feynman theorem. We investigate the system in real as well as momentum space and we find an inCDW phase with wave vector $\textbf{Q} = (\pi, Q_y) = (Q_x, \pi)$. Our numerical results show that the arising of an inCDW state depends on parameters, such as the magnitude of the inCDW and CDW interactions, band filling, hybridization and the relative depth of the bands. In general, inCDW tends to emerge at low temperatures, away from half-filling. We also show that, whether the CDW ordering is commensurate or incommensurate, large values of the relative depth between bands may suppress it. We discuss how each parameter of the model affects the emergence of an inCDW phase.

cond-mat.str-el

Beyond Grounding: Extracting Fine-Grained Event Hierarchies Across Modalities

Events describe happenings in our world that are of importance. Naturally, understanding events mentioned in multimedia content and how they are related forms an important way of comprehending our world. Existing literature can infer if events across textual and visual (video) domains are identical (via grounding) and thus, on the same semantic level. However, grounding fails to capture the intricate cross-event relations that exist due to the same events being referred to on many semantic levels. For example, in Figure 1, the abstract event of "war" manifests at a lower semantic level through subevents "tanks firing" (in video) and airplane "shot" (in text), leading to a hierarchical, multimodal relationship between the events. In this paper, we propose the task of extracting event hierarchies from multimodal (video and text) data to capture how the same event manifests itself in different modalities at different semantic levels. This reveals the structure of events and is critical to understanding them. To support research on this task, we introduce the Multimodal Hierarchical Events (MultiHiEve) dataset. Unlike prior video-language datasets, MultiHiEve is composed of news video-article pairs, which makes it rich in event hierarchies. We densely annotate a part of the dataset to construct the test benchmark. We show the limitations of state-of-the-art unimodal and multimodal baselines on this task. Further, we address these limitations via a new weakly supervised model, leveraging only unannotated video-article pairs from MultiHiEve. We perform a thorough evaluation of our proposed method which demonstrates improved performance on this task and highlight opportunities for future research.

cs.CV

Fine-Grained Visual Entailment

Visual entailment is a recently proposed multimodal reasoning task where the goal is to predict the logical relationship of a piece of text to an image. In this paper, we propose an extension of this task, where the goal is to predict the logical relationship of fine-grained knowledge elements within a piece of text to an image. Unlike prior work, our method is inherently explainable and makes logical predictions at different levels of granularity. Because we lack fine-grained labels to train our method, we propose a novel multi-instance learning approach which learns a fine-grained labeling using only sample-level supervision. We also impose novel semantic structural constraints which ensure that fine-grained predictions are internally semantically consistent. We evaluate our method on a new dataset of manually annotated knowledge elements and show that our method achieves 68.18\% accuracy at this challenging task while significantly outperforming several strong baselines. Finally, we present extensive qualitative results illustrating our method's predictions and the visual evidence our method relied on. Our code and annotated dataset can be found here: https://github.com/SkrighYZ/FGVE.

cs.CV

Joint Multimedia Event Extraction from Video and Article

Visual and textual modalities contribute complementary information about events described in multimedia documents. Videos contain rich dynamics and detailed unfoldings of events, while text describes more high-level and abstract concepts. However, existing event extraction methods either do not handle video or solely target video while ignoring other modalities. In contrast, we propose the first approach to jointly extract events from video and text articles. We introduce the new task of Video MultiMedia Event Extraction (Video M2E2) and propose two novel components to build the first system towards this task. First, we propose the first self-supervised multimodal event coreference model that can determine coreference between video events and text events without any manually annotated pairs. Second, we introduce the first multimodal transformer which extracts structured event information jointly from both videos and text documents. We also construct and will publicly release a new benchmark of video-article pairs, consisting of 860 video-article pairs with extensive annotations for evaluating methods on this task. Our experimental results demonstrate the effectiveness of our proposed method on our new benchmark dataset. We achieve 6.0% and 5.8% absolute F-score gain on multimodal event coreference resolution and multimedia event extraction.

cs.CV

Interplay between charge density wave and superconductivity in multi-band systems with inter-band Coulomb interaction

In this work we study the competition or coexistence between charge density wave (CDW) and superconductivity (SC) in a two-band model system in a square lattice. One of the bands has a net attractive interaction ($J_d$) that is responsible for SC. The model includes on-site Coulomb repulsion between quasi-particles in different bands ($U_{dc}$) and the hybridization ($V$) between them. We are interested in describing inter-metallic systems with a $d$-band of moderately correlated electrons, for which a mean-field approximation is adequate, coexisting with a large $sp$-band. For simplicity, all interactions and the hybridization $V$ are considered site-independent. We obtain the eigenvalues of the Hamiltonian numerically and minimize the free energy density with respect to the relevant parameters to obtain the phase diagrams as function of $J_d$, $U_{dc}$, $V$, composition ($n_{\mathrm{tot}}$) and the relative depth of the bands ($\epsilon_{d0}$). We consider two types of superconducting ground states coexisting with the CDW. One is a homogeneous ground state and the other is a pair density wave where the SC order parameter has the same spatial modulation of the CDW. Our results show that the CDW and SC orders compete, but depending on the parameters of the model these phases may coexist. The model reproduces most of the experimental features of high dimensionality ($d>1$) metals with competing CDW and SC states, including the existence of first and second-order phase transitions in their phase diagrams.

cond-mat.str-el

Learning to Transfer Visual Effects from Videos to Images

We study the problem of animating images by transferring spatio-temporal visual effects (such as melting) from a collection of videos. We tackle two primary challenges in visual effect transfer: 1) how to capture the effect we wish to distill; and 2) how to ensure that only the effect, rather than content or artistic style, is transferred from the source videos to the input image. To address the first challenge, we evaluate five loss functions; the most promising one encourages the generated animations to have similar optical flow and texture motions as the source videos. To address the second challenge, we only allow our model to move existing image pixels from the previous frame, rather than predicting unconstrained pixel values. This forces any visual effects to occur using the input image's pixels, preventing unwanted artistic style or content from the source video from appearing in the output. We evaluate our method in objective and subjective settings, and show interesting qualitative results which demonstrate objects undergoing atypical transformations, such as making a face melt or a deer bloom.

cs.CV

A Scalable Real-Time Architecture for Neural Oscillation Detection and Phase-Specific Stimulation

Oscillations in the local field potential (LFP) of the brain are key signatures of neural information processing. Perturbing these oscillations at specific phases in order to alter neural information processing is an area of active research. Existing systems for phase-specific brain stimulation typically either do not offer real-time timing guarantees (desktop computer based systems) or require extensive programming of vendor-specific equipment. This work presents a real-time detection system architecture that is platform-agnostic and that scales to thousands of recording channels, validated using a proof-of-concept microcontroller-based implementation.

eess.SP

Preserving Semantic Neighborhoods for Robust Cross-modal Retrieval

The abundance of multimodal data (e.g. social media posts) has inspired interest in cross-modal retrieval methods. Popular approaches rely on a variety of metric learning losses, which prescribe what the proximity of image and text should be, in the learned space. However, most prior methods have focused on the case where image and text convey redundant information; in contrast, real-world image-text pairs convey complementary information with little overlap. Further, images in news articles and media portray topics in a visually diverse fashion; thus, we need to take special care to ensure a meaningful image representation. We propose novel within-modality losses which encourage semantic coherency in both the text and image subspaces, which does not necessarily align with visual coherency. Our method ensures that not only are paired images and texts close, but the expected image-image and text-text relationships are also observed. Our approach improves the results of cross-modal retrieval on four datasets compared to five baselines.

cs.CV

Predicting the Politics of an Image Using Webly Supervised Data

The news media shape public opinion, and often, the visual bias they contain is evident for human observers. This bias can be inferred from how different media sources portray different subjects or topics. In this paper, we model visual political bias in contemporary media sources at scale, using webly supervised data. We collect a dataset of over one million unique images and associated news articles from left- and right-leaning news sources, and develop a method to predict the image's political leaning. This problem is particularly challenging because of the enormous intra-class visual and semantic diversity of our data. We propose a two-stage method to tackle this problem. In the first stage, the model is forced to learn relevant visual concepts that, when joined with document embeddings computed from articles paired with the images, enable the model to predict bias. In the second stage, we remove the requirement of the text domain and train a visual classifier from the features of the former model. We show this two-stage approach facilitates learning and outperforms several strong baselines. We also present extensive qualitative results demonstrating the nuances of the data.

cs.LG

Influence of the symmetry of the hybridization on the critical temperature of multi-band superconductors

In this work we study a two-band model of a superconductor in a square lattice. One band is narrow in energy and includes local Coulomb correlations between its quasi-particles. Pairing occurs in this band due to nearest neighbor attractive interactions. Extended s-wave, as well as d-wave symmetries of the superconducting order parameter are considered. The correlated electrons hybridize with those in another, wide conduction band through a k-dependent mixing, with even or odd parity depending on the nature of the orbitals. The many-body problem is treated within a slave-boson approach that has proved adequate to deal with the strong electronic correlations that are assumed here. Since applied pressure changes mostly the ratio between hybridization and bandwidths, we can use this ratio as a control parameter to obtain the phase diagrams of the model. We find that for a wide range of parameters, the critical temperature increases as a function of hybridization (pressure), with a region of first-order transitions. When frustration is introduced it gives rise to a stable superconducting phase. We find that superconductivity can be suppressed for specific values of band-filling due to the Coulomb repulsion. We show how pressure, composition and strength of correlations affect the superconductivity for different symmetries of the order parameter and the hybridization.

cond-mat.supr-con