arXiv ScienceSearch

arXiv subjects

Han Fu

Publications and source records attributed to Han Fu.

At least 19 recordsLinked to original sources

DELE-w0.5: Inferring Action from Future Latent State for Robotic Manipulation

World-Action Models (WAMs) build robot control on video-generation backbones, which jointly predict dense future visual trajectories and robot actions. We argue that video generation is an unnecessary intermediate objective for world-action modeling. For robotic manipulation, the goal of a world model is not to reproduce how the world looks at every intermediate moment, but to predict the state that the world will reach after an action is executed. The intermediate frames only describe the visual transition between physical states, which consumes substantial model capacity and computation, but do not directly specify the physical outcome that the robot action is intended to produce. In this paper, we propose DELE-w0.5, which infers robot actions from predicted future states without relying on video generation. Concretely, DELE-w0.5 infers the action sequence from its corresponding compact future latent state. The future latent state captures the action-relevant physical outcome of robot interaction and serves as an explicit bridge between world modeling and action generation. The core design principle of DELE-w0.5 is to model how the physical world changes under robot actions, rather than how its visual appearance evolves frame by frame. This formulation removes the high-dimensional visual redundancy introduced by dense video representations, and it therefore enables cheaper training and low-latency inference. Across 640 real-robot trials on four long-horizon manipulation tasks, our DELE-w0.5 achieves the best performance among all compared policies, attaining 62.5% overall full-task success and 81.3% macro ordered-stage progress. It outperforms the strongest baseline by 32.5 percentage points in full-task success and 20.1 percentage points in macro progress.

cs.RO

CloudCons: A Comprehensive End-to-End Benchmark for Cloud Resource Consolidation

Driven by conservative over-provisioning to guarantee service reliability, resource utilization in cloud data centers remains at low levels. To mitigate this, the forecast-then-optimize paradigm has emerged to optimize consolidation by anticipating future demands. While emerging time series foundation models promise to enhance this paradigm through zero-shot generalization, existing benchmarks focus solely on prediction error metrics. The actual decision utility of these advanced models remains unverified, rendering their practical value for downstream tasks uncertain. To bridge this gap, we propose CloudCons, a comprehensive end-to-end benchmark designed to evaluate forecasting models within the specific context of cloud resource consolidation. We build high-quality datasets that cover diverse workloads from Huawei Cloud, Microsoft Azure, and Google Borg, capturing distinct service characteristics ranging from synchronized diurnal rhythms to stochastic, pulse-like bursts and high-frequency noise. We conduct an extensive evaluation of statistical, deep learning, and foundation models. Our experiments reveal a pivotal finding: while foundation models demonstrate superior zero-shot forecasting accuracy, this advantage does not inherently translate into better decision utility. Of practical significance, we systematically analyze how the selection of predictive quantiles acts as a critical lever. We provide actionable guidelines for calibrating these selections to balance the trade-off between resource efficiency and service reliability, offering vital insights for real-world deployment decisions.

cs.AI

TSFMAudit: Data Contamination Auditing in Forecasting Time Series Foundation Models

Time series foundation models (TSFMs) are increasingly pretrained on large corpora, raising concerns that evaluation datasets may have been exposed during pretraining and thus yield overly optimistic performance estimates. Auditing such contamination is challenging in time series because signals are continuous and heterogeneous, and often lack corpus documentation. To the best of our knowledge, this is the first work to study pretraining contamination auditing for TSFMs. We formalize the problem of pretraining contamination auditing for TSFMs and propose TSFMAudit, a method based on probe adaptation dynamics. Our key intuition is that contamination manifests as unusually efficient adaptation: after a fine tuning probe, contaminated datasets tend to exhibit faster loss reduction with smaller backbone movement. We evaluate TSFMAudit on 6 TSFMs and 187 datasets using documented training source evidence as supervision, and compare against 10 competitive baselines adapted from the LLM literature.

cs.LG

Delta-Adapter: Scalable Exemplar-Based Image Editing with Single-Pair Supervision

Exemplar-based image editing applies a transformation defined by a source-target image pair to a new query image. Existing methods rely on a pair-of-pairs supervision paradigm, requiring two image pairs sharing the same edit semantics to learn the target transformation. This constraint makes training data difficult to curate at scale and limits generalization across diverse edit types. We propose Delta-Adapter, a method that learns transferable editing semantics under single-pair supervision, requiring no textual guidance. Rather than directly exposing the exemplar pair to the model, we leverage a pre-trained vision encoder to extract a semantic delta that encodes the visual transformation between the two images. This semantic delta is injected into a pre-trained image editing model via a Perceiver-based adapter. Since the target image is never directly visible to the model, it can serve as the prediction target, enabling single-pair supervision without requiring additional exemplar pairs. This formulation allows us to leverage existing large-scale editing datasets for training. To further promote faithful transformation transfer, we introduce a semantic delta consistency loss that aligns the semantic change of the generated output with the ground-truth semantic delta extracted from the exemplar pair. Extensive experiments demonstrate that Delta-Adapter consistently improves both editing accuracy and content consistency over four strong baselines on seen editing tasks, while also generalizing more effectively to unseen editing tasks. Code will be available at https://delta-adapter.github.io.

cs.CV

Where did we fail? -- Reproducing build failures in embedded open source software

Due to hardware-software co-development in embedded systems, continuous integration (CI) builds frequently fail because of complex cross-compilation, board configurations, and toolchain constraints. Although CI build logs contain valuable diagnostic information, they are short-lived and difficult to reuse due to heterogeneous runners, toolchains, and log formats. To address these challenges, we present PhantomRun, a unified abstraction layer and publicly reusable dataset that standardizes the retrieval, storage, and reproduction of CI build logs and metadata. Across 4628 failing CI runs, we reconstructed 91.8% of builds and preserved execution outcomes in 98% of evaluated cases. PhantomRun provides two core capabilities: retrieving the build log of any commit and faithfully re-executing the corresponding build in a controlled environment. By exposing all build artifacts and metadata in a uniform, machine-readable format, PhantomRun enables reproducible and longitudinal studies of CI failures. An empirical evaluation shows that reproduced builds closely match their originals, typically differing only in timestamps or minor nondeterministic reordering, demonstrating the feasibility of large-scale historical CI reconstruction.

cs.SE

PhantomRun: Auto Repair of Compilation Errors in Embedded Open Source Software

Continuous Integration (CI) pipelines for embedded software sometimes fail during compilation, consuming significant developer time for debugging. We study four major open-source embedded system projects, spanning over 4000 build failures from the project's CI runs. We find that hardware dependencies account for the majority of compilation failures, followed by syntax errors and build-script issues. Most repairs need relatively small changes, making automated repair potentially suitable as long as the diverse setups and lack of test data can be handled. In this paper, we present PhantomRun, an automated framework that leverages large language models (LLMs) to generate and validate fixes for CI compilation failures. The framework addresses the challenge of diverse build infrastructures and tool chains across embedded system projects by providing an adaptation layer for GitHub Actions and GitLab CI and four different build systems. PhantomRun utilizes build logs, source code, historical fixes, and compiler error messages to synthesize fixes using LLMs. Our evaluations show that PhantomRun successfully repairs up to 45% of CI compilation failures across the targeted projects, demonstrating the viability of LLM-based repairs for embedded-system CI pipelines.

cs.SE

Enhanced superconducting diode effect in hybrid Josephson junctions

The superconducting diode effect (SDE) has recently been observed in various systems, sparking interest in novel superconducting devices and offering a new platform to probe intrinsic material properties. Josephson junctions with strong Rashba spin-orbit coupling have exhibited nonreciprocal critical currents under applied magnetic fields. In this work, we investigate the SDE in Josephson junctions incorporating periodic hole arrays patterned into the superconducting leads on InAs heterostructures with epitaxial aluminum. We observe an enhanced diode effect when a top gate depletes the 2DEG in the region of the hole arrays, while preserving the overall supercurrent. Theoretical analysis shows that the physics behind this phenomenon is the increased difference of transparency between different bands in the junction. These results highlight a new pathway for engineering and controlling nonreciprocal superconducting transport in hybrid systems.

cond-mat.mes-hall

Quench induced collective excitations: from breathing to acoustic modes

In trapped Bose-Einstein condensates, interaction quenches which are abrupt changes of the interaction strength typically implemented via Feshbach tuning, are a practical and widely used protocol to address far-from-equilibrium collective modes. Using both numerical Gross Pitaevskii and analytical schemes we study these interaction-quench-induced collective modes in a harmonically trapped two-dimensional Bose--Einstein condensate contrasting the behavior found at low and high energies. In the low-lying regime, we characterize realistic circumstances in which there is a breakdown of the expected scale invariance so that the collective excitations follow hydrodynamic theory instead of the predictions given by SO(2,1) conformal symmetry. In the high energy regime, we focus on important trap effects associated with acoustic oscillations which have been of interest experimentally. This comprehensive analysis of the collective excitations in trapped two-dimensional Bose-Einstein condensates is experimentally accessible. Through their frequencies and damping, this reflects an important built-in spectroscopy of such many-body states.

cond-mat.quant-gas

Auto-repair without test cases: How LLMs fix compilation errors in large industrial embedded code

The co-development of hardware and software in industrial embedded systems frequently leads to compilation errors during continuous integration (CI). Automated repair of such failures is promising, but existing techniques rely on test cases, which are not available for non-compilable code. We employ an automated repair approach for compilation errors driven by large language models (LLMs). Our study encompasses the collection of more than 40000 commits from the product's source code. We assess the performance of an industrial CI system enhanced by four state-of-the-art LLMs, comparing their outcomes with manual corrections provided by human programmers. LLM-equipped CI systems can resolve up to 63 % of the compilation errors in our baseline dataset. Among the fixes associated with successful CI builds, 83 % are deemed reasonable. Moreover, LLMs significantly reduce debugging time, with the majority of successful cases completed within 8 minutes, compared to hours typically required for manual debugging.

cs.SE

VisionTS++: Cross-Modal Time Series Foundation Model with Continual Pre-trained Vision Backbones

Recent studies have indicated that vision models pre-trained on images can serve as time series foundation models (TSFMs) by reformulating time series forecasting (TSF) as image reconstruction. However, effective cross-modal transfer from vision to time series remains challenging due to three discrepancies: (1) the data-modality gap between structured, bounded image data and unbounded, heterogeneous time series; (2) the multivariate-forecasting gap between fixed RGB-three-channel vision models and time series with arbitrary numbers of variates; and (3) the probabilistic-forecasting gap between the deterministic outputs of vision models and the requirement for uncertainty-aware probabilistic predictions. To bridge these gaps, we propose VisonTS++, a TSFM based on continual pre-training of a vision model on large-scale time series. Our approach introduces three key innovations: (1) vision-model-based filtering to identify high-quality sequences to stabilize pre-training and mitigate modality gap; (2) colorized multivariate conversion, encoding multivariate series as multi-subfigure RGB images to enhance cross-variate modeling; (3) multi-quantile forecasting, using parallel reconstruction heads to generate quantile forecasts without parametric assumptions. Experiments show that VisionTS++ achieves state-of-the-art performance in both in-distribution and out-of-distribution forecasting, outperforming specialized TSFMs by 6%-44% in MSE reduction and ranking first in GIFT-Eval benchmark which comprises 23 datasets across 7 domains. Our work demonstrates that with appropriate adaptation, vision models can effectively generalize to TSF, thus advancing the pursuit of universal TSFMs. Code is available at https://github.com/HALF111/VisionTSpp.

cs.CV

The Power of Architecture: Deep Dive into Transformer Architectures for Long-Term Time Series Forecasting

Transformer-based models have recently become dominant in Long-term Time Series Forecasting (LTSF), yet the variations in their architecture, such as encoder-only, encoder-decoder, and decoder-only designs, raise a crucial question: What Transformer architecture works best for LTSF tasks? However, existing models are often tightly coupled with various time-series-specific designs, making it difficult to isolate the impact of the architecture itself. To address this, we propose a novel taxonomy that disentangles these designs, enabling clearer and more unified comparisons of Transformer architectures. Our taxonomy considers key aspects such as attention mechanisms, forecasting aggregations, forecasting paradigms, and normalization layers. Through extensive experiments, we uncover several key insights: bi-directional attention with joint-attention is most effective; more complete forecasting aggregation improves performance; and the direct-mapping paradigm outperforms autoregressive approaches. Furthermore, our combined model, utilizing optimal architectural choices, consistently outperforms several existing models, reinforcing the validity of our conclusions. We hope these findings offer valuable guidance for future research on Transformer architectural designs in LTSF. Our code is available at https://github.com/HALF111/TSF_architecture.

cs.LG

Mathematical modelling to inform outbreak response vaccination

Mathematical models are established tools to assist in outbreak response. They help characterise complex patterns in disease spread, simulate control options to assist public health authorities in decision-making, and longer-term operational and financial planning. In the context of vaccine-preventable diseases (VPDs), vaccines are one of the most-cost effective outbreak response interventions, with the potential to avert significant morbidity and mortality through timely delivery. Models can contribute to the design of vaccine response by investigating the importance of timeliness, identifying high-risk areas, prioritising the use of limited vaccine supply, highlighting surveillance gaps and reporting, and determining the short- and long-term benefits. In this review, we examine how models have been used to inform vaccine response for 10 VPDs, and provide additional insights into the challenges of outbreak response modelling, such as data gaps, key vaccine-specific considerations, and communication between modellers and stakeholders. We illustrate that while models are key to policy-oriented outbreak vaccine response, they can only be as good as the surveillance data that inform them.

q-bio.PE

In industrial embedded software, are some compilation errors easier to localize and fix than others?

Industrial embedded systems often require specialized hardware. However, software engineers have access to such domain-specific hardware only at the continuous integration (CI) stage and have to use simulated hardware otherwise. This results in a higher proportion of compilation errors at the CI stage than in other types of systems, warranting a deeper study. To this end, we create a CI diagnostics solution called ``Shadow Job'' that analyzes our industrial CI system. We collected over 40000 builds from 4 projects from the product source code and categorized the compilation errors into 14 error types, showing that the five most common ones comprise 89 % of all compilation errors. Additionally, we analyze the resolution time, size, and distance for each error type, to see if different types of compilation errors are easier to localize or repair than others. Our results show that the resolution time, size, and distance are independent of each other. Our research also provides insights into the human effort required to fix the most common industrial compilation errors. We also identify the most promising directions for future research on fault localization.

cs.SE

Gate tunable enhancement of supercurrent in hybrid planar Josephson junctions

Planar Josephson junctions (JJs) have emerged as a promising platform for the realization of topological superconductivity and Majorana zero modes. To obtain robust quasi one-dimensional (1D) topological superconducting states using planar JJs, limiting the number of 1D Andreev bound states' subbands that can be present, and increasing the size of the topological superconducting gap, are two fundamental challenges. It has been suggested that both problems can be addressed by properly designing the interfaces between the JJ's normal region and the superconducting leads. We fabricated Josephson junctions with periodic hole structures on the superconducting contact leads on InAs heterostructures with epitaxial superconducting Al. By depleting the chemical potential inside the hole region with a top gate, we observed an enhancement of the supercurrent across the junction. Such an enhancement is reproduced in theoretical simulations. The theoretical analysis shows that the enhancement of the JJ's critical current is achieved when the hole depletion is such to optimize the matching of quasiparticles' wave-function at the normal/superconductor interface. These results show how the combination of carefully designed patterns for the Al coverage, and external gates, can be successfully used to tune the density and wave functions' profiles in the normal region of the JJ, and therefore open a new avenue to tune some of the critical properties, such as number of subbands and size of the topological gap, that must be optimized to obtain robust quasi 1D superconducting states supporting Majorana bound states.

cond-mat.mes-hall

PEMT: Multi-Task Correlation Guided Mixture-of-Experts Enables Parameter-Efficient Transfer Learning

Parameter-efficient fine-tuning (PEFT) has emerged as an effective method for adapting pre-trained language models to various tasks efficiently. Recently, there has been a growing interest in transferring knowledge from one or multiple tasks to the downstream target task to achieve performance improvements. However, current approaches typically either train adapters on individual tasks or distill shared knowledge from source tasks, failing to fully exploit task-specific knowledge and the correlation between source and target tasks. To overcome these limitations, we propose PEMT, a novel parameter-efficient fine-tuning framework based on multi-task transfer learning. PEMT extends the mixture-of-experts (MoE) framework to capture the transferable knowledge as a weighted combination of adapters trained on source tasks. These weights are determined by a gated unit, measuring the correlation between the target and each source task using task description prompt vectors. To fully exploit the task-specific knowledge, we also propose the Task Sparsity Loss to improve the sparsity of the gated unit. We conduct experiments on a broad range of tasks over 17 datasets. The experimental results demonstrate our PEMT yields stable improvements over full fine-tuning, and state-of-the-art PEFT and knowledge transferring methods on various tasks. The results highlight the effectiveness of our method which is capable of sufficiently exploiting the knowledge and correlation features across multiple tasks.

cs.CL

Calibration of Time-Series Forecasting: Detecting and Adapting Context-Driven Distribution Shift

Recent years have witnessed the success of introducing deep learning models to time series forecasting. From a data generation perspective, we illustrate that existing models are susceptible to distribution shifts driven by temporal contexts, whether observed or unobserved. Such context-driven distribution shift (CDS) introduces biases in predictions within specific contexts and poses challenges for conventional training paradigms. In this paper, we introduce a universal calibration methodology for the detection and adaptation of CDS with a trained model. To this end, we propose a novel CDS detector, termed the "residual-based CDS detector" or "Reconditionor", which quantifies the model's vulnerability to CDS by evaluating the mutual information between prediction residuals and their corresponding contexts. A high Reconditionor score indicates a severe susceptibility, thereby necessitating model adaptation. In this circumstance, we put forth a straightforward yet potent adapter framework for model calibration, termed the "sample-level contextualized adapter" or "SOLID". This framework involves the curation of a contextually similar dataset to the provided test sample and the subsequent fine-tuning of the model's prediction layer with a limited number of steps. Our theoretical analysis demonstrates that this adaptation strategy can achieve an optimal bias-variance trade-off. Notably, our proposed Reconditionor and SOLID are model-agnostic and readily adaptable to a wide range of models. Extensive experiments show that SOLID consistently enhances the performance of current forecasting models on real-world datasets, especially on cases with substantial CDS detected by the proposed Reconditionor, thus validating the effectiveness of the calibration approach.

cs.LG

Simulating Cosmological Evolution by Quantum Quench of an Atomic BEC

In cosmological evolution, it is the homogeneous scalar field (inflaton) that drives the universe to expand isotropically and to generate standard model particles. However, to simulate cosmology, atomic gas research has focused on the dynamics of Bose-Einstein condensates (BEC) with continuously applied forces. In this paper we argue a complementary approach needs also to be pursued; we, thus, consider the analogue BEC experiments in a non-driven, closed atomic system. We implement this using a BEC in an optical lattice which, after a quench, freely transitions from an unstable to a stable state. This dynamical evolution displays the counterpart "preheating", "reheating" and "thermalization" phases of cosmology. Importantly, our studies of these analogue processes yield tractable analytic models. Of great utility to the cold atom community, such understanding elucidates the dynamics of non-adiabatic condensate preparation.

cond-mat.quant-gas

Dynamical preparation of an atomic condensate in a Hofstadter band

The creation of a Hamiltonian in the quantum regime which has non-trivial topological features is a central goal of the cold-atom community, enabling widespread exploration of novel phases of quantum matter. A general scheme to synthesize such Hamiltonians is based on dynamical modulation of optical lattices which thereby generate vector potentials. At the same time the modulation can lead to heating and serious difficulties with equilibration. Here we show that these challenges can be overcome by demonstrating how a Hofstadter Bose-Einstein condensate (BEC) can be dynamically realized, using experimental protocols. From Gross-Pitaevskii simulations our study reveals a complex, multistage evolution; this includes a chaotic intermediate "heating" stage followed by a spontaneous reentrance to the BEC. The observed behavior is reminiscent of evolution in cosmological models.

cond-mat.quant-gas