arXiv ScienceSearch

arXiv subjects

Junhao Liu

Publications and source records attributed to Junhao Liu.

At least 19 recordsLinked to original sources

Turbulence Cascade in Cygnus X Revealed by Multi-point VDF Method

Turbulence plays a crucial role in regulating star formation activities within molecular clouds, yet few methods can directly reveal its properties and underlying processes. We use molecular line data from the Nobeyama 45m Cygnus X CO Survey to study the turbulence properties and their relationships with star-forming activities and/or other non-thermal motions. In this work, we apply the multi-point velocity dispersion function (VDF), rather than direct linewidth measurements, to investigate the non-thermal properties of molecular cloud motions. We filter out the large-scale ordered structure and isolate a relatively small-scale turbulence component. Through the Friends In Velocity (FIVe) algorithm, we identify 10 substructures of the clouds and derive the turbulent properties of each cloud using the VDF method. We find that both the cloud-complex regions and the 10 velocity substructures exhibit turbulence correlation lengths of $\sim 2$--5 pc. This plateau scale suggests a parsec-scale turbulence correlation or driving scale in Cygnus X. Below this scale, the rising VDFs trace the velocity scaling of the turbulent cascade, whereas larger-scale VDF variations likely reflect cloud-scale motions. The comparison between cloud complexes and substructures further suggests that, in observational data, the VDF may constrain the turbulence correlation scale more robustly than the turbulence velocity dispersion.

astro-ph.GA

Magnetic Fields in Massive Star-forming Regions (MagMaR). IX. Radiative Torque Alignment and Disruption in NGC6334I

Intense radiation from high-mass stars is expected to significantly affect dust grain alignment and evolution through RAdiative Torques (RATs). We investigate this effect in a massive star-forming region, NGC6334I, using 1.2 mm dust continuum polarization observations from the Atacama Large Millimeter/submillimeter Array. The polarization fraction spans from $\lesssim1\%$ to $\sim10\%$ and decreases with increasing column density, remaining below $2\%$ in dense cores despite high temperatures ($\sim100$ K), where efficient grain alignment by RATs is expected. We investigate how grain alignment, grain growth, grain disruption, B-field tangling, and local physical conditions affect the polarization properties of MM1, MM2, MM3, and their surroundings. Polarization angle dispersion shows that B-field tangling contributes to depolarization at moderate densities but cannot fully explain the lowest polarization fractions. Using RAT-based grain alignment and polarization modeling, we find that reduced alignment efficiency and high optical depth reproduce the low polarization in the densest regions. MM2 shows evidence of grain growth, with maximum grain sizes $a_{\max}\sim0.35-1.0~\mu$m, while MM1 exhibits smaller values of $\sim0.35-0.50~\mu$m. Accounting for optical depth increases the inferred grain sizes in MM1 to $\sim1.0-2.0~\mu$m. Analytical estimates of radiative torque disruption from the intense outburst suggest that micron-sized grains in high-temperature, moderate-density regions can fragment into submicron grains. Alternatively, high optical depth may also explain the low polarization in the densest regions even in the presence of micron-sized grains. Incorporating the B-field inclination effect indicates a transition from predominantly plane-of-sky fields at low densities to more line-of-sight-aligned configurations at high densities.

astro-ph.GA

Learning When to Trust via Selective Context Preference Optimization

Language models increasingly condition their answers on external signals, and a single misleading one can turn a correct answer wrong. The obvious remedy, training models to resist such signals, hides a failure mode: a model that ignores all context looks robust yet is useless when the context is worth trusting. We recast the problem as selective trust and introduce MIST, a human-annotated benchmark that renders each reasoning item under four matched conditions (clean, misleading, correct-context, and irrelevant-context), together with SC2W, a paired metric counting how often a misleading signal flips a clean-correct answer to wrong. Across a comprehensive benchmark study, we observe that such a susceptibility is universal. We then propose SCOPE, which mines clean-correct/misleading-wrong failures and optimizes a standard Direct Preference Optimization (DPO) objective over matched preference pairs balanced equally across all four conditions, rather than over misleading items alone. Our approach substantially reduces SC2W on popular open-sourced models while preserving accuracy when the added context is clean, correct, or irrelevant. With this work, we argue that models should be judged on selective trust, not on resistance alone.

cs.CL

DualAnchor: Preserving Language Priors and Improving Lexical Fidelity in Gloss-Free Sign Language Translation

Recent advances in large language models (LLMs) have led sign language translation (SLT), the task of converting sign-language videos into spoken-language text, to increasingly adopt LLMs as textual backbones. However, despite their strong language modeling capabilities, existing LLM-based SLT methods often undermine rather than exploit this language prior, producing disfluent translations, a failure we term language-prior degradation. Meanwhile, existing methods typically align videos and text at the sentence level, which does not ensure accurate lexical details and creates a lexical fidelity gap. To address both issues, we propose DualAnchor, a gloss-free LLM-based SLT training framework that couples two complementary anchors for linguistically fluent and visually faithful generation. Token-level Prior Anchoring (TPA) preserves the LLM's language prior by regularizing the multimodal decoder at each decoding step toward the next-token distribution of a frozen LLM conditioned on the same autoregressive prefix. Optimal Transport Alignment (OTA) improves lexical fidelity by formulating visual-textual matching as entropy-regularized partial optimal transport, with Sinkhorn optimization inducing a soft alignment between visual tokens and textual content tokens under a cosine cost. DualAnchor achieves strong overall performance on both PHOENIX-2014T and CSL-Daily. Targeted analyses attribute these gains to the complementary effects of the two anchors: TPA improves fluency, whereas OTA reduces fine-grained lexical errors.

cs.CL

Think, Plan, Paint: Layout-Aware Reasoning for Controllable Image Generation in Unified Models

Unified Multimodal Large Language Models (MLLMs) offer a promising paradigm for unifying visual understanding and generation, yet they still struggle to follow complex spatial instructions and logical constraints in controllable image generation. To address this gap, we present ATLAS, a unified framework that equips MLLMs with a human-like "Think, Plan, and Paint" paradigm. We adopt layout as the shared representation that connects the three stages, enabling the model to reason about spatial requirements, plan explicit object arrangements, and render the final image. We further improve plan-to-image fidelity with reinforcement-learning-based layout alignment. We instantiate ATLAS at 7B and 80B scales, achieving state-of-the-art performance among MLLMs on image generation benchmarks and an average 65.31% improvement over existing layout-based unified MLLMs. On spatially related tasks, ATLAS obtains an average 23.06% gain over the base models. Through the same layout interface, ATLAS also supports instruction-guided editing and multimodal grounding. We further introduce ATLAS-Reasoning, a benchmark for evaluating generation under complex spatial instructions.

cs.CV

ALMA observations of Magnetic Fields in the Massive Star-forming Region IRAS 18360-0537

Assessing the significance of magnetic fields in high-mass star formation remains one of the most challenging topics in astrophysics. In this study, we present full polarization observations obtained from the Atacama Large Millimeter/Submillimeter Array (ALMA) of the high-mass star-forming region IRAS18360-0537. The polarized dust emission at 1.3 mm reveals a clear hourglass-shaped morphology of the magnetic field. Interestingly, the magnetic field orientation is nearly perpendicular to both the outflow and core rotation axes, while it aligns with the elongation of the core. This orientation poses challenges for interpretation, particularly in light of the strong magnetic field strength estimated using the Davis-Chandrasekhar-Fermi method. Several scenarios provide insights into the underlying reasons for this magnetic field morphology. A clear velocity gradient seen in high-density tracing of molecular spectral lines indicates that the core is fast-rotating. The curved outskirts of the magnetic fields coincide with the outflow cavity, suggesting a possible influence from the outflow. The accretion flows along the core's elongation are also notable. Our study shows that the morphology of the magnetic field is probably highly influenced by the gas bulk motions.

astro-ph.GA

One Framework for All: Cross-Modal Membership Inference for Generative Models

Large generative models across text-to-text, text-to-image, and image-to-text modalities have been shown to pose significant privacy risks. One fundamental threat is membership inference attacks (MIA), which aim to determine whether a given data point was used in a model's training set. Although prior work has investigated MIAs against these three classes of generative models, existing approaches treat them in isolation and are not cross-applicable, thereby limiting their real-world utility. To address this limitation, we present the first comprehensive study of a unified membership inference framework that applies across text-to-text, text-to-image, and image-to-text modalities. Our approach is grounded in a key modality-agnostic observation: the output distribution of a generative model can approximate its training data distribution. Leveraging this property, we model the distributions of model-generated outputs and auxiliary non-member samples in a shared embedding space, and perform membership inference via likelihood ratio testing. We conduct extensive experiments in a strict black-box setting under both partial-knowledge and zero-knowledge threat models, and evaluate membership inference against both fine-tuning and pre-training data. Experimental results demonstrate our approach's superior performance in comparison to existing state-of-the-art methods, which are typically optimized for a single model class.

cs.LG

Fully coherent short wavelength free-electron laser driven by a single sub-microjoule seed

High-repetition-rate, fully coherent extreme-ultraviolet (EUV) and X-ray free-electron lasers (FELs) are essential for advanced time-resolved ultrafast spectroscopies. While external seeding serves as the standard technique to achieve precise temporal coherence, conventional methods demand hundred-megawatt peak-power laser systems. Furthermore, advanced configurations like echo-enabled harmonic generation (EEHG) introduce the severe complexities of dual-laser synchronization. Together, these requirements fundamentally restrict operations to kilohertz repetition rates and compromise overall system stability. Here, we experimentally demonstrate a fully coherent EEHG-FEL driven by a single, sub-microjoule seed laser. By employing a direct-amplification enabled harmonic generation technique, we utilize an initial 0.4 microJ (2 MW peak power) ultraviolet seed to directly drive coherent lasing at nanometer wavelengths. By eliminating the need for extreme peak powers and multiple synchronized lasers, this approach significantly simplifies the seeding architecture and provides a practical and robust pathway toward megahertz-class, fully coherent EUV and X-ray light sources.

physics.acc-ph

Guiding LLM-based Loop Invariant Synthesis via Feedback on Local Reasoning Errors

We propose a novel framework that provides constructive feedback to an LLM in the "guess-and-check" paradigm by formally verifying its own thinking process and detecting local reasoning errors. We apply this framework to the loop invariant synthesis problem. We prompt the model to produce a step-by-step natural language proof justifying its thinking process for the failed verification condition of its generated loop invariants. Then, we use an LLM to translate the reasoning steps into first-order logic implications, which can be checked automatically. An invalid implication pinpoints the exact logical flaw in the LLM's thinking process, which we then use to construct targeted feedback for refinement. We have implemented our approach in a tool called LORIS and evaluated it on a main benchmark suite of 460 C programs and an additional benchmark suite of 50 C programs each of which involves non-linear properties. On the main benchmark suite, LORIS solved 445 of the programs, and achieved an overall success rate of $93.1\%$. LORIS also demonstrates robustness on the challenging non-linear benchmark suite.

cs.PL

Safety Context Injection: Inference-Time Safety Alignment via Static Filtering and Agentic Analysis

Large Reasoning Models (LRMs) improve performance on complex tasks, but they also make safety control harder at deployment time. In black-box settings, defenders cannot modify model weights and must instead intervene at inference time. This setting creates three practical challenges: harmful intent may be hidden by educational or role-play framing, deep safety analysis can introduce non-trivial latency, and long adversarial contexts can dilute the local cues that simpler filters rely on. These challenges can expose an apparent thinking--output gap, where the model appears cautious during reasoning but still produces an unsafe final answer. To address this problem, we propose Safety Context Injection (SCI), an inference-time framework that separates safety assessment from task generation and prepends a structured external risk report as injected safety context for the protected model. The framework is instantiated in two complementary variants: Static Model Filtering (SMF), a lightweight one-pass guard for fast deployment, and Dynamic Agents Filtering (DAF), an agentic-loop-based analyzer that iteratively gathers and synthesizes evidence for ambiguous or long-context attacks. Across AdvBench and GPTFuzz, spanning base and reasoning models under five jailbreak families, both variants reduce attack success rate and toxicity in the evaluated settings. SMF offers an efficient low-latency option, while DAF is more effective when harmful intent is semantically disguised or dispersed across long contexts.

cs.CR

FL-Sailer: Efficient and Privacy-Preserving Federated Learning for Scalable Single-Cell Epigenetic Data Analysis via Adaptive Sampling

Single-cell ATAC-seq (scATAC-seq) enables high-resolution mapping of chromatin accessibility, yet privacy regulations and data size constraints hinder multi-institutional sharing. Federated learning (FL) offers a privacy-preserving alternative, but faces three fundamental barriers in scATAC-seq analysis: ultra-high dimensionality, extreme sparsity, and severe cross-institutional heterogeneity. We propose FL-Sailer, the first FL framework designed for scATAC-seq data. FL-Sailer integrates two key innovations: (i) adaptive leverage score sampling, which selects biologically interpretable features while reducing dimensionality by 80%, and (ii) an invariant VAE architecture, which disentangles biological signals from technical confounders via mutual information minimization. We provide a convergence guarantee, showing that FL-Sailer converges to an approximate solution of the original high-dimensional problem with bounded error. Extensive experiments on synthetic and real epigenomic datasets demonstrate that FL-Sailer not only enables previously infeasible multi-institutional collaborations but also surpasses centralized methods by leveraging adaptive sampling as an implicit regularizer to suppress technical noise. Our work establishes that federated learning, when tailored to domain-specific challenges, can become a superior paradigm for collaborative epigenomic research.

cs.LG

A programmable stellarator-tokamak hybrid for million-scale magnetic-configuration discovery

Tokamaks and stellarators are the leading magnetic-confinement concepts for fusion, but they rely on complementary design principles. Tokamaks use simple axisymmetric coils and plasma current, whereas stellarators use externally generated three-dimensional fields for steady-state operation. Here, we propose a programmable stellarator--tokamak hybrid that uses a fixed set of simple planar coils to access a broad magnetic-configuration space. The device adds 288 dipole-field coils to a tokamak-like coil set, with only six independent coil geometries required by symmetry. By programming coil currents, the same hardware generates more than 1.66 million optimized stellarator configurations spanning quasi-axisymmetry, quasi-helical symmetry, and quasi-isodynamicity, as well as tokamak-relevant three-dimensional perturbations. Representative configurations exhibit nested magnetic surfaces, low neoclassical transport, and favorable energetic-particle confinement. This approach enables rapid magnetic-configuration discovery without hardware redesign.

physics.plasm-ph

Phase-Stable Self-Modulation for GHz Continuous-Wave Ultrafast X-Ray Free-Electron Lasers

High-brightness femtosecond-to-attosecond pulses are indispensable for probing electron dynamics on their fundamental temporal scales. X-ray free-electron lasers (XFELs) at high repetition rates will facilitate high-statistics measurements and time-resolved studies that were previously inaccessible. Although energy recovery linacs (ERLs) are well suited for high-repetition-rate operation, their relatively low peak current poses a major challenge for generating intense ultrashort X-ray pulses. Here, we propose a completely laser-free scheme that fundamentally overcomes this bottleneck through a continuous, phase-stable self-modulation process. By interacting with its own coherently emitted terahertz radiation within a helical wiggler, the electron bunch naturally accumulates a robust, few-cycle energy modulation in its core, even when starting with the intrinsically low peak current typical of ERLs. A downstream dispersion chicane subsequently converts this energy modulation into an isolated, exceptionally sharp current spike. Start-to-end simulations based on a 1~GeV ERL light source demonstrate the feasibility of generating isolated soft X-ray pulses with an average peak power exceeding 4~GW and a pulse duration of about 1~fs at an unprecedented 1.3~GHz repetition rate. The proposed scheme offers a highly practical pathway for advancing ultrafast X-ray generation into the true continuous-wave regime, with transformative implications for the development of next-generation coherent light sources.

physics.acc-ph

Mott-Derived Local Moments and Kondo Hybridization in a d-electron Kagome lattice

Unlike canonical Kondo lattices in f-electron systems, where localized f orbitalsnaturally provide local moments, d-electron Kondo lattices require a distinct mechanism for local-moment formation. However, the study of d-electron Kondo lattices in bulk materials remains far from settled, particularly with regard to the microscopic origin of the local moments. Here, we report a microscopic mechanism for this process in the bilayer kagome metal CsCr6Sb6, where strong correlations drive a Mott splitting of the kagome flat band to supply the requisite local moments. By combining STM/STS and ARPES, we resolve a spectroscopic hierarchy between high-energy correlation effects and low temperature hybridization. Low-temperature STS reveals a robust asymmetric suppression of the density of states near EF that is well captured phenomenologically by a Fano-type lineshape, while ARPES detects a sharp quasiparticlepeak near EF. These low-energy signatures evolveon the same temperature scale and disappear upon warming, consistent with the onset of Kondo hybridization. At the same time, STS resolves symmetric humps at approximately +-50 mV and ARPES identifies a weakly dispersive feature around 50 meV below EF; unlike the near-EF hybridization signatures, these features persist to substantially higher temperatures. This separation of energy and temperature scales supports a two-stage picture in which a kagome flat band first undergoes correlation-driven splitting into lower and upper Hubbard bands, and the occupied lower Hubbard band supplies the local moments that later hybridize with itinerant electrons at lower temperature. Our results therefore move beyond the phenomenology of a kagome Kondo lattice candidate and instead provide a microscopic spectroscopic picture linking Mottness to Kondo hybridization in a frustrated d-electron system.

cond-mat.str-el

WASD: Locating Critical Neurons as Sufficient Conditions for Explaining and Controlling LLM Behavior

Precise behavioral control of large language models (LLMs) is critical for complex applications. However, existing methods often incur high training costs, lack natural language controllability, or compromise semantic coherence. To bridge this gap, we propose WASD (unWeaving Actionable Sufficient Directives), a novel framework that explains model behavior by identifying sufficient neural conditions for token generation. Our method represents candidate conditions as neuron-activation predicates and iteratively searches for a minimal set that guarantees the current output under input perturbations. Experiments on SST-2 and CounterFact with the Gemma-2-2B model demonstrate that our approach produces explanations that are more stable, accurate, and concise than conventional attribution graphs. Moreover, through a case study on controlling cross-lingual output generation, we validated the practical effectiveness of WASD in controlling model behavior.

cs.CL

The dominance of turbulence over magnetism in the formation of massive star cluster seeds

High-mass stars form in protoclusters, where gravo-magnetic processes shape collapsing clouds and clumps to be elongated preferentially perpendicular to magnetic (B) fields. Yet it remains unclear whether gravo-magnetic processes still govern the formation of smaller-scale condensations in massive-star-forming protoclusters, which are crucial for understanding the stellar initial mass function and multiplicity. Here we report the first statistical evidence that the condensation elongations are preferentially aligned with local B fields, based on high-resolution data from the largest dust polarization survey toward 30 massive star-forming regions with the Atacama Large Millimeter/submillimeter Array (ALMA). Our clustered massive star formation simulations reveal that this more parallel alignment is exclusively observed in models where initial turbulence dominates B fields. In contrast, models with initial B fields dominating turbulence distinctly exhibit a more perpendicular alignment. The comparison between observations and simulations suggests that turbulence could play a more important role than B fields in the formation of condensations in the context of clustered massive star formation, contradicting the prediction of classical magnetically regulated models. Moreover, we find a possibly turbulence-induced preferential misalignment between the B field and rotation axis of condensations, which may potentially reduce the magnetic braking efficiency and facilitate the formation of large protostellar disks. Our findings indicate that turbulence could be critical in determining the initial stellar properties.

astro-ph.GA

SC-Arena: A Natural Language Benchmark for Single-Cell Reasoning with Knowledge-Augmented Evaluation

Large language models (LLMs) are increasingly applied in scientific research, offering new capabilities for knowledge discovery and reasoning. In single-cell biology, however, evaluation practices for both general and specialized LLMs remain inadequate: existing benchmarks are fragmented across tasks, adopt formats such as multiple-choice classification that diverge from real-world usage, and rely on metrics lacking interpretability and biological grounding. We present SC-ARENA, a natural language evaluation framework tailored to single-cell foundation models. SC-ARENA formalizes a virtual cell abstraction that unifies evaluation targets by representing both intrinsic attributes and gene-level interactions. Within this paradigm, we define five natural language tasks (cell type annotation, captioning, generation, perturbation prediction, and scientific QA) that probe core reasoning capabilities in cellular biology. To overcome the limitations of brittle string-matching metrics, we introduce knowledge-augmented evaluation, which incorporates external ontologies, marker databases, and scientific literature to support biologically faithful and interpretable judgments. Experiments and analysis across both general-purpose and domain-specialized LLMs demonstrate that (i) under the Virtual Cell unified evaluation paradigm, current models achieve uneven performance on biologically complex tasks, particularly those demanding mechanistic or causal understanding; and (ii) our knowledge-augmented evaluation framework ensures biological correctness, provides interpretable, evidence-grounded rationales, and achieves high discriminative capacity, overcoming the brittleness and opacity of conventional metrics. SC-Arena thus provides a unified and interpretable framework for assessing LLMs in single-cell biology, pointing toward the development of biology-aligned, generalizable foundation models.

cs.AI

Focus-LIME: Surgical Interpretation of Long-Context Large Language Models via Proxy-Based Neighborhood Selection

As Large Language Models (LLMs) scale to handle massive context windows, achieving surgical feature-level interpretation is essential for high-stakes tasks like legal auditing and code debugging. However, existing local model-agnostic explanation methods face a critical dilemma in these scenarios: feature-based methods suffer from attribution dilution due to high feature dimensionality, thus failing to provide faithful explanations. In this paper, we propose Focus-LIME, a coarse-to-fine framework designed to restore the tractability of surgical interpretation. Focus-LIME utilizes a proxy model to curate the perturbation neighborhood, allowing the target model to perform fine-grained attribution exclusively within the optimized context. Empirical evaluations on long-context benchmarks demonstrate that our method makes surgical explanations practicable and provides faithful explanations to users.

cs.CL