arXiv ScienceSearch

arXiv subjects

Qing Zhang

Publications and source records attributed to Qing Zhang.

At least 19 recordsLinked to original sources

Beyond "Made with AI": Visualizing Provenance Density to Mitigate the Transparency Penalty

As generative AI makes polished prose cheap to produce, users can no longer rely on fluency as a proxy for truth. We call this failure mode the Fluency Trap: users trust fluent hallucinations while also discounting accurate content once it is disclosed as AI-generated. Binary ``Made with AI'' labels respond with authorship disclosure, but they do not show what supports a claim. We propose Provenance Density, an evidence-visualization interface that shows the density of verified claims in a text. In a user study with 81 participants, an idealized Provenance Density interface produced a large discernment gap between truth and fabrication ($+4.15$ points, $d=1.82$), whereas participants given no signal showed no detectable discrimination. A technical audit with 200 samples shows that retrieval density alone is insufficient; unexpectedly, the Consistency Veto carries most of the discriminative signal on dynamic queries. As AI-generated content becomes indistinguishable from human writing, effective transparency must move from authorship disclosure toward evidence visualization.

cs.AI

TrapVLA: Trapping Vision-Language-Action Models in Configured Failure Modes

This work introduces Configured Failure Trapping, a novel backdoor attack task against Vision-Language-Action (VLA) models, which aims to activate attacks through stealthy textual triggers and induce configured failure modes. Unlike prior backdoor attacks that treat any task failure as a successful attack, Configured Failure Trapping requires the attacker to control how the robot fails (e.g., causing the robot to grasp with a specified positional offset), making it substantially more challenging and hard to detect. To support the new task, we propose an effective data engine for synthesizing high-quality target trajectories and an automated suite for measuring configured-failure fidelity. Then, based on this foundation, we construct two new benchmarks, namely Trap-LIBERO and Trap-RoboTwin, that instantiate Configured Failure Trapping across four representative failure modes. To address this task, we identify sparse action deviation as a critical challenge and accordingly propose a novel method named TrapVLA, which explicitly learns trigger-induced action residuals to steer the policy toward the configured failure behavior. Extensive experiments across simulation benchmarks and real-world robotic settings show that TrapVLA effectively injects configured failure modes into VLA models while largely preserving performance on clean data. Project page: https://john-liua.github.io/TrapVLA/

cs.RO

Spin-group theory on Edelstein effect and spin-orbit torque in Collinear Ferromagnets

Current-induced spin-orbit torques (SOTs) are central to the electrical manipulation of magnetic order in spintronic devices. In transition-metal/collinear ferromagnet bilayers, field-like and damping-like torques have been described only phenomenologically via the spin or orbital Hall effect, lacking a rigorous symmetry-based foundation. The precise role of spin-orbit coupling (SOC) in both the Edelstein effect and SOTs has remained unresolved. Here we develop a spin-group symmetry theory for the Edelstein effect and SOTs in collinear ferromagnets, treating SOC as a symmetry-breaking perturbation. For 4mm (C4v) point group symmetry, we derive the full forms of field-like and damping-like torques, which arise predominantly from first- and second-order SOC. We further show that SOTs in both orbital-Hall-dominated Ti/Ni and spin-Hall-dominated Pt/CoFe bilayers originate at first-order SOC. Taking the 3m (C3v) torque as a paradigmatic example, we elucidate the role of second- and higher-order SOC torques in field-free switching of perpendicular magnetic anisotropy. Remarkably, in PtMnSb, we demonstrate that SOTs under certain point group symmetries deviate from the conventional form: zeroth- and first-order SOC contributions vanish identically, with the leading SOT emerging at second order. All symmetry-based predictions from spin-group theory are in excellent quantitative agreement with first-principles calculations. Our work establishes a unified symmetry framework for the microscopic understanding of the Edelstein effect and current-induced spin torques in ferromagnetic systems.

cond-mat.mtrl-sci

Near-Field Velocity Estimation and Doppler-Aware Localization in OFDM Massive MIMO

In Orthogonal Frequency Division Multiplexing (OFDM)-based massive Multiple-Input Multiple-Output (MIMO) near-field (NF) sensing, target motion induces an antenna-dependent bistatic Doppler variation across the array aperture. Ignoring this spatial Doppler variation leads to a model mismatch that degrades NF localization. In this paper, we propose a low-complexity recursive framework for joint radial/transverse velocity estimation and Doppler-aware localization. Initialized by a constant-Doppler coarse localization, the method alternates between closed-form Least Squares Estimator (LSE)-based velocity estimation and antenna-dependent Doppler-aware localization refinement. Simulation and measurement results demonstrate the effectiveness of the proposed framework against two benchmark methods. Compared with a low-complexity constant-Doppler baseline method, the proposed algorithm improves range, angle, and radial velocity estimation results, while also enabling transverse velocity estimation. In the measurement results, the overall localization error decreases from 0.268 m to 0.064 m. The radial and transverse velocity estimation errors are 0.032 m/s and 0.069 m/s, respectively. Compared with a high-complexity exhaustive four-dimensional (4D) Maximum Likelihood Estimator (MLE), the proposed method achieves comparable velocity estimation results while yielding a more accurate localization result when the 4D MLE has a practical finite search grid.

eess.SP

Towards Reliable Stain Transfer: An Iterative Data-Model Co-Optimization Framework Based on Multimodal Expert-Guided Assessment

Histopathological examination primarily relies on hematoxylin and eosin (H&E) and immunohistochemistry (IHC) staining. Although IHC provides critical molecular information, it is costly and requires specialized expertise. Stain transfer provides an efficient alternative by computationally generating IHC from H&E images, but remains challenged by unified and interpretable modeling for heterogeneous biomarkers under pixel-unaligned supervision. We propose DMCoStain, a novel Data-Model Co-optimization framework for Stain transfer. It iteratively co-refines training data and model capability, improving staining accuracy and interpretability in both pathological and structural consistency. To refine training data in a clinically meaningful manner, it incorporates the Multimodal Expert-Guided Finer Selection (MEGFS) strategy, built upon a pioneering IHC-positive-expression (IPE) vision-language model (VLM) that emulates pathologist reasoning. To support MEGFS, we construct ImmunoInstruction, the first large-scale IPE instruction-following dataset with 150K VQA samples. Extensive experiments on multiple tissues and biomarkers demonstrate that DMCoStain achieves state-of-the-art (SOTA) accuracy. This paradigm offers strong practical value, and MEGFS also functions as a specialized evaluation tool for future model development. Dataset, code, and more details are in https://github.com/SikangSHU/DMCoStain.

cs.CV

Spin Hall Effect in Collinear Ferromagnets from Spin-Group Symmetry

Magnetic materials support both time-reversal-even (T-even) and time-reversal-odd (T-odd) spin Hall currents, yet their underlying microscopic origins remain elusive. Here, we elucidate the spin Hall effect (SHE) in collinear ferromagnets by treating spin-orbit coupling (SOC) as a perturbation that breaks spin-group symmetry, thereby revealing how magnetic order activates distinct spin Hall response. To first order in SOC, we identify two dominant T-even SHE mechanisms: a magnetization-independent conventional contribution and a magnetization-dependent channel associated with anomalous Hall charge transport. At the same order, the leading T-odd magnetic spin Hall effect (MSHE) originates from the exchange interaction between the conventional spin current and the local magnetization. At second order in SOC, we further uncover a distinct T-odd planar spin Hall mechanism. Our spin-symmetry analysis is corroborated by first-principles calculations, which reveal a pronounced anisotropic magnetic spin Hall effect whose magnitude can be comparable to the T-even spin Hall conductivity (SHC) when the magnetic moment is tilted away from the principal crystallographic axes. These findings clarify the microscopic origins of the SHC in collinear ferromagnets and pave the way for ferromagnet-based spin current sources with versatile properties in spintronic applications.

cond-mat.mes-hall

Dzyaloshinskii-Moriya Gradients Unlock Topological Dimensional Reduction in Magnetic Hopfions

Three-dimensional magnetic solitons retain topological protection only while their spin field remains continuous. Here we show that chemical inhomogeneity can break this protection in a controlled way, converting a hopfion-like toroidal texture into an effectively two-dimensional skyrmion string. Tilt-dependent Lorentz transmission electron microscopy, electron energy-loss spectroscopy, and micromagnetic simulations of pristine, uniformly Y-doped, and nonuniformly Y-doped disordered TiO2 nanoparticles embedded in a FeCrNiMn host reveal two regimes. Uniform Y doping enriches Ti3+/oxygen-vacancy localization centers and stabilizes closed rings with Hopf invariant QH approximately 1. Nonuniform Y doping forms a Ti4+-rich, vacancy-depleted boundary that creates a sharp q=D/(2A) gradient and a weak-moment leakage channel. This coupled mismatch torque and continuity leakage split the ring, leaving a skyrmion-string remnant and a field-sensitive helicity texture.

cond-mat.other

A Longitudinal Study of Android Apps Signing Key Protection

Android app signing relies on developer-managed credentials, making secure key protection essential for the integrity of the software supply chain. A recent platform key leakage incident involving two major OEM manufacturers demonstrates that even robustly designed signing mechanisms can be compromised due to developers' oversight. In this work, we conduct a longitudinal ecosystem study to characterize this threat by mining public repositories for Android signing credentials, recovering compromised keys via exposed passwords, and matching them against signatures from over 4,000 apps collected from major stores and OEM system images. Our analysis identifies 5,673 compromised keystores on GitHub and 26 unique certificates linked to 278 real-world apps. These include 26 third-party apps in public app stores and 252 preinstalled apps from seven manufacturers, collectively affecting over 10 billion users. We demonstrate the practical exploitability of these leaks through a proof-of-concept app replacement attack and identify spillover risks in non-smartphone platforms, including a popular automotive head-unit platform installed in over 1,100 vehicle models. Our results reveal that signing-key mismanagement is a systemic risk, underscoring the need for a more rigorous key-management support in Android release engineering and distribution infrastructures.

cs.CR

Structural Energy Guidance for View-Consistent Text-to-3D Generation

Text-to-3D generation based on diffusion models often suffers from the Janus problem, leading to inconsistent geometry across viewpoints. This work identifies viewpoint bias in 2D diffusion priors as the main cause and proposes Structural Energy-Guided Sampling (SEGS), a training-free and plug-and-play framework to improve multi-view consistency. SEGS constructs a structural energy in the PCA subspace of U-Net features and injects its gradient into the denoising process. It can be easily integrated into SDS/VSD pipelines without retraining. Experiments show that SEGS reduces the Janus Rate by about 10% on average and improves View-CS scores across multiple baselines, including DreamFusion, Magic3D, and LucidDreamer. This method effectively alleviates viewpoint artifacts while preserving appearance fidelity, providing a flexible solution for high-quality text-to-3D content generation.

cs.CV

Learning Action Manifold with Multi-view Latent Priors for Robotic Manipulation

This paper tackles spatial perception and manipulation challenges in Vision-Language-Action (VLA) models. To address depth ambiguity from monocular input, we leverage a pre-trained multi-view diffusion model to synthesize latent novel views and propose a Geometry-Guided Gated Transformer (G3T) that aligns multi-view features under 3D geometric guidance while adaptively filtering occlusion noise. To improve action learning efficiency, we introduce Action Manifold Learning (AML), which directly predicts actions on the valid action manifold, bypassing inefficient regression of unstructured targets like noise or velocity. Experiments on LIBERO, RoboTwin 2.0, and real-robot tasks show our method achieves superior success rate and robustness over SOTA baselines. Project page: https://junjxiao.github.io/Multi-view-VLA.github.io/.

cs.RO

RoboBlockly Studio: Conversational Block Programming with Embodied Robot Feedback for Computational Thinking

Computational thinking (CT) is increasingly promoted as a core literacy, yet learners and teachers face challenges in connecting abstract program logic to meaningful outcomes. We design and evaluate RoboBlockly Studio, an integrated interactive system that combines block-based programming, a conversational AI teaching agent, and embodied robot execution. RoboBlockly Studio creates a tight iterative loop of authoring, running, observing, and revising. Informed by interviews with five programming teachers, the system was designed to support four goals: (1) preserving learner agency in computational thinking, (2) making program behavior transparent and interpretable, (3) grounding programming in embodied, classroom-aligned tasks, and (4) scaffolding reflection through pedagogically grounded AI dialogue. We deployed RoboBlockly Studio with 32 high school students, observing how robot and AI feedback influenced students' interactions with code, reflections on problem-solving strategies, and understanding of CT concepts. We discuss design insights and implications for creating interactive, embodied learning environments that integrate AI and robotics to support CT learning in computing education.

cs.HC

Reconfigurable ultrafast perovskite polariton logic gates via nonlinear dynamics

Exciton-polaritons provide a great platform for developing ultrafast all-optical logic gates for quantum and optical chips. However, progress toward practical polariton logic remains limited due to incomplete logical functionality on a single device. Herein, we present a single-device perovskite polariton platform enabling reconfigurable, ultrafast logic gates with functional completeness. The device consists of an optically trapped perovskite microwire, generating well-controlled non-equilibrium polariton condensation states for multiple logic operation channels. By tailoring the power of signal and gate beams, the same device is programmed to execute three basic Boolean functions (AND,OR,and NOT) and a high-order XOR function with a high on/off ratio of 21 dB, and a fast response time 6.7 ps. The reconfigurability arises from the selective activation of different nonlinear responses of polariton condensates, including amplification, seeding state transitions, and nonlinear interaction. These results provide valuable insights for advancing exciton-polariton logic gates.

physics.optics

Generative Texture Filtering

We present a generative method for texture filtering, which exhibits surprisingly good performance and generalizability. Our core idea is to empower texture filtering by taking full advantage of the strong learned image prior of pre-trained generative models. To this end, we propose to fine-tune a pre-trained generative model via a two-stage strategy. Specifically, we first conduct supervised fine-tuning on a very small set of paired images, and then perform reinforcement fine-tuning on a large-scale unlabeled dataset under the guidance of a reward function that quantifies the quality of texture removal and structure preservation. Extensive experiments show that our method clearly outperforms previous methods, and is effective to deal with previously challenging cases. Our code is available at https://github.com/OnlyZZZZ/Generative_Texture_Filtering.

cs.CV

Observation of field-odd and field-free superconducting diode effects in $\mathrm{Mo}_2\mathrm{C}$ nanoflakes

The superconducting diode effect (SDE) enables nonreciprocal supercurrent flow, holding immense potential for ultra-low-power quantum electronics. Intrinsic SDE typically requires materials with inherent symmetry breakings. Here, we report the discovery of SDE in chemical vapor deposition-grown molybdenum carbide ($\mathrm{Mo}_2\mathrm{C}$) nanoflakes, a material traditionally considered centrosymmetric. Strikingly, this system uniquely hosts both field-odd and field-free SDEs. Transport measurements reveal a field-odd SDE with tunable efficiency exceeding 40% at 4 K under a perpendicular in-plane magnetic field. In a separate sample, a robust field-free SDE persists under zero-field and field-coolings. Out-of-plane field sweeps confirm the intrinsic nature of these phenomena. We propose that domain-boundary supercurrents or charge density wave-like orders drive this unexpected combination of symmetry breakings. Our findings establish air-stable $\mathrm{Mo}_2\mathrm{C}$ as an ideal platform for nonreciprocal superconducting electronics operating at liquid-helium temperatures, expanding the search for SDE into nominally centrosymmetric superconductors.

cond-mat.supr-con

SGS-Intrinsic: Semantic-Invariant Gaussian Splatting for Sparse-View Indoor Inverse Rendering

We present SGS-Intrinsic, an indoor inverse rendering framework that works well for sparse-view images. Unlike existing 3D Gaussian Splatting (3DGS) based methods that focus on object-centric reconstruction and fail to work under sparse view settings, our method allows to achieve high-quality geometry reconstruction and accurate disentanglement of material and illumination. The core idea is to construct a dense and geometry-consistent Gaussian semantic field guided by semantic and geometric priors, providing a reliable foundation for subsequent inverse rendering. Building upon this, we perform material-illumination disentanglement by combining a hybrid illumination model and material prior to effectively capture illumination-material interactions. To mitigate the impact of cast shadows and enhance the robustness of material recovery, we introduce illumination-invariant material constraint together with a deshadowing model. Extensive experiments on benchmark datasets show that our method consistently improves both reconstruction fidelity and inverse rendering quality over existing 3DGS-based inverse rendering approaches. Our code is available at https://github.com/GrumpySloths/SGS_Intrinsic.github.io.

cs.CV

You Only Erase Once: Erasing Anything without Bringing Unexpected Content

We present YOEO, an approach for object erasure. Unlike recent diffusion-based methods which struggle to erase target objects without generating unexpected content within the masked regions due to lack of sufficient paired training data and explicit constraint on content generation, our method allows to produce high-quality object erasure results free of unwanted objects or artifacts while faithfully preserving the overall context coherence to the surrounding content. We achieve this goal by training an object erasure diffusion model on unpaired data containing only large-scale real-world images, under the supervision of a sundries detector and a context coherence loss that are built upon an entity segmentation model. To enable more efficient training and inference, a diffusion distillation strategy is employed to train for a few-step erasure diffusion model. Extensive experiments show that our method outperforms the state-of-the-art object erasure methods. Code will be available at https://zyxunh.github.io/YOEO-ProjectPage/.

cs.CV

Learning Explicit Continuous Motion Representation for Dynamic Gaussian Splatting from Monocular Videos

We present an approach for high-quality dynamic Gaussian Splatting from monocular videos. To this end, we in this work go one step further beyond previous methods to explicitly model continuous position and orientation deformation of dynamic Gaussians, using an SE(3) B-spline motion bases with a compact set of control points. To improve computational efficiency while enhancing the ability to model complex motions, an adaptive control mechanism is devised to dynamically adjust the number of motion bases and control points. Besides, we develop a soft segment reconstruction strategy to mitigate long-interval motion interference, and employ a multi-view diffusion model to provide multi-view cues for avoiding overfitting to training views. Extensive experiments demonstrate that our method outperforms state-of-the-art methods in novel view synthesis. Our code is available at https://github.com/hhhddddddd/se3bsplinegs.

cs.CV

ARIADNE: A Perception-Reasoning Synergy Framework for Trustworthy Coronary Angiography Analysis

Conventional pixel-wise loss functions fail to enforce topological constraints in coronary vessel segmentation, producing fragmented vascular trees despite high pixel-level accuracy. We present ARIADNE, a two-stage framework coupling preference-aligned perception with RL-based diagnostic reasoning for topologically coherent stenosis detection. The perception module employs DPO to fine-tune the Sa2VA vision-language foundation model using Betti number constraints as preference signals, aligning the policy toward geometrically complete vessel structures rather than pixel-wise overlap metrics. The reasoning module formulates stenosis localization as a Markov Decision Process with an explicit rejection mechanism that autonomously defers ambiguous anatomical candidates such as bifurcations and vessel crossings, shifting from coverage maximization to reliability optimization. On 1,400 clinical angiograms, ARIADNE achieves state-of-the-art centerline Dice of 0.838, reduces false positives by 41% compared to geometric baselines. External validation on multi-center benchmarks ARCADE and XCAD confirms generalization across acquisition protocols. This represents the first application of DPO for topological alignment in medical imaging, demonstrating that preference-based learning over structural constraints mitigates topological violations while maintaining diagnostic sensitivity in interventional cardiology workflows.

cs.CV