arXiv ScienceSearch

arXiv subjects

Junchi Chen

Publications and source records attributed to Junchi Chen.

7 recordsLinked to original sources

Multi2AV-Safety: Benchmarking Safety in Multimodal-to-Audio-Video Generation

Audio-video generation is rapidly moving from prompt-driven synthesis toward multimodal conditioning, where text, images, audio, and video can jointly shape the generated output. This shift changes the nature of safety evaluation: harmful intent may no longer reside in any single input, but instead emerge from how otherwise benign or weakly harmful conditions interact across modalities and time. Existing safety benchmarks, however, remain largely prompt-centric or tied to fixed conditioning interfaces, leaving such compositional risks difficult to study systematically. To bridge this gap, we introduce Multi2AV-Safety, the first safety benchmark, to the best of our knowledge, to cover all 11 non-singleton T/I/A/V conditioning configurations for audio-video generation, comprising 11,024 attack instances. Evaluation on Multi2AV-Safety reveals systematic weaknesses in representative multimodal safety guards across attack mechanisms and harm-evidence structures. Our evaluation reveals two complementary failure modes: harmful semantics can emerge from the combination of individually benign inputs, while explicit harmful cues can become harder to detect when mixed with benign multimodal context. Together, these results identify \emph{compositional risk perception} as a central capability gap in safeguarding multimodal-conditioned audio-video generation: current safety guards fail to reliably integrate safety evidence across modalities and time, even when all conditioning inputs are observable. The dataset will be publicly released in October 2026.

cs.AI

DT-Guard: Intent-Driven Reasoning-Active Training for Reasoning-Free LLM Safety Guardrail

Large language models deployed in open-world applications require safety guardrails that are both robust to complex risks and efficient enough for low-latency runtime moderation. Existing guardrails face a practical trade-off between lightweight classification-based models, which are efficient but often struggle with concealed intent, ambiguous semantics, and borderline safety decisions, and reasoning-based guards, which improve judgment quality but introduce additional token generation and inference latency. We present DT-Guard, a content safety guardrail model based on a Reasoning-Active Training, Reasoning-Free Inference paradigm. The key idea is to use reasoning supervision during training while emitting only structured safety labels at inference time. DT-Guard formulates safety judgment as a progressive decision process, Intent - Category - Safety, and constructs an intent-driven dataset with intent labels, risk categories, safety labels, and structured reasoning trajectories. To further improve hard-case robustness, we propose Rollout-Guided Progressive Hard-Case Optimization (RG-PHO), which uses multi-rollout consistency to identify stably mastered, persistently failed, and preference-unstable samples, and applies targeted supervised and preference optimization accordingly. At inference time, DT-Guard directly generates structured labels without explicit reasoning traces, preserving deployment efficiency. Experiments on prompt-side and response-side safety benchmarks show that DT-Guard achieves average F1 scores of 0.886 and 0.870, respectively. With only a 4B backbone, it reaches a dual-side average F1 of 0.878, outperforming strong 8B guardrail baselines. These results demonstrate that reasoning supervision can be effectively internalized into low-latency safety discrimination.

cs.AI

SeClaw: Spec-Driven Security Task Synthesis for Evaluating Autonomous Agents

Autonomous LLM agents increasingly operate in stateful environments where they access tools, files, memory, and external services. While such capabilities enable complex real-world workflows, they also introduce security risks that are difficult to capture with existing evaluations. Current agent security benchmarks often rely on manually curated tasks, provide limited coverage of emerging threats, and focus primarily on final outcomes rather than the execution processes that lead to unsafe behavior. We introduce SeClaw, a framework that combines specification-driven security task synthesis with execution-based security evaluation for Autonomous agents. Spec-driven security task synthesis enables scalable and controllable construction of security tasks from structured risk specifications, while SeClaw docker provides a standardized testbed for evaluating agent behavior under diverse safety-risk scenarios. The benchmark covers risks arising from resources, user tasks, environments, and intrinsic agent behaviors, and supports trajectory-aware assessment of unsafe actions beyond final responses. By bridging systematic task synthesis and reproducible security evaluation, SeClaw provides a practical foundation for measuring, diagnosing, and comparing security failures in autonomous LLM agents. The code is available at https://github.com/seclaw-eval/seclaw-eval.

cs.CR

Equilibrium Thermochemistry and Crystallographic Morphology of Manganese Sulfide Nanocrystals

Manganese sulfide (MnS) is a p-type magnetic semiconductor whose physicochemical properties are sensitive to nanocrystal (NC) morphology, yet the thermodynamic driving forces governing morphology across MnS polymorphs remain poorly understood. Here, we use density functional theory (DFT) to predict the equilibrium morphologies of rock salt (RS), zinc blende (ZB), and wurtzite (WZ) MnS NCs as a function of the relative chemical potential of sulfur, $\Delta \mu_{S}$. Benchmarking against Heyd$\unicode{x2013}$Scuseria$\unicode{x2013}$Ernzerhof (HSE06) hybrid functional calculations reveals that the r$^2$SCAN meta-generalized gradient approximation reproduces experimental lattice constants and thermochemical reaction energies but underestimates S-terminated polar surface energies by up to a factor of five; applying a Hubbard $U$ correction (r$^2$SCAN+$U$, $U = 2.7$ eV) to the Mn 3d states brings the results into close agreement with HSE06. Using the validated r$^2$SCAN+$U$ framework with the Gibbs$\unicode{x2013}$Wulff theorem, we predict that RS-MnS NCs favor nanocubes across nearly the entire stability window, ZB-MnS NCs transform from rhombic dodecahedra (Mn-rich) to polyhedra with 16 triangular faces (S-rich), and WZ-MnS NCs adopt rod-like morphologies with $\Delta \mu_{S}$-sensitive base truncation. Synthesized RS-MnS NCs confirm the predicted cubic morphology, and high-temperature oxidative solution calorimetry yields an apparent surface energy of 1.15 $\pm$ 0.38 J$\cdot$m$^{-2}$, higher than the theoretical equilibrium value (0.42$\unicode{x2013}$0.43 J$\cdot$m$^{-2}$) due to high-index facet exposure, surface area uncertainty, and non-ideal surface configurations in real samples. This work establishes a framework for predicting the equilibrium morphologies of metal chalcogenide NCs.

cond-mat.mtrl-sci

GuardTrace-VL: Detecting Unsafe Multimodel Reasoning via Iterative Safety Supervision

Multimodal large reasoning models (MLRMs) are increasingly deployed for vision-language tasks that produce explicit intermediate rationales. However, reasoning traces can contain unsafe content even when the final answer is non-harmful, creating deployment risks. Existing multimodal safety guards primarily evaluate only the input question and the final answer, neglecting the intermediate reasoning process. This oversight allows undetected harm, such as biased inferences or policy-violating use of visual context, to emerge during reasoning. We introduce GuardTrace-VL, a vision-aware safety auditor that monitors the full Question-Thinking-Answer (QTA) pipeline via joint image-text analysis, enabling detection of unsafe content as it emerges in the reasoning stage. To support training and evaluation, we construct the GuardTrace dataset, which is generated through diverse prompting strategies and refined via a MLRM- and human-based voting and verification pipeline. Furthermore, we propose a three-stage progressive training scheme combined with the data refinement process, enabling the model to learn nuanced and context-dependent safety preferences according to different risk levels. On our proposed test set covering both in-domain and out-of-domain scenarios, GuardTrace-VL model achieves an F1 score of 93.1% on unsafe reasoning detection tasks, representing a 13.5% improvement in F1 score compared to the previous strongest multimodal safety defense methods. The codes will be made publicly available.

cs.CV

FreeBird.jl: An Extensible Toolbox for Simulating Interfacial Phase Equilibria

We present FreeBird, an extensible Julia-based platform for computational studies of phase equilibria at generic interfaces. The package supports a range of system configurations, from atomistic solid surfaces to coarse-grained lattice$-$gas models, with energies evaluated using classical interatomic potentials or lattice Hamiltonians. Both atomistic and lattice systems accommodate single- or multi-component mixtures with flexibly definable surface and lattice geometries. Implemented sampling algorithms include nested sampling, Wang$-$Landau sampling, Metropolis Monte Carlo, and, for tractable lattice systems, exact enumeration. Leveraging Julia's type hierarchies and multiple dispatch, FreeBird provides a modular interface that allows seamless integration of system definitions, energy evaluators, and sampling schemes. Designed for flexibility, extensibility, and performance, FreeBird offers a versatile framework for exploring the thermodynamics of interfacial phenomena.

cond-mat.stat-mech

High-quality femtosecond laser surface micro/nano-structuring assisted by a thin frost layer

Femtosecond laser ablation has been demonstrated to be a versatile tool to produce micro/nanoscale features with high precision and accuracy. However, the use of high laser fluence to increase the ablation efficiency usually results in unwanted effects, such as redeposition of debris, formation of recast layer and heat-affected zone in or around the ablation craters. Here we circumvent this limitation by exploiting a thin frost layer with a thickness of tens of microns, which can be directly formed by the condensation of water vapor from the air onto the exposed surface whose temperature is below the freezing point. When femtosecond laser beam is focused onto the target surface covered with a thin frost layer, only the local frost layer around the laser-irradiated spot melts into water, helping to boost ablation efficiency, suppress the recast layer and reduce the heat-affect zone, while the remaining frost layer can prevent ablation debris from adhering to the target surface. By this frost-assisted strategy, high-quality surface micro/nano-structures are successfully achieved on both plane and curved surfaces at high laser fluences, and the mechanism behind the formation of high-spatial-frequency (HSF) laser induced periodic surface structures (LIPSSs) on silicon is discussed.

physics.optics