arXiv ScienceSearch

arXiv subjects

Liyan Liu

Publications and source records attributed to Liyan Liu.

7 recordsLinked to original sources

DeepVoyager-VL: Incentivizing Vision-in-the-Loop Search for Long-Horizon Multimodal Agents

Multimodal large language models (MLLMs) have advanced visual understanding and reasoning, yet their static parametric knowledge limits their ability to address knowledge-intensive and dynamically evolving open-world problems. To move beyond this limitation, multimodal deep search has emerged as a key direction for open-world information access, evolving from single-turn factual retrieval toward long-horizon, multi-turn search guided by visual evidence. However, existing methods typically confine vision to the input or answer stage, overlooking its role in intermediate reasoning, and lack designs tailored to long-horizon interaction. Consequently, visual evidence rarely drives continued retrieval, constraining both interaction depth and reasoning span. To address these limitations, we propose DeepVoyager-VL, a long-horizon multimodal deep-search framework for vision-in-the-loop search. Specifically, we construct a multimodal event graph to drive data synthesis, yielding problems with intermediate visual dependencies and long reasoning chains. We then design an agent framework for active visual acquisition and on-demand image loading. Finally, we fine-tune models on the synthesized data without reinforcement learning. Extensive experiments across ten multimodal search benchmarks demonstrate the effectiveness of our method.

cs.CV

AgentOmnia: Scaling Agentic Models for Full-Scenario Applications

Large language model agents have advanced rapidly, yet progress remains fragmented across domains, capabilities, task difficulty, and interaction settings. We frame this as full-scenario agentic scaling and present AgentOmnia, a framework coordinating task-space definition, data synthesis, post-training, evaluation, and improvement across To-Consumer (ToC), To-Business (ToB), and To-Employee (ToE) applications. An extensible Domain x Capability x Atomic Difficulty taxonomy aligns these stages and enables fine-grained diagnosis with OmniaBench. AgentOmnia combines bidirectional environment-task synthesis with tool-dependency, program-structured, and solver-based pipelines, constructing 5,018 stateful environments with 255,375 tools and 52,361 tasks. Programs, solvers, and verifiers provide correctness signals, while supervised fine-tuning, online agentic reinforcement learning, and a rollback curriculum support post-training. Evaluation failures translate into Product Requirement Documents (PRDs) for targeted self-evolution. Starting from Qwen3-30B-A3B-Thinking-2507, AgentOmnia raises the pass rate on the OmniaBench challenging subset from 9.16% to 37.11% and the macro-average across OmniaBench, $\tau^2$-Bench, DeepPlanning, and VitaBench from 22.86% to 41.69%. Under a unified protocol,it leads the evaluated agentic post-trained baselines on OmniaBench and retains the highest four-benchmark macro-average. It also surpasses Qwen3-235B-A22B-Thinking-2507 on all four benchmarks and exceeds Qwen3.5-35B-A3B on the macro-average. Gains span three application splits, ten capability dimensions, eight atomic-difficulty factors, and 76 of 90 level-1 domains, indicating broad rather than category-specific improvement. A one-round study provides initial evidence for PRD-guided self-evolution, motivating validation at larger scales and in industrial settings.

cs.AI

SearchArt: Training Long-Horizon Search Agent with Scalable Synthetic and Verified Task

Recent advances in large language models (LLMs) have enabled search agents to autonomously tackle complex tasks across extended search and reasoning horizons. However, training effective search agents remains challenging due to the lack of scalable and long-horizon tasks, and the difficulty of evaluating and correcting intermediate reasoning and tool-use behaviors. We introduce SearchArt, a scalable framework for training long-horizon search agents through verification-driven task synthesis and a multi-stage post-training pipeline. SearchArt constructs large-scale datasets for complex search-, research- and user-oriented tasks by synthesizing diverse information-seeking QA pairs and corresponding search trajectories from web documents and automatically generated evidence graphs. To ensure the reliability of the synthesized data, we design a verification pipeline that jointly evaluates QA consistency, trajectory quality, and the relevance of retrieved evidence. The verified trajectories are subsequently used in a multi-stage training process comprising supervised fine-tuning and reinforcement learning-based policy optimization. Search agents trained with SearchArt exhibit adaptive search planning, iterative evidence aggregation, and complex reasoning over extended interaction horizons. Experimental results demonstrate that, with only (Qwen3.5-) 27B parameters, SearchArt scores 74.39 on BrowseComp-ZH, 70.06 on BrowseComp, and 52.55 on Deepresearch-bench, matching or surpassing frontier closed-source agents on both deepsearch and deepresearch benchmarks.

cs.IR

A nonextensive approach for the instability of current-driven ion-acoustic waves in space plasma

The instability of current-driven ion-acoustic waves in the collisionless magnetic-field-free space plasma is investigated by using a nonextensive approach. The ions and the electrons are thought of in the power-law distributions that can be described by the generalized q-Maxwellian velocity distribution and are considered with the different nonextensive q-parameters. The generalized q-wave frequency and the generalized instability q-growth rate for the ion-acoustic waves are derived. The numerical results show that the nonextensive effects on the ion-acoustic waves are not apparent when the electron temperature is much more than the ion temperature, but they are salient when the electron temperature is not much more than the ion temperature. As compared with the electrons, the ions play a dominant role in the nonextensive effects.

physics.plasm-ph

Ion acoustic waves in the plasma with the power-law q-distribution in nonextensive statistics

We investigate the dispersion relation and Landau damping of ion acoustic waves in the collisionless magnetic-field-free plasma if it is described by the nonextensive q-distributions of Tsallis statistics. We show that the increased numbers of superthermal particles and low velocity particles can explain the strengthened and weakened modes of Landau damping, respectively, with the q-distribution. When the ion temperature is equal to the electron temperature, the weakly damped waves are found to be the distributions with small values of q.

physics.plasm-ph

Energy fluctuations and the ensemble equivalence in Tsallis statistics

We investigate the general property of the energy fluctuation for the canonical ensemble in Tsallis statistics and the ensemble equivalence. By taking the ideal gas and the non-interacting harmonic oscillators as examples, we show that, when the particle number N is large enough, the relative fluctuation of the energy is proportional to 1/N in the new statistics, instead of square root of 1/N in Boltzmann-Gibbs statistics. Thus the equivalence between the microcanonical and the canonical ensemble still holds in Tsallis statistics.

cond-mat.stat-mech

Stability analysis of the classical ideal gas in nonextensive statistics and the negative specific heat

We present a stability analysis of the classical ideal gas in a new theory of nonextensive statistics and use the theory to understand the phenomena of negative specific heat in some self-gravitating systems. The stability analysis is made on the basis of the second variation of Tsallis entropy. It is shown that the system is thermodynamically unstable if the nonextensive parameter is q>5/3, which is exactly equivalent to the condition of appearance of the negative specific heat.

cond-mat.stat-mech