arXiv ScienceSearch

arXiv subjects

Junwei Yu

Publications and source records attributed to Junwei Yu.

At least 19 recordsLinked to original sources

Open-Ended CT Volume Segmentation with Weak Supervision from Language

We introduce a method for training a text-conditioned segmentation model for CT scans, which combines voxel-level supervision with coarse but scalable slice-level supervision from reports. We extract, from a large database of scan-report pairs, descriptions of findings with indices of slices where those findings occur. We then finetune a general-purpose 2D image segmentation model, SAM3, with standard segmentation losses from strongly labeled data and with a slice-level classification loss from the extracted weak supervision. Our results on the ReXGroundingCT dataset illustrate that this strategy improves the segmentation dice score: from an 8% relative gain when there are 1000 fully labeled volumes to 22% when there are 250 fully labeled volumes.

cs.CV

On the emergence of dead cores in elliptic systems with sublinear competitive interactions

In this paper, we study semilinear elliptic systems with sublinear coupling terms, showing that, under competition-type interactions, solutions typically have dead cores, i.e., they vanish on certain open subsets of the domain. We apply our results to a large class of solutions treated in the literature, for instance to ground states and least energy sign-changing solutions, under Dirichlet boundary conditions or in the whole space. The proofs are based on a general result stating that subsolutions to a certain sublinear equation have dead cores, and on uniform H\"older bounds.

math.AP

Beyond Averaging in John Ellipsoid Approximation: High-Accuracy Algorithms in the Leverage-Score Model

The John ellipsoid of a symmetric polytope $P=\{\mathbf{x}\in\mathbb{R}^d:\|\mathbf{A}\mathbf{x}\|_\infty\le1\}$, $\mathbf{A}\in\mathbb{R}^{n\times d}$, is computed by a long line of leverage-score algorithms, from Cohen, Cousins, Lee and Yang (COLT 2019) to its successors [WY24, CLS+25], all reaching a $(1+\varepsilon)$-approximation in $\Theta(\varepsilon^{-1}\log(n/d))$ iterations. We separate this complexity into three costs the modern line conflates (certification, identification, and accuracy) and locate the historical $\varepsilon^{-1}$ in the first alone. In the equivalent D-optimal-design form $\min_{\mathbf{p}\in\Delta_n}-\log\det(\sum_i p_i\mathbf{a}_i\mathbf{a}_i^\top)$, the leverage-score oracle is exactly the first-order oracle and the $(1+\varepsilon)$-John guarantee the Frank-Wolfe gap $g(\mathbf{p})\le\varepsilon d$; through this dictionary the costs come apart. The $\varepsilon^{-1}$ is a certification artifact: the uniform average of the iterates, the certificate used throughout the line, has gap exactly $\Theta(1/T)$, however cheap each iteration is made. Pointed instead at the last iterate the same oracle is fast: a warm-started accelerated method reaches the guarantee in $C(\mathbf{A})+O(\sqrt{\kappa}\log(1/\varepsilon))$ queries after an $\varepsilon$-independent setup $C(\mathbf{A})$, and once the optimal face is identified the facial problem is an unconstrained self-concordant minimization whose Hessian the oracle recovers exactly, so damped Newton needs only $O(\log\log(1/\varepsilon))$ steps, for a total of $C(\mathbf{A})+O(d^2\log\log(1/\varepsilon))$ queries. The accuracy dependence is thus doubly logarithmic after an $\varepsilon$-independent, condition-dependent setup; the open problem is the remaining identification cost (a condition-free bound on reaching the optimal face) and lower bounds. Accuracy is not the obstruction.

math.OC

Stateful Visual Encoders for Vision-Language Models

Vision-language models (VLMs) are increasingly used in multi-image, multi-turn agentic settings where decisions depend on visual changes. However, in existing open-weight VLMs, visual comparisons happen only inside the language model, while the visual encoder itself remains stateless: each image is encoded independently, without access to the prior visual context. As a result, small but task-critical changes may be attenuated before the language model has a chance to compare them, especially when those changes do not affect the high-level semantics of the scene. We introduce a Stateful Visual Encoder, which conditions each visual representation on prior visual features. Under supervised finetuning, VLMs equipped with stateful encoders achieve consistent improvements on controlled tasks involving cross-image spatial aggregation, multi-object visual differencing, and visual trajectory behavior cloning. These improvements are consistent across input resolutions, language model sizes, and VLM backbones. Finally, we validate our model on real-world tasks, including longitudinal radiology, fine-grained image comparison, and remote sensing, where stateful encoders consistently improve generalist VLM baselines and can match or surpass specialized models in selected domains. Project page: https://statefulvisualencoders.github.io/

cs.CV

Structural Feature Engineering for Generative Engine Optimization: How Content Structure Shapes Citation Behavior

The proliferation of AI-powered search engines has shifted information discovery from traditional link-based retrieval to direct answer generation with selective source citation, creating new challenges for content visibility. While existing Generative Engine Optimization (GEO) approaches focus primarily on semantic content modification, the role of structural features in influencing citation behavior remains underexplored. In this paper, we propose GEO-SFE, a systematic framework for structural feature engineering in generative engine optimization. Our approach decomposes content structure into three hierarchical levels: macro-structure (document architecture), meso-structure (information chunking), and micro-structure (visual emphasis), and models their impact on citation probability across different generative engine architectures. We develop architecture-aware optimization strategies and predictive models that preserve semantic integrity while improving structural effectiveness. Experimental evaluation across six mainstream generative engines demonstrates consistent improvements in citation rate (17.3 percent) and subjective quality (18.5 percent), validating the effectiveness and generalizability of the proposed framework. This work establishes structural optimization as a foundational component of GEO, providing a data-driven methodology for enhancing content visibility in LLM-powered information ecosystems.

cs.CL

Decentralized Multi-Agent System with Trust-Aware Communication

The emergence of Large Language Models (LLMs) is rapidly accelerating the development of autonomous multi-agent systems (MAS), paving the way for the Internet of Agents. However, traditional centralized MAS architectures present significant challenges, including single points of failure, vulnerability to censorship, inherent scalability limitations, and critical trust issues. We propose a novel Decentralized Multi-Agent System (DMAS) architecture designed to overcome these fundamental problems by enabling trust-aware, scalable, and censorship-resistant interactions among autonomous agents. Our DMAS features a decentralized agent runtime underpinned by a blockchain-based architecture. We formalize a trust-aware communication protocol that leverages cryptographic primitives and on-chain operations to provide security properties: verifiable interaction cycles, communication integrity, authenticity, non-repudiation, and conditional confidentiality, which we further substantiate through a comprehensive security analysis. Our performance analysis validates the DMAS as a scalable and efficient solution for building trustworthy multi-agent systems.

cs.MA

Normalized solutions for the Sobolev critical Schr\"{o}dinger equation with trapping potential

We study the existence and multiplicity of positive normalized solutions with prescribed $L^{2}$-norm for the Sobolev critical Schr\"odinger equation $-\Delta U + V(x) U = \lambda U + |U|^{2^*-2} U$ in $\mathbb{R}^N$, $\int_{\mathbb{R}^N} U^2\,dx = \rho^2$, where $N \ge 3$, $V\ge 0$ is a trapping potential, $\lambda \in \mathbb{R}$ and $2^*=\frac{2N}{N-2}$. Our first result is that the existence of local minimum solutions for $\rho \in (0, \rho^*)$, for some suitable $\rho^* > 0$, under appropriate assumptions on the potential. These solutions correspond to ground states. Our second result concerns the existence of mountain pass solutions, under the same assumptions.

math.AP

UnSAMv2: Self-Supervised Learning Enables Segment Anything at Any Granularity

The Segment Anything Model (SAM) family has become a widely adopted vision foundation model, but its ability to control segmentation granularity remains limited. Users often need to refine results manually - by adding more prompts or selecting from pre-generated masks - to achieve the desired level of detail. This process can be ambiguous, as the same prompt may correspond to several plausible masks, and collecting dense annotations across all granularities is prohibitively expensive, making supervised solutions infeasible. To address this limitation, we introduce UnSAMv2, which enables segment anything at any granularity without human annotations. UnSAMv2 extends the divide-and-conquer strategy of UnSAM by discovering abundant mask-granularity pairs and introducing a novel granularity control embedding that enables precise, continuous control over segmentation scale. Remarkably, with only $6$K unlabeled images and $0.02\%$ additional parameters, UnSAMv2 substantially enhances SAM-2, achieving segment anything at any granularity across interactive, whole-image, and video segmentation tasks. Evaluated on over $11$ benchmarks, UnSAMv2 improves $\text{NoC}_{90}$ (5.69 $\rightarrow$ 4.75), 1-IoU (58.0 $\rightarrow$ 73.1), and $\text{AR}_{1000}$ (49.6 $\rightarrow$ 68.3), showing that small amounts of unlabeled data with a granularity-aware self-supervised learning method can unlock the potential of vision foundation models.

cs.CV

Energy local minimizers for the nonlinear Schr\"{o}dinger equation on product spaces

We investigate the existence of local minimizers with prescribed $L^2$-norm for the energy functional associated to the mass-supercritical nonlinear Schr\"{o}dinger equation on the product space $\mathbb{R}^N \times M^k$, where $(M^k,g)$ is a compact Riemannian manifold, thus complementing the study of the mass-subcritical case performed by Terracini, Tzvetkov and Visciglia in [\emph{Anal. PDE} 2014, arXiv:1205.0342]. First we prove that, for small $L^2$-mass, the problem admits local minimizers. Next, we show that when the $L^2$-norm is sufficiently small, the local minimizers are constants along $M^k$, and they coincide with those of the corresponding problem on $\mathbb{R}^N$. Finally, under certain conditions, we show that the local minimizers obtained above are nontrivial along $M^k$. The latter situation occurs, for instance, for every $M^k$ of dimension $k\ge 2$, with the choice of an appropriate metric $\hat g$, and in $\mathbb{R}\times\mathbb{S}^k$, $k\ge 3$, where $\mathbb{S}^k$ is endowed with the standard round metric.

math.AP

Normalized solutions for the nonlinear Schr\"odinger equation with potential: the purely Sobolev critical case

We study the existence and multiplicity of positive solutions in $H^1(\mathbb{R}^N)$, $N\ge3$, with prescribed $L^2$-norm, for the (stationary) nonlinear Schr\"odinger equation with Sobolev critical power nonlinearity. It is well known that, in the free case, the associated energy functional has a mountain pass geometry on the $L^2$-sphere. This boils down, in higher dimensions, to the existence of a mountain pass solution which is (a suitable scaling of) the Aubin-Talenti function. In this paper, we consider the same problem, in presence of a weakly attractive, possibly irregular, potential, wondering (i) whether a local minimum solution appears, thus providing an orbitally stable family of solitons, and (ii) if the existence of a mountain-pass solution persists. We provide positive answers, depending on suitable assumptions on the potential and on the mass value. Moreover, by the Hopf-Cole transform, we give some applications of our results to the existence of multiple solutions to ergodic Mean Field Games systems with potential and quadratic Hamiltonian.

math.AP

DynTaskMAS: A Dynamic Task Graph-driven Framework for Asynchronous and Parallel LLM-based Multi-Agent Systems

The emergence of Large Language Models (LLMs) in Multi-Agent Systems (MAS) has opened new possibilities for artificial intelligence, yet current implementations face significant challenges in resource management, task coordination, and system efficiency. While existing frameworks demonstrate the potential of LLM-based agents in collaborative problem-solving, they often lack sophisticated mechanisms for parallel execution and dynamic task management. This paper introduces DynTaskMAS, a novel framework that orchestrates asynchronous and parallel operations in LLM-based MAS through dynamic task graphs. The framework features four key innovations: (1) a Dynamic Task Graph Generator that intelligently decomposes complex tasks while maintaining logical dependencies, (2) an Asynchronous Parallel Execution Engine that optimizes resource utilization through efficient task scheduling, (3) a Semantic-Aware Context Management System that enables efficient information sharing among agents, and (4) an Adaptive Workflow Manager that dynamically optimizes system performance. Experimental evaluations demonstrate that DynTaskMAS achieves significant improvements over traditional approaches: a 21-33% reduction in execution time across task complexities (with higher gains for more complex tasks), a 35.4% improvement in resource utilization (from 65% to 88%), and near-linear throughput scaling up to 16 concurrent agents (3.47X improvement for 4X agents). Our framework establishes a foundation for building scalable, high-performance LLM-based multi-agent systems capable of handling complex, dynamic tasks efficiently.

cs.MA

Quantum Speedups for Approximating the John Ellipsoid

In 1948, Fritz John proposed a theorem stating that every convex body has a unique maximal volume inscribed ellipsoid, known as the John ellipsoid. The John ellipsoid has become fundamental in mathematics, with extensive applications in high-dimensional sampling, linear programming, and machine learning. Designing faster algorithms to compute the John ellipsoid is therefore an important and emerging problem. In [Cohen, Cousins, Lee, Yang COLT 2019], they established an algorithm for approximating the John ellipsoid for a symmetric convex polytope defined by a matrix $A \in \mathbb{R}^{n \times d}$ with a time complexity of $O(nd^2)$. This was later improved to $O(\text{nnz}(A) + d^\omega)$ by [Song, Yang, Yang, Zhou 2022], where $\text{nnz}(A)$ is the number of nonzero entries of $A$ and $\omega$ is the matrix multiplication exponent. Currently $\omega \approx 2.371$ [Alman, Duan, Williams, Xu, Xu, Zhou 2024]. In this work, we present the first quantum algorithm that computes the John ellipsoid utilizing recent advances in quantum algorithms for spectral approximation and leverage score approximation, running in $O(\sqrt{n}d^{1.5} + d^\omega)$ time. In the tall matrix regime, our algorithm achieves quadratic speedup, resulting in a sublinear running time and significantly outperforming the current best classical algorithms.

cs.DS

Fast John Ellipsoid Computation with Differential Privacy Optimization

Determining the John ellipsoid - the largest volume ellipsoid contained within a convex polytope - is a fundamental problem with applications in machine learning, optimization, and data analytics. Recent work has developed fast algorithms for approximating the John ellipsoid using sketching and leverage score sampling techniques. However, these algorithms do not provide privacy guarantees for sensitive input data. In this paper, we present the first differentially private algorithm for fast John ellipsoid computation. Our method integrates noise perturbation with sketching and leverages score sampling to achieve both efficiency and privacy. We prove that (1) our algorithm provides $(\epsilon,\delta)$-differential privacy and the privacy guarantee holds for neighboring datasets that are $\epsilon_0$-close, allowing flexibility in the privacy definition; (2) our algorithm still converges to a $(1+\xi)$-approximation of the optimal John ellipsoid in $\Theta(\xi^{-2}(\log(n/\delta_0) + (L\epsilon_0)^{-2}))$ iterations where $n$ is the number of data point, $L$ is the Lipschitz constant, $\delta_0$ is the failure probability, and $\epsilon_0$ is the closeness of neighboring input datasets. Our theoretical analysis demonstrates the algorithm's convergence and privacy properties, providing a robust approach for balancing utility and privacy in John ellipsoid computation. This is the first differentially private algorithm for fast John ellipsoid computation, opening avenues for future research in privacy-preserving optimization techniques.

cs.DS

Normalized solutions for Sobolev critical Schr\"{o}dinger equations on bounded domains

We study the existence and multiplicity of positive solutions with prescribed $L^2$-norm for the Sobolev critical Schr\"odinger equation on a bounded domain $\Omega\subset\mathbb{R}^N$, $N\ge3$: \[ -\Delta U = \lambda U + U^{2^{*}-1},\qquad U\in H^1_0(\Omega),\qquad \int_\Omega U^2\,dx = \rho^{2}, \] where $2^*=\frac{2N}{N-2}$. First, we consider a general bounded domain $\Omega$ in dimension $N\ge3$, with a restriction, only in dimension $N=3$, involving its inradius and first Dirichlet eigenvalue. In this general case we show the existence of a mountain pass solution on the $L^2$-sphere, for $\rho$ belonging to a subset of positive measure of the interval $(0,\rho^{**})$, for a suitable threshold $\rho^{**}>0$. Next, assuming that $\Omega$ is star-shaped, we extend the previous result to all values $\rho\in(0,\rho^{**})$. With respect to that of local minimizers, already known in the literature, the existence of mountain pass solutions in the Sobolev critical case is much more elusive. In particular, our proofs are based on the sharp analysis of the bounded Palais-Smale sequences, provided by a nonstandard adaptation of the Struwe monotonicity trick, that we develop.

math.AP

Parallelizing quantum simulation with decision diagrams

Recent technological advancements show promise in leveraging quantum mechanical phenomena for computation. This brings substantial speed-ups to problems that are once considered to be intractable in the classical world. However, the physical realization of quantum computers is still far away from us, and a majority of research work is done using quantum simulators running on classical computers. Classical computers face a critical obstacle in simulating quantum algorithms. Quantum states reside in a Hilbert space whose size grows exponentially to the number of subsystems, i.e., qubits. As a result, the straightforward statevector approach does not scale due to the exponential growth of the memory requirement. Decision diagrams have gained attention in recent years for representing quantum states and operations in quantum simulations. The main advantage of this approach is its ability to exploit redundancy. However, mainstream quantum simulators still rely on statevectors or tensor networks. We consider the absence of decision diagrams due to the lack of parallelization strategies. This work explores several strategies for parallelizing decision diagram operations, specifically for quantum simulations. We propose optimal parallelization strategies. Based on the experiment results, our parallelization strategy achieves a 2-3 times faster simulation of Grover's algorithm and random circuits than the state-of-the-art single-thread DD-based simulator DDSIM.

quant-ph

Normalized solutions for a Choquard equation with exponential growth in $\mathbb{R}^{2}$

In this paper, we study the existence of normalized solutions to the following nonlinear Choquard equation with exponential growth \begin{align*} \left\{ \begin{aligned} &-\Delta u+\lambda u=(I_{\alpha}\ast F(u))f(u), \quad \quad \hbox{in }\mathbb{R}^{2},\\ &\int_{\mathbb{R}^{2}}|u|^{2}dx=a^{2}, \end{aligned} \right. \end{align*} where $a>0$ is prescribed, $\lambda\in \mathbb{R}$, $\alpha\in(0,2)$, $I_{\alpha}$ denotes the Riesz potential, $\ast$ indicates the convolution operator, the function $f(t)$ has exponential growth in $\mathbb{R}^{2}$ and $F(t)=\int^{t}_{0}f(\tau)d\tau$. Using the Pohozaev manifold and variational methods, we establish the existence of normalized solutions to the above problem.

math.AP

Normalized solutions for Schr\"{o}dinger systems in dimension two

In this paper, we study the existence of normalized solutions to the following nonlinear Schr\"{o}dinger systems with exponential growth \begin{align*} \left\{ \begin{aligned} &-\Delta u+\lambda_{1}u=H_{u}(u,v), \quad \quad \hbox{in }\mathbb{R}^{2},\\ &-\Delta v+\lambda_{2} v=H_{v}(u,v), \quad \quad \hbox{in }\mathbb{R}^{2},\\ &\int_{\mathbb{R}^{2}}|u|^{2}dx=a^{2},\quad \int_{\mathbb{R}^{2}}|v|^{2}dx=b^{2}, \end{aligned} \right. \end{align*} where $a,b>0$ are prescribed, $\lambda_{1},\lambda_{2}\in \mathbb{R}$ and the functions $H_{u},H_{v}$ are partial derivatives of a Carath\'{e}odory function $H$ with $H_{u},H_{v}$ have exponential growth in $\mathbb{R}^{2}$. Our main results are totally new for Schr\"{o}dinger systems in $\mathbb{R}^{2}$. Using the Pohozaev manifold and variational methods, we establish the existence of normalized solutions to the above problem.

math.AP

Existence of solution for a class of fractional Hamiltonian-type elliptic systems with exponential critical growth in R

In this paper, we study the following class of fractional Hamiltonian systems: \begin{eqnarray*} \begin{aligned}\displaystyle \left\{ \arraycolsep=1.5pt \begin{array}{ll} (-\Delta)^{\frac{1}{2}} u + u = \Big(I_{\mu_{1}}\ast G(v)\Big)g(v) \ \ \ & \mbox{in} \ \mathbb{R},\\[2mm] (-\Delta)^{\frac{1}{2}} v + v = \Big(I_{\mu_{2}}\ast F(u)\Big)f(u) \ \ \ & \mbox{in} \ \mathbb{R}, \end{array} \right. \end{aligned} \end{eqnarray*} where $(-\Delta)^{\frac{1}{2}}$ is the square root Laplacian operator, $\mu_{1},\mu_{2}\in(0,1)$, $I_{\mu_{1}},I_{\mu_{2}}$ denote the Riesz potential, $\ast$ indicates the convolution operator, $F(s),G(s)$ are the primitive of $f(s),g(s)$ with $f(s),g(s)$ have exponential growth in $\mathbb{R}$. Using the linking theorem and variational methods, we establish the existence of at least one positive solution to the above problem.

math.AP