arXiv ScienceSearch

arXiv subjects

Qiang Ren

Publications and source records attributed to Qiang Ren.

12 recordsLinked to original sources

Solutions with clustering concentration layers to the Ambrosetti-Prodi type problem

We consider the following Ambrosetti-Prodi type problem \begin{equation} \left\{\begin{array}{ll} -\mathrm{div} (A(x)\nabla u)=|u|^p-t\mathbfΨ(x), &\mbox{in $Ω$,} \\ u=0, & \mbox{on $\partial Ω$}, \end{array} \right. \end{equation} where $Ω\subset \mathbb{R}^2$, $t>0$, $p>3$ and $\mathbfΨ$ is an eigenfunction corresponding to the first eigenvalue of the following operator \[\mathfrak{L}(u)=-\mathrm{div} (A(x)\nabla u).\] Moreover, $A(x)=\{A_{ij}(x)\}_{2\times 2}$ is a symmetric positive defined matrix function. Let $Γ\subset Ω$ be a closed curve and also a non-degenerate critical point of the functional \[\mathcal{K}(Γ)=\int_Γ\mathbfΨ^{\frac{p+3}{2p}}dvol_{\mathfrak{g}},\] where $\mathfrak{g}(X,Y)=\langle A^*X,Y\rangle$ is a Riemannian metric on $\mathbb{R}^2$ and $A^*$ is the adjoint matrix for $A$. We prove that there exists a sequence of $t=t_l\to +\infty$ such that this problem has solutions $u_{t_l}$ with clustering concentration layers directed along $Γ$.

math.AP

MiroThinker: Pushing the Performance Boundaries of Open-Source Research Agents via Model, Context, and Interactive Scaling

We present MiroThinker v1.0, an open-source research agent designed to advance tool-augmented reasoning and information-seeking capabilities. Unlike previous agents that only scale up model size or context length, MiroThinker explores interaction scaling at the model level, systematically training the model to handle deeper and more frequent agent-environment interactions as a third dimension of performance improvement. Unlike LLM test-time scaling, which operates in isolation and risks degradation with longer reasoning chains, interactive scaling leverages environment feedback and external information acquisition to correct errors and refine trajectories. Through reinforcement learning, the model achieves efficient interaction scaling: with a 256K context window, it can perform up to 600 tool calls per task, enabling sustained multi-turn reasoning and complex real-world research workflows. Across four representative benchmarks-GAIA, HLE, BrowseComp, and BrowseComp-ZH-the 72B variant achieves up to 81.9%, 37.7%, 47.1%, and 55.6% accuracy respectively, surpassing previous open-source agents and approaching commercial counterparts such as GPT-5-high. Our analysis reveals that MiroThinker benefits from interactive scaling consistently: research performance improves predictably as the model engages in deeper and more frequent agent-environment interactions, demonstrating that interaction depth exhibits scaling behaviors analogous to model size and context length. These findings establish interaction scaling as a third critical dimension for building next-generation open research agents, complementing model capacity and context windows.

cs.CL

MiroEval: Benchmarking Multimodal Deep Research Agents in Process and Outcome

Recent progress in deep research systems has been impressive, but evaluation still lags behind real user needs. Existing benchmarks predominantly assess final reports using fixed rubrics, failing to evaluate the underlying research process. Most also offer limited multimodal coverage, rely on synthetic tasks that do not reflect real-world query complexity, and cannot be refreshed as knowledge evolves. To address these gaps, we introduce MiroEval, a benchmark and evaluation framework for deep research systems. The benchmark comprises 100 tasks (70 text-only, 30 multimodal), all grounded in real user needs and constructed via a dual-path pipeline that supports periodic updates, enabling a live and evolving setting. The proposed evaluation suite assesses deep research systems along three complementary dimensions: adaptive synthesis quality evaluation with task-specific rubrics, agentic factuality verification via active retrieval and reasoning over both web sources and multimodal attachments, and process-centric evaluation audits how the system searches, reasons, and refines throughout its investigation. Evaluation across 13 systems yields three principal findings: the three evaluation dimensions capture complementary aspects of system capability, with each revealing distinct strengths and weaknesses across systems; process quality serves as a reliable predictor of overall outcome while revealing weaknesses invisible to output-level metrics; and multimodal tasks pose substantially greater challenges, with most systems declining by 3 to 10 points. The MiroThinker series achieves the most balanced performance, with MiroThinker-H1 ranking the highest overall in both settings. Human verification and robustness results confirm the reliability of the benchmark and evaluation framework. MiroEval provides a holistic diagnostic tool for the next generation of deep research agents.

cs.AI

A GPU-based Monte Carlo framework for IMRT QA using EPID transit dosimetry

Purpose: We presented a GPU-based MC framework, ARCHER-EPID, specifically designed for EPID transit dosimetry, with improving accuracy and efficiency. Methods: A comprehensive MC framework was developed to perform full radiation transport simulations through three distinct zones: a detailed linear accelerator head model, a CT-based patient/phantom geometry, and a realistic, multi-layered EPID model. To convert the simulated absorbed dose to a realistic detector signal, a dose-response correction model was implemented. The framework was validated by comparing simulations against experimental measurements for 25 IMRT fields delivered to both a solid water phantom and a anthropomorphic phantom. Agreement was quantified using Gamma analysis. Results: The GPU-accelerated ARCHER-EPID framework can complete the simulation for a complex IMRT field in about 90 seconds. A 2D correction factor lookup table is generated by parameterizing radiological thickness and effective field size to account for the EPID's energy-dependent response. The data revealed that for small fields, beam hardening is the dominant effect, while for large fields, the contribution from patient-generated scatter overwhelms this effect. The average 2D gamma passing rates (3%/3 mm criteria) between simulation and measurements are 98.43% for the solid water phantom and 97.86% for the anthropomorphic phantom, respectively. Visual comparison of the images and dose profiles between simulation and measurements show a high degree of agreement. Conclusions: We have successfully developed and validated a GPU-based MC framework that provides gold-standard accuracy for EPID transit dosimetry in radiotherapy. The results demonstrate that our proposed method has potential for routine application in PSQA.

physics.med-ph

Sign-changing concentration phenomena of an anisotropic sinh-Poisson type equation with a Hardy or Hénon term

We consider the following anisotropic sinh-Poisson tpye equation with a Hardy or Hénon term: $$-\mathrm{div} (a(x)\nabla u)+ a(x)u=\varepsilon^2a(x)|x-q|^{2α}(e^u-e^{-u}) \quad\mathrm{in}\quad Ω,$$ $$\frac{\partial u}{\partial n}=0,\quad \mathrm{on}\quad \partialΩ,$$ where $\varepsilon>0$, $q\in \barΩ\subset \mathbb{R}^2$, $α\in(-1,\infty)- \mathbb{N}$, $Ω\subset \mathbb{R}^2$ is a smooth bounded domain, $n$ is the unit outward normal vector of $\partial Ω$ and $a(x)$ is a smooth positive function defined on $\barΩ$. From finite dimensional reduction method, we proved that this problem has a sequence of sign-changing solutions with arbitrarily many interior spikes accumulating to $q$, provided $q\in Ω$ is a local maximizer of $a(x)$. However, if $q\in \partial Ω$ is a strict local maximum point of $a(x)$ and satisfies $\langle \nabla a(q),n \rangle=0$, we proved that this problem has a family of sign-changing solutions with arbitrarily many mixed interior and boundary spikes accumulating to $q$. Under the same condition, we could also construct a sequence of blow-up solutions for the following problem $$ -\mathrm{div} (a(x)\nabla u)+ a(x)u=\varepsilon^2a(x)|x-q|^{2α}e^u\quad \mathrm{in} \quadΩ,$$ $$\frac{\partial u}{\partial n}=0, \quad \mathrm{on}\quad \partialΩ.$$

math.AP

High Noise Immune Time-domain Inversion via Cascade Network (TICaN) for Complex Scatterers

In this paper, a high noise immune time-domain inversion cascade network (TICaN) is proposed to reconstruct scatterers from the measured electromagnetic fields. The TICaN is comprised of a denoising block aiming at improving the signal-to-noise ratio, and an inversion block to reconstruct the electromagnetic properties from the raw time-domain measurements. The scatterers investigated in this study include complicated geometry shapes and high contrast, which cover the stratum layer, lossy medium and hyperfine structure, etc. After being well trained, the performance of the TICaN is evaluated from the perspective of accuracy, noise-immunity, computational acceleration, and generalizability. It can be proven that the proposed framework can realize high-precision inversion under high-intensity noise environments. Compared with traditional reconstruction methods, TICaN avoids the tedious iterative calculation by utilizing the parallel computing ability of GPU and thus significantly reduce the computing time. Besides, the proposed TICaN has certain generalization ability in reconstructing the unknown scatterers such as the famous Austria rings. Herein, it is confident that the proposed TICaN will serve as a new path for real-time quantitative microwave imaging for various practical scenarios.

eess.SP

Progress in Water-Based Metamaterial Absorber: A Review

Increasing attention on microwave ultra-broadband metamaterial absorbers has been paid due to their promising applications. While most microwave ultra-broadband metamaterial absorbers developed so far are based on metallic resonant structures, dispersive dielectric water-based metamaterial opens a simpler and more versatile route for the construction of polarization- and angle- insensitive ultra-broadband absorption. Here, we review the recent progresses of water-based metamaterial absorber by providing an illustration of the mechanisms to realize ultra-broadband, tunable and multi-functional absorption. We also address the further development direction and some potential novel applications.

physics.app-ph

Byzantine-Robust Federated Learning via Credibility Assessment on Non-IID Data

Federated learning is a novel framework that enables resource-constrained edge devices to jointly learn a model, which solves the problem of data protection and data islands. However, standard federated learning is vulnerable to Byzantine attacks, which will cause the global model to be manipulated by the attacker or fail to converge. On non-iid data, the current methods are not effective in defensing against Byzantine attacks. In this paper, we propose a Byzantine-robust framework for federated learning via credibility assessment on non-iid data (BRCA). Credibility assessment is designed to detect Byzantine attacks by combing adaptive anomaly detection model and data verification. Specially, an adaptive mechanism is incorporated into the anomaly detection model for the training and prediction of the model. Simultaneously, a unified update algorithm is given to guarantee that the global model has a consistent direction. On non-iid data, our experiments demonstrate that the BRCA is more robust to Byzantine attacks compared with conventional methods

cs.LG

Predicting Surface Heat Flux on Complex Systems via Conv-LSTM

Existing algorithms with iterations as the principle for 3D inverse heat conduction problems (IHCPs) are usually time-consuming. With the recent advancements in deep learning techniques, it is possible to apply the neural network to compute IHCPs. In this paper, a new framework based on Convolutional-LSTM is introduced to predict the transient heat flux via measured temperature. The inverse heat conduction models concerned in this work have 3D complex structures with non-linear boundary conditions and thermophysical parameters. In order to reach high precision, a forward solver based on the finite element method is utilized to generate sufficient data for training. The fully trained framework can provide accurate predictions efficiently once the measured temperature and models are acquired. It is believed that the proposed framework offers a new pattern for real-time heat flux inversion.

cs.CE

Adaptive Routing Between Capsules

Capsule network is the most recent exciting advancement in the deep learning field and represents positional information by stacking features into vectors. The dynamic routing algorithm is used in the capsule network, however, there are some disadvantages such as the inability to stack multiple layers and a large amount of computation. In this paper, we propose an adaptive routing algorithm that can solve the problems mentioned above. First, the low-layer capsules adaptively adjust their direction and length in the routing algorithm and removing the influence of the coupling coefficient on the gradient propagation, so that the network can work when stacked in multiple layers. Then, the iterative process of routing is simplified to reduce the amount of computation and we introduce the gradient coefficient $λ$. Further, we tested the performance of our proposed adaptive routing algorithm on CIFAR10, Fashion-MNIST, SVHN and MNIST, while achieving better results than the dynamic routing algorithm.

cs.LG

Grouping Capsules Based Different Types

Capsule network was introduced as a new architecture of neural networks, it encoding features as capsules to overcome the lacking of equivariant in the convolutional neural networks. It uses dynamic routing algorithm to train parameters in different capsule layers, but the dynamic routing algorithm need to be improved. In this paper, we propose a novel capsule network architecture and discussed the effect of initialization method of the coupling coefficient $c_{ij}$ on the model. First, we analyze the rate of change of the initial value of $c_{ij}$ when the dynamic routing algorithm iterates. The larger the initial value of $c_{ij}$, the better effect of the model. Then, we proposed improvement that training different types of capsules by grouping capsules based different types. And this improvement can adjust the initial value of $c_{ij}$ to make it more suitable. We experimented with our improvements on some computer vision datasets and achieved better results than the original capsule network

cs.CV

Multi-bump solutions for fractional Nirenberg problem

We consider the multi-bump solutions of the following fractional Nirenberg problem \begin{equation}\label{01} (-Δ)^s u=K(x)u^{\frac{n+2s}{n-2s}}, \;\;\;\;u>0\;\;\text{ in }\mathbb{R}^n, \end{equation} where $s\in (0,1)$ and $n>2+2s$. If $K$ is a periodic function in some $k$ variables with $1\leq k<\frac{n-2s}2$, we proved that \eqref{01} has multi-bump solutions with bumps clustered on some lattice points in $\mathbb{R}^k$ via Lyapunov-Schmidt reduction. It is also established that the equation \eqref{01} has an infinite-many-bump solutions with bumps clustered on some lattice points in $\mathbb{R}^n$ which is isomorphic to $\mathbb{Z}_+^k$.

math.AP