arXiv ScienceSearch

arXiv subjects

Linyu Miao

Publications and source records attributed to Linyu Miao.

3 recordsLinked to original sources

Classical Solutions for a Finite-Horizon Exit-Time Problem with Degenerate Diffusion Control

We study finite-horizon exit-time control of a one-dimensional affine diffusion, with an unbounded control acting on both drift and volatility. Since an admissible control can cancel the instantaneous volatility, the associated HJB equation is not uniformly parabolic. Moreover, the location of the degeneracy is not prescribed by the model coefficients but depends on the unknown value function and optimal feedback. Rather than relying on a viscosity-solution formulation, we work under suitable structural assumptions and prove that the value function is a $C^{1,2}$ classical solution throughout the interior of the state domain and is strictly convex in the state variable. Moreover, we construct a locally Lipschitz optimal feedback, explicitly identify the degenerate set as a curve, and show that the closed-loop drift on this curve equals the time derivative of its state coordinate. Finally, we prove invariance and inaccessibility before exit and derive exact squared-logarithmic asymptotics for approach probabilities.

math.OC

Global Sobolev convergence of Howard iteration for parabolic Bellman equations with controlled diffusion

We prove global Sobolev convergence of Howard policy iteration for finite-horizon Bellman equations with control-dependent diffusion. In one dimension, the result holds for bounded measurable coefficients under uniform ellipticity, without a large discount, a short horizon, a small diffusion perturbation, or regularity assumptions on the improving policies. We also obtain multidimensional extensions under structural conditions on the diffusion matrix. The key observation is that vanishing policy improvements force the Bellman residual to converge in measure; higher integrability then yields strong convergence and constructs the unique strong solution. We extend the argument to entropy-regularized models, including zero-temperature limits, and establish quadratic convergence for special one-dimensional models.

math.AP

ACADREASON: Exploring the Limits of Reasoning Models with Academic Research Problems

In recent years, the research focus of large language models (LLMs) and agents has shifted increasingly from demonstrating novel capabilities to complex reasoning and tackling challenging tasks. However, existing evaluations focus mainly on math/code contests or general tasks, while existing multi-domain academic benchmarks lack sufficient reasoning depth, leaving the field without a rigorous benchmark for high-level reasoning. To fill this gap, we introduce the Acadreason benchmark, designed to evaluate the ability of LLMs and agents to acquire and reason over academic knowledge. It consists of 50 expert-annotated academic problems across five high-reasoning domains, including computer science, economics, law, mathematics, and philosophy. All questions are sourced from top-tier publications in recent years and undergo rigorous annotation and quality control to ensure they are both challenging and answerable. We conduct systematic evaluations of over 10 mainstream LLMs and agents. The results show that most LLMs scored below 20 points, with even the cutting-edge GPT-5 achieving only 16 points. While agents achieved higher scores, none exceeded 40 points. This demonstrates the current capability gap between LLMs and agents in super-intelligent academic research tasks and highlights the challenges of Acadreason.

cs.CL