arXiv ScienceSearch

arXiv subjects

Zijie Lin

Publications and source records attributed to Zijie Lin.

At least 19 recordsLinked to original sources

Just-In-Time Reinforcement Learning: Continual Learning in LLM Agents Without Gradient Updates

While Large Language Model (LLM) agents excel at general tasks, they inherently struggle with continual adaptation due to the frozen weights after deployment. Conventional reinforcement learning (RL) offers a solution but incurs prohibitive computational costs and the risk of catastrophic forgetting. We introduce Just-In-Time Reinforcement Learning (JitRL), a training-free framework that enables test-time policy optimization without any gradient updates. JitRL maintains a dynamic, non-parametric memory of experiences and retrieves relevant trajectories to estimate action advantages on-the-fly. These estimates are then used to directly modulate the LLM's output logits. We theoretically prove that this additive update rule is the exact closed-form solution to the KL-constrained policy optimization objective. Extensive experiments on WebArena and Jericho demonstrate that JitRL establishes a new state-of-the-art among training-free methods. Crucially, JitRL outperforms the performance of computationally expensive fine-tuning methods (e.g., WebRL) while reducing monetary costs by over 30 times, offering a scalable path for continual learning agents. The code is available at https://github.com/liushiliushi/JitRL.

cs.LG

Enhancing Automated Paper Reproduction via Prompt-Free Collaborative Agents

Automated paper reproduction has emerged as a promising approach to accelerate scientific research, employing multi-step workflow frameworks to systematically convert academic papers into executable code. However, existing frameworks often lack mechanisms to verify and refine the outputs at each generation step, or rely heavily on manually designed prompts for self-refinement, which limits their adaptability and scalability. To address these limitations, we propose a prompt-free collaborative agent framework that automatically enhances the quality of paper-to-code generation. Our approach employs two collaborative agents: a verification agent that examines whether the outputs at each step satisfy the requirements specified in the corresponding system prompt, and a refinement agent that revises the outputs based on the identified issues. Unlike previous methods that require human experts to craft specific refinement prompts for each step, our framework achieves automatic verification and improvement by leveraging only the original system prompts. We integrate our collaborative agents into the Paper2Code framework and conduct comprehensive experiments on PaperBench Code-Dev and Paper2CodeBench datasets. Experimental results demonstrate that our approach significantly improves the accuracy and completeness of reproduced code, achieving performance gains of approximately 15\% and 13\%, respectively, compared to the baseline without our agents. Furthermore, comparative experiments against Self-Refine validate the robustness and consistency of our prompt-free approach across different datasets.

cs.AI

Minimal PI-systems with all points are multiply minimal

We construct a minimal subshift \((X^{*},\sigma)\) that serves as an open proximal extension of its maximal equicontinuous factor. We establish that every point in this subshift is multiply recurrent minimal. This work solves an open problem raised by Huang, Shao and Ye regarding the existence of minimal PI-systems such that each point is multiply minimal.

math.DS

Enhancing Multi-Agent Debate System Performance via Confidence Expression

Generative Large Language Models (LLMs) have demonstrated remarkable performance across a wide range of tasks. Recent research has introduced Multi-Agent Debate (MAD) systems, which leverage multiple LLMs to simulate human debate and thereby improve task performance. However, while some LLMs may possess superior knowledge or reasoning capabilities for specific tasks, they often struggle to clearly communicate this advantage during debates, in part due to a lack of confidence expression. Moreover, inappropriate confidence expression can cause agents in MAD systems to either stubbornly maintain incorrect beliefs or converge prematurely on suboptimal answers, ultimately reducing debate effectiveness and overall system performance. To address these challenges, we propose incorporating confidence expression into MAD systems to allow LLMs to explicitly communicate their confidence levels. To validate this approach, we develop ConfMAD, a MAD framework that integrates confidence expression throughout the debate process. Experimental results demonstrate the effectiveness of our method, and we further analyze how confidence influences debate dynamics, offering insights into the design of confidence-aware MAD systems.

cs.CL

IGD: Token Decisiveness Modeling via Information Gain in LLMs for Personalized Recommendation

Large Language Models (LLMs) have shown strong potential for recommendation by framing item prediction as a token-by-token language generation task. However, existing methods treat all item tokens equally, simply pursuing likelihood maximization during both optimization and decoding. This overlooks crucial token-level differences in decisiveness-many tokens contribute little to item discrimination yet can dominate optimization or decoding. To quantify token decisiveness, we propose a novel perspective that models item generation as a decision process, measuring token decisiveness by the Information Gain (IG) each token provides in reducing uncertainty about the generated item. Our empirical analysis reveals that most tokens have low IG but often correspond to high logits, disproportionately influencing training loss and decoding, which may impair model performance. Building on these insights, we introduce an Information Gain-based Decisiveness-aware Token handling (IGD) strategy that integrates token decisiveness into both tuning and decoding. Specifically, IGD downweights low-IG tokens during tuning and rebalances decoding to emphasize tokens with high IG. In this way, IGD moves beyond pure likelihood maximization, effectively prioritizing high-decisiveness tokens. Extensive experiments on four benchmark datasets with two LLM backbones demonstrate that IGD consistently improves recommendation accuracy, achieving significant gains on widely used ranking metrics compared to strong baselines.

cs.CL

Improving Phishing Email Detection Performance of Small Large Language Models

Large language models(LLMs) have demonstrated remarkable performance on many natural language processing(NLP) tasks and have been employed in phishing email detection research. However, in current studies, well-performing LLMs typically contain billions or even tens of billions of parameters, requiring enormous computational resources. To reduce computational costs, we investigated the effectiveness of small-parameter LLMs for phishing email detection. These LLMs have around 3 billion parameters and can run on consumer-grade GPUs. However, small LLMs often perform poorly in phishing email detection task. To address these issues, we designed a set of methods including Prompt Engineering, Explanation Augmented Fine-tuning, and Model Ensemble to improve phishing email detection capabilities of small LLMs. We validated the effectiveness of our approach through experiments, significantly improving both accuracy and F1 score on the SpamAssassin and CEAS\_08 datasets. Furthermore, the fine-tuned models demonstrated strong transferability, achieving robust performance across multiple unseen phishing datasets, outperforming traditional baselines and approaching standard-sized LLMs.

cs.CL

AutoP2C: An LLM-Based Agent Framework for Code Repository Generation from Multimodal Content in Academic Papers

Machine Learning (ML) research is spread through academic papers featuring rich multimodal content, including text, diagrams, and tabular results. However, translating these multimodal elements into executable code remains a challenging and time-consuming process that requires substantial ML expertise. We introduce ``Paper-to-Code'' (P2C), a novel task that transforms the multimodal content of scientific publications into fully executable code repositories, which extends beyond the existing formulation of code generation that merely converts textual descriptions into isolated code snippets. To automate the P2C process, we propose AutoP2C, a multi-agent framework based on large language models that processes both textual and visual content from research papers to generate complete code repositories. Specifically, AutoP2C contains four stages: (1) repository blueprint extraction from established codebases, (2) multimodal content parsing that integrates information from text, equations, and figures, (3) hierarchical task decomposition for structured code generation, and (4) iterative feedback-driven debugging to ensure functionality and performance. Evaluation on a benchmark of eight research papers demonstrates the effectiveness of AutoP2C, which can successfully generate executable code repositories for all eight papers, while OpenAI-o1 or DeepSeek-R1 can only produce runnable code for one paper. The code is available at https://github.com/shoushouyu/Automated-Paper-to-Code.

cs.SE

Completely Li--Yorke chaotic homeomorphisms with positive entropy

It is an open problem whether a homeomorphism on a compact metric space satisfying that each proper pair is either positively or negatively Li--Yorke, called completely Li--Yorke chaotic, can have positive entropy. In the present paper, an affirmative answer to this question is given. In fact, for each homeomorphism $T$ with positive entropy such that each proper pair is not two-sided asymptotic, a completely Li--Yorke chaotic homeomorphism with positive entropy associated with the given homeomorphism can be constructed.

math.DS

Group Extensions for Random Shifts of Finite Type

Symbolic dynamical theory plays an important role in the research of amenability with a countable group. Motivated by the deep results of Dougall and Sharp, we study the group extensions for topologically mixing random shifts of finite type. For a countable group $G$, we consider the potential connections between relative Gurevi\v{c} pressure (entropy), the spectral radius of random Perron-Frobenius operator and amenability of $G$. Given $G^{\rm ab}$ by the abelianization of $G$ where $G^{\rm ab}=G/[G,G]$, we consider the random group extensions of random shifts of finite type between $G$ and $G^{\rm ab}$. It can be proved that the relative Gurevi\v{c} entropy of random group $G$ extensions is equal to the relative Gurevi\v{c} entropy of random group $G^{\rm ab}$ extensions if and only if $G$ is amenable. Moreover, we establish the relativized variational principle and discuss the unique equilibrium state for random group $\mathbb{Z}^{d}$ extensions.

math.DS

Weighted Topological Entropy of Random Dynamical Systems

Let $f_{i},i=1,2$ be continuous bundle random dynamical systems over an ergodic compact metric system $(\Omega,\mathcal{F},\mathbb{P},\vartheta)$. Assume that ${\bf a}=(a_{1},a_{2})\in\mathbb{R}^{2}$ with $a_{1}>0$ and $a_{2}\geq0$, $f_{2}$ is a factor of $f_{1}$ with a factor map $\Pi:\Omega\times X_{1}\rightarrow\Omega\times X_{2}$. We define the ${\bf a}$-weighted Bowen topological entropy of $h^{{\bf a}}(\omega,f_{1},X_{1})$ of $f_{1}$ with respect to $\omega\in \Omega$. It is shown that the quality $h^{{\bf a}}(\omega,f_{1},X_{1})$ is measurable in $\Omega$, and denoted that $h^{{\bf a}}(f_{1},\Omega\times X_{1})$ is the integration of $h^{{\bf a}}(\omega,f_{1},X_{1})$ against $\mathbb{P}$. We prove the following variational principle: \begin{align*} h^{{\bf a}}(f_{1},\Omega\times X_{1})=\sup\left\{a_{1}h_{\mu}^{(r)}(f_{1})+a_{2}h_{\mu\circ\Pi^{-1}}^{(r)}(f_{2})\right\}, \end{align*} where the supremum is taken over the set of all $\mu\in\mathcal{M}_{\mathbb{P}}^{1}(\Omega\times X_{1},f_{1})$. In the case of random dynamical systems with an ergodic and compact driving system, this gives an affirmative answer to the question posed by Feng and Huang [Variational principle for weighted topological pressure, J. Math. Pures Appl. 106 (2016), 411-452]. It also generalizes the relativized variational principle for fiber topological entropy, and provides a topological extension of Hausdorff dimension of invariant sets and random measures on the $2$-torus $\mathbb{T}^{2}$. In addition, the Shannon-McMillan-Breiman theorem, Brin-Katok local entropy formula and Katok entropy formula of weighted measure-theoretic entropy for random dynamical systems are also established in this paper.

math.DS

3D large-scale fused silica microfluidic chips enabled by hybrid laser microfabrication for continuous-flow UV photochemical synthesis

We demonstrate a hybrid laser microfabrication approach, which combines the technical merits of ultrafast laser-assisted chemical etching and carbon dioxide laser-induced in-situ melting, for centimeter-scale and bonding-free fabrication of 3D complex hollow microstructures in fused silica glass. With the developed approach, large-scale fused silica microfluidic chips with integrated 3D cascaded micromixing units can be reliably manufactured. High-performance on-chip mixing and continuous-flow photochemical synthesis under UV LEDs irradiation at ~280 nm were demonstrated using the manufactured chip, indicating a powerful capability for versatile fabrication of highly transparent all-glass microfluidic reactors for on-chip photochemical synthesis.

physics.app-ph

High pointwise emergence and Katok's conjecture for systems with non-uniform structure

Recently, Kiriki, Nakano and Soma introduced a concept called pointwise emergence as a new quantitative perspective into the study of non-existence of averages for dynamical systems. In the present paper, we consider the set of points with high pointwise emergence for systems with non-uniform structure and prove that this set carries full topological pressure. For the proof of this result, we show that such systems have ergodic measures of arbitrary intermediate pressures.

math.DS

Mean Li--Yorke chaos and multifractal analysis on subshifts

In the present paper, we use the generalized multifractal framework introduced by Olsen to study the Bowen entropy and packing entropy of historic sets with typical weights over aperiodic and irreducible shifts of finite type. Following those results and a transfer from almost everywhere to everywhere, we show that for each point $\omega$ in a irreducible shift of finite type $\Sigma_A$, the Bowen entropy of the set consisting of all the points that are mean Li-Yorke pairs with $\omega$ is $0$, and its packing entropy is full. This result is beyond the ergodic theory. Also, by the transfer from almost everywhere to everywhere, we show that for each point $\omega$ in a irreducible shift of finite type $\Sigma_A$, the Bowen entropy of the set consisting of all the points that are Li-Yorke pairs with $\omega$ is full. This result is also beyond the ergodic theory.

math.DS

Equilibrium states which are not Gibbs measure on hereditary subshifts

In this paper, we consider which kind of invariant measure on hereditary subshifts is not Gibbs measure. For the hereditary closure of a subshift $(X,S)$, we prove that in some situation, the invariant measure $ν*B_{p,1-p}$ can not be a Gibbs measure where $ν$ is an invariant measure on $(X,S)$. As an application, we show that for some $\B$-free subshifts, the unique equilibrium state $ν_η*B_{p,1-p}$ is not Gibbs measure.

math.DS

Shadowing and mixing on systems of countable group actions

Let $(X,G,Φ)$ be a dynamical system, where $X$ is compact Hausdorff space, and $G$ is a countable discrete group. We investigate shadowing property and mixing between subshifts and general dynamical systems. For the shadowing property, fix some finite subset $S\subset G$. We prove that if $X$ is totally disconnected, then $Φ$ has $S$-shadowing property if and only if $(X,G,Φ)$ is conjugate to an inverse limit of a sequence of shifts of finite type which satisfies Mittag-Leffler condition. Also, suppose that $X$ is metric space (may be not totally disconnected), we prove that if $Φ$ has $S$-shadowing property, then $(X,G,Φ)$ is a factor of an inverse limit of a sequence of shifts of finite type by a factor map which almost lifts pseudo-orbit for $S$. On the other hand, let property $P$ be one of the following property: transitivity, minimal, totally transitivity, weakly mixing, mixing, and specification property. We prove that if $X$ is totally disconnected, then $Φ$ has property $P$ if and only if $(X,G,Φ)$ is conjugate to an inverse limit of an inverse system that consists of subshifts with property $P$ which satisfies Mittag-Leffler condition. Also, for the case of metric space (may be not totally disconnected), if property $P$ is not minimal or specification property, we prove that $Φ$ has property $P$ if and only if $(X,G,Φ)$ is a factor of an inverse limit of a sequence of subshifts with property $P$ which satisfies Mittag-Leffler condition.

math.DS

Hausdorff measure of sets of distributional chaotic pairs for shift maps

Let $σ_{K}: \sum_{K}\rightarrow \sum_{K}$ be a shift map. For an interval $[p,q]\subset[0,1]$, let $D_{σ_{K}}([p,q])$ denote the set of pairs for which the density spectrum of the $ε$-approach time set equals $[p,q]$ when $ε$ is small and $E_{σ_{K}}([p,q])$ the set of pairs for which the density spectrum of the $ε$-approach time set converges to $[p,q]$ when $ε\rightarrow 0^+$. Then $\dim_{H} D_{σ_{K}}([p,q])=\dim_{H} E_{σ_{K}}([p,q])=2-q$. Moreover, $\mathscr{H}^{2-q}(E_{σ_{K}}([p,q]))=1$ when $q=0$ and $\mathscr{H}^{2-q}(E_{σ_{K}}([p,q]))=+\infty$ when $q>0$. Meanwhile, $\mathscr{H}^{2-q}(D_{σ_{K}}([p,q]))=+\infty$ when $q=1$ and $\mathscr{H}^{2-q}(D_{σ_{K}}([p,q]))=0$ when $q<1$.

math.DS

Freeform microfluidic networks encapsulated in laser printed three-dimensional macro-scale glass objects

Large-scale microfluidic microsystems with complex three-dimensional (3D) configurations are highly in demand by both fundamental research and industrial application, holding the potentials for fostering a wide range of innovative applications such as lab-on-a-chip and organ-on-a-chip as well as continuous-flow manufacturing of fine chemicals. However, freeform fabrication of such systems remains challenging for most of the current fabrication techniques in terms of fabrication resolution, flexibility, and achievable footprint size. Here, we report ultrashort pulse laser microfabrication of freeform microfluidic circuits with high aspect ratios and tunable diameters embedded in 3D printed glass objects. We achieve uniform microfluidic channel diameter by carefully distributing a string of extra access ports along the microfluidic channels for avoiding the over-etching in the thin microfluidic channels. After the chemical etching is completed, the extra access ports are sealed using carbon dioxide laser induced localized glass melting. We demonstrate a model hand of fused silica with a size of ~3 cm * 2.7 cm * 1.1 cm in which the whole blood vessel system is encapsulated.

physics.app-ph

Bowen entropy of generic point for fixed-point free flows

Let $(X, ϕ)$ be a compact metric flow without fixed points. We will be concerned with the entropy of flows which takes into consideration all possible reparametrizations of the flows. In this paper, by establishing the Brin-Katok's entropy formula for flows without fixed points in non-ergodic case, we prove the following result: for an ergodic $ϕ$-invariant measure $μ$, $$ h_{top}^{B}(ϕ, G_μ(ϕ))=h_μ(ϕ_{1}),$$ where $G_μ(ϕ)$ is the set of generic points for $μ$ and $h_{top}^{B}(ϕ, G_μ(ϕ))$ is the Bowen entropy on $G_μ(ϕ)$. This extends the classical result of Bowen in 1973 to fixed-point free flows. Moreover, we show that the Bowen entropy can be determined via the local entropies of measures.

math.DS