arXiv ScienceSearch

arXiv subjects

Yujie Liu

Publications and source records attributed to Yujie Liu.

At least 19 recordsLinked to original sources

Finite-Time Minimax Bounds and an Optimal Lyapunov Policy in Queueing Control

We introduce an original minimax framework for finite-time performance analysis in queueing control and propose a surprisingly simple Lyapunov-based scheduling policy with superior finite-time performance. The framework quantitatively characterizes how the expected total queue length scales with key system parameters, including the capacity of the scheduling set and the variability of arrivals and departures across queues. This characterization provides a systematic quantitative basis for evaluating and comparing scheduling policies in the finite-time regime, including nonstationary settings under certain assumptions on the model, and shows that the proposed policy provably and empirically outperforms the classical MaxWeight strategy in finite time. Within this framework, we establish three main sets of results. First, we derive minimax lower bounds on the expected total queue length for parallel-queue scheduling via a novel Brownian coupling argument. Second, we propose a new policy, LyapOpt, which minimizes the full quadratic Lyapunov drift-capturing both first- and second-order terms-and achieves optimal finite-time performance under the dominated region condition in heavy traffic while retaining classical stability guarantees. Third, we identify a key limitation of the classical MaxWeight policy, which optimizes only the first-order drift: its finite-time performance depends suboptimally on system parameters, leading to substantially larger backlogs in explicitly characterized settings. Together, these results delineate the scope and limitations of classical drift-based scheduling and motivate new queueing-control methods with rigorous finite-time guarantees.

math.OC

DAEP: Difficulty-Aware Evidence Planning for Medical Video Corpus Temporal Answer Grounding

We describe DAEP, team BIGC's submission to NLPCC 2026 Shared Task 1 Track 3: Difficulty-Aware Temporal Answer Grounding in Video Corpus (DA-TAGVC). The task requires retrieving the target video from 50 candidates and localizing the answer-supporting span. DAEP ranks videos with subtitle, visual, and procedural-context evidence, expands high-scoring anchors into temporal spans, and reranks spans for final output. Its main design is to convert the task-provided simple/complex input label into an inference-time evidence plan controlling modality weights, Top-K aggregation, boundary threshold, expansion length, and reranking strength. In the official evaluation, BIGC ranks first among ten systems with an Average score of 0.2728. Validation ablations show that visual evidence, procedural context, and difficulty-aware planning improve ranking quality, with the largest gain on complex questions.

cs.CV

Quantum Compressed Sensing CT Reconstruction Algorithm Based on Penalized Weighted Least Squares and Guided Total Variation

Objective. Existing quadratic unconstrained binary optimization (QUBO)-based sparse-view computed tomography (CT) reconstruction neglects photon-counting statistics and anatomical heterogeneity. We address both limitations within the QUBO framework.Approach. We propose a quantum compressed-sensing CT method combining penalized weighted least squares (PWLS) and guided total variation (GTV). PWLS weights projection residuals by photon-count reliability, whereas GTV uses gradients from a prior image reconstructed by the simultaneous algebraic reconstruction technique (SART) to preserve edges and suppress noise in homogeneous regions. After binary encoding, both terms form a unified QUBO model. Experiments used four 40 times 40 CT images under a 10-view fan-beam geometry with Poisson noise. Comparisons included conventional reconstruction methods, QUBO variants, gradient descent, simulated annealing, and a D-Wave hybrid quantum-classical solver.Main results. PWLS-GTV achieved the best reconstruction quality across all cases. In the representative chest case, it reached a peak signal-to-noise ratio (PSNR) of 36.64 dB, compared with 22.48 dB for SART, the best conventional baseline. GTV consistently outperformed conventional total variation. Simulated annealing and the D-Wave hybrid solver produced similar reconstructions, whereas gradient descent was ineffective. Repeated hybrid-solver runs showed stable performance.Significance. The framework incorporates photon-statistical weighting and structure-guided regularization into QUBO-based CT reconstruction without changing its quadratic form, providing a proof of concept for quantum-assisted sparse-view CT reconstruction.

cs.CV

Dual-Mapping Sparse Vector Coding for Phase Noise-Resilient Short-Packet Transmission

Sparse vector transmission (SVT) has emerged as a promising technique for ultra-reliable low-latency short-packet communications. However, existing SVT schemes typically assume negligible phase noise (PN), an assumption that rarely holds in practical wireless systems. In this paper, a dual-mapping sparse vector coding (DM-SVC) scheme is proposed for short-packet communications subject to PN. In DM-SVC, pilot symbols are mapped onto multiple non-zero blocks and data symbols onto isolated non-zero elements within a single sparse vector, thereby enabling pilot-data separation through distinct sparsity patterns rather than explicit resource partitioning. Moreover, the indices of pilot blocks convey additional information bits, further improving spectral efficiency. A basis expansion model is adopted to represent the PN process, substantially reducing the number of parameters to be estimated. Furthermore, an iterative joint PN estimation and data decoding algorithm is developed, where pilot block indices are first detected exploiting block-sparse priors, after which PN estimation and data decoding proceed iteratively. Simulation results show that DM-SVC could achieve block error rate performance close to that of perfect PN compensation, while offering improved spectral efficiency and reduced codebook storage overhead compared to state-of-the-art SVT schemes.

eess.SP

Quantum CT via Dynamic Interval Encoding and Prior-Balanced QUBO Reconstruction

Quadratic unconstrained binary optimization (QUBO)-based quantum computed tomography (CT) casts reconstruction as a binary quadratic problem for quantum annealing and hybrid quantum--classical solvers. For grayscale CT, however, image encoding is constrained by the binary-variable budget: fixed global bit-plane encodings increase QUBO size and coupling complexity as gray-level precision improves, whereas low-bit encodings introduce quantization error. We propose a QUBO-based grayscale CT reconstruction framework that combines dynamic interval encoding with prior-balanced optimization. Each refinement round encodes active pixels only within local gray-level intervals around the current estimate, and a boundary-hit-guided update rule adaptively switches between search expansion and local refinement. To improve optimization stability, the method balances projection-domain data consistency and an edge-preserving quadratic prior before forming the final QUBO. Sparse-view and limited-angle fan-beam CT experiments show that the proposed method recovers structures and gray-level distributions more faithfully than the evaluated analytic, iterative, variational, and representation-based baselines. Expressivity analysis and ablation studies further indicate that the improvement mainly arises from effective gray-level representation through dynamic local encoding and more stable data-fidelity--prior coupling. Experiments on the D-Wave hybrid binary quadratic model (BQM) solver further demonstrate that the formulation is executable on a hardware-backed hybrid quantum--classical backend.

cs.CV

Projection-Volume Fidelity Divergence: Diagnosing and Controlling Optimization Drift in Sparse-View 3D Gaussian Tomography

Sparse-view computed tomography is a severely ill-posed inverse problem, where recent 3D Gaussian Splatting methods offer an efficient explicit representation for tomographic reconstruction. However, we find that projection-domain optimization can be misleading in this setting: the rendered projections may continue to improve while the reconstructed volume deteriorates. We identify this failure mode as Projection-Volume Fidelity Divergence (PVFD), a representation-level optimization drift caused by anisotropic Gaussian deformation and view-specific primitive co-adaptation under sparse Radon constraints. To characterize this behavior, we introduce geometry- and volume-level diagnostics that measure needle-like Gaussian degeneration and the stability of the voxelized density field. Based on these observations, we propose LADES, a ground-truth-free optimization controller for sparse-view Gaussian tomography. LADES combines Linearly Annealed Dropout, which applies strong stochastic masking in early training to disrupt premature primitive co-adaptation and gradually restores full capacity for structural consolidation, with Structure-Aware Early Stopping, which terminates densification according to the saturation of Gaussian population growth rather than validation PSNR. Experiments on sparse-view CT reconstruction show that LADES improves volumetric fidelity, suppresses structural degeneration, and substantially reduces training time while maintaining competitive projection accuracy. These results suggest that robust Gaussian-based tomography requires monitoring and controlling volumetric structure, rather than optimizing projection fit alone.

cs.CV

Marginal Advantage Accumulation for Memory-Driven Agent Self-Evolution

In batch-style trace distillation, the same memory operation may receive contradictory feedback across different batches. Existing methods lack a cross-batch, operation-level evidence accumulation mechanism, making it impossible to distinguish stably effective operations from accidental hits. This paper formalizes the requirement as two structural conditions, alignability and comparability, and proposes Marginal Advantage Accumulation (MAA). MAA constructs differential signals to make them comparable across batches, accumulates signed evidence per operation via EMA, and ensures cross-batch traceability through semantic identity merging. As a post-processing architecture, MAA achieves the best results in 14 out of 16 settings across 4 benchmarks and 4 target models, consistently outperforming existing batch-level distillation baselines and matching or surpassing online alternatives in most settings, while reducing optimization-phase token consumption by approximately 75%.

cs.LG

DFT-s-OFDM with Chirping for Integrated Sensing and Communications in 6G and Beyond

The sixth generation (6G) of mobile communications and beyond is expected to enable advanced functionalities, such as integrated sensing and communication (ISAC), while involving diverse terminal/user equipment types from terrestrial to non-terrestrial networks. As waveforms are acknowledged as a fundamental technology driving 6G and beyond, this article presents a contribution in this technical domain. First, it provides an overview of several standardized communication waveforms, as well as chirp-based waveforms for radar sensing and Internet of Things (IoT) applications. This article then presents single-carrier chirping waveform: discrete Fourier transform spread orthogonal frequency division multiplexing (DFT-s-OFDM) with chirping. Its fundamental principles, key properties, performances, and advantages are examined from both communication and sensing perspectives. Finally, several future research directions are outlined to further explore its potential and opportunities for ISAC.

eess.SP

ResearchBench: Benchmarking LLMs in Scientific Discovery via Inspiration-Based Task Decomposition

Large language models (LLMs) have shown potential in assisting scientific research, yet their ability to discover high-quality research hypotheses remains unexamined due to the lack of a dedicated benchmark. To address this gap, we introduce the first large-scale benchmark for evaluating LLMs on a sufficient set of scientific discovery sub-tasks-inspiration retrieval, hypothesis composition, and hypothesis ranking-where sufficient means that perfectly solving these sub-tasks perfectly solves the overall discovery task. We develop an automated LLM-based framework that extracts critical components-research questions, background surveys, inspirations, and hypotheses-from papers across 12 disciplines, with expert validation confirming its accuracy. To prevent data contamination, we focus exclusively on publications from 2024 onward, ensuring minimal overlap with LLM pretraining data; our automated framework further enables automatic extraction of even more recent papers as LLM pretraining cutoffs advance, supporting scalable and contamination-free automatic renewal of this discovery benchmark. Our evaluation shows that, across disciplines, LLMs excel at inspiration retrieval-an out-of-distribution task-suggesting their ability to surface novel knowledge associations.

cs.CL

DyBBT: Dynamic Balance via Bandit-inspired Targeting for Dialog Policy with Cognitive Dual-Systems

Task oriented dialog systems often rely on static exploration strategies that do not adapt to dynamic dialog contexts, leading to inefficient exploration and suboptimal performance. We propose DyBBT, a novel dialog policy learning framework that formalizes the exploration challenge through a structured cognitive state space capturing dialog progression, user uncertainty, and slot dependency. DyBBT proposes a bandit inspired meta-controller that dynamically switches between a fast intuitive inference (System 1) and a slow deliberative reasoner (System 2) based on real-time cognitive states and visitation counts. Extensive experiments on single- and multi-domain benchmarks show that DyBBT achieves state-of-the-art performance in success rate, efficiency, and generalization, with human evaluations confirming its decisions are well aligned with expert judgment.

cs.CL

HiCoLoRA: Addressing Context-Prompt Misalignment via Hierarchical Collaborative LoRA for Zero-Shot DST

Zero-shot Dialog State Tracking (zs-DST) is essential for enabling Task-Oriented Dialog Systems (TODs) to generalize to new domains without costly data annotation. A central challenge lies in the semantic misalignment between dynamic dialog contexts and static prompts, leading to inflexible cross-layer coordination, domain interference, and catastrophic forgetting. To tackle this, we propose Hierarchical Collaborative Low-Rank Adaptation (HiCoLoRA), a framework that enhances zero-shot slot inference through robust prompt alignment. It features a hierarchical LoRA architecture for dynamic layer-specific processing (combining lower-layer heuristic grouping and higher-layer full interaction), integrates Spectral Joint Domain-Slot Clustering to identify transferable associations (feeding an Adaptive Linear Fusion Mechanism), and employs Semantic-Enhanced SVD Initialization (SemSVD-Init) to preserve pre-trained knowledge. Experiments on multi-domain datasets MultiWOZ and SGD show that HiCoLoRA outperforms baselines, achieving SOTA in zs-DST. Code is available at https://github.com/carsonz/HiCoLoRA.

cs.CL

DarwinTOD: LLM-driven Lifelong Self-evolution for Task-oriented Dialog Systems

Traditional task-oriented dialog systems are unable to evolve from ongoing interactions or adapt to new domains after deployment, that is a critical limitation in real-world dynamic environments. Continual learning approaches depend on episodic retraining with human curated data, failing to achieve autonomy lifelong improvement. While evolutionary computation and LLM driven self improvement offer promising mechanisms for dialog optimization, they lack a unified framework for holistic, iterative strategy refinement. To bridge this gap, we propose DarwinTOD, a lifelong self evolving dialog framework that systematically integrates these two paradigms, enabling continuous strategy optimization from a zero-shot base without task specific fine-tuning. DarwinTOD maintains an Evolvable Strategy Bank and operates through a dual-loop process: online multi-agent dialog execution with peer critique, and offline structured evolutionary operations that refine the strategy bank using accumulated feedback. This closed-loop design enables autonomous continuous improvement without human intervention. Extensive experiments show that DarwinTOD surpasses previous state-of-the-art methods and exhibits continuous performance gains throughout evolution. Our work provides a novel framework for building dialog systems with lifelong self evolution capabilities.

cs.MA

A 44-minute periodic radio transient in a supernova remnant

Long-period radio transients (LPTs) are a newly discovered class of radio emitters with periods ranging from minutes to hours. The astrophysical nature remains undetermined, particularly of LPTs with no detectable companions. We report the first evidence for a plausible supernova remnant (SNR) association with an LPT (DART J1832-0911, 2656.23+-0.15 s period), which supports a neutron star origin of such objects. The dispersion measure of this LPT, SNR's CO emission and HI absorption, and low probability of chance of alignment with field pulsars are all consistent with such an association. The source displays either phase-locked circular or nearly 100\% linear polarization, indicating its strong and geometrically stable magnetic field. No detectable optical counterpart was found, even with a 10m-class telescope. The SNR association and the stable polarization suggest that DART J1832-0911 most likely originates from a young neutron star, whose spin could have been braked by supernova's fallback materials. This discovery provides critical insights into the nature of ultra-long period transients and their link to stellar remnants.

astro-ph.HE

LLM-driven discovery for carbon allotropes with bond-network entropy

The discovery of novel carbon allotropes with tailored thermal and mechanical properties is critical for advanced thermal management. However, exploring the vast configurational space of carbon using \textit{ab initio} calculations remains computationally prohibitive. Driven by the rich topological landscape of carbon, where the competition between $sp, sp^2,$ and $sp^3$ hybridization states dictates material performance, we establish a closed-loop AI framework to explore this complex configurational space. We introduce a hybridization entropy descriptor to guide the search beyond conventional forms. Here, we establish a closed-loop AI framework that synergizes a Large Language Model (LLM) for structural generation with a Machine Learning Potential (MLP) for accelerated evaluation. Leveraging CrystaLLM to generate candidates and an iteratively refined MLP for high-fidelity validation, we screened thousands of structures to identify several stable allotropes with exotic properties. Specifically, we report ``yne-diamond C$_{12}$'' and ``yne-hex-diamond C$_{8}$'', which exhibit extreme thermal anisotropy and ultralow in-plane shear stiffness arising from their mixed $sp$-$sp^3$ hybridization. Furthermore, we discovered a complex $sp$-$sp^2$-$sp^3$ hybridized C$_{12}$ phase that combines metallic conductivity with an anomalous negative Poisson's ratio. Notably, we identified a superhard phase (C16_3) possessing a calculated Vickers hardness (103.3 GPa) exceeding that of diamond 96 GPa). Microscopic analysis reveals that thermal transport in these materials is governed by the interplay between rigid frameworks and flexible linkers. This work expands the known carbon phase space and demonstrates the efficacy of coupling generative AI with machine learning potentials for the accelerated inverse design of functional materials.

cond-mat.mtrl-sci

Dual-Mapping Sparse Vector Transmission for Short Packet URLLC

Sparse vector coding (SVC) is a promising short-packet transmission method for ultra reliable low latency communication (URLLC) in next generation communication systems. In this paper, a dual-mapping SVC (DM-SVC) based short packet transmission scheme is proposed to further enhance the transmission performance of SVC. The core idea behind the proposed scheme lies in mapping the transmitted information bits onto sparse vectors via block and single-element sparse mappings. The block sparse mapping pattern is able to concentrate the transmit power in a small number of non-zero blocks thus improving the decoding accuracy, while the single-element sparse mapping pattern ensures that the code length does not increase dramatically with the number of transmitted information bits. At the receiver, a two-stage decoding algorithm is proposed to sequentially identify non-zero block indexes and single-element non-zero indexes. Extensive simulation results verify that proposed DM-SVC scheme outperforms the existing SVC schemes in terms of block error rate and spectral efficiency.

eess.SP

SciEvalKit: An Open-source Evaluation Toolkit for Scientific General Intelligence

We introduce SciEvalKit, a unified benchmarking toolkit designed to evaluate AI models for science across a broad range of scientific disciplines and task capabilities. Unlike general-purpose evaluation platforms, SciEvalKit focuses on the core competencies of scientific intelligence, including Scientific Multimodal Perception, Scientific Multimodal Reasoning, Scientific Multimodal Understanding, Scientific Symbolic Reasoning, Scientific Code Generation, Science Hypothesis Generation and Scientific Knowledge Understanding. It supports six major scientific domains, spanning from physics and chemistry to astronomy and materials science. SciEvalKit builds a foundation of expert-grade scientific benchmarks, curated from real-world, domain-specific datasets, ensuring that tasks reflect authentic scientific challenges. The toolkit features a flexible, extensible evaluation pipeline that enables batch evaluation across models and datasets, supports custom model and dataset integration, and provides transparent, reproducible, and comparable results. By bridging capability-based evaluation and disciplinary diversity, SciEvalKit offers a standardized yet customizable infrastructure to benchmark the next generation of scientific foundation models and intelligent agents. The toolkit is open-sourced and actively maintained to foster community-driven development and progress in AI4Science.

cs.AI

Origin of shallow n-type doping in AlN and Al-rich AlGaN

Achieving efficient n-type doping in AlN, a representative ultrawide bandgap (UWBG) semiconductor, remains a longstanding challenge that limits its application in high-power electronics and deep-ultraviolet optoelectronics. Conventional dopants in AlN often introduce deep levels or form compensating complexes, leading to low free-carrier concentrations. In this work, we combine first-principles defect calculations with a structural search method tailored to explore metastable configurations to systematically investigate donor-type defects in AlN. Our results reveal that the aluminum interstitial ($Al_i$) can exhibit shallow-donor behavior in specific metastable configurations that were previously overlooked. This discovery expands the understanding of n-type dopability in AlN, and highlights the critical role of metastable defects in modulating electronic properties.

cond-mat.mtrl-sci

Metavalent Bonding-Induced Phonon Hardening and Giant Anharmonicity in BeO

The search for materials with intrinsically low thermal conductivity ($κ_L$) is critical for energy applications, yet conventional descriptors often fail to capture the complex interplay between bonding and lattice dynamics. Here, first-principles calculations are used to contrast the thermal transport in covalent zincblende (zb) and metavalent rocksalt (rs) BeO. We find that the metavalent bonding in rs-BeO enhances lattice anharmonicity, activating multi-phonon scattering channels and suppressing phonon transport. This results in an ultralow $κ_L$ of 24 W m$^{-1}$ K$^{-1}$ at 300 K, starkly contrasting with the zb phase (357 W m$^{-1}$ K$^{-1}$). Accurately modeling such strongly anharmonic systems requires explicit inclusion of temperature-dependent phonon renormalization and four-phonon scattering. These contributions, negligible in zb-BeO, are essential for high-precision calculations of the severely suppressed $κ_L$ in rs-BeO. Finally, we identify three key indicators to guide the discovery of metavalently bonded, incipient-metallic materials: (i) an NaCl-type crystal structure, (ii) large Grüneisen parameters ($\textgreater$2), and (iii) a breakdown of the Lyddane-Sachs-Teller relation. These findings provide microscopic insight into thermal transport suppression by metavalent bonding and offer a predictive framework for identifying promising thermoelectrics and phase-change materials.

cond-mat.mtrl-sci