arXiv ScienceSearch

arXiv subjects

Yan Dai

Publications and source records attributed to Yan Dai.

At least 19 recordsLinked to original sources

Decreasing Digital Distraction in College Students: Associated Online Learning Strategies Identified by Unsupervised Data Mining Approaches

The proliferation of digital tools in education offers numerous benefits but also introduces significant challenges, notably digital distractions that hinder academic performance, especially in online learning contexts. This study employed unsupervised data mining techniques, specifically association rule mining and clustering analysis, to identify effective learning strategies associated with lower levels of digital distractions among college students. Data from 530 participants revealed that self-regulated learning strategies (i.e., goal setting, environment structuring, and time management) co-occurred most consistently with lower digital distractions. Additionally, learner-instructor and learner-content engagement strategies, as well as technical competencies, also tended to appear in the same profiles as lower distraction. Interestingly, reliance on peer help-seeking and learner-learner engagement strategies appeared less often in those lower distraction profiles. These findings offer actionable implications for educators to design targeted interventions that foster focused and productive online learning environments.

cs.CY

On the Riemann Boundary Value Problem for Poly- and Meta-hyperanalytic Function Spaces over d-summable Curves

Hyperanalytic functions, in the sense established by the mathematician Avron Douglis, are Douglis algebra-valued functions defined via a hypercomplex structure rather than the standard Cauchy-Riemann equations characteristic of traditional complex analysis. The classes of polyhyperanalytic and meta-hyperanalytic functions represent advanced generalizations of Douglis's analysis. They are employed in the study of partial differential equations and elasticity, extending the concept of the classical holomorphic function through higher-order iterations and non-homogeneous terms. The aim of this work is to find solvability conditions for a fundamental Riemann-type boundary value problem for spaces of poly-hyperanalytic and meta-hyperanalytic functions defined on an open, bounded, simply connected subset of the complex plane, where the boundary need only be a closed d-summable curve. In fractal geometry, d-summability is a geometric property used to define the boundaries of fractal domains, enabling advanced mathematical integration and calculus on complex structures defined by Jenny Harrison and Alec Norton.

math.CV

Physics-Assisted Deep Learning Denoising for Stabilized IMPULSED dMRI Microenvironment Parameter Fitting

Diffusion-weighted MRI (dMRI) is a powerful tool for quantifying cellular microenvironment parameters. This study proposes a physics-assisted deep learning (DL)-based denoising framework designed to enhance dMRI signal quality and improve the robustness of subsequent biophysical model fitting. A dataset of paired noise-free and Rician-noise-corrupted dMRI signals was generated using the IMPULSED-dMRI signal model. Three denoising architectures were evaluated: Convolutional Neural Networks (CNN), Multilayer Perceptron (MLP), and Long Short-Term Memory (LSTM) networks. Denoised signals were then fitted to estimate cell diameter $d$, intracellular volume fraction $V_{\mathrm{in}}$, and extracellular apparent diffusion coefficient $D_\mathrm{ex}$. DL-based processing substantially improved dMRI signal denoising. The MLP and LSTM achieved similar performance, with the LSTM slightly better overall, and both outperformed the CNN. In the subsequent model fitting step, the LSTM produced modest reductions in parameter MAE. The dominant benefit was fitting stabilization, with the overall fitting failure rate reduced from 57.6\% to 17.7\%. The proposed framework improves dMRI signal quality and stabilizes subsequent IMPULSED-based microenvironmental parameter fitting.

physics.med-ph

An integrated diffusion-weighted imaging processing and interpretation platform for MR-guided radiotherapy

Background: Magnetic resonance imaging-guided linear accelerators (MR-Linacs) allow diffusion-weighted imaging (DWI) to be acquired at every treatment fraction, but converting these low-signal-to-noise-ratio acquisitions into clinical decisions requires both reliable quantitative processing and an interpretation that reconciles a scattered and often contradictory literature. Purpose: To describe and evaluate an integrated, web-based platform that carries raw MR-Linac DWI to a structured, literature-grounded clinical interpretation, and to assess its retrieval-augmented generation (RAG) interpretation module by independent expert rating. Methods: The platform couples a deep-learning processing pipeline, comprising distortion correction, denoising, and intravoxel incoherent motion (IVIM)/apparent diffusion coefficient (ADC) fitting, with longitudinal region-of-interest analysis and a RAG interpretation agent. The agent reasons over a two-layer knowledge base of curated publications (a structured catalog index plus line-indexed full text), delegates arithmetic to deterministic tools, and is designed to trace each statement to a source document, section, and line range. One medical physicist and one physician independently rated the agent's reports for nine longitudinal glioblastoma cases on a 1-5 scale across three metrics: clinical-reasoning soundness, literature-citation quality, and overall clinical utility. Results: Across 54 ratings, the pooled mean was 4.65 +/- 0.80, with 93% of ratings >= 4; metric means were 4.6 (reasoning), 4.5 (citation), and 4.8 (utility), and raters agreed within one point on 85% of paired ratings. Conclusions: A single platform can integrate MR-Linac DWI post-processing with traceable, expert-evaluated clinical interpretation, while highlighting the safeguards needed to verify LLM-generated reasoning in radiation oncology.

physics.med-ph

Policy Regret for Embedding Model Routing: Contextual Bandits with Low-Rank Experts

Modern recommendation systems increasingly rely on dynamically routing diverse queries to multiple embedding models. Despite its practical significance, this problem remains poorly understood under realistic conditions like adversarial queries, bandit feedback, and limited observability of models. We formalize embedding model routing as an adversarial contextual linear bandit with low-rank experts, where contexts are queries, actions are items, and experts are the embedding models working on low-rank latent representation spaces. We first establish that standard regret notions suffer from structural misspecification or statistical intractability, and we identify a log-quadratic policy class that is expressive enough to capture query-dependent model routing, yet structured enough to allow efficient online learning. Second, we propose a policy gradient algorithm called Hypentropy Policy Gradient (HPG). It provably adapts to the unknown low-rank structure under incomplete information and attains $\tilde{\mathcal O}(s\sqrt{M T})$ linearized policy regret -- where $s, M$, and $T$ are the intrinsic rank of the experts, the number of models, and the number of rounds -- thus avoiding a curse of dimensionality. Finally, we also provide an computationally efficient and parameter-free implementation of HPG.

cs.LG

Market Design for AI: Beyond the Copyright Binary

How can we design a market of human-generated content for use in training AI models that both enables technological progress and preserves individual incentives for high-quality content creation? Existing approaches take polar positions: a "free-for-all" model based on fair use and a "strong intellectual property rights" model. We show that both fail: Free-for-all does not compensate creators, and---by modeling as a static Stackelberg game---strong intellectual property rights also underpower creative incentives. We find this especially true for more innovative creators, a phenomenon we term the "originality penalty." Extending this insight to a dynamic model, we find another market failure undermining AI model performance, even for an initially good model: Such a model induces greater reliance by humans on AI-assisted creation, resulting in homogenized content feeding back into training, which degrades the model performance---a "curse of precision." We further propose a market design with a data intermediary negotiating collectively with the AI firm and subsidizing innovative contributions, thus restoring efficiency.

econ.TH

Investigating the Uncertainty of Cellular Microenvironment Parameter Estimations via Diffusion MRI Cytometry

This study aims to identify cell microenvironment parameters that can be robustly estimated from IMPULSED diffusion MRI signals and to develop a reliable mapping-based estimation framework. Diffusion MRI signals were simulated using the established IMPULSED model with one pulsed gradient spin echo sequence and two oscillating gradient spin echo sequences at different frequencies. Five cellular parameters were considered: cell diameter ($d$), intracellular diffusion coefficient ($D_{in}$), intracellular volume fraction ($V_{in}$), extracellular diffusion coefficient ($D_{ex}$), and the frequency-dependent slope of $D_{ex}$ ($\beta_{ex}$). Parameter uncertainty was quantified using Jacobian-based sensitivity analysis at an SNR of 30, representing clinically achievable conditions on a 1.5T MRI scanner. To enable direct parameter mapping, signals were logarithmically transformed, reduced in dimension using principal component analysis, and then used to estimate parameters with linear regression, fourth-order polynomial regression, and a fully connected four-layer neural network. Model validation was performed in vitro using MC38 cell lines. Uncertainty analysis identified $d$, $V_{in}$, and $D_{ex}$ as robustly derivable parameters, each with relative uncertainty below 1.0. Among the tested models, the four-layer neural network performed best, with mean absolute errors of 1.7 $\mu$m for $d$, 5.06% for $V_{in}$, and 0.28 $\mu$m$^2$/ms for $D_{ex}$. In vitro validation showed a 6.7% error in cell diameter estimation. These results demonstrate that IMPULSED dMRI can support robust estimation of key cell microenvironment parameters and provide a practical framework for noninvasive assessment of tumor microenvironment changes during radiation therapy response monitoring.

physics.med-ph

Optimizing IMPULSED Acquisition Protocols for Clinical 3T Scanners Through Bayesian Experimental Design

To optimize diffusion MRI acquisition protocols for IMPULSED model at clinical 3T scanner using Bayesian experimental design, enabling accurate cellular-scale parameter estimation under realistic scan time and scanner hardware constraints. Expected Information Gain (EIG) was used as the optimization objective to maximize the information content of acquired measurements for IMPULSED model fitting. Bayesian optimization with Gaussian process surrogates efficiently searched the high-dimensional acquisition parameter space, including pulse types (PGSE, OGSEn1, and OGSEn2), diffusion times, and b-values. Optimized protocols were systematically evaluated against a heuristically designed baseline protocol through simulation studies assessing classification accuracy and parameter estimation performance across SNR levels of 5-40. Robustness to optimization assumptions was examined by varying prior distributions and assumed SNR. In-vivo validation was performed using canine tumor data acquired at 3T. The optimized protocol eliminated OGSEn2 acquisitions, concentrated measurements at high b-values, employing concurrently optimized diffusion timing. Compared to the baseline protocol, the optimized design achieved superior classification accuracy for distinguishing cell populations and reduced parameter estimation error across biologically relevant parameter ranges at various SNRs. Performance advantages were consistent across diverse optimization scenarios, demonstrating robustness to prior knowledge and noise assumptions. In-vivo parameter maps showed substantially improved quality and smoothness. Bayesian optimization substantially improves IMPULSED acquisition design for clinical 3T scanners. This principled, algorithm-agnostic framework enables accurate diffusion MRI cytometry under clinical constraints, with potential applications to tumor characterization and treatment monitoring.

physics.med-ph

R-Index: A Robust Metric for IVIM Parameter Estimation on Clinical MRI Scanners

Background: Intravoxel Incoherent Motion (IVIM) model characterizes both water diffusion and perfusion in tissues, providing quantitative biomarkers valuable for tumor tissue characterization. However, parameter estimation based on this model is challenging due to its ill-posed nature, resulting in poor reproducibility, particularly at low signal to noise ratios (SNRs) in a clinic scenario. Purpose: This study analyzes the uncertainty of IVIM model fitting, quantifies parameter collinearity, and introduces a new index with enhanced robustness to enhance clinical applicability of the IVIM model. Study Type: Prospective. Population: One healthy volunteer. Field Strength/Sequence: 1.5T; single-shot EPI DWI. Assessment: The probability distributions of estimated IVIM parameters were evaluated across a clinically relevant range. Collinearity among parameters was assessed and a new metric, the R-index, was proposed. The R-index linearly combines individual IVIM parameters to mitigate collinearity and reduce estimation uncertainty. Simulation and a volunteer study was conducted to validate the presence of parameter collinearity and to assess the robustness of the R-index. Statistical Tests: N/A Results: In simulation studies with a typical clinical setting (SNR = 20), normalized IVIM parameters exhibited mean standard deviations ranging from 0.107 to 0.269, while the R-index showed a reduced deviation of 0.064. Repeated scans in a healthy volunteer confirmed the presence of parameter collinearity, with 32% of voxels exhibiting statistically significant correlations (p < 0.05) among fitted IVIM parameters, and a mean Pearson correlation coefficient of r = -0.96. Data Conclusion: The R-index provides a robust metric for IVIM model fitting under low SNR conditions typical of clinical MRI, offering improved reproducibility and potential for broader clinical applicability.

physics.med-ph

SAM2-Aug: Prior knowledge-based Augmentation for Target Volume Auto-Segmentation in Adaptive Radiation Therapy Using Segment Anything Model 2

Purpose: Accurate tumor segmentation is vital for adaptive radiation therapy (ART) but remains time-consuming and user-dependent. Segment Anything Model 2 (SAM2) shows promise for prompt-based segmentation but struggles with tumor accuracy. We propose prior knowledge-based augmentation strategies to enhance SAM2 for ART. Methods: Two strategies were introduced to improve SAM2: (1) using prior MR images and annotations as contextual inputs, and (2) improving prompt robustness via random bounding box expansion and mask erosion/dilation. The resulting model, SAM2-Aug, was fine-tuned and tested on the One-Seq-Liver dataset (115 MRIs from 31 liver cancer patients), and evaluated without retraining on Mix-Seq-Abdomen (88 MRIs, 28 patients) and Mix-Seq-Brain (86 MRIs, 37 patients). Results: SAM2-Aug outperformed convolutional, transformer-based, and prompt-driven models across all datasets, achieving Dice scores of 0.86(liver), 0.89(abdomen), and 0.90(brain). It demonstrated strong generalization across tumor types and imaging sequences, with improved performance in boundary-sensitive metrics. Conclusions: Incorporating prior images and enhancing prompt diversity significantly boosts segmentation accuracy and generalizability. SAM2-Aug offers a robust, efficient solution for tumor segmentation in ART. Code and models will be released at https://github.com/apple1986/SAM2-Aug.

eess.IV

Efficiency, Feasibility, and Incentive-Awareness in Constrained Online Resource Allocation

We study the dynamic allocation of indivisible resources to strategic agents under long-term constraints, where the planner aims to maximize social welfare, satisfy multiple constraints, and elicit near-truthful reports. We find standard primal-dual methods fragile in this setting: agents easily manipulate their reports to distort dual variables, sacrificing social efficiency for individual utility. To address this, we propose the Incentive-Aware Primal-Dual (IAPD) framework. On the primal side, we integrate three components to suppress manipulation: a VCG-based payment neutralizes immediate misreporting benefits, while epoch-based lazy updates and random exploration together ensure potential future gains are outweighed by immediate penalties. On the dual side, to overcome a learning barrier due to lazy updates -- which we call the "price of incentives" -- we design a novel optimistic online learning algorithm, O-FTRL-FP. It utilizes a fixed-point oracle to resolve the circular dependency between optimistic dual variables and the resulting allocations. Ultimately, our mechanism attains $\tilde{\mathcal O}(\sqrt T)$ social welfare regret, satisfies all long-term constraints, and induces a near-truthful equilibrium. It also smoothly generalizes to multi-unit multi-demand allocation problems. Notably, this $\tilde{\mathcal O}(\sqrt T)$ regret near-matches the non-strategic $\Omega(\sqrt T)$ lower bound, demonstrating that incentive-awareness can be accommodated at nearly no cost.

cs.GT

Non-Monetary Mechanism Design without Priors: Achieving Efficiency via Adaptive Costly Audits

We study repeated resource allocation with strategic agents, where monetary transfers are disallowed and the planner has no prior information on agents' utility distributions. Inspired by the costly state verification literature, we assume the planner can request costly audits on the winning agent after allocation, revealing their true utility but without the ability to revoke the allocation. We design a mechanism achieving $T$-independent $\mathcal O(K^2)$ regret in social welfare while requesting $\mathcal O(K^3 \log T)$ audits in expectation, where $K$ is the number of agents and $T$ is the number of rounds. We further show an $\Omega(K)$ lower bound on the regret and an $\Omega(1)$ lower bound on the number of audits required for low regret. We also generalize our mechanism and analysis to imperfect audit models. Algorithmically, we show that incentivizing truthful behavior relies on accurately estimating agents' truthful winning probability online. To achieve this, we impose future punishments via adaptive audits; we also introduce an incentive-aligned flagging component allowing agents to flag biased estimates, which we prove is in their best interest. Analytically, without distributional information, the revelation principle cannot dictate a truth-telling equilibrium. Instead, we characterize a Perfect Bayesian Equilibrium via a reduction to an auxiliary game with only benign strategies. The technical tools developed herein can be of independent interest for other robust mechanism design problems where the revelation principle is inapplicable.

cs.GT

The Theoretical Study of $\Sigma^{+} p \to\Lambda a_0^{+} p$ Reaction

We conducted a theoretical study on the process $\Sigma^{+} p \to\Lambda a_0^{+} p$ based on an effective Lagrangian approach. This model encompasses the excitation of intermediate states leading to the production of $\Delta(1920)$ through $\pi^{+}$ and $K^{+}$ meson exchanges between the initial $\Sigma^{+}$ baryon and the initial proton $p$, as well as the generation of $\Delta(1940)$ via $\pi^{+}$ and $\rho^{+}$ meson exchanges. We provide predictions for the total and differential cross sections and discuss the potential impacts of cutoff parameters, off-shell effects, and branching ratios on the $\Delta^{\ast} \to a_0 p$ decay. Given the dominance of $\Delta(1940)$ in the region close to the reaction threshold, this reaction is considered an ideal platform for deeply exploring the unique properties of the $\Delta(1940)$ resonance. Through this approach, we can gain a more precise understanding of the intrinsic characteristics of this particle.

hep-ph

uniINF: Best-of-Both-Worlds Algorithm for Parameter-Free Heavy-Tailed MABs

In this paper, we present a novel algorithm, uniINF, for the Heavy-Tailed Multi-Armed Bandits (HTMAB) problem, demonstrating robustness and adaptability in both stochastic and adversarial environments. Unlike the stochastic MAB setting where loss distributions are stationary with time, our study extends to the adversarial setup, where losses are generated from heavy-tailed distributions that depend on both arms and time. Our novel algorithm `uniINF` enjoys the so-called Best-of-Both-Worlds (BoBW) property, performing optimally in both stochastic and adversarial environments without knowing the exact environment type. Moreover, our algorithm also possesses a Parameter-Free feature, i.e., it operates without the need of knowing the heavy-tail parameters $(\sigma, \alpha)$ a-priori. To be precise, uniINF ensures nearly-optimal regret in both stochastic and adversarial environments, matching the corresponding lower bounds when $(\sigma, \alpha)$ is known (up to logarithmic factors). To our knowledge, uniINF is the first parameter-free algorithm to achieve the BoBW property for the heavy-tailed MAB problem. Technically, we develop innovative techniques to achieve BoBW guarantees for Parameter-Free HTMABs, including a refined analysis for the dynamics of log-barrier, an auto-balancing learning rate scheduling scheme, an adaptive skipping-clipping loss tuning technique, and a stopping-time analysis for logarithmic regret.

cs.LG

Adversarial Network Optimization under Bandit Feedback: Maximizing Utility in Non-Stationary Multi-Hop Networks

Stochastic Network Optimization (SNO) concerns scheduling in stochastic queueing systems. It has been widely studied in network theory. Classical SNO algorithms require network conditions to be stationary with time, which fails to capture the non-stationary components in many real-world scenarios. Many existing algorithms also assume knowledge of network conditions before decision, which rules out applications where unpredictability presents. Motivated by these issues, we consider Adversarial Network Optimization (ANO) under bandit feedback. Specifically, we consider the task of *i)* maximizing some unknown and time-varying utility function associated to scheduler's actions, where *ii)* the underlying network is a non-stationary multi-hop one whose conditions change arbitrarily with time, and *iii)* only bandit feedback (effect of actually deployed actions) is revealed after decisions. Our proposed `UMO2` algorithm ensures network stability and also matches the utility maximization performance of any "mildly varying" reference policy up to a polynomially decaying gap. To our knowledge, no previous ANO algorithm handled multi-hop networks or achieved utility guarantees under bandit feedback, whereas ours can do both. Technically, our method builds upon a novel integration of online learning into Lyapunov analyses: To handle complex inter-dependencies among queues in multi-hop networks, we propose meticulous techniques to balance online learning and Lyapunov arguments. To tackle the learning obstacles due to potentially unbounded queue sizes, we design a new online linear optimization algorithm that automatically adapts to loss magnitudes. To maximize utility, we propose a bandit convex optimization algorithm with novel queue-dependent learning rate scheduling that suites drastically varying queue lengths. Our new insights in online learning can be of independent interest.

math.OC

Refined Sample Complexity for Markov Games with Independent Linear Function Approximation

Markov Games (MG) is an important model for Multi-Agent Reinforcement Learning (MARL). It was long believed that the "curse of multi-agents" (i.e., the algorithmic performance drops exponentially with the number of agents) is unavoidable until several recent works (Daskalakis et al., 2023; Cui et al., 2023; Wang et al., 2023). While these works resolved the curse of multi-agents, when the state spaces are prohibitively large and (linear) function approximations are deployed, they either had a slower convergence rate of $O(T^{-1/4})$ or brought a polynomial dependency on the number of actions $A_{\max}$ -- which is avoidable in single-agent cases even when the loss functions can arbitrarily vary with time. This paper first refines the AVLPR framework by Wang et al. (2023), with an insight of designing *data-dependent* (i.e., stochastic) pessimistic estimation of the sub-optimality gap, allowing a broader choice of plug-in algorithms. When specialized to MGs with independent linear function approximations, we propose novel *action-dependent bonuses* to cover occasionally extreme estimation errors. With the help of state-of-the-art techniques from the single-agent RL literature, we give the first algorithm that tackles the curse of multi-agents, attains the optimal $O(T^{-1/2})$ convergence rate, and avoids $\text{poly}(A_{\max})$ dependency simultaneously.

cs.LG

Understanding Adam Optimizer via Online Learning of Updates: Adam is FTRL in Disguise

Despite the success of the Adam optimizer in practice, the theoretical understanding of its algorithmic components still remains limited. In particular, most existing analyses of Adam show the convergence rate that can be simply achieved by non-adative algorithms like SGD. In this work, we provide a different perspective based on online learning that underscores the importance of Adam's algorithmic components. Inspired by Cutkosky et al. (2023), we consider the framework called online learning of updates/increments, where we choose the updates/increments of an optimizer based on an online learner. With this framework, the design of a good optimizer is reduced to the design of a good online learner. Our main observation is that Adam corresponds to a principled online learning framework called Follow-the-Regularized-Leader (FTRL). Building on this observation, we study the benefits of its algorithmic components from the online learning perspective.

cs.LG

The Crucial Role of Normalization in Sharpness-Aware Minimization

Sharpness-Aware Minimization (SAM) is a recently proposed gradient-based optimizer (Foret et al., ICLR 2021) that greatly improves the prediction performance of deep neural networks. Consequently, there has been a surge of interest in explaining its empirical success. We focus, in particular, on understanding the role played by normalization, a key component of the SAM updates. We theoretically and empirically study the effect of normalization in SAM for both convex and non-convex functions, revealing two key roles played by normalization: i) it helps in stabilizing the algorithm; and ii) it enables the algorithm to drift along a continuum (manifold) of minima -- a property identified by recent theoretical works that is the key to better performance. We further argue that these two properties of normalization make SAM robust against the choice of hyper-parameters, supporting the practicality of SAM. Our conclusions are backed by various experiments.

cs.LG