arXiv ScienceSearch

arXiv subjects

Yichen Luo

Publications and source records attributed to Yichen Luo.

8 recordsLinked to original sources

Mitigating Gradient Pathology in PINNs through Aligned Constraint

While Physics-Informed Neural Networks (PINNs) are powerful for solving Partial Differential Equations (PDEs), their training is often paralyzed by gradient pathology. The gradients from the PDE residuals and boundary constraints oppose each other, trapping the model in local minima. Current solutions, such as adaptive weighting or hard constraints, either fail to fundamentally resolve this ill-conditioning or are limited to simple geometries. In this study, we systematically analyze the possible causes of this gradient pathology from the perspectives of loss landscapes and optimization dynamics. Based on the obtained conclusion, we propose Constraint-Aligned loss with Manifold Lifting (CAML). By reformulating all zeroth-order terms into aligned constraints, our method effectively mitigates gradient conflicts. In addition, we introduce a delay factor to help the optimizer skip the high-curvature area. Experiments demonstrate that our CAML significantly enhances numerical stability and efficiency in highly complex PINN problems. Our code is open-sourced on https://github.com/YichenLuo-0/CAML.

cs.LG

De-Idealizing De-Idealization: Beyond Full Reversal

There is a question of whether de-idealization is needed for justified use of -- for 'checking' -- idealizations. We argue that the standard philosophical account of de-idealization has become too idealized, but that this does not preclude the possibility of justificatory practices which show how models can be used to make inferences about the world. In turn, motivated by examples in physics, we provide a more expansive and practice-driven account of de-idealization by relaxing the standards for closeness to more realistic theoretical items, identifying at least three kinds of procedures for de-idealization: intra-model, inter-model, and measurement de-idealizations. These examples highlight how idealizations can be -- and indeed have been -- scrutinized within physics without appealing to the philosopher's idealized notion of de-idealization.

physics.hist-ph

FlyAOC: Evaluating Agentic Ontology Curation of Drosophila Scientific Knowledge Bases

Scientific knowledge bases accelerate discovery by curating findings from primary literature into structured, queryable formats for both human researchers and emerging AI systems. Maintaining these resources requires expert curators to search relevant papers, reconcile evidence across documents, and produce ontology-grounded annotations - a workflow that existing benchmarks, focused on isolated subtasks like named entity recognition or relation extraction, do not capture. We present FlyBench to evaluate AI agents on end-to-end agentic ontology curation from scientific literature. Given only a gene symbol, agents must search and read from a corpus of 16,898 full-text papers to produce structured annotations: Gene Ontology terms describing function, expression patterns, and historical synonyms linking decades of nomenclature. The benchmark includes 7,397 expert-curated annotations across 100 genes drawn from FlyBase, the Drosophila (fruit fly) knowledge base. We evaluate four baseline agent architectures: memorization, fixed pipeline, single-agent, and multi-agent. We find that architectural choices significantly impact performance, with multi-agent designs outperforming simpler alternatives, yet scaling backbone models yields diminishing returns. All baselines leave substantial room for improvement. Our analysis surfaces several findings to guide future development; for example, agents primarily use retrieval to confirm parametric knowledge rather than discover new information. We hope FlyBench will drive progress on retrieval-augmented scientific reasoning, a capability with broad applications across scientific domains.

cs.AI

Resisting Manipulative Bots in Meme Coin Copy Trading: A Multi-Agent Approach with Chain-of-Thought Reasoning

Copy trading has become the dominant entry strategy in meme coin markets. However, due to the market's extremely illiquid and volatile nature, the strategy exposes an exploitable attack surface: adversaries deploy manipulative bots to front-run trades, conceal positions, and fabricate sentiment, systematically extracting value from na\"ive copiers at scale. Despite its prevalence, bot-driven manipulation remains largely unexplored, and no robust defensive framework exists. We propose a manipulation-resistant copy-trading system based on a multi-agent architecture powered by a multi-modal large language model (LLM) and chain-of-thought (CoT) reasoning. Our approach outperforms zero-shot and most statistic-driven baselines in prediction accuracy as well as all baselines in economic performance, achieving an average copier return of 3% per meme coin investment under realistic market frictions. Overall, our results demonstrate the effectiveness of agent-based defenses and predictability of trader profitability in adversarial meme coin markets, providing a practical foundation for robust copy trading.

cs.AI

Stretchable and self-adhesive triboelectric sensor for real-time musculoskeletal monitoring and personalized recovery

Recent advances in medical diagnostics have highlighted the importance of wearable technologies for continuous and real-time physiological monitoring. In this study, we introduce a flexible, self-powered triboelectric nanogenerator (MB-TENG) engineered from commercially available medical elastic bandages for biomechanical sensing during rehabilitation and gait analysis. Leveraging the porous and skin-friendly properties of the bandage combined with a PTFE film, the MB-TENG delivers robust electrical performance, achieving a peak open-circuit voltage (VOC) of 122~V, a short-circuit current (ISC) of 25~$\mu$A, and a transferred charge (QSC) of 110~nC, while maintaining long-term stability across 40{,}000 mechanical cycles. Its inherent self-adhesive property allows for multi-layer assembly without extra bonding agents, and mechanical stretching enhances output, enabling dual configurability. A stacked design further improves the power capacity, supporting applications in wearable medical electronics. The MB-TENG device seamlessly conforms to joint surfaces and foot regions, providing accurate detection of motion states and abnormal gait patterns. These features underscore the MB-TENG's potential as a low-cost, scalable platform for personalized rehabilitation, injury monitoring, and early musculoskeletal diagnosis.

physics.med-ph

MMET: A Multi-Input and Multi-Scale Transformer for Efficient PDEs Solving

Partial Differential Equations (PDEs) are fundamental for modeling physical systems, yet solving them in a generic and efficient manner using machine learning-based approaches remains challenging due to limited multi-input and multi-scale generalization capabilities, as well as high computational costs. This paper proposes the Multi-input and Multi-scale Efficient Transformer (MMET), a novel framework designed to address the above challenges. MMET decouples mesh and query points as two sequences and feeds them into the encoder and decoder, respectively, and uses a Gated Condition Embedding (GCE) layer to embed input variables or functions with varying dimensions, enabling effective solutions for multi-scale and multi-input problems. Additionally, a Hilbert curve-based reserialization and patch embedding mechanism decrease the input length. This significantly reduces the computational cost when dealing with large-scale geometric models. These innovations enable efficient representations and support multi-scale resolution queries for large-scale and multi-input PDE problems. Experimental evaluations on diverse benchmarks spanning different physical fields demonstrate that MMET outperforms SOTA methods in both accuracy and computational efficiency. This work highlights the potential of MMET as a robust and scalable solution for real-time PDE solving in engineering and physics-based applications, paving the way for future explorations into pre-trained large-scale models in specific domains. This work is open-sourced at https://github.com/YichenLuo-0/MMET.

cs.LG

LLM-Powered Multi-Agent System for Automated Crypto Portfolio Management

Cryptocurrency portfolio management requires the fusion of heterogeneous multi-modal signals, including structured price and on-chain time series, unstructured news text, and technical indicators, under high-volatility and real-time constraints. While deep learning approaches show predictive capability, their opacity limits practical adoption, and single large language model (LLM) agents struggle to process the breadth of modality-specific inputs needed for robust decision-making. We propose a multi-agent system (MAS) framework in which three modality-specialised agents, a Crypto Agent for market dynamics, a News Agent for weekly news sentiment, and a Trading Agent for signal fusion and portfolio execution, decompose the task across three communication architectures: hierarchical, collaborative, and debate. We evaluate four capability configurations: zero-shot, chain-of-thought (CoT), retrieval-augmented generation (RAG), and skill-augmented. In a 52-week backtest over calendar year 2025 across the top 15 L1 blockchain native cryptocurrencies by market capitalisation as of January 2025, the best configuration, Hierarchical (Skill), achieves a cumulative return of 133.52% and a Sharpe ratio of 1.502, outperforming single-agent variants, passive benchmarks, and deep learning baselines. An ablation study identifies the Crypto Agent as the most critical component, with its removal reducing cumulative return by 42.57 percentage points. A cross-model comparison further shows that MAS outperforms the single-agent baseline under GPT-4o, GPT-5, and Claude Sonnet 4.5, suggesting that the benefit of multi-agent coordination is model-agnostic. Unlike black-box deep learning models, every portfolio decision is traceable to explicit agent reasoning, offering an interpretable and effective approach to multi-modal cryptocurrency portfolio management.

q-fin.TR

Piercing the Veil of TVL: DeFi Reappraised

Total value locked (TVL) is widely used to measure the size and popularity of decentralized finance (DeFi). However, TVL can be easily manipulated and inflated through "double counting" activities such as wrapping and leveraging. As existing methodologies addressing double counting are inconsistent and flawed, we propose a new framework, termed "total value redeemable (TVR)", to assess the true underlying value of DeFi. Our formal analysis reveals how DeFi's complex network spreads financial contagion via derivative tokens, increasing TVL's sensitivity to external shocks. To quantify double counting, we construct the DeFi multiplier, which mirrors the money multiplier in traditional finance (TradFi). This measurement reveals substantial double counting in DeFi, finding that the gap between TVL and TVR reached \$139.87 billion during the peak of DeFi activity on December 2, 2021, with a TVL-to-TVR ratio of approximately 2. We conduct sensitivity tests to evaluate the stability of TVL compared to TVR, demonstrating the former's significantly higher level of instability than the latter, especially during market downturns: A 25% decline in the price of Ether (ETH) leads to a \$1 billion greater non-linear decrease in TVL compared to TVR via the liquidations triggered by derivative tokens. We also document that the DeFi money multiplier is positively correlated with crypto market indicators and negatively correlated with macroeconomic indicators. Overall, our findings suggest that TVR is more reliable and stable than TVL.

q-fin.GN