arXiv ScienceSearch

arXiv subjects

Jiaming Ma

Publications and source records attributed to Jiaming Ma.

12 recordsLinked to original sources

Science Earth: Towards A Planet-Scale Operating System for AI-Native Scientific Discovery

Scientific discovery demands intelligence, perseverance, and serendipity across vast search spaces. Today, top scientific capabilities remain siloed--one AI system for biological analysis, another for clinical reasoning, mathematical derivation, or materials simulation--and no pre-designed team can anticipate every skill a question will need. Science Earth is a planet-scale scientific runtime in which any capability--a simulation cluster, a wet-lab robot, a proof engine, a single-cell pipeline--can connect to any other, with collaboration structure emerging from the question itself. Its underlying EACN protocol lets capabilities discover one another, negotiate task ownership, and adjudicate across incompatible evidentiary standards without prior knowledge of who will meet whom. This shifts the organizing challenge from workflow design to open-ended connectivity. Two runs validate this under structurally distinct conditions. In a trans-Pacific higher-order Kuramoto synchronization study, agents identified and corrected a closure-ratio assumption in Ott-Antonsen analytic theory that fails outside the Lorentzian limit, within thirty minutes. In an eight-agent single-cell run on the 4.88M-cell Kang 2024 pan-cancer atlas, heterogeneous capabilities coupled over a 64.9-hour window with one structural external instruction, producing three new result layers and anchoring findings against an independent wet-lab study on an adjacent CCR8- TIGIT+ Treg subset. These cases are a first empirical reading, not a benchmark sweep. They show that when AI capabilities are truly connectable and coordination emerges from the problem, scientific reasoning becomes a distributed, self-correcting process--a step towards scaling AI-native discovery to the planet.

cs.AI

ResearchEVO: An End-to-End Framework for Automated Scientific Discovery and Documentation

An important recurring pattern in scientific breakthroughs is a two-stage process: an initial phase of undirected experimentation that yields an unexpected finding, followed by a retrospective phase that explains why the finding works and situates it within existing theory. We present ResearchEVO, an end-to-end framework that computationally instantiates this discover-then-explain paradigm. The Evolution Phase employs LLM-guided bi-dimensional co-evolution -- simultaneously optimizing both algorithmic logic and overall architecture -- to search the space of code implementations purely by fitness, without requiring any understanding of the solutions it produces. The Writing Phase then takes the best-performing algorithm and autonomously generates a complete, publication-ready research paper through sentence-level retrieval-augmented generation with explicit anti-hallucination verification and automated experiment design. To our knowledge, ResearchEVO is the first system to cover this full pipeline end to end: no prior work jointly performs principled algorithm evolution and literature-grounded scientific documentation. We validate the framework on two cross-disciplinary scientific problems -- Quantum Error Correction using real Google quantum hardware data, and Physics-Informed Neural Networks -- where the Evolution Phase discovered human-interpretable algorithmic mechanisms that had not been previously proposed in the respective domain literatures. In both cases, the Writing Phase autonomously produced compilable LaTeX manuscripts that correctly grounded these blind discoveries in existing theory via RAG, with zero fabricated citations.

cs.AI

A Level Set Method with Secant Iterations for the Least-Squares Constrained Nuclear Norm Minimization

We present an efficient algorithm for least-squares constrained nuclear norm minimization, a computationally challenging problem with broad applications. Our approach combines a level set method with secant iterations and a proximal generation method. As a key theoretical contribution, we establish the nonsingularity of the Clarke generalized Jacobian for a general class of projection norm functions over closed convex sets. This property and the (strong) semismoothness of our value function yield fast local convergence of the secant method. For the resulting nuclear norm regularized subproblems, we develop a proximal generation method that exploits low-rank structures without compromising convergence. Extensive numerical experiments demonstrate the superior performance of our approach compared to state-of-the-art methods.

math.OC

PHAT: Modeling Period Heterogeneity for Multivariate Time Series Forecasting

While existing multivariate time series forecasting models have advanced significantly in modeling periodicity, they largely neglect the periodic heterogeneity common in real-world data, where variables exhibit distinct and dynamically changing periods. To effectively capture this periodic heterogeneity, we propose PHAT (Period Heterogeneity-Aware Transformer). Specifically, PHAT arranges multivariate inputs into a three-dimensional "periodic bucket" tensor, where the dimensions correspond to variable group characteristics with similar periodicity, time steps aligned by phase, and offsets within the period. By restricting interactions within buckets and masking cross-bucket connections, PHAT effectively avoids interference from inconsistent periods. We also propose a positive-negative attention mechanism, which captures periodic dependencies from two perspectives: periodic alignment and periodic deviation. Additionally, the periodic alignment attention scores are decomposed into positive and negative components, with a modulation term encoding periodic priors. This modulation constrains the attention mechanism to more faithfully reflect the underlying periodic trends. A mathematical explanation is provided to support this property. We evaluate PHAT comprehensively on 14 real-world datasets against 18 baselines, and the results show that it significantly outperforms existing methods, achieving highly competitive forecasting performance. Our sources is available at GitHub.

cs.LG

To See Far, Look Close: Evolutionary Forecasting for Long-term Time Series

The prevailing Direct Forecasting (DF) paradigm dominates Long-term Time Series Forecasting (LTSF) by forcing models to predict the entire future horizon in a single forward pass. While efficient, this rigid coupling of output and evaluation horizons necessitates computationally prohibitive re-training for every target horizon. In this work, we uncover a counter-intuitive optimization anomaly: models trained on short horizons-when coupled with our proposed Evolutionary Forecasting (EF) paradigm-significantly outperform those trained directly on long horizons. We attribute this success to the mitigation of a fundamental optimization pathology inherent in DF, where conflicting gradients from distant futures cripple the learning of local dynamics. We establish EF as a unified generative framework, proving that DF is merely a degenerate special case of EF. Extensive experiments demonstrate that a singular EF model surpasses task-specific DF ensembles across standard benchmarks and exhibits robust asymptotic stability in extreme extrapolation. This work propels a paradigm shift in LTSF: moving from passive Static Mapping to autonomous Evolutionary Reasoning.

cs.LG

A General ReLearner: Empowering Spatiotemporal Prediction by Re-learning Input-label Residual

Prevailing spatiotemporal prediction models typically operate under a forward (unidirectional) learning paradigm, in which models extract spatiotemporal features from historical observation input and map them to target spatiotemporal space for future forecasting (label). However, these models frequently exhibit suboptimal performance when spatiotemporal discrepancies exist between inputs and labels, for instance, when nodes with similar time-series inputs manifest distinct future labels, or vice versa. To address this limitation, we propose explicitly incorporating label features during the training phase. Specifically, we introduce the Spatiotemporal Residual Theorem, which generalizes the conventional unidirectional spatiotemporal prediction paradigm into a bidirectional learning framework. Building upon this theoretical foundation, we design an universal module, termed ReLearner, which seamlessly augments Spatiotemporal Neural Networks (STNNs) with a bidirectional learning capability via an auxiliary inverse learning process. In this process, the model relearns the spatiotemporal feature residuals between input data and future data. The proposed ReLearner comprises two critical components: (1) a Residual Learning Module, designed to effectively disentangle spatiotemporal feature discrepancies between input and label representations; and (2) a Residual Smoothing Module, employed to smooth residual terms and facilitate stable convergence. Extensive experiments conducted on 11 real-world datasets across 14 backbone models demonstrate that ReLearner significantly enhances the predictive performance of existing STNNs.Our code is available on GitHub.

cs.LG

The Aubin Property for Generalized Equations over $C^2$-cone Reducible Sets

This paper establishes the equivalence between the Aubin property and strong regularity for generalized equations over $C^2$-cone reducible sets. This result resolves a long-standing question in variational analysis and extends the classical equivalence theorem for polyhedral sets to a broad class of non-polyhedral sets. Our proof strategy departs from traditional variational techniques, integrating insights from convex geometry with powerful tools from algebraic topology. At the heart of our analysis is a novel index theorem for a class of functions involving metric projections onto arbitrary closed convex sets. Its proof exploits the geometry of normal cones and topological degree theory. We then use a lift of the diffeomorphism provided by the $C^2$-cone reduction to reduce the original generalized equations to the setting of the index theorem. The homological inverse mapping theorem then yields the desired strong regularity. We show that the $C^2$-cone reducibility assumption cannot be removed in general by constructing a three-dimensional semialgebraic counterexample. This result unifies and extends existing stability results for conventional nonlinear programming, nonlinear second-order cone programming, and nonlinear semidefinite programming under a single general framework, and removes the local optimality assumption imposed in the existing non-polyhedral results.

math.OC

On the $p$-order Semismoothness of the Metric Projection onto Slices of the Positive Semidefinite Cone

The metric projection onto the positive semidefinite (PSD) cone is strongly semismooth, a property that guarantees local quadratic convergence for many powerful algorithms in semidefinite programming. In this paper, we investigate whether this essential property holds for the metric projection onto an affine slice of the PSD cone, which is the operator implicitly used by many algorithms that handle linear constraints directly. Although this property is known to be preserved for the second-order cone, we conclusively demonstrate that this is not the case for the PSD cone. Specifically, we provide a constructive example that for any $p > 0$, there exists an affine slice of a PSD cone for which the metric projection operator fails to be $p$-order semismooth. This finding establishes a fundamental difference between the geometry of the second-order cone and the PSD cone and necessitates new approaches for both analysis and algorithm design for linear semidefinite programming problems.

math.OC

Spatiotemporal Causal Decoupling Model for Air Quality Forecasting

Due to the profound impact of air pollution on human health, livelihoods, and economic development, air quality forecasting is of paramount significance. Initially, we employ the causal graph method to scrutinize the constraints of existing research in comprehensively modeling the causal relationships between the air quality index (AQI) and meteorological features. In order to enhance prediction accuracy, we introduce a novel air quality forecasting model, AirCade, which incorporates a causal decoupling approach. AirCade leverages a spatiotemporal module in conjunction with knowledge embedding techniques to capture the internal dynamics of AQI. Subsequently, a causal decoupling module is proposed to disentangle synchronous causality from past AQI and meteorological features, followed by the dissemination of acquired knowledge to future time steps to enhance performance. Additionally, we introduce a causal intervention mechanism to explicitly represent the uncertainty of future meteorological features, thereby bolstering the model's robustness. Our evaluation of AirCade on an open-source air quality dataset demonstrates over 20\% relative improvement over state-of-the-art models.

cs.AI

All-water supercapacitor enabled by 1-nm clay channels

Water confined to channels one nanometer thick exhibits electrochemical behavior distinct from bulk water, including enhanced protonic conductivity and large dielectric anisotropy. Here, we exploit these characteristics to design a scalable electrochemical energy-storage system ("blue capacitor") constructed entirely from naturally abundant materials. By assembling layered clays and conductive graphene, we produce 1-nm-thick channels in which confined water acts as the sole electrolyte. We systematically study different clay types, the electrode composition, and separator thickness using complementary physicochemical and electrochemical techniques. The device operates stably up to 1.6 V, achieves specific capacitances of up to 40 F/g, nearly 100% coulombic efficiency, and stable performance over more than 60,000 charge-discharge cycles. Structural and dynamic analyses validate the device architecture, water purity, and proton transport in the nanopores. These results demonstrate that nanoconfined water can function as an electrolyte in a macroscopic electrochemical device, providing a platform for exploring sustainable aqueous energy-storage systems.

cond-mat.soft

Nanostructured Fe2O3/CuxO Heterojunction for Enhanced Solar Redox Flow Battery Performance

Solar redox flow batteries (SRFB) have received much attention as an alternative integrated technology for simultaneous conversion and storage of solar energy. Yet, the photocatalytic efficiency of semiconductor-based single photoelectrode, such as hematite, remains low due to the trade-off between fast electron hole recombination and insufficient light utilization, as well as inferior reaction kinetics at the solid/liquid interface. Herein, we present an {\alpha}-Fe2O3/CuxO p-n junction, coupled with a readily scalable nanostructure, that increases the electrochemically active sites and improves charge separation. Thanks to light-assisted scanning electrochemical microscopy (Photo-SECM), we elucidate the morphology-dependent carrier transfer process involved in the photo-oxidation reaction at a {\alpha}-Fe2O3 photoanode. The optimized nanostructured is then exploited in the {\alpha}-Fe2O3/CuxO p-n junction, achieving an outstanding unbiased photocurrent density of 0.46 mA/cm2, solar-to-chemical (STC) efficiency over 0.35% and a stable photocharge-discharge cycling. The average solar-to-output energy efficiency (SOEE) for this unassisted {\alpha}-Fe2O3-based SRFB system reaches 0.18%, comparable to previously reported DSSC-assisted hematite SRFBs. The use of earth-abundant materials and the compatibility with scalable nanostructuring and heterojunction preparation techniques, offer promising opportunities for cost-effective device deployment in real-world applications.

physics.chem-ph

Randomly Projected Convex Clustering Model: Motivation, Realization, and Cluster Recovery Guarantees

In this paper, we propose a randomly projected convex clustering model for clustering a collection of $n$ high dimensional data points in $\mathbb{R}^d$ with $K$ hidden clusters. Compared to the convex clustering model for clustering original data with dimension $d$, we prove that, under some mild conditions, the perfect recovery of the cluster membership assignments of the convex clustering model, if exists, can be preserved by the randomly projected convex clustering model with embedding dimension $m = O(\epsilon^{-2}\log(n))$, where $0 < \epsilon < 1$ is some given parameter. We further prove that the embedding dimension can be improved to be $O(\epsilon^{-2}\log(K))$, which is independent of the number of data points. Extensive numerical experiment results will be presented in this paper to demonstrate the robustness and superior performance of the randomly projected convex clustering model. The numerical results presented in this paper also demonstrate that the randomly projected convex clustering model can outperform the randomly projected K-means model in practice.

cs.LG