arXiv ScienceSearch

arXiv subjects

Sam Yang

Publications and source records attributed to Sam Yang.

6 recordsLinked to original sources

An end-to-end differentiable transient vapor-compression framework for automated machine sizing and unified optimal control

Accelerating the electrification of thermal energy requires vapor-compression heat pumps capable of dynamic, grid-responsive operation. However, equipment engineering remains fragmented across static rating-point selection, stiff multi-phase transient simulation, and gradient-based optimal control. Here, we present an end-to-end differentiable, finite-volume vapor-compression framework implemented natively in JAX that automates machine sizing directly from stated thermal duties and unifies dynamic simulation with predictive control under a single compiled residual $\dot{y}={f}(t,{y},{u})$. Thermodynamic evaluations bypass runtime root-finding via bilinear $(p,h)$ manifolds pre-flashed from Helmholtz equations of state, enabling analytical forward-mode automatic differentiation. Mass conservation across multi-phase coils is strictly preserved by incorporating both $(\partial\rho/\partial p)_h$ and $(\partial\rho/\partial h)_p$ partial derivatives into the dynamic pressure differential equation. The sizer directly inverts compressor displacement, electronic expansion valve area, and heat-exchanger tube counts via four-point cycle synthesis and $\varepsilon$-NTU matching using the identical polytropic compressor map. Crucially, the compiled physics kernel is shared symmetrically between $L$-stable TR-BDF2 stiff integration and implicit-Euler Model Predictive Control (MPC), eliminating plant-controller surrogate mismatch. Validated against open-access experimental benchmarks without parameter fitting, the framework predicts cooling capacity with $7.37\%$ MAPE across 16 mini-split operational runs and bounds on-period cooling error within $1.19\%$--$1.62\%$ on utility-scale Hardware-in-the-Loop traces. This work provides an open-source, differentiable foundation for automated machine synthesis, dynamic grid orchestration, and gradient-based hardware-control co-design.

eess.SY

Implicit-adjoint finite-volume topology optimization of two-dimensional conjugate heat transfer

Designing compact, high-efficiency thermal architectures requires resolving the competing demands of solid conduction, fluid convection, and flow resistance within highly constrained physical envelopes. Density-based topology optimization provides a systematic framework for synthesizing these coupled layouts, yet the reproducibility and numerical stability of the resulting designs depend critically on the underlying discrete solvers and adjoint sensitivity mechanics. In this work, we present a transparent, self-contained two-dimensional finite-volume formulation on a staggered Marker-and-Cell grid for conjugate heat transfer governed by design-dependent energy transport coupled to Stokes--Brinkman or Darcy flow at fixed solid volume. To prevent spurious artificial thermal sources in porous, weakly compressible Brinkman domains, the discrete advection operator is constructed to satisfy the identity $\mathbf{u}\cdot\nabla T=\nabla\cdot(\mathbf{u}T)-T(\nabla\cdot\mathbf{u})$ cellwise, ensuring that uniform temperature fields remain exact discrete nullspaces even under inexact continuity satisfaction. Reverse-mode derivatives are evaluated via the implicit function theorem rather than unrolled iterative loops, yielding exact discrete adjoints with bounded memory requirements. The discrete operators are systematically validated through the method of manufactured solutions and directional Taylor remainder tests. Four representative thermofluid design benchmarks are optimized using a projected-gradient scheme with $\beta$-continuation, wherein candidate iterates are accepted and published only upon satisfying rigorous, predeclared gates on residual convergence, mass conservation, volume feasibility, and numerical finiteness. The resulting formulation provides an inspectable, deterministic reference stack for verifiable conjugate thermofluidic topology optimization.

math.NA

Multi-fidelity power flow solver

We propose a multi-fidelity neural network (MFNN) tailored for rapid high-dimensional grid power flow simulations and contingency analysis with scarce high-fidelity contingency data. The proposed model comprises two networks -- the first one trained on DC approximation as low-fidelity data and coupled to a high-fidelity neural net trained on both low- and high-fidelity power flow data. Each network features a latent module which parametrizes the model by a discrete grid topology vector for generalization (e.g., $n$ power lines with $k$ disconnections or contingencies, if any), and the targeted high-fidelity output is a weighted sum of linear and nonlinear functions. We tested the model on 14- and 118-bus test cases and evaluated its performance based on the $n-k$ power flow prediction accuracy with respect to imbalanced contingency data and high-to-low-fidelity sample ratio. The results presented herein demonstrate MFNN's potential and its limits with up to two orders of magnitude faster and more accurate power flow solutions than DC approximation.

cs.LG

MLPro: A System for Hosting Crowdsourced Machine Learning Challenges for Open-Ended Research Problems

The task of developing a machine learning (ML) model for a particular problem is inherently open-ended, and there is an unbounded set of possible solutions. Steps of the ML development pipeline, such as feature engineering, loss function specification, data imputation, and dimensionality reduction, require the engineer to consider an extensive and often infinite array of possibilities. Successfully identifying high-performing solutions for an unfamiliar dataset or problem requires a mix of mathematical prowess and creativity applied towards inventing and repurposing novel ML methods. Here, we explore the feasibility of hosting crowdsourced ML challenges to facilitate a breadth-first exploration of open-ended research problems, thereby expanding the search space of problem solutions beyond what a typical ML team could viably investigate. We develop MLPro, a system which combines the notion of open-ended ML coding problems with the concept of an automatic online code judging platform. To conduct a pilot evaluation of this paradigm, we crowdsource several open-ended ML challenges to ML and data science practitioners. We describe results from two separate challenges. We find that for sufficiently unconstrained and complex problems, many experts submit similar solutions, but some experts provide unique solutions which outperform the "typical" solution class. We suggest that automated expert crowdsourcing systems such as MLPro have the potential to accelerate ML engineering creativity.

cs.HC

Understanding Mobile GUI: from Pixel-Words to Screen-Sentences

The ubiquity of mobile phones makes mobile GUI understanding an important task. Most previous works in this domain require human-created metadata of screens (e.g. View Hierarchy) during inference, which unfortunately is often not available or reliable enough for GUI understanding. Inspired by the impressive success of Transformers in NLP tasks, targeting for purely vision-based GUI understanding, we extend the concepts of Words/Sentence to Pixel-Words/Screen-Sentence, and propose a mobile GUI understanding architecture: Pixel-Words to Screen-Sentence (PW2SS). In analogy to the individual Words, we define the Pixel-Words as atomic visual components (text and graphic components), which are visually consistent and semantically clear across screenshots of a large variety of design styles. The Pixel-Words extracted from a screenshot are aggregated into Screen-Sentence with a Screen Transformer proposed to model their relations. Since the Pixel-Words are defined as atomic visual components, the ambiguity between their visual appearance and semantics is dramatically reduced. We are able to make use of metadata available in training data to auto-generate high-quality annotations for Pixel-Words. A dataset, RICO-PW, of screenshots with Pixel-Words annotations is built based on the public RICO dataset, which will be released to help to address the lack of high-quality training data in this area. We train a detector to extract Pixel-Words from screenshots on this dataset and achieve metadata-free GUI understanding during inference. We conduct experiments and show that Pixel-Words can be well extracted on RICO-PW and well generalized to a new dataset, P2S-UI, collected by ourselves. The effectiveness of PW2SS is further verified in the GUI understanding tasks including relation prediction, clickability prediction, screen retrieval, and app type classification.

cs.CV

High-throughput screening of encapsulated islets using wide-field lens-free on-chip imaging

Islet microencapsulation is a promising solution to diabetes treatment, but its quality control based on manual microscopic inspection is extremely low-throughput, highly variable and laborious. This study presents a high-throughput islet-encapsulation quality screening system based on lens-free on-chip imaging with a wide field-of-view of 18.15 cm^2, which is more than 100 times larger than that of a lens-based optical microscope, enabling it to image and analyze ~8,000 microcapsules in a single frame. Custom-written image reconstruction and processing software provides the user with clinically important information, such as microcapsule count, size, intactness, and information on whether each capsule contains an islet. This high-throughput and cost-effective platform can be useful for researchers to develop better encapsulation protocols as well as perform quality control prior to transplantation.

physics.ins-det