arXiv ScienceSearch

arXiv subjects

Piyush Kumar

Publications and source records attributed to Piyush Kumar.

At least 19 recordsLinked to original sources

Muonium dynamics as a probe for depth-resolved properties of 4H-SiC

This study establishes a baseline for muonium (Mu) charge-exchange dynamics in n-type 4H-SiC through a detailed low-energy muon spin rotation (LE-uSR) investigation. Epitaxially grown and ion-implanted samples with nitrogen and phosphorus donors were characterized to assess the effect of carrier concentration and doping method on defect formation. LE-uSR enabled nanometer scale depth profiling of near-surface and implanted regions, revealing variations in charge carrier concentration due to fixed surface charges. The temperature dependence of the diamagnetic fraction and phase provided direct evidence of the Mu0 to Mu- transition, with extracted activation energies consistent with known donor ionization energies. Additionally, high-field uSR was used to analyze the Mu dynamics, and Monte-Carlo simulations to model the Mu0 electron capture process. The simulation results offer a quantitative method to extract free electron concentrations from LE-uSR data, enhancing its capability to characterize the activation of dopants and carrier depth profiles. We demonstrate that LE-uSR is a powerful depth-resolved tool that can provide insights for optimizing the fabrication of reliable SiC devices for power electronics.

cond-mat.mtrl-sci

A Process-Aware Hybrid Si/IGO Monolithic-3D 6T SRAM with BEOL Pass-Gates for the 2nm Node

We propose a monolithic-3D (M3D) 6T SRAM at the 2nm node, integrating BEOL IGO pass-gates (PGs) with an all-silicon nanosheet latch, buried power rails (BPRs), and Ru interconnects. TCAD calibrated to a state-of-the-art double-gate IGO transistor with a tri-layer HfO$_2$/ZrO$_2$/HfO$_2$ (HZH) gate stack is combined with virtual fabrication and 3D parasitic extraction to realize the first process-aware layout of this topology. A novel neighbor-cell shared source/drain (S/D) bitline (BL) design enlarges the IGO contact area to mitigate contact resistance and restore PG drive without area penalty. The resulting cell achieves a 25\% footprint reduction vs the high-performance (HP) 122 Si baseline while maintaining robust static noise margin (SNM) over a wide supply voltage range. At the 128$\times$256 subarray-level, it reduces write delay by 42.2\% and EDP by 9.7\% compared to the high-density (HD) 111 Si baseline, owing to reduced cell parasitics and wordline (WL) loading from the smaller footprint.

cs.AR

Prior-Fitted Functional Flow: In-Context Generative Models for Pharmacokinetics

We introduce Prior-Fitted Functional Flows, a generative foundation model for pharmacokinetics that enables zero-shot population synthesis and individual forecasting without manual parameter tuning. We learn functional vector fields, explicitly conditioned on the sparse, irregular data of an entire study population. This enables the generation of coherent virtual cohorts as well as forecasting of partially observed patient trajectories with calibrated uncertainty. We construct a new open-access literature corpus to inform our priors, and demonstrate state-of-the-art predictive accuracy on extensive real-world datasets.

cs.LG

A Diffusion-Driven Fine-Grained Nodule Synthesis Framework for Enhanced Lung Nodule Detection from Chest Radiographs

Early detection of lung cancer in chest radiographs (CXRs) is crucial for improving patient outcomes, yet nodule detection remains challenging due to their subtle appearance and variability in radiological characteristics like size, texture, and boundary. For robust analysis, this diversity must be well represented in training datasets for deep learning based Computer-Assisted Diagnosis (CAD) systems. However, assembling such datasets is costly and often impractical, motivating the need for realistic synthetic data generation. Existing methods lack fine-grained control over synthetic nodule generation, limiting their utility in addressing data scarcity. This paper proposes a novel diffusion-based framework with low-rank adaptation (LoRA) adapters for characteristic controlled nodule synthesis on CXRs. We begin by addressing size and shape control through nodule mask conditioned training of the base diffusion model. To achieve individual characteristic control, we train separate LoRA modules, each dedicated to a specific radiological feature. However, since nodules rarely exhibit isolated characteristics, effective multi-characteristic control requires a balanced integration of features. We address this by leveraging the dynamic composability of LoRAs and revisiting existing merging strategies. Building on this, we identify two key issues, overlapping attention regions and non-orthogonal parameter spaces. To overcome these limitations, we introduce a novel orthogonality loss term during LoRA composition training. Extensive experiments on both in-house and public datasets demonstrate improved downstream nodule detection. Radiologist evaluations confirm the fine-grained controllability of our generated nodules, and across multiple quantitative metrics, our method surpasses existing nodule generation approaches for CXRs.

cs.CV

Modeling and Optimization of Two-Terminal Spin-Orbit-Torque MRAM

This paper presents physical modeling and benchmarking for two-terminal spin-orbit torque magnetic random-access memory (2T-SOT-MRAM). The results indicate that the common SOT materials that provide only in-plane torque can provide little to no improvement over spin-transfer-torque (STT) MRAM in terms of write energy. However, emerging SOT materials that provide out-of-plane torques with efficiencies as small as 0.1 can result in significant improvements in the write energy for such 2-terminal devices, especially when the magnet lateral dimensions are scaled down to 30 or 20 nm. Additionally, a novel 2T-SOT MRAM device is proposed that can increase the path electrons pass through the SOT layer; hence, increasing the generated spin current and the energy efficiency of the device. Our benchmarking results indicate that an out-of-plane SOT efficiency of 0.051 for 20nm wide devices can result in write energies approaching SRAM at the 7nm technology node.

cond-mat.mes-hall

Finite density QCD phase structure from strangeness fluctuations

Charting the phase diagram of Quantum Chromodynamics (QCD) at large density is a challenging task due to the complex action problem in lattice simulations. Through simulations at imaginary baryon chemical potential $\mu_B$ we observe that, if the strangeness neutrality condition is imposed, both the strangeness chemical potential $\mu_S/\mu_B$ and the strangeness susceptibility $\chi_2^S$ take on constant values at the chiral transition for varying $\mu_B$. We present new lattice data to extrapolate contours of constant $\mu_S/\mu_B$ or $\chi_2^S$ to finite baryon chemical potential. We argue that they are good proxies for the QCD crossover because, as we show, they are only mildly influenced by criticality and by finite volume effects. We obtain continuum limits for these proxies up to $\mu_B = 400$ MeV, through a next-to-next-to-leading order (N$^2$LO) Taylor expansion based on large-statistics data on $16^3 \times 8$, $20^3 \times 10$ and $24^3 \times 12$ lattices with our 4HEX improved staggered action. We show that these are in excellent agreement with existing results for the chiral transition and, strikingly, also with analogous contours obtained with the hadron resonance gas (HRG) model. On the $16^3 \times 8$ lattice, we carry out the expansion up to next-to-next-to-next-to-next-to-leading order (N$^4$LO), and extend the extrapolation beyond $\mu_B=500$ MeV, again finding perfect agreement with the HRG model. This suggests that the crossover line constructed from this proxy starts deviating from the chemical freeze-out line near $\mu_B\approx500$ MeV, as expected but not yet observed.

hep-lat

Hybrid Integration of InGaN Lasers in a Foundry-Fabricated Visible-Light Photonics Platform

Visible-spectrum photonic integrated circuits (PICs) present compact and scalable solutions for emerging technologies including quantum computing, biosensing, and virtual/augmented reality. Realizing their full potential requires the development of scalable visible-light-source integration methods compatible with high-volume manufacturing and capable of delivering high optical coupling efficiencies. Here, we demonstrate passive-alignment flip-chip bonding of 450-nm InGaN laser diodes onto a foundry-fabricated visible-light silicon (Si) photonics platform with silicon nitride (SiN) waveguides, thermo-optic (TO) devices, and photodetectors. Hybrid laser integration is realized using a sub-micron-precision die bonder equipped with a vision alignment system and a heatable pickup tool, allowing independent placement of multiple lasers onto a single Si chip. Co-design of the lasers and Si photonics, with lithographically defined alignment marks and mechanical stoppers, enables precise postbonding alignment. Efficient optical coupling between lasers and the SiN waveguides is demonstrated, with a minimum measured coupling loss of 1.1 dB. We achieve a maximum on-chip optical power of 60.7 mW and an on-chip wall-plug efficiency of 7.8%, the highest reported for hybrid-integrated visible-spectrum lasers, to our knowledge. An active PIC is also shown, integrating a bonded laser, an on-chip photodetector for power monitoring, and a thermo-optic switch for optical routing and variable attenuation. Overall, this work highlights passive-alignment flip-chip bonding as a practical, high-performance approach for integrating lasers onto visible-spectrum PICs. We envision that continued refinement of this technique within our photonics platform will support increasingly complex PICs with integrated lasers spanning the visible spectrum.

physics.optics

High-precision baryon number cumulants from lattice QCD in a finite box: cumulant ratios, Lee-Yang zeros and critical endpoint predictions

We have performed high-statistics lattice simulations using 4HEX improved staggered fermions on $16^3 \times 8$ lattices. We calculated the Taylor expansion coefficients of the pressure with respect to the baryochemical potential to the tenth order at zero, and fourth order at purely imaginary chemical potentials. We used this data to construct rational function approximations of the free energy. We use a rational ansatz that explicitly satisfies the charge conjugation symmetry and the Roberge-Weiss periodicity, which are exact properties of the QCD free energy. We use this ansatz to estimate the position of Lee-Yang zeros in the complex chemical potential plane. The temperature dependence of the imaginary part of the Lee-Yang zeros is then fitted with ans\"atze motivated by the universal behavior of the free energy near a 3D Ising critical point. In principle, this allows one to estimate the temperature of the critical endpoint. We consider several sources of systematic errors. On this single lattice spacing we find that with $84\%$ probability, the chiral critical endpoint is either below $103$~MeV temperature or it does not exist. We also identify some caveats of the method, which do not disappear even with the extremely high statistics of this present study. We discuss to what extent these can be eliminated by future high statistics lattice analyses.

hep-lat

Computational Aerothermal Framework and Analysis of Stetson Mach 6 Blunt Cone

Accurately predicting aerothermal behavior is paramount for the effective design of hypersonic vehicles, as aerodynamic heating plays a pivotal role in influencing performance metrics and structural integrity. This study introduces a computational aerothermal framework and analyzes a blunt cone subjected to Mach 6 conditions, drawing inspiration from Stetson foundational experimental work published in 1983. While the findings offer significant insights into the phenomena at play, the study highlights an urgent necessity for integrating chemical kinetics to comprehensively capture non-equilibrium effects, thereby enhancing the predictive accuracy of computational fluid dynamics (CFD) simulations. This research implements a one-way coupling method between CFD simulations and heat conduction analysis, facilitating a thorough investigation of surface heat transfer characteristics. The numerical results elucidate discrete roughness elements' impact on surface heating and fluid dynamics within high-speed airflow. Furthermore, the investigation underscores the critical importance of accounting for non-equilibrium thermochemical effects in aerothermal modeling to bolster the accuracy of high-enthalpy flow simulations. By refining predictive computational tools and deepening understanding of hypersonic aerothermal mechanisms, this research lays a robust groundwork for future experimental and computational endeavors, significantly contributing to advancing high-speed flight applications.

physics.flu-dyn

Content Addressable Memory Design with Reference Resistor for Improved Search Resolution

Despite the parallel in-memory search capabilities of content addressable memories (CAMs), their use in applications is constrained by their limited resolution that worsens as they are scaled to larger arrays or advanced nodes. In this work we present experimental results for a novel back-end-of-line compatible reference resistive device that can significantly improve the search resolution of CAMs implemented with CMOS and beyond-CMOS technologies to less than or equal to 5-bits.

cs.ET

Portable Lattice QCD implementation based on OpenCL

The presence of GPU from different vendors demands the Lattice QCD codes to support multiple architectures. To this end, Open Computing Language (OpenCL) is one of the viable frameworks for writing a portable code. It is of interest to find out how the OpenCL implementation performs as compared to the code based on a dedicated programming interface such as CUDA for Nvidia GPUs. We have developed an OpenCL backend for our already existing code of the Wuppertal-Budapest collaboration. In this contribution, we show benchmarks of the most time consuming part of the numerical simulation, namely, the inversion of the Dirac operator. We present the code performance on the JUWELS and LUMI Supercomputers based on Nvidia and AMD graphics cards, respectively, and compare with the CUDA backend implementation.

hep-lat

Computational Analysis of the Temperature Profile Developed for a Hot Zone of 2500{\deg}C in an Induction Furnace

Temperature gradients developed at ultra-high temperatures create a challenge for temperature measurements that are required for material processing. At ultra-high temperatures, the components of the system can react and change phases depending on their thermodynamic stability. These reactions change the system's physical properties, such as thermal conductivity and fluidity. This phenomenon complicates the extrapolation of temperature measurements, as they depend on the thermal conductivity of multiple insulating layers. The proposed model is an induction furnace employing an electromagnetic field to generate heat reaching 2500 degrees Celsius. A heat transfer simulation applying the finite element method determined temperatures and verified experimentally at key locations on the surface of the experimental setup within the furnace. The computed temperature profile of cylindrical graphite crucibles embedded in a larger cylindrical graphite body surrounded by zirconia grog is determined. Compared to experimental results, the simulation showed a percentage error of approximately 3.4 percent, confirming its accuracy.

physics.comp-ph

Partition of Unity Physics-Informed Neural Networks (POU-PINNs): An Unsupervised Framework for Physics-Informed Domain Decomposition and Mixtures of Experts

Physics-informed neural networks (PINNs) commonly address ill-posed inverse problems by uncovering unknown physics. This study presents a novel unsupervised learning framework that identifies spatial subdomains with specific governing physics. It uses the partition of unity networks (POUs) to divide the space into subdomains, assigning unique nonlinear model parameters to each, which are integrated into the physics model. A vital feature of this method is a physics residual-based loss function that detects variations in physical properties without requiring labeled data. This approach enables the discovery of spatial decompositions and nonlinear parameters in partial differential equations (PDEs), optimizing the solution space by dividing it into subdomains and improving accuracy. Its effectiveness is demonstrated through applications in porous media thermal ablation and ice-sheet modeling, showcasing its potential for tackling real-world physics challenges.

cs.LG

Designing an Optimal Scoop for Holloman High-Speed Test Track Water Braking Mechanism using Computational Fluid Dynamics

Specializing in high-speed testing, Holloman High-Speed Test Track (HHSTT) uses water braking to stop vehicles on the test track. This method takes advantage of the higher density of water, compared to air, to increase braking capability through momentum exchange by increasing the water content in that section at the end of the track. By studying water braking using computational fluid dynamics (CFD), the forces acting on tracked vehicles can be approximated and prepared before actual testing through numerical simulations. In this study, emphasis will be placed on the brake component of the tracked sled, which is responsible for interacting with water to brake. By discretizing a volume space around our brake, we accelerate the water and air to simulate the brake coupling relatively. The multiphase flow model uses the governing equations of the gas and liquid phases with the finite volume method to perform 3D simulations. By adjusting the air and water inlet velocity, it is possible to simulate HHSTT sled tests at various operating speeds.

physics.flu-dyn

Surface Structuring of Patterned 4H-SiC Surfaces Using a SiC/Si/SiC Sandwich Approach

Mesa- and trench-patterned surfaces of 4H-SiC(0001) 4{\textdegree}off wafers were structured in macrosteps using Si melting in a SiC-Si-SiC sandwich configuration. Si spreading difficulties were observed in the case of trench-patterned samples while the attempts on mesa-patterned ones were more successful. In the latter case, parallel macrosteps were formed on both the dry-etched and unetched areas though these macrosteps rarely cross the patterns edges. The proposed mechanism involved preferential etching at Si-C bilayer step edges and fast lateral propagation along the [1120] direction.

cond-mat.mtrl-sci

Computational Investigation of Roughness Effects on Boundary Layer Transition for Stetson's Blunt Cone at Mach 6

In this aerothermal study, we performed a two-dimensional steady-state Computational Fluid Dynamics (CFD) and heat conduction simulation at Mach 6. The key to our methodology was a one-way coupling between CFD surface temperature as a boundary condition and the calculation of the heat transfer flux and temperatures inside the solid stainless-steel body of a nose geometry. This approach allowed us to gain insight into surface heat transfer signatures with corresponding fluid flow regimes, such as the one experienced in laminar fluid flow. We have also examined this heat transfer under roughness values encountered in Stetson's studies at the Wright-Patterson Air Force Base Ludwig tube. To validate our findings, we have performed this type of work on a blunt cone, specifically for the U.S. Air Force. The research focuses on predicting transition onset using laminar correlations derived from Stetson's experimental studies, examining the role of discrete roughness elements. Findings emphasize the importance of incorporating non-equilibrium effects in future computational frameworks to enhance predictive accuracy for high-speed aerodynamic applications.

physics.flu-dyn

A 5T-2MTJ STT-assisted Spin Orbit Torque based Ternary Content Addressable Memory for Hardware Accelerators

In this work, we present a novel non-volatile spin transfer torque (STT) assisted spin-orbit torque (SOT) based ternary content addressable memory (TCAM) with 5 transistors and 2 magnetic tunnel junctions (MTJs). We perform a comprehensive study of the proposed design from the device-level to application-level. At the device-level, various write characteristics such as write error rate, time, and current have been obtained using micromagnetic simulations. The array-level search and write performance have been evaluated based on SPICE circuit simulations with layout extracted parasitics for bitcells while also accounting for the impact of interconnect parasitics at the 7nm technology node. A search error rate of 3.9x10^-11 is projected for exact search while accounting for various sources of variation in the design. In addition, the resolution of the search operation is quantified under various scenarios to understand the achievable quality of the approximate search operations. Application-level performance and accuracy of the proposed design have been evaluated and benchmarked against other state-of-the-art CAM designs in the context of a CAM-based recommendation system.

cs.ET

Empowering Abilities: Increasing Representation of Students with Disabilities in the STEM Field

The ExploreSTEM Summer Camps 2023 were designed to deliver inclusive STEM education to students aged 14 to 22 years with disabilities. This paper presents a thorough examination of the 2023 camp program, emphasizing the pivotal role of inclusive STEM education in potentially shaping students' personal and academic trajectories. The curriculum encompassed four weeklong fundamental STEM domains: Internet of Things (IoT), Computational Engineering, Artificial Intelligence (AI), and Augmented and Virtual Reality (AR/VR). Within Camp 1, students actively engaged with Dash robots, employing dedicated programming environments to command actions and gather sensor data, fostering interactions with the IoT platform and facilitating seamless data transmission. Camp 2 was dedicated to acquainting students with foundational computational engineering principles, establishing a robust framework for comprehending intricate engineering concepts. Camp 3 commenced with insightful presentations elucidating AI applications across multifaceted industries, including engineering, healthcare, and education, illuminating AI's pervasive influence on contemporary society. The primary aim of Camp 4 was to introduce students to the immersive domains of AR and VR, showcasing their applications beyond conventional STEM disciplines into everyday life experiences. The amalgamation of informative presentations, interactive activities, and a nurturing learning environment cultivated an engaging and enriching experience for all participants. By embracing inclusivity and harnessing innovative pedagogical approaches, the ExploreSTEM Summer Camps empowered students to explore, innovate, and excel within the dynamic realm of STEM education.

econ.GN