arXiv ScienceSearch

arXiv subjects

Jinhui Wang

Publications and source records attributed to Jinhui Wang.

12 recordsLinked to original sources

SPORT: Spherical-PSNR-Optimized tRuncaTion for Power-Efficient 360-Degree Video Systems

Memory bandwidth accounts for 30-40% of total power consumption in standalone virtual reality (VR) headsets, yet existing systems typically store the entire 360-degree frame at a uniform resolution regardless of viewer gaze. This paper presents SPORT (Spherical-PSNR Optimized tRuncaTion), a bit-truncation framework that reduces display-path memory power by storing only the most significant bits of pixels outside the user's field of view (FoV). Specifically, a new bit-truncation framework is developed to use weighted-to-spherically-uniform PSNR (WS-PSNR) directly in the optimization constraint, eliminating the metric inconsistency that arises when standard PSNR is used for a WS-PSNR quality target. Also, gaze-predictive tile classification compensates for the 9.33 ms end-to-end pipeline latency, reducing boundary misclassifications by 5.2 percentage points at a cost of only 0.01 ms. In addition, the developed SPORT-B variant, which keeps the FoV lossless, achieves 47.9% memory power saving and 47.9% bandwidth reduction across different 4K video sequences while satisfying all three per-region WS-PSNR thresholds and maintaining SSIM = 1.000 in the attended region. The full adaptive variant SPORT-A reaches 51.6% power saving, 3.1percentage points more than a PSNR-based optimizer at equal measured quality. SPORT is validated on the TrunMEM360 flexible SRAM Application-Specific Integrated Circuit (ASIC) fabricated in SkyWater 130 nm CMOS, confirming byte-exact silicon-software agreement, with WS-PSNR and SSIM matching within 0.1 dB and 0.001. CACTI-based analysis confirms 48.72% DRAM leakage reduction and 36.4%/36.7% read/write energy reduction. The total motion-to-photon latency of 9.33 ms satisfies the 20 ms VR comfort budget with a 53.3% safety margin.

cs.AR

SN 2024abfl: A Low-Luminosity Type IIP Supernova in NGC 2146 from a Low-Mass Red Supergiant Progenitor

Type IIP supernovae (SNe IIP) exhibit a significant diversity in their explosion properties, yet the physical mechanisms driving this diversity remain unknown. In this work, we present photometric and spectroscopic observations of SN 2024abfl, a SN IIP in NGC 2146 with a directly detected red supergiant (RSG) progenitor. We find it has a low plateau luminosity ($M_V \sim -15$ mag) and a relatively long plateau length ($\sim 126.5$ days). By fitting a semi-analytical model, we estimated a $^{56}$Ni mass of $\sim 0.009 M_\odot$, an initial kinetic energy of $\sim 0.42$ foe, an initial thermal energy of $\sim 0.03$ foe and an ejecta mass of $\sim 8.3 M_\odot$. The spectral evolution of SN 2024abfl is similar to those of other SNe IIP, except for much lower ejecta velocities at similar epochs. At later epochs, we find a relatively high-velocity H$\alpha$ absorption feature at $\sim -4000$ km s$^{-1}$, possibly due to a fast-moving plume of matter in the inner ejecta, and two emission features at $\pm 2000$ km s$^{-1}$, possibly caused by CSM interaction. We estimate the progenitor mass to be $\le 15 M_\odot$ based on nebular spectra. We conclude that SN 2024abfl is a low-luminosity SN IIP originating from a low-mass RSG progenitor.

astro-ph.HE

Sneak Path Current Modeling in Memristor Crossbar Arrays for Analog In-Memory Computing

Memristor crossbar arrays have emerged as a key component for next-generation non-volatile memories, artificial neural networks, and analog in-memory computing (IMC) systems. By minimizing data transfer between the processor and memory, they offer substantial energy savings. However, a major design challenge in memristor crossbar arrays is the presence of sneak path currents, which degrade electrical performance, reduce noise margins, and limit reliable operations. This work presents a closed-form analytical framework based on 1.4nm technology for accurately estimating sneak path currents in memristor crossbar arrays. The proposed model captures the interdependence of key design parameters in memristor crossbar arrays, including array size, ON/OFF ratio of memristors, read voltage, and interconnect conditions, through mathematically derived relationships. It supports various practical configurations, such as different data patterns and connection strategies, enabling rapid and comprehensive sneak path current modeling. The sensitivity analysis includes how design parameters influence sneak path current and noise margin loss, underscoring the trade-offs involved in scaling crossbar arrays. Validation through SPICE simulations shows that the model achieves an error of less than 10.9% while being up to 4784 times faster than full circuit simulations. This analytical framework offers a powerful tool for quantitative assessment and pre-design/real-time optimization of memristor-based analog in-memory computing (IMC) architectures.

cs.ET

Flexible Bit-Truncation Memory for Approximate Applications on the Edge

Bit truncation has demonstrated great potential to enable run-time quality-power adaptive data storage, thereby optimizing the power/energy efficiency of approximate applications and supporting their deployment in edge environments. However, existing bit-truncation memories require custom designs for a specific application. In this paper, we present a novel bit-truncation memory with full adaptation flexibility, which can truncate any number of data bits at run time to meet different quality and power trade-off requirements for various approximate applications. The developed bit-truncation memory has been applied to two representative data-intensive approximate applications: video processing and deep learning. Our experiments show that the proposed memory can support three different video applications (including luminance-aware, content-aware, and region-of-interest-aware) with enhanced power efficiency (up to 47.02% power savings) as compared to state-of-the-art. In addition, the proposed memory achieves significant (up to 51.69%) power savings for both baseline and pruned lightweight deep learning models, respectively, with a low implementation cost (2.89% silicon area overhead).

cs.AR

Stable self-charged perovskite quantum rods for liquid laser with near-zero threshold

Colloidal quantum dots (QDs) are promising optical gain materials that require further threshold reduction to realize their full potential. While QD charging theoretically reduces the threshold to zero, its effectiveness has been limited by strong Auger recombination and unstable charging. Here we theoretically reveal the optimal combination of charging number and Auger recombination to minimize the lasing threshold. Experimentally, we develop stable self-charged perovskite quantum rods (QRs) as an alternative to QDs via state engineering and Mn-doping strategy. An unprecedented two-order-of-magnitude reduction in nonradiative Auger recombination enables QRs to support a sufficient charging number of up to 6. The QR liquid lasing is then achieved with a near-zero threshold of 0.098 using quasi-continuous pumping of nanosecond pulses, which is the lowest threshold among all reported QD lasers. These achievements demonstrate the potential of the specially engineered QRs as an excellent gain media and pave the way for their prospective applications.

cond-mat.mes-hall

Machine Learning-Driven Student Performance Prediction for Enhancing Tiered Instruction

Student performance prediction is one of the most important subjects in educational data mining. As a modern technology, machine learning offers powerful capabilities in feature extraction and data modeling, providing essential support for diverse application scenarios, as evidenced by recent studies confirming its effectiveness in educational data mining. However, despite extensive prediction experiments, machine learning methods have not been effectively integrated into practical teaching strategies, hindering their application in modern education. In addition, massive features as input variables for machine learning algorithms often leads to information redundancy, which can negatively impact prediction accuracy. Therefore, how to effectively use machine learning methods to predict student performance and integrate the prediction results with actual teaching scenarios is a worthy research subject. To this end, this study integrates the results of machine learning-based student performance prediction with tiered instruction, aiming to enhance student outcomes in target course, which is significant for the application of educational data mining in contemporary teaching scenarios. Specifically, we collect original educational data and perform feature selection to reduce information redundancy. Then, the performance of five representative machine learning methods is analyzed and discussed with Random Forest showing the best performance. Furthermore, based on the results of the classification of students, tiered instruction is applied accordingly, and different teaching objectives and contents are set for all levels of students. The comparison of teaching outcomes between the control and experimental classes, along with the analysis of questionnaire results, demonstrates the effectiveness of the proposed framework.

cs.LG

On the asymptotics of 3+1D cosmologies with bounded scalar potential and isometry group forming 2-dimensional orbits

We study the onset of inflation in 3+1 dimensional cosmologies with an inflationary potential $U$ satisfying $0 < \Lambda_1 \leq U \leq \Lambda_2$, matter satisfying the dominant and strong energy conditions, and with spatial slices that can be foliated by 2-dimensional surfaces that are orbits under an isometry group. Assuming an initial Cauchy slice with positive mean curvature everywhere, we show, via mean curvature flow, that there exists a family of spatial slices parameterized by $\lambda$, whose volume grows between the flat slicings in de Sitter spaces with cosmological constants $\Lambda_1$ and $\Lambda_2$. In particular, inflationary expansion indeed occurs in this setting with inhomogeneous initial conditions. Finally, we apply this "inflationary time coordinate" $\lambda$ to study asymptotics of the variation in the metric, the average stress-energy tensor, and the dynamics of an inflaton field on a spatial slice.

hep-th

An Adaptive Interpolation Scheme for Wideband Frequency Sweep in Electromagnetic Simulations

An adaptive interpolation scheme is proposed to accurately calculate the wideband responses in electromagnetic simulations. In the proposed scheme, the sampling points are first carefully divided into several groups based on their responses to avoid the Runge phenomenon and the error fluctuations, and then different interpolation strategies are used to calculate the responses in the whole frequency band. If the relative error does not satisfy the predefined threshold in a specific frequency band, it will be refined until the error criteria is met. The detailed error analysis is also presented to verify the accuracy of the interpolation scheme. At last, two numerical examples including the antenna radiation and the filter simulation are carried out to validate its accuracy and efficiency.

cs.CE

Improving DNN Fault Tolerance using Weight Pruning and Differential Crossbar Mapping for ReRAM-based Edge AI

Recent research demonstrated the promise of using resistive random access memory (ReRAM) as an emerging technology to perform inherently parallel analog domain in-situ matrix-vector multiplication -- the intensive and key computation in deep neural networks (DNNs). However, hardware failure, such as stuck-at-fault defects, is one of the main concerns that impedes the ReRAM devices to be a feasible solution for real implementations. The existing solutions to address this issue usually require an optimization to be conducted for each individual device, which is impractical for mass-produced products (e.g., IoT devices). In this paper, we rethink the value of weight pruning in ReRAM-based DNN design from the perspective of model fault tolerance. And a differential mapping scheme is proposed to improve the fault tolerance under a high stuck-on fault rate. Our method can tolerate almost an order of magnitude higher failure rate than the traditional two-column method in representative DNN tasks. More importantly, our method does not require extra hardware cost compared to the traditional two-column mapping scheme. The improvement is universal and does not require the optimization process for each individual device.

cs.LG

Identification and Development of Therapeutics for COVID-19

After emerging in China in late 2019, the novel Severe acute respiratory syndrome-like coronavirus 2 (SARS-CoV-2) spread worldwide and as of early 2021, continues to significantly impact most countries. Only a small number of coronaviruses are known to infect humans, and only two are associated with the severe outcomes associated with SARS-CoV-2: Severe acute respiratory syndrome-related coronavirus, a closely related species of SARS-CoV-2 that emerged in 2002, and Middle East respiratory syndrome-related coronavirus, which emerged in 2012. Both of these previous epidemics were controlled fairly rapidly through public health measures, and no vaccines or robust therapeutic interventions were identified. However, previous insights into the immune response to coronaviruses gained during the outbreaks of severe acute respiratory syndrome (SARS) and Middle East respiratory syndrome (MERS) have proved beneficial to identifying approaches to the treatment and prophylaxis of novel coronavirus disease 2019 (COVID-19). A number of potential therapeutics against SARS-CoV-2 and the resultant COVID-19 illness were rapidly identified, leading to a large number of clinical trials investigating a variety of possible therapeutic approaches being initiated early on in the pandemic. As a result, a small number of therapeutics have already been authorized by regulatory agencies such as the Food and Drug Administration (FDA) in the United States, and many other therapeutics remain under investigation. Here, we describe a range of approaches for the treatment of COVID-19, along with their proposed mechanisms of action and the current status of clinical investigation into each candidate. The status of these investigations will continue to evolve, and this review will be updated as progress is made.

q-bio.QM

Pathogenesis, Symptomatology, and Transmission of SARS-CoV-2 through Analysis of Viral Genomics and Structure

The novel coronavirus SARS-CoV-2, which emerged in late 2019, has since spread around the world and infected hundreds of millions of people with coronavirus disease 2019 (COVID-19). While this viral species was unknown prior to January 2020, its similarity to other coronaviruses that infect humans has allowed for rapid insight into the mechanisms that it uses to infect human hosts, as well as the ways in which the human immune system can respond. Here, we contextualize SARS-CoV-2 among other coronaviruses and identify what is known and what can be inferred about its behavior once inside a human host. Because the genomic content of coronaviruses, which specifies the virus's structure, is highly conserved, early genomic analysis provided a significant head start in predicting viral pathogenesis and in understanding potential differences among variants. The pathogenesis of the virus offers insights into symptomatology, transmission, and individual susceptibility. Additionally, prior research into interactions between the human immune system and coronaviruses has identified how these viruses can evade the immune system's protective mechanisms. We also explore systems-level research into the regulatory and proteomic effects of SARS-CoV-2 infection and the immune response. Understanding the structure and behavior of the virus serves to contextualize the many facets of the COVID-19 pandemic and can influence efforts to control the virus and treat the disease.

q-bio.QM

Anomaly Detection with Tensor Networks

Originating from condensed matter physics, tensor networks are compact representations of high-dimensional tensors. In this paper, the prowess of tensor networks is demonstrated on the particular task of one-class anomaly detection. We exploit the memory and computational efficiency of tensor networks to learn a linear transformation over a space with dimension exponential in the number of original features. The linearity of our model enables us to ensure a tight fit around training instances by penalizing the model's global tendency to a predict normality via its Frobenius norm---a task that is infeasible for most deep learning models. Our method outperforms deep and classical algorithms on tabular datasets and produces competitive results on image datasets, despite not exploiting the locality of images.

cs.LG