arXiv Science⌕ Search

arXiv subjects

Yu Luo

Publications and source records attributed to Yu Luo.

At least 145 records · Page 8Linked to original sources

Enhancing In-Situ Structural Health Monitoring through RF Energy-Powered Sensor Nodes and Mobile Platform

This research contributes to long-term structural health monitoring (SHM) by exploring radio frequency energy-powered sensor nodes (RF-SNs) embedded in concrete. Unlike traditional in-situ monitoring systems relying on batteries or wire-connected power sources, the RF-SN captures radio energy from a mobile radio transmitter for sensing and communication. This offers a cost-effective solution for consistent in-situ perception. To optimize the system performance across various situations, we've explored both active and passive communication methods. For the active RF-SN, we implement a specialized control circuit enabling the node to transmit data through ZigBee protocol at low incident power. For the passive RF-SN, radio energy is not only for power but also as a carrier signal, with data conveyed by modulating the amplitude of the backscattered radio wave. To address the challenge of significant attenuation of the backscattering signal in concrete, we utilize a square chirp-based modulation scheme for passive communication. This scheme allows the receiver to successfully decode the data even under a negative signal-to-noise ratio (SNR) condition. The experimental results indicate that an active RF-SN embedded in concrete at a depth of 13.5 cm can be effectively powered by a 915MHz mobile radio transmitter with an effective isotropic radiated power (EIRP) of 32.5dBm. This setup allows the RF-SN to send over 1 kilobyte of data within 10 seconds, with an additional 1.7 kilobytes every 1.6 seconds of extra charging. For the passive RF-SN buried at the same depth, continuous data transmission at a rate of 224 bps with a 3% bit error rate (BER) is achieved when the EIRP of the transmitter is 23.6 dBm.

eess.SY↗

An Efficient Sparse Inference Software Accelerator for Transformer-based Language Models on CPUs

In recent years, Transformer-based language models have become the standard approach for natural language processing tasks. However, stringent throughput and latency requirements in industrial applications are limiting their adoption. To mitigate the gap, model compression techniques such as structured pruning are being used to improve inference efficiency. However, most existing neural network inference runtimes lack adequate support for structured sparsity. In this paper, we propose an efficient sparse deep learning inference software stack for Transformer-based language models where the weights are pruned with constant block size. Our sparse software accelerator leverages Intel Deep Learning Boost to maximize the performance of sparse matrix - dense matrix multiplication (commonly abbreviated as SpMM) on CPUs. Our SpMM kernel outperforms the existing sparse libraries (oneMKL, TVM, and LIBXSMM) by an order of magnitude on a wide range of GEMM shapes under 5 representative sparsity ratios (70%, 75%, 80%, 85%, 90%). Moreover, our SpMM kernel shows up to 5x speedup over dense GEMM kernel of oneDNN, a well-optimized dense library widely used in industry. We apply our sparse accelerator on widely-used Transformer-based language models including Bert-Mini, DistilBERT, Bert-Base, and BERT-Large. Our sparse inference software shows up to 1.5x speedup over Neural Magic's Deepsparse under same configurations on Xeon on Amazon Web Services under proxy production latency constraints. We also compare our solution with two framework-based inference solutions, ONNX Runtime and PyTorch, and demonstrate up to 37x speedup over ONNX Runtime and 345x over PyTorch on Xeon under the latency constraints. All the source code is publicly available on Github: https://github.com/intel/intel-extension-for-transformers.

cs.LG↗

Physics-data-driven intelligent optimization for large-scale meta-devices

Meta-devices have gained significant attention and have been widely utilized in optical systems for focusing and imaging, owing to their lightweight, high-integration, and exceptional-flexibility capabilities. However, based on the assumption of local phase approximation, traditional design method neglect the local lattice coupling effect between adjacent meta-atoms, thus harming the practical performance of meta-devices. Using physics-driven or data-driven optimization algorithms can effectively solve the aforementioned problems. Nevertheless, both of the methods either involve considerable time costs or require a substantial amount of data sets. Here, we propose a physics-data-driven approach based "intelligent optimizer" that enables us to adaptively modify the sizes of the studied meta-atom according to the sizes of its surrounding ones. Such a scheme allows to mitigate the undesired local lattice coupling effect, and the proposed network model works well on thousands of datasets with a validation loss of 3*10-3. Experimental results show that the 1-mm-diameter metalens designed with the "intelligent optimizer" possesses a relative focusing efficiency of 93.4% (as compared to ideal focusing) and a Strehl ratio of 0.94. In contrast to the previous inverse design method, our method significantly boosts designing efficiency with five orders of magnitude reduction in time. Our design approach may sets a new paradigm for devising large-scale meta-devices.

physics.optics↗

Blockchain-based Decentralized Co-governance: Innovations and Solutions for Sustainable Crowdfunding

This thesis provides an in-depth exploration of the Decentralized Co-governance Crowdfunding (DCC) Ecosystem, a novel solution addressing prevailing challenges in conventional crowdfunding methods faced by MSMEs and innovative projects. Among the problems it seeks to mitigate are high transaction costs, lack of transparency, fraud, and inefficient resource allocation. Leveraging a comprehensive review of the existing literature on crowdfunding economic activities and blockchain's impact on organizational governance, we propose a transformative socio-economic model based on digital tokens and decentralized co-governance. This ecosystem is marked by a tripartite community structure - the Labor, Capital, and Governance communities - each contributing uniquely to the ecosystem's operation. Our research unfolds the evolution of the DCC ecosystem through distinct phases, offering a novel understanding of socioeconomic dynamics in a decentralized digital world. It also delves into the intricate governance mechanism of the ecosystem, ensuring integrity, fairness, and a balanced distribution of value and wealth.

cs.CY↗

Perception and Semantic Aware Regularization for Sequential Confidence Calibration

Deep sequence recognition (DSR) models receive increasing attention due to their superior application to various applications. Most DSR models use merely the target sequences as supervision without considering other related sequences, leading to over-confidence in their predictions. The DSR models trained with label smoothing regularize labels by equally and independently smoothing each token, reallocating a small value to other tokens for mitigating overconfidence. However, they do not consider tokens/sequences correlations that may provide more effective information to regularize training and thus lead to sub-optimal performance. In this work, we find tokens/sequences with high perception and semantic correlations with the target ones contain more correlated and effective information and thus facilitate more effective regularization. To this end, we propose a Perception and Semantic aware Sequence Regularization framework, which explore perceptively and semantically correlated tokens/sequences as regularization. Specifically, we introduce a semantic context-free recognition and a language model to acquire similar sequences with high perceptive similarities and semantic correlation, respectively. Moreover, over-confidence degree varies across samples according to their difficulties. Thus, we further design an adaptive calibration intensity module to compute a difficulty score for each samples to obtain finer-grained regularization. Extensive experiments on canonical sequence recognition tasks, including scene text and speech recognition, demonstrate that our method sets novel state-of-the-art results. Code is available at https://github.com/husterpzh/PSSR.

cs.CV↗

A set of moment tensor potentials for zirconium with increasing complexity

Machine learning force fields (MLFFs) are an increasingly popular choice for atomistic simulations due to their high fidelity and improvable nature. Here, we propose a hybrid small-cell approach that combines attributes of both offline and active learning to systematically expand a quantum mechanical (QM) database while constructing MLFFs with increasing model complexity. Our MLFFs employ the moment tensor potential formalism. During this process, we quantitatively assessed structural properties, elastic properties, dimer potential energies, melting temperatures, phase stability, point defect formation energies, point defect migration energies, free surface energies, and generalized stacking fault (GSF) energies of Zr as predicted by our MLFFs. Unsurprisingly, model complexity has a positive correlation with prediction accuracy. We also find that the MLFFs wee able to predict the properties of out-of-sample configurations without directly including these specific configurations in the training dataset. Additionally, we generated 100 MLFFs of high complexity (1513 parameters each) that reached different local optima during training. Their predictions cluster around the benchmark DFT values, but subtle physical features such as the location of local minima on the GSFE surface are washed out by statistical noise.

physics.comp-ph↗

Local Group Dwarf Galaxy Detection Limit in the CSST survey

We predict the dwarf galaxy detection limits for the upcoming Chinese Space Station Telescope (CSST) survey that will cover 17,500 deg$^{2}$ of the sky with a wide field of view of 1.1 deg$^2$. The point-source depth reaches 26.3 mag in the $g$ band and 25.9 mag in the $i$ band. Constructing mock survey data based on the designed photometric bands, we estimate the recovery rate of artificial dwarf galaxies from mock point-source photometric catalogues. The detection of these artificial dwarf galaxies is strongly dependent on their distance, magnitude and size, in agreement with searches in current surveys. We expect CSST to enable the detection of dwarf galaxies with $M_V = -3.0$ and $μ_{250} = 32.0$ mag/arcsec$^2$ (surface-brightness limit for a system of half-light radius $r_{\rm h}$ = 250 pc at 400 kpc, and $M_V = -4.9$ and $μ_{250} = 30.5$ mag/arcsec$^2$ around the Andromeda galaxy. Beyond the Local Group, the CSST survey will achieve $M_V = -5.8$, and $μ_{250}$ = 29.7 mag/arcsec$^2$ in the distance range of 1--2 Mpc, opening up an exciting discovery space for faint field dwarf galaxies. With its optical bands, wide survey footprint, and space resolution, CSST will undoubtedly expand our knowledge of low-mass dwarf galaxies to an unprecedented volume.

astro-ph.GA↗

Context-Aware Selective Label Smoothing for Calibrating Sequence Recognition Model

Despite the success of deep neural network (DNN) on sequential data (i.e., scene text and speech) recognition, it suffers from the over-confidence problem mainly due to overfitting in training with the cross-entropy loss, which may make the decision-making less reliable. Confidence calibration has been recently proposed as one effective solution to this problem. Nevertheless, the majority of existing confidence calibration methods aims at non-sequential data, which is limited if directly applied to sequential data since the intrinsic contextual dependency in sequences or the class-specific statistical prior is seldom exploited. To the end, we propose a Context-Aware Selective Label Smoothing (CASLS) method for calibrating sequential data. The proposed CASLS fully leverages the contextual dependency in sequences to construct confusion matrices of contextual prediction statistics over different classes. Class-specific error rates are then used to adjust the weights of smoothing strength in order to achieve adaptive calibration. Experimental results on sequence recognition tasks, including scene text recognition and speech recognition, demonstrate that our method can achieve the state-of-the-art performance.

cs.AI↗

Exceptional Laurent biorthogonal polynomials through spectral transformations of generalized eigenvalue problems

A formulation is given for the spectral transformation of the generalized eigenvalue problem through the decomposition of the second-order differential operators. This allows us to construct some Laurent biorthogonal polynomial systems with gaps in the degree of the polynomial sequence. These correspond to an exceptional-type extension of the orthogonal polynomials, as an extension of the Laurent biorthogonal polynomials. Specifically, we construct the exceptional extension of the Hendriksen-van Rossum polynomials, which are biorthogonal analogs of the classical orthogonal polynomials. Similar to the cases of exceptional extensions of classical orthogonal polynomials, both of state-deletion and state-addition occur.

math.CA↗

Can we constrain galaxy geometry parameters using spatially integrated SED fitting?

Sophisticated spectral energy distribution (SED) models describe dust attenuation and emission using geometry parameters. This treatment is natural since dust effects are driven by the underlying star-dust geometry in galaxies. An example is the Starduster SED model, which divides a galaxy into a stellar disk, a stellar bulge, and a dust disk. This work utilises the Starduster SED model to study the efficacy of inferring geometry parameters using spatially integrated SED fitting. Our method fits the SED model to mock photometry produced by combining a semi-analytic model with the same SED model. Our fitting results imply that the disk radius can be constrained, while the inclination angle, dust disk to stellar disk radius ratio, bulge radius and intrinsic bulge to total luminosity ratio are unconstrained, even though 21 filters from UV to FIR are used. We also study the impact of S/N, finding that the increase of S/N (up to 80) brings limited improvements to the results. We provide a detailed discussion to explain these findings, and point out the implications for models with more general geometry.

astro-ph.GA↗

Effects of galaxy intrinsic alignment on weak lensing peak statistics

The galaxy intrinsic alignment (IA) is a dominant source of systematics in weak lensing (WL) studies. In this paper, by employing large simulations with semi-analytical galaxy formation, we investigate the IA effects on WL peak statistics. Different simulated source galaxy samples of different redshift distributions are constructed, where both WL shear and IA signals are included. Convergence reconstruction and peak statistics are then performed for these samples. Our results show that the IA effects on peak abundances mainly consist of two aspects. One is the additional contribution from IA to the shape noise. The other is from the satellite IA that can affect the peak signals from their host clusters significantly. The latter depends on the level of inclusion in a shear sample of the satellite galaxies of the clusters that contribute to WL peaks, and thus is sensitive to the redshift distribution of source galaxies. We pay particular attention to satellite IA and adjust it artificially in the simulations to analyze the dependence of the satellite IA impacts on its strength. This information can potentially be incorporated into the modeling of WL peak abundances, especially for high peaks physically originated from massive clusters of galaxies, and thus to mitigate the IA systematics on the cosmological constraints derived from WL peaks.

astro-ph.CO↗

Type-enriched Hierarchical Contrastive Strategy for Fine-Grained Entity Typing

Fine-grained entity typing (FET) aims to deduce specific semantic types of the entity mentions in text. Modern methods for FET mainly focus on learning what a certain type looks like. And few works directly model the type differences, that is, let models know the extent that one type is different from others. To alleviate this problem, we propose a type-enriched hierarchical contrastive strategy for FET. Our method can directly model the differences between hierarchical types and improve the ability to distinguish multi-grained similar types. On the one hand, we embed type into entity contexts to make type information directly perceptible. On the other hand, we design a constrained contrastive strategy on the hierarchical structure to directly model the type differences, which can simultaneously perceive the distinguishability between types at different granularity. Experimental results on three benchmarks, BBN, OntoNotes, and FIGER show that our method achieves significant performance on FET by effectively modeling type differences.

cs.CL↗

ChiQA: A Large Scale Image-based Real-World Question Answering Dataset for Multi-Modal Understanding

Visual question answering is an important task in both natural language and vision understanding. However, in most of the public visual question answering datasets such as VQA, CLEVR, the questions are human generated that specific to the given image, such as `What color are her eyes?'. The human generated crowdsourcing questions are relatively simple and sometimes have the bias toward certain entities or attributes. In this paper, we introduce a new question answering dataset based on image-ChiQA. It contains the real-world queries issued by internet users, combined with several related open-domain images. The system should determine whether the image could answer the question or not. Different from previous VQA datasets, the questions are real-world image-independent queries that are more various and unbiased. Compared with previous image-retrieval or image-caption datasets, the ChiQA not only measures the relatedness but also measures the answerability, which demands more fine-grained vision and language reasoning. ChiQA contains more than 40K questions and more than 200K question-images pairs. A three-level 2/1/0 label is assigned to each pair indicating perfect answer, partially answer and irrelevant. Data analysis shows ChiQA requires a deep understanding of both language and vision, including grounding, comparisons, and reading. We evaluate several state-of-the-art visual-language models such as ALBEF, demonstrating that there is still a large room for improvements on ChiQA.

cs.CL↗

An Extended Halo-based Group/Cluster finder: application to the DESI legacy imaging surveys DR8

We extend the halo-based group finder developed by \citet[][]{Yang2005a} to use data {\it simultaneously} with either photometric or spectroscopic redshifts. A mock galaxy redshift survey constructed from a high-resolution N-body simulation is used to evaluate the performance of this extended group finder. For galaxies with magnitude ${\rm z\le 21}$ and redshift $0<z\le 1.0$ in the DESI legacy imaging surveys (the Legacy Surveys), our group finder successfully identifies more than 60\% of the members in about $90\%$ of halos with mass $\ga 10^{12.5}\msunh$. Detected groups with mass $\ga 10^{12.0}\msunh$ have a purity (the fraction of true groups) greater than 90\%. The halo mass assigned to each group has an uncertainty of about 0.2 dex at the high mass end $\ga 10^{13.5}\msunh$ and 0.40 dex at the low mass end. Groups with more than 10 members have a redshift accuracy of $\sim 0.008$. We apply this group finder to the Legacy Surveys DR8 and find 5.2 Million groups with at least 3 members. About 387,000 of these groups have at least 10 members. The resulting catalog containing 3D coordinates, richness, halo masses, and total group luminosities, is made publicly available.

astro-ph.GA↗

Quenching of Massive Disk Galaxies in the IllustrisTNG Simulation

A rare population of massive disk galaxies have been found to invade the red sequence dominated by early-type galaxies. These red/quenched massive disk galaxies have recently gained great interest into their formation and origins. The usually proposed quenching mechanisms, such as bar quenching and environment quenching, seem not suitable for those bulge-less quenched disks in low-density environment. In this paper, we use the IllustrisTNG-300 simulation to investigate the formation of massive quenched central disk galaxies. It is found that these galaxies contain less gas and harbor giant supermassive black holes(SMBHs) (above $ 10^{8}M_{\odot}$) than their star forming counterparts. By tracing their formation history, we found that quenched disk galaxies formed early and preserved disk morphology for cosmological time scales. They have experienced less than one major merger on average and it is mainly mini-mergers (mass ratio $<$1/10) that contribute to the growth of their SMBHs. In the Illustris-TNG simulation the black hole feedback mode switches from thermal to kinetic feedback when the black hole mass is more massive than $\sim 10^{8}M_{\odot}$, which is more efficient to eject gas outside of the galaxy and to suppress further cooling of hot gaseous halo. We conclude that kinetic AGN feedback in massive red/quenched disk galaxy is the dominant quenching mechanism.

astro-ph.GA↗

Spatio-Temporal Federated Learning for Massive Wireless Edge Networks

This paper presents a novel approach to conduct highly efficient federated learning (FL) over a massive wireless edge network, where an edge server and numerous mobile devices (clients) jointly learn a global model without transporting the huge amount of data collected by the mobile devices to the edge server. The proposed FL approach is referred to as spatio-temporal FL (STFL), which jointly exploits the spatial and temporal correlations between the learning updates from different mobile devices scheduled to join STFL in various training epochs. The STFL model not only represents the realistic intermittent learning behavior from the edge server to the mobile devices due to data delivery outage, but also features a mechanism of compensating loss learning updates in order to mitigate the impacts of intermittent learning. An analytical framework of STFL is proposed and employed to study the learning capability of STFL via its convergence performance. In particular, we have assessed the impact of data delivery outage, intermittent learning mitigation, and statistical heterogeneity of datasets on the convergence performance of STFL. The results provide crucial insights into the design and analysis of STFL-based wireless networks.

cs.LG↗

Projective robustness for quantum channels and measurements and their operational significance

Recently, the projective robustness of quantum states has been introduced in [arXiv:2109.04481(2021)]. It shows that the projective robustness is a useful resource monotone and can comprehensively characterize capabilities and limitations of probabilistic protocols manipulating quantum resources deterministically. In this paper, we will extend the projective robustness to any convex resource theories of quantum channels and measurements. First, We introduce the projective robustness of quantum channels and prove that it satisfies some good properties, especially sub- or supermultiplicativity under any free quantum process. Moreover, we use the projective robustness of channels to give lower bounds on the errors and overheads in any channel resource distillation. Meanwhile, we show that the projective robustness of channels quantifies the maximal advantage that a given channel outperforms all free channels in simultaneous discrimination and exclusion of a fixed state ensemble. Second, we define the projective robustness of quantum measurements and prove that it exactly quantifies the maximal advantage that a given measurement provides over all free measurements in simultaneous discrimination and exclusion of two fixed state ensembles. Finally, within a specific channel resource setting based on measurement incompatibility, we show that the projective robustness of quantum channels coincides with the projective robustness of measurement incompatibility.

quant-ph↗

The Formation of M101-alike Galaxies in the Cold Dark Matter Model

The population of satellite galaxies in a host galaxy is a combination of the cumulative accretion of subhaloes and their associated star formation efficiencies, therefore, the luminosity distribution of satellites provides valuable information of both dark matter properties and star formation physics. Recently, the luminosity function of satellites in nearby Milky Way-mass galaxies has been well measured to satellites as faint as Leo I with $M_{V} \sim -8$. In addition to the finding of the diversity in the satellite luminosity functions, it has been noticed that there is a big gap among the magnitude of satellites in some host galaxies, such as M101, where the gap is around 5 in magnitude, noticeably larger than the prediction from the halo abundance matching method. The reason of this gap is still unknown. In this paper, we use a semi-analytical model of galaxy formation, combined with high-resolution N-body simulation, to investigate the probability and origin of such big gap in M101-alike galaxies. We found that, although M101 analogues are very rare with probability of \sim 0.1%-0.2% in the local universe, their formation is a natural outcome of the CDM model. The gap in magnitude is mainly due to the mass of the accreted subhaloes, not from the stochastic star formation in them. We also found that the gap is correlated with the total satellite mass and host halo mass. By tracing the formation history of M101 type galaxies, we find that they likely formed after $z \sim 1$ due to the newly accreted bright satellites. The gap is not in a stable state, and it will disappear in ~7 Gyr due to mergers of bright satellites with the central galaxy.

astro-ph.GA↗