arXiv ScienceSearch

arXiv subjects

Xiantao Jiang

Publications and source records attributed to Xiantao Jiang.

5 recordsLinked to original sources

CPR-IE:A Compression-Prediction-Resource Intelligence Efficiency Metric

Comparing intelligent systems under deployment constraints requires more than predictiveaccuracy.This paper develops Compression-Prediction-Resource Intelligence Efficiency (CPR-IE) as a protocol-relative ordering by representational economy, predictive quality, and resourceburden. The analysis separates two questions-how raw resource consumption is represented, andhow the resulting attributes are aggregated. Proportional-increment composition uniquely yieldslogarithmic cumulative burden, and context-independent ratio response yields power responsesto compression, prediction, and burden; with reference normalization the representation is I(C,P,T).We prove Pareto consistency, unit invariance, boundary behavior, trade-off identities, ranking-stability regions, and cross-task aggregation. A translog parent model makes interaction restrictions explicit, and further results establish cardinal and ordinal identification, sub-Gaussianfinite-sample ranking guarantees, robust selection under exponent uncertainty, and deterministicregret bounds. Minimum description length, algorithmic complexity, proper scoring rules, varia-tional inference, and Landauer's principle motivate measurement choices but do not entail theformula. CPR-IE is a constructed efficiency representation, not a universal law or a definition ofintelligence itself.

cs.AI

QuIDE: Mastering the Quantized Intelligence Trade-off via Active Optimization

There is currently no unified metric for evaluating the efficiency of quantized neural networks. We propose QuIDE, built around the Intelligence Index I = (C x P)/log_2(T+1), which collapses the compression-accuracy-latency trade-off into a single score. Experiments across six settings -- SimpleCNN (MNIST, CIFAR), ResNet-18 (ImageNet-1K), and Llama-3-8B -- show a task-dependent Pareto Knee. 4-bit quantization is optimal for MNIST and large LLMs, while 8-bit is the sweet spot for complex CNN tasks (ResNet-18 on ImageNet), where 4-bit PTQ collapses accuracy catastrophically. The accuracy-gated variant I' correctly flags these non-viable configurations that the raw I would reward. QuIDE provides a reproducible evaluation protocol and a ready-to-use fitness function for mixed-precision search.

cs.LG

Doppler-induced continuous spectral broadening of ultraviolet lasers

We propose a compact scheme based on ultrafast-rotating phase plates (URPPs) to achieve continuous spectral broadening of ultraviolet lasers. The rapid rotation elements behave as a random oscillator which induces Doppler frequency shift into the ultraviolet lasers. As an example, for a disk-shaped phase plate, with the beam acting on the edge at a radius of 10 cm, a rotation frequency of 1 kHz, and a phase-element size of 10 nm, the continuous spectral broadening reaches 0.07%. Further increasing the rotation speed or reducing the phase-element can lead to greater spectral broadening. When multiple URPPs are arranged in series, the superimposed spatiotemporal modulation further enhances the continuous spectral broadening and achieves more effective speckle smoothing. The scheme is applicable to broadening the independent spectrum of optical frequency combs as well as to the mitigation of laser-plasma instabilities in inertial fusion energy.

physics.optics

Leveraging Cross-Attention Transformer and Multi-Feature Fusion for Cross-Linguistic Speech Emotion Recognition

Speech Emotion Recognition (SER) plays a crucial role in enhancing human-computer interaction. Cross-Linguistic SER (CLSER) has been a challenging research problem due to significant variability in linguistic and acoustic features of different languages. In this study, we propose a novel approach HuMP-CAT, which combines HuBERT, MFCC, and prosodic characteristics. These features are fused using a cross-attention transformer (CAT) mechanism during feature extraction. Transfer learning is applied to gain from a source emotional speech dataset to the target corpus for emotion recognition. We use IEMOCAP as the source dataset to train the source model and evaluate the proposed method on seven datasets in five languages (e.g., English, German, Spanish, Italian, and Chinese). We show that, by fine-tuning the source model with a small portion of speech from the target datasets, HuMP-CAT achieves an average accuracy of 78.75% across the seven datasets, with notable performance of 88.69% on EMODB (German language) and 79.48% on EMOVO (Italian language). Our extensive evaluation demonstrates that HuMP-CAT outperforms existing methods across multiple target languages.

eess.AS

Multicolor Graphdiyne Random Lasers

By breaking the restriction of mirrors, random lasers from a disordered medium have found unique applications spanning from displays, spectroscopy, biomedical treatments, to Li-Fi.Gain media in the form of two-dimension with distinct physical and chemical properties may lead to the next-generation of random lasers. Graphdiyne, a 2D graphene allotrope with intrigued carbon hybridization, atomic lattice, and optoelectronic properties, has attracted increasing attention recently. Herein, the photon emission characteristics and photo-carrier dynamics in graphdiyne are systematically studied, and the multicolor random lasers have been unprecedently realized using graphdiyne nanosheets as the gain. Considering the well bio-compatibility of graphdiyne, these results may look ahead a plethora of potential applications in the nanotechnology platform based on graphdiyne.

physics.optics