arXiv ScienceSearch

arXiv subjects

Mathilde Hary

Publications and source records attributed to Mathilde Hary.

5 recordsLinked to original sources

Power law scaling for classification accuracy in physical neural networks

Physical neural networks (PNNs) harness the intrinsic complexity of physical systems to perform neural computation, potentially at speeds and energy efficiencies inaccessible to conventional digital hardware. Yet, a principled framework for quantifying and predicting their computing accuracy across diverse substrates has remained elusive. Here we introduce the Hotelling Trace Criterion (HTC), a task-conditioned measure of PNN- state separability that can be evaluated without training. We demonstrate that it predicts PNN classification performance with high fidelity across highly nonlinear optical fibres, vertical-cavity surface-emitting lasers, and coupled nonlinear oscillator networks, for benchmark tasks of different difficulty. Classification loss follows a power law in HTC, with Pearson correlation coefficients exceeding 0.99 for MNIST and $\approx$0.97 for Fashion-MNIST, noteworthy experimental and simulated data from physically distinct systems collapse onto a single scaling curve determined by the task rather than the substrate. Applying HTC layer-by-layer during training further reveals that gradient-based optimisation distributes representational capacity unevenly across PNN layers, providing a quantitative diagnostic of training and architecture efficiency invisible to standard loss monitoring. Crucially, once the scaling exponent is established from a small number of trained calibration systems, all further performance predictions require no training since performance can be derived from the much more efficient HTC measurement. These results establish HTC as a substrate-agnostic figure of merit for comparing and scaling PNNs, advancing the field further towards a complete theory connecting fundamental hardware parameters to task performance through universal scaling laws.

cs.ET

Limits of nonlinear and dispersive fiber propagation for an optical fiber-based extreme learning machine

We report a generalized nonlinear Schr\"odinger equation simulation model of an extreme learning machine (ELM) based on optical fiber propagation. Using the MNIST handwritten digit dataset as a benchmark, we study how accuracy depends on propagation dynamics, as well as parameters governing spectral encoding, readout, and noise. For this dataset and with quantum noise limited input, test accuracies of : over 91% and 93% are found for propagation in the anomalous and normal dispersion regimes respectively. Our results also suggest that quantum noise on the input pulses introduces an intrinsic penalty to ELM performance.

physics.optics

Principles and Metrics of Extreme Learning Machines Using a Highly Nonlinear Fiber

Optical computing offers potential for ultra high-speed and low latency computation by leveraging the intrinsic properties of light. Here, we explore the use of highly nonlinear optical fibers (HNLFs) as platforms for optical computing based on the concept of Extreme Learning Machines. Task-independent evaluations are introduced to the field for the first time and focus on the fundamental metrics of effective dimensionality and consistency, which we experimentally characterize for different nonlinear and dispersive conditions. We show that input power and fiber characteristics significantly influence the dimensionality of the computational system, with longer fibers and higher dispersion producing up to 100 principal components (PCs) at input power levels of 30 mW, where the PC correspond to the linearly independent dimensions of the system. The spectral distribution of the PC's eigenvectors reveals that the high-dimensional dynamics facilitating computing through dimensionality expansion are located within 40~nm of the pump wavelength at 1560~nm, providing general insight for computing with nonlinear Schr\"odinger equation systems. Task-dependent results demonstrate the effectiveness of HNLFs in classifying MNIST dataset images. Using input data compression through PC analysis, we inject MNIST images of various input dimensionality into the system and study the impact of input power upon classification accuracy. At optimized power levels we achieve a classification test accuracy of 88\%, significantly surpassing the baseline of 83.7\% from linear systems. Noteworthy, we find that best performance is not obtained at maximal input power, i.e. maximal system dimensionality, but at more than one order of magnitude lower. The same is confirmed regarding the MNIST image's compression, where accuracy is substantially improved when strongly compressing the image to less than 50 PCs.

physics.optics

Tailored supercontinuum generation using genetic algorithmoptimized Fourier domain pulse shaping

We report the generation of spectrally-tailored supercontinuum using Fourier-domain pulse shaping of femtosecond pulses injected into a highly nonlinear fiber controlled by a genetic algorithm. User-selectable spectral enhancement is demonstrated over the 1550-2000~nm wavelength range, with the ability to both select a target central wavelength and a target bandwidth in the range 1--5~nm. The spectral enhancement factor relative to unshaped input pulses is typically $\sim$5--20 in the range 1550--1800~nm and increases for longer wavelengths, exceeding a factor of 160 around 2000~nm. We also demonstrate results where the genetic algorithm is applied to the enhancement of up to four wavelengths simultaneously.

physics.optics

A feed-forward neural network as a nonlinear dynamics integrator for supercontinuum generation

The nonlinear propagation of ultrashort pulses in optical fiber depends sensitively on both input pulse and fiber parameters. As a result, optimizing propagation for specific applications generally requires time-consuming simulations based on sequential integration of the generalized nonlinear Schr\"odinger equation (GNLSE). Here, we train a feed-forward neural network to learn the differential propagation dynamics of the GNLSE, allowing emulation of direct numerical integration of fiber propagation, and particularly the highly complex case of supercontinuum generation. Comparison with a recurrent neural network shows that the feed-forward approach yields faster training and computation, and reduced memory requirements. The approach is generic and can be extended to other physical systems.

physics.comp-ph