2.1ETJun 30
Power law scaling for classification accuracy in physical neural networksAndrei V. Ermolaev, Mathilde Hary, Anas Skalli et al.
Physical neural networks (PNNs) harness the intrinsic complexity of physical systems to perform neural computation, potentially at speeds and energy efficiencies inaccessible to conventional digital hardware. Yet, a principled framework for quantifying and predicting their computing accuracy across diverse substrates has remained elusive. Here we introduce the Hotelling Trace Criterion (HTC), a task-conditioned measure of PNN- state separability that can be evaluated without training. We demonstrate that it predicts PNN classification performance with high fidelity across highly nonlinear optical fibres, vertical-cavity surface-emitting lasers, and coupled nonlinear oscillator networks, for benchmark tasks of different difficulty. Classification loss follows a power law in HTC, with Pearson correlation coefficients exceeding 0.99 for MNIST and $\approx$0.97 for Fashion-MNIST, noteworthy experimental and simulated data from physically distinct systems collapse onto a single scaling curve determined by the task rather than the substrate. Applying HTC layer-by-layer during training further reveals that gradient-based optimisation distributes representational capacity unevenly across PNN layers, providing a quantitative diagnostic of training and architecture efficiency invisible to standard loss monitoring. Crucially, once the scaling exponent is established from a small number of trained calibration systems, all further performance predictions require no training since performance can be derived from the much more efficient HTC measurement. These results establish HTC as a substrate-agnostic figure of merit for comparing and scaling PNNs, advancing the field further towards a complete theory connecting fundamental hardware parameters to task performance through universal scaling laws.
13.0LGMar 21, 2025
Model-free front-to-end training of a large high performance laser neural networkAnas Skalli, Satoshi Sunada, Mirko Goldmann et al.
Artificial neural networks (ANNs), have become ubiquitous and revolutionized many applications ranging from computer vision to medical diagnoses. However, they offer a fundamentally connectionist and distributed approach to computing, in stark contrast to classical computers that use the von Neumann architecture. This distinction has sparked renewed interest in developing unconventional hardware to support more efficient implementations of ANNs, rather than merely emulating them on traditional systems. Photonics stands out as a particularly promising platform, providing scalability, high speed, energy efficiency, and the ability for parallel information processing. However, fully realized autonomous optical neural networks (ONNs) with in-situ learning capabilities are still rare. In this work, we demonstrate a fully autonomous and parallel ONN using a multimode vertical cavity surface emitting laser (VCSEL) using off-the-shelf components. Our ONN is highly efficient and is scalable both in network size and inference bandwidth towards the GHz range. High performance hardware-compatible optimization algorithms are necessary in order to minimize reliance on external von Neumann computers to fully exploit the potential of ONNs. As such we present and extensively study several algorithms which are broadly compatible with a wide range of systems. We then apply these algorithms to optimize our ONN, and benchmark them using the MNIST dataset. We show that our ONN can achieve high accuracy and convergence efficiency, even under limited hardware resources. Crucially, we compare these different algorithms in terms of scaling and optimization efficiency in term of convergence time which is crucial when working with limited external resources. Our work provides some guidance for the design of future ONNs as well as a simple and flexible way to train them.