Training spiking multi-layer networks with surrogate gradients on an analog neuromorphic substrate

Cramer, Benjamin; Billaudelle, Sebastian; Kanya, Simeon; Leibfried, Aron; Grübl, Andreas; Karasenko, Vitali; Pehle, Christian; Schreiber, Korbinian; Stradmann, Yannik; Weis, Johannes; Schemmel, Johannes; Zenke, Friedemann

Computer Science > Neural and Evolutionary Computing

arXiv:2006.07239v1 (cs)

[Submitted on 12 Jun 2020 (this version), latest version 20 May 2021 (v3)]

Title:Training spiking multi-layer networks with surrogate gradients on an analog neuromorphic substrate

Authors:Benjamin Cramer, Sebastian Billaudelle, Simeon Kanya, Aron Leibfried, Andreas Grübl, Vitali Karasenko, Christian Pehle, Korbinian Schreiber, Yannik Stradmann, Johannes Weis, Johannes Schemmel, Friedemann Zenke

View PDF

Abstract:Spiking neural networks are nature's solution for parallel information processing with high temporal precision at a low metabolic energy cost. To that end, biological neurons integrate inputs as an analog sum and communicate their outputs digitally as spikes, i.e., sparse binary events in time. These architectural principles can be mirrored effectively in analog neuromorphic hardware. Nevertheless, training spiking neural networks with sparse activity on hardware devices remains a major challenge. Primarily this is due to the lack of suitable training methods that take into account device-specific imperfections and operate at the level of individual spikes instead of firing rates. To tackle this issue, we developed a hardware-in-the-loop strategy to train multi-layer spiking networks using surrogate gradients on the analog BrainScales-2 chip. Specifically, we used the hardware to compute the forward pass of the network, while the backward pass was computed in software. We evaluated our approach on downscaled 16x16 versions of the MNIST and the fashion MNIST datasets in which spike latencies encoded pixel intensities. The analog neuromorphic substrate closely matched the performance of equivalently sized networks implemented in software. It is capable of processing 70 k patterns per second with a power consumption of less than 300 mW. Added activity regularization resulted in sparse network activity with about 20 spikes per input, at little to no reduction in classification performance. Thus, overall, our work demonstrates low-energy spiking network processing on an analog neuromorphic substrate and sets several new benchmarks for hardware systems in terms of classification accuracy, processing speed, and efficiency. Importantly, our work emphasizes the value of hardware-in-the-loop training and paves the way toward energy-efficient information processing on non-von-Neumann architectures.

Subjects:	Neural and Evolutionary Computing (cs.NE); Emerging Technologies (cs.ET); Machine Learning (cs.LG); Neurons and Cognition (q-bio.NC); Machine Learning (stat.ML)
Cite as:	arXiv:2006.07239 [cs.NE]
	(or arXiv:2006.07239v1 [cs.NE] for this version)
	https://doi.org/10.48550/arXiv.2006.07239

Submission history

From: Sebastian Billaudelle [view email]
[v1] Fri, 12 Jun 2020 14:45:12 UTC (5,707 KB)
[v2] Mon, 15 Mar 2021 17:52:22 UTC (1,928 KB)
[v3] Thu, 20 May 2021 14:13:26 UTC (916 KB)

Computer Science > Neural and Evolutionary Computing

Title:Training spiking multi-layer networks with surrogate gradients on an analog neuromorphic substrate

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Neural and Evolutionary Computing

Title:Training spiking multi-layer networks with surrogate gradients on an analog neuromorphic substrate

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators