通过大幅降低梯度估计偏差将平衡传播扩展到深度卷积神经网络

Scaling Equilibrium Propagation to Deep ConvNets by Drastically Reducing Its Gradient Estimator Bias.

作者信息

Laborieux Axel, Ernoult Maxence, Scellier Benjamin, Bengio Yoshua, Grollier Julie, Querlioz Damien

机构信息

Université Paris-Saclay, CNRS, Centre de Nanosciences et de Nanotechnologies, Palaiseau, France.

Unité Mixte de Physique, CNRS, Thales, Université Paris-Saclay, Palaiseau, France.

出版信息

Front Neurosci. 2021 Feb 18;15:633674. doi: 10.3389/fnins.2021.633674. eCollection 2021.

DOI:10.3389/fnins.2021.633674

PMID:33679315

原文链接:https://pmc.ncbi.nlm.nih.gov/articles/PMC7930909/

Abstract

Equilibrium Propagation is a biologically-inspired algorithm that trains convergent recurrent neural networks with a local learning rule. This approach constitutes a major lead to allow learning-capable neuromophic systems and comes with strong theoretical guarantees. Equilibrium propagation operates in two phases, during which the network is let to evolve freely and then "nudged" toward a target; the weights of the network are then updated based solely on the states of the neurons that they connect. The weight updates of Equilibrium Propagation have been shown mathematically to approach those provided by Backpropagation Through Time (BPTT), the mainstream approach to train recurrent neural networks, when nudging is performed with infinitely small strength. In practice, however, the standard implementation of Equilibrium Propagation does not scale to visual tasks harder than MNIST. In this work, we show that a bias in the gradient estimate of equilibrium propagation, inherent in the use of finite nudging, is responsible for this phenomenon and that canceling it allows training deep convolutional neural networks. We show that this bias can be greatly reduced by using symmetric nudging (a positive nudging and a negative one). We also generalize Equilibrium Propagation to the case of cross-entropy loss (by opposition to squared error). As a result of these advances, we are able to achieve a test error of 11.7% on CIFAR-10, which approaches the one achieved by BPTT and provides a major improvement with respect to the standard Equilibrium Propagation that gives 86% test error. We also apply these techniques to train an architecture with unidirectional forward and backward connections, yielding a 13.2% test error. These results highlight equilibrium propagation as a compelling biologically-plausible approach to compute error gradients in deep neuromorphic systems.

摘要

平衡传播是一种受生物启发的算法，它使用局部学习规则训练收敛递归神经网络。这种方法是实现具有学习能力的神经形态系统的一项重大突破，并具有强大的理论保障。平衡传播分两个阶段运行，在此期间网络可自由演化，然后“微调”至目标；网络权重随后仅根据它们所连接神经元的状态进行更新。数学证明，当以无限小的强度进行微调时，平衡传播的权重更新接近通过时间反向传播（BPTT）（训练递归神经网络的主流方法）所提供的更新。然而在实践中，平衡传播的标准实现无法扩展到比MNIST更难的视觉任务。在这项工作中，我们表明，有限微调中固有的平衡传播梯度估计偏差是造成这种现象的原因，消除该偏差可实现深度卷积神经网络的训练。我们表明，使用对称微调（正向和负向微调）可大幅减少这种偏差。我们还将平衡传播推广到交叉熵损失的情况（与平方误差相对）。由于这些进展，我们在CIFAR-10上实现了11.7%的测试误差，接近BPTT所达到的误差，并相对于给出86%测试误差的标准平衡传播有了重大改进。我们还应用这些技术训练了一种具有单向向前和向后连接的架构，测试误差为13.2%。这些结果突出了平衡传播作为一种在深度神经形态系统中计算误差梯度的引人注目的生物学上合理的方法。

https://cdn.ncbi.nlm.nih.gov/pmc/blobs/9b2b/7930909/0549f13bc48b/fnins-15-633674-g0001.jpg

相似文献

Scaling Equilibrium Propagation to Deep ConvNets by Drastically Reducing Its Gradient Estimator Bias.

Front Neurosci. 2021 Feb 18;15:633674. doi: 10.3389/fnins.2021.633674. eCollection 2021.

Equilibrium Propagation: Bridging the Gap between Energy-Based Models and Backpropagation.

Front Comput Neurosci. 2017 May 4;11:24. doi: 10.3389/fncom.2017.00024. eCollection 2017.

Equilibrium Propagation for Memristor-Based Recurrent Neural Networks.

Front Neurosci. 2020 Mar 24;14:240. doi: 10.3389/fnins.2020.00240. eCollection 2020.

Biologically Plausible Training Mechanisms for Self-Supervised Learning in Deep Networks.

Front Comput Neurosci. 2022 Mar 21;16:789253. doi: 10.3389/fncom.2022.789253. eCollection 2022.

Layer-Skipping Connections Improve the Effectiveness of Equilibrium Propagation on Layered Networks.

Front Comput Neurosci. 2021 May 17;15:627357. doi: 10.3389/fncom.2021.627357. eCollection 2021.

Biologically-inspired neuronal adaptation improves learning in neural networks.

Commun Integr Biol. 2023 Jan 17;16(1):2163131. doi: 10.1080/19420889.2022.2163131. eCollection 2023.

Biologically plausible deep learning - But how far can we go with shallow networks?

Neural Netw. 2019 Oct;118:90-101. doi: 10.1016/j.neunet.2019.06.001. Epub 2019 Jun 20.

Deep Learning With Asymmetric Connections and Hebbian Updates.

Front Comput Neurosci. 2019 Apr 4;13:18. doi: 10.3389/fncom.2019.00018. eCollection 2019.

Tuning Convolutional Spiking Neural Network With Biologically Plausible Reward Propagation.

IEEE Trans Neural Netw Learn Syst. 2022 Dec;33(12):7621-7631. doi: 10.1109/TNNLS.2021.3085966. Epub 2022 Nov 30.

Deep convolutional neural network and IoT technology for healthcare.

Digit Health. 2024 Jan 17;10:20552076231220123. doi: 10.1177/20552076231220123. eCollection 2024 Jan-Dec.

引用本文的文献

Self-Contrastive Forward-Forward algorithm.

Nat Commun. 2025 Jul 1;16(1):5978. doi: 10.1038/s41467-025-61037-0.

Training an Ising machine with equilibrium propagation.

Nat Commun. 2024 Apr 30;15(1):3671. doi: 10.1038/s41467-024-46879-4.

Memristor Crossbar Circuits Implementing Equilibrium Propagation for On-Device Learning.

Micromachines (Basel). 2023 Jul 3;14(7):1367. doi: 10.3390/mi14071367.

Energy-based analog neural network framework.

Front Comput Neurosci. 2023 Mar 3;17:1114651. doi: 10.3389/fncom.2023.1114651. eCollection 2023.

Cognitive and plastic recurrent neural network clock model for the judgment of time and its variations.

Sci Rep. 2023 Mar 8;13(1):3852. doi: 10.1038/s41598-023-30894-4.

Biologically-inspired neuronal adaptation improves learning in neural networks.

Commun Integr Biol. 2023 Jan 17;16(1):2163131. doi: 10.1080/19420889.2022.2163131. eCollection 2023.

Combining backpropagation with Equilibrium Propagation to improve an Actor-Critic reinforcement learning framework.

Front Comput Neurosci. 2022 Aug 23;16:980613. doi: 10.3389/fncom.2022.980613. eCollection 2022.

Neurons learn by predicting future activity.

Nat Mach Intell. 2022 Jan;4(1):62-72. doi: 10.1038/s42256-021-00430-y. Epub 2022 Jan 25.

Deep physical neural networks trained with backpropagation.

Nature. 2022 Jan;601(7894):549-555. doi: 10.1038/s41586-021-04223-6. Epub 2022 Jan 26.

Cell-type-specific neuromodulation guides synaptic credit assignment in a spiking neural network.

Proc Natl Acad Sci U S A. 2021 Dec 21;118(51). doi: 10.1073/pnas.2111821118.

本文引用的文献

Burst-dependent synaptic plasticity can coordinate learning in hierarchical circuits.

Nat Neurosci. 2021 Jul;24(7):1010-1019. doi: 10.1038/s41593-021-00857-x. Epub 2021 May 13.

EqSpike: spike-driven equilibrium propagation for neuromorphic implementations.

iScience. 2021 Feb 20;24(3):102222. doi: 10.1016/j.isci.2021.102222. eCollection 2021 Mar 19.

Backpropagation and the brain.

Nat Rev Neurosci. 2020 Jun;21(6):335-346. doi: 10.1038/s41583-020-0277-3. Epub 2020 Apr 17.

Equilibrium Propagation for Memristor-Based Recurrent Neural Networks.

Front Neurosci. 2020 Mar 24;14:240. doi: 10.3389/fnins.2020.00240. eCollection 2020.

Digital Biologically Plausible Implementation of Binarized Neural Networks With Differential Hafnium Oxide Resistive Memory Arrays.

Front Neurosci. 2020 Jan 9;13:1383. doi: 10.3389/fnins.2019.01383. eCollection 2019.

A deep learning framework for neuroscience.

Nat Neurosci. 2019 Nov;22(11):1761-1770. doi: 10.1038/s41593-019-0520-2. Epub 2019 Oct 28.

Equivalence of Equilibrium Propagation and Recurrent Backpropagation.

Neural Comput. 2019 Feb;31(2):312-329. doi: 10.1162/neco_a_01160. Epub 2018 Dec 21.

Equilibrium Propagation: Bridging the Gap between Energy-Based Models and Backpropagation.

Front Comput Neurosci. 2017 May 4;11:24. doi: 10.3389/fncom.2017.00024. eCollection 2017.

Random synaptic feedback weights support error backpropagation for deep learning.

Nat Commun. 2016 Nov 8;7:13276. doi: 10.1038/ncomms13276.

The graph neural network model.

IEEE Trans Neural Netw. 2009 Jan;20(1):61-80. doi: 10.1109/TNN.2008.2005605. Epub 2008 Dec 9.

文献AI研究员

20分钟写一篇综述，助力文献阅读效率提升50倍。

立即体验

用中文搜PubMed

大模型驱动的PubMed中文搜索引擎

马上搜索

文档翻译

学术文献翻译模型，支持多种主流文档格式。

立即体验

通过大幅降低梯度估计偏差将平衡传播扩展到深度卷积神经网络

Scaling Equilibrium Propagation to Deep ConvNets by Drastically Reducing Its Gradient Estimator Bias.

作者信息

机构信息

出版信息

相似文献

引用本文的文献

本文引用的文献

文献AI研究员

用中文搜PubMed

文档翻译

Suppr 超能文献

相似文献

引用本文的文献

本文引用的文献