一种由单层感知器组成的非常简单的通用逼近器的学习规则。

Auer Peter, Burgsteiner Harald, Maass Wolfgang

Chair for Information Technology, University of Leoben, Franz-Josef-Strasse 18, A-8700 Leoben, Austria.

Neural Netw. 2008 Jun;21(5):786-95. doi: 10.1016/j.neunet.2007.12.036. Epub 2007 Dec 31.

One may argue that the simplest type of neural networks beyond a single perceptron is an array of several perceptrons in parallel. In spite of their simplicity, such circuits can compute any Boolean function if one views the majority of the binary perceptron outputs as the binary output of the parallel perceptron, and they are universal approximators for arbitrary continuous functions with values in [0,1] if one views the fraction of perceptrons that output 1 as the analog output of the parallel perceptron. Note that in contrast to the familiar model of a "multi-layer perceptron" the parallel perceptron that we consider here has just binary values as outputs of gates on the hidden layer. For a long time one has thought that there exists no competitive learning algorithm for these extremely simple neural networks, which also came to be known as committee machines. It is commonly assumed that one has to replace the hard threshold gates on the hidden layer by sigmoidal gates (or RBF-gates) and that one has to tune the weights on at least two successive layers in order to achieve satisfactory learning results for any class of neural networks that yield universal approximators. We show that this assumption is not true, by exhibiting a simple learning algorithm for parallel perceptrons - the parallel delta rule (p-delta rule). In contrast to backprop for multi-layer perceptrons, the p-delta rule only has to tune a single layer of weights, and it does not require the computation and communication of analog values with high precision. Reduced communication also distinguishes our new learning rule from other learning rules for parallel perceptrons such as MADALINE. Obviously these features make the p-delta rule attractive as a biologically more realistic alternative to backprop in biological neural circuits, but also for implementations in special purpose hardware. We show that the p-delta rule also implements gradient descent-with regard to a suitable error measure-although it does not require to compute derivatives. Furthermore it is shown through experiments on common real-world benchmark datasets that its performance is competitive with that of other learning approaches from neural networks and machine learning. It has recently been shown [Anthony, M. (2007). On the generalization error of fixed combinations of classifiers. Journal of Computer and System Sciences 73(5), 725-734; Anthony, M. (2004). On learning a function of perceptrons. In Proceedings of the 2004 IEEE international joint conference on neural networks (pp. 967-972): Vol. 2] that one can also prove quite satisfactory bounds for the generalization error of this new learning rule.

有人可能会说，除了单个感知器之外，最简单的神经网络类型是由几个感知器并行组成的阵列。尽管它们很简单，但如果将大多数二元感知器的输出视为并行感知器的二元输出，这样的电路可以计算任何布尔函数；如果将输出为1的感知器的比例视为并行感知器的模拟输出，那么它们对于取值在[0,1]之间的任意连续函数都是通用逼近器。需要注意的是，与常见的“多层感知器”模型不同，我们这里考虑的并行感知器在隐藏层上的门的输出只有二元值。长期以来，人们一直认为对于这些极其简单的神经网络不存在竞争性学习算法，这些神经网络后来也被称为委员会机器。通常认为必须用Sigmoid门（或径向基函数门）替换隐藏层上的硬阈值门，并且必须在至少两个连续层上调整权重，以便对于任何能产生通用逼近器的神经网络类别都能获得令人满意的学习结果。我们通过展示一种用于并行感知器的简单学习算法——并行增量规则（p - delta规则），表明这个假设是不正确的。与多层感知器的反向传播不同，p - delta规则只需要调整单层权重，并且不需要高精度地计算和通信模拟值。减少的通信量也使我们的新学习规则与其他用于并行感知器的学习规则（如MADALINE）有所区别。显然，这些特性使得p - delta规则作为生物神经回路中比反向传播在生物学上更现实的替代方案，以及在专用硬件中的实现都具有吸引力。我们表明，p - delta规则在合适的误差度量下也实现了梯度下降，尽管它不需要计算导数。此外，通过在常见的真实世界基准数据集上的实验表明，它的性能与神经网络和机器学习中的其他学习方法具有竞争力。最近有研究表明[安东尼，M.（2007年）。关于分类器固定组合的泛化误差。《计算机与系统科学杂志》73(5)，725 - 734；安东尼，M.（2004年）。关于学习感知器的函数。在2004年IEEE国际神经网络联合会议论文集（第967 - 972页）：第2卷]，对于这个新学习规则的泛化误差也可以证明相当令人满意的界。

相似文献

A learning rule for very simple universal approximators consisting of a single layer of perceptrons.

Neural Netw. 2008 Jun;21(5):786-95. doi: 10.1016/j.neunet.2007.12.036. Epub 2007 Dec 31.

On the classification capability of sign-constrained perceptrons.

Neural Comput. 2008 Jan;20(1):288-309. doi: 10.1162/neco.2008.20.1.288.

A new backpropagation learning algorithm for layered neural networks with nondifferentiable units.

Neural Comput. 2007 May;19(5):1422-35. doi: 10.1162/neco.2007.19.5.1422.

Direct parallel perceptrons (DPPs): fast analytical calculation of the parallel perceptrons weights with margin control for classification tasks.

IEEE Trans Neural Netw. 2011 Nov;22(11):1837-48. doi: 10.1109/TNN.2011.2169086. Epub 2011 Oct 6.

A forecast-based STDP rule suitable for neuromorphic implementation.

Neural Netw. 2012 Aug;32:3-14. doi: 10.1016/j.neunet.2012.02.018. Epub 2012 Feb 14.

Minimization of error functionals over perceptron networks.

Neural Comput. 2008 Jan;20(1):252-70. doi: 10.1162/neco.2008.20.1.252.

Comparison of universal approximators incorporating partial monotonicity by structure.

Neural Netw. 2010 May;23(4):471-5. doi: 10.1016/j.neunet.2009.09.002. Epub 2009 Sep 17.

On the computational power of threshold circuits with sparse activity.

Neural Comput. 2006 Dec;18(12):2994-3008. doi: 10.1162/neco.2006.18.12.2994.

An integral upper bound for neural network approximation.

Neural Comput. 2009 Oct;21(10):2970-89. doi: 10.1162/neco.2009.04-08-745.

Bounds on the number of hidden neurons in three-layer binary neural networks.

Neural Netw. 2003 Sep;16(7):995-1002. doi: 10.1016/S0893-6080(03)00006-6.

引用本文的文献

High-rate leading spikes in propagating spike sequences predict seizure outcome in surgical patients with temporal lobe epilepsy.

Brain Commun. 2023 Oct 24;5(6):fcad289. doi: 10.1093/braincomms/fcad289. eCollection 2023.

A Cascade BP Neural Network Tuned PID Controller for a High-Voltage Cable-Stripping Robot.

Micromachines (Basel). 2023 Mar 20;14(3):689. doi: 10.3390/mi14030689.

Machine learning application in personalised lung cancer recurrence and survivability prediction.

Comput Struct Biotechnol J. 2022 Apr 4;20:1811-1820. doi: 10.1016/j.csbj.2022.03.035. eCollection 2022.

A Complex-Valued Oscillatory Neural Network for Storage and Retrieval of Multidimensional Aperiodic Signals.

Front Comput Neurosci. 2021 May 24;15:551111. doi: 10.3389/fncom.2021.551111. eCollection 2021.

Surrogate models based on machine learning methods for parameter estimation of left ventricular myocardium.

R Soc Open Sci. 2021 Jan 13;8(1):201121. doi: 10.1098/rsos.201121. eCollection 2021 Jan.

Bio-Inspired Evolutionary Model of Spiking Neural Networks in Ionic Liquid Space.

Front Neurosci. 2019 Nov 8;13:1085. doi: 10.3389/fnins.2019.01085. eCollection 2019.

Prediction of the Tensile Response of Carbon Black Filled Rubber Blends by Artificial Neural Network.

Polymers (Basel). 2018 Jun 9;10(6):644. doi: 10.3390/polym10060644.

Modeling the Temperature Dependence of Dynamic Mechanical Properties and Visco-Elastic Behavior of Thermoplastic Polyurethane Using Artificial Neural Network.

Polymers (Basel). 2017 Oct 18;9(10):519. doi: 10.3390/polym9100519.

An Oscillatory Neural Autoencoder Based on Frequency Modulation and Multiplexing.

Front Comput Neurosci. 2018 Jul 10;12:52. doi: 10.3389/fncom.2018.00052. eCollection 2018.

SuperSpike: Supervised Learning in Multilayer Spiking Neural Networks.

Neural Comput. 2018 Jun;30(6):1514-1541. doi: 10.1162/neco_a_01086. Epub 2018 Apr 13.

Suppr 超能文献

核心技术专利：CN118964589B侵权必究

相似文献

A learning rule for very simple universal approximators consisting of a single layer of perceptrons.

Neural Netw. 2008 Jun;21(5):786-95. doi: 10.1016/j.neunet.2007.12.036. Epub 2007 Dec 31.

On the classification capability of sign-constrained perceptrons.

Neural Comput. 2008 Jan;20(1):288-309. doi: 10.1162/neco.2008.20.1.288.

A new backpropagation learning algorithm for layered neural networks with nondifferentiable units.

Neural Comput. 2007 May;19(5):1422-35. doi: 10.1162/neco.2007.19.5.1422.

Direct parallel perceptrons (DPPs): fast analytical calculation of the parallel perceptrons weights with margin control for classification tasks.

IEEE Trans Neural Netw. 2011 Nov;22(11):1837-48. doi: 10.1109/TNN.2011.2169086. Epub 2011 Oct 6.

A forecast-based STDP rule suitable for neuromorphic implementation.

Neural Netw. 2012 Aug;32:3-14. doi: 10.1016/j.neunet.2012.02.018. Epub 2012 Feb 14.

Minimization of error functionals over perceptron networks.

Neural Comput. 2008 Jan;20(1):252-70. doi: 10.1162/neco.2008.20.1.252.

Comparison of universal approximators incorporating partial monotonicity by structure.

Neural Netw. 2010 May;23(4):471-5. doi: 10.1016/j.neunet.2009.09.002. Epub 2009 Sep 17.

On the computational power of threshold circuits with sparse activity.

Neural Comput. 2006 Dec;18(12):2994-3008. doi: 10.1162/neco.2006.18.12.2994.

An integral upper bound for neural network approximation.

Neural Comput. 2009 Oct;21(10):2970-89. doi: 10.1162/neco.2009.04-08-745.

Bounds on the number of hidden neurons in three-layer binary neural networks.

Neural Netw. 2003 Sep;16(7):995-1002. doi: 10.1016/S0893-6080(03)00006-6.

引用本文的文献

High-rate leading spikes in propagating spike sequences predict seizure outcome in surgical patients with temporal lobe epilepsy.

Brain Commun. 2023 Oct 24;5(6):fcad289. doi: 10.1093/braincomms/fcad289. eCollection 2023.

A Cascade BP Neural Network Tuned PID Controller for a High-Voltage Cable-Stripping Robot.

Micromachines (Basel). 2023 Mar 20;14(3):689. doi: 10.3390/mi14030689.

Machine learning application in personalised lung cancer recurrence and survivability prediction.

Comput Struct Biotechnol J. 2022 Apr 4;20:1811-1820. doi: 10.1016/j.csbj.2022.03.035. eCollection 2022.

A Complex-Valued Oscillatory Neural Network for Storage and Retrieval of Multidimensional Aperiodic Signals.

Front Comput Neurosci. 2021 May 24;15:551111. doi: 10.3389/fncom.2021.551111. eCollection 2021.

Surrogate models based on machine learning methods for parameter estimation of left ventricular myocardium.

R Soc Open Sci. 2021 Jan 13;8(1):201121. doi: 10.1098/rsos.201121. eCollection 2021 Jan.

Bio-Inspired Evolutionary Model of Spiking Neural Networks in Ionic Liquid Space.

Front Neurosci. 2019 Nov 8;13:1085. doi: 10.3389/fnins.2019.01085. eCollection 2019.

Prediction of the Tensile Response of Carbon Black Filled Rubber Blends by Artificial Neural Network.

Polymers (Basel). 2018 Jun 9;10(6):644. doi: 10.3390/polym10060644.

Modeling the Temperature Dependence of Dynamic Mechanical Properties and Visco-Elastic Behavior of Thermoplastic Polyurethane Using Artificial Neural Network.

Polymers (Basel). 2017 Oct 18;9(10):519. doi: 10.3390/polym9100519.

An Oscillatory Neural Autoencoder Based on Frequency Modulation and Multiplexing.

Front Comput Neurosci. 2018 Jul 10;12:52. doi: 10.3389/fncom.2018.00052. eCollection 2018.

SuperSpike: Supervised Learning in Multilayer Spiking Neural Networks.

Neural Comput. 2018 Jun;30(6):1514-1541. doi: 10.1162/neco_a_01086. Epub 2018 Apr 13.

A learning rule for very simple universal approximators consisting of a single layer of perceptrons.

作者信息

机构信息

出版信息

相似文献

引用本文的文献

文献AI研究员

用中文搜PubMed

文档翻译

Suppr 超能文献

相似文献

引用本文的文献