偏差在神经网络中的作用是什么?

我知道梯度下降和反向传播算法。我不明白的是:什么时候使用偏见是重要的，你如何使用它?

例如，在映射AND函数时，当我使用两个输入和一个输出时，它不会给出正确的权重。然而，当我使用三个输入(其中一个是偏差)时，它给出了正确的权重。

当前回答

Two different kinds of parameters can be adjusted during the training of an ANN, the weights and the value in the activation functions. This is impractical and it would be easier if only one of the parameters should be adjusted. To cope with this problem a bias neuron is invented. The bias neuron lies in one layer, is connected to all the neurons in the next layer, but none in the previous layer and it always emits 1. Since the bias neuron emits 1 the weights, connected to the bias neuron, are added directly to the combined sum of the other weights (equation 2.1), just like the t value in the activation functions.1

它不实用的原因是，您同时调整权重和值，因此对权重的任何更改都会抵消对先前数据实例有用的值的更改……在不改变值的情况下添加偏置神经元可以让你控制层的行为。

此外，偏差允许您使用单个神经网络来表示类似的情况。考虑由以下神经网络表示的AND布尔函数:

(来源:aihorizon.com)

W0对应于b。 W1对应x1。 W2对应于x2。

A single perceptron can be used to represent many boolean functions. For example, if we assume boolean values of 1 (true) and -1 (false), then one way to use a two-input perceptron to implement the AND function is to set the weights w0 = -3, and w1 = w2 = .5. This perceptron can be made to represent the OR function instead by altering the threshold to w0 = -.3. In fact, AND and OR can be viewed as special cases of m-of-n functions: that is, functions where at least m of the n inputs to the perceptron must be true. The OR function corresponds to m = 1 and the AND function to m = n. Any m-of-n function is easily represented using a perceptron by setting all input weights to the same value (e.g., 0.5) and then setting the threshold w0 accordingly. Perceptrons can represent all of the primitive boolean functions AND, OR, NAND ( 1 AND), and NOR ( 1 OR). Machine Learning- Tom Mitchell)

阈值是偏置，w0是与偏置/阈值神经元相关的权重。

2010-03-19 21:38:55

其他回答

在我的硕士论文中的几个实验中(例如第59页)，我发现偏差可能对第一层很重要，但特别是在最后的完全连接层，它似乎没有发挥很大的作用。

这可能高度依赖于网络架构/数据集。

2017-08-01 17:09:42

简单来说，如果你有y=w1*x，其中y是你的输出，w1是权重，想象一个条件，x=0，那么y=w1*x等于0。

如果你想要更新你的权重，你必须计算delw=target-y的变化量，其中target是你的目标输出。在这种情况下，'delw'将不会改变，因为y被计算为0。所以，假设你可以添加一些额外的值，这将有助于y = w1x + w01，其中偏差=1，权重可以调整以获得正确的偏差。考虑下面的例子。

就直线斜率而言，截距是线性方程的一种特殊形式。

Y = mx + b

检查图像

图像