【问题标题】:How do you change rank of tf.random_normal as a shape您如何将 tf.random_normal 的等级更改为形状
【发布时间】:2017-07-05 21:05:07
【问题描述】:

我是 tensorflow 的新手,我正在学习 sentdex 的教程。 无论我解决了多少语法问题,我都会不断收到相同的错误。

ValueError: Shape must be rank 1 but is rank 0 for 
'random_normal_7/RandomStandardNormal' (op: 'RandomStandardNormal') 
with input shapes: []

我相信问题就在这里,但我不知道如何解决它。

def neural_network_model(data):
hidden_1_layer = {'weights': tf.Variable(tf.random_normal([784, 
n_nodes_hl1])),
                  'biases': 
tf.Variable(tf.random_normal([n_nodes_hl1]))}

hidden_2_layer = {'weights': tf.Variable(tf.random_normal([n_nodes_hl1, 
n_nodes_hl2])),
                  'biases': 
tf.Variable(tf.random_normal([n_nodes_hl2]))}

hidden_3_layer = {'weights': tf.Variable(tf.random_normal([n_nodes_hl2, 
n_nodes_hl3])),
                  'biases': 
tf.Variable(tf.random_normal([n_nodes_hl3]))}

output_layer = {'weights': tf.Variable(tf.random_normal([n_nodes_hl3, 
n_classes])),
                'biases': tf.Variable(tf.random_normal(n_classes))}

我的整个代码是

import tensorflow as tf
from tensorflow.examples.tutorials.mnist import input_data

mnist = input_data.read_data_sets("/tmp/ data/", one_hot=True)

n_nodes_hl1 = 500
n_nodes_hl2 = 500
n_nodes_hl3 = 500

n_classes = 10
batch_size = 100

x = tf.placeholder('float', [None, 784])
y = tf.placeholder('float')


def neural_network_model(data):
hidden_1_layer = {'weights': tf.Variable(tf.random_normal([784, 
n_nodes_hl1])),
                  'biases': 
tf.Variable(tf.random_normal([n_nodes_hl1]))}

hidden_2_layer = {'weights': tf.Variable(tf.random_normal([n_nodes_hl1, 
n_nodes_hl2])),
                  'biases': 
tf.Variable(tf.random_normal([n_nodes_hl2]))}

hidden_3_layer = {'weights': tf.Variable(tf.random_normal([n_nodes_hl2, 
n_nodes_hl3])),
                  'biases': 
tf.Variable(tf.random_normal([n_nodes_hl3]))}

output_layer = {'weights': tf.Variable(tf.random_normal([n_nodes_hl3, 
n_classes])),
                'biases': tf.Variable(tf.random_normal(n_classes))}

l1 = tf.add(tf.matmul(data, hidden_1_layer['weights']), 
hidden_1_layer['biases'])
l1 = tf.nn.relu(l1)

l2 = tf.add(tf.matmul(data, hidden_2_layer['weights']), 
hidden_2_layer['biases'])
l2 = tf.nn.relu(l2)

l3 = tf.add(tf.matmul(data, hidden_3_layer['weights']), 
hidden_3_layer['biases'])
l3 = tf.nn.relu(l3)

output = tf.matmul(l3, output_layer['weights']) + 
output_layer['biases']

return output


def train_neural_network(x):
prediction = neural_network_model(x)
cost = tf.reduce_mean(tf.nn.softmax_cross_entropy_with_logits
(logits=prediction, labels=y))
optimizer = tf.train.AdamOptimizer().minimize(cost)

hm_epochs = 10

with tf.Session() as sess:
    sess.run(tf.global_variables_initializer())

    for epoch in range(hm_epochs):
        epoch_loss = 0
        for _ in range(int(mnist.train.num_examples / batch_size)):
            epoch_x, epoch_y = mnist.train.next_batch(batch_size)
            _, c = sess.run([optimizer, cost], feed_dict={x: epoch_x,
y: epoch_y})
            epoch_loss += c
        print('Epoch', epoch, 'completed out of', hm_epochs, 'loss:', 
epoch_loss)

    correct = tf.equal(tf.argmax(prediction, 1), tf.argmax(y, 1))
    accuracy = tf.reduce_mean(tf.cast(correct, 'float'))
    print('Accuracy:', accuracy.eval({x: mnist.test.images, y: 
mnist.test.labels}))


train_neural_network(x)

【问题讨论】:

  • 你在你的 output_layer 偏差中尝试过tf.Variable(tf.random_normal([n_classes]) 吗?它似乎缺少代码中的括号

标签: python python-3.x tensorflow traceback


【解决方案1】:

tf.random_normal()shape)的第一个参数必须是一维张量或整数列表,表示随机张量每个维度的长度。假设n_classes 是一个整数,将tf.random_normal(n_classes) 替换为tf.random_normal([n_classes]) 应该可以修复错误。

【讨论】:

  • 这绝对解决了错误。但是一个新的出现了。我得到的不是等级问题,而是尺寸问题。错误是“ValueError:尺寸必须相等,但是对于输入形状为 [?,784]、[500,500] 的 'MatMul_1'(操作:'MatMul')的尺寸必须是 784 和 500。”
  • 我认为问题在于这一行:l2 = tf.add(tf.matmul(data, hidden_2_layer['weights']),您将data 乘以第 2 层的权重,但您应该将l1 乘以层的权重2(第3层的定义也有同样的bug)。
  • 我认为这不是问题所在。相反,我用加号替换了权重和偏差之间的层“,”。解决了巨大的回溯问题。现在我收到一个类型错误:l1 = tf.add(tf.matmul(data, hidden_​​1_layer['weights']) + hidden_​​1_layer['biases']) TypeError: add() missing 1 required positional argument: 'y'
【解决方案2】:

您必须从第 1 层连接到第 2 层,好像您将数据传递到所有层而不连接它们, 你的代码

l1 = tf.add(tf.matmul(data, hidden_1_layer['weights']), 
hidden_1_layer['biases'])
l1 = tf.nn.relu(l1)

l2 = tf.add(tf.matmul(data, hidden_2_layer['weights']), 
hidden_2_layer['biases'])
l2 = tf.nn.relu(l2)

l3 = tf.add(tf.matmul(data, hidden_3_layer['weights']), 
hidden_3_layer['biases'])
l3 = tf.nn.relu(l3)

output = tf.matmul(l3, output_layer['weights']) + 
output_layer['biases']

更正的代码:

l1 = tf.add(tf.matmul(data, hidden_1_layer['weights']), 
hidden_1_layer['biases'])
l1 = tf.nn.relu(l1)

l2 = tf.add(tf.matmul(l1, hidden_2_layer['weights']), 
hidden_2_layer['biases'])
l2 = tf.nn.relu(l2)

l3 = tf.add(tf.matmul(l2, hidden_3_layer['weights']), 
hidden_3_layer['biases'])
l3 = tf.nn.relu(l3)

output = tf.matmul(l3, output_layer['weights']) + 
output_layer['biases']

【讨论】:

    猜你喜欢
    • 2020-07-05
    • 1970-01-01
    • 1970-01-01
    • 1970-01-01
    • 2023-03-29
    • 1970-01-01
    • 2014-04-09
    • 1970-01-01
    • 2022-11-29
    相关资源
    最近更新 更多