由李飞飞团队开发的

依然是基于encoder-decoder模型进行改编

 

图像生成文本(四) —— Show and Tell模型

使用了googlenet

相对于Multi-modal模型,其图像特征只使用了一次

图像生成文本(四) —— Show and Tell模型

 

 

与Encoder-Decoder的区别

由GooLeNet替换了Encoder,由GooLeNet得到hn

图像生成文本(四) —— Show and Tell模型

 

 

 

 

相关文章:

  • 2021-07-01
  • 2021-04-12
  • 2022-12-23
  • 2021-10-28
  • 2022-12-23
  • 2021-04-18
  • 2021-04-16
猜你喜欢
  • 2021-07-24
  • 2021-09-23
  • 2021-05-01
  • 2021-06-15
  • 2022-01-12
  • 2022-12-23
  • 2021-06-29
相关资源
相似解决方案