ADP 前沿学习

← 板块一 · 研究前沿

Sequential Neural Models with Stochastic Layers

Fraccaro, Sønderby, Paquet, Winther · stat.ML,cs.LG · 2016-11-13 · 原文

How can we efficiently propagate uncertainty in a latent state representation with recurrent neural networks? This paper introduces stochastic recurrent neural networks which glue a deterministic recurrent neural network and a state space model together to form a stochastic and sequential neural generative model. The clear separation of deterministic and stochastic layers allows a structured variational inference network to track the factorization of the model's posterior distribution. By retaining both the nonlinear recursive structure of a recurrent neural network and averaging over the uncertainty in a latent path, like a state space model, we improve the state of the art results on the Blizzard and TIMIT speech modeling data sets by a large margin, while achieving comparable performances to competing methods on polyphonic music modeling.

🔮 让 ChatGPT 全网深度追问

讲义

讲义·推断 依据「原文」自动生成的结构化摘要(推断),非原文表述;以原文为准。

1. 人话版

How can we efficiently propagate uncertainty in a latent state representation with recurrent neural networks?

This paper introduces stochastic recurrent neural networks which glue a deterministic recurrent neural network and a state space model together to form a stochastic and sequential neural generative model.

2. 领域脉络

本文类目:stat.ML、cs.LG,属于其所在研究脉络的最新进展。

3. 机制拆解

By retaining both the nonlinear recursive structure of a recurrent neural network and averaging over the uncertainty in a latent path, like a state space model, we improve the state of the art results on the Blizzard and TIMIT speech modeling data sets by a large margin, while achieving comparable performances to competing methods on polyphonic music modeling.

4. 证据与数字

摘要未给出量化结果——留意原文的实验与数据。

5. 反例与边界

The clear separation of deterministic and stochastic layers allows a structured variational inference network to track the factorization of the model's posterior distribution.

6. 跨领域连接与意外收获

横跨 2 个类目(stat.ML、cs.LG),关注其在你兴趣板块间的迁移面。

7. 可复用方法

把本文机制与你手头项目对照,找一个两周内能验证的最小实验。

8. 术语表

精读时把不熟的术语记入此处,作为下次回忆的锚点。