Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Refined Gate: A Simple and Effective Gating Mechanism for Recurrent Units

Feb 26, 2020

Zhanzhan Cheng, Yunlu Xu, Mingjian Cheng, Yu Qiao, Shiliang Pu, Yi Niu, Fei Wu

Figure 1 for Refined Gate: A Simple and Effective Gating Mechanism for Recurrent Units

Figure 2 for Refined Gate: A Simple and Effective Gating Mechanism for Recurrent Units

Figure 3 for Refined Gate: A Simple and Effective Gating Mechanism for Recurrent Units

Figure 4 for Refined Gate: A Simple and Effective Gating Mechanism for Recurrent Units

Share this with someone who'll enjoy it:

Abstract:Recurrent neural network (RNN) has been widely studied in sequence learning tasks, while the mainstream models (e.g., LSTM and GRU) rely on the gating mechanism (in control of how information flows between hidden states). However, the vanilla gates in RNN (e.g. the input gate in LSTM) suffer from the problem of gate undertraining mainly due to the saturating activation functions, which may result in failures of learning gating roles and thus the weak performance. In this paper, we propose a new gating mechanism within general gated recurrent neural networks to handle this issue. Specifically, the proposed gates directly short connect the extracted input features to the outputs of vanilla gates, denoted as refined gates. The refining mechanism allows enhancing gradient back-propagation as well as extending the gating activation scope, which, although simple, can guide RNN to reach possibly deeper minima. We verify the proposed gating mechanism on three popular types of gated RNNs including LSTM, GRU and MGU. Extensive experiments on 3 synthetic tasks, 3 language modeling tasks and 5 scene text recognition benchmarks demonstrate the effectiveness of our method.

View paper on

Share this with someone who'll enjoy it:

Title:Refined Gate: A Simple and Effective Gating Mechanism for Recurrent Units

Paper and Code