Adaptive Computation Time for Recurrent Neural Networks (ACT) is one of the most promising architectures for variable computation. ACT adapts to the input sequence by being able to look at each sample more than once, and learn how many times it should do it. In this paper, we compare ACT to Repeat-RNN, a novel architecture based on repeating each sample a fixed number of times. We found surprising results, where Repeat-RNN performs as good as ACT in the selected tasks. Source code in TensorFlow and PyTorch is publicly available at this https URL
Paper
References (7)
Similar papers
Adaptive Computation Time for Recurrent Neural NetworksAlex Graves2016 · arXiv.org · 785 citations In Library
Differentiable Adaptive Computation Time for Visual ReasoningCristóbal Eyzaguirre, Álvaro Soto2020 In Library
Layer Flexible Adaptive Computation TimeLida Zhang, Abdolghani Ebrahimi, Diego Klabjan2021 In Library