Fetching the paper…

Efficient Sequence Training of Attention Models using Approximative Recombination · Around