Accelerating NMT Batched Beam Decoding with LMBR Posteriors for Deployment

Gonzalo Iglesias,William Tambellini,Adrià De Gispert,Eva Hasler,Bill Byrne
DOI: https://doi.org/10.48550/arXiv.1804.11324
2018-05-01
Abstract:We describe a batched beam decoding algorithm for NMT with LMBR n-gram posteriors, showing that LMBR techniques still yield gains on top of the best recently reported results with Transformers. We also discuss acceleration strategies for deployment, and the effect of the beam size and batching on memory and speed.
Computation and Language
What problem does this paper attempt to address?