Abstract:Consistency is a key requirement of high-quality translation. It is especially important to adhere to pre-approved terminology and adapt to corrected translations in domain-specific projects. Machine translation (MT) has achieved significant progress in the area of domain adaptation. However, real-time adaptation remains challenging. Large-scale language models (LLMs) have recently shown interesting capabilities of in-context learning, where they learn to replicate certain input-output text generation patterns, without further fine-tuning. By feeding an LLM at inference time with a prompt that consists of a list of translation pairs, it can then simulate the domain and style characteristics. This work aims to investigate how we can utilize in-context learning to improve real-time adaptive MT. Our extensive experiments show promising results at translation time. For example, GPT-3.5 can adapt to a set of in-domain sentence pairs and/or terminology while translating a new sentence. We observe that the translation quality with few-shot in-context learning can surpass that of strong encoder-decoder MT systems, especially for high-resource languages. Moreover, we investigate whether we can combine MT from strong encoder-decoder models with fuzzy matches, which can further improve translation quality, especially for less supported languages. We conduct our experiments across five diverse language pairs, namely English-to-Arabic (EN-AR), English-to-Chinese (EN-ZH), English-to-French (EN-FR), English-to-Kinyarwanda (EN-RW), and English-to-Spanish (EN-ES).

Learning Domain Specific Language Models for Automatic Speech Recognition through Machine Translation

Language Model Bootstrapping Using Neural Machine Translation For Conversational Speech Recognition

Towards ASR Robust Spoken Language Understanding Through In-Context Learning With Word Confusion Networks

Language Modeling, Lexical Translation, Reordering: The Training Process of NMT through the Lens of Classical SMT

Language Model-Driven Unsupervised Neural Machine Translation

Adaptation of Language Models for SMT Using Neural Networks with Topic Information.

Language-agnostic Multilingual Modeling

Exploiting Language Relatedness in Machine Translation Through Domain Adaptation Techniques

Salute the Classic: Revisiting Challenges of Machine Translation in the Age of Large Language Models

An Investigation On Statistical Machine Translation With Neural Language Models

Adaptive Machine Translation with Large Language Models

Language Modelling Approaches to Adaptive Machine Translation

Configurable Multilingual ASR with Speech Summary Representations

Using Language Models to Disambiguate Lexical Choices in Translation

Unified Model Learning for Various Neural Machine Translation

Learning not to Discriminate: Task Agnostic Learning for Improving Monolingual and Code-switched Speech Recognition

Transfer learning of language-independent end-to-end ASR with language model fusion

Neural Machine Translation model for University Email Application

A Deep Learning System for Domain-specific Speech Recognition

Leveraging native language information for improved accented speech recognition

FC-MTLF: A Fine- and Coarse-grained Multi-Task Learning Framework for Cross-Lingual Spoken Language Understanding.