REVIEW 10 cited by
Adaptive Machine Translation with Large Language Models
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Consistency is a key requirement of high-quality translation. It is especially important to adhere to pre-approved terminology and adapt to corrected translations in domain-specific projects. Machine translation (MT) has achieved significant progress in the area of domain adaptation. However, real-time adaptation remains challenging. Large-scale language models (LLMs) have recently shown interesting capabilities of in-context learning, where they learn to replicate certain input-output text generation patterns, without further fine-tuning. By feeding an LLM at inference time with a prompt that consists of a list of translation pairs, it can then simulate the domain and style characteristics. This work aims to investigate how we can utilize in-context learning to improve real-time adaptive MT. Our extensive experiments show promising results at translation time. For example, LLMs can adapt to a set of in-domain sentence pairs and/or terminology while translating a new sentence. We observe that the translation quality with few-shot in-context learning can surpass that of strong encoder-decoder MT systems, especially for high-resource languages. Moreover, we investigate whether we can combine MT from strong encoder-decoder models with fuzzy matches, which can further improve translation quality, especially for less supported languages. We conduct our experiments across five diverse language pairs, namely English-to-Arabic (EN-AR), English-to-Chinese (EN-ZH), English-to-French (EN-FR), English-to-Kinyarwanda (EN-RW), and English-to-Spanish (EN-ES).
Forward citations
Cited by 10 Pith papers
-
Preperiodic points, finiteness, and structures of semigroups of algebraic morphisms
The paper proves finiteness and structural results for preperiodic points of algebraic morphisms, including Burnside-type and Northcott-type theorems.
-
Combining the Best of Both Worlds: A Method for Hybrid NMT and LLM Translation
A learned source-feature decider routes each sentence to either an NMT model or an LLM, improving average translation quality over both single systems and a QE-based baseline while using the LLM for only about 20-30% ...
-
Knowledge is Power: Harnessing Large Language Models for Enhanced Cognitive Diagnosis
A two-stage framework uses LLM-generated text diagnoses plus contrastive and mask-reconstruction alignment to improve cognitive diagnosis models, with reported gains on four education datasets.
-
MageBench: Bridging Large Multimodal Models to Agents
MageBench introduces a 483-scenario benchmark showing current large multimodal models are far weaker than humans at agent tasks requiring continuous visual feedback and planning.
-
Quantum Graph Transformer for NLP Sentiment Classification
A hybrid quantum-classical graph transformer for sentiment classification reports higher accuracy and better sample efficiency than a classical graph transformer on five small benchmark datasets.
-
Multilingual Prompt Engineering in Large Language Models: A Survey Across NLP Tasks
A survey that categorizes multilingual prompting techniques by NLP task and language family, and designates potential state-of-the-art prompting methods for each dataset.
-
Large Language Models for Persian $ \leftrightarrow $ English Idiom Translation
A new resource of 2,200 Persian idioms and two 200-sentence benchmarks show Claude-3.5-Sonnet leads idiom translation accuracy for Persian-English, with hybrid LLM-plus-NMT setups aiding weaker models in English-to-Persian.
-
Lexicography Saves Lives (LSL): Automatically Translating Suicide-Related Language
The paper reports a 200-language machine-translated suicide lexicon and a pilot five-language human evaluation, but releases no data and contains inconsistent evaluation scores.
-
Video Diffusion Transformers are In-Context Learners
Concatenating multiple video clips into one input and fine-tuning a LoRA adapter lets a pretrained video diffusion transformer produce consistent multi-scene videos from a single prompt.
-
Demystifying ChatGPT: How It Masters Genre Recognition
A benchmark reports that ChatGPT outperforms other LLMs and traditional classifiers on multi-label movie genre prediction, but the result is undermined by likely pretraining contamination and weak baselines.
Discussion (0). Continue with ORCID to comment.