Pith. sign in

REVIEW 5 cited by

CometKiwi: IST-Unbabel 2022 Submission for the Quality Estimation Shared Task

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2209.06243 v1 pith:BQZESRYQ submitted 2022-09-13 cs.CL cs.LG

classification cs.CLcs.LG
keywords qualitytasksword-levelestimationlanguagepairsresultssentence
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We present the joint contribution of IST and Unbabel to the WMT 2022 Shared Task on Quality Estimation (QE). Our team participated on all three subtasks: (i) Sentence and Word-level Quality Prediction; (ii) Explainable QE; and (iii) Critical Error Detection. For all tasks we build on top of the COMET framework, connecting it with the predictor-estimator architecture of OpenKiwi, and equipping it with a word-level sequence tagger and an explanation extractor. Our results suggest that incorporating references during pretraining improves performance across several language pairs on downstream tasks, and that jointly training with sentence and word-level objectives yields a further boost. Furthermore, combining attention and gradient information proved to be the top strategy for extracting good explanations of sentence-level QE models. Overall, our submissions achieved the best results for all three tasks for almost all language pairs by a considerable margin.

Discussion (0). Sign in to comment.

Forward citations

Cited by 5 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. OpenAlex reports about 33 citations worldwide. Full citation record

  1. Towards Style Alignment in Cross-Cultural Translation

    cs.CL 2025-06 conditional novelty 6.0 of 10

    LLMs systematically reduce politeness, intimacy, and formality variation in translation, and a retrieval-augmented prompting method that supplies native style exemplars improves style alignment without hurting content...

  2. Hunyuan-MT Technical Report

    cs.CL 2025-09 conditional novelty 5.0 of 10

    Hunyuan-MT and Chimera, a 7B open-source translation model and its multi-candidate fusion variant, claim state-of-the-art multilingual translation including Mandarin to minority languages, with open weights.

  3. Text2Cypher Across Languages: Evaluating and Finetuning LLMs

    cs.CL 2025-06 conditional novelty 5.0 of 10

    A new multilingual Text2Cypher benchmark shows LLMs rank English highest, Spanish next, and Turkish lowest, and multilingual finetuning narrows the language gap more than English-only finetuning.

  4. RIVAL: Reinforcement Learning with Iterative and Adversarial Optimization for Machine Translation

    cs.CL 2025-06 conditional novelty 5.0 of 10

    RIVAL iteratively re-trains a reward model adversarially against the current translator and adds a BLEU-predicting head, improving in-domain WMT and subtitle translation over SFT baselines.

  5. BeaverTalk: Oregon State University's IWSLT 2025 Simultaneous Speech Translation System

    cs.CL 2025-05 conditional novelty 4.0 of 10

    BeaverTalk combines VAD segmentation, Whisper ASR, and a LoRA-fine-tuned Gemma 3 with a single-sentence memory bank to achieve BLEU 24.64 to 37.23 on ACL 60/60 across two language pairs and two latency regimes.

Pith tools