Pith. sign in

Paper Citation Record · LEDGER

MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2504.10160.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.10160 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:14:01.716379Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T02:56:29.964672Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a11ac186-d2f9-44e8-a5ee-99cf35f7bd09 · inbound

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs cites this paper.

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-22T01:20:52.107060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-22T01:16:51.288077Z digest=sha256:b5c938982d3167d174b236ec70aa0cc691fa7986550990c7b2afec599fcff075

Observation 25423da9-622c-4e82-b194-c93f4694c9a8 · inbound

MT$^{3}$: Scaling MLLM-based Text Image Machine Translation via Multi-Task Reinforcement Learning cites this paper.

MT$^{3}$: Scaling MLLM-based Text Image Machine Translation via Multi-Task Reinforcement Learning MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:14:01.716379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:14:01.716379Z digest=sha256:9a5c89aaa5dc1d7e17d961e42b23442ad28490fb2c248e8b99183d4156c2883d

Observation 5a8bfdb9-eae5-411a-88b5-c68f5282623a · inbound

TAT-R1: Terminology-Aware Translation with Reinforcement Learning and Word Alignment cites this paper.

TAT-R1: Terminology-Aware Translation with Reinforcement Learning and Word Alignment MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:43:09.416914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:43:09.416914Z digest=sha256:e0233469d9541bcb5b30c864db94f7e5cca96633ead947a0360df2745946281f

Observation 44331071-acc5-481f-9936-d2f72986b32e · inbound

RIVAL: Reinforcement Learning with Iterative and Adversarial Optimization for Machine Translation cites this paper.

RIVAL: Reinforcement Learning with Iterative and Adversarial Optimization for Machine Translation MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T10:31:58.538085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:31:58.538085Z digest=sha256:7603220ae04a7c832145558a49d975efbda48374e68347fdb39af70470ceb0fe

Observation be7e92a0-697a-410c-9ae5-6d8d640e559c · inbound

TACTIC: Translation Agents with Cognitive-Theoretic Interactive Collaboration cites this paper.

TACTIC: Translation Agents with Cognitive-Theoretic Interactive Collaboration MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:18:23.641438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:18:23.641438Z digest=sha256:6ce242755ddd71fbd400766b5ce8922d43b6ccd8a53e8ed82ac912d3dfe0afb8

Observation cdd2b602-83a4-4d2f-b263-da4843ac94ad · inbound

Seed LiveInterpret 2.0: End-to-end Simultaneous Speech-to-speech Translation with Your Voice cites this paper.

Seed LiveInterpret 2.0: End-to-end Simultaneous Speech-to-speech Translation with Your Voice MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T14:52:42.976548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:52:42.976548Z digest=sha256:a0b60dfd68cae2bea1e0b3442d31a9852a3ace56b7f0a61dfd844e97cd2a5649

Observation a8a4ed51-9823-426e-998a-af213afef551 · inbound

Datasets and Recipes for Video Temporal Grounding via Reinforcement Learning cites this paper.

Datasets and Recipes for Video Temporal Grounding via Reinforcement Learning MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T14:44:26.794508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:44:26.794508Z digest=sha256:d2d0517ad435b9d14909cc7b89a1e312e0a1d41fd3e23e5285302b9d147855d1

Observation 506720b2-0c4a-4ad0-8c3c-17e67987324b · inbound

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning cites this paper.

AutoTIR: Autonomous Tools Integrated Reasoning via Reinforcement Learning MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T12:23:55.081531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:23:55.081531Z digest=sha256:44586b4df1b257cad61b1fab53b0701e03ad3ee673075f6cc1adcd2e05db4b0f

Observation 11ae3987-9429-45c0-8c1e-f426e34f4347 · inbound

ReflectMT: Internalizing Reflection for Efficient and High-Quality Machine Translation cites this paper.

ReflectMT: Internalizing Reflection for Efficient and High-Quality Machine Translation MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning

Reference 50

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T02:48:27.350622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-05-10T02:45:59.473429Z digest=sha256:3924216aa76dd2119d093260f3f57d3e0bfd3d253f487a1c0292f9b5ae5b596f

Observation 6fc77cda-fb07-475d-936a-f88158169fda · inbound

Small RL Controller, Large Language Model: RL-Guided Adaptive Sampling for Test-Time Scaling cites this paper.

Small RL Controller, Large Language Model: RL-Guided Adaptive Sampling for Test-Time Scaling MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:56:29.966343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-28T10:25:10.559953Z digest=sha256:d06d81ccdf0df79594386e8a456382d6098b43bde184d7e8d635ef04c91440bf

Observation 5e04a4c3-3cf7-4b43-aa8c-bdfec9010982 · inbound

Reasoning Before Translation: Enhancing Legal Machine Translation with Structured Reasoning cites this paper.

Reasoning Before Translation: Enhancing Legal Machine Translation with Structured Reasoning MT-R1-Zero: Advancing LLM-based Machine Translation via R1-Zero-like Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T13:15:35.246083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:15:35.246083Z digest=sha256:4468eee7859a3bdfd33cd43d0f3dcce32d8f4bc4932a3d3d4172d15ec574269d