Pith. sign in

Paper Citation Record · LEDGER

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning

As of 21 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 1 inbound Pith citation observation for arXiv:2507.07306.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.07306 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:48:51.453332Z

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-08T19:34:02.860270Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-09T05:50:25.837175Z

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a9bd8f55-93ce-4bbc-87d3-99ed0d237ec9 · outbound

This paper cites 0:00:01,229.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning 0:00:01,229

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:52.564849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T18:48:51.266705Z digest=sha256:095d2d769da220c41b81eb1f2c63f458c47c54b85a7be86ee49c0cd50413f6b8

Observation fbc4129c-cc86-4fa1-8c96-27f2c8d7ac72 · outbound

This paper cites - Translate into natural, fluent Simplified Chinese.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning - Translate into natural, fluent Simplified Chinese

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:52.301134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T18:48:51.354932Z digest=sha256:a8e485a1545456b5209582024dda8a396948ef7283f337978fd2fa6dadd9e164

Observation 49dbd9a3-455b-4399-bfc5-c3124132ee7f · outbound

This paper cites Qwen2-Audio Technical Report.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning Qwen2-Audio Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:50.400665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:50.400665Z digest=sha256:f858e5e95db36016d053055d2ccee55830802d099762710899c77c2c1921aee2

Observation eade5100-c891-44b2-bbc9-5497669b11b7 · outbound

This paper cites Generative Multi-Modal Knowledge Retrieval with Large Language Models.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning Generative Multi-Modal Knowledge Retrieval with Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:50.667682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:50.667682Z digest=sha256:fb415b0721868c1ce843a5f664f52f4361c21416df646badbaa59d1b88d00623

Observation 40b6a302-24ba-4964-af30-91310563ccd1 · outbound

This paper cites Low-Resource Machine Translation through Retrieval-Augmented LLM Prompting: A Study on the Mambai Language.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning Low-Resource Machine Translation through Retrieval-Augmented LLM Prompting: A Study on the Mambai Language

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:50.814311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:50.814311Z digest=sha256:049fa65921c02987674673406f18e4a906e950b88ae1e7d5ad45376e82c27e23

Observation 507589dc-73e6-4233-8273-186ec0158572 · outbound

This paper cites A Survey on Multi-modal Machine Translation: Tasks, Methods and Challenges.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning A Survey on Multi-modal Machine Translation: Tasks, Methods and Challenges

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:51.081327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:51.081327Z digest=sha256:6e6f4f5644c653ddcb2b487abf015ba0ee0f99e3bcd2a95d1479c762e3ed0be0

Observation 756f39ab-2379-4120-8206-c9aba4205673 · outbound

This paper cites 翻译提供的视频中的说话内容到中文。只需要输出翻译内容原文,不要输出任何解释。.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning 翻译提供的视频中的说话内容到中文。只需要输出翻译内容原文,不要输出任何解释。

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:52.090711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T18:48:51.453332Z digest=sha256:f51f3ec2f46ad2ddf65fbd5abcda7091d84f24d40c10d03a75dd87c67ca8d4a3

Observation 6c36e146-7be7-4520-bc03-e7fef0e2f1da · outbound

This paper cites Google's Neural Machine Translation System: Bridging the Gap between Human and Machine Translation.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning Google's Neural Machine Translation System: Bridging the Gap between Human and Machine Translation

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:51.181622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:51.181622Z digest=sha256:d797e0cbb2063275656d9ffc796a72845329f8e9c68678a7ebf6b399dd3dace1

Observation 60b8b508-fee2-4c63-a1d5-756ac76760a8 · outbound

This paper cites How to Design Translation Prompts for ChatGPT: An Empirical Study.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning How to Design Translation Prompts for ChatGPT: An Empirical Study

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:50.559294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:50.559294Z digest=sha256:7bd2faca3e9b6761a9ff6445f66dc12245dcb7469556b9185f46f968cb558d3f

Observation a052cd3b-a9f7-4471-9a7d-561997f55ad8 · outbound

This paper cites Thibault Sellam, Dipanjan Das, and Ankur P Parikh.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning Thibault Sellam, Dipanjan Das, and Ankur P Parikh

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:50.931379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:50.931379Z digest=sha256:160f23ab742b213eda34bd8e9c121dda87ac97fa43b53b2d0d03e64fa0bd84bf

Observation 29858307-8c49-4cdb-a116-dd6566757533 · outbound

This paper cites Retrieving Examples from Memory for Retrieval Augmented Neural Machine Translation: A Systematic Comparison.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning Retrieving Examples from Memory for Retrieval Augmented Neural Machine Translation: A Systematic Comparison

Reference 2024

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T18:48:51.849051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T18:48:50.286314Z digest=sha256:d2861e3818da022360490eb18c3f971e34dc36c192804008c5db07c0914c2443

Observation 5bd26649-41f8-44b1-a26e-b69870a8c3f1 · outbound

This paper cites Qwen2.5-VL Technical Report.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning Qwen2.5-VL Technical Report

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:50.189858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:50.189858Z digest=sha256:32c8261b4afc62754f159d33e36c91f9377b8f87ba90517b3b6faaa6e99bae37

Pith citing papers

Observation 59257bbb-5a73-44df-9dbc-4678817bf33b · inbound

VIDA: A dataset for Visually Dependent Ambiguity in Multimodal Machine Translation cites this paper.

VIDA: A dataset for Visually Dependent Ambiguity in Multimodal Machine Translation ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-09T05:50:25.838847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-08T19:34:02.860270Z digest=sha256:3f8db7cadf2a184c78445a9cfd7ee5d57c20ccd010baa15bb8dbc61230b77f99