Pith. sign in

Paper Citation Record · LEDGER

Length Aware Speech Translation for Video Dubbing

As of 8 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 1 inbound Pith citation observation for arXiv:2506.00740.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.00740 v1

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:02:22.477354Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:02:20.273561Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T12:02:22.672497Z

Reference resolution

24 of 24 outbound references displayed

  • verified exact1
  • verified fuzzy21
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 50be5103-5baf-4227-a80e-2b28d4433ed9 · outbound

This paper cites Length Aware Speech Translation for Video Dubbing.

Length Aware Speech Translation for Video Dubbing Length Aware Speech Translation for Video Dubbing

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:02:22.724435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:20.273561Z digest=sha256:3693486909e694ae83eec3c3eb8291af8706540df42905cb816bc8fcc7833c16

Observation d6cd933b-2236-41c7-b09a-e344d07b06c1 · outbound

This paper cites The duration of translated audio is influ- enced by: (a) the length of the translated text, and (b) the dura- tion model within the text-to-speech (TTS) system.

Length Aware Speech Translation for Video Dubbing The duration of translated audio is influ- enced by: (a) the length of the translated text, and (b) the dura- tion model within the text-to-speech (TTS) system

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:02:26.774957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:20.320178Z digest=sha256:a71aa815cacb30f3cf80a7c064d5e9445b0479359db34859d180bc3303488a2f

Observation 55c0d34f-7e87-4ae5-8f9e-480ebab415d9 · outbound

This paper cites Model and Data The ST model used in our experiments is multilingual and jointly trained on Spanish (ES) and Korean (KO) data.

Length Aware Speech Translation for Video Dubbing Model and Data The ST model used in our experiments is multilingual and jointly trained on Spanish (ES) and Korean (KO) data

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:02:26.482554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:20.493478Z digest=sha256:bd7d394dbf3066d508e5b2bbded939fe851b9b06a2b54ef52cebaf18e5739a9c

Observation 9336c538-4a13-4902-a877-f02753dfcf8f · outbound

This paper cites an unresolved cited work.

Length Aware Speech Translation for Video Dubbing Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:02:26.610354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:20.403503Z digest=sha256:16d0592a7d23183673a2a6b588e694dd898fc1e75c7b7a4eb753caa0193cd475

Observation e5019cc5-db24-48f8-85f7-84aae49d17e8 · outbound

This paper cites an unresolved cited work.

Length Aware Speech Translation for Video Dubbing Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:02:26.355715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:20.581943Z digest=sha256:8c6ab978b021f58223e584d03d154efc9d7ab7cbb77143302ac2439546b52a70

Observation 2af8228e-6317-4b13-8f9d-f1e6e8ee17af · outbound

This paper cites Our approach leverages predefined length control tokens to generate transla- tions of varying lengths—short, normal, and long—while main- taining high translation quality.

Length Aware Speech Translation for Video Dubbing Our approach leverages predefined length control tokens to generate transla- tions of varying lengths—short, normal, and long—while main- taining high translation quality

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:02:26.158753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:20.678246Z digest=sha256:3e71d15c30d441e7fb398629465b0a8bbfd48962a2672963c855e9077ab2654b

Observation d44a2f01-7e25-4985-b926-9ae24a81c927 · outbound

This paper cites Leveraging weakly supervised data to improve end-to-end speech-to-text translation,.

Length Aware Speech Translation for Video Dubbing Leveraging weakly supervised data to improve end-to-end speech-to-text translation,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:02:25.962999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:20.739344Z digest=sha256:448a841419abba4162680f21c25c59e42f6a6284074076ae9f6adeadf5b9cb62

Observation e4c0e5fa-9e0c-4500-8491-37fb26954bee · outbound

This paper cites Large-scale stream- ing end-to-end speech translation with neural transducers,.

Length Aware Speech Translation for Video Dubbing Large-scale stream- ing end-to-end speech translation with neural transducers,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:02:25.805883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:20.844471Z digest=sha256:4651bc89c575131487d26e7ee135b74ae1a4fdd2c478a58e477b26b9a440986e

Observation 5f522bd4-46cd-490f-911b-a374363298d9 · outbound

This paper cites Revisiting end-to-end speech-to-text translation from scratch,.

Length Aware Speech Translation for Video Dubbing Revisiting end-to-end speech-to-text translation from scratch,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:02:25.677659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:20.972785Z digest=sha256:cb4c87730a275877c08b8abcc86612a69617eda8ad1069fdccc8f191887d1a53

Observation 80d6cb22-325f-485a-a43f-4405985b0eea · outbound

This paper cites Videodubber: machine translation with speech-aware length control for video dubbing,.

Length Aware Speech Translation for Video Dubbing Videodubber: machine translation with speech-aware length control for video dubbing,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:02:25.469350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:21.090318Z digest=sha256:cd23bae719c2c3b30182a468a8739e73a7175be5590693c1ed688a0b7f0a059b

Observation 6f8690b8-3172-4649-860e-18a7fc8f9f60 · outbound

This paper cites Controlling machine translation for multiple attributes with additive interven- tions,.

Length Aware Speech Translation for Video Dubbing Controlling machine translation for multiple attributes with additive interven- tions,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:02:25.307976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:21.144276Z digest=sha256:8d6339399e9e71dfad9aea4079ce73f608711f2c05f6152d390f28ce4dc79434

Observation c891b55d-409f-43c4-9d31-54d915e32e42 · outbound

This paper cites Is 42 the answer to ev- erything in subtitling-oriented speech translation?.

Length Aware Speech Translation for Video Dubbing Is 42 the answer to ev- erything in subtitling-oriented speech translation?

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:02:25.122444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:21.231253Z digest=sha256:a2f03fcf1531fe98165ca993acd3b328c8694284267d6a978671820a8c64eaad

Observation 77e30b5e-5581-4f16-8001-2021b968ed32 · outbound

This paper cites HW-TSC’s participa- tion in the IWSLT 2022 isometric spoken language translation,.

Length Aware Speech Translation for Video Dubbing HW-TSC’s participa- tion in the IWSLT 2022 isometric spoken language translation,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:02:24.977450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:21.321075Z digest=sha256:ee947856591c17410663dde9034a1e2d822126f81673932c1c9a6e6b4d59ad77

Observation 786a7a35-b98e-4a3d-b333-e4cdf9587e0e · outbound

This paper cites Adapting end-to-end speech recognition for readable subtitles,.

Length Aware Speech Translation for Video Dubbing Adapting end-to-end speech recognition for readable subtitles,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:02:24.782957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:21.444140Z digest=sha256:f29fb5e0107895ad11bb8ee7702335e241b120c57dbeffd325af219df43d95d1

Observation 94eab385-7f58-4ef7-a5a4-0a720508dfa6 · outbound

This paper cites Length- aware NMT and adaptive duration for automatic dubbing,.

Length Aware Speech Translation for Video Dubbing Length- aware NMT and adaptive duration for automatic dubbing,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:02:24.579489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:21.522466Z digest=sha256:aa91d2450e3ce7ec6f3d7efd39a144142b7706bf8685c7bfb35fb1493c0c61ad

Observation 63532acc-da9b-4a64-8027-a88178be48d5 · outbound

This paper cites Machine translation verbosity control for automatic dubbing,.

Length Aware Speech Translation for Video Dubbing Machine translation verbosity control for automatic dubbing,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:02:24.232143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:21.641051Z digest=sha256:17667dfeb10456915001b66dba295217337045269d77ce4afcf272b7bdf90545

Observation 04a7a3fd-aef3-43e6-bca2-b6d024091488 · outbound

This paper cites Isochrony-aware neural machine translation for automatic dub- bing,.

Length Aware Speech Translation for Video Dubbing Isochrony-aware neural machine translation for automatic dub- bing,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:02:24.056077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:21.720649Z digest=sha256:60a62a3aa508139a7710ae76fdabc42e0ab96e3d918d1ca11729c78fdebe09a1

Observation 1e40c525-7a80-446e-9302-88fa9f79df3c · outbound

This paper cites Isometric neural machine translation us- ing phoneme count ratio reward-based reinforcement learning,.

Length Aware Speech Translation for Video Dubbing Isometric neural machine translation us- ing phoneme count ratio reward-based reinforcement learning,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:02:23.923490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:21.815996Z digest=sha256:01d603d052e372b912f054ef09a6b6ca72a3e544ab7f5176ba2b406f506bc589

Observation c164e01c-07f4-4b11-ab56-17cadb7bef10 · outbound

This paper cites FastSpeech 2: Fast and high-quality end-to-end text to speech,.

Length Aware Speech Translation for Video Dubbing FastSpeech 2: Fast and high-quality end-to-end text to speech,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:02:23.764767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:21.932443Z digest=sha256:7f35518c113221b211959cb6bbc503a4b0dadf9b17adb39b8ba5f989eee6a38b

Observation 16f9d6ea-13d9-43c2-a79f-a6925629b3d5 · outbound

This paper cites Conformer: Convolution-augmented transformer for speech recognition,.

Length Aware Speech Translation for Video Dubbing Conformer: Convolution-augmented transformer for speech recognition,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:02:23.548438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:22.031806Z digest=sha256:3bde34f30184e8081b595407fc86a36f0df42d5db08a92c14ff0a0a201916140

Observation 1c399852-892d-4cbc-b3d9-4033bca30e9c · outbound

This paper cites Hybrid CTC/attention architecture for end-to-end speech recog- nition,.

Length Aware Speech Translation for Video Dubbing Hybrid CTC/attention architecture for end-to-end speech recog- nition,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:02:23.409250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:22.090541Z digest=sha256:a358ee17c3c487e5801d9fbf7517e9631419cabe421039ff18ce6c29b64ff9ba

Observation 0b7d8599-7d6a-4c05-a552-fd47b114cc47 · outbound

This paper cites Fleurs: Few-shot learning evaluation of universal representations of speech,.

Length Aware Speech Translation for Video Dubbing Fleurs: Few-shot learning evaluation of universal representations of speech,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:02:23.206909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:22.186666Z digest=sha256:37ee7cb1c73eedf7d1cc1e25c56656340081bee88ccfed14664b8d29b1407820

Observation b412060a-49bc-47aa-a46f-c407124203a7 · outbound

This paper cites LeanSpeech: The Microsoft lightweight speech synthesis system for limmits challenge 2023,.

Length Aware Speech Translation for Video Dubbing LeanSpeech: The Microsoft lightweight speech synthesis system for limmits challenge 2023,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:02:23.010430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:22.337542Z digest=sha256:e7c57a6999fee2f0eb8447c9efc249e5ddf86021f91d008a72d1a44988b85b7d

Observation 930721e1-96fd-4a3e-a177-4d0affc1bfeb · outbound

This paper cites A call for clarity in reporting BLEU scores,.

Length Aware Speech Translation for Video Dubbing A call for clarity in reporting BLEU scores,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:02:22.869228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:22.477354Z digest=sha256:d237020d8934784399f4f7075e9d9ceb17c85b7c0c288eb833356d8d234416ba

Pith citing papers

Observation 50be5103-5baf-4227-a80e-2b28d4433ed9 · inbound

Length Aware Speech Translation for Video Dubbing cites this paper.

Length Aware Speech Translation for Video Dubbing Length Aware Speech Translation for Video Dubbing

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:02:22.724435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:02:20.273561Z digest=sha256:3693486909e694ae83eec3c3eb8291af8706540df42905cb816bc8fcc7833c16