Pith. sign in

Paper Citation Record · LEDGER

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation

As of 8 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 1 inbound Pith citation observation for arXiv:2607.03657.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.03657 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-12T00:53:05.419742Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-12T00:53:05.419742Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

42 of 42 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved42
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 44f31bfa-a977-4f26-af64-c025b218ebd9 · outbound

This paper cites Sign Language Trans- lation (SLT) aims to bridge communication gaps by trans- lating sign-language videos into spoken sentences.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Sign Language Trans- lation (SLT) aims to bridge communication gaps by trans- lating sign-language videos into spoken sentences

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:f996e27c0bb715b28cab51e8f916131dd1bf143fbb49107f7969db53fc51e546

Observation 904aa2ba-8e55-437c-8f3a-8f9c91a4b9e6 · outbound

This paper cites ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:fd78f940247b343b80ad7a130eb5d8ae2fb183cee560ee9075727f9a69c72a26

Observation eeddb7e1-89c9-4e83-82c4-fe481fe0d951 · outbound

This paper cites SignLLM [5] maps sign videos to discrete tokens aligned with LLMs.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation SignLLM [5] maps sign videos to discrete tokens aligned with LLMs

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:596156f2f2eea14a795f78e7dc9aadd49e9b2e41f4c0a1d99852c07b3b681f05

Observation 9d36b98f-90a7-47a3-bc64-605afec9d8ea · outbound

This paper cites Re- cent work leverageslarge-scale pretraining and multimodal LLMs.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Re- cent work leverageslarge-scale pretraining and multimodal LLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:0325ee1f3521680bdaae23df3739c07f7c29c6451db74862e2de0c0557182bb5

Observation 2d32709d-b02d-45a6-b90b-42852d90b0fb · outbound

This paper cites Translate the given sentence into<language>.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Translate the given sentence into<language>

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:7af972c534cc16641e35b3196028f46ce3ac274a13969eb8954a4bea6c769e24

Observation df985ce5-4bd3-4222-9df0-a871dd45dd92 · outbound

This paper cites Experimental Setup Datasets.We evaluate on two benchmark SLT datasets: PHOENIX14T [25] and CSL-Daily [26].

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Experimental Setup Datasets.We evaluate on two benchmark SLT datasets: PHOENIX14T [25] and CSL-Daily [26]

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:bd1b09ff171a8b6ef6ef9957f548261c3a171f957ab3bcf5a9397fd72d4729ce

Observation c81409e8-09fd-4e08-a012-c74491eedae3 · outbound

This paper cites an unresolved cited work.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:715a4b2ebc6a4979ec1c7de98c6c3026a251e947703fe61ea67359d9622b21f9

Observation 642b8bcb-b2c0-4ebe-aef9-20b67b7de116 · outbound

This paper cites an unresolved cited work.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:8f48b24325556ec57a5581d66ecef4e0790676b8ad68c2f42980d0a2897982c6

Observation 45e7175b-e3e7-4d58-8488-ebfc66214358 · outbound

This paper cites Sign language transformers: Joint end-to-end sign language recognition and translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Sign language transformers: Joint end-to-end sign language recognition and translation,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:35207ffb6fc4016287e7ed2492fc4aeb45666419767f0a7b1a42a7c16d1d75c1

Observation 1266af36-d757-4300-b91f-8f7aa15ddb6a · outbound

This paper cites Factorized Learning Assisted with Large Language Model for Gloss-free Sign Language Translation.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Factorized Learning Assisted with Large Language Model for Gloss-free Sign Language Translation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:5f9f376b4f2da20345ec1b8a072a209af9d2bca5300b6efec420fe16124751ae

Observation 21f54591-24a9-4b3b-9817-13aa0b883718 · outbound

This paper cites Gloss-free sign language translation: Improving from visual-language pretraining,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Gloss-free sign language translation: Improving from visual-language pretraining,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:166ae8d3bf014ac580623926588cae24aab71b738ea6f4186d52b04eb1d1e1cd

Observation abf0959e-9119-4624-ad2f-a1b7f504f03c · outbound

This paper cites Leveraging the power of mllms for gloss-free sign language transla- tion,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Leveraging the power of mllms for gloss-free sign language transla- tion,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:5c2566bf21dfcd96255c222b28381be1190a285350ce6370c7d023ac4c926a5a

Observation 53095427-6741-48d3-8723-c0745c10db7b · outbound

This paper cites Llms are good sign language translators,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Llms are good sign language translators,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:aaba063f99b194d3d76552bfff74febd920335a74451c6bf00c1cd28baa0d518

Observation dbfc33fa-6e84-47ad-9f2a-ec8cdfbf987c · outbound

This paper cites Lost in translation, found in context: Sign language translation with contextual cues,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Lost in translation, found in context: Sign language translation with contextual cues,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:fd79a3f2ec6f8d5ba11a51029950980160c69b93b390b3e85b4391702404d2bd

Observation f5e1c48a-cb80-4ae4-bc19-befe5bae6ba5 · outbound

This paper cites Sign2GPT: Leveraging Large Language Models for Gloss-Free Sign Language Translation.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Sign2GPT: Leveraging Large Language Models for Gloss-Free Sign Language Translation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:e12432c1f2085e253f35ce31db5033a4f24edf97eba7c6e41ba41f4f7f48f25c

Observation e9d58d38-1a59-4ecb-b248-bf0117adca00 · outbound

This paper cites Tspnet: Hierarchical feature learning via temporal semantic pyramid for sign language translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Tspnet: Hierarchical feature learning via temporal semantic pyramid for sign language translation,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:4d1b0196cb29c6c18c56bd81f8d6c02f69b406fcdb4c85a387e0e8fd3f4a9ce7

Observation 4b58aec4-71d2-41e9-aac2-f72c8c68e281 · outbound

This paper cites Conditional sen- tence generation and cross-modal reranking for sign lan- guage translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Conditional sen- tence generation and cross-modal reranking for sign lan- guage translation,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:2cd62af604dff21d24f9f71d998b1d0985273a92f2343fb5a7c8762045106bb2

Observation 0b2a8b37-1673-4aeb-b5e4-3ae2e8ec5e11 · outbound

This paper cites A token-level contrastive framework for sign language translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation A token-level contrastive framework for sign language translation,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:bd6984fa46eab1f92906c7eb83e70328c7b19446811a3e5a7877243444ca4bfe

Observation 0dc7d26f-fe5a-474f-a41b-5cd7620aeb47 · outbound

This paper cites Gloss attention for gloss-free sign language translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Gloss attention for gloss-free sign language translation,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:26b2c038df526d9499b8dd9a279e391be7799b4a1c00214316936662eba13b03

Observation ca9a3f79-3091-4547-9b77-5bad231a13da · outbound

This paper cites Visual alignment pre- training for sign language translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Visual alignment pre- training for sign language translation,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:bf3e3bc256d4603c8f5d87e461852bca9a7c923c6e4d344d2a1ffbcc449235e9

Observation 4b541d6c-79b4-4367-a7d8-495388c5ac6d · outbound

This paper cites An Efficient Sign Language Translation Using Spatial Configuration and Motion Dynamics with LLMs.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation An Efficient Sign Language Translation Using Spatial Configuration and Motion Dynamics with LLMs

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:02261cd06db084f6aef38a3b68e177136ba88ac3e9f8000f45a24e0e3f603c62

Observation 58457a83-56e3-413b-9060-9485926f7cc6 · outbound

This paper cites Better sign language translation with stmc-transformer,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Better sign language translation with stmc-transformer,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:08677c0bc3e3af2348805ac27db55ec7d6d3fa7493cccf003e5fee4379105670

Observation e2d25ca2-ac59-4de3-9a14-e6329f930d5d · outbound

This paper cites Lost in translation, found in embeddings: Sign language translation and alignment,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Lost in translation, found in embeddings: Sign language translation and alignment,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:2ebaa5b3fa48d900f73211bb3c0a1e72811426341f0932c76e93747609468b04

Observation 56c40903-7d60-4aff-8480-32c7eba70a65 · outbound

This paper cites Multimodal sign language recognition via temporal deformable convolu- tional sequence learning.,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Multimodal sign language recognition via temporal deformable convolu- tional sequence learning.,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:76ae332dbbb7e00bc6d39c81c1b79cc3f987d87f9006d1e8a9e4d978c85b7408

Observation 2468cb21-2394-467c-9cb3-81cdd806bb20 · outbound

This paper cites Multi-channel transformers for multi-articulatory sign language translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Multi-channel transformers for multi-articulatory sign language translation,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:b550517b9cd9b422fadd27f10f567200730870cd4b5ab59aedc09e4de45a7551

Observation 5d8d1f0f-2055-45ea-8387-e00c117fe462 · outbound

This paper cites Spatial- temporal multi-cue network for sign language recogni- tion and translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Spatial- temporal multi-cue network for sign language recogni- tion and translation,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:fa7abbe2643b25d4a34d5a57b855acf24e201a90af4b124b336dc84b0f7460f7

Observation 61c3a763-a5c2-4172-86aa-1f2ef5fe52c9 · outbound

This paper cites Graph- based multimodal sequential embedding for sign lan- guage translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Graph- based multimodal sequential embedding for sign lan- guage translation,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:8490f9e5143ba05862d03530a3302fd2cca728807e6faa67f522dd421dc02557

Observation 565ab00d-b204-4af9-9c4b-e5fdf8c5744c · outbound

This paper cites Skeleton-aware neural sign language translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Skeleton-aware neural sign language translation,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:c3c1fdc68cb8d1d7f9ce34ce266b9068beb55a204722e5c72da6e649bcf085c3

Observation 5c2752ac-d1c2-4699-a948-3259f7b9a05f · outbound

This paper cites Two-stream net- work for sign language recognition and translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Two-stream net- work for sign language recognition and translation,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:c403ac079373fa8304761445891c8f5a1f1ca47cf552dd9d9f71413833041d7c

Observation 0bf15414-6b57-4033-9535-3d91fd91406d · outbound

This paper cites Ustm: Unified spa- tial and temporal modeling for continuous sign language recognition,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Ustm: Unified spa- tial and temporal modeling for continuous sign language recognition,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:ea2a6cb5b20651bd5b11ac325f0fb051a62952723b548da4249323f5881387e9

Observation d51bd3aa-e5ff-4acf-8d1d-e14eebc5a5e9 · outbound

This paper cites Openpose: Re- altime multi-person 2d pose estimation using part affin- ity fields,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Openpose: Re- altime multi-person 2d pose estimation using part affin- ity fields,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:1c3df23180ef1490de8421ec0dbd96d4f454f497da3e6f74bd37c8a2ddc26843

Observation 005b2082-0c03-48e2-8127-b7701bc67a87 · outbound

This paper cites Temporal convolutional networks for action segmentation and de- tection,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Temporal convolutional networks for action segmentation and de- tection,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:01a37fb055b4e7c65c2e071ef774f9c47a00e58b900df64664f5a883f74bb070

Observation 4cc18bda-1be3-4653-8566-decea2ce0c1b · outbound

This paper cites Neural sign language translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Neural sign language translation,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:746cf90dbdd05ab65cd20eb1091caa840cd578aee8bd7a561b91482d85f89a95

Observation d932677a-5331-4ba5-a516-5f5aced9e2ed · outbound

This paper cites Improving sign language translation with monolingual data by sign back-translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Improving sign language translation with monolingual data by sign back-translation,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:c39188d613519004d467853984eec0542218fd1a0451208268ea88a00ab22ce3

Observation ab9a8202-c176-4fa8-9314-30088f71a9c9 · outbound

This paper cites Stochastic transformer networks with linear compet- ing units: Application to end-to-end sl translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Stochastic transformer networks with linear compet- ing units: Application to end-to-end sl translation,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:d538c24cf89e987d2295758cf849b35bc95d6e2ab95ae7a62a7f14977db2f434

Observation 4e631fb4-d64c-49c8-b802-af43b1aed53f · outbound

This paper cites Mska: Multi- stream keypoint attention network for sign language recognition and translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Mska: Multi- stream keypoint attention network for sign language recognition and translation,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:a9e2eb15a8fbba5948cfc4e49b6b6ed1120a2c409840854d1b49ffa811fb7e82

Observation 789c2d50-9bc0-45b0-925c-bdb1339e3292 · outbound

This paper cites Cross-modality Data Augmentation for End-to-End Sign Language Translation.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Cross-modality Data Augmentation for End-to-End Sign Language Translation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:a8a8af095b45c6f49255d26fe3adcdb20da0f20bd55c3634d741cc2071263aa9

Observation 490c6297-7a99-4be4-a976-c157a1ff7f45 · outbound

This paper cites Crosslingual generalization through multitask finetun- ing,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Crosslingual generalization through multitask finetun- ing,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:f1bc2e316a3b2fbfe6fb8b7f7656fcbf204a09c1492937095e4e896afc3c6c17

Observation c4114bb0-25e7-41a2-9793-96604dbc18e1 · outbound

This paper cites Lora: Low-rank adaptation of large language models,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Lora: Low-rank adaptation of large language models,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:c08f0950f19b01b0a95560f9b057a6b5f77e01e7d4a78f338905c9b5562526ee

Observation 9ae2e96d-8417-4cd5-b12c-19048fb1af02 · outbound

This paper cites Scaling instruction-finetuned language models,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Scaling instruction-finetuned language models,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:7db8581d069aa87fcc048d5b22db8e5f2db87727bafb5871e806f2dedff9be49

Observation 82606eee-2de5-4d61-bb27-1329984ec715 · outbound

This paper cites Multilingual denois- ing pre-training for neural machine translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Multilingual denois- ing pre-training for neural machine translation,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:ed8665692dcc51c406e713d4331434452ac226c6e64d21ff989b18d7832481e2

Observation 9795a18a-e8a6-4886-b498-f3686aa20b6c · outbound

This paper cites Ararea- soner: Evaluating reasoning-based llms for arabic nlp,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Ararea- soner: Evaluating reasoning-based llms for arabic nlp,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:175062cf92905ff294172bd8155cd4b737bf7a920051c8cf60b7b2c53259010b

Pith citing papers

Observation 904aa2ba-8e55-437c-8f3a-8f9c91a4b9e6 · inbound

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation cites this paper.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:fd78f940247b343b80ad7a130eb5d8ae2fb183cee560ee9075727f9a69c72a26