Pith. sign in

Paper Citation Record · LEDGER

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation

As of 21 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 1 inbound Pith citation observation for arXiv:2607.03657.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.03657 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-12T00:53:05.419742Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-12T00:53:05.419742Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

42 of 42 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved42
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 44f31bfa-a977-4f26-af64-c025b218ebd9 · outbound

This paper cites Sign Language Trans- lation (SLT) aims to bridge communication gaps by trans- lating sign-language videos into spoken sentences.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Sign Language Trans- lation (SLT) aims to bridge communication gaps by trans- lating sign-language videos into spoken sentences

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:26460adfde207e28c7ef807df2fe6892aba49144c338a90caa0ccb53d732333b

Observation 904aa2ba-8e55-437c-8f3a-8f9c91a4b9e6 · outbound

This paper cites ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:dff1ebf28c581f5806baa06add93926f23388711e8ba04f5aa4880699c77ba79

Observation eeddb7e1-89c9-4e83-82c4-fe481fe0d951 · outbound

This paper cites SignLLM [5] maps sign videos to discrete tokens aligned with LLMs.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation SignLLM [5] maps sign videos to discrete tokens aligned with LLMs

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:7d92c1676605fa10f77c2e364556522bc1e0508b4858b3d2724c472383cdda15

Observation 9d36b98f-90a7-47a3-bc64-605afec9d8ea · outbound

This paper cites Re- cent work leverageslarge-scale pretraining and multimodal LLMs.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Re- cent work leverageslarge-scale pretraining and multimodal LLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:8e312bd4390011e59b5968ae04f60fb53054358a13831b2db5951b010c48f8f8

Observation 2d32709d-b02d-45a6-b90b-42852d90b0fb · outbound

This paper cites Translate the given sentence into<language>.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Translate the given sentence into<language>

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:9e9a3e980d20794c7dc38a6041f0eaacda6db00d5d4eb286b577e598bba5ebfe

Observation df985ce5-4bd3-4222-9df0-a871dd45dd92 · outbound

This paper cites Experimental Setup Datasets.We evaluate on two benchmark SLT datasets: PHOENIX14T [25] and CSL-Daily [26].

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Experimental Setup Datasets.We evaluate on two benchmark SLT datasets: PHOENIX14T [25] and CSL-Daily [26]

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:b7e1645bee418915f7253548c00c3a91c16ba13f37dd73b9cf80b5b64907dca0

Observation c81409e8-09fd-4e08-a012-c74491eedae3 · outbound

This paper cites an unresolved cited work.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:9522b69eddc6e617e9d42e8d8d4cd360f7b02fa6796385daf3c15f15ff57620a

Observation 642b8bcb-b2c0-4ebe-aef9-20b67b7de116 · outbound

This paper cites an unresolved cited work.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:a7f7876828603c7166b25edeace08ba8cf1a389d359701a3d405cde63d200d5c

Observation 45e7175b-e3e7-4d58-8488-ebfc66214358 · outbound

This paper cites Sign language transformers: Joint end-to-end sign language recognition and translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Sign language transformers: Joint end-to-end sign language recognition and translation,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:8763954f84bbf74ae5dafcb3364a58d88bf98536e0730ee1e39e4923c5e9a31b

Observation 1266af36-d757-4300-b91f-8f7aa15ddb6a · outbound

This paper cites Factorized Learning Assisted with Large Language Model for Gloss-free Sign Language Translation.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Factorized Learning Assisted with Large Language Model for Gloss-free Sign Language Translation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:c2d00870d00c02a6ec56b905bcf8a28dc7c2fb2e1fc72704a44edd0981749bd0

Observation 21f54591-24a9-4b3b-9817-13aa0b883718 · outbound

This paper cites Gloss-free sign language translation: Improving from visual-language pretraining,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Gloss-free sign language translation: Improving from visual-language pretraining,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:5dd5307edcd18f460883590eea5d673f3eb48209b11a78491860eae5421b68c2

Observation abf0959e-9119-4624-ad2f-a1b7f504f03c · outbound

This paper cites Leveraging the power of mllms for gloss-free sign language transla- tion,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Leveraging the power of mllms for gloss-free sign language transla- tion,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:748f453c29f6c02100575322b3e27ea54d2a714869fae8fcc35591ea2395f4bf

Observation 53095427-6741-48d3-8723-c0745c10db7b · outbound

This paper cites Llms are good sign language translators,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Llms are good sign language translators,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:e337f59e1d358f667383c984d059da818b6a193f3a2567c7b51ee18acf64c6f0

Observation dbfc33fa-6e84-47ad-9f2a-ec8cdfbf987c · outbound

This paper cites Lost in translation, found in context: Sign language translation with contextual cues,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Lost in translation, found in context: Sign language translation with contextual cues,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:62ea7321dec6f9cfae3757984296b3713dff2954989effb8738e06d73f3143a0

Observation f5e1c48a-cb80-4ae4-bc19-befe5bae6ba5 · outbound

This paper cites Sign2GPT: Leveraging Large Language Models for Gloss-Free Sign Language Translation.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Sign2GPT: Leveraging Large Language Models for Gloss-Free Sign Language Translation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:5b056d561da345bdf5c5430f6d9f4362c299a99b4b0accf86d99b78075e49d04

Observation e9d58d38-1a59-4ecb-b248-bf0117adca00 · outbound

This paper cites Tspnet: Hierarchical feature learning via temporal semantic pyramid for sign language translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Tspnet: Hierarchical feature learning via temporal semantic pyramid for sign language translation,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:bd0893e47a32eade6122ea9c0aa104d2229dfd2d7d61893a9fa5c820437cbff6

Observation 4b58aec4-71d2-41e9-aac2-f72c8c68e281 · outbound

This paper cites Conditional sen- tence generation and cross-modal reranking for sign lan- guage translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Conditional sen- tence generation and cross-modal reranking for sign lan- guage translation,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:5abd2c65e5d6ffb60f0c14c504d9a91a9f3c4c672bdce476b9ba52cafbcad402

Observation 0b2a8b37-1673-4aeb-b5e4-3ae2e8ec5e11 · outbound

This paper cites A token-level contrastive framework for sign language translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation A token-level contrastive framework for sign language translation,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:322eea61643d83d93bc8b4e451af26d090c3299652a9f389435b2fc4b9b2a316

Observation 0dc7d26f-fe5a-474f-a41b-5cd7620aeb47 · outbound

This paper cites Gloss attention for gloss-free sign language translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Gloss attention for gloss-free sign language translation,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:aab9d96e93539b9a75f29a37e7319044438656c071f329e2b4c8d22bebd3768e

Observation ca9a3f79-3091-4547-9b77-5bad231a13da · outbound

This paper cites Visual alignment pre- training for sign language translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Visual alignment pre- training for sign language translation,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:fe6aaf20c8c9b198ca338e6d90f6a8278a6b942977ca48ea36f7558d11a60615

Observation 4b541d6c-79b4-4367-a7d8-495388c5ac6d · outbound

This paper cites An Efficient Sign Language Translation Using Spatial Configuration and Motion Dynamics with LLMs.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation An Efficient Sign Language Translation Using Spatial Configuration and Motion Dynamics with LLMs

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:c65fc7865c67646245b8398bec8088de78a6be74a2e27bb80e976d8b65a6f4ec

Observation 58457a83-56e3-413b-9060-9485926f7cc6 · outbound

This paper cites Better sign language translation with stmc-transformer,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Better sign language translation with stmc-transformer,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:9a287566b57fbd3d54521cf231f90979ccd175c78bd415209604bba774821f64

Observation e2d25ca2-ac59-4de3-9a14-e6329f930d5d · outbound

This paper cites Lost in translation, found in embeddings: Sign language translation and alignment,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Lost in translation, found in embeddings: Sign language translation and alignment,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:6a2e273f9c7c09bb531df0f09559be83f6e0a98f0b5794be8daf3ac9d7a93c8b

Observation 56c40903-7d60-4aff-8480-32c7eba70a65 · outbound

This paper cites Multimodal sign language recognition via temporal deformable convolu- tional sequence learning.,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Multimodal sign language recognition via temporal deformable convolu- tional sequence learning.,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:aee6dabadb6938028357edee9e7196d0beb1c32d70041a737a9eb3ed16450293

Observation 2468cb21-2394-467c-9cb3-81cdd806bb20 · outbound

This paper cites Multi-channel transformers for multi-articulatory sign language translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Multi-channel transformers for multi-articulatory sign language translation,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:08b375d162b62c3031a853a56acdedbebc82d386370d458d4fbef7396ff80544

Observation 5d8d1f0f-2055-45ea-8387-e00c117fe462 · outbound

This paper cites Spatial- temporal multi-cue network for sign language recogni- tion and translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Spatial- temporal multi-cue network for sign language recogni- tion and translation,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:25452cea46749f23af0adcbf3a211ff760ac58f8dd82608e9043a16e89ff0bd0

Observation 61c3a763-a5c2-4172-86aa-1f2ef5fe52c9 · outbound

This paper cites Graph- based multimodal sequential embedding for sign lan- guage translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Graph- based multimodal sequential embedding for sign lan- guage translation,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:7d643230de4c6a9a10ee04e3577daa469d67a63202ef45f5ae1286f63068234b

Observation 565ab00d-b204-4af9-9c4b-e5fdf8c5744c · outbound

This paper cites Skeleton-aware neural sign language translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Skeleton-aware neural sign language translation,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:7c3681459ac99e5538e487eb5707d8ba3028ab8b25205a58780534911fc552b5

Observation 5c2752ac-d1c2-4699-a948-3259f7b9a05f · outbound

This paper cites Two-stream net- work for sign language recognition and translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Two-stream net- work for sign language recognition and translation,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:fc260640568d857d0da4b06798e83897b521b66ad162f813418b2b24c1a4f0da

Observation 0bf15414-6b57-4033-9535-3d91fd91406d · outbound

This paper cites Ustm: Unified spa- tial and temporal modeling for continuous sign language recognition,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Ustm: Unified spa- tial and temporal modeling for continuous sign language recognition,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:f6a5020d85d0829ef93723be08c26533d2407e85117e5a4be3b7a2291694aade

Observation d51bd3aa-e5ff-4acf-8d1d-e14eebc5a5e9 · outbound

This paper cites Openpose: Re- altime multi-person 2d pose estimation using part affin- ity fields,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Openpose: Re- altime multi-person 2d pose estimation using part affin- ity fields,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:96095acc299e2d53ea29e71a4a3a00843869ebdedc1989ba987d627fa7309701

Observation 005b2082-0c03-48e2-8127-b7701bc67a87 · outbound

This paper cites Temporal convolutional networks for action segmentation and de- tection,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Temporal convolutional networks for action segmentation and de- tection,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:88384f935f35daaf0efac5623f46365efd6ad9eacd3270fb381db2ad78a09dd3

Observation 4cc18bda-1be3-4653-8566-decea2ce0c1b · outbound

This paper cites Neural sign language translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Neural sign language translation,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:fe2d912b8dc2948cad45b1e6489f57016d4472b9c24cf54e6a0a275439add5d5

Observation d932677a-5331-4ba5-a516-5f5aced9e2ed · outbound

This paper cites Improving sign language translation with monolingual data by sign back-translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Improving sign language translation with monolingual data by sign back-translation,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:28692bf4f4f3ea791088398c9b81884f0e2ed3b3b70bcfff9706d06cbd2d1eb6

Observation ab9a8202-c176-4fa8-9314-30088f71a9c9 · outbound

This paper cites Stochastic transformer networks with linear compet- ing units: Application to end-to-end sl translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Stochastic transformer networks with linear compet- ing units: Application to end-to-end sl translation,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:e20592cc27b89f4b052a96ef5bf89e204836f7171e0a5310b649d70b11089b77

Observation 4e631fb4-d64c-49c8-b802-af43b1aed53f · outbound

This paper cites Mska: Multi- stream keypoint attention network for sign language recognition and translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Mska: Multi- stream keypoint attention network for sign language recognition and translation,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:574ee720359a230833541b00070c882b0f2402a516223bb429ee71341648e619

Observation 789c2d50-9bc0-45b0-925c-bdb1339e3292 · outbound

This paper cites Cross-modality Data Augmentation for End-to-End Sign Language Translation.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Cross-modality Data Augmentation for End-to-End Sign Language Translation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:361db84bec98f77f4efa5f70edf75088f1d9cbb3558e01c98bba0dc85ab6579c

Observation 490c6297-7a99-4be4-a976-c157a1ff7f45 · outbound

This paper cites Crosslingual generalization through multitask finetun- ing,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Crosslingual generalization through multitask finetun- ing,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:9863abbfce4ab7b5dca2ec7d544a872372b0a305a8daffeffebfcb6160314afa

Observation c4114bb0-25e7-41a2-9793-96604dbc18e1 · outbound

This paper cites Lora: Low-rank adaptation of large language models,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Lora: Low-rank adaptation of large language models,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:eca8cfe07fdfbb5971673f8f17f987a6e6fb7aef7ad60f312422b8a00dc1ec8b

Observation 9ae2e96d-8417-4cd5-b12c-19048fb1af02 · outbound

This paper cites Scaling instruction-finetuned language models,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Scaling instruction-finetuned language models,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:beb16fcc6d3b2c81e583d0035321cf71c2e98948965df367446ef719b3572510

Observation 82606eee-2de5-4d61-bb27-1329984ec715 · outbound

This paper cites Multilingual denois- ing pre-training for neural machine translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Multilingual denois- ing pre-training for neural machine translation,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:3e37856e411fccc18b658ded73d66fcea5c9008063e05340716b1eb6f0f7cf50

Observation 9795a18a-e8a6-4886-b498-f3686aa20b6c · outbound

This paper cites Ararea- soner: Evaluating reasoning-based llms for arabic nlp,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Ararea- soner: Evaluating reasoning-based llms for arabic nlp,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:475dd7fc6bf6e368386ac26114ac506a89a009dcbcc5412cc532304a42ffaca7

Pith citing papers

Observation 904aa2ba-8e55-437c-8f3a-8f9c91a4b9e6 · inbound

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation cites this paper.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:dff1ebf28c581f5806baa06add93926f23388711e8ba04f5aa4880699c77ba79