Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-12T00:53:05.419742Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 1 inbound Pith citation observation for arXiv:2607.03657.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-12T00:53:05.419742Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-12T00:53:05.419742Z
A source-named dated measurement, never combined with another source.
Source: cited_works
42 of 42 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 44f31bfa-a977-4f26-af64-c025b218ebd9 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Sign Language Trans- lation (SLT) aims to bridge communication gaps by trans- lating sign-language videos into spoken sentences
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 904aa2ba-8e55-437c-8f3a-8f9c91a4b9e6 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eeddb7e1-89c9-4e83-82c4-fe481fe0d951 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation SignLLM [5] maps sign videos to discrete tokens aligned with LLMs
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d36b98f-90a7-47a3-bc64-605afec9d8ea · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Re- cent work leverageslarge-scale pretraining and multimodal LLMs
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d32709d-b02d-45a6-b90b-42852d90b0fb · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Translate the given sentence into<language>
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df985ce5-4bd3-4222-9df0-a871dd45dd92 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Experimental Setup Datasets.We evaluate on two benchmark SLT datasets: PHOENIX14T [25] and CSL-Daily [26]
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c81409e8-09fd-4e08-a012-c74491eedae3 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Unresolved cited work
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 642b8bcb-b2c0-4ebe-aef9-20b67b7de116 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45e7175b-e3e7-4d58-8488-ebfc66214358 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Sign language transformers: Joint end-to-end sign language recognition and translation,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1266af36-d757-4300-b91f-8f7aa15ddb6a · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Factorized Learning Assisted with Large Language Model for Gloss-free Sign Language Translation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21f54591-24a9-4b3b-9817-13aa0b883718 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Gloss-free sign language translation: Improving from visual-language pretraining,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abf0959e-9119-4624-ad2f-a1b7f504f03c · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Leveraging the power of mllms for gloss-free sign language transla- tion,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53095427-6741-48d3-8723-c0745c10db7b · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Llms are good sign language translators,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbfc33fa-6e84-47ad-9f2a-ec8cdfbf987c · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Lost in translation, found in context: Sign language translation with contextual cues,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5e1c48a-cb80-4ae4-bc19-befe5bae6ba5 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Sign2GPT: Leveraging Large Language Models for Gloss-Free Sign Language Translation
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9d58d38-1a59-4ecb-b248-bf0117adca00 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Tspnet: Hierarchical feature learning via temporal semantic pyramid for sign language translation,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b58aec4-71d2-41e9-aac2-f72c8c68e281 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Conditional sen- tence generation and cross-modal reranking for sign lan- guage translation,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b2a8b37-1673-4aeb-b5e4-3ae2e8ec5e11 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation A token-level contrastive framework for sign language translation,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0dc7d26f-fe5a-474f-a41b-5cd7620aeb47 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Gloss attention for gloss-free sign language translation,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca9a3f79-3091-4547-9b77-5bad231a13da · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Visual alignment pre- training for sign language translation,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b541d6c-79b4-4367-a7d8-495388c5ac6d · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation An Efficient Sign Language Translation Using Spatial Configuration and Motion Dynamics with LLMs
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58457a83-56e3-413b-9060-9485926f7cc6 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Better sign language translation with stmc-transformer,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2d25ca2-ac59-4de3-9a14-e6329f930d5d · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Lost in translation, found in embeddings: Sign language translation and alignment,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56c40903-7d60-4aff-8480-32c7eba70a65 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Multimodal sign language recognition via temporal deformable convolu- tional sequence learning.,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2468cb21-2394-467c-9cb3-81cdd806bb20 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Multi-channel transformers for multi-articulatory sign language translation,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d8d1f0f-2055-45ea-8387-e00c117fe462 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Spatial- temporal multi-cue network for sign language recogni- tion and translation,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61c3a763-a5c2-4172-86aa-1f2ef5fe52c9 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Graph- based multimodal sequential embedding for sign lan- guage translation,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 565ab00d-b204-4af9-9c4b-e5fdf8c5744c · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Skeleton-aware neural sign language translation,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c2752ac-d1c2-4699-a948-3259f7b9a05f · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Two-stream net- work for sign language recognition and translation,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bf15414-6b57-4033-9535-3d91fd91406d · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Ustm: Unified spa- tial and temporal modeling for continuous sign language recognition,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d51bd3aa-e5ff-4acf-8d1d-e14eebc5a5e9 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Openpose: Re- altime multi-person 2d pose estimation using part affin- ity fields,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 005b2082-0c03-48e2-8127-b7701bc67a87 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Temporal convolutional networks for action segmentation and de- tection,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cc18bda-1be3-4653-8566-decea2ce0c1b · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Neural sign language translation,
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d932677a-5331-4ba5-a516-5f5aced9e2ed · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Improving sign language translation with monolingual data by sign back-translation,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab9a8202-c176-4fa8-9314-30088f71a9c9 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Stochastic transformer networks with linear compet- ing units: Application to end-to-end sl translation,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e631fb4-d64c-49c8-b802-af43b1aed53f · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Mska: Multi- stream keypoint attention network for sign language recognition and translation,
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 789c2d50-9bc0-45b0-925c-bdb1339e3292 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Cross-modality Data Augmentation for End-to-End Sign Language Translation
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 490c6297-7a99-4be4-a976-c157a1ff7f45 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Crosslingual generalization through multitask finetun- ing,
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4114bb0-25e7-41a2-9793-96604dbc18e1 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Lora: Low-rank adaptation of large language models,
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ae2e96d-8417-4cd5-b12c-19048fb1af02 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Scaling instruction-finetuned language models,
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82606eee-2de5-4d61-bb27-1329984ec715 · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Multilingual denois- ing pre-training for neural machine translation,
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9795a18a-e8a6-4886-b498-f3686aa20b6c · outbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Ararea- soner: Evaluating reasoning-based llms for arabic nlp,
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 904aa2ba-8e55-437c-8f3a-8f9c91a4b9e6 · inbound
ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.