Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T10:59:50.316255Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 0 inbound Pith citation observations for arXiv:2504.16761.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T10:59:50.316255Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
42 of 42 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2e1cb311-37ca-4359-b927-c7ab66e21405 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Deep learning approaches on image captioning: A review,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7b0b336-bd7a-41e4-a778-0eb9cd830dfb · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism From methods to datasets: A survey on image-caption generators,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 87fdbd8f-dce6-4cf8-a912-db90c75a14e4 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism A survey on vision transformer,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffe5c0a4-8251-4e05-87b5-52ec63548431 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Text augmentation using bert for image captioning,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation c73c811d-153d-4a05-9a24-71dcd9445b7f · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Contrastive language- image pre-training with knowledge graphs,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 796b4b33-25ed-4c19-875f-c9efd9dcf15b · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Entangled transformer for image captioning,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 50f91054-d0dc-4760-8b62-d836717ef205 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Meshed-memory transformer for image captioning,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ae568fc-702d-49ea-a29f-d76e962bfea7 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Multimodal transformer with multi- view visual representation for image captioning,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01d9fa9e-2d74-4b48-9bff-46ff8475adf6 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Learning transferable visual models from natural language supervision,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f91ecc6-80b0-4a38-9fe3-dbf7a1a2c04a · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9df9c8cd-5337-4a79-860e-83625bfa41cb · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Samt- generator: A second-attention for image captioning based on multi-stage transformer network,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation c7c0c7ab-4bfc-4e79-b76c-8d6c95c89611 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism S2 transformer for image captioning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 835c1840-d3b7-49a5-a270-64fb83d1497e · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Ca-captioner: A novel concentrated attention for image captioning,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 53e8609a-4a5e-44a0-a4a2-b47043953969 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Image cap- tioning using transformer-based double attention network,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation f6d96a6a-8c9c-40b5-89c7-3410ac9c3490 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism With a little help from your own past: Prototypical memory networks for image captioning,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation a9809b41-cca4-45cf-9dd5-5876bed97e91 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Haav: Hierarchical aggregation of augmented views for image captioning,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 9a3bed3e-b3c6-46b0-bdc9-31a73a0dd0b1 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Dual vision transformer,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation fe0f43ed-15bd-42ac-b9c9-6adbc4859024 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Spt: Spatial pyramid transformer for image captioning,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation e15e0415-71c3-4b20-b798-97944ec6a639 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Show and tell: A neural image caption generator,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 430aab14-aa6e-4b49-b049-e57d8578222d · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a037e398-a076-4bfd-8ecf-67c5544adaec · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism A topic-based multi-channel attention model under hybrid mode for image caption,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 9bfd8312-5085-4675-b998-0214f19a2c6e · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Transformer model incorporating local graph semantic attention for image caption,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 343b0f64-5407-40fa-b452-aae5d9a44804 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Dynamic-balanced double-attention fusion for image captioning,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 74635107-53e3-4e22-8f17-3d180e6f0165 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Improving image captioning by leveraging intra-and inter-layer global representation in transformer network,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 0a3f1b45-9da4-40b4-a68a-9f61f8e7dd69 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Geometry attention transformer with position-aware lstms for image captioning,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 540ef87b-320b-4f75-85fd-47163a910185 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Vision-enhanced and consensus- aware transformer for image captioning,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2e2bed9-5fae-4ef8-a674-cdf1488301c1 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism A novel cross-fusion method of different types of features for image captioning,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 19673f21-c8db-439d-bf9b-9379c1a07100 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Transformer- based local-global guidance for image captioning,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 1f05575c-ef93-461c-92bb-6fa746ef6e96 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Deep Captioning with Multimodal Recurrent Neural Networks (m-RNN)
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 047cc0eb-30a7-45a8-8844-a72b6c9746ff · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Re- view networks for caption generation,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2539ab21-cf4c-4897-af52-1c6f36889711 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Semantic compositional networks for visual captioning,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation d3c9ac04-1a96-4481-91db-d5ecbfc99201 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Knowing when to look: Adaptive attention via a visual sentinel for image captioning,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 595b87ea-c349-4863-a98d-25d380a7d535 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Self- critical sequence training for image captioning,
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad6fd4ab-682f-4b6a-9040-8e04d72780f4 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Gatecap: Gated spatial and semantic attention model for image captioning,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 8df0a377-a171-43a8-bd92-ff17314d8f73 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Boosting image captioning with attributes,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 60dbf035-3441-4005-a563-33243def2958 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Bottom-up and top-down attention for image captioning and visual question answering,
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 212a3d8d-5a4a-4043-8042-a318f2f89b10 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Recurrent fusion network for image captioning,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 42127f5c-5451-4bfd-b3d8-0973b8a6c1b3 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Exploring visual relationship for image captioning,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation f280bec8-dabd-4509-b03a-5d18894cfa55 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Auto-encoding scene graphs for image captioning,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 1c866516-02e6-4314-b6e8-65afb39f6aad · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Attention on attention for image captioning,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation d2ab1097-61cb-4866-aae6-f7430c5de033 · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Image captioning using vision encoder decoder model,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 4a49fa3d-0689-4799-897d-a967c9f2188b · outbound
Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Optimal trans- formers based image captioning using beam search,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
No inbound Pith citation observations are available.