Pith. sign in

Paper Citation Record · LEDGER

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism

As of 21 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 0 inbound Pith citation observations for arXiv:2504.16761.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.16761 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:59:50.316255Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

42 of 42 outbound references displayed

  • verified exact0
  • verified fuzzy29
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2e1cb311-37ca-4359-b927-c7ab66e21405 · outbound

This paper cites Deep learning approaches on image captioning: A review,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Deep learning approaches on image captioning: A review,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T10:59:50.120841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:59:50.120841Z digest=sha256:ffa6098269b091364e1080ce3fa96651be001bb8e66b4bffcc6fc703756dcd27

Observation b7b0b336-bd7a-41e4-a778-0eb9cd830dfb · outbound

This paper cites From methods to datasets: A survey on image-caption generators,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism From methods to datasets: A survey on image-caption generators,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.893753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.127279Z digest=sha256:ef650117699b0b38be6f562f1c5088becffcd85185b0f46ffc880b9e53c74ee7

Observation 87fdbd8f-dce6-4cf8-a912-db90c75a14e4 · outbound

This paper cites A survey on vision transformer,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism A survey on vision transformer,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T10:59:50.132579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:59:50.132579Z digest=sha256:a6606f37320f0959f5402b7164e50f379750564237c60b45e191a0f3fb0d55a8

Observation ffe5c0a4-8251-4e05-87b5-52ec63548431 · outbound

This paper cites Text augmentation using bert for image captioning,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Text augmentation using bert for image captioning,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.870814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.137872Z digest=sha256:3bb1d0b7165e0c274211f6b0a25e0322cd87a09c364dea75331cd640f0761bdc

Observation c73c811d-153d-4a05-9a24-71dcd9445b7f · outbound

This paper cites Contrastive language- image pre-training with knowledge graphs,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Contrastive language- image pre-training with knowledge graphs,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.856515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.143478Z digest=sha256:3bbb0276d9cde16d56e1a32347c092b949313a0a0b3c241a9a326dc64ceaf7dd

Observation 796b4b33-25ed-4c19-875f-c9efd9dcf15b · outbound

This paper cites Entangled transformer for image captioning,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Entangled transformer for image captioning,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.841837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.148733Z digest=sha256:06c216c9e514ae92d1c0bfa8a0d09716e0ac4f35000172236d6df236bfc209eb

Observation 50f91054-d0dc-4760-8b62-d836717ef205 · outbound

This paper cites Meshed-memory transformer for image captioning,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Meshed-memory transformer for image captioning,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T10:59:50.154037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:59:50.154037Z digest=sha256:5561d3b37f82d7a5cbd88de72cd3750e78df66dd0d7370e2906c44a9d7d30cb0

Observation 2ae568fc-702d-49ea-a29f-d76e962bfea7 · outbound

This paper cites Multimodal transformer with multi- view visual representation for image captioning,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Multimodal transformer with multi- view visual representation for image captioning,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T10:59:50.159043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:59:50.159043Z digest=sha256:580e55f3c5608d0c9cb17f5f4f5f7b4ce0bd23f2c921428c2fc6a0984a894a96

Observation 01d9fa9e-2d74-4b48-9bff-46ff8475adf6 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Learning transferable visual models from natural language supervision,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T10:59:50.163495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:59:50.163495Z digest=sha256:e8e130142c49c247cc1cbf1edede777d81763d55fc7dfd0c6e7b4609d7b5e01f

Observation 6f91ecc6-80b0-4a38-9fe3-dbf7a1a2c04a · outbound

This paper cites MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T10:59:50.168472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:59:50.168472Z digest=sha256:970104d56e51d77e79e4f502f0a16bd69eaba43d081ef68439affa14324363a6

Observation 9df9c8cd-5337-4a79-860e-83625bfa41cb · outbound

This paper cites Samt- generator: A second-attention for image captioning based on multi-stage transformer network,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Samt- generator: A second-attention for image captioning based on multi-stage transformer network,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.799177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.174099Z digest=sha256:5ea2408e01d147b5402a458cd18d19c144c2f79b5821033d16a60bac7738e9e0

Observation c7c0c7ab-4bfc-4e79-b76c-8d6c95c89611 · outbound

This paper cites S2 transformer for image captioning.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism S2 transformer for image captioning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.785470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.178706Z digest=sha256:cc4d7a1052363228e66098fb040191ae0c60354c0416ec8868cab5192c8be3a7

Observation 835c1840-d3b7-49a5-a270-64fb83d1497e · outbound

This paper cites Ca-captioner: A novel concentrated attention for image captioning,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Ca-captioner: A novel concentrated attention for image captioning,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.770415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.182811Z digest=sha256:ab000ddcfca7ba2e1330caaf17e64bc6bb9c2cd462fe5669a160b1c805dccc38

Observation 53e8609a-4a5e-44a0-a4a2-b47043953969 · outbound

This paper cites Image cap- tioning using transformer-based double attention network,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Image cap- tioning using transformer-based double attention network,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.755184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.187255Z digest=sha256:625457771e273c7f7d152b6f88e94c574df94c7ec49cd5a7a537a8d30935b1fa

Observation f6d96a6a-8c9c-40b5-89c7-3410ac9c3490 · outbound

This paper cites With a little help from your own past: Prototypical memory networks for image captioning,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism With a little help from your own past: Prototypical memory networks for image captioning,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.739343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.192488Z digest=sha256:e7e37f2b5f0f96af51555cebb0a2f9c21430ee6d5ca1d8bc7ad252f1da1e4914

Observation a9809b41-cca4-45cf-9dd5-5876bed97e91 · outbound

This paper cites Haav: Hierarchical aggregation of augmented views for image captioning,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Haav: Hierarchical aggregation of augmented views for image captioning,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.725022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.198070Z digest=sha256:504222a754b0c82c2c6f2022cff69db405102e4a601c370b0fa064c56f65fbc0

Observation 9a3bed3e-b3c6-46b0-bdc9-31a73a0dd0b1 · outbound

This paper cites Dual vision transformer,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Dual vision transformer,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.709713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.202445Z digest=sha256:e4eb8df2aa75674bdbe5dad8466689c590257d7a107385d6539168b61f09d619

Observation fe0f43ed-15bd-42ac-b9c9-6adbc4859024 · outbound

This paper cites Spt: Spatial pyramid transformer for image captioning,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Spt: Spatial pyramid transformer for image captioning,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.693549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.206883Z digest=sha256:e4da5914d615d110ec76c8adce19931f815c91197c631e681a96b49a0a6593e3

Observation e15e0415-71c3-4b20-b798-97944ec6a639 · outbound

This paper cites Show and tell: A neural image caption generator,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Show and tell: A neural image caption generator,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T10:59:50.211563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:59:50.211563Z digest=sha256:56c5e9e472a702674afa11a11cb96571da174b4443d43f660e296095e22f29ad

Observation 430aab14-aa6e-4b49-b049-e57d8578222d · outbound

This paper cites Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T10:59:50.216074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:59:50.216074Z digest=sha256:96bfcec14973b1eee3967dfc5213e1d829f7048a41bb35465bcaacbe0cc32c85

Observation a037e398-a076-4bfd-8ecf-67c5544adaec · outbound

This paper cites A topic-based multi-channel attention model under hybrid mode for image caption,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism A topic-based multi-channel attention model under hybrid mode for image caption,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.659361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.220542Z digest=sha256:ac79226a9db350cde9a8a332195a994a37361283773b6930664c720367278f43

Observation 9bfd8312-5085-4675-b998-0214f19a2c6e · outbound

This paper cites Transformer model incorporating local graph semantic attention for image caption,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Transformer model incorporating local graph semantic attention for image caption,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.644315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.224647Z digest=sha256:e1eaa93802a1656c4a6d486d1fffb3e042b899a3c800a0352eeb839f56590478

Observation 343b0f64-5407-40fa-b452-aae5d9a44804 · outbound

This paper cites Dynamic-balanced double-attention fusion for image captioning,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Dynamic-balanced double-attention fusion for image captioning,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.629265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.228795Z digest=sha256:14d8e439b4b16198d054bd886d495edde6c3532bba4a2a24e22b087c1ae5984a

Observation 74635107-53e3-4e22-8f17-3d180e6f0165 · outbound

This paper cites Improving image captioning by leveraging intra-and inter-layer global representation in transformer network,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Improving image captioning by leveraging intra-and inter-layer global representation in transformer network,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.614455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.233363Z digest=sha256:a4b3a411f8f7c60ef7a630ebd300e2500cb3488e6db83665c8f587c58fc2abfd

Observation 0a3f1b45-9da4-40b4-a68a-9f61f8e7dd69 · outbound

This paper cites Geometry attention transformer with position-aware lstms for image captioning,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Geometry attention transformer with position-aware lstms for image captioning,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.599268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.238055Z digest=sha256:710fea4d18c6dfc542ca2dad9ec4adfce1b6e720e0ac200654769d047a20b4c6

Observation 540ef87b-320b-4f75-85fd-47163a910185 · outbound

This paper cites Vision-enhanced and consensus- aware transformer for image captioning,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Vision-enhanced and consensus- aware transformer for image captioning,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-16T10:59:50.243481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:59:50.243481Z digest=sha256:fcdc3bdd9ccc8009ab4b0255439d087be43d1854c60066f96e94cdbefca93437

Observation f2e2bed9-5fae-4ef8-a674-cdf1488301c1 · outbound

This paper cites A novel cross-fusion method of different types of features for image captioning,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism A novel cross-fusion method of different types of features for image captioning,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.576910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.248490Z digest=sha256:681e970bb5f8b7f785b81affb3f9cd476c09138ac99474e1d4210c483378b10a

Observation 19673f21-c8db-439d-bf9b-9379c1a07100 · outbound

This paper cites Transformer- based local-global guidance for image captioning,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Transformer- based local-global guidance for image captioning,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.562334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.253464Z digest=sha256:dd7ed9dc4797440ff9241ef082d8dff4d42cbd7d70d384e47065b45ac97bdc88

Observation 1f05575c-ef93-461c-92bb-6fa746ef6e96 · outbound

This paper cites Deep Captioning with Multimodal Recurrent Neural Networks (m-RNN).

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Deep Captioning with Multimodal Recurrent Neural Networks (m-RNN)

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T10:59:50.258035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:59:50.258035Z digest=sha256:34b2efc329a19c7d115f916b5b157f35181530cf0248888304fc29f7e60efbc4

Observation 047cc0eb-30a7-45a8-8844-a72b6c9746ff · outbound

This paper cites Re- view networks for caption generation,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Re- view networks for caption generation,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.547899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.262490Z digest=sha256:b813031e9ce41ea36063efbfcf7d5cb67c146895df63ceaa58707e3b147d6630

Observation 2539ab21-cf4c-4897-af52-1c6f36889711 · outbound

This paper cites Semantic compositional networks for visual captioning,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Semantic compositional networks for visual captioning,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.532672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.266726Z digest=sha256:a2f5f2523b9173b27e3cd907de0d7361bd2b62db0c3b2b1f40e154260c20d626

Observation d3c9ac04-1a96-4481-91db-d5ecbfc99201 · outbound

This paper cites Knowing when to look: Adaptive attention via a visual sentinel for image captioning,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Knowing when to look: Adaptive attention via a visual sentinel for image captioning,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-16T10:59:50.271127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:59:50.271127Z digest=sha256:a223c6afbeaae78644eb6622a026a4d7ff727e89d6846158eb082c3e7a49a578

Observation 595b87ea-c349-4863-a98d-25d380a7d535 · outbound

This paper cites Self- critical sequence training for image captioning,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Self- critical sequence training for image captioning,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-16T10:59:50.275779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:59:50.275779Z digest=sha256:6cfe7d2b48d3253e2b0edba2331a6a81db62b16530fe76fed3b1ffb57c2331a0

Observation ad6fd4ab-682f-4b6a-9040-8e04d72780f4 · outbound

This paper cites Gatecap: Gated spatial and semantic attention model for image captioning,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Gatecap: Gated spatial and semantic attention model for image captioning,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.499479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.281535Z digest=sha256:ae76e73fd010176cd6a9360d44b48940a74d5c55453b3b011f6ff8b56bdc179f

Observation 8df0a377-a171-43a8-bd92-ff17314d8f73 · outbound

This paper cites Boosting image captioning with attributes,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Boosting image captioning with attributes,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.484054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.285739Z digest=sha256:40de8a084512999b49e2c7bc0f4dbb6e2df656834b44f2e72b635bdd6927af85

Observation 60dbf035-3441-4005-a563-33243def2958 · outbound

This paper cites Bottom-up and top-down attention for image captioning and visual question answering,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Bottom-up and top-down attention for image captioning and visual question answering,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-16T10:59:50.290298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:59:50.290298Z digest=sha256:c3c0744461ee35ebc9405effe41f9fa124cf709ddf780fb9397d5a2b83c8dbe0

Observation 212a3d8d-5a4a-4043-8042-a318f2f89b10 · outbound

This paper cites Recurrent fusion network for image captioning,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Recurrent fusion network for image captioning,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.459200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.294473Z digest=sha256:7d07c1685a01ac7dba1dc3076a16c4f50eb10178b715b3e57d683287ad5fc584

Observation 42127f5c-5451-4bfd-b3d8-0973b8a6c1b3 · outbound

This paper cites Exploring visual relationship for image captioning,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Exploring visual relationship for image captioning,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.443247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.298918Z digest=sha256:8a177c789976d95aac1203f497010744e2f1a2d362f017c8db56fe97e5bf15fd

Observation f280bec8-dabd-4509-b03a-5d18894cfa55 · outbound

This paper cites Auto-encoding scene graphs for image captioning,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Auto-encoding scene graphs for image captioning,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.428184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.303322Z digest=sha256:60e794f72886148550c1adcb8e53a16c138cc2ec4d26ab0518982b72e651e0fd

Observation 1c866516-02e6-4314-b6e8-65afb39f6aad · outbound

This paper cites Attention on attention for image captioning,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Attention on attention for image captioning,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.413370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.307714Z digest=sha256:8aa0df3d6fa27a48988cb06a1a6df87a7fa5ff5ef15a18f01df973e0080298b5

Observation d2ab1097-61cb-4866-aae6-f7430c5de033 · outbound

This paper cites Image captioning using vision encoder decoder model,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Image captioning using vision encoder decoder model,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.398123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.312099Z digest=sha256:cb303f01d04ac23f62a76b360adc8b2f5ecd43ca457289a042e51da29cf643d0

Observation 4a49fa3d-0689-4799-897d-a967c9f2188b · outbound

This paper cites Optimal trans- formers based image captioning using beam search,.

Tri-FusionNet: Enhancing Image Description Generation with Transformer-based Fusion Network and Dual Attention Mechanism Optimal trans- formers based image captioning using beam search,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:59:50.383025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T10:59:50.316255Z digest=sha256:b3aca675367c1186198dde5f5c2768bf8870fdb044a1f87c395d948250f53467

Pith citing papers

No inbound Pith citation observations are available.