Pith. sign in

Paper Citation Record · LEDGER

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension

As of 23 August 2026, this Paper Citation Record lists 68 of 68 outbound references and 0 inbound Pith citation observations for arXiv:2508.16300.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.16300 v1

Coverage vector

measured 68 of 68 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T17:28:34.786737Z

measured 68 of 68 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

68 of 68 outbound references displayed

  • verified exact4
  • verified fuzzy28
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch15

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation be4374cc-24fa-4261-a4b3-f665b584041f · outbound

This paper cites Hierarchical interactive multimodal transformer for aspect-based multimodal sentiment analysis.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Hierarchical interactive multimodal transformer for aspect-based multimodal sentiment analysis

Reference 1

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T17:28:40.565758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:26.949273Z digest=sha256:9127826394cd858bb029d248e7ca8a32062d69207333340f690e8dda2e855fe7

Observation 01724b23-5e4f-48b2-bfb9-1e29293230df · outbound

This paper cites Majumder, D.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Majumder, D

Reference 2

Resolution
verified exact
doi, observed 2026-08-05T17:28:35.898728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:26.993330Z digest=sha256:08bc15f1196bd63e77ffaf9490a4c36bd9218e9b9e8e2368cb3270bc0c25c5a0

Observation 5b8ebf5b-e84e-4b1f-84db-d3b5d90355c6 · outbound

This paper cites Modeling high-order relationships: Brain-inspired hypergraph-induced multimodal-multitask framework for semantic comprehension.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Modeling high-order relationships: Brain-inspired hypergraph-induced multimodal-multitask framework for semantic comprehension

Reference 3

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T17:28:40.294291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:27.126036Z digest=sha256:8089836ff088236de1b6cca86d230d2815c2b46c61f1b1b70b4e71ef7d7fd55c

Observation 2ff6181c-4c3a-417a-8a6d-8d775c4c2e06 · outbound

This paper cites an unresolved cited work.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T17:28:27.230588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:28:27.230588Z digest=sha256:a370c760f1f2b2d61d41d9388f8cf8c14f8551fdd369190e6da82afdbea9c1f1

Observation 74e22e8a-b3c9-432f-b547-fd8cb4285ec3 · outbound

This paper cites Nagendra Kumar.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Nagendra Kumar

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T17:28:27.288567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:28:27.288567Z digest=sha256:a527248ef3841cb7bd8491f0153ec91c6e5e06f87c8d1883cec148528335ff2b

Observation 9c41e7e4-f6b1-43ea-bbc3-3ac481da0b45 · outbound

This paper cites Multimodal sentiment analysis: A systematic review of history, datasets, multimodal fusion methods, applications, challenges and future directions.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Multimodal sentiment analysis: A systematic review of history, datasets, multimodal fusion methods, applications, challenges and future directions

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:47.451697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:27.389955Z digest=sha256:72793df659265c00510acedca8874c97f41852f989c20a87114fa74fda8d40ab

Observation fdc2f273-7c2c-4146-9f35-4089eb55e666 · outbound

This paper cites A multitask framework for sentiment, emotion and sarcasm aware cyberbullying detection from multi-modal code-mixed memes.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension A multitask framework for sentiment, emotion and sarcasm aware cyberbullying detection from multi-modal code-mixed memes

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T17:28:27.500562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:28:27.500562Z digest=sha256:22570464c9167501f4fe1a24c66490c37b215dab09c98559d47fdc31f2dd3899

Observation 564ee581-28fc-4ce4-aad0-247f90a9ed90 · outbound

This paper cites Multimodal hate speech detection via multi-scale visual kernels and knowledge distillation architecture.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Multimodal hate speech detection via multi-scale visual kernels and knowledge distillation architecture

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T17:28:27.640635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:28:27.640635Z digest=sha256:e182c30e5ba8fbdda36e474440480327eb456dce6dd4033f5807ada85668b395

Observation 6b388f22-fb20-4904-b832-f340231e067d · outbound

This paper cites an unresolved cited work.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T17:28:27.731597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:28:27.731597Z digest=sha256:3b93b85af02cf42702cc913779d2368d53937bc8d3ebb650dc08805bed6c8118

Observation 04673c59-3f4e-4098-a53c-f2b2a04cd883 · outbound

This paper cites o nig, Eva-Maria Me ner, Alan Cowen, Erik Cambria, and Bj\.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension o nig, Eva-Maria Me ner, Alan Cowen, Erik Cambria, and Bj\

Reference 10

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T17:28:39.756848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:27.842324Z digest=sha256:267c9d5a4e2098600ae324875a4294cf71b90c7317396f86cfa4470e768e67fa

Observation c9819b07-e734-4bb8-9f97-3625e2c1b2ae · outbound

This paper cites Shamim Hossain.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Shamim Hossain

Reference 11

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T17:28:39.463735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:27.892572Z digest=sha256:215130c1b5f06f96c8a8f3a0cbf4d914c08b3d22bbe2b0bb679bfdece611a595

Observation 093ba0b1-bd67-44f1-b294-84be189f5967 · outbound

This paper cites Mahaemosen: Towards emotion-aware multimodal marathi sentiment analysis.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Mahaemosen: Towards emotion-aware multimodal marathi sentiment analysis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T17:28:27.964049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:28:27.964049Z digest=sha256:b09e8411f66e5fa67c06c0e04686f9ad82f0ea434f170fd1c4a2dc4df1d9e917

Observation e390a28a-2958-4c18-8f01-72d7d59db27c · outbound

This paper cites Zico Kolter, Louis-Philippe Morency, and Ruslan Salakhutdinov.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Zico Kolter, Louis-Philippe Morency, and Ruslan Salakhutdinov

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T17:28:28.088813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:28:28.088813Z digest=sha256:5fd66592dfb637892c05e071f2c19ed871bd556e6e6ae81d1ac95cb2bc0b9308

Observation 9a34af92-b10a-4744-8037-7a1e77ebf504 · outbound

This paper cites A hybrid deep neural network for multimodal personalized hashtag recommendation.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension A hybrid deep neural network for multimodal personalized hashtag recommendation

Reference 14

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T17:28:39.235938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:28.169810Z digest=sha256:1ecbb036b45a86a02f838c15175bcac92ba284797fc16e2960566abf67439aa4

Observation b9ac6cfc-fffe-4428-9098-a38e02c2b8a3 · outbound

This paper cites All-but-the-top: Simple and effective post-processing for word representations.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension All-but-the-top: Simple and effective post-processing for word representations

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:47.164159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:28.305696Z digest=sha256:0a9639945bf42c528df776982df41a80df57ba96c982f3b2b18db9683c58229f

Observation 4ad202e1-69c2-4aa5-b4d6-3cfbaa21a905 · outbound

This paper cites Multimodal information bottleneck: Learning minimal sufficient unimodal and multimodal representations.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Multimodal information bottleneck: Learning minimal sufficient unimodal and multimodal representations

Reference 16

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T17:28:38.971053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:28.368540Z digest=sha256:33c3245974598313f6f2eb3b2170eda30ed87b9f67059947015fd9f016896d82

Observation a51bd3b6-74fe-43ee-9d79-ae047248317f · outbound

This paper cites Divide, conquer and combine: Hierarchical feature fusion network with local and global perspectives for multimodal affective computing.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Divide, conquer and combine: Hierarchical feature fusion network with local and global perspectives for multimodal affective computing

Reference 17

Resolution
verified exact
doi, observed 2026-08-05T17:28:35.590373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:28.495499Z digest=sha256:0d65535e4e3718ba0f8cb2ea996fcb1c9559ef588266a55412a2fe35ad8319f7

Observation e3e71fea-3c0a-45b6-8d4b-5df8408bc3da · outbound

This paper cites User-aware multilingual abusive content detection in social media.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension User-aware multilingual abusive content detection in social media

Reference 18

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T17:28:38.736153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:28.595515Z digest=sha256:66def65abb01a7fc4eac110c0cd13fd0c66d15a596fda5f5b4ba0f46fe4c481b

Observation b0cad9ab-b82c-4268-8996-752e941ccdf3 · outbound

This paper cites Hierarchical attention networks for document classification.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Hierarchical attention networks for document classification

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T17:28:28.722586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:28:28.722586Z digest=sha256:619382e1fefaea2f9c07312ee07ca64222714439820e85b3d419b4f3d5b6dcd7

Observation 15466535-ce0f-4e02-8b65-605a62bf0cd0 · outbound

This paper cites Bi-stream graph learning based multimodal fusion for emotion recognition in conversation.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Bi-stream graph learning based multimodal fusion for emotion recognition in conversation

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:46.892516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:28.790796Z digest=sha256:ddd0fecf30d3e38d70b20190f16c6c16aae57794ecf5fdb90ef7697e959ee487

Observation 5d56dde1-2b4e-4c8d-83c4-96691d2d0ecf · outbound

This paper cites Integrating gin-based multimodal feature transformation and multi-feature combination voting for irony-aware cyberbullying detection.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Integrating gin-based multimodal feature transformation and multi-feature combination voting for irony-aware cyberbullying detection

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:46.594510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:28.892860Z digest=sha256:e25d156a08904366d5920b5b2af19d72782eed9128a4c8bc7124e1e3a6e9b312

Observation 397bb00f-3cb9-4c3f-a0a3-2aa32c4ace00 · outbound

This paper cites Gcnet: Graph completion network for incomplete multimodal learning in conversation.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Gcnet: Graph completion network for incomplete multimodal learning in conversation

Reference 22

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T17:28:38.492259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:28.992521Z digest=sha256:95c67c7c1b2ec185a0ed651c188a246f17571d177fa2d588fed659ce72cef87b

Observation 5f94835b-5cfd-4a92-ab14-5fd05b7bed0c · outbound

This paper cites Graphcfc: A directed graph based cross-modal feature complementation approach for multimodal conversational emotion recognition.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Graphcfc: A directed graph based cross-modal feature complementation approach for multimodal conversational emotion recognition

Reference 23

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T17:28:38.184348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:29.109337Z digest=sha256:52c061fdf96379004c363a8f5a8c960df43a1eb2b385034964dd2e1e7a0bc0bb

Observation bd73436d-9c3d-414a-80e7-ab6d9c2c594d · outbound

This paper cites Graph neural network meets sparse representation: Graph sparse neural networks via exclusive group lasso.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Graph neural network meets sparse representation: Graph sparse neural networks via exclusive group lasso

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T17:28:29.192636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:28:29.192636Z digest=sha256:bb702153cf3de8848261476ec8149a002df3cc399ed675796f7f1b00950f1a48

Observation 7447f28d-af05-4d87-b9ec-1603edb7d170 · outbound

This paper cites Heterogeneous graph contrastive learning network for personalized micro-video recommendation.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Heterogeneous graph contrastive learning network for personalized micro-video recommendation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T17:28:29.292326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:28:29.292326Z digest=sha256:779af6a650693e4045268391c1259e15126f3d43ebd37e1a69aa648a46b22c59

Observation eb66893a-d5d6-47d2-9ef6-0afe6b2b62dd · outbound

This paper cites Hgber: Heterogeneous graph neural network with bidirectional encoding representation.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Hgber: Heterogeneous graph neural network with bidirectional encoding representation

Reference 26

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T17:28:37.786366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:29.383851Z digest=sha256:036156e144316df84d7c55a7d33ca77c997567c7b4cc2925a0c0aeaddcecd7b9

Observation 01dc5d7e-53ae-42cc-ac2a-989914579904 · outbound

This paper cites Align before attend: Aligning visual and textual features for multimodal hateful content detection.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Align before attend: Aligning visual and textual features for multimodal hateful content detection

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:46.371425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:29.519439Z digest=sha256:9ea0138ac1e741765fe5b2a0a0d3514a025dd947baae144569c4f1c77e3219c7

Observation b5e205ac-fe4a-46af-ac1b-e5ac96c40e97 · outbound

This paper cites Hatefusion: Harnessing attention-based techniques for enhanced filtering and detection of implicit hate speech.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Hatefusion: Harnessing attention-based techniques for enhanced filtering and detection of implicit hate speech

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:46.087582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:29.625251Z digest=sha256:9d8d42b206324b0b75cdd6b819047f9ebaa2e2d9f3d130ad63967b0d3f70a0f0

Observation 42c02b23-a851-4898-bb29-70ab4a5a8038 · outbound

This paper cites Mimicking the brain’s cognition of sarcasm from multidisciplines for twitter sarcasm detection.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Mimicking the brain’s cognition of sarcasm from multidisciplines for twitter sarcasm detection

Reference 29

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T17:28:37.487407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:29.772895Z digest=sha256:0eca871da947c5332e062e588eaca2debf9977b314d6e19d914fa60938a378d1

Observation 793f3212-3940-44aa-a437-bc5f6752b543 · outbound

This paper cites A multitask learning model for multimodal sarcasm, sentiment and emotion recognition in conversations.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension A multitask learning model for multimodal sarcasm, sentiment and emotion recognition in conversations

Reference 30

Resolution
verified exact
doi, observed 2026-08-05T17:28:35.336745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:29.892863Z digest=sha256:58b933569095446a537805807a8dc06d51d3512175111a411de4928a5edacbcd

Observation 2197dc2d-0034-4754-b5bd-50e66cc7abab · outbound

This paper cites COMMA - DEER : CO mmon-sense aware multimodal multitask approach for detection of emotion and emotional reasoning in conversations.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension COMMA - DEER : CO mmon-sense aware multimodal multitask approach for detection of emotion and emotional reasoning in conversations

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:45.900463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:30.008067Z digest=sha256:e8a3642c5e1a289c1d014b031c62f565ee4266414274ae1415c02e042d25b1f0

Observation 88e4aecc-ccc5-4796-b5f5-c64592f2570c · outbound

This paper cites Progressive modality reinforcement for human multimodal emotion recognition from unaligned multimodal sequences.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Progressive modality reinforcement for human multimodal emotion recognition from unaligned multimodal sequences

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:45.636576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:30.132714Z digest=sha256:7d3f6685bd68ee8e8b5783a1457a5923d8ca10ced6ff7fdf9a791c7a3a1bbb81

Observation e6f17e11-67ba-4e99-b219-ad07607c5b96 · outbound

This paper cites Shah, Surendrabikram Thapa, Usman Naseem, and Mehwish Nasim.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Shah, Surendrabikram Thapa, Usman Naseem, and Mehwish Nasim

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:45.387701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:30.249534Z digest=sha256:076c509e3272f729ce69a631707c64626e0f56774d193b4e6198aaaa53fa198d

Observation 61e04b1d-b2f8-4d16-8a98-526f6ea630d3 · outbound

This paper cites Universal multimodal representation for language understanding.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Universal multimodal representation for language understanding

Reference 34

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T17:28:37.180059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:30.358426Z digest=sha256:eb1e845ffa6359267a81b358568dac40f6c428d13c56ce0576f78e118294f733

Observation e693e27f-a06a-4489-8329-e6685425a2cd · outbound

This paper cites EDA : Easy data augmentation techniques for boosting performance on text classification tasks.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension EDA : Easy data augmentation techniques for boosting performance on text classification tasks

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T17:28:30.524917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:28:30.524917Z digest=sha256:bbaf7a4d766bb8f1dcc0ca213ba466fde8a96a98b458cc38bc5cba9dfff2040e

Observation 8c54ac1f-3ddf-481b-9a05-f3618ce704f4 · outbound

This paper cites Robust training under linguistic adversity.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Robust training under linguistic adversity

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:45.167930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:30.682259Z digest=sha256:cdf9d07844a89aa3bfef4764cfd27061698497f03d54a9a2b5b706442a5523da

Observation aea2f406-9ce3-4cc1-b285-a16f454e9e84 · outbound

This paper cites T iny BERT : Distilling BERT for natural language understanding.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension T iny BERT : Distilling BERT for natural language understanding

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T17:28:30.828519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:28:30.828519Z digest=sha256:bf964430ee23d08813609d8c5b88ac9ba997e99b4172da0dddf6c83c11bf920c

Observation 9116f451-3aeb-4739-83d0-d65b5a8e839b · outbound

This paper cites Transformer-based multimodal emotional perception for dynamic facial expression recognition in the wild.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Transformer-based multimodal emotional perception for dynamic facial expression recognition in the wild

Reference 38

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T17:28:36.855785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:30.952663Z digest=sha256:9fdb017d0ebb28ee21b498ee776150efbf8373656ec9881e9783c2c0c94e8eac

Observation 1f45d571-392c-43e1-b9da-9ceb9ac4cdd9 · outbound

This paper cites Deep residual learning for image recognition.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Deep residual learning for image recognition

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:44.900742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:31.040010Z digest=sha256:ef3b54c71fcd2112835f216fadf0bb1177cfa705d0d07c47582decabcbd25753

Observation 4b676f98-061c-4b2c-8486-250799fcdfcd · outbound

This paper cites Random erasing data augmentation.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Random erasing data augmentation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T17:28:31.132184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:28:31.132184Z digest=sha256:0b155728a7cb1ee21fb2bae5d4138dd3aafb56ca1cdd07bbe71b29b183f56553

Observation 006e3eb3-4c7b-4649-a831-136320432591 · outbound

This paper cites Free-form image inpainting with gated convolution.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Free-form image inpainting with gated convolution

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:44.668008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:31.244350Z digest=sha256:9a7ac2638b147b38154103a3de3ba22b943b501bbc0094172840899a68926783

Observation 40e76ab8-462a-4d33-810d-6db957a2d64f · outbound

This paper cites Learning transferable visual models from natural language supervision.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Learning transferable visual models from natural language supervision

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T17:28:31.359124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:28:31.359124Z digest=sha256:9ee7907caf9c31bdb6fa90fb10b3fc0b717d01fd1bbfe29302b28c85f39a303e

Observation a777b1fe-7d73-4889-8db7-56a2bcf60df1 · outbound

This paper cites Bert: Pre-training of deep bidirectional transformers for language understanding.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Bert: Pre-training of deep bidirectional transformers for language understanding

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:44.422577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:31.456991Z digest=sha256:a38267c12b5853e7218c17677e0c493d9e8bdb3b11bfc14d1a12b223f16bdaee

Observation de93674d-f4ac-4729-9e74-b149727fee6f · outbound

This paper cites Detectron2.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Detectron2

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T17:28:31.550703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:28:31.550703Z digest=sha256:0ac643f8f914db269b05d799e42b6cdd4245e621102b04c99ce9e838b06a1edc

Observation 3c6d6e6e-2610-4bcf-85a6-2df4a34f21c0 · outbound

This paper cites Mask r-cnn.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Mask r-cnn

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-05T17:28:31.637706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:28:31.637706Z digest=sha256:8a3e379dcd4d5dc03af8ce99416e9e644c3c3f2910c43633d8af660d600135cf

Observation b8c4e34e-e8e3-4234-8c0c-c35feee98106 · outbound

This paper cites The stanford corenlp natural language processing toolkit.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension The stanford corenlp natural language processing toolkit

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:44.202736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:31.716737Z digest=sha256:93442a01d630b6b622e2d6cef198bd5cbc9586325f99f903f91beb2ed05bdf6c

Observation 9e63d700-78d9-4772-ab27-9a0fff2fced9 · outbound

This paper cites Hierarchical attention-enhanced contextual capsulenet for multilingual hope speech detection.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Hierarchical attention-enhanced contextual capsulenet for multilingual hope speech detection

Reference 47

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T17:28:36.552257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:31.862851Z digest=sha256:9bfb284c04782912388b92fbcf3b947505f2865621e73870f5f4e9aa538dcf03

Observation 9b2f99c2-2afe-4ce5-9c3b-69ec97fc0369 · outbound

This paper cites A context-aware attention and graph neural network-based multimodal framework for misogyny detection.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension A context-aware attention and graph neural network-based multimodal framework for misogyny detection

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T17:28:31.964267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:28:31.964267Z digest=sha256:cb418854033e73e38b528c9b70e6f6372bb27d07696e6aba761e95d394f9b55c

Observation fb12d003-3d0d-4dab-852e-49612cf4f777 · outbound

This paper cites Inductive representation learning on large graphs.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Inductive representation learning on large graphs

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T17:28:32.113447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:28:32.113447Z digest=sha256:878a7bd8e0f5e13456b5a8e79b537f8c8f9bc2578730f34a1f527149ad2e2194

Observation ae714211-872c-4b3e-8040-7329cbed2cf2 · outbound

This paper cites o rn Gamb \.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension o rn Gamb \

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:43.941268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:32.268171Z digest=sha256:488d7de21ff37525689c275ffffd7f9459b5811da869938b9298b1f70fddf554

Observation 3447a8f4-61d4-4d8e-9fd3-8258dfcc1d2b · outbound

This paper cites Exploring hate speech detection in multimodal publications.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Exploring hate speech detection in multimodal publications

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:43.661398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:32.416931Z digest=sha256:79d25f2b2d2d3da7ac7e015d60780652f82a006dea3ec86070b6a9fcb24bc0bb

Observation 243c8dd2-6692-429c-a455-2cad8b39a4a6 · outbound

This paper cites Detecting harmful memes and their targets.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Detecting harmful memes and their targets

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:43.401537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:32.520722Z digest=sha256:043cea6fef655cc5d044d5b40ca8cf0b49b1a78c4686c3d58a9ab02579b09535

Observation 27d84f48-4f76-4744-b16c-c8e907f161d1 · outbound

This paper cites Multimodal sentiment analysis: A systematic review of history, datasets, multimodal fusion methods, applications, challenges and future directions.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Multimodal sentiment analysis: A systematic review of history, datasets, multimodal fusion methods, applications, challenges and future directions

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:43.183004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:32.651914Z digest=sha256:4870922a03144d9bd3af2c9e7e0ff6a241b6e29aea61f396deca967dff1b4aec

Observation 8f1bfae2-8af9-417f-853c-830b7cb11493 · outbound

This paper cites Multi-modality cross attention network for image and sentence matching.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Multi-modality cross attention network for image and sentence matching

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:42.869144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:32.796395Z digest=sha256:f3b4ab2603a442d711c614251630e58635b408b384b7dd22aa74b03814378c86

Observation ed5ab989-8189-40e5-a369-cf37a57977c9 · outbound

This paper cites Attention is all you need.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Attention is all you need

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-05T17:28:32.915608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:28:32.915608Z digest=sha256:5a1dbf418e71159cbeca31ea6112d72814bf653655aaff77faa17d68c088df3a

Observation 88ffb1d7-f4e9-4640-821a-4785e78d6f0f · outbound

This paper cites VisualBERT: A Simple and Performant Baseline for Vision and Language.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension VisualBERT: A Simple and Performant Baseline for Vision and Language

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-05T17:28:33.109227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:28:33.109227Z digest=sha256:2af39dd4cf878b218c0d70b2b7b1459c802c5c09d1a6f37c3d5cdb913db416a7

Observation 056351ba-7473-42a3-92fa-3ba63aec023b · outbound

This paper cites Vilt: Vision-and-language transformer without convolution or region supervision.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Vilt: Vision-and-language transformer without convolution or region supervision

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:42.672020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:33.261714Z digest=sha256:22e231f7dbaed871ef9bb64babb2b0b141d35899a2d6b1e5bd98f3880961bbe2

Observation 0dc46a00-bb76-49eb-b407-f962ea666404 · outbound

This paper cites Disentangling hate in online memes.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Disentangling hate in online memes

Reference 58

Resolution
metadata mismatch
raw_fallback, observed 2026-08-05T17:28:36.228564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:33.368405Z digest=sha256:fdd93a12962ea7afd20ab708f3bb5f5f3e8d5ef8b67dc89141591455a86f7574

Observation 961ee4b8-dff6-4586-9bee-ac24a8b4d502 · outbound

This paper cites Momenta: A multimodal framework for detecting harmful memes and their targets.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Momenta: A multimodal framework for detecting harmful memes and their targets

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:42.436263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:33.496937Z digest=sha256:408e0378430a2c237fa6ca8b4e36a6541cc1079f2ee0d9eee93790a1a1c7da7c

Observation cbcd9a0c-bf1a-48db-b8d9-dbd927d7a163 · outbound

This paper cites Beneath the surface: Unveiling harmful memes with multimodal reasoning distilled from large language models.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Beneath the surface: Unveiling harmful memes with multimodal reasoning distilled from large language models

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:42.228048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:33.624440Z digest=sha256:ab24d28993b6b3fbbcc0bd4ff1a35086d5ef4c1ca10962376f58415cca0af64a

Observation 53d8dcfe-1998-4ec5-a39b-1857f24fbf0f · outbound

This paper cites Multi-interactive memory network for aspect based multimodal sentiment analysis.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Multi-interactive memory network for aspect based multimodal sentiment analysis

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:41.999315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:33.768791Z digest=sha256:a6b47ab3875de5f25fd8cb32fededce85edbae7cbd937ef79fbf4d25955f3241

Observation 8a71dabe-df25-4b50-b4ea-3c145ac6ae4c · outbound

This paper cites Modeling intra and inter-modality incongruity for multi-modal sarcasm detection.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Modeling intra and inter-modality incongruity for multi-modal sarcasm detection

Reference 62

Resolution
verified exact
doi, observed 2026-08-05T17:28:35.025480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:33.878344Z digest=sha256:9109b6963f57cc3fb3170c61e7629fba60bacb350c6a9b8213452e85811e51d2

Observation da95eacf-7fbd-40e8-bdf8-d736c56eab24 · outbound

This paper cites Ensemble pretrained models for multimodal sentiment analysis using textual and video data fusion.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Ensemble pretrained models for multimodal sentiment analysis using textual and video data fusion

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:41.729481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:34.051871Z digest=sha256:3175127c22f74d960ef494a71255dd757bd49ea4150bc69660ae60220b4658eb

Observation e9240ae4-57b6-43b0-b582-fe046e9c972c · outbound

This paper cites O zlem \.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension O zlem \

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:41.466584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:34.257495Z digest=sha256:0c22815a063210ffc787c1c06afe0b7f29301ff3fa815ea0df3b123393133ba1

Observation 5e436228-f8f7-4fdc-b625-efac1e429b9e · outbound

This paper cites Mmffhs: Multi-modal feature fusion for hate speech detection on social media.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Mmffhs: Multi-modal feature fusion for hate speech detection on social media

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:41.221052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:34.364474Z digest=sha256:d6bce68574020077c3ed1c98389226cdec71cefe19bd8822ec48e5eb26a9d5c9

Observation 62e03b64-6555-4be9-ae61-239fc548cb0f · outbound

This paper cites Emotion-aware multimodal fusion for meme emotion detection.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension Emotion-aware multimodal fusion for meme emotion detection

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:40.998154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:34.487110Z digest=sha256:5f656ef9f679ce5fc6368578809347b00526d017299eac2653535f49ad3cb6a4

Observation 54cc7061-679e-449a-8c51-5a14fea35ca1 · outbound

This paper cites C hat G P T 4o.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension C hat G P T 4o

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T17:28:40.774444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-05T17:28:34.638087Z digest=sha256:4b6b37e16fccd3c615857777a3293149a8c407baab6c913645f0f993fd4e827b

Observation 1796b688-55c7-46c6-b600-167e6b0e199c · outbound

This paper cites write newline.

A Multimodal-Multitask Framework with Cross-modal Relation and Hierarchical Interactive Attention for Semantic Comprehension write newline

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-05T17:28:34.786737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:28:34.786737Z digest=sha256:6708d907de6730bfe0503095c000a1e82c622f67317817caa4b2b9da8e8d7fbb

Pith citing papers

No inbound Pith citation observations are available.