Pith. sign in

Paper Citation Record · LEDGER

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion

As of 17 August 2026, this Paper Citation Record lists 27 of 27 outbound references and 2 inbound Pith citation observations for arXiv:2509.17446.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.17446 v4

Coverage vector

measured 27 of 27 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:51:46.996181Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:51:46.891087Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-08T00:49:09.573510Z

Reference resolution

27 of 27 outbound references displayed

  • verified exact1
  • verified fuzzy17
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f528d143-3c65-4591-bbea-a709f30617cc · outbound

This paper cites MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:46.891087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:46.891087Z digest=sha256:94d6bad2d0dc2ef3e9da269dad6e2169bcc466b692923a7c91bf2528170caa8c

Observation 5567a18e-07cc-4aba-ace6-2e1514b45ff2 · outbound

This paper cites Model Overview As illustrated in Fig.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Model Overview As illustrated in Fig

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.546471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:46.896312Z digest=sha256:b8e2c2e2f0a7641738cd2058300842441a67cf6f668b5e36dd2b1836835f5664

Observation 90a68bb4-30f6-4c65-8d03-33eea146f003 · outbound

This paper cites an unresolved cited work.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Unresolved cited work

Reference 3

Resolution
verified exact
raw_fallback, observed 2026-08-15T15:51:47.273979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:46.901222Z digest=sha256:739dc24d8bfd0ff382b6b39cc92f3095e1b5589baf450efc5dbf4fa1d2deb074

Observation b0f42847-ed54-4637-a397-574b2270c050 · outbound

This paper cites an unresolved cited work.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:51:47.532950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:46.906048Z digest=sha256:30b4ef40e279061da494042d4a142835eded148b27071a1edc4633e7b37b7e0d

Observation 2d4d5bcf-f8a5-4021-bbde-ae522dfce194 · outbound

This paper cites an unresolved cited work.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:51:47.519408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:46.910514Z digest=sha256:2af88eb941b598c7da22cea87db5ebc0d3c0601f7df677553886fb0761f9d4ec

Observation e48c8a52-375c-4de8-9d13-b36269182ee9 · outbound

This paper cites Deep Learning Approaches for Multimodal Intent Recognition: A Survey.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Deep Learning Approaches for Multimodal Intent Recognition: A Survey

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:46.914985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:46.914985Z digest=sha256:f653670253cda38a26f461b5061fe97945415c2582bf84443cc8f037fa71dc66

Observation bc6046f5-a54f-421d-9d01-f2a1a218b18c · outbound

This paper cites Temporal working memory: Query- guided segment refinement for enhanced multimodal understanding,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Temporal working memory: Query- guided segment refinement for enhanced multimodal understanding,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.506674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:46.919561Z digest=sha256:6a7df00a147859c707048ed10785d6a285c7bb4d54742cf146391ca7d5aa20e7

Observation 2b22f4fc-0697-4edd-a33c-aa678ca08fa9 · outbound

This paper cites To- wards visual-prompt temporal answer grounding in in- structional video,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion To- wards visual-prompt temporal answer grounding in in- structional video,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.494272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:46.923324Z digest=sha256:56e1303a2facdcbd5ff7a8e0504fe4eaf59685876b3a40032c77fe4f2395e1ac

Observation f4672e7e-dde4-42b1-9222-ebcebf606a84 · outbound

This paper cites Visual document understand- ing and question answering: A multi-agent collabora- tion framework with test-time scaling,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Visual document understand- ing and question answering: A multi-agent collabora- tion framework with test-time scaling,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:46.926883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:46.926883Z digest=sha256:40e30d31310e47aa4c86cc47ac8c49856cb590035a229ab27eae805b19a6b977

Observation de325dd4-b4e6-40af-b9dd-ddb49be69200 · outbound

This paper cites Multimodal transformer for un- aligned multimodal language sequences,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Multimodal transformer for un- aligned multimodal language sequences,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.480557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:46.930606Z digest=sha256:44cdbab14df502b65dcfaf50ceb9df7b4d699fdf342e209c2b38868bd677aa51

Observation 5b2d2083-93f1-4545-9bcc-b0812502239d · outbound

This paper cites Adaptive mul- timodal fusion: Dynamic attention allocation for intent recognition,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Adaptive mul- timodal fusion: Dynamic attention allocation for intent recognition,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.467553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:46.934791Z digest=sha256:62f692a87f9a9ca13830469f624cc21c4bbc2f84f070b74b8c9463e73d273dd7

Observation 6000014f-e138-4efc-b697-2d60ff710782 · outbound

This paper cites Multimodal transformer with multi-scale alignment for multimodal sentiment analysis,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Multimodal transformer with multi-scale alignment for multimodal sentiment analysis,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.454648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:46.938626Z digest=sha256:fcc26107e74573d3a56577b13ca1c8d1ea170df34f941fb3f615ab716ba7e325

Observation 7da4ab5b-a2fa-4629-8a4b-a046ceb2c561 · outbound

This paper cites Improving Multimodal Fusion with Hierarchical Mutual Information Maximization for Multimodal Sentiment Analysis.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Improving Multimodal Fusion with Hierarchical Mutual Information Maximization for Multimodal Sentiment Analysis

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:46.942355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:46.942355Z digest=sha256:a98df8e4bab5ffd454796233b72a7cbd2b33fe166f97c00ae8b11e6b0dc654de

Observation 0ddd00f2-d4fa-4101-9bce-1ed40ecc9cc9 · outbound

This paper cites Dynamic multimodal fu- sion,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Dynamic multimodal fu- sion,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.440593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:46.946140Z digest=sha256:891d393715ed6fee538df29f5554ba3080b79ef2980eae6e5d43338cc786716a

Observation 725afab5-e854-41c6-a824-0e55ab6f22e1 · outbound

This paper cites Token-level contrastive learning with modality-aware prompting for multimodal intent recog- nition,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Token-level contrastive learning with modality-aware prompting for multimodal intent recog- nition,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.427623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:46.949590Z digest=sha256:d67de521f789076f57657ef7bf16af8c5bef4cca6773aaa829ee853ea55351f3

Observation f6560aa6-3759-4685-af4e-8dd0cb368bbd · outbound

This paper cites Factorized Contrastive Learning: Going Beyond Multi-view Redundancy.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Factorized Contrastive Learning: Going Beyond Multi-view Redundancy

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:46.953000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:46.953000Z digest=sha256:932bbd85d46af7e256ec7506fe7877dd3dfdbb7dbeb0b055fa4bea77979f629c

Observation 1df649bb-1446-4120-9b21-e40bca68dbd0 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Representation Learning with Contrastive Predictive Coding

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:46.957027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:46.957027Z digest=sha256:dc3b4f120be0cd9b9356305570f5a611938ea08477aeb4988da2bdbe218ca80d

Observation 6b1392ee-69b7-4f76-9a83-7fa95d545827 · outbound

This paper cites Mag-bert: Multimodal adapta- tion gate bert for multimodal sentiment analysis,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Mag-bert: Multimodal adapta- tion gate bert for multimodal sentiment analysis,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.415072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:46.960653Z digest=sha256:e1424d48e36dbd78bbfbaf8ef67308d4502438dc1c4c0f9d7341e468d349e48e

Observation d6829b96-a73a-483e-b91e-3d4e67226f81 · outbound

This paper cites Super- vised contrastive learning,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Super- vised contrastive learning,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.402249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:46.964160Z digest=sha256:6270f3062b88a8d864b46124c61e28ef2e5d2600df2130fafb672c0dcb3e87cb

Observation 21c51923-216d-4caf-8526-ffb2664bc9f1 · outbound

This paper cites Contextual augmented global contrast for multimodal intent recog- nition,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Contextual augmented global contrast for multimodal intent recog- nition,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.389489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:46.968055Z digest=sha256:60e41f8897c20f75e608390d14c5f560cadbc8d5d9ba1bfbae6d1ae1dfc3384b

Observation 4bc8d6e4-68b3-439c-acc2-7c90437a8740 · outbound

This paper cites Prototypical net- works for few-shot learning,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Prototypical net- works for few-shot learning,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.375484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:46.971689Z digest=sha256:30e1abd41c3b4d3328c5bbea5c3e895d4fac3aa3b2936752014775b11e154fc2

Observation 784331fc-9e48-4d4a-997b-281675232a6b · outbound

This paper cites Mintrec: A new dataset for multimodal intent recognition,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Mintrec: A new dataset for multimodal intent recognition,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.358934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:46.975806Z digest=sha256:f7cbb4b809f36f834253a82b10e57d8b16750577008162a1533a0d182ce5e99b

Observation 933be9cd-5da7-4c99-a6c3-6a247770e382 · outbound

This paper cites Mintrec2.0: A large-scale benchmark dataset for multimodal intent recognition and out-of- scope detection in conversations,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Mintrec2.0: A large-scale benchmark dataset for multimodal intent recognition and out-of- scope detection in conversations,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.345706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:46.980139Z digest=sha256:ea2c22bd914ac5a685e31218e48cc840985b4f4482c49245e549e6d461be6b17

Observation 03986d9a-d258-4798-b590-7c29da0d16be · outbound

This paper cites Understanding con- trastive representation learning through alignment and uniformity on the hypersphere,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Understanding con- trastive representation learning through alignment and uniformity on the hypersphere,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.332076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:46.984268Z digest=sha256:214826449e04d3aaeedc11031ecf5bf7cdd0f9c68c6a681c0f70414fd8818a7b

Observation d07c04cb-57af-4065-ab7c-fe834c1ae367 · outbound

This paper cites Can contrastive learning avoid short- cut solutions?,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Can contrastive learning avoid short- cut solutions?,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.318269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:46.988416Z digest=sha256:7ac787c1d96662ce9ab25dfb7941c8968db0dacf2fc00bc59d1406b2a81d5835

Observation 82b0a686-7890-495e-8ac5-d004893d7a58 · outbound

This paper cites HiCLIP: Contrastive Language-Image Pretraining with Hierarchy-aware Attention.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion HiCLIP: Contrastive Language-Image Pretraining with Hierarchy-aware Attention

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:46.992396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:46.992396Z digest=sha256:993215a64712608a1e27716147dae14dcdf80d7c38db171a41ee77e65d5958ec

Observation ec593fda-efe6-4013-920b-613fda3256ba · outbound

This paper cites Attention bottlenecks for multimodal fusion,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Attention bottlenecks for multimodal fusion,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.304860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T15:51:46.996181Z digest=sha256:d881f9cb3c6793ed6a0b702efce988e04ebc709cca7f07e2f6ec2de2896ec627

Pith citing papers

Observation f528d143-3c65-4591-bbea-a709f30617cc · inbound

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion cites this paper.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:46.891087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:46.891087Z digest=sha256:94d6bad2d0dc2ef3e9da269dad6e2169bcc466b692923a7c91bf2528170caa8c

Observation b3106e20-e87d-4058-8b4d-fc168df99611 · inbound

Modality Agreement- and Conflict-Aware Prototype Hypergraph Learning for Multimodal Intent Understanding cites this paper.

Modality Agreement- and Conflict-Aware Prototype Hypergraph Learning for Multimodal Intent Understanding MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-08T00:49:09.580828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-08T00:49:09.503196Z digest=sha256:e9bdc9af0b07c6f735a866b27c7448d68f53e5af81d7cd8edc52c56ea2460a54