Pith. sign in

Paper Citation Record · LEDGER

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge

As of 18 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 0 inbound Pith citation observations for arXiv:2505.06814.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.06814 v1

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T22:35:59.141011Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

32 of 32 outbound references displayed

  • verified exact1
  • verified fuzzy19
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1c57d5a9-210a-4431-afbf-c0a191e80ee0 · outbound

This paper cites Overview of the nlpcc 2023 shared task: Chinese medical instructional video question answering.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Overview of the nlpcc 2023 shared task: Chinese medical instructional video question answering

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.758298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.022150Z digest=sha256:9005089700b84ffaadb424cc51f1e12c225bec1878f063652d214024a13da55a

Observation 758a6f30-b67a-4025-be8d-c552ed6d11e5 · outbound

This paper cites Overview of the nlpcc 2024 shared task 7: Multi-lingual medical instructional video question answering.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Overview of the nlpcc 2024 shared task 7: Multi-lingual medical instructional video question answering

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.747241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.026946Z digest=sha256:a8f425b1799ee8eeaa6011eaa5c9306a1b7f05d076a5c0b53601e4d41db28491

Observation 1ec040bd-3dfe-41fb-81fa-41ad91841e34 · outbound

This paper cites A systematic review of machine learning applications in infectious disease prediction, diagnosis, and outbreak fore- casting.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge A systematic review of machine learning applications in infectious disease prediction, diagnosis, and outbreak fore- casting

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.735619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.030812Z digest=sha256:2f2631d3e6ae95168c110d4bc133e244476b1d5c2732f49e2b7c3883501f06dc

Observation 0a546ac8-8eee-48a1-8fff-9b5cd3417b1b · outbound

This paper cites Tf-icon: Diffusion-based training-free cross-domain image composition.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Tf-icon: Diffusion-based training-free cross-domain image composition

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.724868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.034940Z digest=sha256:5324ae402ab10c377a77eb5046f7ae5999d0a05a5046e2cabcd9c0b71fe3272c

Observation 95db3e03-0c2f-4493-8de1-2aab886ad0c7 · outbound

This paper cites Ich-prnet: a cross-modal intracerebral haemorrhage prognostic prediction method using joint-attention in- teraction mechanism.Neural Networks, 184:107096, 2025.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Ich-prnet: a cross-modal intracerebral haemorrhage prognostic prediction method using joint-attention in- teraction mechanism.Neural Networks, 184:107096, 2025

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.713670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.038753Z digest=sha256:4b4739d7cd3da177782b04b1f20e6431d382ed1d7c196f2538ef3e58c5c512b1

Observation 966b5fce-d1e8-48f7-94c8-36a309855057 · outbound

This paper cites Ich-scnet: Intracerebral hemorrhage segmentation and prognosis classification network using clip-guided sam mecha- nism.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Ich-scnet: Intracerebral hemorrhage segmentation and prognosis classification network using clip-guided sam mecha- nism

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.699572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.042707Z digest=sha256:af5919fff09fb3c629c4879c28ae5f62dcf0667347ad4583f4c9502560ade142

Observation 529b38af-485b-449e-8eb8-df14adbd592d · outbound

This paper cites Towards nation-wide analytical healthcare infrastructures: A privacy-preserving augmented knee rehabilitation case study.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Towards nation-wide analytical healthcare infrastructures: A privacy-preserving augmented knee rehabilitation case study

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.046732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.046732Z digest=sha256:3a38e2fa4bb54919158ce5fb27380d38e5ee94ef27099a04eb4eff54339e74e8

Observation 7479754d-4183-4d75-8184-07cae77cb536 · outbound

This paper cites Learning to Unify Audio, Visual and Text for Audio-Enhanced Multilingual Visual Answer Localization.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Learning to Unify Audio, Visual and Text for Audio-Enhanced Multilingual Visual Answer Localization

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.050670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.050670Z digest=sha256:74d801c989bdb943b86440ede2503a24b8fdccc8d6ecff3e8d449fea0b6cea68

Observation 1391dcac-a8f6-49ae-8eea-57d2cdf5a872 · outbound

This paper cites Prism: Self-pruning intrinsic selection method for training-free multi- modal data selection, 2025.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Prism: Self-pruning intrinsic selection method for training-free multi- modal data selection, 2025

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.688250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.054396Z digest=sha256:bd6b0f528b34f99a024ddcfbca097878af56c37f0cca22ae810d6bc93e4aa78b

Observation e3abe435-08eb-4872-b31f-5df9a368e361 · outbound

This paper cites Robust Watermarking Using Generative Priors Against Image Editing: From Benchmarking to Advances.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Robust Watermarking Using Generative Priors Against Image Editing: From Benchmarking to Advances

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.057796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.057796Z digest=sha256:a1478c1a20ca1136744d10594b144f2e14431dfa662c5b61810527d5c23a3e07

Observation d6dd0299-1b94-4dc0-8017-ff7b32d66fd3 · outbound

This paper cites Set You Straight: Auto-Steering Denoising Trajectories to Sidestep Unwanted Concepts.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Set You Straight: Auto-Steering Denoising Trajectories to Sidestep Unwanted Concepts

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.061770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.061770Z digest=sha256:5403f60ea82039d17f21705604190159db043a06f55a337773123a3e85b61f1f

Observation fd829d8b-1740-4d6b-abbe-bace547192c5 · outbound

This paper cites LingYi: Medical Conversational Question Answering System based on Multi-modal Knowledge Graphs.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge LingYi: Medical Conversational Question Answering System based on Multi-modal Knowledge Graphs

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-15T22:35:59.447503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.066362Z digest=sha256:e52f959ae5ac39969721787c6fcd17d4da15cce83ae87db88743f9b2df636abe

Observation dfd89b6e-3d66-49a3-8568-501979759302 · outbound

This paper cites Llava steering: Visual instruction tuning with 500x fewer parameters through modality linear representation-steering, 2025.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Llava steering: Visual instruction tuning with 500x fewer parameters through modality linear representation-steering, 2025

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.676218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.070339Z digest=sha256:40e5f7bb6f5f0a9ecb265cd9b7d938b32073e4ab1b21864853a584c955f47fca

Observation de4d2db6-26ba-49f5-a727-64c64760e57b · outbound

This paper cites Visual answer localization with cross-modal mutual knowledge transfer.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Visual answer localization with cross-modal mutual knowledge transfer

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.663265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.073882Z digest=sha256:3ad0969759bea7c7e0695e7f2398476de6713e3f7019bd6a0d74a698bb4bb46b

Observation 76ecb771-b6a3-4096-9fe8-2b86cdc612c4 · outbound

This paper cites Enhancing thyroid disease prediction using ma- chine learning: A comparative study of ensemble models and class balancing tech- niques.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Enhancing thyroid disease prediction using ma- chine learning: A comparative study of ensemble models and class balancing tech- niques

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.650841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.077678Z digest=sha256:5d19d4ccad91adc110d95b9ab7ad89f00db003131831ba77bfeefcb5ccf0dc0e

Observation e16aaf9b-4a09-4897-a875-e1fcc5435eb5 · outbound

This paper cites A generative adversarial network-based investor sentiment in- dicator: Superior predictability for the stock market.Mathematics, 13(9):1476, 2025.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge A generative adversarial network-based investor sentiment in- dicator: Superior predictability for the stock market.Mathematics, 13(9):1476, 2025

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.638327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.081388Z digest=sha256:6897a3d91b5001079870e4b21be1a8cbc66d695136167706207b81e5d8638a12

Observation c8055437-5f2e-4fee-9844-4432911f1809 · outbound

This paper cites Multimodal high- order relationship inference network for fashion compatibility modeling in internet of multimedia things.IEEE Internet of Things Journal, 11(1):353–365, 2024.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Multimodal high- order relationship inference network for fashion compatibility modeling in internet of multimedia things.IEEE Internet of Things Journal, 11(1):353–365, 2024

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.626294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.085164Z digest=sha256:c7cd3fa9480ec9ab1a0512fffea5e15df99db174b9024b5b6da68e0540bd1368

Observation c811c3bf-269a-451c-9600-d48be8e12205 · outbound

This paper cites De- tection of ai deepfake and fraud in online payments using gan-based models.arXiv preprint arXiv:2501.07033, 2025.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge De- tection of ai deepfake and fraud in online payments using gan-based models.arXiv preprint arXiv:2501.07033, 2025

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.088942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.088942Z digest=sha256:924ba7b7a52c65c33e2717af495ff292a5bb0341ebbc6c0f30d0ccbe728edce1

Observation fa1f9d64-991a-4edc-a3cc-ec90b10a80b0 · outbound

This paper cites Enhancing intent understanding for ambiguous prompts through human-machine co-adaptation.arXiv preprint arXiv:2501.15167, 2025.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Enhancing intent understanding for ambiguous prompts through human-machine co-adaptation.arXiv preprint arXiv:2501.15167, 2025

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.092542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.092542Z digest=sha256:daa5c61c5fcd11cbe15e690a93397c2593df036df433d3641028bb81ac8412c9

Observation e0228f52-1048-4cd5-8f3b-f16524a5ec4e · outbound

This paper cites Mace: Mass concept erasure in diffusion models.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Mace: Mass concept erasure in diffusion models

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.614545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.096530Z digest=sha256:1d4dd44b02effca239ef20c3651b6bcb569a3d5daac43645d0405f44940ab38b

Observation bebc369c-b979-4d3a-ba0e-d7bd624b5273 · outbound

This paper cites Score: Story coherence and retrieval enhancement for ai narratives.arXiv preprint arXiv:2503.23512, 2025.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Score: Story coherence and retrieval enhancement for ai narratives.arXiv preprint arXiv:2503.23512, 2025

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.100140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.100140Z digest=sha256:86785ca821f5274193fe7bff34ef52ea08e3fb30418274f5996fcd11fa290f79

Observation fc1cfef9-cc4d-4a33-8afd-ea729d6d5e98 · outbound

This paper cites Learning to locate visual answer in video corpus using question.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Learning to locate visual answer in video corpus using question

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.602358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.103699Z digest=sha256:010c5dd213a4b9e573810647fba74888aa2a873f949ffe70969235d46651f0a6

Observation 7293766e-9a1c-4814-8131-12dc5b30f0ce · outbound

This paper cites Improving multilingual temporal answering grounding in single video via llm-based translation and ocr enhancement.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Improving multilingual temporal answering grounding in single video via llm-based translation and ocr enhancement

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.590814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.107376Z digest=sha256:0bdffa7c456d35d97d367d1ba56054876b3cb7b820286dcc71aceed0acff6ed2

Observation 43e49dc4-b3b9-47b6-b4d7-ab929f706fd0 · outbound

This paper cites Improv- ing cross-modal visual answer localization in chinese medical instructional video using language prompts.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Improv- ing cross-modal visual answer localization in chinese medical instructional video using language prompts

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.577486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.111232Z digest=sha256:c13e2f3e512e6c8ce6d6c9cca38104717eecd3f4fe227e56e5096720715ce09f

Observation c0fa08cd-de37-4a06-878c-37d0b2f5ba7f · outbound

This paper cites Mqua: Multi-level query-video augmentation for multilingual video corpus retrieval.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Mqua: Multi-level query-video augmentation for multilingual video corpus retrieval

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.114991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.114991Z digest=sha256:d3af855972900bc1b5e48120aab50994e9a6c47d8d4321b4103fddfd591054ba

Observation 883cf69e-05ae-46b5-87a7-4fbd6afe1f1a · outbound

This paper cites A two-stage chinese medical video retrieval framework with llm.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge A two-stage chinese medical video retrieval framework with llm

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.118451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.118451Z digest=sha256:376463c532e8c86ddda00078320227b6686dfb75843a958ad6d27d83202735d0

Observation e9606ac6-6f53-4326-8a77-c80581d56911 · outbound

This paper cites Mul- tilingual temporal answer grounding in video corpus with enhanced visual-textual integration.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Mul- tilingual temporal answer grounding in video corpus with enhanced visual-textual integration

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.121971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.121971Z digest=sha256:06f08f16af1b76d4fc3656580d33fc54a8e24c9add7b46a8753bd7cee26c5e1f

Observation bb847df7-720d-48d3-a765-bc12eac5f669 · outbound

This paper cites A uni- fied framework for optimizing video corpus retrieval and temporal answer ground- ing: fine-grained modality alignment and local-global optimization.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge A uni- fied framework for optimizing video corpus retrieval and temporal answer ground- ing: fine-grained modality alignment and local-global optimization

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.125926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.125926Z digest=sha256:0cf358659eb54e02ffebd4a482c5874918db20e4c659e73ca9f36e055904a841

Observation accf6a10-c5ee-4d6a-83d1-11f3fb446d26 · outbound

This paper cites Correlation-aware cross-modal attention net- work for fashion compatibility modeling in ugc systems.ACM Transactions on Multimedia Computing, Communications and Applications, 2024.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Correlation-aware cross-modal attention net- work for fashion compatibility modeling in ugc systems.ACM Transactions on Multimedia Computing, Communications and Applications, 2024

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.535640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.129694Z digest=sha256:7e13cd2dea7b90d3d6f72c36c902e807223a5680be5765982620fcde87c2bd64

Observation d20edbe7-de37-424a-9a99-ef5de3e73c9e · outbound

This paper cites Language-agnostic BERT Sentence Embedding.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Language-agnostic BERT Sentence Embedding

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.133119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.133119Z digest=sha256:3c01bb5a4647f9224edd12a9b2a9682801b46a52a9c8512671deaf7232a67ac2

Observation fcea76bc-44ba-48eb-8632-e924b1305c27 · outbound

This paper cites Vpai lab at medvidqa 2022: a two-stage cross-modal fusion method for medical instructional video classi- fication.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Vpai lab at medvidqa 2022: a two-stage cross-modal fusion method for medical instructional video classi- fication

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.524131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.137319Z digest=sha256:5685c9e9a0d1853cac62fb38b830f62e180b9f4ac2da5db24a00ee6d4cb0fbee

Observation f198ac09-494c-44a9-926b-43bef7043ebf · outbound

This paper cites Category-aware multimodal attention network for fashion compatibility modeling.IEEE Transac- tions on Multimedia, 25:9120–9131, 2023.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Category-aware multimodal attention network for fashion compatibility modeling.IEEE Transac- tions on Multimedia, 25:9120–9131, 2023

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.512412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.141011Z digest=sha256:c77cdebe9363f8f43f18b69f1f3801922e12ba38f836512538f01f0979cdde17

Pith citing papers

No inbound Pith citation observations are available.