Pith. sign in

Paper Citation Record · LEDGER

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction

As of 17 August 2026, this Paper Citation Record lists 70 of 70 outbound references and 3 inbound Pith citation observations for arXiv:2606.08566.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.08566 v1

Coverage vector

measured 70 of 70 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-27T18:53:57.300531Z

measured 73 of 73 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T19:59:19.860529Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

70 of 70 outbound references displayed

  • verified exact3
  • verified fuzzy0
  • unresolved66
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 592df09a-62b9-4d39-977c-40a35a77d76d · outbound

This paper cites Mdkat: Multimodal decoupling with knowledge aggregation and transfer for video emotion recognition,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Mdkat: Multimodal decoupling with knowledge aggregation and transfer for video emotion recognition,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:9df7134418e24da204211d97ca63794bed7979a4318432e9877d72c0ae5451dc

Observation a9520a1d-ce63-419b-ad84-3fc801c0d4b3 · outbound

This paper cites Feature evaluation and joint interaction for audio-visual emotion recognition,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Feature evaluation and joint interaction for audio-visual emotion recognition,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:7bba2f5d9c115a4771a652e244fbb687f7b176b926cc4e1da7ed3e53a23368c7

Observation 10d641bd-8eb1-4af1-9371-f9f0cbb1e085 · outbound

This paper cites Glove: Global vectors for word representation,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Glove: Global vectors for word representation,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:3def63d33633bcaea81a01abbb4c540c874b4b2540c3852b62b38a2b88adf3e8

Observation 465dab71-b3bd-4f72-8dde-ce4e3b4c9949 · outbound

This paper cites Weakly supervised text-based actor-action video segmentation by clip-level multi-instance learning,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Weakly supervised text-based actor-action video segmentation by clip-level multi-instance learning,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:42c9c8ec0c183bbd3046c006997a3cf6d4544b8f331f9d2418970d8d7870fafd

Observation 89156e58-57c0-4662-8bef-4df5b18d54d7 · outbound

This paper cites Graph mixture of experts and memory-augmented routers for multivariate time series anomaly detec- tion,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Graph mixture of experts and memory-augmented routers for multivariate time series anomaly detec- tion,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:a7cf012e9769d8f91d2b515ae60eca84ecf9688e22478986dd22855190d094bf

Observation ccbac76a-39bc-4c79-ad36-043967f366ae · outbound

This paper cites GPT-4o System Card.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction GPT-4o System Card

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-07-02T22:27:25.971841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:5a22e376d9826aeabeda56a56b52e2675a8d0e5288a6b0dddde576920eb9d395

Observation 5476de3a-4007-4713-b712-c8c019c6e852 · outbound

This paper cites Ecpec: Emotion-cause pair extraction in conversations,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Ecpec: Emotion-cause pair extraction in conversations,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:e893a15df7d908b6d8cc45ad47f2fe06d589b656c04c31c292692db5e307090e

Observation 357faf6b-6cec-4f41-8e95-64fffb99ef5d · outbound

This paper cites Multi- round mutual emotion-cause pair extraction for emotion-attributed video captioning,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Multi- round mutual emotion-cause pair extraction for emotion-attributed video captioning,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:4cd4258538df3eb821a455b8564e5cb178d01bac3b42696bf4b9c95bd94a3aeb

Observation 550a8329-a2c6-4058-8e98-ccfe3326dd2d · outbound

This paper cites Global-view and speaker-aware emotion cause extraction in conversations,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Global-view and speaker-aware emotion cause extraction in conversations,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:9c5db14ca01b722b5cfe6643783839d0c2970ea52ef192021917e5207d401f6d

Observation 0ab00d3f-fa43-4639-8fa1-48fc08464af2 · outbound

This paper cites Multimodal emotion- cause pair extraction with holistic interaction and label constraint,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Multimodal emotion- cause pair extraction with holistic interaction and label constraint,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:0d0552d79f9196702dc242fde82288b4c8a61376bea8fd03b5143d22a4e6ca9f

Observation 14b3edcd-bef0-476f-af15-c1f97bc05aad · outbound

This paper cites Reconstruction network for video captioning,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Reconstruction network for video captioning,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:ba9424469a3991c350d1862ebf6b442efb8af7827fc1b48914270d63b09636af

Observation cb8ed6e9-0f93-4ee0-a706-21a8440474ee · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Adam: A Method for Stochastic Optimization

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-02T22:27:25.979582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:f34a00f6effe63c6384c345f569ee26c77531703ea3f60aee070c13a7c182d32

Observation e8587893-5761-4214-903d-e7022f36bf2b · outbound

This paper cites Enhancing emotion-cause pair extraction in conversations via center event detection and reasoning,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Enhancing emotion-cause pair extraction in conversations via center event detection and reasoning,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:b729ad2d0260f007de9767ce8a41d0148bc5fdd169c9f77c0d5d219937655cc0

Observation 4eaf246b-8fe6-45d5-9e78-91c3cc83b81f · outbound

This paper cites Prompting video-language foundation models with domain-specific fine-grained heuristics for video question answering,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Prompting video-language foundation models with domain-specific fine-grained heuristics for video question answering,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:7d6dcb734dd8cfff3f8d1eec9939c5b929324675216a4986117fca01c896db93

Observation bf8100fe-2677-4c30-97d6-90eda88882d4 · outbound

This paper cites Meteor: An automatic metric for mt evalua- tion with improved correlation with human judgments,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Meteor: An automatic metric for mt evalua- tion with improved correlation with human judgments,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:ad5cce8da3dd12b409eb88abb9f40b38979d960b5327921f3bb5bdf6c902c369

Observation 09ee6374-7aa1-4fc3-8392-775cbc48e1e3 · outbound

This paper cites Lora: Low-rank adaptation of large language models,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Lora: Low-rank adaptation of large language models,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:1e37c78f16db53b632562918c0760fc2f13334a32ce03544e5435bac8e452a05

Observation 52cf7334-9ad3-4671-857f-d1845087f8d7 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction LoRA: Low-Rank Adaptation of Large Language Models

Reference 17

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T22:27:25.977100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:821715c4077f961f2b65441519f302956e7edd952e171ffd27add633bed77857

Observation 1e1c7547-9adb-4b34-99bd-1779156efc9e · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Exploring the limits of transfer learning with a unified text-to-text transformer,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:b13008ad34c04991150560aab8824af1ef5dbaa9000b534816b969b69a31385d

Observation ddc21777-16ca-41d0-8cbf-29bf4495ea22 · outbound

This paper cites Learning probabilistic presence-absence evidence for weakly-supervised audio-visual event perception,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Learning probabilistic presence-absence evidence for weakly-supervised audio-visual event perception,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:a0daeb044b7623c8200f5671f8958d623c4095844408c50fc588febc9c6cd1da

Observation 8f6d043c-06e7-424f-af30-ae954c78f60f · outbound

This paper cites Expllm: Towards chain of thought for facial expression recognition,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Expllm: Towards chain of thought for facial expression recognition,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:9a0fb01f7baf181273b99769927c0477e4ad4b5454f70ab017c11f82a652dad1

Observation 8e900c65-4caf-4b48-9236-62c42662d63e · outbound

This paper cites Benchmarking micro- action recognition: Dataset, method, and application,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Benchmarking micro- action recognition: Dataset, method, and application,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:fa27cfed2356a546e91d684d2a89165734bef9e8385dbdd5b77b2bdf272b80e7

Observation bc5e89fb-4b1b-4ade-b632-56837bb26b92 · outbound

This paper cites Contextual attention network for emotional video captioning,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Contextual attention network for emotional video captioning,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:b37d469158ae67d8ded73fba9cb25a8394ebd668a230a9457fb08b8720497803

Observation 6ee6c3bb-369b-4f0d-aeef-688827c5d124 · outbound

This paper cites Observe before generate: Emotion-cause aware video caption for multimodal emotion cause gen- eration in conversations,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Observe before generate: Emotion-cause aware video caption for multimodal emotion cause gen- eration in conversations,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:a1a2c34e41c5a2e8684f8d5ef677f8241dc3a9d4e157f772e5bdedb72bf759f2

Observation d370b28d-9371-469b-801b-9a4f8d8c1e53 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Learning transferable visual models from natural language supervision,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:5a35e0d239fbf9d26775b71ff9285797234760f33cde4ff333a9783f76c3c9e8

Observation 52b67f99-e9b3-4c85-9475-ad18d6bf1711 · outbound

This paper cites Cross-modal coherence-enhanced feedback prompting for news captioning,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Cross-modal coherence-enhanced feedback prompting for news captioning,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:c50633b93dd08206d856394e2968005e47c688dae1756e2b6c6955ecd1825efc

Observation ddce4da4-0de3-4173-89d5-4ac28ebcf786 · outbound

This paper cites Cider: Consensus- based image description evaluation,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Cider: Consensus- based image description evaluation,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:cea4c59e1f4ae8a9f7b78ad3c21879f3cfdd7a76f6405bafbbf08dea7ba071ae

Observation 4a73872e-72ad-443c-b33e-ea616e1c17cc · outbound

This paper cites Semantic grouping network for video captioning,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Semantic grouping network for video captioning,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:40eed25114b2e3118707057d8ccfb092970a48c34b1c92f9acf56f0d827478ab

Observation e329f73f-c7f2-4dbe-817f-23b5e2550064 · outbound

This paper cites Rule-driven news captioning,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Rule-driven news captioning,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:aba2b52e34891ef2db1d72f24025ead4812da0db868c6c1c35c0815868f48b0c

Observation e33f43ca-6c4f-4074-affc-d1a334066fb4 · outbound

This paper cites Eliciting in-context learning in vision-language models for videos through curated data distributional properties,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Eliciting in-context learning in vision-language models for videos through curated data distributional properties,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:c0bd55d90c6d8ba97aaca3e111d9b59de1b4fd8ed593e344b531e96a15a74af6

Observation 1f8ef91f-4d3e-411d-bc4c-a5fab71d6307 · outbound

This paper cites A versatile multimodal learning framework for zero-shot emotion recognition,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction A versatile multimodal learning framework for zero-shot emotion recognition,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:6b234298660c2c6e4cbf0342a535ecb4b7155dcc5a30a9e98aa24eda132addae

Observation 93be78e6-8ada-4b33-aaf9-33959f4de274 · outbound

This paper cites Cascade cross-modal attention network for video actor and action segmentation from a sentence,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Cascade cross-modal attention network for video actor and action segmentation from a sentence,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:7ed55d6ed146935bcd4da0f487610f19e8fc895ddfc78ce5b09705aeb5e4bffe

Observation 51c51306-eb7c-410a-a876-d9a9b72ded4e · outbound

This paper cites Emotion-cause pair extraction: A new task to emotion analysis in texts,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Emotion-cause pair extraction: A new task to emotion analysis in texts,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:3c38e0284d6e5d9666c14892772500e07d635af99d15564412d9db001361d65c

Observation 40800fc4-c3e5-4751-8452-f0935f7ba9fa · outbound

This paper cites Collecting highly parallel data for paraphrase evaluation,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Collecting highly parallel data for paraphrase evaluation,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:f0e42d0d1eb4f88422796b9c0a5e988c4e848389d9b30489fe2d5ba2718749de

Observation c924b1db-9ea0-465c-a82a-e9c15305fadf · outbound

This paper cites From coarse to fine: A distillation method for fine-grained emotion-causal span pair extraction in conversation,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction From coarse to fine: A distillation method for fine-grained emotion-causal span pair extraction in conversation,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:ed399864b4477440724ae2d7e1eee15d8ffda30cdbaeb4d6e2c74d0f0fe9403d

Observation 714a911b-6f89-499a-8fdf-a2a5f34cd025 · outbound

This paper cites From extraction to generation: multimodal emotion-cause pair generation in conversations,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction From extraction to generation: multimodal emotion-cause pair generation in conversations,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:33e9a8539528bb85d1a2c97f7366b2140b61cc1c4cb04870a0dfb3323a981922

Observation c4cc4550-738e-426d-875a-526872d03bb8 · outbound

This paper cites Improving image captioning via predicting structured concepts,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Improving image captioning via predicting structured concepts,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:09a3c129f70634b4568b2c0c4596373fa32277c18a67474290ecfded6f9eea78

Observation a81cdc17-dd6b-46bb-9a87-d74328cb3014 · outbound

This paper cites Bootstrapping large language models for radiology report generation,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Bootstrapping large language models for radiology report generation,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:40afe63e06059d9c8c332102b8997b51b47958167e7630cb7c89e766534b32e9

Observation cd5acfcd-393c-4181-890a-4c0335f0bcf5 · outbound

This paper cites VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-07-02T22:27:25.974508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:3a047a6b03fc883eb4fc0b0feba2b472abb3b827a22def55742c44a8a1f765db

Observation 6bf84976-00ca-495e-82cd-197ee6e6fed7 · outbound

This paper cites Improving radiology report generation with d 2-net: When diffusion meets dis- criminator,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Improving radiology report generation with d 2-net: When diffusion meets dis- criminator,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:f0c70b52bf7b01afae211b60dd65492bf6573e7f33bc038c28e1f8e40dbfd785

Observation 89c4b020-c1c7-44e7-90d4-ca7abeb2d1a8 · outbound

This paper cites Improving radiology report generation with multi-grained abnormality prediction,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Improving radiology report generation with multi-grained abnormality prediction,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:cb1ce7d6b711968c378dde138ca5bc5a54a290ee84896eca75f22267d2ab2f2e

Observation d3058131-ee76-439d-9a64-5900f2ac94f2 · outbound

This paper cites Enriched image cap- tioning based on knowledge divergence and focus,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Enriched image cap- tioning based on knowledge divergence and focus,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:083664d51fe8ee4a5f39cfe08ef6e57b2d75adc9096a12b8cb07063ec9e3b27f

Observation 134c1104-5ab5-47e7-a899-c93bdb75713c · outbound

This paper cites Emotional video captioning with vision-based emotion interpretation network,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Emotional video captioning with vision-based emotion interpretation network,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:9a5fc228cea5213625195751da443f1918d43764789baceac5d362a957538986

Observation 66caae5c-97a8-4ef8-b70e-0f26ac6dbc68 · outbound

This paper cites Emotion- prior awareness network for emotional video captioning,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Emotion- prior awareness network for emotional video captioning,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:7b7640d592112d3011e86c150e220efa3bf984479a1f33479c1a1fd7e7858765

Observation 59614382-fde1-450a-ab52-fa6a62e0cd88 · outbound

This paper cites Combatting data imbalance and noise in micro-action recognition,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Combatting data imbalance and noise in micro-action recognition,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:57bd5e41f45ab2ea693fbfc2a2923505250bec6e6dabc785811711b5fcd5ff41

Observation d9a9ee7a-c80d-49e1-80a7-8dae8c425877 · outbound

This paper cites Eliciting in-context learning in vision-language models for videos through curated data distributional properties,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Eliciting in-context learning in vision-language models for videos through curated data distributional properties,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:62c5ee635bfefc1b086d58d59adef2fca4dcf684fb02e9ca3f8c4a33b5a70f40

Observation 2e9a73e7-761d-4fd5-9ee7-b4b721e91040 · outbound

This paper cites Obtaining reliable human ratings of valence, arousal, and dominance for 20,000 english words,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Obtaining reliable human ratings of valence, arousal, and dominance for 20,000 english words,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:b680faf3f20c1242a93aa9c00c062cf36db0b075f1023c755f07e44258e58227

Observation 961a1ab3-11eb-408e-a420-07008ebe47d4 · outbound

This paper cites Linguistic-aware patch slimming framework for fine-grained cross-modal alignment,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Linguistic-aware patch slimming framework for fine-grained cross-modal alignment,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:914bec2b111e42cee39d1084a63ff0cb2c3611c40cdd33e255ef18ede4a034d8

Observation 21dc0f3a-87b9-4028-a323-387ab8b1fa25 · outbound

This paper cites Emotion-oriented cross-modal prompting and alignment for human- centric emotional video captioning,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Emotion-oriented cross-modal prompting and alignment for human- centric emotional video captioning,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:0dbb0a30b96d6ffdcc4f18717248c5fa71f97d00c807cc40adbcda916bcea933

Observation 3ce3a258-af13-4c16-8b20-037ea4dade00 · outbound

This paper cites Dual-path collaborative generation network for emotional video captioning,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Dual-path collaborative generation network for emotional video captioning,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:b59ed0665b413bb5904c6b8dde53045e03ef90ea9361a8ce6af2948f75843765

Observation 5b05a019-abee-4f49-9969-aad0aa093855 · outbound

This paper cites Rouge: A package for automatic evaluation of summaries,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Rouge: A package for automatic evaluation of summaries,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:d7a5576dd9c5f40b4b7079a5e7bf2585e8bc1008163b1d639955adb58848ef8f

Observation fe8bb236-7672-4e39-b7e8-9d1cd63f69b9 · outbound

This paper cites A knowledge-guided graph attention network for emotion-cause pair ex- traction,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction A knowledge-guided graph attention network for emotion-cause pair ex- traction,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:c32399f737136d1b059a5cf370c7e3b03e96602466412751d7c167c825ef7de6

Observation a48cf7b1-e8c9-4de4-94a2-a9ff311074f4 · outbound

This paper cites A comprehen- sive survey of 3d dense captioning: Localizing and describing objects in 3d scenes,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction A comprehen- sive survey of 3d dense captioning: Localizing and describing objects in 3d scenes,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:352ac5fa96115baf5c37d1628376746a6d93cd602e56e5b15e03429693f51a1b

Observation 318f8aa0-b129-4aea-a06e-4720125b2b76 · outbound

This paper cites Predicting emotions in user-generated videos,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Predicting emotions in user-generated videos,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:851869c03ea92448d150f06b41be815dcf9d351f128eb1ece6cee72f66e88f69

Observation 4e2ff6a0-8b09-4040-b0fc-c7f14bb13a38 · outbound

This paper cites Multi-attention network for compressed video referring object segmentation,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Multi-attention network for compressed video referring object segmentation,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:723678f52815c259144b07c0be737605367cf7baeb29f8dd18d348e27dc2ddfa

Observation 69c70fba-1c49-467c-9157-2cab1641f4f1 · outbound

This paper cites Towards efficient partially relevant video retrieval with active moment discovering,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Towards efficient partially relevant video retrieval with active moment discovering,

Reference 55

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:28a77decd84c98ac43c8aeeea0cd970e021900c71205c604aa2967e5b473e9b8

Observation f03ab210-8308-4660-acb6-b751abe24ca8 · outbound

This paper cites Vectorized evidential learning for weakly- supervised temporal action localization,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Vectorized evidential learning for weakly- supervised temporal action localization,

Reference 56

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:8a6fccda1804d8e37e341f3df2fa3a204ead2fdcbbf382478d1eeee967c7fc6d

Observation 2a7c8efd-e2a6-4049-81d5-61cc737c4199 · outbound

This paper cites Sentiment-oriented transformer- based variational autoencoder network for live video commenting,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Sentiment-oriented transformer- based variational autoencoder network for live video commenting,

Reference 57

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:89eaf89aaf34a1f6366b815dd6a6aea9154c4765fd57a59e4ee66a4f985a68c8

Observation 7fa533f6-2f98-45b7-a27e-5afb452db884 · outbound

This paper cites Prompting few-shot multi- hop question generation via comprehending type-aware semantics,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Prompting few-shot multi- hop question generation via comprehending type-aware semantics,

Reference 58

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:f5034444b778bad50832e016e7babd7008cc846fbf2476483193b2857eb39b5f

Observation 08f17e4f-7d30-47c3-9aae-0ad120a8f86c · outbound

This paper cites Affectnet+: A database for enhancing facial expression recognition with soft-labels,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Affectnet+: A database for enhancing facial expression recognition with soft-labels,

Reference 59

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:e49aa5263943cd04add42f3e05de1b5748b42e69cd75bdba141764f5c506b705

Observation 75904bad-a358-46dc-b846-65d92f1369dd · outbound

This paper cites Emotion expression with fact transfer for video description,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Emotion expression with fact transfer for video description,

Reference 60

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:b972e67082882b409972d8af2154a1298414bc8ac2db35a90333cd3bed9bb9e2

Observation 41108cfc-9607-4209-924d-39d3ea81ca56 · outbound

This paper cites Graph-based multimodal sequential embedding for sign language translation,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Graph-based multimodal sequential embedding for sign language translation,

Reference 61

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:9224ac70a6ffb1d54a88325cad9ea4f7de41bf83b535d62df44b4c96dc35deba

Observation 75872f37-f97c-4c54-8349-4d822cee9796 · outbound

This paper cites Boost tracking by natural language with prompt-guided grounding,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Boost tracking by natural language with prompt-guided grounding,

Reference 62

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:7707e3c37e408dfe01d9320750aac80d4ed75162f95ab33751424d7b165840c8

Observation 182605e2-ccab-4724-bd15-d3a1ffbe766d · outbound

This paper cites Multimodal emotion- cause pair extraction in conversations,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Multimodal emotion- cause pair extraction in conversations,

Reference 63

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:4fbbef82c43c7927a5334e777557d20cc66f2d98955dfef3f6be82ad3db8cb2a

Observation ef5c5014-5569-4cad-aed6-285abcf533f3 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Bleu: a method for automatic evaluation of machine translation,

Reference 64

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:56f48a9414f76d0b9aa8bfc95cdd9c399fa1dd6be41d3df631d3e56d0e2e6e25

Observation 3e13477b-133b-4019-a24e-5ef05d5b5d04 · outbound

This paper cites Syntax- guided hierarchical attention network for video captioning,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Syntax- guided hierarchical attention network for video captioning,

Reference 65

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:7d496f62eeac308badc07e3be3c1c9e9c54a6cf681ea8dfb6183a9a177649342

Observation dd43d5e1-f421-4fcd-a462-45bd678f316b · outbound

This paper cites Enhanced generative framework with llms for multimodal emotion-cause pair extraction in conversations,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Enhanced generative framework with llms for multimodal emotion-cause pair extraction in conversations,

Reference 66

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:dbb8b0f79c51408b6804bf7c351f44060e44ea3da13fdb64ee24d546ed017d0e

Observation e9323464-c30f-4b8b-a39f-0eb4d83caaec · outbound

This paper cites Improving video summarization by exploring the coherence between corresponding captions,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Improving video summarization by exploring the coherence between corresponding captions,

Reference 67

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:9a6b575fb309a70364268c7b6dfaab2c81694cf86baf3dbbe36a5a905c4d9bbb

Observation 9cd40461-af6a-47c8-9ad0-f74ea2c030de · outbound

This paper cites Emotion prediction oriented method with multiple supervisions for emotion-cause pair extraction,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Emotion prediction oriented method with multiple supervisions for emotion-cause pair extraction,

Reference 68

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:be75823c2c7a82b2d6bf5f5614d168bf173a5df30965f6831f56d14406b21518

Observation 31e6d2c5-35ea-4588-89e7-e2db095fd644 · outbound

This paper cites Subjective- objective emotion correlated generation network for subjective video captioning,.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction Subjective- objective emotion correlated generation network for subjective video captioning,

Reference 69

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:8db6f0d65f3fe2931c0fa62525fd3bbbe2ae2cd564232b9e702b456f1a4304ba

Observation b83f9170-34c2-4e6a-a9ea-b0670e15d5e3 · outbound

This paper cites He was a post-doctor with the School of Information Science and Technology, University of Science and Technology of China, from 2022 to 2024.

Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction He was a post-doctor with the School of Information Science and Technology, University of Science and Technology of China, from 2022 to 2024

Reference 70

Resolution
unresolved
no resolver link, observed 2026-06-27T18:53:57.300531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T18:53:57.300531Z digest=sha256:e4722b533c0a661a9ff3c7767b3ad95a2b72c5ff74eeb88c8470a7f672d0ad72

Pith citing papers

Observation 932542d8-a3a2-4169-a22f-4abbbbbbeea3 · inbound

EmoStyle: Affective Conditioning of Style-Specialist Experts for Emotional Image Generation cites this paper.

EmoStyle: Affective Conditioning of Style-Specialist Experts for Emotional Image Generation Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-14T13:48:36.708969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T13:48:36.708969Z digest=sha256:d98663435dbc164acfae9c878ef62cb7fd8278abd65334be17d0ac5c9a581a77

Observation cef63516-d1ba-46d6-a4d9-4013f1b08080 · inbound

Geometry-aware Gaussian Prior and Axial Attention for Cervical Cytology Image Classification cites this paper.

Geometry-aware Gaussian Prior and Axial Attention for Cervical Cytology Image Classification Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction

Reference 46

Resolution
unresolved
no resolver link, observed 2026-07-14T12:58:07.349876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:58:07.349876Z digest=sha256:b961a5aa2dad97fab4865d9c410941d9c9c3d45767fa22af29c9ecf00d973e99

Observation 827fc6d4-a6fb-425d-aae6-910f65ee1df3 · inbound

HTT-Net: Hierarchical Text-guided Transition Modeling for Surgical Video Phase Recognition cites this paper.

HTT-Net: Hierarchical Text-guided Transition Modeling for Surgical Video Phase Recognition Towards Accurate Emotion-Attributed Video Captioning via Fine-grained Emotion-Cause Pair Extraction

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-01T19:59:19.860529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T19:59:19.860529Z digest=sha256:5dad1ffaa99d8944838492e5a24c141de79d0db82667768673fb0263f1f044ad