Pith. sign in

Paper Citation Record · LEDGER

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation

As of 23 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 0 inbound Pith citation observations for arXiv:2505.18168.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.18168 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:44:29.296392Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

58 of 58 outbound references displayed

  • verified exact0
  • verified fuzzy25
  • unresolved33
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fd585619-4881-4431-a9bd-fe730379ea18 · outbound

This paper cites A comprehensive review of facial expression recognition techniques.Multimedia Systems, 29(1):73–103, 2023.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation A comprehensive review of facial expression recognition techniques.Multimedia Systems, 29(1):73–103, 2023

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.448939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:28.894066Z digest=sha256:7db5929b9196371949beda988fe861cf8063c7bef01a1751f442e9c42a29ff24

Observation f4c81d0c-efae-426b-84af-0d4c07a6e459 · outbound

This paper cites Llama 3 model card.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Llama 3 model card

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:28.900814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:28.900814Z digest=sha256:f52e30b70f50f5d7d315ee5c20d1d807b1bf0400c77a0bf49ffe184b6a4c0ca4

Observation f9b005ec-5505-41f0-9761-d91329bffaf0 · outbound

This paper cites Claude-3.5.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Claude-3.5

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:28.909847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:28.909847Z digest=sha256:465c0bfa93eab94b22110aa7661cd66be4959e67a8a3ee1cc1529ce298de5708

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:28.915865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:28.915865Z digest=sha256:82c883b0be963fe5b3f4884fe486a087ac826d4589e77b987501f762596feb7d

Observation efa85399-bc12-401f-8eb7-10c7d6329949 · outbound

This paper cites Video-based facial micro-expression analysis: A survey of datasets, features and algorithms.IEEE transactions on pattern analysis and machine intelligence, 44(9):5826–5846, 2021.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Video-based facial micro-expression analysis: A survey of datasets, features and algorithms.IEEE transactions on pattern analysis and machine intelligence, 44(9):5826–5846, 2021

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.393790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:28.923638Z digest=sha256:ac8aedfacd7d141ec95f64a8af2e3e9be1ed5d741eae97012feb93c395a28492

Observation f1a0ea46-782d-486d-9fd2-3cc0bb322616 · outbound

This paper cites Knowledge-driven self-supervised representa- tion learning for facial action unit recognition.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Knowledge-driven self-supervised representa- tion learning for facial action unit recognition

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.373312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:28.932420Z digest=sha256:04653f21dfcd08acdb719c1a3c4402d06e94fba0639e5cf2ebb5a2b58f81c485

Observation 64dff31f-2c95-413c-bc29-a408f7a74e10 · outbound

This paper cites MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:28.941071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:28.941071Z digest=sha256:3acd4305b83a758e628083392ed03931638c8456fa256563cf946db131382904

Observation 6df65fbd-105c-42ed-bcb9-d5ad5940160a · outbound

This paper cites How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:28.948070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:28.948070Z digest=sha256:11ac86a296b69a91f8487f256833b26a598bdc62ea5f500aa1a5985dda34cadb

Observation c113b1d2-cb90-4053-9e67-7f9642cd5069 · outbound

This paper cites Emotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Emotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:28.953395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:28.953395Z digest=sha256:6fc4a52fd3c0fafdd152252066357692edbae4d595baef63c1d78616df32fe06

Observation 6480a781-f225-4d80-bd6a-e0962068275f · outbound

This paper cites Knowledge augmented deep neural networks for joint facial expression and action unit recognition.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Knowledge augmented deep neural networks for joint facial expression and action unit recognition

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.357090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:28.959369Z digest=sha256:78915b7858582da77cc63b54e61648a1e301c4437cff67bfae282f145dfd9a96

Observation 0b40f4f7-85c6-414f-93d8-9a224ae0e826 · outbound

This paper cites Knowledge augmented deep neural networks for joint facial expression and action unit recognition.Advances in Neural Information Processing Systems, 33:14338–14349, 2020.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Knowledge augmented deep neural networks for joint facial expression and action unit recognition.Advances in Neural Information Processing Systems, 33:14338–14349, 2020

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.338683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:28.965617Z digest=sha256:eda42a20568679b14ae6ea9b097a121acae063691b09a0b810d66348763d47df

Observation 70b427b1-5d35-4377-8650-259d7ecd28e4 · outbound

This paper cites Decomposition of uncertainty in bayesian deep learning for efficient and risk-sensitive learning.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Decomposition of uncertainty in bayesian deep learning for efficient and risk-sensitive learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.319853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:28.971272Z digest=sha256:3abeefcddf1081b4babd6a1c719fa0f2c3a690693bede3504c9b94794558f823

Observation ad635b35-ec6a-4bfd-963f-c24dba3eaed2 · outbound

This paper cites Consulting Psychologists Press, 1978.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Consulting Psychologists Press, 1978

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.302560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:28.976805Z digest=sha256:b854089ee2b9dd1d76c338ca7e13d9c1178c5b5f8b9a5d3e4d5c9014696da4b9

Observation e0164aa9-a2d8-4f3b-8ba0-e1a5767d1518 · outbound

This paper cites Universals and cultural differences in the judgments of facial expressions of emotion.Journal of personality and social psychology, 53(4):712, 1987.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Universals and cultural differences in the judgments of facial expressions of emotion.Journal of personality and social psychology, 53(4):712, 1987

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.284906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:28.981797Z digest=sha256:45ef96e15c228c340424069798446cb2cbc6d5d251b011fda9dcb8a03a66f2cf

Observation 75196822-5c1b-464b-9ded-20435cad4d14 · outbound

This paper cites Oxford University Press, USA, 1997.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Oxford University Press, USA, 1997

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.267463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:28.988199Z digest=sha256:03bab020f6d95c54b640a6b4f0e697ab805b7ae8624c3eb30f5139ff327af021

Observation 14a49383-e8f3-456d-8012-77ed47e06e97 · outbound

This paper cites Selfme: Self-supervised motion learning for micro-expression recognition.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Selfme: Self-supervised motion learning for micro-expression recognition

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.249255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:28.995942Z digest=sha256:5aee509f29a29d8852585b580f17bac4f6c47320bce6c5d49b9d3453ab37fff8

Observation 3da9ee51-14c9-4a74-bb20-d6377e3075db · outbound

This paper cites StimuVAR: Spatiotemporal Stimuli-aware Video Affective Reasoning with Multimodal Large Language Models.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation StimuVAR: Spatiotemporal Stimuli-aware Video Affective Reasoning with Multimodal Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.001846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.001846Z digest=sha256:f9dd1eb265fc6da8f219df67bec82386b2504fc8d07ac6cdaa6dd35cdd817f50

Observation 3a11e821-1bf5-4677-beb8-bdae00a73d9f · outbound

This paper cites Opera: Alleviating hallucination in multi-modal large language models via over-trust penalty and retrospection- allocation.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Opera: Alleviating hallucination in multi-modal large language models via over-trust penalty and retrospection- allocation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.225745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:29.007507Z digest=sha256:2b49e651ae1bbb4136e0fafee321f2d8e9b709c7189b126cae1532b1eac1ff77

Observation 57c4d9c6-abde-4ca6-a54f-8f3c02e4eb99 · outbound

This paper cites Disentangling identity and pose for facial ex- pression recognition.IEEE Transactions on Affective Computing, 13(4):1868–1878, 2022.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Disentangling identity and pose for facial ex- pression recognition.IEEE Transactions on Affective Computing, 13(4):1868–1878, 2022

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.205050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:29.013511Z digest=sha256:bb9f9ffa3d3c9a0d74d1d468b3197ed4bbfb313237835438295444c7e84742ce

Observation 745b3899-9528-4f52-b83f-f36bb18e897b · outbound

This paper cites Scaling Laws for Neural Language Models.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Scaling Laws for Neural Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.020890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.020890Z digest=sha256:06a87c281ef39279a18beea0378645d57cadbf0b7db96c3697c16a3e3b2465ce

Observation 4dbe5b7c-5255-4780-858c-4b1e63e723ed · outbound

This paper cites What uncertainties do we need in bayesian deep learning for computer vision?Advances in neural information processing systems, 30, 2017.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation What uncertainties do we need in bayesian deep learning for computer vision?Advances in neural information processing systems, 30, 2017

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.027335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.027335Z digest=sha256:72d25d896d75973e778da8dd1273f272b8e48fe438d880d25109ab505bfa45ad

Observation 14e3cbe0-f036-49c6-9812-b2f827fa4ebe · outbound

This paper cites Expression, Affect, Action Unit Recognition: Aff-Wild2, Multi-Task Learning and ArcFace.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Expression, Affect, Action Unit Recognition: Aff-Wild2, Multi-Task Learning and ArcFace

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.032686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.032686Z digest=sha256:d7876ee8aaa91b1ecb3696c42cb106b0f0ab84a2a41a767aa03cff5553d72bcf

Observation e131d91b-6d0a-42ba-9509-7a97a75e051a · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation LLaVA-OneVision: Easy Visual Task Transfer

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.038401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.038401Z digest=sha256:727fe900c8402099f43a336cb248d03f55c4bb5fc78b914d3be4d2868f5c2307

Observation b9df44d5-067c-460b-a979-2be5b9b89ed7 · outbound

This paper cites LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.051757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.051757Z digest=sha256:4e490eafe09591ababb49ac1ac42a62e5c9417f199e94477d4a44946f5df6044

Observation f7e0f455-66ff-43f4-b495-efc396b13ea6 · outbound

This paper cites Deep facial expression recognition: A survey.IEEE transactions on affective computing, 13(3):1195–1215, 2020.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Deep facial expression recognition: A survey.IEEE transactions on affective computing, 13(3):1195–1215, 2020

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.168425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:29.057474Z digest=sha256:ed4f00ae404d2e8deee0caba76ca96ca5aa37869694691dd16e92a79aebbac9a

Observation fc6dc9f8-2aa3-4003-a90e-44dd3b59918c · outbound

This paper cites Reliable crowdsourcing and deep locality-preserving learning for expression recognition in the wild.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Reliable crowdsourcing and deep locality-preserving learning for expression recognition in the wild

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.062498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.062498Z digest=sha256:fcdc966d697eac2a1fef149cd0395f827091312753573a3aefa09b065a77054f

Observation 9be963ef-447c-4508-bf1d-b1e0c7f15dce · outbound

This paper cites AffectGPT: Dataset and Framework for Explainable Multimodal Emotion Recognition.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation AffectGPT: Dataset and Framework for Explainable Multimodal Emotion Recognition

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.077312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.077312Z digest=sha256:85bad16f29d93d0f3073436abfe9cee6d1230b2221a1cbb7889d07335df1beb8

Observation 98a81cd0-be0a-462b-ab95-3b99c413f73b · outbound

This paper cites Gpt-4v with emotion: A zero-shot benchmark for generalized emotion recognition.Information Fusion, 108:102367, 2024.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Gpt-4v with emotion: A zero-shot benchmark for generalized emotion recognition.Information Fusion, 108:102367, 2024

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.132961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:29.084190Z digest=sha256:17da13cc1fb8a2ec5c009e5b8718a5cc4806ab6fd26254176246b1766cb9e260

Observation a712980b-a266-4421-a2fc-2ed05ad21a9f · outbound

This paper cites Video-LLaVA: Learning United Visual Representation by Alignment Before Projection.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Video-LLaVA: Learning United Visual Representation by Alignment Before Projection

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.089156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.089156Z digest=sha256:1677b76cc99a0d51385a715cc1816e15021b3113c4cdf39ab3c68e3deeace1b0

Observation dbf13a1f-c0d8-4dbe-b698-d95295c02381 · outbound

This paper cites an unresolved cited work.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:44:30.113934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:29.094388Z digest=sha256:5e8aa68ddf9c5e2759e21ff6a3528468e921e0d385f16c254cd6444deece57b7

Observation 51b2a338-d5b0-4a45-bdc6-eebed7f88f3c · outbound

This paper cites Improved baselines with visual instruction tuning.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Improved baselines with visual instruction tuning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.099555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.099555Z digest=sha256:fd96af79dc05953abf27e0c94fe137a078b09aba1ea1ed69a303bff87124828c

Observation e3975bf9-e31e-4837-b8cf-bd9194cfc3c3 · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36, 2024.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Visual instruction tuning.Advances in neural information processing systems, 36, 2024

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.105455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.105455Z digest=sha256:1c45894e17b4becf16c3acaa0a2f893f64fe6c1623b1b9fd77105d8d298ec959

Observation bcefcfd6-9b7b-433c-84e8-a4efc059aea2 · outbound

This paper cites Facial expressions elicit multiplexed perceptions of emotion categories and dimensions.Current Biology, 32(1):200–209, 2022.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Facial expressions elicit multiplexed perceptions of emotion categories and dimensions.Current Biology, 32(1):200–209, 2022

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.047620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:29.113610Z digest=sha256:ca748703b36fae8871a95561a7f0a717db8f2105cb002e9ae7ba8b3b1528a3b6

Observation 77c80aab-deea-4ab0-b651-fb41d2a2aedb · outbound

This paper cites Mafw: A large-scale, multi-modal, compound affective database for dynamic facial expression recognition in the wild.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Mafw: A large-scale, multi-modal, compound affective database for dynamic facial expression recognition in the wild

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.121054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.121054Z digest=sha256:fb005fd7f6b00403d3f2f969e359fc1c0d672b0f7c5a4be6695c6d96f6d729fb

Observation 125f7e72-fa0f-4036-9d25-3013e45fa241 · outbound

This paper cites an unresolved cited work.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:44:30.018532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:29.127122Z digest=sha256:b6c5294b0148a59dc876fec5aa9f5307cf1ebec6137f5c39415e3c5847873402

Observation 79a82334-a25c-44e0-b9bd-6e1dac34184c · outbound

This paper cites Disfa: A spontaneous facial action intensity database.IEEE Transactions on Affective Computing, 4(2):151–160, 2013.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Disfa: A spontaneous facial action intensity database.IEEE Transactions on Affective Computing, 4(2):151–160, 2013

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.996354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:29.136191Z digest=sha256:d33df20240b197988354367b4d65fe05b4896838fb18bf70c0efbb79bedf9ada

Observation 8054be0b-be9f-4196-a4e4-dd3a699d5405 · outbound

This paper cites Affectnet: A database for facial expression, valence, and arousal computing in the wild.IEEE Transactions on Affective Computing, 10(1):18–31, 2017.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Affectnet: A database for facial expression, valence, and arousal computing in the wild.IEEE Transactions on Affective Computing, 10(1):18–31, 2017

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.976436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:29.141926Z digest=sha256:3e020acbb03304b81497fff14641fa6ac0244f6b359b2003ca881f5200ef0f5f

Observation 979104c4-dc8e-4e2f-90d1-45aeafc68fa1 · outbound

This paper cites Multi-label co- regularization for semi-supervised facial action unit recognition.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Multi-label co- regularization for semi-supervised facial action unit recognition

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.955566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:29.146858Z digest=sha256:16349c3b7d0da963c0b4fddc1c8a641bf84790a511ab0f600608f7376e4a2c0b

Observation 5901f188-0ce3-4085-9712-84e3446d8ec2 · outbound

This paper cites an unresolved cited work.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.152898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.152898Z digest=sha256:97442b589744221130b667146777f5e396e5dac2b65b0d81ca4c6fe09d6bea59

Observation 2eda21f1-d106-4c47-ab59-be000aac1d17 · outbound

This paper cites Hello gpt-4o.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Hello gpt-4o

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.158753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.158753Z digest=sha256:389e3f9dbdd4a7abdf8f3b1f3b18ed9a02db2afaaf1d26b6d70b23e2d015fc6e

Observation a10c2b5b-932d-419c-b748-35c1eab6f891 · outbound

This paper cites A unified and interpretable emotion representation and expression generation.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation A unified and interpretable emotion representation and expression generation

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.908899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:29.165525Z digest=sha256:bb5d226972466720de553533017f86f86b6d306772141bd03c8e531fdb2d78c5

Observation d5cd58f2-4a73-469d-982c-3bc941549031 · outbound

This paper cites A circumplex model of affect.Journal of personality and social psychology, 39(6):1161, 1980.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation A circumplex model of affect.Journal of personality and social psychology, 39(6):1161, 1980

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.178279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.178279Z digest=sha256:9d054bec884ab51de57f6a7025d841fc48e79a89eb24f93601a7e170ea015818

Observation edb27f16-0e2b-4d94-a18e-edc32ab125a2 · outbound

This paper cites Uncertain graph neural networks for facial action unit detection.Proceedings of the AAAI Conference on Artificial Intelligence, 35(7):5993–6001, 2021.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Uncertain graph neural networks for facial action unit detection.Proceedings of the AAAI Conference on Artificial Intelligence, 35(7):5993–6001, 2021

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.875105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:29.187860Z digest=sha256:9fe4790f0ebe897d21a199840cad7bb07c495e3cdce48c94d274224e6fc1fbc9

Observation a8f47576-8091-4484-9c08-adbeb174d240 · outbound

This paper cites Hybrid message passing with performance-driven structures for facial action unit detection.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Hybrid message passing with performance-driven structures for facial action unit detection

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.854631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:29.194648Z digest=sha256:b327e6b8330e30e79168e78b5f44274737c145680d9b4d280388505096d79dcd

Observation d3416001-0052-425e-a3a4-23903835f89c · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Gemini: A Family of Highly Capable Multimodal Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.199625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.199625Z digest=sha256:d39c94907e88625faa6dc40248ebdf6b82e402b61d08fc8a2b9f5bac38209f9d

Observation 3fe20492-2509-484f-9a1b-b6ab00c93242 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.207681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.207681Z digest=sha256:2b6b6450620a0b2b2c5bc5649a5420b6836fc1f51c5c0b3fdccb3484a3426c86

Observation 3aae1250-6c06-48f5-b557-3bba229d829b · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation LLaMA: Open and Efficient Foundation Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.212990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.212990Z digest=sha256:38c22dccf9abf0c44ecba98e2c62cf62603ed876a1ddc7cf0f00882f03a82326

Observation 0dd3b1f7-89e4-4717-a477-52605470f60e · outbound

This paper cites Rethinking the learning paradigm for dynamic facial expression recognition.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Rethinking the learning paradigm for dynamic facial expression recognition

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.831395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:29.218238Z digest=sha256:aece3ae498c156c584c4246ff683216f31c64465c977a1cc1d3863c0f4cc9994

Observation 35934a95-d9e1-42a0-a6dc-1c0ce49d7288 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.223723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.223723Z digest=sha256:cf5f9371c57da70434cea676351f28269eafc687d2b89287d8cad19a847d6338

Observation de36e934-6782-42fa-9aee-d879762dbda5 · outbound

This paper cites Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.233811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.233811Z digest=sha256:8acd23dcfb85cd54b93a94c315d79bc3dd1dc6e21491a4180495c9dd574f3950

Observation f76c93ae-0f6e-42bc-ae1f-c6ae17b41707 · outbound

This paper cites Emovit: Revolutionizing emotion insights with visual instruction tuning.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Emovit: Revolutionizing emotion insights with visual instruction tuning

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.240611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.240611Z digest=sha256:caae065ad4610842ff43e8f0565a7bd22bb282e315629b5e50bfac1a05649406

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.247106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.247106Z digest=sha256:610c506f20a1abeff06e0d2892e49d62175885f58541dcb60f5d50a1d85d8cbc

Observation c69de7ab-bd21-4e78-969f-d0d6795b13a4 · outbound

This paper cites Robust emotion recognition in context debiasing.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Robust emotion recognition in context debiasing

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.793904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:29.255486Z digest=sha256:dbe282a025fa64c5617b00bb8ebd9e34374af4cfa391bd8c46288101ce7253ad

Observation 492de562-dca7-4574-bf19-7a55fa8adf0d · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.262550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.262550Z digest=sha256:c42a12f01970ca1b58402d0d180dd52c5066e5c2556bb8803eb2bdc5ff6f0f97

Observation aebe6239-57eb-4344-ada0-bd843454f20a · outbound

This paper cites Sigmoid loss for language image pre-training.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Sigmoid loss for language image pre-training

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.276459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.276459Z digest=sha256:3441f25f74a5dcb7b505fd012c5d34eb308bfc031d89aaa2bea7a106c0b3b216

Observation c66b5d0a-dbb2-4932-9697-f7fb8ebd4415 · outbound

This paper cites A high-resolution spontaneous 3d dynamic facial expression database.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation A high-resolution spontaneous 3d dynamic facial expression database

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.761442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:29.284777Z digest=sha256:6d858d59056246f8901f5c55386009f2f5ad80e198851bb6a048f304e21a4cec

Observation 45d9dd06-f5d7-46be-809d-79e895bd6a15 · outbound

This paper cites Bp4d-spontaneous: a high-resolution spontaneous 3d dynamic facial expression database.Image and Vision Computing, 32(10):692–706, 2014.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Bp4d-spontaneous: a high-resolution spontaneous 3d dynamic facial expression database.Image and Vision Computing, 32(10):692–706, 2014

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.290214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.290214Z digest=sha256:319336b45845ffa94468c4d99a1c4410abd5134fb757780739ed6ac495c92b0e

Observation d99c2338-975d-412d-81a8-97d9fad1dd0c · outbound

This paper cites Khfa: Knowledge-driven hierarchical feature alignment framework for subject- invariant facial action unit detection.IEEE Transactions on Instrumentation and Measurement, 2024.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Khfa: Knowledge-driven hierarchical feature alignment framework for subject- invariant facial action unit detection.IEEE Transactions on Instrumentation and Measurement, 2024

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.730776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:44:29.296392Z digest=sha256:1e932d5182aa37e36edbf02a9b3a5e0cd382ad1bf13e28bd8ac75b000756b034

Pith citing papers

No inbound Pith citation observations are available.