Pith. sign in

Paper Citation Record · LEDGER

MEGL: Multimodal Explanation-Guided Learning

As of 13 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 1 inbound Pith citation observation for arXiv:2411.13053.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.13053 v1

Coverage vector

measured 62 of 62 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T16:56:11.028503Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T06:28:45.967569Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

62 of 62 outbound references displayed

  • verified exact4
  • verified fuzzy25
  • unresolved33
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 15353b77-9544-449e-a4cb-8f1c24276c75 · outbound

This paper cites GPT-4 Technical Report.

MEGL: Multimodal Explanation-Guided Learning GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.660284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.660284Z digest=sha256:42b923c71c75e0ada00ea6e049ddda4a61d77162718be5d37c6536c4222a0ad5

Observation 7bc41f5c-d495-4dd5-b73c-28586f8d5158 · outbound

This paper cites Tbexplain: A text-based explanation method for scene classification models with the statistical prediction correction.

MEGL: Multimodal Explanation-Guided Learning Tbexplain: A text-based explanation method for scene classification models with the statistical prediction correction

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:12.283009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.666878Z digest=sha256:d21de4108fe780bdfd1b9d187cf3de3041173d93cc0dc2af604a695e0abce9a1

Observation c2879fd1-ee4f-4a37-9ce1-23997739be21 · outbound

This paper cites Spice: Semantic propositional image cap- tion evaluation.

MEGL: Multimodal Explanation-Guided Learning Spice: Semantic propositional image cap- tion evaluation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.672926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.672926Z digest=sha256:d804a8c00b9f15b016f4ccb32ea05e747e69fba029c706667c248084dcc3505a

Observation 85ae0cc5-ccd5-4c48-8a86-e94b200db6ec · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

MEGL: Multimodal Explanation-Guided Learning Gemini: A Family of Highly Capable Multimodal Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.678779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.678779Z digest=sha256:39f11090b03ed02f3d6fc7a0900b6f2d7385f69f532254a33e7b10248d6038ee

Observation 7cbe616d-b9f4-492e-8ad3-60cf4d072cb0 · outbound

This paper cites Explainable artificial intelligence (xai): Concepts, taxonomies, opportunities and challenges toward responsible ai.

MEGL: Multimodal Explanation-Guided Learning Explainable artificial intelligence (xai): Concepts, taxonomies, opportunities and challenges toward responsible ai

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:12.251668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.684903Z digest=sha256:c88de1b0c87665ed12983e05826ab56416d97ca848576ddf6cf59478a50e9cb2

Observation 899117b8-c0e6-4a98-bbd2-124306b764f5 · outbound

This paper cites Meteor: An automatic metric for mt evaluation with improved correlation with hu- man judgments.

MEGL: Multimodal Explanation-Guided Learning Meteor: An automatic metric for mt evaluation with improved correlation with hu- man judgments

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.690613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.690613Z digest=sha256:36c95f07ade2eb69eaba3f7c83190026f7afbc7243d03e588653c60170feef77

Observation c13859aa-b564-466c-bdb5-1a309aa9c406 · outbound

This paper cites Let there be a clock on the beach: Reducing object halluci- nation in image captioning.

MEGL: Multimodal Explanation-Guided Learning Let there be a clock on the beach: Reducing object halluci- nation in image captioning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.696412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.696412Z digest=sha256:10e63af34d0e74980ab246ba42cc543b8858b6f3c2e681a6aa784984ec02902d

Observation 47983ac8-7e1d-4da0-8261-7d561c1babbe · outbound

This paper cites A survey on xai and nat- ural language explanations.

MEGL: Multimodal Explanation-Guided Learning A survey on xai and nat- ural language explanations

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:12.205236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.702334Z digest=sha256:fe8a31179856254dd44d352830f13431804c17177d3e398bb6359f5ae1cc24ff

Observation 146450f3-937b-4809-8650-eb4620a0957f · outbound

This paper cites Machine learning interpretability: A survey on methods and metrics.

MEGL: Multimodal Explanation-Guided Learning Machine learning interpretability: A survey on methods and metrics

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:12.187087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.708048Z digest=sha256:e14876b2bb965d031bf6a7780b510262b9628ab5eaaf129a3c8015fc79504695

Observation 037c5f3f-1d90-41af-b474-af404cbb615f · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

MEGL: Multimodal Explanation-Guided Learning An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.714945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.714945Z digest=sha256:68d537585be654d75f73b4a6cf1c7ba2261d388c47e4ba78ea27cd4e56ac47d4

Observation 81534ad3-e0c2-4a00-96fd-896ba1c66cc3 · outbound

This paper cites Techniques for in- terpretable machine learning.

MEGL: Multimodal Explanation-Guided Learning Techniques for in- terpretable machine learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:12.162112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.721914Z digest=sha256:665729a7aa2cb096c687d30d636ac9e1e021a752813272f19f5d9cc648c4725f

Observation 5c894918-599e-46b5-b109-6de32f0c9d50 · outbound

This paper cites Learn- ing credible deep neural networks with rationale regulariza- tion.

MEGL: Multimodal Explanation-Guided Learning Learn- ing credible deep neural networks with rationale regulariza- tion

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:12.137278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.727380Z digest=sha256:7a820dabcd762a9d84ba81180067f338b9854b1db0b5e2adf9b0c62ee6246741

Observation 2b78c56b-0dac-48a6-897f-7cd57aa8e746 · outbound

This paper cites Attention branch network: Learning of attention mechanism for visual explanation.

MEGL: Multimodal Explanation-Guided Learning Attention branch network: Learning of attention mechanism for visual explanation

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:12.113801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.734558Z digest=sha256:c322ac5b62ef7b63059028b457e6a1c311ed8c24507834386d07b303b7a97e0c

Observation 56e56819-2781-4067-adbc-30e54b80ca64 · outbound

This paper cites Res: A robust framework for guiding visual explanation.

MEGL: Multimodal Explanation-Guided Learning Res: A robust framework for guiding visual explanation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:12.092452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.742038Z digest=sha256:6ee038cba06063da564e4fa9c7b03abd1bca92fa3c4bd57cbbc48723ec1fa94e

Observation 1dcb86e1-9f93-434b-8216-1e828afeccba · outbound

This paper cites Going beyond xai: A system- atic survey for explanation-guided learning.

MEGL: Multimodal Explanation-Guided Learning Going beyond xai: A system- atic survey for explanation-guided learning

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:12.069196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.748124Z digest=sha256:2ef7677014edeedf5e53e754666ce1ba26f51b0bdc6ae87a7706abdda4ddc77e

Observation 98b98a3c-287c-483d-b648-635f1c4001c6 · outbound

This paper cites Don't trust your eyes: on the (un)reliability of feature visualizations.

MEGL: Multimodal Explanation-Guided Learning Don't trust your eyes: on the (un)reliability of feature visualizations

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.753790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.753790Z digest=sha256:163ed61cabbdb8007a54a0c5fc57abb6ef3ea7ca86e36d8f5ddd7654a1e22866

Observation 2adce892-e148-4d9e-8697-cc188f5c7927 · outbound

This paper cites Essa: Explanation iterative supervision via saliency-guided data augmentation.

MEGL: Multimodal Explanation-Guided Learning Essa: Explanation iterative supervision via saliency-guided data augmentation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:12.047347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.759684Z digest=sha256:3c3a01302d221a86e2f84a42e7f668ee9ef3c77100a88fe94253ced6cb1837df

Observation 59d4c10b-5fa6-4c89-8878-19e29c5c76f5 · outbound

This paper cites XAI-CLASS: Explanation-Enhanced Text Classification with Extremely Weak Supervision.

MEGL: Multimodal Explanation-Guided Learning XAI-CLASS: Explanation-Enhanced Text Classification with Extremely Weak Supervision

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-12T16:56:11.467685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.765942Z digest=sha256:3370f7eaf50e82d5cc1ce0e285553b1859eb9c377f686628437c40af7ed5cfd0

Observation 4a7d89d4-2ae6-44bd-b453-e73740ec871f · outbound

This paper cites Generating Faithful and Salient Text from Multimodal Data.

MEGL: Multimodal Explanation-Guided Learning Generating Faithful and Salient Text from Multimodal Data

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-12T16:56:11.439452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.771646Z digest=sha256:3c59b451cb286362350ac7020ed3ac6e113aa50c837655f84164d776c62c1360

Observation e8d0caaf-df93-42d4-82a6-ad7854bcfee4 · outbound

This paper cites Deep residual learning for image recognition.

MEGL: Multimodal Explanation-Guided Learning Deep residual learning for image recognition

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:12.024021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.777520Z digest=sha256:d40e03c53e52cf79893b48ea3b8639daf5cda8b9eb94faab6cc6c803591e4a00

Observation 602e1ae9-3e55-452f-9dc1-5b8b856d4ebf · outbound

This paper cites Generating vi- sual explanations.

MEGL: Multimodal Explanation-Guided Learning Generating vi- sual explanations

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:12.003327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.783772Z digest=sha256:1d42b446bb0f0904a890e6905804c19ccbf6d383b9fe754f3bd7842118e60826

Observation ebc38e85-ddbb-421a-9b7f-b89d83ba596a · outbound

This paper cites CLIPScore: A Reference-free Evaluation Metric for Image Captioning.

MEGL: Multimodal Explanation-Guided Learning CLIPScore: A Reference-free Evaluation Metric for Image Captioning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.790044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.790044Z digest=sha256:0ab869dca6d2e6daa276a92c202235de82a65e7805b51dbef8395cd1a72b500e

Observation 6f254c4a-8c19-4c69-94de-478503745cfa · outbound

This paper cites Large Language Models Are Reasoning Teachers.

MEGL: Multimodal Explanation-Guided Learning Large Language Models Are Reasoning Teachers

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.796057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.796057Z digest=sha256:bb8491113fddf7cecf2b3a2ca292cec33bad56caccdca9a85c92c3a0ecfca89b

Observation 4040d0eb-80ed-45ee-9dec-41b38168d5b6 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

MEGL: Multimodal Explanation-Guided Learning LoRA: Low-Rank Adaptation of Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.802153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.802153Z digest=sha256:9c23f6a2ba56f138a339d438f04bbbcb4be832c1263f074bca96a4cb0cea43c6

Observation a1f7f630-1207-4b51-856b-69c96ac39fa8 · outbound

This paper cites Bliva: A simple multimodal llm for better handling of text-rich visual questions.

MEGL: Multimodal Explanation-Guided Learning Bliva: A simple multimodal llm for better handling of text-rich visual questions

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:11.985871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.808892Z digest=sha256:7f5c2044b21c6b8c54412748a84a25c846aa0342917695433e74145774c0394e

Observation 3eae2815-0a6b-4f7d-a3dc-bdc6484c5a4d · outbound

This paper cites Survey of hallucination in natural language generation.

MEGL: Multimodal Explanation-Guided Learning Survey of hallucination in natural language generation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.815283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.815283Z digest=sha256:15a7c09112e4a02b8d44c7239ccec6930d1877a3e072f14c9794aca6aabc5e94

Observation 3ed97673-8634-4388-83d0-6c2f2545e45b · outbound

This paper cites Hallucination augmented contrastive learn- ing for multimodal large language model.

MEGL: Multimodal Explanation-Guided Learning Hallucination augmented contrastive learn- ing for multimodal large language model

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.821291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.821291Z digest=sha256:793e5d69931b752e834598a31418335313e5f9f2d0425c6d37c546cf2036dfc4

Observation 6a51ba39-de51-4193-b766-2d78bd219634 · outbound

This paper cites Deep learning.

MEGL: Multimodal Explanation-Guided Learning Deep learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.826822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.826822Z digest=sha256:70b443f404688a0527bec9da0a120d3c96e4af088ec13ea066b160b0d5b8994f

Observation 83afcede-b744-43ed-9f0e-210b7b8556ca · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

MEGL: Multimodal Explanation-Guided Learning Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.834757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.834757Z digest=sha256:e44570a92687d766191d7196e5672a58e0b31e99d0037008bf68c39dd7d1f61a

Observation 122bbb4a-7801-47f5-9096-3a3bda804545 · outbound

This paper cites Symbolic Chain-of-Thought Distillation: Small Models Can Also "Think" Step-by-Step.

MEGL: Multimodal Explanation-Guided Learning Symbolic Chain-of-Thought Distillation: Small Models Can Also "Think" Step-by-Step

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.840711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.840711Z digest=sha256:c5d8f349c295027a4071f84064c0062f10f92340b23ea710db816a1f77c48712

Observation df77f620-7aea-43c3-b353-2b6e26c8d8fb · outbound

This paper cites Vqa-e: Explaining, elaborating, and enhancing your answers for visual questions.

MEGL: Multimodal Explanation-Guided Learning Vqa-e: Explaining, elaborating, and enhancing your answers for visual questions

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:11.897684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.848190Z digest=sha256:f1870ca15d0336b4b721064fc4825a5fb57354b0b51e432047399a6ca8d8cf0f

Observation 9e4357b6-e46c-4471-9d75-c0509a676a57 · outbound

This paper cites Explanations from Large Language Models Make Small Reasoners Better.

MEGL: Multimodal Explanation-Guided Learning Explanations from Large Language Models Make Small Reasoners Better

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.853956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.853956Z digest=sha256:f290699726ba661d3728143118e4a49598120038fc2de693526723ff9865fbac

Observation 22fa143a-0efc-4d6e-bb66-03f38bea3a63 · outbound

This paper cites Rouge: A package for automatic evaluation of summaries.

MEGL: Multimodal Explanation-Guided Learning Rouge: A package for automatic evaluation of summaries

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.859671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.859671Z digest=sha256:84a3d33ff2a2feb4b4e78b85e574e7d3f8b0dd75dd2cdf9f98a6a1454ae0b77c

Observation 76fc5d5c-7857-4505-be5f-ac1d1a78d020 · outbound

This paper cites Visual instruction tuning.

MEGL: Multimodal Explanation-Guided Learning Visual instruction tuning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.865125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.865125Z digest=sha256:09ca4e37ae75d12968587834220a9fbfefdceb41c5b8a7de118d772be0c1e1aa

Observation 6052a327-dbc4-4348-9e1d-73070e0366ab · outbound

This paper cites A Unified Approach to Interpreting Model Predictions.

MEGL: Multimodal Explanation-Guided Learning A Unified Approach to Interpreting Model Predictions

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.870584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.870584Z digest=sha256:aa8122eb405fa483fc4b9ad4f0a1516bf374a5f63c00b183f59c88a6534548aa

Observation 56ec39f9-2eb3-41e4-a6e5-82335b3ca43f · outbound

This paper cites Teaching Small Language Models to Reason.

MEGL: Multimodal Explanation-Guided Learning Teaching Small Language Models to Reason

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.876678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.876678Z digest=sha256:d384a7a7a0a2b7ffdcaf69aec786b878e240789c6b22f4dc8b0d1d84740adb1a

Observation 1e715a4a-d54c-47a3-bfcb-b2e6ff77c77d · outbound

This paper cites Natural Language Rationales with Full-Stack Visual Reasoning: From Pixels to Semantic Frames to Commonsense Graphs.

MEGL: Multimodal Explanation-Guided Learning Natural Language Rationales with Full-Stack Visual Reasoning: From Pixels to Semantic Frames to Commonsense Graphs

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-08-12T16:56:11.267610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.883474Z digest=sha256:a31e367f60ff795460f558e5b5236ff380726a2221d2da796711496b3cd78330

Observation c352a93d-f575-4683-8bf0-48133222af74 · outbound

This paper cites VALE: A Multimodal Visual and Language Explanation Framework for Image Classifiers using eXplainable AI and Language Models.

MEGL: Multimodal Explanation-Guided Learning VALE: A Multimodal Visual and Language Explanation Framework for Image Classifiers using eXplainable AI and Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.890450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.890450Z digest=sha256:de7e38d3129313cf354a0d159aa76cc668b11e735f33339b93948366df044cb5

Observation 30dbf13b-3c89-4e29-9d18-8e048b0419ca · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

MEGL: Multimodal Explanation-Guided Learning Bleu: a method for automatic evaluation of machine translation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.896773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.896773Z digest=sha256:127bfa720e8a8b61bca5675a94d246548f5a4d828f06c9d3a6bea3281aa29e16

Observation 13fe9371-f332-4959-a8b8-41f54d512a0b · outbound

This paper cites Multimodal explanations: Justifying deci- sions and pointing to the evidence.

MEGL: Multimodal Explanation-Guided Learning Multimodal explanations: Justifying deci- sions and pointing to the evidence

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:11.842324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.902344Z digest=sha256:393567801eea994489e253939b1ce2e2e71524f3e7435cfab52ff0d34f7fdc99

Observation c40bf59f-cab8-4229-a7bd-67ad88f0e56d · outbound

This paper cites Ro- bust explanations for visual question answering.

MEGL: Multimodal Explanation-Guided Learning Ro- bust explanations for visual question answering

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:11.819499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.909821Z digest=sha256:0b16fcdc73b2530577f89464ce1b8538d4cd3d67e4211b98f54c73ae471f040e

Observation 95c2a448-9cdd-4635-a578-3b026d7ae609 · outbound

This paper cites RISE: Randomized Input Sampling for Explanation of Black-box Models.

MEGL: Multimodal Explanation-Guided Learning RISE: Randomized Input Sampling for Explanation of Black-box Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.916178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.916178Z digest=sha256:d3f6c26433a10fc174a719b67a1a0645898088a638d61f020bd31119a4adc3d4

Observation fc169658-621c-4730-a695-2163b296de94 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

MEGL: Multimodal Explanation-Guided Learning Learning transferable visual models from natural language supervi- sion

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.922390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.922390Z digest=sha256:9434621a6c1ea647b4786804be7427cab420bb21b1a2ee5171d60af72d450dc5

Observation b20adb0b-4d6c-44ad-aa25-2dea72338326 · outbound

This paper cites A First Look: Towards Explainable TextVQA Models via Visual and Textual Explanations.

MEGL: Multimodal Explanation-Guided Learning A First Look: Towards Explainable TextVQA Models via Visual and Textual Explanations

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.927640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.927640Z digest=sha256:f1f1a95d53d025d512fd12cfd86c14f592478e4fa68ed13d9f86f98041cfc379

Observation a2f4a05d-5fc4-444a-8401-7d5effe0365a · outbound

This paper cites Interpretations are useful: penalizing explanations to align neural networks with prior knowledge.

MEGL: Multimodal Explanation-Guided Learning Interpretations are useful: penalizing explanations to align neural networks with prior knowledge

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:11.784062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.934197Z digest=sha256:37ce6be64e7716f5fb794747d251a808271ef6ea984aebb7dc043840a29147cc

Observation 4e6b52c1-2b6f-4c3d-81fb-1679a94ef007 · outbound

This paper cites Grad-CAM: Why did you say that?.

MEGL: Multimodal Explanation-Guided Learning Grad-CAM: Why did you say that?

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.939679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.939679Z digest=sha256:966c4baa4c6867becfc91ca62d585400b7a1827e29beaf2079537942c78d3afd

Observation 26b83f98-3f27-4dd7-a827-dc7ff3e54edd · outbound

This paper cites 10 Grad-cam: Visual explanations from deep networks via gradient-based localization.

MEGL: Multimodal Explanation-Guided Learning 10 Grad-cam: Visual explanations from deep networks via gradient-based localization

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:11.764337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.945055Z digest=sha256:641d8cbffd288ab254b6ad1e298c6d22262c2bc680df3011066c0ac2d65cdd6d

Observation 0a306b1d-06ae-4e15-9e91-67bf3d10a69c · outbound

This paper cites Human-ai interactive and continuous sensemaking: A case study of image classification using scribble attention maps.

MEGL: Multimodal Explanation-Guided Learning Human-ai interactive and continuous sensemaking: A case study of image classification using scribble attention maps

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:11.738853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.950683Z digest=sha256:e49e057106507eb3e41a9ebc19a9f99cada2006587030af2a0a86ad6e66ca57d

Observation ea4f856a-b39e-4fa9-8f67-64adee87dcbd · outbound

This paper cites A review of taxonomies of explainable ar- tificial intelligence (xai) methods.

MEGL: Multimodal Explanation-Guided Learning A review of taxonomies of explainable ar- tificial intelligence (xai) methods

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:11.715500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.955766Z digest=sha256:26a8824cf1d80350f44d2ad2d93f39485ec2d56fde6546bfe0f7728f1f4a411e

Observation 47b93eac-a449-49e2-86ea-6d7f6c393203 · outbound

This paper cites Axiomatic attribution for deep networks.

MEGL: Multimodal Explanation-Guided Learning Axiomatic attribution for deep networks

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.961575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.961575Z digest=sha256:1cdf8667285e9e2a436f6509a0863a1c3f33d8d42fc04a2a43db4aad3de4666d

Observation 5d3324e8-4fbf-4783-9765-740e1caff284 · outbound

This paper cites Robustness May Be at Odds with Accuracy.

MEGL: Multimodal Explanation-Guided Learning Robustness May Be at Odds with Accuracy

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.967542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.967542Z digest=sha256:d00c46e3fa8d15a8dfd7e365ac39fbd1682d9798be17ad50439d724e9a18585c

Observation 4f7ba613-df0a-45b4-b7a7-b434be152d0b · outbound

This paper cites Explainable artificial intel- ligence (xai) in deep learning-based medical image analysis.

MEGL: Multimodal Explanation-Guided Learning Explainable artificial intel- ligence (xai) in deep learning-based medical image analysis

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:11.676349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.973194Z digest=sha256:c430b52121b77f3ecbef608fd46cf697efd1082629c57dee3fd5f040d60b2c32

Observation e5034e88-8d5f-4f70-abd9-e7aa8747268a · outbound

This paper cites Cider: Consensus-based image description evalua- tion.

MEGL: Multimodal Explanation-Guided Learning Cider: Consensus-based image description evalua- tion

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.978628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.978628Z digest=sha256:1aaa5faef13103ff873a862c347f232f0f6912291ef91575efa2a2d5a6f8e21e

Observation ee4ee791-962b-4d0f-942f-74020e33c9b6 · outbound

This paper cites Faithful Multimodal Explanation for Visual Question Answering.

MEGL: Multimodal Explanation-Guided Learning Faithful Multimodal Explanation for Visual Question Answering

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:10.983765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:10.983765Z digest=sha256:12a1d47c08b706e0154fabff2a06c6349e7230aff34bb41ecee251ea54733a7a

Observation 22389f54-2720-41f5-87f6-600ed372e1df · outbound

This paper cites Generating deep networks explanations with ro- bust attribution alignment.

MEGL: Multimodal Explanation-Guided Learning Generating deep networks explanations with ro- bust attribution alignment

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:11.645013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.989040Z digest=sha256:7822dd86eee35313b431d64afd43ac4df0e9fe0765c1b77967aca2ff9fc31805

Observation 3385db47-7dfb-4201-9f7f-fa0d384329c2 · outbound

This paper cites Overlooked trustworthiness of saliency maps.

MEGL: Multimodal Explanation-Guided Learning Overlooked trustworthiness of saliency maps

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:11.625479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.994196Z digest=sha256:29a5184e9d67bcf6d46e8f1181ae8b09c493cf4d64fe4b005299eed295cd49ea

Observation 9009f5c6-94ed-4024-a78b-d3704f8243e8 · outbound

This paper cites Rationale- augmented convolutional neural networks for text classifica- tion.

MEGL: Multimodal Explanation-Guided Learning Rationale- augmented convolutional neural networks for text classifica- tion

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:11.602286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:10.999277Z digest=sha256:16f1e058ec0549fc73c6bf725d3829bd2d172097817256eef4d0a0ad8656f5f3

Observation 7a6b2583-3e64-4014-bb1e-e766050e673c · outbound

This paper cites Magi: Multi-annotated explanation-guided learning.

MEGL: Multimodal Explanation-Guided Learning Magi: Multi-annotated explanation-guided learning

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:56:11.580092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:11.004590Z digest=sha256:f1ad510634389f48a76c989f7af3adff0474bde8bf9f45033211c1c034221031

Observation 8bf2ff17-5b71-4991-baa1-284d17bae969 · outbound

This paper cites Multimodal Chain-of-Thought Reasoning in Language Models.

MEGL: Multimodal Explanation-Guided Learning Multimodal Chain-of-Thought Reasoning in Language Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:11.010184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:11.010184Z digest=sha256:3e9933614fb9a6922c3450c5c5894674378c4e357d0a4b2ad62f47dee686a3cc

Observation e1490300-7fa8-440c-8905-86b3073586b7 · outbound

This paper cites Large Language Models are In-context Teachers for Knowledge Reasoning.

MEGL: Multimodal Explanation-Guided Learning Large Language Models are In-context Teachers for Knowledge Reasoning

Reference 60

Resolution
verified exact
local_arxiv, observed 2026-08-12T16:56:11.102813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T16:56:11.016101Z digest=sha256:31e356cf539dd5676eee7e1673172cff9c600a4e98e2fe4b4ee5429df4b85f6d

Observation 7c2640ea-c721-494c-ad29-d2389b7c5fed · outbound

This paper cites Judging llm-as-a-judge with mt-bench and chatbot arena.Advances in Neural Information Processing Systems, 36:46595–46623, 2023.

MEGL: Multimodal Explanation-Guided Learning Judging llm-as-a-judge with mt-bench and chatbot arena.Advances in Neural Information Processing Systems, 36:46595–46623, 2023

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:11.022673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:11.022673Z digest=sha256:2f27c4d7886aa56b05f553a84c3f8e470ade2e82c3961c820e7952f8c623d2a9

Observation 17642f7c-15d3-42c4-8bf7-dba8934526a6 · outbound

This paper cites Analyzing and Mitigating Object Hallucination in Large Vision-Language Models.

MEGL: Multimodal Explanation-Guided Learning Analyzing and Mitigating Object Hallucination in Large Vision-Language Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-12T16:56:11.028503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:56:11.028503Z digest=sha256:69fe0d47a8fd7e7781f68e44dacfb1d969f209a8ca4014b609506a54640dea0e

Pith citing papers

Observation 59674582-494d-4957-a588-f117f2b3c38e · inbound

Where Not to Learn: Prior-Aligned Training with Subset-based Attribution Constraints for Reliable Decision-Making cites this paper.

Where Not to Learn: Prior-Aligned Training with Subset-based Attribution Constraints for Reliable Decision-Making MEGL: Multimodal Explanation-Guided Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T06:28:45.967569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:28:45.967569Z digest=sha256:2ad4e18e12a10544200ed62e7c5044d11c47ff1ac08949ef4131149ce6e1f5b3