Pith. sign in

Paper Citation Record · LEDGER

Temporal Working Memory: Query-Guided Segment Refinement for Enhanced Multimodal Understanding

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2502.06020.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.06020 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:50:40.666741Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-19T14:12:21.927254Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1eaadc4b-f034-40b9-b47f-ab33f38d371f · inbound

Music Audio-Visual Question Answering Requires Specialized Multimodal Designs cites this paper.

Music Audio-Visual Question Answering Requires Specialized Multimodal Designs Temporal Working Memory: Query-Guided Segment Refinement for Enhanced Multimodal Understanding

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-19T14:12:21.932953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T14:11:52.011626Z digest=sha256:5d3f0dc801aa7228cc619813ae4a33b86223d7db5702cfa02c289ed3a9ffb4b7

Observation d3ac8f25-d2ca-464d-974d-b906056695fb · inbound

Learning Sparsity for Effective and Efficient Music Performance Question Answering cites this paper.

Learning Sparsity for Effective and Efficient Music Performance Question Answering Temporal Working Memory: Query-Guided Segment Refinement for Enhanced Multimodal Understanding

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:50:40.666741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:50:40.666741Z digest=sha256:70c50c6687f6f5956e98f8a153b7d3c6840b48a6ec15e4e790586a74ec5dcd44

Observation 1e268c2b-dda3-4adb-ae2f-9bc1235da902 · inbound

FakeSV-VLM: Taming VLM for Detecting Fake Short-Video News via Progressive Mixture-Of-Experts Adapter cites this paper.

FakeSV-VLM: Taming VLM for Detecting Fake Short-Video News via Progressive Mixture-Of-Experts Adapter Temporal Working Memory: Query-Guided Segment Refinement for Enhanced Multimodal Understanding

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T15:41:50.873521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:41:50.873521Z digest=sha256:a2ab6e7d846dfe7fcc617340678cd135f4f584ab0c38ae90bd7ca37098b9501e

Observation b8fddfc9-11a4-477f-9b31-f006d6e697db · inbound

GDLLM: A Global Distance-aware Modeling Approach Based on Large Language Models for Event Temporal Relation Extraction cites this paper.

GDLLM: A Global Distance-aware Modeling Approach Based on Large Language Models for Event Temporal Relation Extraction Temporal Working Memory: Query-Guided Segment Refinement for Enhanced Multimodal Understanding

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T14:51:28.377811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:51:28.377811Z digest=sha256:18233a5783bb17994a93c28e0ef6dbe047e37ff2d83584a5c9ef61bd2bbfa9f1

Observation fead5bf8-9fc4-4c25-a540-c56a9117b3f8 · inbound

Multi-Modal Machine Learning Framework for Predicting Early Recurrence of Brain Tumors Using MRI and Clinical Biomarkers cites this paper.

Multi-Modal Machine Learning Framework for Predicting Early Recurrence of Brain Tumors Using MRI and Clinical Biomarkers Temporal Working Memory: Query-Guided Segment Refinement for Enhanced Multimodal Understanding

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T12:53:01.820900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:53:01.820900Z digest=sha256:a2cc4ab7aa00d3660022e5b1e1319709c8fdfb168960af8f3a50a43790f377c3

Observation a4f4a1e9-f92b-473d-927b-b41d5be5d383 · inbound

A Multimodal Deep Learning Framework for Early Diagnosis of Liver Cancer via Optimized BiLSTM-AM-VMD Architecture cites this paper.

A Multimodal Deep Learning Framework for Early Diagnosis of Liver Cancer via Optimized BiLSTM-AM-VMD Architecture Temporal Working Memory: Query-Guided Segment Refinement for Enhanced Multimodal Understanding

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T12:53:14.693916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:53:14.693916Z digest=sha256:1a10287a64db0fad1e276553555f0c144610940e886f496780aec356c7205784

Observation b107f77a-1c8a-42a2-9d10-8475833fc3d6 · inbound

OpenVision 2: A Family of Generative Pretrained Visual Encoders for Multimodal Learning cites this paper.

OpenVision 2: A Family of Generative Pretrained Visual Encoders for Multimodal Learning Temporal Working Memory: Query-Guided Segment Refinement for Enhanced Multimodal Understanding

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T12:22:32.652497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:22:32.652497Z digest=sha256:c3df8d11f5c6b6bab3440f20b59c8ec0451d32ed40f15b87e9af5fe809c899e7

Observation 0156f84b-e68e-4724-b256-7f155f5145a3 · inbound

Language-Guided Long Horizon Manipulation with LLM-based Planning and Visual Perception cites this paper.

Language-Guided Long Horizon Manipulation with LLM-based Planning and Visual Perception Temporal Working Memory: Query-Guided Segment Refinement for Enhanced Multimodal Understanding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T11:40:42.676909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:40:42.676909Z digest=sha256:8c5a175fa109dacd8b98c613c7b015aff2a6a9f51a0b4a293b9feadfffc71116