Pith. sign in

Paper Citation Record · LEDGER

LLaVA-MR: Large Language-and-Vision Assistant for Video Moment Retrieval

As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2411.14505.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.14505 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:32:52.859271Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T17:27:14.988080Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a8c5eb12-65ec-487d-8628-b0f2bed9b148 · inbound

DisTime: Distribution-based Time Representation for Video Large Language Models cites this paper.

DisTime: Distribution-based Time Representation for Video Large Language Models LLaVA-MR: Large Language-and-Vision Assistant for Video Moment Retrieval

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:32:52.859271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:32:52.859271Z digest=sha256:2c0b8dbe3ae410c7a6c5ec60dc16c02f08fac495ff95cf8d202a9e836d477e7a

Observation 34b3b1a3-7981-425a-a102-7663f633fb81 · inbound

Sparse-Dense Side-Tuner for efficient Video Temporal Grounding cites this paper.

Sparse-Dense Side-Tuner for efficient Video Temporal Grounding LLaVA-MR: Large Language-and-Vision Assistant for Video Moment Retrieval

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:02.771931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:02.771931Z digest=sha256:bd276e678e86896d4a24737412b926604319aa53e37c1cf4c34515b81c3f05fc

Observation e4436d2b-52f0-4d91-9374-14e51158659b · inbound

A Survey on Video Temporal Grounding with Multimodal Large Language Model cites this paper.

A Survey on Video Temporal Grounding with Multimodal Large Language Model LLaVA-MR: Large Language-and-Vision Assistant for Video Moment Retrieval

Reference 106

Resolution
unresolved
no resolver link, observed 2026-08-05T23:32:17.938393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:32:17.938393Z digest=sha256:73a44373bbdbb9e65b9f77c5156ef79d1d5c992bad1ca5b23851b3520d9f80eb

Observation 66daadb5-120d-4ba9-985d-3d3e1ae18f45 · inbound

SMART: Shot-Aware Multimodal Video Moment Retrieval with Audio-Enhanced MLLM cites this paper.

SMART: Shot-Aware Multimodal Video Moment Retrieval with Audio-Enhanced MLLM LLaVA-MR: Large Language-and-Vision Assistant for Video Moment Retrieval

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T21:42:45.907432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T21:42:45.907432Z digest=sha256:c6e765747a2e09d91d4fcca0bd427246074740feea6218582a8311499051fb74

Observation 4b3d4251-7538-4694-85e7-b6fa0d1c2d9c · inbound

SiMing-Bench: Evaluating Procedural Correctness from Continuous Interactions in Clinical Skill Videos cites this paper.

SiMing-Bench: Evaluating Procedural Correctness from Continuous Interactions in Clinical Skill Videos LLaVA-MR: Large Language-and-Vision Assistant for Video Moment Retrieval

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:30:58.574381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-10T17:37:40.373211Z digest=sha256:8e096f9199a8bbe9662b86fbb8bc2a980062d51796bb6757c511fdf7d6f4d05a

Observation 905d3179-9bae-429f-9fb5-bafc0ad4fb4e · inbound

Towards One-to-Many Temporal Grounding cites this paper.

Towards One-to-Many Temporal Grounding LLaVA-MR: Large Language-and-Vision Assistant for Video Moment Retrieval

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:16:57.671960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-06-28T02:11:48.455492Z digest=sha256:419d1ddd439b96fbbde956803a155610dcdcd04e832bebd9306df2d1fa1dc902

Observation 0b2b1842-8324-47dd-9e39-7e9d9215ef89 · inbound

Watch, Remember, Reason: Human-View Video Understanding with MLLMs cites this paper.

Watch, Remember, Reason: Human-View Video Understanding with MLLMs LLaVA-MR: Large Language-and-Vision Assistant for Video Moment Retrieval

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:27:14.990641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-27T22:00:28.350003Z digest=sha256:6ed6548caf130c0a1f702f009d68df412735deca5d98bd97599e7c4141d78b7f