Pith. sign in

Paper Citation Record · LEDGER

Rec-GPT4V: Multimodal Recommendation with Large Vision-Language Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2402.08670.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.08670 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:29:14.045791Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T03:47:35.690878Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0433936a-d9d1-4d8a-a280-72766676c584 · inbound

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing cites this paper.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Rec-GPT4V: Multimodal Recommendation with Large Vision-Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:14.045791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:14.045791Z digest=sha256:a9d94e90447386050bbf1f6c583a8419e7f93448f54701f3b578e585756f03cf

Observation 9d33fdb6-dbd7-4f14-85b8-8af1915b79ae · inbound

RAG-VisualRec: An Open Resource for Vision- and Text-Enhanced Retrieval-Augmented Generation in Recommendation cites this paper.

RAG-VisualRec: An Open Resource for Vision- and Text-Enhanced Retrieval-Augmented Generation in Recommendation Rec-GPT4V: Multimodal Recommendation with Large Vision-Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:53.255664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:53.255664Z digest=sha256:d436696a8db5252817d1ac16c3128bd6c81d2d6fffc0dcb831eef9525deab6f3

Observation bf739d6f-05fd-40e5-bf9b-2567a8531986 · inbound

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation cites this paper.

ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation Rec-GPT4V: Multimodal Recommendation with Large Vision-Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T00:52:20.947043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:52:20.947043Z digest=sha256:f69dc06e4d09c2e391a975eb5f4677ac296319792af4207646dee432af3a8f99

Observation f3e2742e-08d1-44f6-a583-a34889447b78 · inbound

A Survey on Generative Recommendation: Data, Model, and Tasks cites this paper.

A Survey on Generative Recommendation: Data, Model, and Tasks Rec-GPT4V: Multimodal Recommendation with Large Vision-Language Models

Reference 111

Resolution
verified exact
arxiv_id, observed 2026-05-18T03:50:51.982452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T03:47:08.208082Z digest=sha256:edd5656041f09c78a91c3c3e3d9e9a135e505abbf2b17ed4b3063a830f08a4c3

Observation 2871a6b5-c309-4703-a1dc-e008b13dc720 · inbound

Timeripple: Accelerating vDiTs by Understanding the Spatio-Temporal Correlations in Latent Space cites this paper.

Timeripple: Accelerating vDiTs by Understanding the Spatio-Temporal Correlations in Latent Space Rec-GPT4V: Multimodal Recommendation with Large Vision-Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T22:12:44.510309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:12:44.510309Z digest=sha256:1cb85fc32be82ef7018bf923ce889e02a15241c9095ceb9dee5fb8abc600e108

Observation 7c9e4ad9-71ea-4b6b-9a04-4448460489cd · inbound

Multimodal Large Language Models with Adaptive Preference Optimization for Sequential Recommendation cites this paper.

Multimodal Large Language Models with Adaptive Preference Optimization for Sequential Recommendation Rec-GPT4V: Multimodal Recommendation with Large Vision-Language Models

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-17T06:19:09.826762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T06:16:53.140133Z digest=sha256:d001df5cc464fc5c9332bc302877d219c540e5f79e6ecba9edddd0ff09b62259

Observation 34a2ac48-1407-46ee-a8bf-d52b7aec0837 · inbound

Frozen LVLMs for Micro-Video Recommendation: A Systematic Study of Feature Extraction and Fusion cites this paper.

Frozen LVLMs for Micro-Video Recommendation: A Systematic Study of Feature Extraction and Fusion Rec-GPT4V: Multimodal Recommendation with Large Vision-Language Models

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:11:13.440417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T20:09:45.442172Z digest=sha256:d3a28d58c33af786c4e679be9a566b6e774a734cc74bdd6ad1b612628fa8800a

Observation 6755ffc3-da50-47f0-834b-b0a193a223ef · inbound

Multimodal Generative Engine Optimization: Rank Manipulation for Vision-Language Model Rankers cites this paper.

Multimodal Generative Engine Optimization: Rank Manipulation for Vision-Language Model Rankers Rec-GPT4V: Multimodal Recommendation with Large Vision-Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T09:54:29.234625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:54:29.234625Z digest=sha256:7b87872516cb3a61785ed0c45ee176364f0c1123d4ecd88126ac8bea1941dea8

Observation 46d7228a-5ada-4a28-8e4b-1b4bea9765fa · inbound

RecGOAT: Graph Optimal Adaptive Transport for LLM-Enhanced Multimodal Recommendation with Dual Semantic Alignment cites this paper.

RecGOAT: Graph Optimal Adaptive Transport for LLM-Enhanced Multimodal Recommendation with Dual Semantic Alignment Rec-GPT4V: Multimodal Recommendation with Large Vision-Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T06:01:36.678015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:01:36.678015Z digest=sha256:6b0421d7a1d95a4dd00f788fc8e207e6748aeceea4315c66d4a91cfc4b098ecd

Observation e17f65a3-4179-40d0-9524-8d096f297189 · inbound

TimeMM: Time-as-Operator Spectral Filtering for Dynamic Multimodal Recommendation cites this paper.

TimeMM: Time-as-Operator Spectral Filtering for Dynamic Multimodal Recommendation Rec-GPT4V: Multimodal Recommendation with Large Vision-Language Models

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:06:26.097461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-07T12:54:56.401350Z digest=sha256:381d325a3eac19b2cb431dd36b112099e5f7027eaa19586bfa24f1439efe70a4

Observation 0df56292-c49d-4d32-8800-b7ad80052d00 · inbound

Agent4POI: Agentic Context-Conditioned Affordance Reasoning for Multimodal Point-of-Interest Recommendation cites this paper.

Agent4POI: Agentic Context-Conditioned Affordance Reasoning for Multimodal Point-of-Interest Recommendation Rec-GPT4V: Multimodal Recommendation with Large Vision-Language Models

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-19T17:37:41.401097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T17:36:33.004577Z digest=sha256:fbd4515f113a6f11f4ab335f21df00178db486e192dd57ab924c8e882b8cac4e

Observation 5ee6975e-0442-4c6d-8667-fbdee60a8529 · inbound

The 2nd EReL@MIR Workshop on Efficient Representation Learning for Multimodal Information Retrieval cites this paper.

The 2nd EReL@MIR Workshop on Efficient Representation Learning for Multimodal Information Retrieval Rec-GPT4V: Multimodal Recommendation with Large Vision-Language Models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-29T16:23:39.956120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T15:49:45.212333Z digest=sha256:0ac3454c5220cf1f743aaf5872c63e45e0bc550768eea73ebdee4230ae059412

Observation 21581710-ad7f-429e-8d65-b10585e2b104 · inbound

Popcorn: A Configurable Benchmark for Visual Evidence in Multimodal Movie Recommendation cites this paper.

Popcorn: A Configurable Benchmark for Visual Evidence in Multimodal Movie Recommendation Rec-GPT4V: Multimodal Recommendation with Large Vision-Language Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:47:35.692763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T14:42:10.975540Z digest=sha256:38251342c141f5a78a34d7af11b85956620520d0f0893de30eda217cdaeda08f