Pith. sign in

Paper Citation Record · LEDGER

MINER: Mining Multimodal Internal Representation for Efficient Retrieval

As of 6 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2605.06460.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.06460 v1

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-08T12:40:33.437364Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

36 of 36 outbound references displayed

  • verified exact8
  • verified fuzzy28
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b364d428-618e-4eb9-b70c-69d4e37fb54d · outbound

This paper cites Deep learning based visually rich document content understanding: A survey.Artificial Intelligence Review.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval Deep learning based visually rich document content understanding: A survey.Artificial Intelligence Review

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.079777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:639ec6641f43118d30b414aa1713cbc9f413679e471de879d59dec5a54b8a2c5

Observation a63ee639-ec7b-4ce1-b565-79f6c87f42c4 · outbound

This paper cites Scaling Beyond Context: A Survey of Multimodal Retrieval-Augmented Generation for Document Understanding.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval Scaling Beyond Context: A Survey of Multimodal Retrieval-Augmented Generation for Document Understanding

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-11T19:06:10.589628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:68e08ea918cdc3b83bf42990e061f66c95acc307d15c1b012043ea84653e6234

Observation 7ad879c2-10ae-4111-a5ce-7c699c159407 · outbound

This paper cites Unifying multi- modal retrieval via document screenshot embedding.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval Unifying multi- modal retrieval via document screenshot embedding

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.040593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:5f55624f9a2e740d88753b6a147e8295550c82d6da8ce7d5d9bfa2b92ad5ce78

Observation 8df7abf2-1aa5-4a59-8b5e-75cb1b1d9fec · outbound

This paper cites ColPali: Efficient document retrieval with vision language models.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval ColPali: Efficient document retrieval with vision language models

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.061813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:f5947387cb9bffeccbf2f9139d061386dbd3da282957953658050315a1a8396a

Observation 3ded030d-45a7-4c6a-b7a9-2583602d6493 · outbound

This paper cites VisRAG: Vision-based retrieval-augmented generation on multi-modality documents.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval VisRAG: Vision-based retrieval-augmented generation on multi-modality documents

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.076834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:d26f0e1c2550dd3ddec9649ef9ce1fcb93702792645027db8f8501241705419c

Observation 7df24362-045a-4e27-a641-e987c24ce510 · outbound

This paper cites ColBERT: Efficient and effective passage search via con- textualized late interaction over BERT.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval ColBERT: Efficient and effective passage search via con- textualized late interaction over BERT

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.009922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:9787828111f8d6d8c8dec92b10b5ebcb4d9336ee6cc294675b6aa2146d5f9864

Observation 1fd29631-da25-40e0-aecc-f5855cff1106 · outbound

This paper cites Towards storage-efficient visual document retrieval: An empirical study on reducing patch-level embeddings.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval Towards storage-efficient visual document retrieval: An empirical study on reducing patch-level embeddings

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.056250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:d970fa1a8024668bf29c7b88622df60de1b1fad1e9c816ee5a04ac58ea18eb6a

Observation fb1fdc9b-da19-4d63-8617-26689da130a9 · outbound

This paper cites ColBERTv2: Effective and efficient retrieval via lightweight late interaction.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval ColBERTv2: Effective and efficient retrieval via lightweight late interaction

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.050651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:9ae22752c62c37dd7061673238abbbad8e1127c7338c6f5a4a10b8ae46867db3

Observation 2be3f939-e063-4f4a-a22b-835b307a8389 · outbound

This paper cites Dense passage retrieval for open-domain question answering.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval Dense passage retrieval for open-domain question answering

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.035355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:4d12cf3afed9dd15e069c18c1fd80b60a13ab91b457e335d6a0b8be6cd824ca3

Observation bc30b4e0-14eb-4e9a-98dc-00900c61e9cd · outbound

This paper cites MM-Embed: Universal multimodal retrieval with multimodal LLMs.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval MM-Embed: Universal multimodal retrieval with multimodal LLMs

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.030104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:912330a2fbcd5b03e5aa59520ed1e1ce212d5779294b896b49b6555dec0c97f8

Observation b45cbb90-c25c-4801-bfc1-b46dcf26ff3b · outbound

This paper cites jina- embeddings-v4: Universal embeddings for multimodal multilingual retrieval.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval jina- embeddings-v4: Universal embeddings for multimodal multilingual retrieval

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.047981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:137464e9de067d9370ee5ff420d690d9546cb678fc4531f668f572d41939d46c

Observation 5bb823bd-bfa3-4ec5-bcfb-4a0f98610587 · outbound

This paper cites Head2toe: Utilizing intermediate representations for better transfer learning.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval Head2toe: Utilizing intermediate representations for better transfer learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.012065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:9ad76235300dc1fc69c217cf0e7a899258cacc2f93d2b64d3b7f03ee96c1a944

Observation 639746ab-a7a8-4ebb-afa0-f0911a06c0cc · outbound

This paper cites Visual query tuning: Towards effective usage of intermediate representations for parameter and memory efficient transfer learning.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval Visual query tuning: Towards effective usage of intermediate representations for parameter and memory efficient transfer learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.019733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:f9d695a6b84de2436d2551392ddfff86c3e4c7e5d6ba453816cbc5b437b033a0

Observation d9b176f5-756b-48a5-ab48-ed1f55ab409f · outbound

This paper cites Parameter-efficient and memory-efficient tuning for vision transformer: a disentangled approach.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval Parameter-efficient and memory-efficient tuning for vision transformer: a disentangled approach

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.058973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:20983fad90c3397c9643ccd2fbb0ba5af2e5c23c21c35ae7751391658fac14e1

Observation 0c59dd77-a11a-4f25-a87b-fc74b0b0aa16 · outbound

This paper cites SPIN: Sparsifying and integrating internal neurons in large language models for text classification.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval SPIN: Sparsifying and integrating internal neurons in large language models for text classification

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.045573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:977359155e71fc86bbeb16de5a7e9bb3f6a51a06467183878496d9c9d616222b

Observation 85e0dbd5-f316-43eb-a156-6af8652f9f7f · outbound

This paper cites LLM Safety From Within: Detecting Harmful Content with Internal Representations.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval LLM Safety From Within: Detecting Harmful Content with Internal Representations

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-11T19:06:10.607800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:2c7103ceac9a0a7103b81eacc16d6590a565524a2031b15e90f130dfab460a9d

Observation cae10cba-748b-49ea-aa5c-0859c201015d · outbound

This paper cites Perception Encoder: The best visual embeddings are not at the output of the network.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval Perception Encoder: The best visual embeddings are not at the output of the network

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.064845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:59fb03c70e6c19a22c04f54e6c313a48d8d7bfd7140bdbcb162125308d824246

Observation 2a3a2e76-12f4-4460-8a38-1ac05f3bc343 · outbound

This paper cites Similarity of neural network representations revisited.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval Similarity of neural network representations revisited

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.014306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:8e72866f7fd5a79c3a48a0dc8d72896e278135318faffb71d25a6a2ea3191a90

Observation ee942beb-4920-48c3-91e3-391dca4426db · outbound

This paper cites Vidore benchmark v2: Raising the bar for visual retrieval.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval Vidore benchmark v2: Raising the bar for visual retrieval

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:06:10.633306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:c843bffbf28c3074fae460b7f454f77605a8fe1039ddfa475740281503e5a18c

Observation 6e9f785d-a7b4-487c-8c30-41dbf2678dfb · outbound

This paper cites ViDoRe V3: A Comprehensive Evaluation of Retrieval Augmented Generation in Complex Real-World Scenarios.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval ViDoRe V3: A Comprehensive Evaluation of Retrieval Augmented Generation in Complex Real-World Scenarios

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-05-11T19:06:10.624346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:a2c9bfe4fceb75df43f37f8fd7970cd9fccdcbd30619affc45a7eb455964c1c2

Observation 0fa20434-17c7-40c1-bdc0-d38219fd7208 · outbound

This paper cites MMDocIR: Benchmarking multimodal retrieval for long documents.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval MMDocIR: Benchmarking multimodal retrieval for long documents

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.016663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:94df3236ac636cc4cd5f3d029a357155d8fd0443fde5126773658f21b2a50825

Observation e0b5a8c2-a50f-4d35-8fd4-0f029e609d28 · outbound

This paper cites VisR-Bench: An Empirical Study on Visual Retrieval-Augmented Generation for Multilingual Long Document Understanding.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval VisR-Bench: An Empirical Study on Visual Retrieval-Augmented Generation for Multilingual Long Document Understanding

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:06:10.643227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:4d3e4c37dffb875808fec5b988a6b7b9c3847d7621dee4d6bd5b140ddb0902de

Observation e86a0f5b-1bf2-4657-8c20-b4aeb28e8a24 · outbound

This paper cites Reproducibility, replicability, and insights into visual document retrieval with late interaction.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval Reproducibility, replicability, and insights into visual document retrieval with late interaction

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.032563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:60115e07053d9232b3e2c9b04bc0f7d59253e29cd9151fc83b91bdf184fcce81

Observation 16b7248c-5f4e-401c-b9ee-30af9ccb3083 · outbound

This paper cites Fine-grained late-interaction multi-modal retrieval for retrieval augmented visual question answering.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval Fine-grained late-interaction multi-modal retrieval for retrieval augmented visual question answering

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.043078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:6fbbb0d7c9cb9ff2c3685391b7a473bc5352b0f3b95bd0c477b50c8663a38f23

Observation d311d56e-3222-4553-a9d1-6d2df0e9102c · outbound

This paper cites Eager embed v1: Multimodal dense embeddings for retrieval.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval Eager embed v1: Multimodal dense embeddings for retrieval

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.038119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:b1985d0bc9d09d266b40cdcfcad45ef02a447349cd7edd4da1623ddd3d0f6c77

Observation a735328e-7234-4fa5-b445-8b921d56cbcc · outbound

This paper cites Investigating multi-layer representations for dense passage retrieval.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval Investigating multi-layer representations for dense passage retrieval

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.067596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:9dd5c6529536f2cc0be33102f8eabcfde9b2a5f2d6f2c46c0e0f91aa5bc681c9

Observation 4d7edc2b-ef31-4c3e-b2c7-0e5a977e11f5 · outbound

This paper cites MoCa: Modality-aware Continual Pre-training Makes Better Bidirectional Multimodal Embeddings.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval MoCa: Modality-aware Continual Pre-training Makes Better Bidirectional Multimodal Embeddings

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:06:10.555032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:4202f383b8a126d4f2ef704a04d76842bd43ed41cc77c4e8bd19e5a4184efd66

Observation 033d49f1-62d3-4ebf-9ab3-68b7d9941a7f · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval Representation Learning with Contrastive Predictive Coding

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-05-11T19:06:10.600681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:a299589a065590886d0ae49750cd6403ea89ff0f6e6017226084813cf8daf70a

Observation 5f283085-22f1-4440-bdb4-82956a5613b0 · outbound

This paper cites Regression shrinkage and selection via the lasso.Journal of the Royal Statistical Society: Series B (Methodological), 58(1):267–288, 01 1996.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval Regression shrinkage and selection via the lasso.Journal of the Royal Statistical Society: Series B (Methodological), 58(1):267–288, 01 1996

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.025435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:f53644fad0c7c08aae14d90ea10b421165f375719878f8ac9237d7fe66e6c527

Observation 2ababb76-457b-4563-9076-23595ba31248 · outbound

This paper cites Guided query refine- ment: Multimodal hybrid retrieval with test-time optimization.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval Guided query refine- ment: Multimodal hybrid retrieval with test-time optimization

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.070230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:6cba46a939f3a58e17ad3fa6b108bb081e06a62d97fc12397d4963e5c300859b

Observation a1eb0210-d2d5-4b96-85b4-c208863b6016 · outbound

This paper cites Cumulated gain-based evaluation of IR techniques.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval Cumulated gain-based evaluation of IR techniques

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.022505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:9bb45e422938e2795317043029c2ac5616bae7529d464482650f1a22befb19bc

Observation 0b1410b6-fa12-4ca8-9eb7-51fe6ca33638 · outbound

This paper cites Decoupled weight decay regularization.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval Decoupled weight decay regularization

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.027901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:f80c7f71314f0e95b3376648ccb857602a3cb322834eb9ece8cea9ce0454aea8

Observation e8b25ded-698b-45bb-ae56-ee49ed0f5477 · outbound

This paper cites Sigmoid loss for language image pre-training.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval Sigmoid loss for language image pre-training

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.053095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:bd702cb54f5148d82ed24b003f79b00b32655da4234e3806de26c1169e5379a0

Observation 650bd11a-5b5d-4e3e-8614-3385312b442a · outbound

This paper cites VLM2Vec: Training vision-language models for massive multimodal embedding tasks.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval VLM2Vec: Training vision-language models for massive multimodal embedding tasks

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.007708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:ff7dde18ca6611ece750c1a3f2e01df863e691582c58404d4424639933920553

Observation bf0c0a53-37cd-48d8-8ec5-4694c4ca566f · outbound

This paper cites Qdrant: Vector similarity search engine and vector database.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval Qdrant: Vector similarity search engine and vector database

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T13:07:48.073622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:7707fec02a4c2dbac6463786bb76b2313cc01559e3de72e5fb376b92a8eba22f

Observation b3f66d4b-07c3-4387-aa62-fcb3ef1fe88f · outbound

This paper cites Qwen3-VL-Embedding and Qwen3-VL-Reranker: A Unified Framework for State-of-the-Art Multimodal Retrieval and Ranking.

MINER: Mining Multimodal Internal Representation for Efficient Retrieval Qwen3-VL-Embedding and Qwen3-VL-Reranker: A Unified Framework for State-of-the-Art Multimodal Retrieval and Ranking

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:36:21.368649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:40:33.437364Z digest=sha256:ee509151cf64f46be1558e5675f506a32e3f87753bb19f0dbf93e8877c67fdd2

Pith citing papers

No inbound Pith citation observations are available.