Pith. sign in

Paper Citation Record · LEDGER

ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 27 inbound Pith citation observations for arXiv:2305.11172.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.11172 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 27 of 27 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:29:22.421366Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T17:18:43.921976Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 23434c5e-bfcc-4060-bbb7-18bcf7196dc3 · inbound

InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks cites this paper.

InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 144

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:46:10.148966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-13T22:46:09.693156Z digest=sha256:57abf831193435f5d8e179babf59ba7bf153ce444e7cbb63b37887f315947bcd

Observation 03570a7f-a732-41a8-b563-f3418d0267d9 · inbound

Chanel-Orderer: A Channel-Ordering Predictor for Tri-Channel Natural Images cites this paper.

Chanel-Orderer: A Channel-Ordering Predictor for Tri-Channel Natural Images ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T16:59:19.058172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:59:19.058172Z digest=sha256:103e5320e9fe17479432c011d7234cbfe10fc63457a7ed68b8765d3307151db5

Observation 9e387e77-8295-4325-8e48-a5f889f18720 · inbound

Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling cites this paper.

Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 248

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:23:58.111387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T13:23:57.588851Z digest=sha256:a86a7de70158bed00e224bdc640b8ac904cd46fa632d26af04a199996d9a92a7

Observation 44690aa2-6b76-4a21-b55d-30015d8ccbb5 · inbound

DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding cites this paper.

DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 88

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:09:25.576298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-11T10:09:21.542356Z digest=sha256:fd519bfeb83970d934d7a45a8499d411ea222d5f274755ae3923fbd8e814224f

Observation 5a5265e9-f206-4211-991e-1848bed6af07 · inbound

VinTAGe: Joint Video and Text Conditioning for Holistic Audio Generation cites this paper.

VinTAGe: Joint Video and Text Conditioning for Holistic Audio Generation ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T15:43:01.059533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:43:01.059533Z digest=sha256:985d0e808f72b41e705afa0e78111644dbd56955b87c57bbf119a550d678d171

Observation fa38f195-8a2e-49b1-9289-e7043eafcfe5 · inbound

Dense Audio-Visual Event Localization under Cross-Modal Consistency and Multi-Temporal Granularity Collaboration cites this paper.

Dense Audio-Visual Event Localization under Cross-Modal Consistency and Multi-Temporal Granularity Collaboration ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T13:56:18.348040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:56:18.348040Z digest=sha256:0d9cea54fe9989b3fa9a014358a6f074934eff979bd4d53ca6b52e37a6657fa9

Observation b5ccff99-e39e-438d-8839-02c93d33ba5a · inbound

GME: Improving Universal Multimodal Retrieval by Multimodal LLMs cites this paper.

GME: Improving Universal Multimodal Retrieval by Multimodal LLMs ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:35:22.570284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T06:35:22.508168Z digest=sha256:8fe9292614db6fb18f682c9fd79ad2d8b49d458afc9f3b1bde22a3eeb4efee6c

Observation ecfef9e5-880e-4904-957b-2da1627fd922 · inbound

Migician: Revealing the Magic of Free-Form Multi-Image Grounding in Multimodal Large Language Models cites this paper.

Migician: Revealing the Magic of Free-Form Multi-Image Grounding in Multimodal Large Language Models ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T21:10:40.653859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:10:40.653859Z digest=sha256:2dd14729b736fd5c344c1af9a6ef06671b2904893fed555384525cc0175638ff

Observation cce99e54-702c-482a-ac93-0483d1a00ce7 · inbound

Multi-Modality Transformer for E-Commerce: Inferring User Purchase Intention to Bridge the Query-Product Gap cites this paper.

Multi-Modality Transformer for E-Commerce: Inferring User Purchase Intention to Bridge the Query-Product Gap ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T17:10:07.179918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:10:07.179918Z digest=sha256:04008a07c4f1c2f47900ee83285fdadacc7a42926808e50551d20a767a836d99

Observation eff5863f-fcb8-42ac-8896-e8c5dfbd2728 · inbound

UniGraph2: Learning a Unified Embedding Space to Bind Multimodal Graphs cites this paper.

UniGraph2: Learning a Unified Embedding Space to Bind Multimodal Graphs ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-09T17:44:03.999167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:44:03.999167Z digest=sha256:453a156574057a889db23f7c91880d771903d93e04c8a0c34524a7199e062149

Observation 2f728b6a-6134-4895-b9e4-11f97a9c7cc5 · inbound

Latent Radiance Fields with 3D-aware 2D Representations cites this paper.

Latent Radiance Fields with 3D-aware 2D Representations ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T20:56:03.358991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:56:03.358991Z digest=sha256:7de1753bcccf9b58bcbad4af696eeb22e502dcb5ee309eef753ce87f39328351

Observation 6cb59208-e077-402e-b7e4-da68ec818a56 · inbound

InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models cites this paper.

InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 123

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:41:08.115462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T13:41:07.991012Z digest=sha256:40ac0200358ab4a96af73fe4faa4452ce53772374b3b8340953bf128e54c65d2

Observation 5371765c-4d78-4de0-b80b-d89ab40d70b8 · inbound

PerceptionLM: Open-Access Data and Models for Detailed Visual Understanding cites this paper.

PerceptionLM: Open-Access Data and Models for Detailed Visual Understanding ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 213

Resolution
unresolved
no resolver link, observed 2026-08-16T12:19:37.465501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:19:37.465501Z digest=sha256:473b1a47012cd4ac1e8683fe311028793b83eff18631d5c1c1d24c2fbbf2463c

Observation 2e801f05-dac8-4e1a-b6a1-b7c3a7e52bba · inbound

Harmony: A Unified Framework for Modality Incremental Learning cites this paper.

Harmony: A Unified Framework for Modality Incremental Learning ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:22.421366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:22.421366Z digest=sha256:794caf7c6ced09293574ddf82c28d9de2fabc84857a391bc52d2fb7857ae9172

Observation da2bad22-91ff-4041-9b98-66f99889b6b8 · inbound

BoundarySeg:An Embarrassingly Simple Method To Boost Medical Image Segmentation Performance for Low Data Regimes cites this paper.

BoundarySeg:An Embarrassingly Simple Method To Boost Medical Image Segmentation Performance for Low Data Regimes ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T21:28:18.656741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:28:18.656741Z digest=sha256:b571da3aa402c4cb74039adbdbd46e0ff106b6b6d19fc2ebd2f7a5307584b49e

Observation 7ba52c1c-6afa-4689-98d0-34c3d6b7211d · inbound

Vision Generalist Model: A Survey cites this paper.

Vision Generalist Model: A Survey ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 172

Resolution
unresolved
no resolver link, observed 2026-08-07T04:44:01.320615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:44:01.320615Z digest=sha256:2ca9a1819fd66232c9eedfb6c535cf6b7262157b274969c8c30f6e2251a501b0

Observation 715e40b1-5038-4a36-ab07-d76d3dd3586a · inbound

InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency cites this paper.

InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 140

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:58:58.839671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T11:58:58.660564Z digest=sha256:925352d44b3afb1ff1e4625ca0dc81aca87686066693c063c33fbdb9fdb22896

Observation ab661a22-76f3-4b5e-8490-5d8190bf5c58 · inbound

Robustifying Diffusion-Denoised Smoothing Against Covariate Shift cites this paper.

Robustifying Diffusion-Denoised Smoothing Against Covariate Shift ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T17:30:02.272524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T17:30:02.272524Z digest=sha256:cf0b319d1c6b77dd3a76ce7f906e5196f6d8762b2b6adc95978332a29459500a

Observation f171fce5-19e5-4ef3-a308-79c2ecb6bd2a · inbound

Multiple Domain Generalization Using Category Information Independent of Domain Differences cites this paper.

Multiple Domain Generalization Using Category Information Independent of Domain Differences ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:00:47.834659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T19:26:13.203582Z digest=sha256:f1308af44d184e16deae3043d224a4693718490cd192e20b30f910684207515c

Observation 5388d1f0-f4e7-44bd-a53a-13efea08fc65 · inbound

Beyond Surface Artifacts: Capturing Shared Latent Forgery Knowledge Across Modalities cites this paper.

Beyond Surface Artifacts: Capturing Shared Latent Forgery Knowledge Across Modalities ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:30:59.423601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T18:06:06.896112Z digest=sha256:4a668a9cae19a0adb9ff2d9d29b0532fca232f0bcfbf1ed0de46882f58369b57

Observation 7713235e-de0e-4902-bc44-860af8ab0d8f · inbound

EmergentBridge: Improving Zero-Shot Cross-Modal Transfer in Unified Multimodal Embedding Models cites this paper.

EmergentBridge: Improving Zero-Shot Cross-Modal Transfer in Unified Multimodal Embedding Models ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:41:02.029953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T15:53:51.162967Z digest=sha256:371c8ac35fd40b7193bac476a7c1c4ce8ef0d281b604710354c194e7b2fbc588

Observation d485db83-ba34-4463-82c9-1f212c2aca0e · inbound

EmergentBridge: Improving Zero-Shot Cross-Modal Transfer in Unified Multimodal Embedding Models cites this paper.

EmergentBridge: Improving Zero-Shot Cross-Modal Transfer in Unified Multimodal Embedding Models ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:17:28.733573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-13T07:16:15.202466Z digest=sha256:2b66f68c30e2f00258e0ee1ce03256c623eb19554d6799988e6caba5bd628904

Observation efd36e58-0199-4508-9216-f1cf7426caf3 · inbound

AeroRAG: Structured Multimodal Retrieval-Augmented LLM for Fine-Grained Aerial Visual Reasoning cites this paper.

AeroRAG: Structured Multimodal Retrieval-Augmented LLM for Fine-Grained Aerial Visual Reasoning ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:40:19.815717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T04:46:22.891034Z digest=sha256:1a278b16d3417b3f38b7d5874277758a5f9612ca2a7582d2205c3bcc2ca0ad28

Observation 2dd659b5-26d1-4c30-8612-d79e28be34f3 · inbound

Grounded 3D-Aware Spatial Vision-Language Modeling cites this paper.

Grounded 3D-Aware Spatial Vision-Language Modeling ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 115

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:13:15.320391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T08:08:36.012761Z digest=sha256:65cab594d5ba82db3669ece0e338b7b1e239eb75f553d6526d148382a044946b

Observation 6cf907f7-ff1f-4148-b51e-4f436dfb80a4 · inbound

Qwen-RobotWorld Technical Report: Unifying Embodied World Modeling through Language-Conditioned Video Generation cites this paper.

Qwen-RobotWorld Technical Report: Unifying Embodied World Modeling through Language-Conditioned Video Generation ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 117

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T17:18:43.923403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-27T04:19:26.332718Z digest=sha256:2d8fb0148528070be13605eac979c1d37c287c69c224c7d45df383f96da594b1

Observation f6cbf66f-a0e5-478d-ac99-a1c044ee36a0 · inbound

Qwen-Audio-VAE Technical Report cites this paper.

Qwen-Audio-VAE Technical Report ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 132

Resolution
unresolved
no resolver link, observed 2026-07-14T03:31:19.309532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T03:31:19.309532Z digest=sha256:683af78dd0988d173f41a3ec0fd3bc12656138726613dcd69379cf2ef673f250

Observation 7a516517-397a-4f39-9f64-924ee457cd8b · inbound

Fusion Embedding: A Unified Embedding Space for Text, Image, Video, and Audio cites this paper.

Fusion Embedding: A Unified Embedding Space for Text, Image, Video, and Audio ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T14:46:37.000739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T14:46:37.000739Z digest=sha256:c731ada672afd2d40a8f73cbbfcfb94fca99d6d01007aefeba780b7e686ba033