Pith. sign in

Paper Citation Record · LEDGER

Location-Aware Fine-Grained Representation Learning for Medical Vision Foundation Models

As of 19 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2608.00976.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.00976 v1

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T00:40:24.500720Z

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

13 of 13 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8d5f67d1-81ce-4b05-bceb-37b6152d331a · outbound

This paper cites Radvlm: A multitask conversational vision-language model for radiology.arXiv preprint arXiv:2502.03333,.

Location-Aware Fine-Grained Representation Learning for Medical Vision Foundation Models Radvlm: A multitask conversational vision-language model for radiology.arXiv preprint arXiv:2502.03333,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T00:40:24.465902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:40:24.465902Z digest=sha256:87bae833f956cbe8e67c62b61cddd0813e453574485de0b0429e7dc35cebb491

Observation a80f62fc-c0c2-4a8f-93d2-63dc92ff2a86 · outbound

This paper cites V-JEPA 2.1: Unlocking Dense Features in Video Self-Supervised Learning.

Location-Aware Fine-Grained Representation Learning for Medical Vision Foundation Models V-JEPA 2.1: Unlocking Dense Features in Video Self-Supervised Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T00:40:24.469395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:40:24.469395Z digest=sha256:9c758d872dfdfe4403380c62f7ad6980b8eb145019090d8ed465163a8e566bdc

Observation 21b53b96-addb-48a1-b0e9-a5374b211a69 · outbound

This paper cites Capabilities of Gemini Models in Medicine.

Location-Aware Fine-Grained Representation Learning for Medical Vision Foundation Models Capabilities of Gemini Models in Medicine

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T00:40:24.473729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:40:24.473729Z digest=sha256:da6553b8a199cafee7eb0e14bf164248e1c408342dbcfd154ef03159b73524ba

Observation da649b29-d4eb-4ebf-90db-635235af74bb · outbound

This paper cites Fleming-vl: Towards universal medical visual reasoning with multimodal llms.arXiv preprint arXiv:2511.00916,.

Location-Aware Fine-Grained Representation Learning for Medical Vision Foundation Models Fleming-vl: Towards universal medical visual reasoning with multimodal llms.arXiv preprint arXiv:2511.00916,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T00:40:24.481101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:40:24.481101Z digest=sha256:bf7c639f9b4ace7120b8d8ce86016201dcf287d8a331d5034fdc254804602104

Observation 3f7c9418-b800-466d-bfbf-02cdaeabc11e · outbound

This paper cites DINOv3.

Location-Aware Fine-Grained Representation Learning for Medical Vision Foundation Models DINOv3

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T00:40:24.484809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:40:24.484809Z digest=sha256:fa61dcdf2ad7344fc5662ed29a1d11ea792a48fb48a46b07896c607e26e1ef05

Observation d68ffc2d-0b80-4d46-8658-6d7da2a3f028 · outbound

This paper cites Gemma 3 Technical Report.

Location-Aware Fine-Grained Representation Learning for Medical Vision Foundation Models Gemma 3 Technical Report

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T00:40:24.488181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:40:24.488181Z digest=sha256:4505816a2b5267c7bb25416b1ddb45051d473c08adf6bb5bb65b8ee76217de08

Observation 39247014-e734-4430-b32e-93b31e87ca27 · outbound

This paper cites Lingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning.

Location-Aware Fine-Grained Representation Learning for Medical Vision Foundation Models Lingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T00:40:24.492327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:40:24.492327Z digest=sha256:fd1acbe57f1a1b94b76cbad2fdfcf144e4f24a41208db2b55983f8c9ad7ffa19

Observation 371fd059-7488-4e7f-94be-3911a5d59320 · outbound

This paper cites SA-Med2D-20M Dataset: Segment Anything in 2D Medical Imaging with 20 Million masks.

Location-Aware Fine-Grained Representation Learning for Medical Vision Foundation Models SA-Med2D-20M Dataset: Segment Anything in 2D Medical Imaging with 20 Million masks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T00:40:24.496366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:40:24.496366Z digest=sha256:301bf7ac4ec65a1ad7f6e4cc54f771486dac93327eca10fcc5676e28d8e9df25

Observation 553027f1-ba0b-4eab-85ad-4b0a68bc7b53 · outbound

This paper cites Chexagent: Towards a foundation model for chest x-ray interpretation.

Location-Aware Fine-Grained Representation Learning for Medical Vision Foundation Models Chexagent: Towards a foundation model for chest x-ray interpretation

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T00:40:24.461873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:40:24.461873Z digest=sha256:569110b1f5a0eb89e8aa1da5071ff6d5fdd72c04f947bea5e12ab76427b16ce5

Observation 3005e101-f44d-4b62-b864-785f92b20018 · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

Location-Aware Fine-Grained Representation Learning for Medical Vision Foundation Models InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T00:40:24.500720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:40:24.500720Z digest=sha256:2b25195b8f48ca626934f4bc7443ce8c08f6e39020441c01be5eb00ceffa30a0

Observation 108d2e2d-be0e-498e-94f5-2175d8c48014 · outbound

This paper cites MedGemma Technical Report.

Location-Aware Fine-Grained Representation Learning for Medical Vision Foundation Models MedGemma Technical Report

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T00:40:24.477286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:40:24.477286Z digest=sha256:ef68997d7daa597570fdc61f6605fd0e2c422259ab2d0059b471ec1dae7d95f5

Observation 6f989b60-66ea-4727-aa32-159957924208 · outbound

This paper cites Mirage: The illusion of visual understanding.arXiv preprint arXiv:2603.21687,.

Location-Aware Fine-Grained Representation Learning for Medical Vision Foundation Models Mirage: The illusion of visual understanding.arXiv preprint arXiv:2603.21687,

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T00:40:24.454599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:40:24.454599Z digest=sha256:6ebd516a74195015ec492e9dc061cb5207368043dfe6bca0cdaf5fcc997aba14

Observation 58329fec-921d-4e35-876a-7cbff2e567cd · outbound

This paper cites MAIRA-2: Grounded Radiology Report Generation.

Location-Aware Fine-Grained Representation Learning for Medical Vision Foundation Models MAIRA-2: Grounded Radiology Report Generation

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-06T00:40:24.458377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:40:24.458377Z digest=sha256:cf5063955d88f68552c7edd2a11122875678d5c918cbc03b4329e8a86c67a798

Pith citing papers

No inbound Pith citation observations are available.