Pith. sign in

Paper Citation Record · LEDGER

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge

As of 23 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 2 inbound Pith citation observations for arXiv:2607.10953.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.10953 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-14T08:06:08.025649Z

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:08:07.587136Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-16T12:16:17.039197Z

Reference resolution

28 of 28 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a65e3056-7772-4d37-b6aa-074b8ff4e05d · outbound

This paper cites Developing generalist foundation models from a multimodal dataset for 3d computed tomography,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Developing generalist foundation models from a multimodal dataset for 3d computed tomography,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:07afe4dbcee0cc4d092eff750365a28e06d56578384a1a3af0e7b4f0f6aaebde

Observation c461f93a-a085-4a86-a4ad-7a3cb034a945 · outbound

This paper cites Merlin: A vision language foundation model for 3d computed tomography,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Merlin: A vision language foundation model for 3d computed tomography,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:4ff973d2961b7fefe20e5f6f4457f0481f18b8838f9665d4e6986c82fba6cb1d

Observation 5a8d7ab7-b4bc-42de-8648-d4c418449175 · outbound

This paper cites Gloria: A multimodal global-local representation learning framework for label- efficient medical image recognition,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Gloria: A multimodal global-local representation learning framework for label- efficient medical image recognition,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:b72b33e31a5a7b47b5de20bf075e47feccda1e318c15c289ac4a5d969f6fc26b

Observation 57c10ef6-df84-4c19-88e5-d028bd0c7072 · outbound

This paper cites Im- proving medical vision-language contrastive pretraining with semantics- aware triage,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Im- proving medical vision-language contrastive pretraining with semantics- aware triage,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:dedf5ef5f0feede063acbd4859fb2c74be60c33f8f582d527bc98b14e0e02b83

Observation 1edc0719-93af-4bbe-b1e1-4c56230814cd · outbound

This paper cites Ecamp: entity-centered context-aware medical vision language pre- training,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Ecamp: entity-centered context-aware medical vision language pre- training,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:507f0cd2b60c2b04c9ecbc75be59d4e716769424da0a0ba1eacc9451d4778ca2

Observation 1d3840f5-ee05-4836-a021-ff7e615b35d3 · outbound

This paper cites Large-scale and fine-grained vision-language pre-training for enhanced ct image understanding,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Large-scale and fine-grained vision-language pre-training for enhanced ct image understanding,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:11ffbd24460d0d0499551dc7de2e49d070ffbc3049a572942baf230bfdf6f022

Observation a0470f05-645d-47b9-b960-0c34b9e1b2ee · outbound

This paper cites MG-3D: Multi-Grained Knowledge-Enhanced 3D Medical Vision-Language Pre-training.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge MG-3D: Multi-Grained Knowledge-Enhanced 3D Medical Vision-Language Pre-training

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:7c5a5c115371f3c3e19cf1d4f22c9cdf6dcdc82a3ba6aabe9e14f3232d68de77

Observation a1ffae66-952a-4075-ab12-3c05eafd025d · outbound

This paper cites Ct-glip: 3d grounded language-image pretraining with ct scans and radi- ology reports for full-body scenarios,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Ct-glip: 3d grounded language-image pretraining with ct scans and radi- ology reports for full-body scenarios,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:8acd7cf932de7cffb8865158921a9de981ed66f59906fe0e64249a5e32131e70

Observation a1c53780-8288-4311-b222-4fea5dbeb249 · outbound

This paper cites Radgraph-xl: A large-scale expert-annotated dataset for entity and re- lation extraction from radiology reports,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Radgraph-xl: A large-scale expert-annotated dataset for entity and re- lation extraction from radiology reports,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:b0599b278b95194b70fde06e5170054aafbba5f3ccd5dc431d113f2c7cabe244

Observation 61a3a22a-814a-4a2a-8642-aafb11cb5e1f · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Learning transferable visual models from natural language supervision,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:d29c437560784f1365656b8930d7fc3ae249a9be10dd6e7924b801b99b9f83da

Observation f441c554-d308-4cab-acbc-1c9246d0685a · outbound

This paper cites Scaling up visual and vision-language representation learning with noisy text supervision,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Scaling up visual and vision-language representation learning with noisy text supervision,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:ad18dab2a99d72916905a7cdbdfecd5f229e6be1526a93364e2ad2b78114e24e

Observation fb92b930-9116-43a3-a6ed-4ff65e79fe21 · outbound

This paper cites Align before fuse: Vision and language representation learning with momentum distillation,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Align before fuse: Vision and language representation learning with momentum distillation,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:184651c9539e4861ddcff819b2ac614f93a36fc65bca43f7d7909e7aa5dc7e7c

Observation 563ec6ae-49a6-4449-9833-0f2924f60a46 · outbound

This paper cites Chestx-ray: Hospital-scale chest x-ray database and benchmarks on weakly-supervised classification and localization of common thorax diseases,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Chestx-ray: Hospital-scale chest x-ray database and benchmarks on weakly-supervised classification and localization of common thorax diseases,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:b70e4ad0c360577cefbc67392900563fe415c871766985a97ab392bcbde71991

Observation bcaf0e72-3cda-4aba-a9bb-335cb81b93ed · outbound

This paper cites Chexpert: A large chest radiograph dataset with uncertainty labels and expert comparison,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Chexpert: A large chest radiograph dataset with uncertainty labels and expert comparison,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:dd31d68a51f47775e38b7e5ac5c932b190eb74f2f85a00354a6a940b7681c2de

Observation 7f0039a6-c1a4-4a27-8728-6767683228fb · outbound

This paper cites Mimic-cxr, a de- identified publicly available database of chest radiographs with free-text reports,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Mimic-cxr, a de- identified publicly available database of chest radiographs with free-text reports,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:815fa2b4df73d3aa9fea440a1076ba207ca389f05985555bcc2b39c1be581661

Observation ab8529fe-2ba5-4c68-b43c-05d2c285226d · outbound

This paper cites Preparing a collection of radiology examinations for distribution and retrieval,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Preparing a collection of radiology examinations for distribution and retrieval,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:8a37632e1df8039357693bed6f549bc0156142505f95a81e5bdb38b6dfaf64bb

Observation 77179951-1e43-4444-ad7c-3dc417f663f7 · outbound

This paper cites Medclip: Contrastive learning from unpaired medical images and text,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Medclip: Contrastive learning from unpaired medical images and text,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:e4da653775b36cbd856472c4858d84e4b8598d6407947e203f16cdd13ba2fea1

Observation 51e05f41-fb42-4ffa-bd15-4140398e762f · outbound

This paper cites Bootstrapping chest ct image understanding by distilling knowledge from x-ray expert models,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Bootstrapping chest ct image understanding by distilling knowledge from x-ray expert models,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:356b3f7523cf74fc3a776042305a3de41f335d175cc845c6958759de8f9a8ed0

Observation d72810f5-d91a-4647-9d67-b1d63c492a9e · outbound

This paper cites Bridged semantic alignment for zero-shot 3d medical image diagnosis,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Bridged semantic alignment for zero-shot 3d medical image diagnosis,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:880343aacc614bf83037b19380ad35897d8caec8cd663cd8a4cf846fdfd6533e

Observation ae87f019-574e-4c84-9996-353943e21ce2 · outbound

This paper cites Joint learning of localized representations from medical images and reports,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Joint learning of localized representations from medical images and reports,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:f0ba3d4fba773a8d09e1e320821a542a1792dc299ab8ad21db9640e8ef51f837

Observation 008ecc4f-6fb3-4fd8-b328-3bf69bdd8e8d · outbound

This paper cites Making the most of text semantics to improve biomedical vision– language processing,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Making the most of text semantics to improve biomedical vision– language processing,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:c1ab269730a43e975219a339eba56878c3ebc7fa646e34fe8e218550bfd6e578

Observation 68c12639-d3f7-4d9b-a0c2-58382eea1417 · outbound

This paper cites Bcnet: Bronchus classification via structure guided represen- tation learning,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Bcnet: Bronchus classification via structure guided represen- tation learning,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:8313d3503ee861eaaed516b4a843bee708bfbb8e04167c20613178fa7eef49e9

Observation c3f70d21-a496-4dfb-9bdd-f27f25110df0 · outbound

This paper cites Boundary as the bridge: Towards heterogeneous partially-labeled medical image segmentation and landmark detection,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Boundary as the bridge: Towards heterogeneous partially-labeled medical image segmentation and landmark detection,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:ce2295f211cf7c6cb40494b5cfc0d80b70396971f3c86c326f156819bd7109f9

Observation 4f276a67-8fed-4325-9d77-099ef5d2705a · outbound

This paper cites Learning spatio-temporal features with 3d residual networks for action recognition,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Learning spatio-temporal features with 3d residual networks for action recognition,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:28da0c35b15f5b4da866cc0ca6864dfda0ee6612bf2af33846ce01275ef0d29f

Observation 8bdc3d90-de79-4dd6-914c-f377c96278c9 · outbound

This paper cites Machine-learning-based multiple abnormality pre- diction with large-scale chest computed tomography volumes,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Machine-learning-based multiple abnormality pre- diction with large-scale chest computed tomography volumes,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:f62f9559dbc7cbf00cf0d0dd6592ed0ed7ce57f5d43bf7ba4cfbeecc1fd1ca49

Observation 8439d96e-8db0-4b6e-a53d-605bac755481 · outbound

This paper cites Compre- hensive language-image pre-training for 3d medical image understand- ing,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Compre- hensive language-image pre-training for 3d medical image understand- ing,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:fc68dbecedb0eaaafb7330f032bdfa1e2023da8303e762074681d5d5481986a9

Observation 79332050-203d-4fe0-b84d-ce9b8e9a1dac · outbound

This paper cites Totalsegmentator: robust segmentation of 104 anatomic structures in ct images,.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Totalsegmentator: robust segmentation of 104 anatomic structures in ct images,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:0b588990acbef48bd71425e15679a765727da5fb97cc9d147c61d94726c52844

Observation 63541012-c3da-42c3-b9a4-4983a6579ce7 · outbound

This paper cites Qwen3 Technical Report.

Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge Qwen3 Technical Report

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-14T08:06:08.025649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:06:08.025649Z digest=sha256:2100d65697b518dd8e7ea8d2e8c8e058417298f5dd82d3aca6a60a0f6dc72e38

Pith citing papers

Observation 64a2a705-fb2f-4170-8573-dccca25557d8 · inbound

OrganLens: Organ-Specific Representation Learning for CT Foundation Models cites this paper.

OrganLens: Organ-Specific Representation Learning for CT Foundation Models Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-01T03:19:39.347401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:19:39.347401Z digest=sha256:efa768d1fe09def324b37f88188296c7e80f975c400dd45bc2380fcec1f49aea

Observation 19799ce4-0921-4edb-9676-e4613b1c38ff · inbound

Auditable agentic AI for evidence-grounded thyroid ultrasound diagnosis and reporting cites this paper.

Auditable agentic AI for evidence-grounded thyroid ultrasound diagnosis and reporting Learning Anatomy-Grounded CT Vision-Language Representations with Organ-Hierarchical Report Knowledge

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-16T00:08:08.327610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-16T00:08:07.587136Z digest=sha256:693df513fe33cffde50aa572d9a15c8c1b8e7642a1a8270e95f81fe2c8c7f8a7