Pith. sign in

Paper Citation Record · LEDGER

MedViLaM: A multimodal large language model with advanced generalizability and explainability for medical data understanding and generation

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2409.19684.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2409.19684 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T17:18:16.391750Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T22:37:26.494552Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b1413c7a-f36c-43d1-a5c8-ca336a0980b3 · inbound

Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation cites this paper.

Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation MedViLaM: A multimodal large language model with advanced generalizability and explainability for medical data understanding and generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T13:51:49.772453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:51:49.772453Z digest=sha256:66c2d001b9d2f7de9f6d7905db24fb95a327cc25748959be6ff1827cdf90808f

Observation 2cbe650e-0647-49b3-9228-3165824af08d · inbound

XrayClaw: Cooperative-Competitive Multi-Agent Alignment for Trustworthy Chest X-ray Diagnosis cites this paper.

XrayClaw: Cooperative-Competitive Multi-Agent Alignment for Trustworthy Chest X-ray Diagnosis MedViLaM: A multimodal large language model with advanced generalizability and explainability for medical data understanding and generation

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-13T21:03:19.890099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T21:00:18.392949Z digest=sha256:6dcc23ebf982def9f6efad53a438c362a01f864cf31a60308b3195ad7d399766

Observation d29316fa-8be9-4c72-b014-903201a6619c · inbound

Beyond Surrogate Gradients: Fully Differentiable Token Pruning for Vision-Language Models cites this paper.

Beyond Surrogate Gradients: Fully Differentiable Token Pruning for Vision-Language Models MedViLaM: A multimodal large language model with advanced generalizability and explainability for medical data understanding and generation

Reference 72

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T13:13:27.437713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T13:07:21.789901Z digest=sha256:066133484b207f73a8404de1976edc55cfbaad204c477c3b414ab3af5235c3c4

Observation aa4ff03c-160e-4ba0-8136-cdd4168ab181 · inbound

A unified multi-task framework enables interpretable chest radiograph analysis cites this paper.

A unified multi-task framework enables interpretable chest radiograph analysis MedViLaM: A multimodal large language model with advanced generalizability and explainability for medical data understanding and generation

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:26:26.612820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T10:56:54.620852Z digest=sha256:5dcf0248dc2bb2a0ffd77bc4306fe087d2bb3aa370e84a6c548787db434d788e

Observation e912aa31-b58d-48ce-9fbc-db27f3656292 · inbound

Learnable Token Sparsification for Efficient Gigapixel Whole Slide Image Reasoning cites this paper.

Learnable Token Sparsification for Efficient Gigapixel Whole Slide Image Reasoning MedViLaM: A multimodal large language model with advanced generalizability and explainability for medical data understanding and generation

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T22:37:26.496193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T18:43:45.530986Z digest=sha256:3e898c97d3ac9a057029b5fceb6e8b3fc68a448a7b852d9b41b7fb081f6b27ab

Observation 66c42414-cee6-4ad4-b5b2-795048e0f53e · inbound

PathSelect: Sequential Token Selection for Whole Slide Pathology cites this paper.

PathSelect: Sequential Token Selection for Whole Slide Pathology MedViLaM: A multimodal large language model with advanced generalizability and explainability for medical data understanding and generation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-30T17:15:34.947525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T17:15:34.947525Z digest=sha256:c0dd7317fe7fd3ba65a0261c9219bc0e93f329f60ad3e267df0ad3e1e0a7232f

Observation d41fddf2-6438-4ff3-b61f-5f0f148e643f · inbound

DiffPrune: differentiable information throttling for token pruning in vision-language models cites this paper.

DiffPrune: differentiable information throttling for token pruning in vision-language models MedViLaM: A multimodal large language model with advanced generalizability and explainability for medical data understanding and generation

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-04T17:18:16.391750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:18:16.391750Z digest=sha256:6c4eb826416f44df00d0beaaccc3489f43c7557316083076a13b4046b0030a63