Pith. sign in

Paper Citation Record · LEDGER

Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts

As of 19 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 1 inbound Pith citation observation for arXiv:2505.08838.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.08838 v2

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T22:02:10.476459Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T16:05:38.824441Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T16:05:39.266280Z

Reference resolution

22 of 22 outbound references displayed

  • verified exact2
  • verified fuzzy2
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0aae5dd9-3380-4448-8881-2bbf1da15a07 · outbound

This paper cites HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale.

Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T22:02:10.384396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:02:10.384396Z digest=sha256:7901c4537775072160e0afcc6e6e254007e6edf82013ff339d44f2cd40af9838

Observation 3b5a57fb-aa68-4ef2-adba-040884db1da1 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T22:02:10.389112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:02:10.389112Z digest=sha256:69b5a1c95bcf182cb0f3171cccb418549e14815b7c81723c61faa6fb4473a1a8

Observation ce2cee0c-47f5-4cee-9338-57753fb0cf2f · outbound

This paper cites In: Burstein, J., Doran, C., Solorio, T.

Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts In: Burstein, J., Doran, C., Solorio, T

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T22:02:10.392843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:02:10.392843Z digest=sha256:c2ecc3a32fabee1bb554dfa66b6064625a5a210cc957362adf8b4e4ac33479e4

Observation 96a23322-e324-43f5-b58f-5daaf22abd4c · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T22:02:10.396677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:02:10.396677Z digest=sha256:c198a724f2e1e62163f504a8bd16867440bee784eb4a0001782004ccf30dfc1d

Observation ba630cdb-5387-4109-a902-3af8551200f0 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts LoRA: Low-Rank Adaptation of Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T22:02:10.401584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:02:10.401584Z digest=sha256:7fb146a08a112e713add60610ceb751c101cdb63029dc1167d23aa6018302019

Observation 4a77b202-7056-4489-9523-f0468f4c5e9b · outbound

This paper cites Breast Ultrasound Report Generation using LangChain.

Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts Breast Ultrasound Report Generation using LangChain

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T22:02:10.406305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:02:10.406305Z digest=sha256:7bc921daf0c34186fefd56b2845d2c8b3d6810355d674b83f3a7114d27553b3e

Observation 991aae31-998d-4bb3-8369-929b234b4bf7 · outbound

This paper cites LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day.

Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T22:02:10.411362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:02:10.411362Z digest=sha256:31fd8aa86fef6581e4b4c65ab168dd3aed6c6f50feb002a7405b1c7441f3232d

Observation acc6979b-fc13-40db-844f-03201b66524d · outbound

This paper cites In: Wang, L., Dou, Q., Fletcher, P.T., Speidel, S., Li, S.

Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts In: Wang, L., Dou, Q., Fletcher, P.T., Speidel, S., Li, S

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:02:11.056188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T22:02:10.415898Z digest=sha256:d5da5edafe593dfbd4e0bb233da44281fff83fa46d7470b99993afb16f2ba70a

Observation 3b2d3a42-377d-492d-ba7a-68d4b0223335 · outbound

This paper cites IEEE Transactions on Medical Imaging44(1), 19–30 (2025).

Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts IEEE Transactions on Medical Imaging44(1), 19–30 (2025)

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T22:02:10.420148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:02:10.420148Z digest=sha256:828e2286f74c45b8575793be935821e56bea12d3958a1871d9d99b78566cc63c

Observation 8eea392b-c4e0-4d37-b2cf-0c0b6a94dabc · outbound

This paper cites In: Text Summarization Branches Out.

Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts In: Text Summarization Branches Out

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T22:02:10.424424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:02:10.424424Z digest=sha256:e742ec2bdd1fd21609b70f0680750260c35c1e456eada22d0462477ce15f88a9

Observation 3343e174-41ce-4c39-8981-3adc0847add0 · outbound

This paper cites Proceedings of the AAAI 10 P.

Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts Proceedings of the AAAI 10 P

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T22:02:10.428526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:02:10.428526Z digest=sha256:a7ac96f7c0d01c98f1f68da71371cede9116eec8dd6c258b227e820ffe53c0f3

Observation 2565770d-3612-4fcf-add1-a310630c47b7 · outbound

This paper cites Visual Instruction Tuning.

Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts Visual Instruction Tuning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T22:02:10.433161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:02:10.433161Z digest=sha256:ed659fc1085d5b823ba85f5a71f2c951db6679aac4c3358df2509f7d696d7649

Observation 32ce6bda-b944-40ee-ac93-e7f453997f0b · outbound

This paper cites In: Proceedings of the 40th Annual Meeting on Asso- ciation for Computational Linguistics.

Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts In: Proceedings of the 40th Annual Meeting on Asso- ciation for Computational Linguistics

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T22:02:10.438097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:02:10.438097Z digest=sha256:ae55cf4df645406b49e58ce6df9e4e1bf5b36fe6f03ab40843079094a7391f60

Observation 9b0f9f87-ab27-40e4-91e4-e5db97786a9d · outbound

This paper cites In: Muresan, S., Nakov, P., Villavicencio, A.

Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts In: Muresan, S., Nakov, P., Villavicencio, A

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T22:02:10.442970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:02:10.442970Z digest=sha256:6db96396e2efb0c25b05c7a83a4d90f687efcd6e40626f5b44e726016383b095

Observation cb855251-7bc2-46d8-9e27-70ab97b7e0ef · outbound

This paper cites IEEE Reviews in Biomedical Engineering pp.

Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts IEEE Reviews in Biomedical Engineering pp

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T22:02:10.446929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:02:10.446929Z digest=sha256:60ff3f68521b0d6d909677bf6232bf42ffd520e484be3486e6f4ae83bf24e23f

Observation d670180c-8447-4805-a1b6-effc8cae80a3 · outbound

This paper cites In: 2015 IEEE Conference on Com- puter Vision and Pattern Recognition (CVPR).

Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts In: 2015 IEEE Conference on Com- puter Vision and Pattern Recognition (CVPR)

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T22:02:10.450930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:02:10.450930Z digest=sha256:e920e07365d7e205c2cd21ff2622a0bab9cc8f4e2be5b9bc6ab6c56f33dbe707

Observation 8932e246-6cc3-496b-bee8-328fa75bc044 · outbound

This paper cites Artificial Intelligence Review57(11) (Sep 2024).

Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts Artificial Intelligence Review57(11) (Sep 2024)

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T22:02:10.454993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:02:10.454993Z digest=sha256:36ec87241bf2a9e2a5496065254896f2c3e5f348d01d3c033aed6b8d54816bad

Observation a29645bd-6f6e-49ef-bafe-30865e66a1e5 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T22:02:10.459364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:02:10.459364Z digest=sha256:46a928df92baa4f2e946c34581ffc653232b4a3812a747485fd1aba4e74ab556

Observation 07c6a001-2935-4249-af39-62ec35240c54 · outbound

This paper cites In: Proceedings of EMNLP 2020: System Demonstra- tions.

Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts In: Proceedings of EMNLP 2020: System Demonstra- tions

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:02:11.032274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T22:02:10.463759Z digest=sha256:3111dc7aa5d67d93c08f66a24edd31e631da3c0a77eda3a8710339c25aa7157d

Observation 219cc6d8-88c4-4e86-b426-683a9dd33df1 · outbound

This paper cites DeltaNet:Conditional Medical Report Generation for COVID-19 Diagnosis.

Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts DeltaNet:Conditional Medical Report Generation for COVID-19 Diagnosis

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-15T22:02:10.639991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T22:02:10.467634Z digest=sha256:61ad95511da89c0d5b203e6c5fab67952462fa0dadbc8e901edcbb776657cd20

Observation 30dcdc89-d4ae-4040-b766-f5e06fd034c6 · outbound

This paper cites Neurocomputing 427, 40–49 (2021).

Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts Neurocomputing 427, 40–49 (2021)

Reference 21

Resolution
verified exact
doi, observed 2026-08-15T22:02:10.511704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T22:02:10.472150Z digest=sha256:ceee1404c12a9d16f72dff33b1602676fcfbe4e942a48b905126320ec5b62d71

Observation 0b2cfd8f-dc8a-4de5-85a7-369ccc2e0974 · outbound

This paper cites Artificial Intelligence in Medicine151, 102846 (2024).

Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts Artificial Intelligence in Medicine151, 102846 (2024)

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T22:02:10.476459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:02:10.476459Z digest=sha256:c1dced0b3048dc871c2174a09fc0a29f63078e1ad8a25d953edc39f11866bc6c

Pith citing papers

Observation b97c11a8-40d4-4461-bf5e-ed3c5e29b4ae · inbound

MedVQA-TREE: A Multimodal Reasoning and Retrieval Framework for Sarcopenia Prediction cites this paper.

MedVQA-TREE: A Multimodal Reasoning and Retrieval Framework for Sarcopenia Prediction Ultrasound Report Generation with Multimodal Large Language Models for Standardized Texts

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-05T16:05:39.271806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T16:05:38.824441Z digest=sha256:5467a57b1bd5d52069f15e9d5b617cb8f83aaba531f3c86bced2649627a33165