Pith. sign in

Paper Citation Record · LEDGER

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks

As of 4 August 2026, this Paper Citation Record lists 100 of 108 outbound references and 2 inbound Pith citation observations for arXiv:2605.10286.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.10286 v1

Coverage vector

measured 100 of 108 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-12T05:10:02.941396Z

measured 102 of 102 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-03T06:30:56.289259+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T00:25:11.457757Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 108 outbound references displayed

  • verified exact44
  • verified fuzzy32
  • unresolved5
  • parse uncertain1
  • malformed identifier1
  • metadata mismatch17

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8590c3d1-714f-49e5-838f-229059352c8a · outbound

This paper cites arXiv.org , author =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks arXiv.org , author =

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.835066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:2ce398c9af5c7ac46cea97070f013acd999de327d2802c502422d4b52fd8782c

Observation 4eab7724-d9e2-4baf-83f9-0af0fd9caa07 · outbound

This paper cites arXiv.org , author =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks arXiv.org , author =

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.832517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:55151e788b4639f3631b9fe71e23f29d528e54eec8a64a9b17c0ef4a8d53e241

Observation ddbf4b11-3073-48db-af19-6805e8da5973 · outbound

This paper cites arXiv.org , author =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks arXiv.org , author =

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.829683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:f0fa076fc542f6001a631de09564cd5a7a1e264e8f87a563d4e8f462a53b611f

Observation c7a09e7e-a47b-4673-8684-1e5cd8aa8f21 · outbound

This paper cites an unresolved cited work.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Unresolved cited work

Reference 4

Resolution
parse uncertain
raw_fallback, observed 2026-05-12T12:06:33.827187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:afc460d597a662ed473016a73e8bdbd833b7ce4f776a85d7c09be42c21f6f958

Observation 7d666398-47d7-46fb-8c6f-ff6adf067ccf · outbound

This paper cites and Gupta, Vinayak and Althoff, Tim and Hartvigsen, Thomas , month = dec, year =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks and Gupta, Vinayak and Althoff, Tim and Hartvigsen, Thomas , month = dec, year =

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.786837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:95197ec0d29fb838460abce5d9c12c00fa51bae08dd830f5b24fa38c545d2ae7

Observation d2c00a8e-21e0-4ebf-9a83-14e8b5e0d1c2 · outbound

This paper cites an unresolved cited work.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-05-12T12:06:33.760461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:48a5810b07ac8332e9dda673db9db72a004037270897c19b0af799866352bcdc

Observation 65402a0d-98ee-4daa-855c-bfdd48ccf726 · outbound

This paper cites Proceedings of the 41st.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Proceedings of the 41st

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.784135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:c2ae734e901792aa8665ce941e95778b8f1b7a4c6a169bde38404b326f71d430

Observation f76aad0a-1487-4909-a43a-1d2b70b8900e · outbound

This paper cites Proceedings of the 10th Machine Learning for Healthcare Conference , year =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Proceedings of the 10th Machine Learning for Healthcare Conference , year =

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.767612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:10e3af2feed12e194d06cafc4e064348f7adb118922d06a45c218d19bea2fe05

Observation dfe0edfd-a566-4238-b994-6038f4a86492 · outbound

This paper cites and Shamout, Farah E.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks and Shamout, Farah E

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.765194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:f2078b7cb4f02489b4a07a2cc6b9b0c40160c0176501bc4895625292d7c71d7d

Observation 956986e3-92d4-40fa-9177-428da5858d64 · outbound

This paper cites Artsi, V.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Artsi, V

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:11:21.766348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:45a117507bf4e95bd9ea935c9b23ec7cfb1e0bbbc43747ffa922b50aadd6ef92

Observation 2f205b71-3d31-42c5-bb6b-166afd34f280 · outbound

This paper cites Proceedings of the 38th.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Proceedings of the 38th

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.778888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:dd0ecd14b1875849ce2983cceb48168efa79aefd2d871aae3f03568217a0d704

Observation cccb34cb-fa14-4b40-84f8-1aefc7e7f144 · outbound

This paper cites Informatics and Health , author =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Informatics and Health , author =

Reference 18

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.966505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:2d5d74771200ac3cef0d4485803b0816ff52d2ad7791e6e38370b5710315c166

Observation 4ebe1f57-2ff1-4135-a8a4-5ad1fd229787 · outbound

This paper cites In: Intelligent Systems and Pattern Recognition.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks In: Intelligent Systems and Pattern Recognition

Reference 19

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.937555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:42c54e2989122023af075c496819b247b0b88b8072beae2fb852abb8dd022017

Observation fc7da582-45dd-41c5-9bd7-12f8a38694a1 · outbound

This paper cites arXiv.org , author =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks arXiv.org , author =

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.799859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:12f8283bf25e4490dba296ecac2544d91c1ac4c2ac7a632210724db02c1ab68a

Observation 54b51020-f2fd-410b-9653-cd87c3bb3286 · outbound

This paper cites arXiv.org , author =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks arXiv.org , author =

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.791673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:93df39720375761636b206c14d60525b105fa95011cfe4282fa31c245342a5bb

Observation 8af205ca-e01a-4bbd-8640-b3743a3499e6 · outbound

This paper cites and Le, Quoc V.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks and Le, Quoc V

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.775516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:d628e365ea5ffab53256ffaff73ea46a9b59b0f9ca044ac997394b77b031e737

Observation 4a85e997-5c4d-4ab5-9eac-42f34287c9ce · outbound

This paper cites Retrieval-augmented generation for knowledge-intensive.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Retrieval-augmented generation for knowledge-intensive

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.781725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:2d51533c7eb0b86e1039d51e9942297683e78f2890d5139b284681f78abd7072

Observation fa95bdd7-ead9-4742-80f6-8061dace8f19 · outbound

This paper cites Intelligence-Based Medicine , author =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Intelligence-Based Medicine , author =

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:11:21.803037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:1ad6ae5a44da42f7bbdf789b6f423a787499d55bb6c23e119058354a24b2d965

Observation 0a0daf09-b80f-4369-a7fd-e668a9407325 · outbound

This paper cites Hospitals , author =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Hospitals , author =

Reference 26

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.796355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:e39534c3f5fee98157d1fd863b4f6006d66952838e4c88c0e0a5122f0aaba17a

Observation 488f1228-f868-4569-a877-7aa79a7f38cc · outbound

This paper cites Proceedings of the 62nd.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Proceedings of the 62nd

Reference 33

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.782852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:03ac1e13b357a64cd6b958bc4733977cc672d471bc15807f1a4dd47350265d6e

Observation a122d888-a304-4e87-867d-8314501f0cac · outbound

This paper cites and Mordatch, Igor , year =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks and Mordatch, Igor , year =

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.789182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:a0a60fb2646554f36830f815458b087d61f9d6a42c25ae7557784a47ccd1849d

Observation 905c4b48-4599-4961-9bd9-330ec76dcd29 · outbound

This paper cites Alistair Johnson, Tom Pollard, Steven Horng, Leo Anthony Celi, and Roger Mark.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Alistair Johnson, Tom Pollard, Steven Horng, Leo Anthony Celi, and Roger Mark

Reference 35

Resolution
metadata mismatch
doi, observed 2026-05-12T05:11:21.778821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:b3749c263baff78481288e77fdd43effcd7a486112a30006d77b057af27c44a6

Observation 240c9a5d-a271-4a57-b63d-5f36682677b0 · outbound

This paper cites and Shamout, Farah E.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks and Shamout, Farah E

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.755061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:cb5fbd6fa837db377da6e8e5a9dac16911c80594b831a76974fca67ec482853e

Observation cd872ba8-7a51-4d5f-8f33-506de68de883 · outbound

This paper cites 2023 , pages =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks 2023 , pages =

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.757832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:1c3905145b2032d8ee721400273d6b74c443a541677487165ef58fd84122f441

Observation f8d8895d-a904-447c-9dec-45b7e839ee71 · outbound

This paper cites an unresolved cited work.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-05-12T12:06:33.772815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:fbbfa2f3c3fb4ef939fb263bb0ed62c575e23a01f8cd251d01f0aae53b3ba930

Observation 40a28629-120d-4401-98aa-a88c6afe9627 · outbound

This paper cites LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:11:21.960748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:165e5609fee62dc5977fd7bfe44c12e38de46ec15e2aba97b1c0113230fbe599

Observation ca4e5bde-3444-4c15-9ffc-148920d52e5d · outbound

This paper cites HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:11:21.929441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:39bad85d8c9f2a4fbcc1c7915f3fcbee45b379fdb6fd6dc90f3bf2dbce0722c8

Observation 52437781-bb0c-4cf3-8340-61032040776b · outbound

This paper cites Qwen2.5-VL Technical Report.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Qwen2.5-VL Technical Report

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-05-12T05:11:21.811686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:2ab474e7a1ca20473a91ee4b9252902746c7a5e6f9abbb10c74eaae686871196

Observation 86aaf8f5-e963-4ae2-9d27-b3378cb925d2 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-05-12T05:11:21.792229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:028b5a1d80b90bef47048a3e7c0f17ac23563743bcf8abb72872a6ff90e720fa

Observation ec767918-802c-497d-a451-f65b1ab606e5 · outbound

This paper cites Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:11:21.948174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:762dd91a28f42a8014976af3106142deeb75c72772ae5b5336cac67b72813711

Observation 56e64423-ac75-48a6-8270-f1c120f20b7b · outbound

This paper cites Language.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Language

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.762843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:84743ac3d340a9aec45a06ec4d082eeaf43eb62a8bcee9d2fc82d3408d0ba108

Observation c77174ce-82b6-4854-867a-1ee4d4761337 · outbound

This paper cites Self-Refine: Iterative Refinement with Self-Feedback.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Self-Refine: Iterative Refinement with Self-Feedback

Reference 47

Resolution
metadata mismatch
local_arxiv, observed 2026-05-12T05:11:21.918839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:c59a3fb309bf4475159ff4a042f9ad96d954d3dc281bfc565d4dcb589318d0f6

Observation f195de01-cbd0-49f6-a370-356ad60be745 · outbound

This paper cites Traj-CoA: Patient Trajectory Modeling via Chain-of-Agents for Lung Cancer Risk Prediction.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Traj-CoA: Patient Trajectory Modeling via Chain-of-Agents for Lung Cancer Risk Prediction

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-20T00:00:25.943863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:8c328a06b700715c49f3c8c43f29f2f1737545ec8b56266df22bdb9fc8b9fe5d

Observation fcf00c6d-9c40-4731-a2a0-1773ae5a915a · outbound

This paper cites Proceedings of the 29th.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Proceedings of the 29th

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.770268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:4fdd7cab41f4678489f1cd8a6daaaa8ade82422e6878ef677184b1104fcf0053

Observation 33a81d47-ec96-4a4e-b5b6-882f7cd473b2 · outbound

This paper cites an unresolved cited work.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-05-12T12:06:33.794149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:99200e90ade518dc6208a32e24133295b7602b2a358061c0ffbac36bba08c1d3

Observation 7f28da05-2d7c-4cd3-80df-b001161a214a · outbound

This paper cites ClinicalBERT: Modeling Clinical Notes and Predicting Hospital Readmission.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks ClinicalBERT: Modeling Clinical Notes and Predicting Hospital Readmission

Reference 54

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T09:26:53.041773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:661c493304e46ddb40bf86516ff635db14f71201e3def939cdd0c93c7f8f2a07

Observation 98ff3b63-8d61-4e03-94ad-e0a2a3ef00fc · outbound

This paper cites an unresolved cited work.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-05-12T12:06:33.796800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:e16cadf8f497ec05ac9b63760350618994781f4e8a1cc70138f88289fda35951

Observation fdef00c9-901d-4245-bd75-1a75f80e2dec · outbound

This paper cites Warren, Lu Cheng, Haidar M.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Warren, Lu Cheng, Haidar M

Reference 56

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:11:21.883326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:8b0340ce31f6baafb0de2f97b043425c6361ca2ef6db9e804adf3a90260c21c6

Observation 9eb3fa15-e481-4e43-9556-5eed798faa8a · outbound

This paper cites Dreher, T.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Dreher, T

Reference 57

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:11:21.773962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:0f0c74c8dc67644f52c6b54386af4e0adedd918489156b6d2989da06ea8aff59

Observation 4b10719a-0204-4f9a-b429-d8178639a960 · outbound

This paper cites Guttag, and Adrian V.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Guttag, and Adrian V

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:11:21.895181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:7e52903c8847fc4112fb6569ccd7d595d7acd20a8ac548bf294ab16c74438b91

Observation 416f1439-4d92-4719-b49c-02220db8ffe6 · outbound

This paper cites Enhancing diagnostic capability with multi-agents conversational large language models , volume =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Enhancing diagnostic capability with multi-agents conversational large language models , volume =

Reference 59

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.821234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:09536425b9459280792fcb40dbbab23cb6b28e03b801b2e181e9953144deb903

Observation 70b9a469-ffcd-4db9-b744-503d77b0c2bb · outbound

This paper cites Proceedings of the 15th.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Proceedings of the 15th

Reference 60

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:11:21.868133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:d37b15112a97daee4d24dda3b3e22c8992b10faae921c185bff77e0b0e63dca3

Observation c98274ef-ae7a-47d8-90fa-23d7b9ae7513 · outbound

This paper cites npj Digital Medicine , publisher =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks npj Digital Medicine , publisher =

Reference 61

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.902727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:2fae7be368e6cbfa228bc57d5be9fd72c129364473ba8273fd934a9879c9b7d9

Observation 27c91e59-f0fe-4ec6-a551-cde81b439828 · outbound

This paper cites Mmedagent: Learning to use medical tools with multi-modal agent.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Mmedagent: Learning to use medical tools with multi-modal agent

Reference 62

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.852360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:d52acef12848f40790f65cdf3ee87d82134e94ceb85eada09604f101a71e2d38

Observation b14b127e-7069-46b9-9cd7-06700aedcde6 · outbound

This paper cites GPT-4 Technical Report.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks GPT-4 Technical Report

Reference 64

Resolution
metadata mismatch
local_arxiv, observed 2026-05-12T05:11:21.750668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:44cff12571fa5db37ff9b8837d2a84cf703c12559599f474c8243da4c524bc41

Observation 01f38518-7563-454c-b0b5-4941b0016f84 · outbound

This paper cites Learning.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Learning

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.808654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:43e246effc34f0020680574a1ec6b8c597f8034a154ae55f3c134b92c9b6819a

Observation 209f307d-e45b-4695-8214-e9cc2fdf4615 · outbound

This paper cites 2025 , journal=.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks 2025 , journal=

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.802570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:e4fc539ff762b3f2e93e43920eb8ca25ca9d654ae1b0b09638c9cba638c26fb9

Observation 71ae8a11-63d2-4edb-b22f-63b0737f842a · outbound

This paper cites MedGemma Technical Report.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks MedGemma Technical Report

Reference 68

Resolution
metadata mismatch
local_arxiv, observed 2026-05-12T05:11:21.714853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:f59ab3a677e81a32993da980e4712d5cb4d7221ebef2a6269c9a746d58730445

Observation 219aebf7-7581-484f-b603-49a233befad6 · outbound

This paper cites Vision Language Models in Medicine.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Vision Language Models in Medicine

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:11:21.670645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:0ca897301b79bed3dddc50742c243ca6fe7c7417ffac1df1b4f05d423d513928

Observation bd347140-57e0-412e-bc3c-a2c4120d35ad · outbound

This paper cites Multimodal.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Multimodal

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.858099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:8d169d506765b78c44709759c00a4858f513794b7348da02d087b28d6aa4fc25

Observation 9abd689d-4421-495c-9b34-1ffda86b1ecc · outbound

This paper cites Guyon and A.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Guyon and A

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.811845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:e8b7133b2cdad91ea7d9ef650179b316e98b03a670df588d2dfd1ef489faef36

Observation 1ad8a2c5-ea43-42fa-8a0f-0aa296708d24 · outbound

This paper cites Guyon and C.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Guyon and C

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.805657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:909e49abbeb12f62b3dd566d72b846c503ab7270e7e24c0ef77b585e6750211d

Observation 505d886b-643f-4855-89c6-9cca37dc03f1 · outbound

This paper cites MIMIC-III, a freely accessible critical care database.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks MIMIC-III, a freely accessible critical care database

Reference 75

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.519615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:305b3b9c3532f8a48fd27e1297ec14623ef7a5786754aa5344606a02de30bada

Observation 68de056f-2ebf-4b76-9ed0-ec803bdc543e · outbound

This paper cites MIMIC-III clinical database.PhysioNet, September 2016.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks MIMIC-III clinical database.PhysioNet, September 2016

Reference 76

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.564005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:8caa18ff9202f5acc4e544dff11e1e4d822df88205d85dcd159401b6752192f8

Observation ced4230b-1fa7-4f48-bcd7-a6e3bc42aecc · outbound

This paper cites Clinical risk prediction using language models: benefits and considerations.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Clinical risk prediction using language models: benefits and considerations

Reference 77

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.593576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:50f21620c4d9cef52b0611ed355b153209f5172a5594fdce53e93874f5e31a40

Observation 67e38d2e-4a99-437e-8656-dfd67e9ce55e · outbound

This paper cites A pragmatic randomized controlled trial of ambient artificial intelligence to improve health practitioner well-being.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks A pragmatic randomized controlled trial of ambient artificial intelligence to improve health practitioner well-being

Reference 78

Resolution
metadata mismatch
doi, observed 2026-05-12T05:11:21.649627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:73e41626fa2b25e78385b1075ad142c391bfdc4fc50d3f2d36cec0ba3c3ed944

Observation 6d4620d1-11c9-432e-9a7e-46450ad3879f · outbound

This paper cites an unresolved cited work.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Unresolved cited work

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:11:21.695874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:41e40abb4944512e09ba721ccbfd9a6b4656cff4eb7ccd2f86b1533ea28eb8ad

Observation e23ac393-0a15-4091-91bb-094016ca50a3 · outbound

This paper cites EHRXQA : A Multi - Modal Question Answering Dataset for Electronic Health Records with Chest X -ray Images.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks EHRXQA : A Multi - Modal Question Answering Dataset for Electronic Health Records with Chest X -ray Images

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.818365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:4f90a4ed35579bc3153c5d7635ef1ef130f1389b8f5ac3de846baf788f33925c

Observation 909dd84a-8f62-49c1-b315-c0228c72bd94 · outbound

This paper cites Qwen2.5-VL Technical Report.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Qwen2.5-VL Technical Report

Reference 81

Resolution
verified exact
local_arxiv, observed 2026-05-12T05:31:25.830002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:e0c89f07855f0883c4087873c63a7ea2dac593e267b653ab3335a59f478dd10a

Observation 68fac71c-0066-4905-949c-40c69955e1a3 · outbound

This paper cites Bicknell, Danner Butler, Sydney Whalen, James Ricks, Cory J.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Bicknell, Danner Butler, Sydney Whalen, James Ricks, Cory J

Reference 82

Resolution
malformed identifier
doi_truncated, observed 2026-05-12T05:11:21.550928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:4b98040438a12e7741ce616f5492a999ecc03ac102d8b6d310067a874bc226ec

Observation 48d31f79-ad17-49fd-8dd3-458f78165c3c · outbound

This paper cites Language Models are Few-Shot Learners.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Language Models are Few-Shot Learners

Reference 83

Resolution
verified exact
local_arxiv, observed 2026-05-12T05:31:25.817942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:da46251c4383ea1ced218cbc7f0fbb1401329d5e5f932ae506248d18e70e7a50

Observation 18b31044-ce80-4cf9-9db5-8a359be30395 · outbound

This paper cites Knowledge and perception of primary care healthcare professionals on the use of artificial intelligence as a healthcare tool.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Knowledge and perception of primary care healthcare professionals on the use of artificial intelligence as a healthcare tool

Reference 84

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.568768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:310bb541aa4b9551177f7c1e6c6836b8c44e8534036227dfe20d3571ffc1a9d9

Observation d0d71ce1-715c-4037-815f-5e25acfa5a8f · outbound

This paper cites Why Do Multi-Agent LLM Systems Fail?.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Why Do Multi-Agent LLM Systems Fail?

Reference 85

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:42:59.242121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:2d566e313939480f186dbeb478dbb7d229d766a9ddbfb8cb4aab6b1ddbf97452

Observation 2b5860c5-77d2-4320-a8c8-80f39e1c6c10 · outbound

This paper cites Multimodal Clinical Benchmark for Emergency Care ( MC - BEC ): A Comprehensive Benchmark for Evaluating Foundation Models in Emergency Medicine.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Multimodal Clinical Benchmark for Emergency Care ( MC - BEC ): A Comprehensive Benchmark for Evaluating Foundation Models in Emergency Medicine

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.824268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:13ff940fb238bbaa01c970efcbaaa923165ee23e1d068ecec7b37be0398e724f

Observation 60bd4685-df30-430c-950e-d8bd3138c0a9 · outbound

This paper cites Multi-modal learning for inpatient length of stay prediction.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Multi-modal learning for inpatient length of stay prediction

Reference 87

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:11:21.577644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:50ace75a455600a90def26383033439ce2dc933ad7445bb7523eca2f1f6ed3ec

Observation afc2741f-a9a7-4cdb-a1c9-2a62f8bc5c6a · outbound

This paper cites HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale

Reference 88

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:31:25.823602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:9a8582b751d804d5834081b3d43132b6fdbc756bc4da3a08da3b25e266576c76

Observation d5247dc3-f81b-4fa4-acac-2b9085bb4953 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 89

Resolution
verified exact
local_arxiv, observed 2026-05-12T05:31:25.864501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:d7936e399a830d4efa5ecf9fbfd64fef56f132a890ef998a48317fff14eea090

Observation c974c379-564d-4b0c-abe6-e3767e0588b8 · outbound

This paper cites Tenenbaum, and Igor Mordatch.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Tenenbaum, and Igor Mordatch

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.820841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:b2bc05f373228e87574a38435c96263dfe21b9e5171b393b4e244f7ab0d572c8

Observation d16899d3-b93c-4195-8e63-4aeef8255d9e · outbound

This paper cites Geras, and Farah E.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Geras, and Farah E

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.837974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:c10be8f681f3ce598cadc81996308f2f11c7fce6a2922f93076afec67cc85835

Observation e536169e-f640-4b34-afcf-23c003c59f29 · outbound

This paper cites Single-agent or Multi-agent Systems? Why Not Both?.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Single-agent or Multi-agent Systems? Why Not Both?

Reference 92

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:31:25.871484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:49922efacc579f408e5feae94e12be3a7f811cd336fc51cf5825413441903c6d

Observation d3257e39-8c25-4988-9763-48db517776f3 · outbound

This paper cites Geras, and Farah E.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Geras, and Farah E

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.860483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:28a9a52f30253223a41800a4ab86a6bd472d6ffe172be96ebcb61658868a4414

Observation 880983be-084e-4578-8b73-e43f98306ae7 · outbound

This paper cites MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework

Reference 94

Resolution
verified exact
local_arxiv, observed 2026-05-12T05:31:25.809518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:1ce2ef77903a74bfa04326fe0939721c5e2f1d3ccb87ffcb863ddfef9051c825

Observation 17a8b695-b08d-4197-916b-0b217f956bfb · outbound

This paper cites MetaPrompting : Learning to Learn Better Prompts.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks MetaPrompting : Learning to Learn Better Prompts

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.855144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:b5d0b3d0735e2b08bd1ac1aca21cc84921185021af99e8b03cfeb9b220006a5f

Observation 94f839cd-c5f8-4e39-a907-5a4ef5585db7 · outbound

This paper cites ClinicalBERT : Modeling Clinical Notes and Predicting Hospital Readmission , April 2019.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks ClinicalBERT : Modeling Clinical Notes and Predicting Hospital Readmission , April 2019

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.849589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:ffceefceeb745a12be2daff9286bfdde69b6f7a7cb4d361210dbccad4120f12b

Observation fc1b640b-e836-4188-8fb8-bb2495924a5d · outbound

This paper cites John Wilbur, Zhe He, R.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks John Wilbur, Zhe He, R

Reference 97

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.644022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:05b6bcc4d5b8da6089cf84eab5ff122e5c2a67a1e9f8fe572d6569f80af16ca5

Observation 78e19f98-e0d7-4cdf-8017-0fe91262e548 · outbound

This paper cites MIMIC - IV - Note : Deidentified free-text clinical notes, 2023 a.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks MIMIC - IV - Note : Deidentified free-text clinical notes, 2023 a

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.852239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:50cf6293ac32dc47d8e753d31fa5f888adc45254afcec493e79ecdda0f534cfd

Observation 7edca90c-5783-400f-9794-93b5a5035b54 · outbound

This paper cites URLhttps://www.nature.com/articles/s41597-019-0322-0.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks URLhttps://www.nature.com/articles/s41597-019-0322-0

Reference 99

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.585214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:3ab5a14ad09b21ca5d631adbd7a111d2ea1b5399d07bfe8ece547569e2496462

Observation a32491b8-8d1d-4e05-9295-0787a40eaf90 · outbound

This paper cites cc/paper_files/paper/2019/file/ ac52c626afc10d4075708ac4c778ddfc-Paper.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks cc/paper_files/paper/2019/file/ ac52c626afc10d4075708ac4c778ddfc-Paper

Reference 100

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.728236Z

Source-reported events for the cited work

correction dated 2023-01-16. Source: crossref record 10.1038/s41597-023-01945-2->10.1038/s41597-022-01899-x:correction, observed 2026-07-11T02:59:03.685257+00:00. This notice travels one citation hop only.

correction dated 2023-04-18. Source: crossref record 10.1038/s41597-023-02136-9->10.1038/s41597-022-01899-x:correction, observed 2026-07-11T02:59:35.186743+00:00. This notice travels one citation hop only.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:cec03339d49db07cf1a42716e566e5731e8c6acdb056bb06e96b3747629aa6b6

Observation e41a7711-bb17-45cd-a70b-fd36762c3e95 · outbound

This paper cites an unresolved cited work.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Unresolved cited work

Reference 101

Resolution
unresolved
raw_fallback, observed 2026-05-12T12:06:33.814562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:77cf7f377d4b54497d368f65df8c946f48d5090940dfb5db7ba6538b242b751c

Observation 8bb30ff7-d8eb-45f4-aa55-0aae2dfc2200 · outbound

This paper cites Voting or Consensus ? Decision - Making in Multi - Agent Debate.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Voting or Consensus ? Decision - Making in Multi - Agent Debate

Reference 102

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.709628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:65144e704ed173995e5805afbaf0dda869168f0b639beac8315a5774fcab46a3

Observation 4c375437-dc30-42cc-8a32-a15a2edceb74 · outbound

This paper cites Vision Language Models in Medicine.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Vision Language Models in Medicine

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:31:25.805519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:7c0966531df9ffec7793c9ef9d7b82d6b340b47f76cf9a8a6a22483c1d172658

Observation eecf895d-669e-4506-8146-44d8c4b6b708 · outbound

This paper cites Clinical Risk Computation by Large Language Models Using Validated Risk Scores.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Clinical Risk Computation by Large Language Models Using Validated Risk Scores

Reference 104

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.660447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:ab350525372276a67254eeea69a63336240b19fbbc805bfbb5bfa21ced1045ca

Observation 623ab034-d58b-450b-8376-039eb80f1745 · outbound

This paper cites Medical transformer for multimodal survival prediction in intensive care: integration of imaging and non-imaging data.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Medical transformer for multimodal survival prediction in intensive care: integration of imaging and non-imaging data

Reference 105

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.664388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:1cc3a87a809776fe6f38812c18d9b47e055d210d063585b920e5f441ddc4229e

Observation 1b77f872-e55c-461f-9748-56824b728b98 · outbound

This paper cites MDAgents : an adaptive collaboration of LLMs for medical decision-making.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks MDAgents : an adaptive collaboration of LLMs for medical decision-making

Reference 106

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.863540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:f749742b30fa76d4540ed23a9c599fb9ae75f6f1fc1fcd2318b698917113f4db

Observation 8cd86349-524f-4bde-b1d0-62c2630dd209 · outbound

This paper cites Towards a Science of Scaling Agent Systems.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Towards a Science of Scaling Agent Systems

Reference 107

Resolution
verified exact
local_arxiv, observed 2026-05-12T05:31:25.833946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:c35074ecdbb6c5febc2ffa2a0b4a48bc5574cef6cae1b256dd347f223757faac

Observation ca0b8734-8b79-49af-8ea6-8bffd0194fde · outbound

This paper cites Bioinformatics , volume =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Bioinformatics , volume =

Reference 108

Resolution
metadata mismatch
doi, observed 2026-05-12T05:11:21.604221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:c1420ffda17956bab24265c04f252f4422f6fbed1a8e3bb3148fedfa43277605

Observation 5b0f7f8c-ea5e-4f72-b302-019b290cd1bb · outbound

This paper cites Learning Missing Modal Electronic Health Records with Unified Multi -modal Data Embedding and Modality - Aware Attention.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Learning Missing Modal Electronic Health Records with Unified Multi -modal Data Embedding and Modality - Aware Attention

Reference 109

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.866329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:f2c1ee939c11ac84d4d7266d705c6978155730b8e9c7820e831d971a8e178bba

Observation d81d31eb-3a70-4359-acd7-6a34611eee8f · outbound

This paper cites A prompt framework for enhancing LLM -based explainability of medical machine learning models: an intensive care unit application.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks A prompt framework for enhancing LLM -based explainability of medical machine learning models: an intensive care unit application

Reference 110

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.634176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:9eeaf944f06b2aaf975e6622127597c5d801138874a0e688ccbdf02a16a9f534

Observation 9c33940e-220e-40d1-9c02-3c07270ed32c · outbound

This paper cites Retrieval-augmented generation for knowledge-intensive NLP tasks.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Retrieval-augmented generation for knowledge-intensive NLP tasks

Reference 111

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:11:21.690309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:5347aff19eaf40e82024d56ba62496ba918832905bf6e05f309dfb23bb8104cd

Observation 5f967fbc-ed65-43e8-b760-3e56eee4a028 · outbound

This paper cites LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day

Reference 112

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:31:25.859967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:449e76d49f5ca75ee7ef8434d484c1640e5fe39dacaaaa09c99349ac9cf2945b

Observation 2a97e58f-bcb2-44cb-a314-fcc5130719e6 · outbound

This paper cites In: Al-Onaizan, Y., Bansal, M., Chen, Y.-N.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks In: Al-Onaizan, Y., Bansal, M., Chen, Y.-N

Reference 113

Resolution
metadata mismatch
doi, observed 2026-05-12T05:11:21.616113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:c16da035e278abdc36759f5fbdeda37b7580ed63ae2c7e53525b47201e1a2cae

Observation 6f9506d9-8c65-4988-b182-b5d58158b3fc · outbound

This paper cites Ovis2.5 Technical Report.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Ovis2.5 Technical Report

Reference 114

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:30:17.288454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:b01e87f3da3770c0dc51342d5fc916104802e00b726a096b30fda70b10359532

Observation 3d247753-e3cc-4085-a355-485dbc5deb45 · outbound

This paper cites Lukac, William Turner, Sitaram Vangala, Aaron T.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Lukac, William Turner, Sitaram Vangala, Aaron T

Reference 115

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.705146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:3e3d2116ed9797b46709663b1b5fb0ea43a376dafbd84bc28153d84804e8529b

Observation 957eaeca-ed5c-487a-b836-f0ae078f6602 · outbound

This paper cites Self-Refine: Iterative Refinement with Self-Feedback.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Self-Refine: Iterative Refinement with Self-Feedback

Reference 116

Resolution
verified exact
local_arxiv, observed 2026-05-12T05:31:25.800945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:e1ffdbb155e114cca59a70655d7090eb0fc747b6b1a449698a7b7e5402ebd315

Observation ed19ff2b-9d15-4ee3-8e36-240d3f797754 · outbound

This paper cites Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine

Reference 117

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:31:25.879478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:b769d4d4a610f02853cad1d7e1e100836910195927376433d37121a2d958ea86

Observation 86e46fa3-88c8-4a47-aaa2-b62cb7255e53 · outbound

This paper cites MedGemma Technical Report.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks MedGemma Technical Report

Reference 118

Resolution
metadata mismatch
local_arxiv, observed 2026-05-12T05:31:25.875394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:4fa7ea8d99bf575b3abd97105995380b27a3dcd1c13a36ec091750b2be4e2bfa

Observation 817be877-f2e7-499a-b59d-a486c346e5d5 · outbound

This paper cites Transforming Healthcare with AI : Promises , Pitfalls , and Pathways Forward.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Transforming Healthcare with AI : Promises , Pitfalls , and Pathways Forward

Reference 119

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.555303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:87974d8b7d3c4dc25fdf164e19766e902fb20c5e58138076d7ee7b2e67c52aad

Observation 5198008f-61c7-48f3-95e0-fc1f6843e3d0 · outbound

This paper cites Large Language Models Encode Clinical Knowledge.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Large Language Models Encode Clinical Knowledge

Reference 120

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.738778Z

Source-reported events for the cited work

correction dated 2023-07-27. Source: crossref record 10.1038/s41586-023-06455-0->10.1038/s41586-023-06291-2:correction, observed 2026-07-11T03:08:19.417011+00:00. This notice travels one citation hop only.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:4114e5523db9e691f427aada2c9cd1bc70bb805edb6042c312653c618ba7034d

Observation 9631a6e4-7c32-4043-aee4-161770b7c85f · outbound

This paper cites Singhal, T.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Singhal, T

Reference 121

Resolution
metadata mismatch
doi, observed 2026-05-12T05:11:21.677554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:47ddb805ce5d657e9090311aef77829933cbe1b0a1fc4724f8835e7d0863f62f

Observation 2baaec85-dd7c-41c4-8242-9cfcad0aae43 · outbound

This paper cites Baxter, Florin Vaida, Amanda Walker, Amy M.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Baxter, Florin Vaida, Amanda Walker, Amy M

Reference 122

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:11:21.640154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:cf2c0695f5dd7faa4443378e55c78493b88c2e645975759d2e05f2f366d5bea2

Pith citing papers

Observation 8453a3f8-b7b9-43d6-8e4d-98d4a510b9fd · inbound

The Path to Self-Evolving Clinical Systems: Scaling Medical Agents from Assistance to Autonomy cites this paper.

The Path to Self-Evolving Clinical Systems: Scaling Medical Agents from Assistance to Autonomy AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks

Reference 111

Resolution
unresolved
no resolver link, observed 2026-07-14T06:30:16.612345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:30:16.612345Z digest=sha256:c2011fa979879404052212c5938f467e58ba4ce445fb1b8726ae7c9b44ec73a8

Observation 135a7814-95d7-487d-8dd8-857f1181faa7 · inbound

Model or Harness? An Interaction-Centric Taxonomy for Localizing Agent Failures cites this paper.

Model or Harness? An Interaction-Centric Taxonomy for Localizing Agent Failures AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-03T00:25:11.457757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:25:11.457757Z digest=sha256:dade10e02697d8354e9b2a6503ec50928226785523274eecb0cdf59b3a5cdc48