Pith. sign in

Paper Citation Record · LEDGER

AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2402.09742.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.09742 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-01T06:07:06.879550Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T09:55:40.453086Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 98b970bf-7059-4d7c-a7fb-c9d08f7c29e9 · inbound

Large Language Model Agent: A Survey on Methodology, Applications and Challenges cites this paper.

Large Language Model Agent: A Survey on Methodology, Applications and Challenges AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

Reference 137

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:52:10.745335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-22T21:51:34.309870Z digest=sha256:07674d44e07dc0673822ea9b9fee75cd4b7dc32879aecf9524af6196acb67176

Observation 89d6027a-0986-455c-9000-fd762434b75c · inbound

A Survey of Scaling in Large Language Model Reasoning cites this paper.

A Survey of Scaling in Large Language Model Reasoning AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.176472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:f72bbe3c12665eae00d110e41deb27f8c7da3b68268252b017d59e3aadf8bb43

Observation 9e7061a3-daf2-4256-af21-4dd6aa659358 · inbound

R2MED: A Benchmark for Reasoning-Driven Medical Retrieval cites this paper.

R2MED: A Benchmark for Reasoning-Driven Medical Retrieval AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T13:54:53.084126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-22T13:52:15.835379Z digest=sha256:9a0cbefe1f095c0213a74a88b968aac22e38c348582fc67f7af2ea32138ecb68

Observation 3b5fe226-53a8-40ed-ab21-ac3cf60e143e · inbound

Real-World Doctor Agent with Proactive Consultation through Multi-Agent Reinforcement Learning cites this paper.

Real-World Doctor Agent with Proactive Consultation through Multi-Agent Reinforcement Learning AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T13:42:19.240874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-19T13:41:30.833840Z digest=sha256:6324034bef602419fe0f8e1a8df06667b1837b568a6fabb2cd07ddc92215463d

Observation e36b5d81-1974-42c3-9fea-ee005346e2db · inbound

TinyTroupe: An LLM-powered Multiagent Persona Simulation Toolkit cites this paper.

TinyTroupe: An LLM-powered Multiagent Persona Simulation Toolkit AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T04:57:04.476242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-19T04:54:38.908451Z digest=sha256:cd5f9ac0cae552a1e1d0332d7f97ecddfb4fa419894f0831237950fbebb2a1f7

Observation a66b77e7-aa7a-4a2f-93cd-4fd24b3ed3fa · inbound

Inflated Excellence or True Performance? Rethinking Medical Diagnostic Benchmarks with Dynamic Evaluation cites this paper.

Inflated Excellence or True Performance? Rethinking Medical Diagnostic Benchmarks with Dynamic Evaluation AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T08:12:29.817129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-18T08:12:02.449352Z digest=sha256:9377281ff82f5b0e80acf6b36554fe7436a6b0b2d4df73f719ff0b9bacdf0e53

Observation 638c0e92-3f86-4a48-abac-54aedd7d2639 · inbound

RE-MCDF: Closed-Loop Multi-Expert LLM Reasoning for Knowledge-Grounded Clinical Diagnosis cites this paper.

RE-MCDF: Closed-Loop Multi-Expert LLM Reasoning for Knowledge-Grounded Clinical Diagnosis AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T08:42:37.034816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T08:40:45.666978Z digest=sha256:3f53e597a3aa6f43f2cde1020ca6bf249dd6d352fffbd41a15e6ee7f7fc5cc87

Observation f74c3cef-a37e-47c7-85ac-30d12d2fa661 · inbound

HealthAgentBench: A Unified Benchmark Suite of Realistic Agentic Healthcare Environments for Challenging Frontier AI Agents cites this paper.

HealthAgentBench: A Unified Benchmark Suite of Realistic Agentic Healthcare Environments for Challenging Frontier AI Agents AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-01T09:55:40.454667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T06:07:06.879550Z digest=sha256:cffc722f66995460f34945dfae1cdfb17e51b512bf27d0f24de84d6be8fcbe28