Pith. sign in

Paper Citation Record · LEDGER

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting

As of 7 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 2 inbound Pith citation observations for arXiv:2507.22902.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.22902 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:08:30.490781Z

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-12T02:21:16.825681Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-12T07:41:42.559848Z

Reference resolution

28 of 28 outbound references displayed

  • verified exact0
  • verified fuzzy17
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f1653e6f-6e4e-43fc-a7bc-d8dfdb901bff · outbound

This paper cites Lower Back Pain.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Lower Back Pain

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:34.156489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:08:28.761844Z digest=sha256:4bf78100d7f04742ed37d96ac26b105de23e684eee8ba4842bf27ad9e8d7dd14

Observation 19090acd-74fb-4ed8-92ef-ec0e82e5544f · outbound

This paper cites Washington, DC: Association of American Medical Colleges, 2021.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Washington, DC: Association of American Medical Colleges, 2021

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:35.590065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:08:27.200231Z digest=sha256:73201d7d4daf2929c7dcfe683ced3c2f71e89770584b38964393e4c9fc62ef2b

Observation ab73543e-c008-4c64-83f1-d66d139d7836 · outbound

This paper cites an unresolved cited work.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T22:08:35.434081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:08:27.268965Z digest=sha256:721dbf33c89fa286379cb4fad66ecfeec95f6382d7b9cebf629df1c48a6c7afb

Observation 7c53d8af-22c7-45c9-9552-887fdef9a442 · outbound

This paper cites Arch Intern Med 18: 1377-1385, 2012.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Arch Intern Med 18: 1377-1385, 2012

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:35.319229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:08:27.384233Z digest=sha256:d63fd59404d76217acd35c34ae38dfc83570dc5ac3ec58ccb96394a997bc843b

Observation 02de3724-d4cf-492e-9ad4-8d2d2ce57be2 · outbound

This paper cites New York: Basic Books, 2019.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting New York: Basic Books, 2019

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:35.050725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:08:27.533897Z digest=sha256:a8e52f9586cea7b722fc957a5ac8066e231e5ab1a60a608916ed3c99c4cfdbba

Observation edce1057-7f50-4b21-8f7e-cd67909ceb1f · outbound

This paper cites NPJ Digit Med 4: 93, 2021.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting NPJ Digit Med 4: 93, 2021

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:34.833681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:08:27.663240Z digest=sha256:a60f32ad26255983fbac0ed996591b94c0636c14d7a8546ad7819a004bcd0b3d

Observation 59068972-6822-4a6c-8934-d1dc6201d54e · outbound

This paper cites Ambient AI Scribing Support: Comparing the Performance of Specialized AI Agentic Architecture to Leading Foundational Models.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Ambient AI Scribing Support: Comparing the Performance of Specialized AI Agentic Architecture to Leading Foundational Models

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T22:08:30.804816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:08:27.815801Z digest=sha256:8423aebe0277a8be8122a623d672540553f1202105f548adf69fbf504d345a1f

Observation 24522d0a-1861-4ffb-9a96-a8e34b8b0448 · outbound

This paper cites AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T22:08:27.981124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:08:27.981124Z digest=sha256:2c4ee86b96f2c8688ef97c3c4a8d704a64f10769515e119c760f75c849195160

Observation 01a3414b-a0c5-43dc-b4f3-82baca1b047b · outbound

This paper cites Improving Clinical Note Generation from Complex Doctor-Patient Conversation.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Improving Clinical Note Generation from Complex Doctor-Patient Conversation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T22:08:28.055295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:08:28.055295Z digest=sha256:698392d7824e1e409fadcaf85d2895f071e93b698b686a59f0511e62ba90070f

Observation b1509b49-a100-40c3-8b12-eb6e16c65fcc · outbound

This paper cites Towards accurate differential diagnosis with large language models.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Towards accurate differential diagnosis with large language models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T22:08:28.172326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:08:28.172326Z digest=sha256:236f773ce0d993fdc8a6a24446d2a81df9c2d97859d6b1f034800e9ef358f505

Observation 7c4ee0ab-67be-499a-b1ad-bfff77976ffe · outbound

This paper cites Ann Intern Med 178:498-506, 2025.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Ann Intern Med 178:498-506, 2025

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:34.615507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:08:28.280044Z digest=sha256:4fb09f21e4ecfe0de8fec35fe77d952385a91112f1629916948158b7f551d4dc

Observation 4b8687d1-1c88-470a-9ae5-eb25ded516ad · outbound

This paper cites LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T22:08:28.413051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:08:28.413051Z digest=sha256:961f7a60c2da3479a36022723039441a50bab9a3a90b6595760aa795085d1313

Observation d7b71641-b679-4d82-b8e9-deb1809ba787 · outbound

This paper cites Assessing the Quality of AI-Generated Clinical Notes: A Validated Evaluation of a Large Language Model Scribe.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Assessing the Quality of AI-Generated Clinical Notes: A Validated Evaluation of a Large Language Model Scribe

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T22:08:28.562592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:08:28.562592Z digest=sha256:b496350b3b99b25f63dfdf0cf0ee9cbb0e789bd4e8eab05e7ef6a06e1184da16

Observation a860dda0-c6f7-4706-adbc-2c7c7d28de7b · outbound

This paper cites clinically consistent.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting clinically consistent

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:34.380831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:08:28.686148Z digest=sha256:e343b4a1f93f7ce0a2d9131d92e8da995bd888ba434b87a5efdaf4e0efcd2172

Observation 5951218f-c975-4f04-9405-8579d66a52fe · outbound

This paper cites Eczema” in one SOAP note would be clinically consistent with a diagnosis of “Atopic Dermatitis.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Eczema” in one SOAP note would be clinically consistent with a diagnosis of “Atopic Dermatitis

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:33.925016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:08:28.887270Z digest=sha256:4a5c31ced3bcc396997962e71b0b0114102333e1a669cc80ce3d530c3ee032a3

Observation d0a10454-1b36-4b30-a557-c104d2d6c523 · outbound

This paper cites Sinusitis.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Sinusitis

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:33.750763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:08:29.050632Z digest=sha256:7ddc70feb6b373ebb5706f891a4e2ed195f4faa1884bed34b6ab0d4cc1477056

Observation 1cd01b5e-d58a-4ad5-8c12-07a9c074d8d2 · outbound

This paper cites gallstones.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting gallstones

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:33.510854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:08:29.181371Z digest=sha256:1035c577aa966270146f359e5386d9431e10d030e16396501c435640e676bec2

Observation 75cb798b-31de-493f-888d-8a467ce7fc29 · outbound

This paper cites clinically consistent.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting clinically consistent

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:33.320166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:08:29.309075Z digest=sha256:075f3a5d7c0d03c7d7103ddc974c11256dbcca72e113b2c644ae312c3bd9bf38

Observation 84fdff10-6ae7-499a-8741-f70684d124a0 · outbound

This paper cites Lower Back Pain.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Lower Back Pain

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:33.119545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:08:29.424324Z digest=sha256:9f3f825137d056f6db5a78ab6ddd7d6594f2eb822ec24e312ee7feb0db3b505d

Observation f14d327c-6303-4522-90ac-6411cbe2fc08 · outbound

This paper cites Eczema” in one SOAP note would be clinically consistent with a diagnosis of “Atopic Dermatitis.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Eczema” in one SOAP note would be clinically consistent with a diagnosis of “Atopic Dermatitis

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:32.866796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:08:29.560209Z digest=sha256:e69d02adeedd9c76717f6bb847c0a9a8995c5c1b58aca95a1845d266eb402660

Observation a25741eb-a066-4976-b93e-a0a36d374796 · outbound

This paper cites Sinusitis.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Sinusitis

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:32.644204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:08:29.668139Z digest=sha256:bd3f4dd6978460309c60c02b9c46322b8eb5942c79164a5f39be10d76ecb65e1

Observation 46f96c63-ad5c-42d5-bfbd-d011cb1daf25 · outbound

This paper cites gallstones.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting gallstones

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:32.410739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:08:29.791790Z digest=sha256:3fad430fd128759f25338e908ee73b82d73219f4a8673fae7c15dd92ffdfc2db

Observation 5ff83ca4-6cea-4236-b5c6-4d109b2853c8 · outbound

This paper cites an unresolved cited work.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T22:08:32.192580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:08:29.948546Z digest=sha256:8f18d74f430d2925e6c3d9309016798e985b1b02306e409d8c3abd24e0c710cb

Observation c597a882-8a9f-44b2-9b21-8291b9485525 · outbound

This paper cites an unresolved cited work.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-06T22:08:31.878329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:08:30.075323Z digest=sha256:a018071c84ecc06eb5aed47b299c6e0c6860386c88e8a4a3217df55205a7ba06

Observation 01a8ec30-1e17-4716-8096-1f4c1e676af8 · outbound

This paper cites an unresolved cited work.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-06T22:08:31.634192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:08:30.184215Z digest=sha256:33ade8d221359fe1a8acad39506f5a5367495871263d6229dbcbe0a2993fd8f3

Observation f7554707-485d-4c9e-80a0-9cbc5eb39e39 · outbound

This paper cites Ibuprofen 600 mg three times a day.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Ibuprofen 600 mg three times a day

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:31.394529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:08:30.242958Z digest=sha256:99277b92743b589c1bbb848c33ce4ea95e71663922e742df42f7b07cc8672c52

Observation 8643c278-7af9-4a5c-b77e-e97d3afd7ce3 · outbound

This paper cites an unresolved cited work.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-06T22:08:31.222079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:08:30.375897Z digest=sha256:7f4d32a493c9ce4b13de7e22a65bf60b981cbaf9f8a7d37433b8e41bb26ae397

Observation 12e8e079-c39b-441c-80f9-e4d537dba161 · outbound

This paper cites Acute Sinusitis.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Acute Sinusitis

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:31.054751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T22:08:30.490781Z digest=sha256:0e82237d1f7f9eb4e33f9b2381215cf395bfd6123e9cc493f2b1625665299565

Pith citing papers

Observation 21cb1cac-046c-4efc-8a34-46571cda31a5 · inbound

SymptomAI: Toward a Conversational AI Agent for Everyday Symptom Assessment cites this paper.

SymptomAI: Toward a Conversational AI Agent for Everyday Symptom Assessment Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:46:53.518176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-07T16:17:39.923337Z digest=sha256:4c93ac80360761ab0f76306199f4e4f23d4fdc64279cecc3cbb3e47c1d8fd4c9

Observation 49492762-1d12-45f2-b3b5-26db7a6acfc9 · inbound

SymptomAI: Toward a Conversational AI Agent for Everyday Symptom Assessment cites this paper.

SymptomAI: Toward a Conversational AI Agent for Everyday Symptom Assessment Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:41:42.609389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T02:21:16.825681Z digest=sha256:77768e16faddae8c260455a09d6650371cf1d0363e976e3c583aeb7d2023b24b