Pith. sign in

Paper Citation Record · LEDGER

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis

As of 20 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 0 inbound Pith citation observations for arXiv:2506.02987.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02987 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:15:38.688229Z

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

12 of 12 outbound references displayed

  • verified exact4
  • verified fuzzy5
  • unresolved3
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation bd23ae5e-f604-45a1-8023-e14b33fa24b9 · outbound

This paper cites Conflict of interest RA declares no competing interest.

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis Conflict of interest RA declares no competing interest

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:40.233145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T11:15:38.087618Z digest=sha256:6eefea49f9ed7f3d9f5cc750f612138e37c482aa4a213ea2e5afc1a771225b9b

Observation 50ec253c-a1d6-4a50-824b-41b164091fe1 · outbound

This paper cites an unresolved cited work.

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis Unresolved cited work

Reference 3

Resolution
verified exact
doi, observed 2026-08-07T11:15:39.875611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T11:15:38.108248Z digest=sha256:0332555b4f0a2e2c12474a908ffc10266c0e078208f3960fa4e4ba50e9c6fa1b

Observation 98e8fb7e-fa1c-474f-a063-8708666bb6e5 · outbound

This paper cites Introducing Claude.

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis Introducing Claude

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:39.725469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T11:15:38.250694Z digest=sha256:610a5298fe323bf45fe789b29120c49a705d996557cd0a5fefe93986b38a9a5a

Observation 752f3d48-72ba-4068-b269-3bcc4864de1b · outbound

This paper cites Grok 3 Beta — The Age of Reasoning Agents.

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis Grok 3 Beta — The Age of Reasoning Agents

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:39.547996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T11:15:38.323288Z digest=sha256:0d1951e6516e0dfa84c7a34acc8af5aa159613c7bf562ec4b31c08bf07b8e7c6

Observation 519ca75b-41e8-4300-b180-714addb07f4c · outbound

This paper cites Introducing Gemini 2.5 Flash, Veo 2, and updates to the Live API.

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis Introducing Gemini 2.5 Flash, Veo 2, and updates to the Live API

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:39.400759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T11:15:38.394339Z digest=sha256:7f6d0dc16c23d9800564ba9c565310b101fd7368b9faf0495430901ed963c9b5

Observation 4a04294b-31b1-4e8d-a244-4f5bc3d2084d · outbound

This paper cites Chatbot Arena LLM Leaderboard.

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis Chatbot Arena LLM Leaderboard

Reference 10

Resolution
verified exact
doi, observed 2026-08-07T11:15:39.146017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T11:15:38.541569Z digest=sha256:348380f81f538e3af69a16b12e89caa8694bc807a3ce4236abe42ea82c470966

Observation ec62ecc9-94aa-4e7a-ab3c-ffa79b972445 · outbound

This paper cites an unresolved cited work.

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis Unresolved cited work

Reference 12

Resolution
verified exact
doi, observed 2026-08-07T11:15:38.842099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T11:15:38.688229Z digest=sha256:ca6e485de7e77c46d926c77ebc9c23aa8c1392518662202df492aadd8619a061

Observation e4c38838-f7b0-4170-bf08-25d8469f0699 · outbound

This paper cites ChatGPT for low- and middle-income countries: a Greek gift? The Lancet Regional Health – Western Pacific December 2023; 41: 100906.

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis ChatGPT for low- and middle-income countries: a Greek gift? The Lancet Regional Health – Western Pacific December 2023; 41: 100906

Reference 781

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:38.129977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:38.129977Z digest=sha256:260078fd134bae82fe1464be4ea1129d35d86258004acd4c42fc706d584ac52c

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:38.163860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:38.163860Z digest=sha256:d97dac3277a3bf220544b0ad4017f128830ade1ba3be49707e108aa21f718796

Observation 070da3af-5e8c-4eb3-959b-fbdd892bbc91 · outbound

This paper cites ChatGPT: the threats to medical education.

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis ChatGPT: the threats to medical education

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:38.119235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:38.119235Z digest=sha256:9e28446d79a2e28bdf85490a991d1f86dfa28893a356bdb55f7316e13949e00d

Observation b9f60f6e-14f3-4b8f-b032-01525aaf0531 · outbound

This paper cites Implications of large language models for clinical practice: ethical analysis through the Principlism framework.

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis Implications of large language models for clinical practice: ethical analysis through the Principlism framework

Reference 2024

Resolution
verified exact
doi, observed 2026-08-07T11:15:38.973023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T11:15:38.615120Z digest=sha256:1f8ec22cbb369d8d12d9cf4ed8fc5734ea3bf98e0a2d0f2dcb9122af739f38e2

Observation 369b1296-8afe-49cf-bedb-94d574bd5866 · outbound

This paper cites Each model was prompted to answer as a GP in the UK and was provided with full question information.

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis Each model was prompted to answer as a GP in the UK and was provided with full question information

Reference 2025

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:40.058811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T11:15:38.097306Z digest=sha256:4cbcd0526e71dd11a9f536f79ad2923a5308a20856ba3224fe638e01f2dad751

Pith citing papers

No inbound Pith citation observations are available.