Pith. sign in

Paper Citation Record · LEDGER

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis

As of 8 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 0 inbound Pith citation observations for arXiv:2506.02987.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02987 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:15:38.688229Z

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

12 of 12 outbound references displayed

  • verified exact4
  • verified fuzzy5
  • unresolved3
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation bd23ae5e-f604-45a1-8023-e14b33fa24b9 · outbound

This paper cites Conflict of interest RA declares no competing interest.

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis Conflict of interest RA declares no competing interest

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:40.233145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:15:38.087618Z digest=sha256:5a2fcab2a7443014195d72c6520a452f3321114074bca3c4598941d93cf7ae0b

Observation 50ec253c-a1d6-4a50-824b-41b164091fe1 · outbound

This paper cites an unresolved cited work.

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis Unresolved cited work

Reference 3

Resolution
verified exact
doi, observed 2026-08-07T11:15:39.875611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:15:38.108248Z digest=sha256:5964bce9a386715b49830507251e7ebc454d21b0329bd6fa1704a1d641226493

Observation 98e8fb7e-fa1c-474f-a063-8708666bb6e5 · outbound

This paper cites Introducing Claude.

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis Introducing Claude

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:39.725469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:15:38.250694Z digest=sha256:425aa00cb3d7fdea80f420f3f0b33e6bf0da90ae5cb383bd03c6214ec8bf8f9e

Observation 752f3d48-72ba-4068-b269-3bcc4864de1b · outbound

This paper cites Grok 3 Beta — The Age of Reasoning Agents.

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis Grok 3 Beta — The Age of Reasoning Agents

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:39.547996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:15:38.323288Z digest=sha256:99c7a42b8407e3deb7f336f99b552d5dfc91b80e0f2cbd4b93163b4b830d9372

Observation 519ca75b-41e8-4300-b180-714addb07f4c · outbound

This paper cites Introducing Gemini 2.5 Flash, Veo 2, and updates to the Live API.

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis Introducing Gemini 2.5 Flash, Veo 2, and updates to the Live API

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:39.400759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:15:38.394339Z digest=sha256:8d6e14000b226a09b32145a8222187337fbf38bef02a9bac7732253589378cee

Observation 4a04294b-31b1-4e8d-a244-4f5bc3d2084d · outbound

This paper cites Chatbot Arena LLM Leaderboard.

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis Chatbot Arena LLM Leaderboard

Reference 10

Resolution
verified exact
doi, observed 2026-08-07T11:15:39.146017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:15:38.541569Z digest=sha256:2d6a0b26746e3d7d9b7c9cdb57eb5ff0844eca33c1f4fe1821881e7bc51e1805

Observation ec62ecc9-94aa-4e7a-ab3c-ffa79b972445 · outbound

This paper cites an unresolved cited work.

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis Unresolved cited work

Reference 12

Resolution
verified exact
doi, observed 2026-08-07T11:15:38.842099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:15:38.688229Z digest=sha256:ac79c818b292111e00611c7ff3b0ce442a46dba62adf189b25b5618bbaf878ff

Observation e4c38838-f7b0-4170-bf08-25d8469f0699 · outbound

This paper cites ChatGPT for low- and middle-income countries: a Greek gift? The Lancet Regional Health – Western Pacific December 2023; 41: 100906.

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis ChatGPT for low- and middle-income countries: a Greek gift? The Lancet Regional Health – Western Pacific December 2023; 41: 100906

Reference 781

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:38.129977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:38.129977Z digest=sha256:260078fd134bae82fe1464be4ea1129d35d86258004acd4c42fc706d584ac52c

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:38.163860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:38.163860Z digest=sha256:7bb7a365b716da8da33769dffcad3970312648e03b4a7ac34cbc0db844b82dca

Observation 070da3af-5e8c-4eb3-959b-fbdd892bbc91 · outbound

This paper cites ChatGPT: the threats to medical education.

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis ChatGPT: the threats to medical education

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T11:15:38.119235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:15:38.119235Z digest=sha256:9e28446d79a2e28bdf85490a991d1f86dfa28893a356bdb55f7316e13949e00d

Observation b9f60f6e-14f3-4b8f-b032-01525aaf0531 · outbound

This paper cites Implications of large language models for clinical practice: ethical analysis through the Principlism framework.

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis Implications of large language models for clinical practice: ethical analysis through the Principlism framework

Reference 2024

Resolution
verified exact
doi, observed 2026-08-07T11:15:38.973023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:15:38.615120Z digest=sha256:ea383861f2f075e36d9a4b5920a213f5d5a41c46579b3776e3c9b124c52952b0

Observation 369b1296-8afe-49cf-bedb-94d574bd5866 · outbound

This paper cites Each model was prompted to answer as a GP in the UK and was provided with full question information.

Performance of leading large language models in May 2025 in Membership of the Royal College of General Practitioners-style examination questions: a cross-sectional analysis Each model was prompted to answer as a GP in the UK and was provided with full question information

Reference 2025

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:15:40.058811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:15:38.097306Z digest=sha256:8dda3a400a7a4da21db5a760909b11bbbad43104f454fffda20222b291252f0a

Pith citing papers

No inbound Pith citation observations are available.