Pith. sign in

Paper Citation Record · LEDGER

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning

As of 8 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 0 inbound Pith citation observations for arXiv:2510.19893.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2510.19893 v2

Coverage vector

measured 29 of 29 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T08:38:31.414748Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

29 of 29 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved27
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 99514345-9b0f-4734-928d-ca401d32c654 · outbound

This paper cites Qwen2.5-VL Technical Report.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning Qwen2.5-VL Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:29.467948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:29.467948Z digest=sha256:403956e9fd92cca6582bd23b558b6aad6aae4cf03a40c0ac83b8919da0c6631b

Observation 78ed3075-01fc-457f-879d-f30e2583fa1d · outbound

This paper cites FAIRWELL: Fair Multimodal Self-Supervised Learning for Wellbeing Prediction.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning FAIRWELL: Fair Multimodal Self-Supervised Learning for Wellbeing Prediction

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:29.603111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:29.603111Z digest=sha256:ed2103f1f409d7865a29e39302614289b52a6b31bcffea6ba5e8519eff736030

Observation e6ec0487-6661-4cba-a35c-939d9bebe63e · outbound

This paper cites Biomedical Visual Instruction Tuning with Clinician Preference Alignment.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning Biomedical Visual Instruction Tuning with Clinician Preference Alignment

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:29.667730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:29.667730Z digest=sha256:5214a4a4a12abad9a6fa9efcd747b0bdaef5f80897b2f24dd3d98cb91a4a5ab1

Observation ac188622-3c7c-4b55-954d-1ffed0c62dd1 · outbound

This paper cites Developing icu clinical behavioral atlas using ambient intelligence and computer vision.NEJM AI, pp.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning Developing icu clinical behavioral atlas using ambient intelligence and computer vision.NEJM AI, pp

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:29.726680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:29.726680Z digest=sha256:d69c639a51242b93ff8141b1678b6fe614d3cb50c7559ecef2f478c0f3f98707

Observation 94173fae-d93c-4419-ac42-442a989c0d82 · outbound

This paper cites Graph-based patient representation for multimodal clinical data: Addressing data heterogeneity.medRxiv, pp.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning Graph-based patient representation for multimodal clinical data: Addressing data heterogeneity.medRxiv, pp

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:29.793418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:29.793418Z digest=sha256:8216072fcbd27cf4bdf656dfa83aa24083a601492e0eb9c24a4bfd1167f73b80

Observation ce73e2e2-8f06-46da-ab60-37eaa7b350e0 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:29.856920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:29.856920Z digest=sha256:dc89530a04b0d1dd44d4a5f5ba9dbe8823a8087ceb8bfc054e803375666b1f0a

Observation d8e58d44-5d98-4311-b033-d3d18261eb11 · outbound

This paper cites Ball, Katie S.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning Ball, Katie S

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:30.029390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:30.029390Z digest=sha256:dbc2808621f54f4b2c2ce4c75d927563d874676ad91d8440570e18d35c27c22a

Observation 5f92d917-d26c-446b-ac1b-7cb18c5f4ded · outbound

This paper cites Seungeun Lee, Yongwon Cho, Yuyoung Ji, Minhyek Jeon, Aram Kim, Byung-Joo Ham, and Yoon- jung Yoonie Joo.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning Seungeun Lee, Yongwon Cho, Yuyoung Ji, Minhyek Jeon, Aram Kim, Byung-Joo Ham, and Yoon- jung Yoonie Joo

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:30.181588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:30.181588Z digest=sha256:589b6cb9bbaf07b7ab0901e18f3ffbf519e14866fb8d2e337639eecd2efa456c

Observation 3eb385f9-3bb3-4e24-b101-efdcb9854dc9 · outbound

This paper cites Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:30.244748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:30.244748Z digest=sha256:48bb08ff02e6211c732b72a525be7596a9ce36ad6abd47d15fa779609d12c9cf

Observation dd33572c-5ff9-47de-a4c9-71f2b0acba49 · outbound

This paper cites URL https://doi.org/10.1016/j.dib.2020.106221.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning URL https://doi.org/10.1016/j.dib.2020.106221

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:30.295869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:30.295869Z digest=sha256:119599d7f149bd66864104ee5b518066f408668a6cf3e86c4311db951342fca3

Observation 81f078de-4b7c-401b-b3cb-9c5b0b4f8058 · outbound

This paper cites Proximal Policy Optimization Algorithms.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning Proximal Policy Optimization Algorithms

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:30.409955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:30.409955Z digest=sha256:2026f10a09eec3f017c3651049920d015ef610e93fb372a5b4fe17cc694cd475

Observation 25bc617e-76de-49eb-b4e6-9bfa91670481 · outbound

This paper cites an unresolved cited work.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning Unresolved cited work

Reference 19

Resolution
verified exact
doi, observed 2026-08-04T08:44:11.969588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-04T08:38:30.555955Z digest=sha256:d46df282c8b23110fd01a6c9ee868991fa00375eb1c9b8a84ddb4a94da27eb85

Observation cd8d6721-e565-4562-bb65-df18011a677f · outbound

This paper cites an unresolved cited work.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning Unresolved cited work

Reference 21

Resolution
malformed identifier
doi_truncated, observed 2026-08-04T08:44:11.614751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-04T08:38:30.681118Z digest=sha256:bcd75510fb68f327b4a68911ec6624b26ad3452bb91cb02ef19e8dadcfebdcf4

Observation e13303e6-4d85-46e5-acf5-4517b00e40bd · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:30.615003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:30.615003Z digest=sha256:59f4955fe7080e5c3b940268df1be3cf1042c99e5e2d552dc7efda45d7ff4de7

Observation 47285338-9244-4228-b949-1e6bbd424d92 · outbound

This paper cites Recorded Feb–May 2021 at Maas- tricht University Medical Center (UMC+); CC BY-NC-ND 4.0.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning Recorded Feb–May 2021 at Maas- tricht University Medical Center (UMC+); CC BY-NC-ND 4.0

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:30.866310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:30.866310Z digest=sha256:71d3b4bd7f4bc5602bf6648463cdb20a6c377f120d42d5bb1f3afddbcdbd33de

Observation d32ae47b-3ce2-4243-bf23-f687b1db9c4a · outbound

This paper cites Multimodal healthcare ai: identifying and designing clinically relevant vision-language applications for radiology.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning Multimodal healthcare ai: identifying and designing clinically relevant vision-language applications for radiology

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:30.917886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:30.917886Z digest=sha256:8f8a218cda35a67bdccd890d95da069da72db4ef0f95681ae44814e7ed02bdfe

Observation ae40702f-afda-40cc-859f-8b711cdcc2e5 · outbound

This paper cites All models are trained with 4 NVIDIA H200 GPUs.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning All models are trained with 4 NVIDIA H200 GPUs

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:31.044763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:31.044763Z digest=sha256:141cbd82627942366dfe1f8397889fdf0b009bebbdcdb6a09d942c0d444c5aa2

Observation 7c40917a-a639-4e1b-bf7a-822c1a59c871 · outbound

This paper cites The dataset includes local labels for bounding boxes; however, we evaluate our models based on the 5 global labels for BI-RADS 1-5.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning The dataset includes local labels for bounding boxes; however, we evaluate our models based on the 5 global labels for BI-RADS 1-5

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:31.144752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:31.144752Z digest=sha256:63ff2f0ce746a1a69e7f6be1d1faa7843d759efcb228ca4c99434d41885697bb

Observation 2e8b9984-5a2a-45dd-a47d-cb7cbcc13ec7 · outbound

This paper cites Malignant.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning Malignant

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:31.296122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:31.296122Z digest=sha256:55da9bcfbc650d9ca4a9f3ece0e624bb258ce17a1a2015e86e8f3283d6f8cd58

Observation e34d0268-07e1-4093-95a6-2a54c7eb2726 · outbound

This paper cites No Hemorrhage.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning No Hemorrhage

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:31.414748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:31.414748Z digest=sha256:7ef5cccc405663c69704d1cf9d51c3b768d1c6a0ae3452b387d1cfb673451719

Observation f177acab-c761-4846-93ed-962db2087851 · outbound

This paper cites The HAM10000 dataset, a large collection of multi-source dermatoscopic images of common pigmented skin lesions.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning The HAM10000 dataset, a large collection of multi-source dermatoscopic images of common pigmented skin lesions

Reference 1953

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:30.788099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:30.788099Z digest=sha256:8911d60a08444fec66c250c9d17169ceb4d228e47f9c61b50a1cdfa600db4906

Observation ada3d564-0d4b-4160-b7a9-bc31280209b2 · outbound

This paper cites MedGemma Technical Report.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning MedGemma Technical Report

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:30.486879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:30.486879Z digest=sha256:5db116e548b954f339f90ca2513d38f237108ebe42709993de44d72501a6dbb3

Observation 9c75a421-4132-4de0-bca3-2ea33992a706 · outbound

This paper cites URL https://doi.org/10.1609/aaai.v33i01.3301590.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning URL https://doi.org/10.1609/aaai.v33i01.3301590

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:30.079948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:30.079948Z digest=sha256:affcfe9a3698514c40ec79b8804c1499d3f553c5a08486da7caf582ec3ac28d5

Observation 3ddeea14-e0cd-4a05-9c4e-d9c59c95fa86 · outbound

This paper cites REINFORCE++: Stabilizing Critic-Free Policy Optimization with Global Advantage Normalization.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning REINFORCE++: Stabilizing Critic-Free Policy Optimization with Global Advantage Normalization

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:29.923727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:29.923727Z digest=sha256:e7a3450aa93e7c695be62144c33bcc43462faf9af50e5780305f3102821b30b9

Observation 54e859b1-f6d4-44fa-a27a-f2d60d178ff2 · outbound

This paper cites Distributionally Robust Neural Networks for Group Shifts: On the Importance of Regularization for Worst-Case Generalization.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning Distributionally Robust Neural Networks for Group Shifts: On the Importance of Regularization for Worst-Case Generalization

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:30.356458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:30.356458Z digest=sha256:badcd68e284d1710ccbadfbfc416924197dcd580aeefc062b8338ad325a99747

Observation 8033e5c2-ab9a-4743-9fcb-142018c0069b · outbound

This paper cites Language Models Get a Gender Makeover: Mitigating Gender Bias with Few-Shot Data Interventions.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning Language Models Get a Gender Makeover: Mitigating Gender Bias with Few-Shot Data Interventions

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:30.738979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:30.738979Z digest=sha256:d2dd6f56a1d4a55afd7fac4329e7b98847d9a082129028653cb09d775845ec06

Observation 95f954f5-f207-484b-a43c-a7c479063c5d · outbound

This paper cites Back to basics: Revisiting reinforce-style optimization for learn- ing from human feedback in llms.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning Back to basics: Revisiting reinforce-style optimization for learn- ing from human feedback in llms

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:29.326888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:29.326888Z digest=sha256:1e4e54badf9241b2c8147464139e0fc722867dfdd5d22fde4b86254d260a8720

Observation efe0f09d-e1de-4798-b828-0515cece9614 · outbound

This paper cites URL https://doi.org/10.18653/v1/2024.acl-long.662.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning URL https://doi.org/10.18653/v1/2024.acl-long.662

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:29.395721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:29.395721Z digest=sha256:a7ea04791a58a1b50cefc4bd3afead20db02776a75e40cc61210c799baaace3f

Observation 56cba47a-9fa3-4fae-9113-f35b1a52479a · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

EQPO: Equitable Group Relative Policy Optimization for Clinical Reasoning Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-04T08:38:29.532523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:38:29.532523Z digest=sha256:cfbc0f08b2474566d63725bc9f267a0878d7a108ea5fc5b50f11edf7eefcb06d

Pith citing papers

No inbound Pith citation observations are available.