Pith. sign in

Paper Citation Record · LEDGER

Aligning Language Models with Selective Prediction

As of 11 August 2026, this Paper Citation Record lists 79 of 79 outbound references and 0 inbound Pith citation observations for arXiv:2607.03528.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.03528 v1

Coverage vector

measured 79 of 79 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-12T01:51:25.883463Z

measured 79 of 79 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

79 of 79 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved78
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation bec28eef-0332-460a-abc3-7b6f0dca8bd7 · outbound

This paper cites Large language models in law: A survey,.

Aligning Language Models with Selective Prediction Large language models in law: A survey,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:dec076a2bb8deb7045bd0066d0f16fe498a4ca0467801fdd5e409791d5e603d9

Observation 399e735f-572a-42cd-bf71-6dc98bd988e1 · outbound

This paper cites Ai-trader: Benchmarking autonomous agents in real-time financial markets,.

Aligning Language Models with Selective Prediction Ai-trader: Benchmarking autonomous agents in real-time financial markets,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:01987ec21537d73c8f791047d507bb3a3468817ed6b3a7025d9e072096a2d744

Observation 2f82fbcf-4e73-42f1-ad29-f95afd75463c · outbound

This paper cites Evaluating large language model workflows in clinical decision support for triage and referral and diagnosis,.

Aligning Language Models with Selective Prediction Evaluating large language model workflows in clinical decision support for triage and referral and diagnosis,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:079ade1d469d2131d3e85296a808fbfa5fdbf9ef30d5a09d8d2a85fec11a590a

Observation 6bcbc1f7-0d38-492a-8436-436c979c5818 · outbound

This paper cites Evaluating large language models for accuracy incentivizes hallucinations,.

Aligning Language Models with Selective Prediction Evaluating large language models for accuracy incentivizes hallucinations,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:c077d2a8815ab85e7b2d88974152d0e9e295f14c33fd3c7ea07919cf042496d4

Observation 4b862ac0-9538-420a-ba92-390ec667910e · outbound

This paper cites On calibration of modern neural networks,.

Aligning Language Models with Selective Prediction On calibration of modern neural networks,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:85e5ba7e34b4bd23a5fd7c5aa50a8ecde5f0cb88b382ee793b92adb56bca6ae0

Observation 5908e105-8391-4047-9328-0422c771ced7 · outbound

This paper cites Measuring calibration in deep learning.

Aligning Language Models with Selective Prediction Measuring calibration in deep learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:12e17ef6952b80fe45f7434568048e25335fd1d32d2aa5727a74253cbf7faa09

Observation 3eaab900-dcc5-41ef-8372-8955ca9c58ae · outbound

This paper cites Calibrated selective classification,.

Aligning Language Models with Selective Prediction Calibrated selective classification,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:b6638e73c30ff0e435b63e1f5d6b1dcd3d48e2322aafeca8801a66542f5a2acc

Observation 0f78c077-9fe9-46f7-ae84-a7307fc22c0a · outbound

This paper cites Selective classification under distribution shifts,.

Aligning Language Models with Selective Prediction Selective classification under distribution shifts,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:f2bfd8b8d67027faefdf79a3a73a3c3d35fa1e9b1f7e5b253ead9c8f82a01f4a

Observation c774baec-4272-49e4-bc4d-143a69575ff6 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Aligning Language Models with Selective Prediction DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:d1baf38206baa4d9984efb5501e2d63b9cd9df4dd7deb297645acb2d1fa95e99

Observation eebb9ee0-1246-4eb6-8769-406802e96599 · outbound

This paper cites Beyond binary rewards: Training LMs to reason about their uncertainty,.

Aligning Language Models with Selective Prediction Beyond binary rewards: Training LMs to reason about their uncertainty,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:b10c1222bf4fd00bc453e3926f3ff162e832dfcf263cbc1d86e015c4cf3543c3

Observation c2df9701-1132-41bd-8a05-dc0b36ce8292 · outbound

This paper cites Deepseek-r1 incentivizes reasoning in llms through reinforcement learning,.

Aligning Language Models with Selective Prediction Deepseek-r1 incentivizes reasoning in llms through reinforcement learning,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:6fe66d3eb910e262f5591e6dfa3923d4e23e7fbc63f845e1d3dc508158867cb5

Observation fbc0cec5-79ca-4751-9f81-30b4ca35eeda · outbound

This paper cites Reinforcement learning with verifiable rewards implicitly incentivizes correct reasoning in base LLMs,.

Aligning Language Models with Selective Prediction Reinforcement learning with verifiable rewards implicitly incentivizes correct reasoning in base LLMs,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:2965241b6e1b13465a5e89e0cf7ad3bebc35eb666e98ec78d11716f63d8cacff

Observation 36c829d5-e77f-4d76-99c0-7efd896e3611 · outbound

This paper cites How do LLMs compute verbal confidence?.

Aligning Language Models with Selective Prediction How do LLMs compute verbal confidence?

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:580ccb9e350abd2dbd7c74afff185ca6426712c35b3b887edaa7a1e53124b982

Observation c4e87e80-47a9-4f5d-9b85-9b1e9e14f768 · outbound

This paper cites Can LLMs express their uncertainty? an empirical evaluation of confidence elicitation in LLMs,.

Aligning Language Models with Selective Prediction Can LLMs express their uncertainty? an empirical evaluation of confidence elicitation in LLMs,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:d3f03b76a97fb1875dee7111cb0bbc3efdee32c9dcec8ca2e566e91973dcf5c4

Observation 91533b61-79f0-40e5-8524-eb79f5e30dff · outbound

This paper cites Taming overconfidence in LLMs: Reward calibration in RLHF,.

Aligning Language Models with Selective Prediction Taming overconfidence in LLMs: Reward calibration in RLHF,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:ea720a949f4ef543a255bda75272d6978e725c5570d9da5114a5dc1402e1c826

Observation 02c50abb-39b1-4511-a967-adc3d9249c57 · outbound

This paper cites Available: https://openreview.net/forum?id=l0tg0jzsdL.

Aligning Language Models with Selective Prediction Available: https://openreview.net/forum?id=l0tg0jzsdL

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:2e78bb070d45d6a669484690f684baed54a7019749947833a808f326280daa83

Observation 157382eb-bbe4-4978-a19b-a62ae6ab16d3 · outbound

This paper cites A novel characterization of the population area under the risk coverage curve (AURC) and rates of finite sample estimators,.

Aligning Language Models with Selective Prediction A novel characterization of the population area under the risk coverage curve (AURC) and rates of finite sample estimators,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:cbd3562e287dcb7ece18052b6874c282078a40ae06522ccbce6c5fff22c03346

Observation 5f539d47-596b-4aec-ae19-746c6ba8d1aa · outbound

This paper cites A survey of confidence estimation and calibration in large language models,.

Aligning Language Models with Selective Prediction A survey of confidence estimation and calibration in large language models,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:e8082403875adbcf50f21203caccf9a8ab357f02b15bf9eccbec1b644b2a608c

Observation e24d06d2-96c1-409f-855f-b56a4ddae7b6 · outbound

This paper cites A survey of uncertainty estimation methods on large language models,.

Aligning Language Models with Selective Prediction A survey of uncertainty estimation methods on large language models,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:a44da5ad5ae85a333c34032b190961b108bce842208db9c5a41a4c797ed9ff25

Observation 9ada13f3-525a-4f01-ae60-1b957770d3ce · outbound

This paper cites What does it take to build a performant selective classifier?.

Aligning Language Models with Selective Prediction What does it take to build a performant selective classifier?

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:384622c09096ff1bc8a9a6774594dcfcfeaca3809eca4a90f9c4e38ebb365f9d

Observation 19f0741a-6bab-4df4-a59b-d652fba0b1e7 · outbound

This paper cites Adaptation with self-evaluation to improve selective prediction in LLMs,.

Aligning Language Models with Selective Prediction Adaptation with self-evaluation to improve selective prediction in LLMs,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:703c65e639a9ae3d48e4862e9b4c6dc99d557fada116d4066567b110628679c2

Observation 776c2baa-5fa0-4dd9-a576-9e886ba32709 · outbound

This paper cites Calibrating LLMs for selective prediction: Balancing coverage and risk,.

Aligning Language Models with Selective Prediction Calibrating LLMs for selective prediction: Balancing coverage and risk,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:f60076700d7f598cf4ab45d07df4d53b0d93b12865540dd6e2e532ce57d4d3e1

Observation 256b359c-5440-432e-bb12-bce912b72a95 · outbound

This paper cites Optimal strategies for reject option classifiers,.

Aligning Language Models with Selective Prediction Optimal strategies for reject option classifiers,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:2c3981213c0ef85daebe85ba89d81f519f75516b18e42b35243957375e8ba06c

Observation b3d407f7-e19a-45ef-aebd-631deab47f1a · outbound

This paper cites Qwen2.5 Technical Report.

Aligning Language Models with Selective Prediction Qwen2.5 Technical Report

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:046387a967a12a57d3c594ef67b094d749a1999e7ec9b4ab72a5415c2b62dfff

Observation f8694dab-80a0-4ba8-845a-4f062a946013 · outbound

This paper cites The llama 3 herd of models,.

Aligning Language Models with Selective Prediction The llama 3 herd of models,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:02c7586185f9d549f010814b31b9dc7d946ffab6b5bb96db6cf6e3d9a922c3e2

Observation 3fc86f36-b6dd-4f64-80f1-0b2a98837980 · outbound

This paper cites The Llama 3 Herd of Models.

Aligning Language Models with Selective Prediction The Llama 3 Herd of Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:8e372a779abcf49f5029651d8cac88930d5b59730612a7fc3120a21d2c1a534a

Observation 700cbd24-69db-430c-9d94-60ffee5476bc · outbound

This paper cites HotpotQA: A dataset for diverse, explainable multi-hop question answering,.

Aligning Language Models with Selective Prediction HotpotQA: A dataset for diverse, explainable multi-hop question answering,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:8ae70bbfd0a1e1deffae24498b23454a6d7bca61c0a2f79c58443b9760d2b9bd

Observation 6a4d4a1d-dfc8-46ab-93ae-5bb9458538c6 · outbound

This paper cites Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models.

Aligning Language Models with Selective Prediction Big-Math: A Large-Scale, High-Quality Math Dataset for Reinforcement Learning in Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:6bc39a83914751db977eba2d00523c9e16e74ae7f2565c6a8394580b44d716e7

Observation 1f3dab7b-1b16-4dfd-9b34-fae242a03c93 · outbound

This paper cites Measuring short-form factuality in large language models.

Aligning Language Models with Selective Prediction Measuring short-form factuality in large language models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:29bf197012c964b29a643b5cc7cc358484bb7090f1dc6fedbb1b93825940617c

Observation e3012517-8558-48d6-b134-d3e03a9b1072 · outbound

This paper cites TriviaQA: A large scale distantly supervised challenge dataset for reading comprehension,.

Aligning Language Models with Selective Prediction TriviaQA: A large scale distantly supervised challenge dataset for reading comprehension,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:892030562236990cddc30c826613f7f19e55d9e52e61b986d64c21aad76d2ee0

Observation 2cdb5e26-c61a-4b73-b9cc-6f2b646c2588 · outbound

This paper cites GPQA: A graduate-level google-proof q&a benchmark,.

Aligning Language Models with Selective Prediction GPQA: A graduate-level google-proof q&a benchmark,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:5c011d33b70d523823948da4cb05662e54775b0cf9dda63b69686c0a6757b568

Observation 4b236d53-dbcb-41af-9f1e-d860d5e360cf · outbound

This paper cites Measuring mathematical problem solving with the MATH dataset,.

Aligning Language Models with Selective Prediction Measuring mathematical problem solving with the MATH dataset,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:43bf67539276329ef521ef487e56c60380a287fc667323534ed9fcb6a72f26d4

Observation a1e0b9de-58e3-45c0-bbeb-afe4fa02ca6d · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Aligning Language Models with Selective Prediction Training Verifiers to Solve Math Word Problems

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:7091129ad20faa1148c42f51fd9e74901c059b61705a4e2f22e1d96e38985fa6

Observation f83d0c91-6836-4b5e-8701-b740fa29ab62 · outbound

This paper cites CommonsenseQA: A question answering challenge targeting commonsense knowledge,.

Aligning Language Models with Selective Prediction CommonsenseQA: A question answering challenge targeting commonsense knowledge,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:aad29c90650d5ff7a8092500ff81a810ea11899d233636d2132e9b72a8b10584

Observation 1fb966ea-bee1-4ed0-94a4-79c52dc31338 · outbound

This paper cites Outcome-based reinforcement learning to predict the future,.

Aligning Language Models with Selective Prediction Outcome-based reinforcement learning to predict the future,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:75bbca6888c79cdeffc7c369c142057be1e3cad9734273b07238951d1bc19267

Observation 25d657a2-15c4-4336-9ef1-edd9c610a392 · outbound

This paper cites Available: https://openreview.net/forum?id=bbhdeL8EUX.

Aligning Language Models with Selective Prediction Available: https://openreview.net/forum?id=bbhdeL8EUX

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:deea295c50dbee55a2a9e98ff320d45b6e6620b36180eeaea2e104bf1d67251d

Observation 92469c7f-8fb8-4488-9773-ed82e3696d71 · outbound

This paper cites What disease does this patient have? a large-scale open domain question answering dataset from medical exams,.

Aligning Language Models with Selective Prediction What disease does this patient have? a large-scale open domain question answering dataset from medical exams,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:ae02550d00b8071fad3bd96a3302cd79a34d9ebca74c72940efbd778d17735f9

Observation 89b878ca-576d-4218-bf78-33d3ddb3e7e3 · outbound

This paper cites Calibrating Verbalized Probabilities for Large Language Models.

Aligning Language Models with Selective Prediction Calibrating Verbalized Probabilities for Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:b14e768b41f6dac7dc21b62f44ff832ce888d2780ebf44d699c4a02381822fe7

Observation 85381dc8-d0b5-4b5e-8a68-5539990767ff · outbound

This paper cites Selectively answering ambiguous questions,.

Aligning Language Models with Selective Prediction Selectively answering ambiguous questions,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:45cc1f7eb2b357f8ea85285c25e4fc191f4f746626e4def7c3f348aa2e856324

Observation 761e4b7b-da68-497f-bfc3-d2e8a461adc6 · outbound

This paper cites Post-abstention: Towards reliably re-attempting the abstained instances in QA,.

Aligning Language Models with Selective Prediction Post-abstention: Towards reliably re-attempting the abstained instances in QA,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:ea36143d44918cb1e5a41f1292a77792c1d60d843335643c971256e946d81973

Observation 645ce5ec-2ee7-4127-a80f-cfd1cc008e62 · outbound

This paper cites Conformal language modeling,.

Aligning Language Models with Selective Prediction Conformal language modeling,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:7804f664335a1db1969f4b42fabc2eb7fdcf3c56f1aa4a35576d3d6df99e4031

Observation 9f7d682d-d0ff-467c-8d35-9f8a56dd3659 · outbound

This paper cites Controllable Text Generation for Large Language Models: A Survey.

Aligning Language Models with Selective Prediction Controllable Text Generation for Large Language Models: A Survey

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:686dd4a0aa431fab3614e7730a02f2b5987e761d57fa01cfbafb00324687e991

Observation 016e8036-a80b-4bd8-a6ca-44dd263b8219 · outbound

This paper cites Just ask for calibration: Strategies for eliciting calibrated confidence scores from language models fine-tuned with human feedback,.

Aligning Language Models with Selective Prediction Just ask for calibration: Strategies for eliciting calibrated confidence scores from language models fine-tuned with human feedback,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:71469f5b8a7ff338fc2c617c41b448882becc9f2bfb40fff4a31d5eb3d5b1d76

Observation 3609a32f-2553-4cbf-b743-5fbc861e0480 · outbound

This paper cites Semantic calibration of LLMs through the lens of temperature scaling,.

Aligning Language Models with Selective Prediction Semantic calibration of LLMs through the lens of temperature scaling,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:b8b5d15ff6455acf2e023870251397dd3dd46f077bab8d1b115977f50b9565bb

Observation 12d3293d-c9f7-4078-8211-7f0689a77774 · outbound

This paper cites Calibrating the confidence of large language models by eliciting fidelity,.

Aligning Language Models with Selective Prediction Calibrating the confidence of large language models by eliciting fidelity,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:0e3626e38e3aa472430c21a6ac11938ed8291be63f534469d9aed57519183587

Observation 14152856-2e82-4974-aa85-792d61984d61 · outbound

This paper cites Calibrating large language models with sample consistency,.

Aligning Language Models with Selective Prediction Calibrating large language models with sample consistency,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:bb1a86fc25c030215be10be11ed624783505a755d0c85832689732ad46d6d15c

Observation 0d9c4862-4cf3-457f-8eae-6b82a7b98ae4 · outbound

This paper cites Open-reasoner-zero: An open source approach to scaling up reinforcement learning on the base model,.

Aligning Language Models with Selective Prediction Open-reasoner-zero: An open source approach to scaling up reinforcement learning on the base model,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:653fc741c0434edb6b0780cedec32bb9cb1ba9527190eb79ab5c3ba3aa84044e

Observation 93c0b0fa-9621-45d6-8aef-b56b0f3682c0 · outbound

This paper cites BNPO: Beta Normalization Policy Optimization.

Aligning Language Models with Selective Prediction BNPO: Beta Normalization Policy Optimization

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:0a4e64e264244925e04c5985618b2e518029a38cfc4b0e1c4c4832c6676780a3

Observation 80602925-b826-4ed2-862f-5351eb3694d6 · outbound

This paper cites LoRA: Low-rank adaptation of large language models,.

Aligning Language Models with Selective Prediction LoRA: Low-rank adaptation of large language models,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:9b353949ffb1d39fb539bb250599d5e3f04d5c4bda27ee3902bf3a49acde4829

Observation baf6e7f4-eec1-4300-a671-a8189169d647 · outbound

This paper cites Lora without regret,.

Aligning Language Models with Selective Prediction Lora without regret,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:8cc321943d17e7ddc3926ab6868b7fc7e85bb0d7821fa728e92f4c8036a78ad9

Observation 02b733a5-ed49-4ec6-a1ed-2350b81d0ec1 · outbound

This paper cites Efficient memory management for large language model serving with pagedattention,.

Aligning Language Models with Selective Prediction Efficient memory management for large language model serving with pagedattention,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:baa97a8daf91b0b07d6a4d916b368e5dfa96bfeac6a0ccba66bc059e80c987b2

Observation 0b1ddfe3-c134-45c5-a84e-698c71b32900 · outbound

This paper cites FlashAttention-2: Faster attention with better parallelism and work partitioning,.

Aligning Language Models with Selective Prediction FlashAttention-2: Faster attention with better parallelism and work partitioning,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:269ac220981a06cc1e67e97aeecc4921e17b65ba2c88a6b4e7dbf119d5dbc330

Observation 996504ec-f614-42a8-8c6f-0a9b2105cf0f · outbound

This paper cites Capabilities of Gemini Models in Medicine.

Aligning Language Models with Selective Prediction Capabilities of Gemini Models in Medicine

Reference 53

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:01307a66387f35ae93317e8878ebe7e24c3448788eedd952f6a06405ea2cd45c

Observation bd0f7b7e-ba27-4d9b-97b7-5ca920f2b8f5 · outbound

This paper cites MedGemma Technical Report.

Aligning Language Models with Selective Prediction MedGemma Technical Report

Reference 54

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:9c644a34b8776b55e74a861e13cdd6ef3fd1c52add2053a338cf8016240cf60b

Observation 64cd9659-eed7-4f07-923e-a4d0a10b4e0c · outbound

This paper cites an unresolved cited work.

Aligning Language Models with Selective Prediction Unresolved cited work

Reference 55

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:f77fefa71fc157d66f96f93d412248c2efe6f99ec70a03a40d29904c311dd0bc

Observation 9f5b1ba6-ff35-49e4-ba49-de19d3c0846a · outbound

This paper cites an unresolved cited work.

Aligning Language Models with Selective Prediction Unresolved cited work

Reference 56

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:28b6bbe46afff15b49080741226623e989f52d1507f1ffe2ad78adc8cf176d97

Observation 3f434036-2f13-4cb1-b571-dd703852b285 · outbound

This paper cites In these cases, it is also okay to have only a small number of uncertainties and then explicitly say that I am unable to spot more uncertainties.

Aligning Language Models with Selective Prediction In these cases, it is also okay to have only a small number of uncertainties and then explicitly say that I am unable to spot more uncertainties

Reference 57

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:3226bc3206d7ae0debf965ec0003c38599383c0a5eac29fb91842e53f6b5706c

Observation 27aeeec9-88df-4db9-bed7-464c3c43faf4 · outbound

This paper cites For example, uncertainties may arise from ambiguities in the question, or from the application of a particular lemma/proof.

Aligning Language Models with Selective Prediction For example, uncertainties may arise from ambiguities in the question, or from the application of a particular lemma/proof

Reference 58

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:7832d2bfa9dbb390f1fba2db1a7bf4fd4094f4297327017d1ab1fb513e337bbd

Observation 3d1e2356-23e2-4021-8f19-8d31cafc4497 · outbound

This paper cites an unresolved cited work.

Aligning Language Models with Selective Prediction Unresolved cited work

Reference 59

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:f9c1f4a302678696576ec2382082e1bfb5f11b67f9e0b5db6abe7098a6c38820

Observation 2ec8dca4-a5eb-4425-ae3d-53486c432d4c · outbound

This paper cites an unresolved cited work.

Aligning Language Models with Selective Prediction Unresolved cited work

Reference 60

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:cba79958f68853e6dd0e6fa1faddf8125f67bce57e7edff0c3756dfcb3e56094

Observation 1e8d68b4-d610-4709-8903-f18c46f0b81f · outbound

This paper cites an unresolved cited work.

Aligning Language Models with Selective Prediction Unresolved cited work

Reference 61

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:0a4fa2c33d00e5200e2f360e4c9c0c9a98f8d71964b1c9b06e19b20335a31da8

Observation ba18d7c5-0126-4301-9104-33545bcb21b5 · outbound

This paper cites This verifier is used for HotPotQA and HotPotQA-Modified.

Aligning Language Models with Selective Prediction This verifier is used for HotPotQA and HotPotQA-Modified

Reference 62

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:7038188d3e32851fc5ee82737bc6a3e774113e2174625f5e5de4abd58e2a41a3

Observation c85f93ce-3f64-43bb-aa4a-0b105ae36b31 · outbound

This paper cites This verifier is used for Math-500, GSM8K, and Big-Math Digits.

Aligning Language Models with Selective Prediction This verifier is used for Math-500, GSM8K, and Big-Math Digits

Reference 63

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:5428b856b30413b3bddfd62c9d567b3fbb2e9947f5555b1b9808966779ef2b74

Observation 61f058d1-998d-4a3e-a07c-22926662d6af · outbound

This paper cites YES” or “NO.

Aligning Language Models with Selective Prediction YES” or “NO

Reference 64

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:44bd3aa9435d4d1d7b675dabcb1487e3c0680d541ff6df2f48360eba7cd6cf4b

Observation d3fecd6c-147a-4ff2-ba5b-df0ff99e3a77 · outbound

This paper cites This modified version systematically varies the availability of supporting evidence by removing zero, one, or both of the key para- graphs required to answer each question.

Aligning Language Models with Selective Prediction This modified version systematically varies the availability of supporting evidence by removing zero, one, or both of the key para- graphs required to answer each question

Reference 65

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:e4c0cd52c6594c8f4c41fd63ec1edaf1cc1a730e1210cef9bfe28a91af9d4c7b

Observation f8a3e2e5-6543-4a68-b92f-cef0805c7f49 · outbound

This paper cites an unresolved cited work.

Aligning Language Models with Selective Prediction Unresolved cited work

Reference 66

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:104ff5aca1c8fe1a62e1028baba9d4b2342d96faa4538c1920ef8d721fd414a3

Observation 117501cd-a0f7-4415-bac2-0a61fd6d9510 · outbound

This paper cites Thus, each question contains 8 paragraphs with both supporting paragraphs present.

Aligning Language Models with Selective Prediction Thus, each question contains 8 paragraphs with both supporting paragraphs present

Reference 67

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:456ca9c0e78c0f710a6825621086d4dd6933f10940498d7a4c14fddb5378a5e9

Observation 09289332-7e68-4868-9870-e985e9410e1e · outbound

This paper cites Correctness is measured using exact-match.

Aligning Language Models with Selective Prediction Correctness is measured using exact-match

Reference 68

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:0ee9c6fa87fd3540dbf00f3bde535d9e77c34555e2830f1913983b670f51b7cf

Observation f2d6c769-46cf-4d89-b455-a9128b9d64e9 · outbound

This paper cites Correctness is measured usingmath-verify.

Aligning Language Models with Selective Prediction Correctness is measured usingmath-verify

Reference 69

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:1e716bcdb568a690a3c2a7722bb5ea70a0891ab7b16ab7bc169e011a8e7a0be4

Observation f1b1ac74-fbd6-463d-a1e2-03da61999975 · outbound

This paper cites Correctness is measured usingmath-verify.

Aligning Language Models with Selective Prediction Correctness is measured usingmath-verify

Reference 70

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:2f476b1de64d0edbf92dec2673a7b76589e9923b4352229d2e4a41ea66eb87bf

Observation 93d37c81-2e19-4beb-a503-d318ca065359 · outbound

This paper cites Correctness is measured usingmath-verify.

Aligning Language Models with Selective Prediction Correctness is measured usingmath-verify

Reference 71

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:8414ba1b0731f33747497cb711fdbc6fa52d355bf6eb0cdbd12b6d71c823b53c

Observation 96a6dcc6-b8ca-48f7-b537-f7fe982686cf · outbound

This paper cites Correctness is measured using LLM-as-a-judge.

Aligning Language Models with Selective Prediction Correctness is measured using LLM-as-a-judge

Reference 72

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:c06a61322a914fba3ab21cbfbb6faf99371ab40b36e59656db327c43ae86882b

Observation ef4c4c29-706d-47a2-8362-abddb5efa5c6 · outbound

This paper cites Correctness is measured using LLM-as-a-judge.

Aligning Language Models with Selective Prediction Correctness is measured using LLM-as-a-judge

Reference 73

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:9c424dbb1dd13d201474e1508002aa506320b345a4aab97f38d4cdef5e5b2d8d

Observation f5336130-c10e-4099-a2ee-72c12181bc58 · outbound

This paper cites Correctness is measured using LLM-as-a-judge.

Aligning Language Models with Selective Prediction Correctness is measured using LLM-as-a-judge

Reference 74

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:da192f15a952e8ac4d965e43d08dd1521c6de2a2f8d42a1b8d6762ec40161e59

Observation 64f70914-6f6d-4187-a804-5991a8ab911d · outbound

This paper cites Correctness is measured using LLM-as-a-judge.

Aligning Language Models with Selective Prediction Correctness is measured using LLM-as-a-judge

Reference 75

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:99be46956b10807345aa3901e4cb02f5f339f174f1a2244dfd534ac5f09a0609

Observation d47453aa-3a5e-4263-90d0-5af5363531a6 · outbound

This paper cites Correctness is measured using LLM-as-a-judge.

Aligning Language Models with Selective Prediction Correctness is measured using LLM-as-a-judge

Reference 76

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:de121e562395888d3b1c5cec1e52890b6984c3388059e7986950b4f2ed291a53

Observation ce152bdc-3929-46ca-b4d1-3166a03cd78c · outbound

This paper cites We refer the reader to Sec.

Aligning Language Models with Selective Prediction We refer the reader to Sec

Reference 77

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:89d703a9d803138a30fe61fe7740765a10449d2bc457e07d03b5265604dbe81e

Observation 32794888-b62e-4905-8be0-5ae10da7a64e · outbound

This paper cites We useM= 10bins.

Aligning Language Models with Selective Prediction We useM= 10bins

Reference 78

Resolution
unresolved
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:9a7dc8f7d9afcc7f8e8dada1c0a7e1f0b017b40bd224680c3f00216919690d4b

Observation 2e75d477-13dc-4219-bd3a-adfcbd97c40f · outbound

This paper cites D. Nitrofurantoin.

Aligning Language Models with Selective Prediction D. Nitrofurantoin

Reference 79

Resolution
malformed identifier
no resolver link, observed 2026-07-12T01:51:25.883463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T01:51:25.883463Z digest=sha256:34240f9575a25fbbb833aba572853308677840dba33f93ae5c8a08c10fb826e0

Pith citing papers

No inbound Pith citation observations are available.