Pith. sign in

Paper Citation Record · LEDGER

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control

As of 8 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 0 inbound Pith citation observations for arXiv:2508.10022.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.10022 v1

Coverage vector

measured 18 of 18 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T23:23:48.288459Z

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

18 of 18 outbound references displayed

  • verified exact0
  • verified fuzzy9
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 76451bdc-a42e-4e6e-8aff-97eba88e3b6e · outbound

This paper cites A Gentle Introduction to Conformal Prediction and Distribution-Free Uncertainty Quantification.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control A Gentle Introduction to Conformal Prediction and Distribution-Free Uncertainty Quantification

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:46.709655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:46.709655Z digest=sha256:8f3aefd3df8657a83fde2798ca25939d5315acb7143e0af3369f6e89857ac683

Observation 2ad85750-b998-4bf0-a679-a343480d6dc3 · outbound

This paper cites PRISM: Self-Pruning Intrinsic Selection Method for Training-Free Multimodal Data Selection.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control PRISM: Self-Pruning Intrinsic Selection Method for Training-Free Multimodal Data Selection

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:46.781501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:46.781501Z digest=sha256:51a3db6deef7b7de74d320b667abe4c08b5d50883825e69c66f9a0374259ef83

Observation 451c4316-055f-48e9-81dd-e292512e1eb6 · outbound

This paper cites LL a VA steering: Visual instruction tuning with 500x fewer parameters through modality linear representation-steering.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control LL a VA steering: Visual instruction tuning with 500x fewer parameters through modality linear representation-steering

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:50.030135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:23:46.863849Z digest=sha256:c2cc9098ee3e4a165ef00dcc2ba160764e5fe5385ce147e0a2791d66fae12d94

Observation dd131b02-8b9c-4054-921c-36c772edf56a · outbound

This paper cites CoT-Kinetics: A Theoretical Modeling Assessing LRM Reasoning Process.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control CoT-Kinetics: A Theoretical Modeling Assessing LRM Reasoning Process

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:46.948042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:46.948042Z digest=sha256:279680b817f3713d08906c7437f9b408d20c8aa4d443de91969111cc24526dbc

Observation eafdf066-90b5-4ed9-836f-ff173eb304e4 · outbound

This paper cites Fedbip: Heterogeneous one-shot federated learning with personalized latent diffusion models.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Fedbip: Heterogeneous one-shot federated learning with personalized latent diffusion models

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:49.846480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:23:47.024290Z digest=sha256:06c9f03b2be54b8c2747b6f066cc70049387384e3e750fbd9458704701691bc0

Observation 6d7ed467-7a4c-4f14-a85c-be2a120b10e3 · outbound

This paper cites Does machine unlearning truly remove model knowledge? a framework for auditing unlearning in llms.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Does machine unlearning truly remove model knowledge? a framework for auditing unlearning in llms

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:47.100379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:47.100379Z digest=sha256:c23ab274a37739b61dfb4c1e144a90106dd6f115ae809440d2b63bbd8f6b6d24

Observation edfb10b0-820c-44f9-b4c9-7dadf750cd4a · outbound

This paper cites Conformal alignment: Knowing when to trust foundation models with guarantees.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Conformal alignment: Knowing when to trust foundation models with guarantees

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:49.621129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:23:47.158934Z digest=sha256:468aeceb06c38eb6e6e9e6420f2547dc39952267944f5be84d20c837024b4ed6

Observation a58d7397-4577-4cc5-9ca4-b5aab31dee69 · outbound

This paper cites Do LLMs Know When to NOT Answer? Investigating Abstention Abilities of Large Language Models.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Do LLMs Know When to NOT Answer? Investigating Abstention Abilities of Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:47.223084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:47.223084Z digest=sha256:5edaa6e3798960a4ffde2d131199341fcf707a30144a73b3693e578aa5df7f0e

Observation c6233eed-af46-4a6d-892b-114be85be1f9 · outbound

This paper cites Backdoor Cleaning without External Guidance in MLLM Fine-tuning.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Backdoor Cleaning without External Guidance in MLLM Fine-tuning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:47.312420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:47.312420Z digest=sha256:d9ecc25060e6478bee7efcd416d4782c4da2d55c291fedb4cc4887cdaab8852a

Observation 3cb58dc8-88a9-4442-b7da-168729b9f851 · outbound

This paper cites Sample then identify: A general framework for risk control and assessment in multimodal large language models.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Sample then identify: A general framework for risk control and assessment in multimodal large language models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:49.430114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:23:47.445094Z digest=sha256:6dda4f22fd00b853b24b7e3367930578d56d434af80cd9d43d75f8e72f7a3953

Observation 043c2906-a24a-4771-827d-a6552001558b · outbound

This paper cites Conformalized multiple testing after data-dependent selection.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Conformalized multiple testing after data-dependent selection

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:49.235604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:23:47.580189Z digest=sha256:297cc7f9acc3ee5d32e2a332262cafe584d5d0d2209b20b3fbe010e83053f42c

Observation 9a6e9177-439e-4cb4-a35b-84edbef7d3dc · outbound

This paper cites ASCD: Attention-Steerable Contrastive Decoding for Reducing Hallucination in MLLM.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control ASCD: Attention-Steerable Contrastive Decoding for Reducing Hallucination in MLLM

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:47.662074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:47.662074Z digest=sha256:1ef657d12e399b39ce19daefabd9e9169e559a972bebcf7832e4f03d4e94fe0f

Observation 5b6655dc-2b0e-4590-9275-709984260001 · outbound

This paper cites Conu: Conformal uncertainty in large language models with correctness coverage guarantees.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Conu: Conformal uncertainty in large language models with correctness coverage guarantees

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:49.052841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:23:47.772855Z digest=sha256:85f349afdae6341519f2149e69b6c42bfca897f5ffbcc4f9d65655931e0c9efe

Observation 48bea371-5e4a-40c3-a8e0-d6ccdac8bdbb · outbound

This paper cites COIN: Uncertainty-Guarding Selective Question Answering for Foundation Models with Provable Risk Guarantees.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control COIN: Uncertainty-Guarding Selective Question Answering for Foundation Models with Provable Risk Guarantees

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:47.832064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:47.832064Z digest=sha256:02d8a7cf0cf260637d27cecc1dc4783a5ce27c0ee6baf1e598335088aa093019

Observation bd6def49-0437-46e9-bdda-7b713f81a5af · outbound

This paper cites Word-sequence entropy: Towards uncertainty estimation in free-form medical question answering applications and beyond.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Word-sequence entropy: Towards uncertainty estimation in free-form medical question answering applications and beyond

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:48.927016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:23:47.962579Z digest=sha256:c9beefcba2f4406598a606ae71be0f8e0cc6e19d9a0e695f979d4e9dd38ad25f

Observation 17bb5d39-e617-4473-acf5-15e5e2599d61 · outbound

This paper cites SC on U : Selective conformal uncertainty in large language models.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control SC on U : Selective conformal uncertainty in large language models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:48.717705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:23:48.072512Z digest=sha256:f2381f52c4814a4615ac3a7af1c5127c77b921c3c9f99f63a170f77f64f73f32

Observation 67504119-5f72-4e32-95b2-fede8e3a3149 · outbound

This paper cites Benchmarking llms via uncertainty quantification.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control Benchmarking llms via uncertainty quantification

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T23:23:48.603235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-05T23:23:48.170468Z digest=sha256:89a5aaf75f65471d304b8f580b93a5833fc65b27effc89ad2d5e380d8eb6574c

Observation 019de6e1-98e3-42a4-a2f2-f2ba73c4ca8f · outbound

This paper cites SPOT! Revisiting Video-Language Models for Event Understanding.

Conformal P-Value in Multiple-Choice Question Answering Tasks with Provable Risk Control SPOT! Revisiting Video-Language Models for Event Understanding

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:48.288459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:23:48.288459Z digest=sha256:f4e021d440e67143a9aff013afd17defa1828f48a1e88f3056a4e1982054c5ad

Pith citing papers

No inbound Pith citation observations are available.