Pith. sign in

Paper Citation Record · LEDGER

Thermometer: Towards Universal Calibration for Large Language Models

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2403.08819.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.08819 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:25:22.485623Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 63a1bb91-648e-4bfd-8ae3-7a60e6c33297 · inbound

Towards Harmonized Uncertainty Estimation for Large Language Models cites this paper.

Towards Harmonized Uncertainty Estimation for Large Language Models Thermometer: Towards Universal Calibration for Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:22.485623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:25:22.485623Z digest=sha256:08a6a3a67f66c9674b16a50063389c94cff00e5b5eb6b62241c0226fce74a06d

Observation 12576635-f91a-4a4f-9475-6f495a25b7b2 · inbound

Towards Objective Fine-tuning: How LLMs' Prior Knowledge Causes Potential Poor Calibration? cites this paper.

Towards Objective Fine-tuning: How LLMs' Prior Knowledge Causes Potential Poor Calibration? Thermometer: Towards Universal Calibration for Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T13:51:12.698057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:51:12.698057Z digest=sha256:9299424e91ae9a057e07afc42005663b682e285c5bf9bf132d3681ac4179369d

Observation 13123c1d-e4be-4616-b5c4-c48acce1a9aa · inbound

Overconfidence in LLM-as-a-Judge: Diagnosis and Confidence-Driven Solution cites this paper.

Overconfidence in LLM-as-a-Judge: Diagnosis and Confidence-Driven Solution Thermometer: Towards Universal Calibration for Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T22:54:25.079218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:54:25.079218Z digest=sha256:ba047e0c88dba52319fec33d82718e067a916b9b97434e90901a98d1bf470a7e

Observation 5b9dc4ac-7e07-47d2-aa6b-df7f43671200 · inbound

Insights into User Interface Innovations from a Design Thinking Workshop at deRSE25 cites this paper.

Insights into User Interface Innovations from a Design Thinking Workshop at deRSE25 Thermometer: Towards Universal Calibration for Large Language Models

Reference 1987

Resolution
unresolved
no resolver link, observed 2026-08-05T16:15:38.387545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:15:38.387545Z digest=sha256:d1f76cb15ea929dd82f01238104261bb0f67cc38298e695a866e92fef264ee97

Observation 5f4aa9d9-897b-4962-9273-2271b2e68222 · inbound

CDE: Curiosity-Driven Exploration for Efficient Reinforcement Learning in Large Language Models cites this paper.

CDE: Curiosity-Driven Exploration for Efficient Reinforcement Learning in Large Language Models Thermometer: Towards Universal Calibration for Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T18:53:02.742636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:53:02.742636Z digest=sha256:a3b8f9b5258322a452dcda60e35834ff0150663f7affe5a737ea88c226e72102

Observation f4904710-1229-4c68-8d01-5b32f88ff9f0 · inbound

The Well-Tempered Classifier: Some Elementary Properties of Temperature Scaling cites this paper.

The Well-Tempered Classifier: Some Elementary Properties of Temperature Scaling Thermometer: Towards Universal Calibration for Large Language Models

Reference 1978

Resolution
unresolved
no resolver link, observed 2026-08-02T23:14:39.507295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:14:39.507295Z digest=sha256:a041b4a201e37b70900a8c7b4fe007e16b9a01600856395320991b28d07912bc

Observation 45844de7-df57-4c1c-9478-0cdc85e30134 · inbound

Beyond Medical Diagnostics: How Medical Multimodal Large Language Models Think in Space cites this paper.

Beyond Medical Diagnostics: How Medical Multimodal Large Language Models Think in Space Thermometer: Towards Universal Calibration for Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T18:16:33.382429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:16:33.382429Z digest=sha256:1bb0f9031c49cf6ef817b13307cfa5a6ae66df01336205cf59f8acb6404da1d4

Observation 4200b305-bba2-4c08-a1c7-30b03acd1c6b · inbound

Overconfidence and Calibration in Medical VQA: Empirical Findings and Hallucination-Aware Mitigation cites this paper.

Overconfidence and Calibration in Medical VQA: Empirical Findings and Hallucination-Aware Mitigation Thermometer: Towards Universal Calibration for Large Language Models

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T21:28:17.720506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T21:24:55.411001Z digest=sha256:a1a9324bc09d65585ae86584b1b668f4275fd4a6e85cae1054ae601c82120c69

Observation 080bbfe9-ff9e-4bed-92d7-1147d9542dfb · inbound

Learning When Not to Decide: A Framework for Overcoming Factual Presumptuousness in AI Adjudication cites this paper.

Learning When Not to Decide: A Framework for Overcoming Factual Presumptuousness in AI Adjudication Thermometer: Towards Universal Calibration for Large Language Models

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-10T02:11:57.293912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T02:08:24.770003Z digest=sha256:56667a216ec87d110228ed61a7aaf15deba1847a608f33dc943b61134c06cc32

Observation 8bcaf9d2-095f-421b-ae27-92c868321fe7 · inbound

Inducing Artificial Uncertainty in Language Models cites this paper.

Inducing Artificial Uncertainty in Language Models Thermometer: Towards Universal Calibration for Large Language Models

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:12:55.364521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-14T20:11:44.878211Z digest=sha256:74baa5a60ce079faec118fefcf63f935dd605ee95246d58974bfee8acbfa152f

Observation b56c1917-6c4e-41e6-bc49-cd7948e331a5 · inbound

Can LLMs Use Linguistic Uncertainty Markers to Reliably Reflect Intrinsic Confidence? cites this paper.

Can LLMs Use Linguistic Uncertainty Markers to Reliably Reflect Intrinsic Confidence? Thermometer: Towards Universal Calibration for Large Language Models

Reference 65

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T12:23:24.418967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T12:18:36.854164Z digest=sha256:334ab576c759b5f257b278055e33d700c81549cc2624de9b6a2c3d7d5291d516

Observation c43a8b4a-a95e-4633-b093-607f86652094 · inbound

Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs cites this paper.

Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs Thermometer: Towards Universal Calibration for Large Language Models

Reference 83

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:35:42.097153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-01T05:22:38.232552Z digest=sha256:15909b167c822dcdedecbeb8c383b417164f5f49e63301a44879de9550f77e4d