Pith. sign in

Paper Citation Record · LEDGER

Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 53 inbound Pith citation observations for arXiv:2307.10236.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.10236 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 53 of 53 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T22:31:28.579519Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T18:08:46.635429Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 18d5c70a-b7a0-47e1-96c4-b420f8693394 · inbound

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models cites this paper.

Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T19:34:41.745863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:34:41.745863Z digest=sha256:30ae620c9a6f40fb6e9338be3227a61b950944a1d69446b3ca62ed2fa3bde8ba

Observation b6e27cdc-0b8a-4f96-ab5a-cddc54e38d76 · inbound

Understanding the Fundamental Design Decisions of Retrieval-Augmented Generation Systems cites this paper.

Understanding the Fundamental Design Decisions of Retrieval-Augmented Generation Systems Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T10:14:22.102051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:14:22.102051Z digest=sha256:9f22959cc4c522228f1171918e555e37d4fcb720f2b0fec532f8132723e01280

Observation 864c8ac3-f57e-4fd3-badf-2acae5754f66 · inbound

VideoICL: Confidence-based Iterative In-context Learning for Out-of-Distribution Video Understanding cites this paper.

VideoICL: Confidence-based Iterative In-context Learning for Out-of-Distribution Video Understanding Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T23:51:14.531598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:51:14.531598Z digest=sha256:da51554dfd0b55bd19213e5bf20debea27b53ddb6f160c9675e15e2eeea664a5

Observation 4b5513f7-13b6-4b3e-b5a9-4b88ba799278 · inbound

CALMM-Drive: Confidence-Aware Autonomous Driving with Large Multimodal Model cites this paper.

CALMM-Drive: Confidence-Aware Autonomous Driving with Large Multimodal Model Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T21:43:21.354377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:43:21.354377Z digest=sha256:61051212e3c8e8fd2ce6e7f1328a4424d2231f0d5375f10ce1afbd21cc5f6d2f

Observation 391b9e34-43a1-42a3-934b-af106e1479fd · inbound

A Survey on Uncertainty Quantification of Large Language Models: Taxonomy, Open Research Challenges, and Future Directions cites this paper.

A Survey on Uncertainty Quantification of Large Language Models: Taxonomy, Open Research Challenges, and Future Directions Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-11T20:37:54.786627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:37:54.786627Z digest=sha256:1db09c8c43f86d118c38ec5ae231742d5febfa2f5a265ef4287df0d482e1680b

Observation 5f558599-8972-4949-a01d-1ef5cfe7249d · inbound

Label-Confidence-Aware Uncertainty Estimation in Natural Language Generation cites this paper.

Label-Confidence-Aware Uncertainty Estimation in Natural Language Generation Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T19:03:01.135848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T19:03:01.135848Z digest=sha256:bd0cb5ab9a1d9540a68d65327daf4c24bdc0f205ef8e415c9b08b9527578cf5b

Observation c9d64545-2256-4248-9cf0-286087b58a29 · inbound

Evolution of Thought: Diverse and High-Quality Reasoning via Multi-Objective Optimization cites this paper.

Evolution of Thought: Diverse and High-Quality Reasoning via Multi-Objective Optimization Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T13:52:26.368862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:52:26.368862Z digest=sha256:84fed57a8b64fb09b317edcbbc8753f67ffe9951ab85589433727f0588c5c34f

Observation 3f6f6ccc-3a78-4c15-a236-945adbca0013 · inbound

Uncertainty-Aware Hybrid Inference with On-Device Small and Remote Large Language Models cites this paper.

Uncertainty-Aware Hybrid Inference with On-Device Small and Remote Large Language Models Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T13:55:45.211650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:55:45.211650Z digest=sha256:dec04f4bd938ede28f6e25a70b95be386c225ed60b1be5b54b5dc12d312cb193

Observation 732f59f5-ed1f-493c-be68-4ba243cc6f26 · inbound

A Survey of Calibration Process for Black-Box LLMs cites this paper.

A Survey of Calibration Process for Black-Box LLMs Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T13:48:40.023939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:48:40.023939Z digest=sha256:47595f98b6da00cbe7fce42fd4b6416c62aa01a948a4ada5666a04147bc4f6f0

Observation e1bfd33f-a38e-4611-8557-d4475653a059 · inbound

Context-DPO: Aligning Language Models for Context-Faithfulness cites this paper.

Context-DPO: Aligning Language Models for Context-Faithfulness Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T13:08:33.665515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:08:33.665515Z digest=sha256:9e0968f7985f16ca519baafc972b6f547743ed74ae29f2d00cd5ab6733d0b81e

Observation 6a926dca-db7b-4f53-b51f-3ff96a785a23 · inbound

Dynamic Attention-Guided Context Decoding for Mitigating Context Faithfulness Hallucinations in Large Language Models cites this paper.

Dynamic Attention-Guided Context Decoding for Mitigating Context Faithfulness Hallucinations in Large Language Models Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T22:42:11.164271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:42:11.164271Z digest=sha256:327ff18556f1c278d9e3dbdcd78e9b851d585f8f197d0d61111769ba657a7f3c

Observation 94ef8d93-cf8c-4978-8c1e-16a2e55e7323 · inbound

Decoding Knowledge in Large Language Models: A Framework for Categorization and Comprehension cites this paper.

Decoding Knowledge in Large Language Models: A Framework for Categorization and Comprehension Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T22:33:24.025218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:33:24.025218Z digest=sha256:40af10f06b86ed1e169f2cf8069112bb19b62867b5356c2152dc9980c2c0359f

Observation 8e12cd05-d4a3-4046-9275-68da9ab511e2 · inbound

Enhancing Uncertainty Modeling with Semantic Graph for Hallucination Detection cites this paper.

Enhancing Uncertainty Modeling with Semantic Graph for Hallucination Detection Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T22:34:21.047700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:34:21.047700Z digest=sha256:d0b707ab0bfbc6dd452bd2a4dbc94aa6bd91c4880a23a9f3aab4d4b3556894dd

Observation b70ef546-6e47-41ef-a251-9c599f7d6340 · inbound

Correctness Assessment of Code Generated by Large Language Models Using Internal Representations cites this paper.

Correctness Assessment of Code Generated by Large Language Models Using Internal Representations Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-10T16:41:31.237054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:41:31.237054Z digest=sha256:3c3eb129ec81898d22f148340e8f3e76584f9d6908f247f26cbc46727c77c417

Observation d61e38a5-dee2-4724-b913-3c5a57d5f793 · inbound

Bridging Internal Probability and Self-Consistency for Effective and Efficient LLM Reasoning cites this paper.

Bridging Internal Probability and Self-Consistency for Effective and Efficient LLM Reasoning Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T18:53:44.196975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:53:44.196975Z digest=sha256:eb13005ecd3f66e149ee3a85f7ae34224824eedbca475a0092fc9f11595534ee

Observation 8c9985b1-0ac7-42d2-844b-922eff5d1786 · inbound

Confident or Seek Stronger: Exploring Uncertainty-Based On-device LLM Routing From Benchmarking to Generalization cites this paper.

Confident or Seek Stronger: Exploring Uncertainty-Based On-device LLM Routing From Benchmarking to Generalization Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T22:49:51.741605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:49:51.741605Z digest=sha256:21237c46237e7b673d8837e09f67ab24a64ca6efb4f5c551a1fa98e2e32ffb18

Observation 04ce6625-c5ca-406c-8ae8-a880fa4b2be6 · inbound

Consensus Entropy: Harnessing Multi-VLM Agreement for Self-Verifying and Self-Improving OCR cites this paper.

Consensus Entropy: Harnessing Multi-VLM Agreement for Self-Verifying and Self-Improving OCR Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-22T20:32:04.647581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-22T20:31:34.074705Z digest=sha256:9cfb5abfc067e47a618abc82a3eede6521b9cde40cde3b747c386587abbd2bed

Observation a6367c03-3368-4dd5-ac2f-b426bf76d09e · inbound

Bridging AI and Carbon Capture: A Dataset for LLMs in Ionic Liquids and CBE Research cites this paper.

Bridging AI and Carbon Capture: A Dataset for LLMs in Ionic Liquids and CBE Research Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T22:31:28.579519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T22:31:28.579519Z digest=sha256:e8e42efb9af4996f7fc5a091a97d0db85d9d59580cca2a522c1013f1ebfdcaa4

Observation 64c5cf64-ab12-424c-af3b-0b74fc54992f · inbound

Communication-Efficient Hybrid Language Model via Uncertainty-Aware Opportunistic and Compressed Transmission cites this paper.

Communication-Efficient Hybrid Language Model via Uncertainty-Aware Opportunistic and Compressed Transmission Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T20:55:06.676835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:55:06.676835Z digest=sha256:fe0b1ee91844b7aeba57b626531973d20f576cd0bb130b50d736904c1b60f8d6

Observation 912e6fa5-4b63-4d10-a79b-15e3f1ec704c · inbound

Divide-Then-Align: Honest Alignment based on the Knowledge Boundary of RAG cites this paper.

Divide-Then-Align: Honest Alignment based on the Knowledge Boundary of RAG Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:50:39.165790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:50:39.165790Z digest=sha256:bad0876d8a0f0e550671f842253028f56844592ffb0ab8228253426d5028c066

Observation 17d49396-5d55-4116-ab29-06355a8eb962 · inbound

Is Your Model Fairly Certain? Uncertainty-Aware Fairness Evaluation for LLMs cites this paper.

Is Your Model Fairly Certain? Uncertainty-Aware Fairness Evaluation for LLMs Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:05.183520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:05.183520Z digest=sha256:a8a867c1d4003cfe980ebb0ecfc0462c1bf5a4e912d635c1d481089cda87476b

Observation d542e00c-9b25-4a32-b4d7-fc025e8ed3db · inbound

HD-NDEs: Neural Differential Equations for Hallucination Detection in LLMs cites this paper.

HD-NDEs: Neural Differential Equations for Hallucination Detection in LLMs Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:34:04.092510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:34:04.092510Z digest=sha256:e369f12e41e43ead9718708669f1b2488f5c9a53b8c8cbbe20327274bb657232

Observation 28f84359-f68c-4e41-9021-fb91d2fd62d0 · inbound

Improving the Calibration of Confidence Scores in Text Generation Using the Output Distribution's Characteristics cites this paper.

Improving the Calibration of Confidence Scores in Text Generation Using the Output Distribution's Characteristics Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:07:28.528388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:07:28.528388Z digest=sha256:7037e7ac9ee00562d3029c5c125bd448c1e6112afe3ce4813ba7bf5b99acfbf3

Observation c4ad12f7-8bf7-4960-95b9-1f84ea729238 · inbound

ChemAU: Harness the Reasoning of LLMs in Chemical Research with Adaptive Uncertainty Estimation cites this paper.

ChemAU: Harness the Reasoning of LLMs in Chemical Research with Adaptive Uncertainty Estimation Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:44.964444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:44.964444Z digest=sha256:563993eae99a34230bdbbe298694cc10ce39087d0aa5c84c19f4e27bd14e8ea6

Observation 94a0b584-6497-4576-a02a-ad089982c8f3 · inbound

Representations of Fact, Fiction and Forecast in Large Language Models: Epistemics and Attitudes cites this paper.

Representations of Fact, Fiction and Forecast in Large Language Models: Epistemics and Attitudes Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T11:44:14.313733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:44:14.313733Z digest=sha256:2f9a1bd5a916e9613d8d5e1184adaf225af1d9e10864609f47da9e5d513e3a7e

Observation 4e92d247-b511-49e5-8697-b48d91bde039 · inbound

Shaking to Reveal: Perturbation-Based Detection of LLM Hallucinations cites this paper.

Shaking to Reveal: Perturbation-Based Detection of LLM Hallucinations Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T11:23:45.835103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:23:45.835103Z digest=sha256:24eec1017ef1327a41ef6a03874537c700e24532e7ccdf4d6f4e0f21f194cfa9

Observation 8509ddc4-7e3f-46b0-81f3-93994956695f · inbound

TRUST: Test-time Resource Utilization for Superior Trustworthiness cites this paper.

TRUST: Test-time Resource Utilization for Superior Trustworthiness Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:08.144917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:07:08.144917Z digest=sha256:ffb3ef20ae6e351d731c844c39741f0568abce271acb1ead3865de019862e88d

Observation 9c9091df-21bc-44f1-a16e-b58d8e39e683 · inbound

Controlling Context: Generative AI at Work in Integrated Circuit Design and Other High-Precision Domains cites this paper.

Controlling Context: Generative AI at Work in Integrated Circuit Design and Other High-Precision Domains Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-15T19:55:12.336238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:55:12.336238Z digest=sha256:536cab70ebd960cfacd350651b60d14130231714158f300ea90ce343afcb8d6f

Observation 2de1bfea-dff1-45d1-be5d-ec1373a574f2 · inbound

WebSailor: Navigating Super-human Reasoning for Web Agent cites this paper.

WebSailor: Navigating Super-human Reasoning for Web Agent Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:37:09.646082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-17T15:37:09.572241Z digest=sha256:ef0a90715738958c06fd9f97898da9c863b2dd2a7bb5a8b12e7d3e44500aaa1f

Observation e8f4b193-3d06-4809-833f-c1f35ce37c68 · inbound

Toward Better Generalisation in Uncertainty Estimators: Leveraging Data-Agnostic Features cites this paper.

Toward Better Generalisation in Uncertainty Estimators: Leveraging Data-Agnostic Features Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T20:02:02.991407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:02:02.991407Z digest=sha256:5d8a49869901b60e323eee77ede7b44c9fa41dc97b6e3fa857f57bb02acd26f8

Observation 4627834b-dcba-4186-bdec-8c6f045205fe · inbound

Evaluating Uncertainty and Quality of Visual Language Action-enabled Robots cites this paper.

Evaluating Uncertainty and Quality of Visual Language Action-enabled Robots Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:25.438745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:02:25.438745Z digest=sha256:94fad3d9983e9d570607d8feffb03c179899ee8dfb92bd657dacb0f3172f3221

Observation c10d5dda-f7ac-43e2-a2b4-7efd07fc44ed · inbound

Confidence Estimation for Text-to-SQL in Large Language Models cites this paper.

Confidence Estimation for Text-to-SQL in Large Language Models Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T22:36:21.614672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:36:21.614672Z digest=sha256:ec3aac7acb33ab5a00dc5f23e94292e8ce5201478ee821cf49b2175f4ee6baa2

Observation c4be402b-50d1-4e14-a95d-45d5e84071d3 · inbound

Can Multiple Responses from an LLM Reveal the Sources of Its Uncertainty? cites this paper.

Can Multiple Responses from an LLM Reveal the Sources of Its Uncertainty? Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-05T14:34:33.108971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:34:33.108971Z digest=sha256:c6212f7fbc6df88f14e21f84158d9392ac2ddb19376ffed920e4498280f73e83

Observation 7d51c7d5-ae77-44e2-81ab-67c99bb17bdb · inbound

Neural Message-Passing on Attention Graphs for Hallucination Detection cites this paper.

Neural Message-Passing on Attention Graphs for Hallucination Detection Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T13:52:09.125465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:52:09.125465Z digest=sha256:c7bcbbf22c04c7e5b703d265de723daf5e6964b58d0ca42c097abda3374f0ba7

Observation d5e5bab8-0f45-4cbc-b79b-b6f0d9d0cdaa · inbound

Robometer: Scaling General-Purpose Robotic Reward Models via Trajectory Comparisons cites this paper.

Robometer: Scaling General-Purpose Robotic Reward Models via Trajectory Comparisons Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 135

Resolution
verified exact
arxiv_id, observed 2026-05-15T18:00:12.654671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T17:59:52.365630Z digest=sha256:4037c49f376db35618e712f642112072ac29f65950e96f52e782b26ec9b6bfb6

Observation 2f02c018-2875-4c24-a5e5-4ac2f5420c11 · inbound

Context Collapse: Barriers to Adoption for Generative AI in Workplace Settings cites this paper.

Context Collapse: Barriers to Adoption for Generative AI in Workplace Settings Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:35:52.626262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T18:58:18.436603Z digest=sha256:788a974291971785a0f0aab56094f7b68504dfee92a80950d1ee5622caf22fec

Observation 059fdd2b-d15a-48aa-8081-7743fdbbdb19 · inbound

Confident in a Confidence Score: Investigating the Sensitivity of Confidence Scores to Supervised Fine-Tuning cites this paper.

Confident in a Confidence Score: Investigating the Sensitivity of Confidence Scores to Supervised Fine-Tuning Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:25:58.665662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-10T17:38:53.434978Z digest=sha256:d3bd237289bd88888b211d034172a858c1a1dcf130178fae4763bcb261ebdf0a

Observation af56679e-e846-44ab-b471-5b935ea7a7e7 · inbound

Learning Uncertainty from Sequential Internal Dispersion in Large Language Models cites this paper.

Learning Uncertainty from Sequential Internal Dispersion in Large Language Models Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-10T08:48:02.443497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-10T08:36:39.242766Z digest=sha256:65448fe1b2a94dc0277247da60f54a122473e306b7b75e7da802ec5c297fcb22

Observation 5af0c885-79f7-496c-8cda-771cc62f48a1 · inbound

Temporal Difference Calibration in Sequential Tasks: Application to Vision-Language-Action Models cites this paper.

Temporal Difference Calibration in Sequential Tasks: Application to Vision-Language-Action Models Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 54

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T13:51:03.364892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-09T23:59:32.295839Z digest=sha256:b45313271ae063c2180c3241ef4f8466745426fe8a9943d1ece0ee17d1fb8422

Observation 55461839-1c00-4185-a985-8a35556bab5b · inbound

LLMs Uncertainty Quantification via Adaptive Conformal Semantic Entropy cites this paper.

LLMs Uncertainty Quantification via Adaptive Conformal Semantic Entropy Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:11:09.088849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-08T17:51:53.511503Z digest=sha256:5acf1c2f1eb97819de6bc947ee703da38d1450a36b22a2dbcc221fcaa32c7096

Observation 421a65a4-8441-4e51-9d2d-4ca75d8643c0 · inbound

LLMs Uncertainty Quantification via Adaptive Conformal Semantic Entropy cites this paper.

LLMs Uncertainty Quantification via Adaptive Conformal Semantic Entropy Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-01T13:15:46.007360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T23:47:33.632049Z digest=sha256:06d787f98bd14ddb4ead51051d3e4de94dbff013903f18b43d5ab0e90d02fb13

Observation 1c2e9598-9e4a-4827-9ff7-6154ca08132c · inbound

BackFlush: Knowledge-Free Backdoor Detection and Elimination with Watermark Preservation in Large Language Models cites this paper.

BackFlush: Knowledge-Free Backdoor Detection and Elimination with Watermark Preservation in Large Language Models Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:02:58.950282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-14T21:01:10.756844Z digest=sha256:88e8be4f14db3803913baed850f010f5dd0956835c2435870cdd597b0af4349f

Observation b33281b2-b7f2-4022-a514-287aa102cfde · inbound

Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming cites this paper.

Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T09:14:45.421169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-22T09:13:45.830192Z digest=sha256:adce369a91514da224679073d5eb76f7123088014747587d804c4d764dfe1cf6

Observation 88a04825-f6ac-45a1-8a0f-92b8eb30babd · inbound

Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming cites this paper.

Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T16:54:58.773116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T16:48:55.898188Z digest=sha256:89d20aa0084c509ddc815890652bacccef7d9e1ddd618b34f13723dcf871a765

Observation 9175b8aa-0484-45f1-8a12-dd820e23d6d1 · inbound

MultiHaluDet: Multilingual Hallucination Detection via LLM Hidden State Probing cites this paper.

MultiHaluDet: Multilingual Hallucination Detection via LLM Hidden State Probing Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T12:34:38.667568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T12:29:59.165791Z digest=sha256:902288145baa37df90e75b9990c76bc278a88fb64bb2fdfbf0aa91578d4318d9

Observation 81b8b31d-b857-49d0-b3eb-435b291e0f35 · inbound

Entropy Distribution as a Fingerprint for Hallucinations in Generative Models cites this paper.

Entropy Distribution as a Fingerprint for Hallucinations in Generative Models Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:13:26.691322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T12:10:43.626846Z digest=sha256:60c7fb758b2ae30ce790b550e2b8ef6719cebb9f9f49c90b6b826a8aca0f034a

Observation 83e58d2a-d1e9-492f-84ad-6cc08aeaebf8 · inbound

Smart Picks in the Dark: Towards Efficient RLVR for Reasoning via Tracing Metacognitive Pivots cites this paper.

Smart Picks in the Dark: Towards Efficient RLVR for Reasoning via Tracing Metacognitive Pivots Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:26:46.219869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T06:55:09.927034Z digest=sha256:20d51d5fb56ed3640cb890afeb35a18c2a835cc689edc879a9862a865717896d

Observation caa66479-c02b-41e8-92a3-5e8caa048f64 · inbound

ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning cites this paper.

ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T21:07:23.911497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-27T19:56:09.820812Z digest=sha256:080b901386c7bc5f2d1f427e6f91f8447f000f0de89a4a6c477c289b717bd201

Observation 7e6d11a9-b4e3-459f-876b-d040a43650e3 · inbound

Quantifying Consistency in LLM Logical Reasoning via Structural Uncertainty cites this paper.

Quantifying Consistency in LLM Logical Reasoning via Structural Uncertainty Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-03T18:08:46.637567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-27T03:18:22.813002Z digest=sha256:0314a4516e968dda6f4e50545315f17b91995ed6542d9c2b15db9591c55fb907

Observation 68dbbede-f71e-4aa9-89a7-9fe7717fcaf9 · inbound

The "I Don't Know" Filter: Enhancing Agentic Reliability in Function Calling cites this paper.

The "I Don't Know" Filter: Enhancing Agentic Reliability in Function Calling Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-11T22:10:22.729188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T22:10:22.729188Z digest=sha256:94a72b60b0f359d9258140ef15410df79623fce6ece88ab26c68ff32bb779082

Observation e2804edd-e922-46dc-b696-e9411f01a820 · inbound

Hallucination Detection in Large Language Models Using Diversion Decoding cites this paper.

Hallucination Detection in Large Language Models Using Diversion Decoding Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-14T11:26:23.545785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T11:26:23.545785Z digest=sha256:66a7b12f1ce5c5140a8c7b144ba33a65b810606afbc0d03891d80c68037119cd

Observation 31f5a8fb-221e-4974-af7b-cc7fbc308b06 · inbound

SAFECAST: Robust Failure Detection for VLA Policies with Contrast-Set Training and Calibration cites this paper.

SAFECAST: Robust Failure Detection for VLA Policies with Contrast-Set Training and Calibration Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T00:10:44.884480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:10:44.884480Z digest=sha256:75a09a802b80111a42e3dd19f51f009e56f8df436044fc30562181bba926ae5e

Observation 3171bdce-0e82-4e16-a457-67913c8494d5 · inbound

From token probabilities to calibrated confidence: An empirical study of mathematical question answering cites this paper.

From token probabilities to calibrated confidence: An empirical study of mathematical question answering Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T00:53:27.941566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T00:53:27.941566Z digest=sha256:f47fd7a7de63e8b0ad021abef41402011aa33a9643a7def079ddc955d16c7172