Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 71 inbound Pith citation observations for arXiv:2305.14975.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:08:24.966348Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-08T10:04:51.524003Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 0d26e54b-5eda-40e3-9677-e2b3743f9b8e · inbound
Functional-level Uncertainty Quantification for Calibrated Fine-tuning on LLMs Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation fc857310-d720-44b5-bb35-893cde8ddb83 · inbound
Is my Meeting Summary Good? Estimating Quality with a Multi-LLM Evaluator Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43137bf2-977c-4039-a8b6-0eb7b2a8ce5e · inbound
A Survey on Uncertainty Quantification of Large Language Models: Taxonomy, Open Research Challenges, and Future Directions Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 203
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5179813-6feb-4f68-9223-48256d41fb10 · inbound
UAlign: Leveraging Uncertainty Estimations for Factuality Alignment on Large Language Models Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf62d123-8a6a-45f8-bc30-bad822d73584 · inbound
Can You Trust LLM Judgments? Reliability of LLM-as-a-Judge Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07f3ca1a-7426-4132-8465-b82fc73e1815 · inbound
Unveiling Uncertainty: A Deep Dive into Calibration and Performance of Multimodal Large Language Models Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87c41a14-f868-400d-80a0-aef4748ac421 · inbound
ReFoRCE: A Text-to-SQL Agent with Self-Refinement, Consensus Enforcement, and Column Exploration Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdee8e7b-d2f3-492e-83e9-f59b50f90375 · inbound
What is a Number, That a Large Language Model May Know It? Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac08aad5-0d5c-4bb6-bb09-9316d8ea88f8 · inbound
AI Alignment at Your Discretion Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 103
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb1c8036-e2fb-453e-9dea-a2e4e2b5f1d0 · inbound
Exploring the Potential for Large Language Models to Demonstrate Rational Probabilistic Beliefs Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f7435d4-a7d6-43bf-8182-a32399b9cf1b · inbound
Guiding VLM Agents with Process Rewards at Inference Time for GUI Navigation Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94dc402d-477a-49c4-ac6b-5235c78a8805 · inbound
Lightweight Latent Verifiers for Efficient Meta-Generation Strategies Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e07cd5e4-985b-4cca-9516-7e2afe1c32c8 · inbound
From Evidence to Belief: A Bayesian Epistemology Approach to Language Models Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 584d44be-ba08-413d-a395-ebfcbd78e2c6 · inbound
How Knowledge Popularity Influences and Enhances LLM Knowledge Boundary Perception Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec1db97f-76d1-473b-bdd5-ef12a1fd1741 · inbound
Grammars of Formal Uncertainty: When to Trust LLMs in Automated Reasoning Tasks Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8abcfe2-2940-4745-b144-bcf9713b4ffc · inbound
Maximizing Confidence Alone Improves Reasoning Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1da2abeb-136f-4094-a287-a324fe04db61 · inbound
Revisiting Uncertainty Estimation and Calibration of Large Language Models Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f12a8b6b-719e-4ae1-96cc-9777a10f2fe7 · inbound
Shaking to Reveal: Perturbation-Based Detection of LLM Hallucinations Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30a7e555-abf9-49aa-ba4c-bc584eb09656 · inbound
SQLens: An End-to-End Framework for Error Detection and Correction in Text-to-SQL Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 619a2cbc-486e-45bc-9b89-0ae08a5ad457 · inbound
AbstentionBench: Reasoning LLMs Fail on Unanswerable Questions Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54c42534-296b-4e84-9d02-2f824336c9ce · inbound
Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know? Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddf5da14-286c-49de-a5d8-df45f05613a0 · inbound
Loki's Dance of Illusions: A Comprehensive Survey of Hallucination in Large Language Models Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8294562d-caf4-45f4-873a-df1bceb7a1dd · inbound
How Overconfidence in Initial Choices and Underconfidence Under Criticism Modulate Change of Mind in Large Language Models Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 119cf71b-9e18-4ad7-9f5f-5f35ee6de10d · inbound
Co-DETECT: Collaborative Discovery of Edge Cases in Text Classification Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06be8adb-b560-492c-8f97-a84a51690bcc · inbound
Uncertainty-Driven Expert Control: Enhancing the Reliability of Medical Vision-Language Models Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e257a96-023e-4eee-8c10-a9eead34d43a · inbound
Counterfactual Probing for Hallucination Detection and Mitigation in Large Language Models Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc4418c2-6516-4fcd-80a3-1e771d10e976 · inbound
Mind the Generation Process: Fine-Grained Confidence Estimation During LLM Generation Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e6f288c-3cec-4927-bf7e-43380ab7134b · inbound
PaVeRL-SQL: Text-to-SQL via Partial-Match Rewards and Verbal Reinforcement Learning Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cedf85ec-f4cd-42ec-b15c-6094bf857b8e · inbound
Inteligencia Artificial jur\'idica y el desaf\'io de la veracidad: an\'alisis de alucinaciones, optimizaci\'on de RAG y principios para una integraci\'on responsable Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08ebd968-1697-4d2e-83ff-3c215f9ea0ec · inbound
LAVA: Language Model Assisted Verbal Autopsy for Cause-of-Death Determination Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6af8266c-4bf5-4722-9483-fc653b770f7d · inbound
Unsupervised Hallucination Detection by Inspecting Reasoning Processes Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19cc7309-0a75-4058-b68a-2f9250200bdb · inbound
HalluField: Detecting LLM Hallucinations via Field-Theoretic Modeling Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89b1e2df-80f6-4170-9ac4-b957549c4d8e · inbound
Enhancing Trustworthy GUI Grounding via Self-Critiqued Reinforcement Learning Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7608833d-b797-46e0-a530-5a89453a8397 · inbound
Can LLMs Estimate Student Struggles? Human-AI Difficulty Alignment with Proficiency Simulation for Item Difficulty Prediction Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 07c74a56-30da-4241-acee-22357dc2f946 · inbound
UCPO: Uncertainty-Aware Policy Optimization Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e84a639-ce49-42db-bf77-d54e9c57ad46 · inbound
Uncertainty-aware Generative Recommendation Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e571b856-0a57-4b15-bfc3-fc1a4353dd58 · inbound
How do LLMs Compute Verbal Confidence Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c860152e-3ff7-4350-a450-6abed1da25e8 · inbound
Causal Evidence that Language Models use Confidence to Drive Behavior Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 5b2b2309-44e6-43e7-b6ff-a8093a8480b3 · inbound
SELFDOUBT: Uncertainty Quantification for Reasoning LLMs via the Hedge-to-Verify Ratio Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4c2fd448-564b-408d-a88a-a2548d5532ba · inbound
Act or Escalate? Evaluating Escalation Behavior in Automation with Language Models Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d08e21b5-500f-47b5-a3d3-d2a31853ccae · inbound
UsefulBench: Towards Decision-Useful Information as a Target for Information Retrieval Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ee48544c-d7c3-47e2-b33b-eb6dc618992b · inbound
Verbal Confidence Saturation in 3-9B Open-Weight Instruction-Tuned LLMs: A Pre-Registered Psychometric Validity Screen Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8113b983-2179-454c-896e-cce94e929d75 · inbound
How LLMs Detect and Correct Their Own Errors: The Role of Internal Confidence Signals Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation de793346-1123-4430-bd84-12ce8735d9e8 · inbound
Confidence Estimation in Automatic Short Answer Grading with LLMs Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ad7b4369-ef84-4889-932a-bf2d6e516a18 · inbound
Confidence Estimation in Automatic Short Answer Grading with LLMs Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e9ed1b3e-a20d-4561-8656-57b2b07cc8e0 · inbound
Zero-Shot Confidence Estimation for Small LLMs: When Supervised Baselines Aren't Worth Training Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 59ac8232-900e-467f-be2e-c0fedc3eede7 · inbound
LLMs are not (consistently) Bayesian: Quantifying internal (in)consistencies of LLMs' probabilistic beliefs Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 091a3177-bbad-4bd9-9919-eb95c3726809 · inbound
Inducing Artificial Uncertainty in Language Models Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8edc7c8f-98b6-4ad1-b7b2-63d67aac3c24 · inbound
Margin-Adaptive Confidence Ranking for Reliable LLM Judgement Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 78049a4b-e569-49a0-ae1c-4b1bc001a279 · inbound
MARGIN: Runtime Confidence Calibration for Multi-Agent Foundation Model Coordination Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 5e85ed73-3521-409a-9b45-06c43adfaae2 · inbound
MARGIN: Runtime Confidence Calibration for Multi-Agent Foundation Model Coordination Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 142b36c8-98e7-4bc7-8993-b252bd342ebc · inbound
MARGIN: Runtime Confidence Calibration for Multi-Agent Foundation Model Coordination Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4f7439e-d337-448d-8cf7-f44492eb1078 · inbound
Confidence Calibration in Large Language Models Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e35011d6-a110-4991-bc9a-add26e706319 · inbound
Functional Entropy: Predicting Functional Correctness in LLM-Generated Code with Uncertainty Quantification Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 9e1c07db-f4f3-4a79-8465-d045fab9c574 · inbound
VLAConf: Calibrated Task-Success Confidence for Vision-Language-Action Models Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a9316994-81e5-4599-ae7a-a6202369a667 · inbound
Beyond Agreement: Scoring Panel-Surfaced Biomedical Entity Candidates for Curator Triage Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c081905b-0e32-479a-a5ba-b78c17033301 · inbound
NBQ: Next-Best-Question for Dynamic Profiling Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 5d0c1da4-8c9c-4093-844b-64238ab22f14 · inbound
Can LLM Rerankers Predict Their Own Ranking Performance? Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 49cea165-9b97-443f-b295-d6c7a0ba96ea · inbound
The Measurement Gap in the Automation of EU Law: Benchmarking Doctrinal Legal Reasoning under the EU AI Act Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e1eed9f8-467b-4df8-8149-7e599f67cbbd · inbound
Confidence Calibration for Multimodal LLMs: An Empirical Study through Medical VQA Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e622431a-d4eb-401d-a148-c66978bcc057 · inbound
Beyond Logprobs: A Multi-Signal Confidence Engine for LLM-Based Document Field Extraction Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 259b9f1f-d222-48c4-baf7-7cb74bc18f70 · inbound
Just how sure are you? Improving Verbalized Uncertainty Calibration in Medical VQA Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b1b13f8b-bcc3-4d38-bb68-fe9cca14be69 · inbound
Reported Confidence in LLMs Tracks Commitment More Than Correctness Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation bf219580-2e9f-420b-9c2e-50f8b5cd14fc · inbound
Scaling with Confidence: Calibrating Confidence of LLMs for Adaptive Test Time Scaling Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a2034867-06e2-4406-8eb4-5e8bc1ce6a27 · inbound
Estimating Uncertainty from Reasoning: A Large-Scale Study of Multi- and Crosslingual MCQA Performance in LLMs Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 02576b22-d371-4fae-a7d4-85bdd210f3b7 · inbound
The Computational Basis of Confidence in Large Language Models Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40fa64b1-f4d3-4ab8-be2c-ecec87abbdde · inbound
Small Vision-Language Models Know When They Are Wrong But Cannot Say So: A Two-Model Study of Stated versus Internal Confidence Under Realistic Image Degradation Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f410d905-d7d6-47c2-9bcf-a59056b4edb9 · inbound
Bigger or Cheaper? Scale and Quantization Effects on Uncertainty Signals in Vision-Language Models Under Image Degradation Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e4b0f74-8b50-49b1-b90b-1807ef6b6d13 · inbound
One Human, $N$ Agents: Audit-Budget Allocation for LLM Agent Fleets under Miscalibrated, Correlated Confidence Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e044ebba-bf2a-40b4-8c3f-3ae1fe444e5b · inbound
CARE: Confidence-Aware Reasoning for Reliable Medical VQA Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 028d454e-2beb-4367-b1fb-0f89be3d99cc · inbound
Explanatory Engagement Under Rare Anomalous Failure: Asymptotic Rarity in Model Behavior (or: The Asymptotic AI) Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.