Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T19:56:09.820812Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:2606.08088.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T19:56:09.820812Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
49 of 49 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 0eac99e5-97d9-40c1-a7d4-60571501673e · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Advances in Neural Information Processing Systems , volume=
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 457f3a84-162b-4381-8e5e-f7a65fda66f9 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning DeepMath-103K: A Large-Scale, Challenging, Decontaminated, and Verifiable Mathematical Dataset for Advancing Reasoning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5e4d0d75-8a83-4d5d-8486-225e33f4ba50 · outbound
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation aa580c9a-1676-4f23-8d9d-52405be4a5f2 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning CodeRL+: Improving Code Generation via Reinforcement with Execution Semantics Alignment
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a4fbc573-08a0-4e9b-bcd5-e54607ad84ea · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning arXiv preprint arXiv:2601.18533 , year=
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f129b06e-b4d6-4fd2-bbd7-2f3857fb4541 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Hansen and Duo Peng and Yuhui Zhang and Alejandro Lozano and Min Woo Sun and Emma Lundberg and Serena Yeung-Levy , year=
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ce1eebc0-fa28-43cd-9131-117f7bf7fd6b · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 64b39f26-c202-4303-b120-3769212a0f2d · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Beyond Binary: Turning Partial Success into Dense Verifiable Rewards for Reinforcement Learning in Code Generation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5795dac7-6c37-4d11-a626-c7a151ce49f5 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Save the good prefix: Precise error penalization via process-supervised rl to enhance llm reasoning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 96734832-56cb-42d8-90c9-2cd72e1a68b6 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Probabilities of Chat LLMs Are Miscalibrated but Still Predict Correctness on Multiple-Choice Q&A
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a4df47c0-54a5-4439-8c05-c1f63bbcbcc6 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Advances in Neural Information Processing Systems , volume=
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81f6aee7-5482-4a80-b541-26cb9dcd9333 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing , pages=
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 718dc540-162a-4b4e-ba19-b8a81123df5c · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Self-Consistency Improves Chain of Thought Reasoning in Language Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation fd4058a2-76b9-4c82-9ac9-bb9209ebcec3 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Proceedings of the 2023 conference on empirical methods in natural language processing , pages=
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c9aaf55-e41e-4f80-9b30-39471778c1b3 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Can AI Assistants Know What They Don't Know?
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 2e75d40d-a96f-491f-82ca-8a19ca6aab03 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Advances in Neural Information Processing Systems , volume=
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1cdb34c0-bea1-4a2e-88bb-51c9e8898260 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Enhancing Confidence Expression in Large Language Models Through Learning from Past Experience
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 69a4c557-380b-43e4-82f2-9f8c87ed3f95 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Advances in neural information processing systems , volume=
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be217a70-f455-46b5-a1e4-3587adc67dca · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning International Conference on Learning Representations , volume=
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8d6598e-6707-4d0e-8607-463a32089b58 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Findings of the Association for Computational Linguistics: ACL 2024 , year=
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59cb0b50-4c60-4f3c-abb6-e8925981db91 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Findings of the Association for Computational Linguistics: EMNLP 2023 , pages=
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c3a458f-77f4-4c05-af5f-d4c24edde5ab · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning arXiv preprint arXiv:2503.02623 (2025)
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 711a585e-f576-4671-8362-1e59c642e8ca · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing , pages=
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 526eddf0-9368-4a52-81aa-eae1d528d6d9 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Findings of the Association for Computational Linguistics: EMNLP 2023 , pages=
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6e5bd6e-77d1-4f5d-80be-4eb688219c2b · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Semantic Uncertainty: Linguistic Invariances for Uncertainty Estimation in Natural Language Generation
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation caa66479-c02b-41e8-92a3-5e8caa048f64 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 6acfe10f-7eb1-411e-b246-8ed429bd522f · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da3faedb-e85a-4033-a526-74de54257a75 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Deep Think with Confidence
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation fa41559d-c7d7-4b17-8d5a-b67068d91256 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Harder is better: Boost- ing mathematical reasoning via difficulty-aware grpo and multi-aspect question reformulation
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation fef274b5-193e-442b-883b-76f7d264321b · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Understanding R1-Zero-Like Training: A Critical Perspective
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation fc381771-0383-4da4-9c5c-5aae583d23cb · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Advances in Neural Information Processing Systems , volume=
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3d1ff10-4d27-43ec-8f1f-fc2b63b4a30a · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Group Sequence Policy Optimization
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e6374cfc-faf8-4d27-bc8b-b7406f62002d · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Soft Adaptive Policy Optimization
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation acf14a0f-0a73-4baa-aec8-9768f1748718 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4c401f2b-8499-4af7-9022-9e5fa945f24c · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Proximal Policy Optimization Algorithms
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 26e49a38-037d-41f7-972b-ebddbf7287e7 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 1934d145-c540-4c41-9459-d1bae4cdbfcb · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Proceedings of the Twentieth European Conference on Computer Systems , pages=
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6df45bed-9891-4866-9394-011220556066 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Qwen3 Technical Report
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 368189ad-bd44-4324-b428-4e84ad633ca6 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 8151db6a-0ee4-488c-92c7-049b71de7618 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 86a42992-4a6b-46bc-957d-59eadd424386 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Measuring Massive Multitask Language Understanding
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation cff1f99b-280b-4be0-a7dd-a7e6ac01f7b0 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Advances in neural information processing systems , volume=
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe411c30-a9a1-4d1f-b629-5f0408d35319 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages=
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f39cbde-8113-4c3d-9d18-7008bbbd13ca · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 9dfb7d77-5ad6-4732-a83e-ee13de50b8cc · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning arXiv preprint arXiv:2503.17736 , year=
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 7a0ee72f-3988-4eb9-9088-b2e897cb7516 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f7e8c7b0-933b-480c-a001-3ca2db9532f1 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Vision-deepresearch benchmark: Rethinking visual and textual search for multimodal large language models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 55701cbf-2b2c-4166-b565-84f8ab479874 · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning arXiv preprint arXiv:2510.01304 , year=
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ffb93de8-b762-45ef-97df-7016bb2954df · outbound
ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning VideoSeeker: Incentivizing Instance-level Video Understanding via Native Agentic Tool Invocation
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
No inbound Pith citation observations are available.