Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T19:01:18.153145Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 100 of 143 outbound references and 0 inbound Pith citation observations for arXiv:2606.00869.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T19:01:18.153145Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
100 of 143 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 6e8ddb13-d959-4422-9b56-a948ff9426aa · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Kimi K2.5: Visual Agentic Intelligence
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3e27f3d8-cc51-4dcd-9db2-313b076fff82 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Qwen3 Technical Report
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f826adc3-6a16-47a0-820c-41ed6a9e9b75 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training DeepSeek-V4: Towards highly efficient million-token context intelligence,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2664db67-8658-47b0-9f2c-8be158890c6a · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training OpenClaw-RL: Train Any Agent Simply by Talking
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 01c10909-202d-4780-a831-d277e9189004 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Magpie: Alignment data synthesis from scratch by prompting aligned LLMs with nothing,
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dda4e049-f471-4511-a460-3987c6c3876d · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Tongyi DeepResearch Technical Report
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0b73d65d-9913-4209-86cf-830410a43f4f · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Language models are few-shot learners,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b034ec21-6d64-4e82-a233-b050464ceda3 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Agent Hospital: A Simulacrum of Hospital with Evolvable Medical Agents
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ec3f4121-2ca0-48b7-b902-31ee327523cb · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Cbr-rag: case-based reasoning for retrieval augmented generation in llms for legal question answering,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1263faa-8c19-466e-8740-322a3025e680 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Improving Retrieval for RAG based Question Answering Models on Financial Documents
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5ee71bc8-a2b4-4e17-88cb-69ce72df285f · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training WildClawBench: A Benchmark for Real-World, Long-Horizon Agent Evaluation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 536ded1c-c08a-4e55-8639-549244faeaeb · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Hallucinations Undermine Trust; Metacognition is a Way Forward
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 48bbdb7d-c90a-475e-b46b-b7eafbf7f55b · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Metacognition and cognitive monitoring: A new area of cognitive-developmental inquiry,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b979ac2-2b79-4fb1-8aaa-4e319465e85f · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Language Models (Mostly) Know What They Know
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ae1f67ed-030e-4eb2-90e2-f674a127993f · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Deepseek-r1 incentivizes reasoning in llms through reinforcement learning,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19f04b23-eba5-4293-a7d0-3b07fc25ec0c · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Are Reasoning Models More Prone to Hallucination?
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b45fbbee-7570-484d-9940-b4e60a599413 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 46c1d457-8f31-445a-aa80-e03da166da28 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Haotian Luo, Li Shen, Haiying He, Yibo Wang, Shi- wei Liu, Wei Li, Naiqiang Tan, Xiaochun Cao, and Dacheng Tao
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bc404058-8563-4522-8e45-f5c31132fa01 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training The hallucination tax of reinforcement finetuning,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9164ef4-4993-4756-b43b-715dcb4ba286 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training A survey on hallucination in large language models: Principles, taxonomy, challenges, and open questions,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ea8474f-eead-4af0-8688-406c3a8d4427 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training A survey of confidence estimation and calibration in large language models,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e54d488d-13ca-4b09-8acb-a52de0d7c7a7 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training R-tuning: Instructing large language models to say ‘i don’t know’,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcde1d15-6972-4ab5-9dcc-adcf6d8f9fc7 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Beyond “i don’t know
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76586b7a-ba5f-4e02-af6c-ef27f171fc42 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Inference-time scaling for generalist reward modeling
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7b1533b5-c7d5-4b7c-b113-75ec43816771 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Let’s verify step by step,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bdf2e05-4eb1-4d78-b2cd-0c719b5ca6bc · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Agent Learning via Early Experience
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 69f4d5fa-622e-4ea7-af9d-c4a5e2a7a8fb · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training General agents need world models,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0148317a-91c0-4eb4-af55-577250628e3e · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Model Spec Midtraining: Improving How Alignment Training Generalizes
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3bf6a049-c66c-470b-bb4d-773a5d06abf7 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Training language models to follow instructions with human feedback,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be427412-84bf-4b43-a144-80da635d787e · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Judging LLM-as-a-Judge with MT-Bench and chatbot arena,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22be0e16-e232-4b4e-ad20-8d2fe9c07c9f · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Large Language Models are not Fair Evaluators
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 96c9eaa7-6613-43ec-8d14-cadead6e7799 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training The False Promise of Imitating Proprietary LLMs
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 98974800-0815-4d80-b88f-fade31ec5c41 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training The Llama 3 Herd of Models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d75d3a30-7166-4658-8725-27c38f14f4ce · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Olmo 3
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1c7dbb91-ded5-459e-b334-c1a6e338d509 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training AbstentionBench: Reasoning LLMs Fail on Unanswerable Questions
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9a9ffb7a-3497-4017-8609-368dc776c869 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Self-refine: Iterative refinement with self-feedback,
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 793a895a-3150-4339-b11d-08e14176726a · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Reflexion: Language agents with verbal reinforcement learning,
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 297b0023-3a7d-453c-bc1a-5b71026d70f3 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Beyond Binary Rewards: Training LMs to Reason About Their Uncertainty
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4009fc6c-eddc-4428-a058-4313f964beae · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training arXiv preprint arXiv:2503.02623 (2025)
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f89db0cc-e6cd-423d-8f01-af7e97b9fdad · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training MASH: Modeling Abstention via Selective Help-Seeking
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 22626431-2395-4ad5-83f5-eeca25ed775e · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training SelectLLM: Query-aware efficient selection algorithm for large language models,
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bb3cf00-a191-4d71-8362-43f8533d3dd1 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Know More, Know Clearer: A Meta-Cognitive Framework for Knowledge Augmentation in Large Language Models
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 99a67573-e027-4352-bcb5-01aeddfabfc4 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 88e877e0-dd6b-4dbf-b637-782b6253c19a · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d8670fb6-0709-4e4b-87e2-16f1c366d255 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Measuring Mathematical Problem Solving With the MATH Dataset
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ed3a809b-6e16-4dfe-9c23-b6b936641b66 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c34744f3-830a-4218-91fc-9156b47103ee · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Solving quantitative reasoning problems with language models,
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef3ecce6-8f6c-48e3-8a58-21d2d154cbff · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training AIMO validation AMC: Problems from AMC 12 2022–2023
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5ceee4f-0029-4968-9fd1-85633bfcbb75 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training AIMO validation AIME: Problems from AIME 2022–2024
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96b2c637-b915-4de9-9ba4-30cf970d492c · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Direct preference optimization: Your language model is secretly a reward model,
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2dd81aa2-bd44-4698-ac8e-88695aac5ab3 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Hybridflow: A flexible and efficient rlhf framework,
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7c0216d-0df6-49d0-af4a-e7b23f66b916 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Efficient memory management for large language model serving with PagedAttention,
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8eaa26c5-9c30-425c-b68c-50887251c3cf · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training DRAGged into Conflicts: Detecting and Addressing Conflicting Sources in Search-Augmented LLMs
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 27c8c8b0-526f-4e3c-a001-10c8cf58049e · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Justrl: Scaling a 1.5 b llm with a simple rl recipe
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8bdf3b5b-6178-47ba-a4a9-9eef55a067c6 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Ur2: Unify rag and reasoning through reinforcement learning,
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30aa5c0e-6835-40d4-adb9-0b46876d988e · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training OpenMathReasoning: A large-scale dataset for mathematical reasoning
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de9bac03-2f72-459d-87e5-4b5db66b4fba · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Deepscaler: Surpassing o1-preview with a 1.5b model by scaling rl
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8754b689-100b-48a9-b502-a305724c3190 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training ALCUNA: Large language models meet new knowledge,
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2596288f-d4fb-4d71-9fd0-1e21afa726fe · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training BBQ: A hand-built bias benchmark for question answering,
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 510c8ede-dcf2-4fb3-93e5-06e22cbfafca · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Beyond the imitation game: Quantifying and extrapolating the capabilities of language models,
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b242b36f-a163-4102-8952-ff737cb88192 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training The Art of Saying No: Contextual Noncompliance in Language Models
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 17ff83c3-bbb6-40a6-9d5f-f4457095cf0e · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Won’t get fooled again: Answering questions with false premises,
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af7f8dda-8e5e-49b8-ab79-01530d7e72f7 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training GPQA: A Graduate-Level Google-Proof Q&A Benchmark
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 390869fe-e783-4c9e-b84d-b8175df683cc · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Training Verifiers to Solve Math Word Problems
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cc734824-a6f8-48ef-9c07-921805a7f29b · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Knowledge of knowledge: Exploring known- unknowns uncertainty with large language models,
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2918d1a-655d-443e-be30-0de0d1c28bf9 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training MediQ: Question- asking LLMs and a benchmark for reliable interactive clinical reasoning,
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f6f3b94-b327-4c5b-8412-19444b39f0f2 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Measuring Massive Multitask Language Understanding
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 05c61c68-1a22-45b7-ac26-9509b07f215d · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Evaluating the moral beliefs encoded in LLMs,
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f0d2a6b-01c7-424c-8bb2-cbfa76c76257 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training MuSiQue: Multihop questions via single-hop question composition,
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4b97692-3725-4f11-adea-1ede892a549b · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training (QA)2: Question answering with questionable assumptions,
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b15a0044-e5cf-4b89-9065-12387670db2d · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training A dataset of information-seeking questions and answers anchored in research papers,
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff3f6037-ebbd-4de5-bb89-69c38c2c9f0c · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training SituatedQA: Incorporating extra-linguistic contexts into QA,
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 020ab70b-345c-4e92-b2d2-d44c20e7eb48 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Know what you don’t know: Unanswerable questions for SQuAD,
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9aca73f5-75ae-4ec9-b117-543e1dce8d0d · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Benchmarking Hallucination in Large Language Models based on Unanswerable Math Word Problem
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 87a0a974-f900-43b5-b4e8-7a7afb72d8e2 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training WorldSense: A Synthetic Benchmark for Grounded Reasoning in Large Language Models
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation abf0102e-ac5e-4a66-b00b-73e2c003fadc · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training A coefficient of agreement for nominal scales,
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc117fe0-dd0b-499f-9a82-f756c76c995a · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training The measurement of observer agreement for categorical data,
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73959039-848e-4cdc-99ba-858d46b7a98a · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Measuring nominal scale agreement among many raters,
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05335121-b50c-4b1a-924d-b1738993691e · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Quagmires in sft-rl post-training: When high sft scores mislead and what to use instead
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f2d39fa9-1d82-46a6-a59a-91ebb8415b88 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Beyond Two-Stage Training: Cooperative SFT and RL for LLM Reasoning
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 42dd68ee-d222-4ae8-a4eb-d2e076113b97 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a6ce62f4-2f0c-435d-9727-edc4d7977589 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Enhancing LLM Reasoning with Iterative DPO: A Comprehensive Empirical Investigation
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d0349acd-d333-4845-8385-e9e4ec6bbcf1 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f52aa683-347a-41b1-b4a1-96e1d3268983 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Tulu 3: Pushing Frontiers in Open Language Model Post-Training
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9bd659e4-fe25-4360-b389-5b52ea14fdef · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Scaling Synthetic Data Creation with 1,000,000,000 Personas
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bef76b20-5982-4376-bab4-4beffeef4bfb · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training The flan collection: Designing data and methods for effective instruction tuning,
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b13a5645-3881-4eef-b975-4b12e6774f7e · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training No robots
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8b8d6ec-418d-442c-8b39-433afeead9dc · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Openassistant conversations-democratizing large language model alignment,
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3887d8e-8486-4f29-8db4-743526df9e16 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Aya dataset: An open-access collection for multilingual instruction tuning,
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65a97c2b-5886-440c-9cd5-8062bf0075e1 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training TableGPT: Towards Unifying Tables, Nature Language and Commands into One GPT
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 10f6f86e-b98a-4838-9fcc-e75f0f78d3f3 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training WildChat: 1M ChatGPT Interaction Logs in the Wild
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ec15ae00-cfa5-4f1a-a14a-e9bf671ad825 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Wizardcoder: Empower- ing code large language models with evol-instruct,
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dcbf4f63-cb42-4545-9a92-c9d2cde371c6 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Wildguard: Open one-stop moderation tools for safety risks, jailbreaks, and refusals of llms,
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 841cab1c-4f80-4c4e-8ac5-8a21afe1e32c · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Wildteaming at scale: From in-the-wild jailbreaks to (adversarially) safer language models,
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11a2b820-3a99-4a15-85ca-7566c57a6205 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Sciriff: A resource to enhance language model instruction-following over scientific literature,
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a889d61-35bc-4534-94a6-cd68941ec8b0 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Unresolved cited work
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 550ad1d4-ac42-4ac7-8583-59784a9471c9 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training Unresolved cited work
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c923e4e-db04-4b16-9032-a24d3db713bf · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training I don’t know
Reference 98
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7df0871e-ed7c-4dce-bdd4-6b37b4390253 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training For base checkpoints, the dataset loader extracts the raw prompt text from the stored chat-style field and tokenizes it directly, instead of applying a chat template
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f1a6f02-4963-4586-8152-bfcfc5519889 · outbound
Enhancing LLM Metacognition via Cognitive Pairwise Training For each promptx, the policy samples a group ofGresponses{y i}G i=1 from vLLM [52]; in the main Math-RL scriptG= 16, temperature is 0.9, andtop-p=0.95
Reference 100
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.