Pith. sign in

Paper Citation Record · LEDGER

Enhancing LLM Metacognition via Cognitive Pairwise Training

As of 4 August 2026, this Paper Citation Record lists 100 of 143 outbound references and 0 inbound Pith citation observations for arXiv:2606.00869.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.00869 v1

Coverage vector

measured 100 of 143 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-28T19:01:18.153145Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 143 outbound references displayed

  • verified exact39
  • verified fuzzy0
  • unresolved55
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch6

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6e8ddb13-d959-4422-9b56-a948ff9426aa · outbound

This paper cites Kimi K2.5: Visual Agentic Intelligence.

Enhancing LLM Metacognition via Cognitive Pairwise Training Kimi K2.5: Visual Agentic Intelligence

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.919403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:9d253d2887ea3da086e4f7b41cd09c3923b19019f21320af6c366e524f62db04

Observation 3e27f3d8-cc51-4dcd-9db2-313b076fff82 · outbound

This paper cites Qwen3 Technical Report.

Enhancing LLM Metacognition via Cognitive Pairwise Training Qwen3 Technical Report

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.921975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:a6e1eec8bcb1a385584dd47dc591796e3b3a4da0f187cfc895dfe92ebdf3af49

Observation f826adc3-6a16-47a0-820c-41ed6a9e9b75 · outbound

This paper cites DeepSeek-V4: Towards highly efficient million-token context intelligence,.

Enhancing LLM Metacognition via Cognitive Pairwise Training DeepSeek-V4: Towards highly efficient million-token context intelligence,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:2594c4ee4cb7a4097ac429e3a987607bef4596047aafbbd35a1127291a9e83ab

Observation 2664db67-8658-47b0-9f2c-8be158890c6a · outbound

This paper cites OpenClaw-RL: Train Any Agent Simply by Talking.

Enhancing LLM Metacognition via Cognitive Pairwise Training OpenClaw-RL: Train Any Agent Simply by Talking

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.914341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:17d6b9538cefff5e27e75aa9bb5eab48ddbcdb290e4f35d6406c2aea52269455

Observation 01c10909-202d-4780-a831-d277e9189004 · outbound

This paper cites Magpie: Alignment data synthesis from scratch by prompting aligned LLMs with nothing,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Magpie: Alignment data synthesis from scratch by prompting aligned LLMs with nothing,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:e433cb0f1616cb6f022f0f14e09f5cbfb1141ada33edaad174d9eaebe91f000d

Observation dda4e049-f471-4511-a460-3987c6c3876d · outbound

This paper cites Tongyi DeepResearch Technical Report.

Enhancing LLM Metacognition via Cognitive Pairwise Training Tongyi DeepResearch Technical Report

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.916799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:bf6d7808ec64eaac2181f732781973b211a8d820ff0ac9f5a1ecfc33733686e8

Observation 0b73d65d-9913-4209-86cf-830410a43f4f · outbound

This paper cites Language models are few-shot learners,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Language models are few-shot learners,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:5fceabe990cb9dd64d389c3bedd6a1c80582ab0d6af371633f75d01c8b002333

Observation b034ec21-6d64-4e82-a233-b050464ceda3 · outbound

This paper cites Agent Hospital: A Simulacrum of Hospital with Evolvable Medical Agents.

Enhancing LLM Metacognition via Cognitive Pairwise Training Agent Hospital: A Simulacrum of Hospital with Evolvable Medical Agents

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:02:33.924827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:caceb866dc7f68ebad61dcdcc8baf485461f738606ab679dee62c834342d2069

Observation ec3f4121-2ca0-48b7-b902-31ee327523cb · outbound

This paper cites Cbr-rag: case-based reasoning for retrieval augmented generation in llms for legal question answering,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Cbr-rag: case-based reasoning for retrieval augmented generation in llms for legal question answering,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:80f9cb2aae828cb8cb562eca062de5fda2df53a8a4d2a5141639320905befaf4

Observation c1263faa-8c19-466e-8740-322a3025e680 · outbound

This paper cites Improving Retrieval for RAG based Question Answering Models on Financial Documents.

Enhancing LLM Metacognition via Cognitive Pairwise Training Improving Retrieval for RAG based Question Answering Models on Financial Documents

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T19:02:33.911742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:fa55e8fb649a663be2bdebb15362949b16255312567cd23c2571601d6641434f

Observation 5ee71bc8-a2b4-4e17-88cb-69ce72df285f · outbound

This paper cites WildClawBench: A Benchmark for Real-World, Long-Horizon Agent Evaluation.

Enhancing LLM Metacognition via Cognitive Pairwise Training WildClawBench: A Benchmark for Real-World, Long-Horizon Agent Evaluation

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.881402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:2a0b9bf21710d327a1b60c801199cdf36ab66462628bcceb6a7471ea240306e6

Observation 536ded1c-c08a-4e55-8639-549244faeaeb · outbound

This paper cites Hallucinations Undermine Trust; Metacognition is a Way Forward.

Enhancing LLM Metacognition via Cognitive Pairwise Training Hallucinations Undermine Trust; Metacognition is a Way Forward

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.880648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:670ec52d778a63b3f81b926139ca9758cf726ec26bb5ebe8f45d5216ec0f8d7b

Observation 48bbdb7d-c90a-475e-b46b-b7eafbf7f55b · outbound

This paper cites Metacognition and cognitive monitoring: A new area of cognitive-developmental inquiry,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Metacognition and cognitive monitoring: A new area of cognitive-developmental inquiry,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:272d854660e5de3ad63569bb2d3f885d5a424e5b267e4669302ad8849a711cd6

Observation 7b979ac2-2b79-4fb1-8aaa-4e319465e85f · outbound

This paper cites Language Models (Mostly) Know What They Know.

Enhancing LLM Metacognition via Cognitive Pairwise Training Language Models (Mostly) Know What They Know

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.876321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:52bb981b259d597f24fb43e5b99f664355da2a61379421097fabb79175eac890

Observation ae1f67ed-030e-4eb2-90e2-f674a127993f · outbound

This paper cites Deepseek-r1 incentivizes reasoning in llms through reinforcement learning,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Deepseek-r1 incentivizes reasoning in llms through reinforcement learning,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:779cd3dc7590bc4c1bc055f3167a2049954b7bb8200c0c2d4ffd43844cc07da2

Observation 19f04b23-eba5-4293-a7d0-3b07fc25ec0c · outbound

This paper cites Are Reasoning Models More Prone to Hallucination?.

Enhancing LLM Metacognition via Cognitive Pairwise Training Are Reasoning Models More Prone to Hallucination?

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:02:33.919996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:f7844c06baa8a090cb1e105ac586b7a360cf5bdaffe546bf5cadcc16f0a95e97

Observation b45fbbee-7570-484d-9940-b4e60a599413 · outbound

This paper cites Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models.

Enhancing LLM Metacognition via Cognitive Pairwise Training Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.917472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:71572dd3b2001193708a82cae9f48ef8585f79179b960861a1b33de67b9ea962

Observation 46c1d457-8f31-445a-aa80-e03da166da28 · outbound

This paper cites Haotian Luo, Li Shen, Haiying He, Yibo Wang, Shi- wei Liu, Wei Li, Naiqiang Tan, Xiaochun Cao, and Dacheng Tao.

Enhancing LLM Metacognition via Cognitive Pairwise Training Haotian Luo, Li Shen, Haiying He, Yibo Wang, Shi- wei Liu, Wei Li, Naiqiang Tan, Xiaochun Cao, and Dacheng Tao

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T19:02:33.930698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:64367eacfdb9ade0a28b61ece9c89fd5a328a30d169dcef6869608ec4c3ac703

Observation bc404058-8563-4522-8e45-f5c31132fa01 · outbound

This paper cites The hallucination tax of reinforcement finetuning,.

Enhancing LLM Metacognition via Cognitive Pairwise Training The hallucination tax of reinforcement finetuning,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:c54a1c298b3cbfb747a4c34365f1a37501eb10db45ccd897d7f80b57475ecd28

Observation b9164ef4-4993-4756-b43b-715dcb4ba286 · outbound

This paper cites A survey on hallucination in large language models: Principles, taxonomy, challenges, and open questions,.

Enhancing LLM Metacognition via Cognitive Pairwise Training A survey on hallucination in large language models: Principles, taxonomy, challenges, and open questions,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:edb4161bd56bd532c3f7d013bc5702c6d79f001028391edca11c9c111b632c07

Observation 3ea8474f-eead-4af0-8688-406c3a8d4427 · outbound

This paper cites A survey of confidence estimation and calibration in large language models,.

Enhancing LLM Metacognition via Cognitive Pairwise Training A survey of confidence estimation and calibration in large language models,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:d949f82bd77f799228f4204766981151c61052a7ab77c599b29758fad6be9335

Observation e54d488d-13ca-4b09-8acb-a52de0d7c7a7 · outbound

This paper cites R-tuning: Instructing large language models to say ‘i don’t know’,.

Enhancing LLM Metacognition via Cognitive Pairwise Training R-tuning: Instructing large language models to say ‘i don’t know’,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:b7efd168aa77a00477c85e4e34e50d6fafbab242f695eb7aaca4e0315e0c24bd

Observation bcde1d15-6972-4ab5-9dcc-adcf6d8f9fc7 · outbound

This paper cites Beyond “i don’t know.

Enhancing LLM Metacognition via Cognitive Pairwise Training Beyond “i don’t know

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:a9ce95a298722496f613f143c430c14c61a2383243af99c888e375d10b980a9d

Observation 76586b7a-ba5f-4e02-af6c-ef27f171fc42 · outbound

This paper cites Inference-time scaling for generalist reward modeling.

Enhancing LLM Metacognition via Cognitive Pairwise Training Inference-time scaling for generalist reward modeling

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:02:33.814905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:e592a8d3d725c652be64e6f7d771d7afacf5337b7c0704c35a6753b2c3618ca6

Observation 7b1533b5-c7d5-4b7c-b113-75ec43816771 · outbound

This paper cites Let’s verify step by step,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Let’s verify step by step,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:b16d22c62a60a7a4988db699bac37ae8a54e1a598f5a67be8b959499e2dbbd19

Observation 6bdf2e05-4eb1-4d78-b2cd-0c719b5ca6bc · outbound

This paper cites Agent Learning via Early Experience.

Enhancing LLM Metacognition via Cognitive Pairwise Training Agent Learning via Early Experience

Reference 26

Resolution
metadata mismatch
local_arxiv, observed 2026-06-28T19:02:33.901234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:a5e414a8a3d5cb23e5d045477482f187b2435ee515d950f6264ef7f88c771500

Observation 69f4d5fa-622e-4ea7-af9d-c4a5e2a7a8fb · outbound

This paper cites General agents need world models,.

Enhancing LLM Metacognition via Cognitive Pairwise Training General agents need world models,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:6e5da231d1bd98fade755ae5afbd9ce80ec72b0a2026cc09907478bf6de22d2a

Observation 0148317a-91c0-4eb4-af55-577250628e3e · outbound

This paper cites Model Spec Midtraining: Improving How Alignment Training Generalizes.

Enhancing LLM Metacognition via Cognitive Pairwise Training Model Spec Midtraining: Improving How Alignment Training Generalizes

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.927866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:d868984cff339ddd46950a78df000c1ec8dca9dd4d75f483a82aba00e2f22a61

Observation 3bf6a049-c66c-470b-bb4d-773a5d06abf7 · outbound

This paper cites Training language models to follow instructions with human feedback,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Training language models to follow instructions with human feedback,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:42d854c63d77ef15b742581d5c7d21f66461df6abf1b08b635ee78c7b78893d2

Observation be427412-84bf-4b43-a144-80da635d787e · outbound

This paper cites Judging LLM-as-a-Judge with MT-Bench and chatbot arena,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Judging LLM-as-a-Judge with MT-Bench and chatbot arena,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:a130a9918c875b8a9b3670015125d82fddb8a55129c4aaf63b24c0ac04b5341e

Observation 22be0e16-e232-4b4e-ad20-8d2fe9c07c9f · outbound

This paper cites Large Language Models are not Fair Evaluators.

Enhancing LLM Metacognition via Cognitive Pairwise Training Large Language Models are not Fair Evaluators

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.911851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:0e98d3b7d0df917b796fe9fde364d19b8b9fca83164cf77c77e586e58ea66e40

Observation 96c9eaa7-6613-43ec-8d14-cadead6e7799 · outbound

This paper cites The False Promise of Imitating Proprietary LLMs.

Enhancing LLM Metacognition via Cognitive Pairwise Training The False Promise of Imitating Proprietary LLMs

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.871338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:145321fdbe9ee673aa4ae5ddcf3a5091b737bd7ba66dbd81f5027cceb79f42e7

Observation 98974800-0815-4d80-b88f-fade31ec5c41 · outbound

This paper cites The Llama 3 Herd of Models.

Enhancing LLM Metacognition via Cognitive Pairwise Training The Llama 3 Herd of Models

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.873974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:89f8ffaf95207c0638cfa10484458f7c03e2154e859b440e8fde3d7025afabf5

Observation d75d3a30-7166-4658-8725-27c38f14f4ce · outbound

This paper cites Olmo 3.

Enhancing LLM Metacognition via Cognitive Pairwise Training Olmo 3

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.903560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:686403a424a6d210fa16e80e98379cc7843198c59e7b56aa2f28dcb088b0baf8

Observation 1c7dbb91-ded5-459e-b334-c1a6e338d509 · outbound

This paper cites AbstentionBench: Reasoning LLMs Fail on Unanswerable Questions.

Enhancing LLM Metacognition via Cognitive Pairwise Training AbstentionBench: Reasoning LLMs Fail on Unanswerable Questions

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:02:33.904071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:aa0a2e672249ad95a9bff83e44de9dd5a3ca2d3b72db7bf2a64582664b886f1c

Observation 9a9ffb7a-3497-4017-8609-368dc776c869 · outbound

This paper cites Self-refine: Iterative refinement with self-feedback,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Self-refine: Iterative refinement with self-feedback,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:9376008227570ccc8962d78576e6f6721e44c29153b9c875bd3e1f724d650186

Observation 793a895a-3150-4339-b11d-08e14176726a · outbound

This paper cites Reflexion: Language agents with verbal reinforcement learning,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Reflexion: Language agents with verbal reinforcement learning,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:fcc356879e109c349177efe4c8c6664446a8b4f3d41962add1e6dc022d51d89d

Observation 297b0023-3a7d-453c-bc1a-5b71026d70f3 · outbound

This paper cites Beyond Binary Rewards: Training LMs to Reason About Their Uncertainty.

Enhancing LLM Metacognition via Cognitive Pairwise Training Beyond Binary Rewards: Training LMs to Reason About Their Uncertainty

Reference 38

Resolution
metadata mismatch
local_arxiv, observed 2026-06-28T19:02:33.868887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:47dc8abf80b847c5ab9b7dfb392a7eb21113af9b68c9f45edf16780e3436f9c8

Observation 4009fc6c-eddc-4428-a058-4313f964beae · outbound

This paper cites arXiv preprint arXiv:2503.02623 (2025).

Enhancing LLM Metacognition via Cognitive Pairwise Training arXiv preprint arXiv:2503.02623 (2025)

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T19:02:33.884024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:d3da693fe974f72f423e8b3bc2a42bcea3eb1ff18dec37cc11a5a469983b21e4

Observation f89db0cc-e6cd-423d-8f01-af7e97b9fdad · outbound

This paper cites MASH: Modeling Abstention via Selective Help-Seeking.

Enhancing LLM Metacognition via Cognitive Pairwise Training MASH: Modeling Abstention via Selective Help-Seeking

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.824556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:989201d090f70c92dfaecc873dfc5cdbbfe697a738ba37f32b6dee766dfb95f6

Observation 22626431-2395-4ad5-83f5-eeca25ed775e · outbound

This paper cites SelectLLM: Query-aware efficient selection algorithm for large language models,.

Enhancing LLM Metacognition via Cognitive Pairwise Training SelectLLM: Query-aware efficient selection algorithm for large language models,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:d9f17bd1cf4c3eeca79d824fd0095c892fe99347a49965c6b8031b8d9258f2aa

Observation 7bb3cf00-a191-4d71-8362-43f8533d3dd1 · outbound

This paper cites Know More, Know Clearer: A Meta-Cognitive Framework for Knowledge Augmentation in Large Language Models.

Enhancing LLM Metacognition via Cognitive Pairwise Training Know More, Know Clearer: A Meta-Cognitive Framework for Knowledge Augmentation in Large Language Models

Reference 42

Resolution
metadata mismatch
local_arxiv, observed 2026-06-28T19:02:33.883180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:a732fd97a8a755130eb51eb5441a6adcece8e1df3f9cb8ffdf63593aa05f27de

Observation 99a67573-e027-4352-bcb5-01aeddfabfc4 · outbound

This paper cites SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild.

Enhancing LLM Metacognition via Cognitive Pairwise Training SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.906505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:85efd9a94b3232d7a2fbd1869c1ceb9ef35b8456e42bd1018ca17e51a121185d

Observation 88e877e0-dd6b-4dbf-b637-782b6253c19a · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

Enhancing LLM Metacognition via Cognitive Pairwise Training DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.841748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:fefaa1d132823791bc491aaf85de95037d85721d2faeced3ed10a871605878f8

Observation d8670fb6-0709-4e4b-87e2-16f1c366d255 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

Enhancing LLM Metacognition via Cognitive Pairwise Training Measuring Mathematical Problem Solving With the MATH Dataset

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.889375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:6a4c0df66dc33cbb0022bce8c2c7da0212e19f4b9f9cf6b0c0ef8b24ff255117

Observation ed3a809b-6e16-4dfe-9c23-b6b936641b66 · outbound

This paper cites OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems.

Enhancing LLM Metacognition via Cognitive Pairwise Training OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.889808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:8dbfcbe7cc4e3b381919cc0cc804a7980b642c9efbdd3cabe6d588214f6c75b9

Observation c34744f3-830a-4218-91fc-9156b47103ee · outbound

This paper cites Solving quantitative reasoning problems with language models,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Solving quantitative reasoning problems with language models,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:e9ebce0d0536b0c12f1db30f056aff871e5d9081924d4d279de744e27d49ea6f

Observation ef3ecce6-8f6c-48e3-8a58-21d2d154cbff · outbound

This paper cites AIMO validation AMC: Problems from AMC 12 2022–2023.

Enhancing LLM Metacognition via Cognitive Pairwise Training AIMO validation AMC: Problems from AMC 12 2022–2023

Reference 48

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:01ca12b97a57555440fc92621a8fb104c0da753975a41d5d5d2bbd3131eac805

Observation b5ceee4f-0029-4968-9fd1-85633bfcbb75 · outbound

This paper cites AIMO validation AIME: Problems from AIME 2022–2024.

Enhancing LLM Metacognition via Cognitive Pairwise Training AIMO validation AIME: Problems from AIME 2022–2024

Reference 49

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:d66835dc1f0f370f2f401292c0ed53a2decc0b003aa63a90d7436f47b8550497

Observation 96b2c637-b915-4de9-9ba4-30cf970d492c · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Direct preference optimization: Your language model is secretly a reward model,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:cfaa592c96da605acb571c8a1a7c43e1eb8686e66691d34f5e8c781e3d6d0f05

Observation 2dd81aa2-bd44-4698-ac8e-88695aac5ab3 · outbound

This paper cites Hybridflow: A flexible and efficient rlhf framework,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Hybridflow: A flexible and efficient rlhf framework,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:af52b189f4da3550a6f7893c019da636f66f3d174b0abab309d8aff2b59ebda6

Observation d7c0216d-0df6-49d0-af4a-e7b23f66b916 · outbound

This paper cites Efficient memory management for large language model serving with PagedAttention,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Efficient memory management for large language model serving with PagedAttention,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:5284e97b1e36014eb2dc7663f68b6d31d6a5d5d136f74221ba8993da3dc60796

Observation 8eaa26c5-9c30-425c-b68c-50887251c3cf · outbound

This paper cites DRAGged into Conflicts: Detecting and Addressing Conflicting Sources in Search-Augmented LLMs.

Enhancing LLM Metacognition via Cognitive Pairwise Training DRAGged into Conflicts: Detecting and Addressing Conflicting Sources in Search-Augmented LLMs

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:02:33.909346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:9bf82338222836cb78d51fdb18301267cdd886d590f1c38a887591508ff2f230

Observation 27c8c8b0-526f-4e3c-a001-10c8cf58049e · outbound

This paper cites Justrl: Scaling a 1.5 b llm with a simple rl recipe.

Enhancing LLM Metacognition via Cognitive Pairwise Training Justrl: Scaling a 1.5 b llm with a simple rl recipe

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:02:33.936053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:dfc77106050262064fd4f1341a2831824b27e5893b5267b96c3e9d5d1fd29617

Observation 8bdf3b5b-6178-47ba-a4a9-9eef55a067c6 · outbound

This paper cites Ur2: Unify rag and reasoning through reinforcement learning,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Ur2: Unify rag and reasoning through reinforcement learning,

Reference 55

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:a727377e2cdf894c6c1ed68e9bd3312544951ff05079a3a6910d9d9cf4fd9014

Observation 30aa5c0e-6835-40d4-adb9-0b46876d988e · outbound

This paper cites OpenMathReasoning: A large-scale dataset for mathematical reasoning.

Enhancing LLM Metacognition via Cognitive Pairwise Training OpenMathReasoning: A large-scale dataset for mathematical reasoning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:cded7adae91cbee0adde9e93b4ad863aa70f61ec12e930eb1551fe014d596c2c

Observation de9bac03-2f72-459d-87e5-4b5db66b4fba · outbound

This paper cites Deepscaler: Surpassing o1-preview with a 1.5b model by scaling rl.

Enhancing LLM Metacognition via Cognitive Pairwise Training Deepscaler: Surpassing o1-preview with a 1.5b model by scaling rl

Reference 57

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:643b8fb6b0591eeba7b046643f069a5d42fb37d918188c8af83ced8bce4f2858

Observation 8754b689-100b-48a9-b502-a305724c3190 · outbound

This paper cites ALCUNA: Large language models meet new knowledge,.

Enhancing LLM Metacognition via Cognitive Pairwise Training ALCUNA: Large language models meet new knowledge,

Reference 58

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:ec5f3cc72f418ce4600b92620ba8cd42539c819e672fce9017277d526982f375

Observation 2596288f-d4fb-4d71-9fd0-1e21afa726fe · outbound

This paper cites BBQ: A hand-built bias benchmark for question answering,.

Enhancing LLM Metacognition via Cognitive Pairwise Training BBQ: A hand-built bias benchmark for question answering,

Reference 59

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:6e82a9c40a593f05634de70f68e48aa9de6449b9f05e77b8df2570f04e6286cc

Observation 510c8ede-dcf2-4fb3-93e5-06e22cbfafca · outbound

This paper cites Beyond the imitation game: Quantifying and extrapolating the capabilities of language models,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Beyond the imitation game: Quantifying and extrapolating the capabilities of language models,

Reference 60

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:9178fbf12df3f6e49b2e8418f688512248b6a6e383cb2b6b8b8156b0f419395b

Observation b242b36f-a163-4102-8952-ff737cb88192 · outbound

This paper cites The Art of Saying No: Contextual Noncompliance in Language Models.

Enhancing LLM Metacognition via Cognitive Pairwise Training The Art of Saying No: Contextual Noncompliance in Language Models

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:02:33.914764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:317fc3431764846b2670e9cda2f36b9330e639c8ae7dde3fc43f4e38a3b4d231

Observation 17ff83c3-bbb6-40a6-9d5f-f4457095cf0e · outbound

This paper cites Won’t get fooled again: Answering questions with false premises,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Won’t get fooled again: Answering questions with false premises,

Reference 62

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:ed3c42ea99eaf080fa42e774ad39b54a7d6be8d0016e303bf6e8b2cfe7197fdc

Observation af7f8dda-8e5e-49b8-ab79-01530d7e72f7 · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

Enhancing LLM Metacognition via Cognitive Pairwise Training GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 63

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.886938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:c976ed7f9bb24c059dcdc2d7f9e5c0f63b29a0c10deaebe12967b3292c7c704d

Observation 390869fe-e783-4c9e-b84d-b8175df683cc · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Enhancing LLM Metacognition via Cognitive Pairwise Training Training Verifiers to Solve Math Word Problems

Reference 64

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.894758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:76f14120dd630499846f13b8a9156c06d7f15b97e23a1ebe9025aff1705ea36b

Observation cc734824-a6f8-48ef-9c07-921805a7f29b · outbound

This paper cites Knowledge of knowledge: Exploring known- unknowns uncertainty with large language models,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Knowledge of knowledge: Exploring known- unknowns uncertainty with large language models,

Reference 65

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:56b37c3b797a75f4da3351b235b75a1e8f98a3884a1b9eb0f88c5d23345a9478

Observation e2918d1a-655d-443e-be30-0de0d1c28bf9 · outbound

This paper cites MediQ: Question- asking LLMs and a benchmark for reliable interactive clinical reasoning,.

Enhancing LLM Metacognition via Cognitive Pairwise Training MediQ: Question- asking LLMs and a benchmark for reliable interactive clinical reasoning,

Reference 66

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:c094a560f7cc696e90117ee2f86c22a4a3198f4cc555b4e8e5c7406abe763514

Observation 1f6f3b94-b327-4c5b-8412-19444b39f0f2 · outbound

This paper cites Measuring Massive Multitask Language Understanding.

Enhancing LLM Metacognition via Cognitive Pairwise Training Measuring Massive Multitask Language Understanding

Reference 67

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.925386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:15dd98f302b22008bb21a62e077885c2abc3518d1a2bb7ac160a0bdf64599f65

Observation 05c61c68-1a22-45b7-ac26-9509b07f215d · outbound

This paper cites Evaluating the moral beliefs encoded in LLMs,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Evaluating the moral beliefs encoded in LLMs,

Reference 68

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:05efcec673b87ab104b679f376cd72c859b9fc29382e76d54bbbf145f4d4e97a

Observation 2f0d2a6b-01c7-424c-8bb2-cbfa76c76257 · outbound

This paper cites MuSiQue: Multihop questions via single-hop question composition,.

Enhancing LLM Metacognition via Cognitive Pairwise Training MuSiQue: Multihop questions via single-hop question composition,

Reference 69

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:98a4382dbb565cac194509f4b2f5468d702288515a1b64d701a52555a00d5239

Observation c4b97692-3725-4f11-adea-1ede892a549b · outbound

This paper cites (QA)2: Question answering with questionable assumptions,.

Enhancing LLM Metacognition via Cognitive Pairwise Training (QA)2: Question answering with questionable assumptions,

Reference 70

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:770197d272c25fcb2891d857685b82126217549b6fed86a27090e165f29eeb09

Observation b15a0044-e5cf-4b89-9065-12387670db2d · outbound

This paper cites A dataset of information-seeking questions and answers anchored in research papers,.

Enhancing LLM Metacognition via Cognitive Pairwise Training A dataset of information-seeking questions and answers anchored in research papers,

Reference 71

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:c3e06e6f253cd336762e46c241a4e175fb4d96459cc7285b850973d809ece4bb

Observation ff3f6037-ebbd-4de5-bb89-69c38c2c9f0c · outbound

This paper cites SituatedQA: Incorporating extra-linguistic contexts into QA,.

Enhancing LLM Metacognition via Cognitive Pairwise Training SituatedQA: Incorporating extra-linguistic contexts into QA,

Reference 72

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:ba81a86b215d385b20c9545c1824e5d8bca07163cae00280aba2692fb5b76744

Observation 020ab70b-345c-4e92-b2d2-d44c20e7eb48 · outbound

This paper cites Know what you don’t know: Unanswerable questions for SQuAD,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Know what you don’t know: Unanswerable questions for SQuAD,

Reference 73

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:f0a9f3829e7eff562520592fc72e81fe4612f1d636013b3b8c3dcea99f1abb16

Observation 9aca73f5-75ae-4ec9-b117-543e1dce8d0d · outbound

This paper cites Benchmarking Hallucination in Large Language Models based on Unanswerable Math Word Problem.

Enhancing LLM Metacognition via Cognitive Pairwise Training Benchmarking Hallucination in Large Language Models based on Unanswerable Math Word Problem

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:02:33.895497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:a0a256edc9a473a91dbfa1b5527bd93a346fe2e975122e4e8b7658da2cc00d87

Observation 87a0a974-f900-43b5-b4e8-7a7afb72d8e2 · outbound

This paper cites WorldSense: A Synthetic Benchmark for Grounded Reasoning in Large Language Models.

Enhancing LLM Metacognition via Cognitive Pairwise Training WorldSense: A Synthetic Benchmark for Grounded Reasoning in Large Language Models

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:02:33.886009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:fadd9aff56e00f9c73cab305df0c0a8cfd5b41f56ec9bfa0476d447855f91674

Observation abf0102e-ac5e-4a66-b00b-73e2c003fadc · outbound

This paper cites A coefficient of agreement for nominal scales,.

Enhancing LLM Metacognition via Cognitive Pairwise Training A coefficient of agreement for nominal scales,

Reference 76

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:4bba092d8484aee3e8c39ce96949c4e3472e8d5e008e6e2de0d8993a16ee1052

Observation fc117fe0-dd0b-499f-9a82-f756c76c995a · outbound

This paper cites The measurement of observer agreement for categorical data,.

Enhancing LLM Metacognition via Cognitive Pairwise Training The measurement of observer agreement for categorical data,

Reference 77

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:1703bbd6d3c41f17594fce2cd4a2d325b164e3ac5262a77711289ee4f666f5dc

Observation 73959039-848e-4cdc-99ba-858d46b7a98a · outbound

This paper cites Measuring nominal scale agreement among many raters,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Measuring nominal scale agreement among many raters,

Reference 78

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:013f02353885a78e096ac4dea3609cbed5be4d3bac3f906e8932554db642d2e5

Observation 05335121-b50c-4b1a-924d-b1738993691e · outbound

This paper cites Quagmires in sft-rl post-training: When high sft scores mislead and what to use instead.

Enhancing LLM Metacognition via Cognitive Pairwise Training Quagmires in sft-rl post-training: When high sft scores mislead and what to use instead

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:02:33.832707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:ea3b9d6453daec5b609d7cf19a5719a829e046868d719ef9612c00f1468252db

Observation f2d39fa9-1d82-46a6-a59a-91ebb8415b88 · outbound

This paper cites Beyond Two-Stage Training: Cooperative SFT and RL for LLM Reasoning.

Enhancing LLM Metacognition via Cognitive Pairwise Training Beyond Two-Stage Training: Cooperative SFT and RL for LLM Reasoning

Reference 80

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.906107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:5e39fabd9723b59514a2d68e28b912d22e23a6d0d8a839a2a3b795af5748142b

Observation 42dd68ee-d222-4ae8-a4eb-d2e076113b97 · outbound

This paper cites The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models.

Enhancing LLM Metacognition via Cognitive Pairwise Training The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models

Reference 81

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.878927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:c33397d5da2eea7c2034eb7231e9fb52edd4c7212ddfcecc5c9d705f89928b07

Observation a6ce62f4-2f0c-435d-9727-edc4d7977589 · outbound

This paper cites Enhancing LLM Reasoning with Iterative DPO: A Comprehensive Empirical Investigation.

Enhancing LLM Metacognition via Cognitive Pairwise Training Enhancing LLM Reasoning with Iterative DPO: A Comprehensive Empirical Investigation

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:02:33.820741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:4215219d2d902eda875982d117463b1c55a4d7eafcc6ae5e2a0e3ea794ea40b9

Observation d0349acd-d333-4845-8385-e9e4ec6bbcf1 · outbound

This paper cites Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs.

Enhancing LLM Metacognition via Cognitive Pairwise Training Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs

Reference 83

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.922491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:3376088af56052acedb4527094d59455a507a30bb36245c913d58284e841bbc7

Observation f52aa683-347a-41b1-b4a1-96e1d3268983 · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

Enhancing LLM Metacognition via Cognitive Pairwise Training Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 84

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.933275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:25ec7cf0c4d058b1e7a8af24169b71e942354ae77b608d1324b71b3a3b480e8d

Observation 9bd659e4-fe25-4360-b389-5b52ea14fdef · outbound

This paper cites Scaling Synthetic Data Creation with 1,000,000,000 Personas.

Enhancing LLM Metacognition via Cognitive Pairwise Training Scaling Synthetic Data Creation with 1,000,000,000 Personas

Reference 85

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.900775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:62524cfcd1500a867e08e7e4588aa05769e27ec8d0add56c3a29caed5bf9e0d0

Observation bef76b20-5982-4376-bab4-4beffeef4bfb · outbound

This paper cites The flan collection: Designing data and methods for effective instruction tuning,.

Enhancing LLM Metacognition via Cognitive Pairwise Training The flan collection: Designing data and methods for effective instruction tuning,

Reference 86

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:d19f6857f99bacc3fc9a52c46e5cffcfe7f994ed7c8b119c30937d2c292994d4

Observation b13a5645-3881-4eef-b975-4b12e6774f7e · outbound

This paper cites No robots.

Enhancing LLM Metacognition via Cognitive Pairwise Training No robots

Reference 87

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:10993d17619a412633ba9f5ad0e4c89e302dabb65ee3f0bfd9267c515f9d9062

Observation f8b8d6ec-418d-442c-8b39-433afeead9dc · outbound

This paper cites Openassistant conversations-democratizing large language model alignment,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Openassistant conversations-democratizing large language model alignment,

Reference 88

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:57d96e78a0a439a174e74af15ad7e5425562c37c6ffd2a356f79405c0f083a97

Observation b3887d8e-8486-4f29-8db4-743526df9e16 · outbound

This paper cites Aya dataset: An open-access collection for multilingual instruction tuning,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Aya dataset: An open-access collection for multilingual instruction tuning,

Reference 89

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:f8e3b9eb78549fc5b5f5a7661ccf36575b15de1528b5cd14b0e68b64f7e8a8f0

Observation 65a97c2b-5886-440c-9cd5-8062bf0075e1 · outbound

This paper cites TableGPT: Towards Unifying Tables, Nature Language and Commands into One GPT.

Enhancing LLM Metacognition via Cognitive Pairwise Training TableGPT: Towards Unifying Tables, Nature Language and Commands into One GPT

Reference 90

Resolution
verified exact
arxiv_id, observed 2026-06-28T19:02:33.892949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:f3a66049381ab9dccc728b92ed0ec3f7e342307a33c4ac816803aae1f86621c8

Observation 10f6f86e-b98a-4838-9fcc-e75f0f78d3f3 · outbound

This paper cites WildChat: 1M ChatGPT Interaction Logs in the Wild.

Enhancing LLM Metacognition via Cognitive Pairwise Training WildChat: 1M ChatGPT Interaction Logs in the Wild

Reference 91

Resolution
verified exact
local_arxiv, observed 2026-06-28T19:02:33.875220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:abb12673df9e0d1954b2f06046ae5e1085da39fe44d2763c83f31c50f52b8a94

Observation ec15ae00-cfa5-4f1a-a14a-e9bf671ad825 · outbound

This paper cites Wizardcoder: Empower- ing code large language models with evol-instruct,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Wizardcoder: Empower- ing code large language models with evol-instruct,

Reference 92

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:3a04b8bac9ba94abfa55f6fc7e4101f81ee736f102562497fd1fc17849f547ab

Observation dcbf4f63-cb42-4545-9a92-c9d2cde371c6 · outbound

This paper cites Wildguard: Open one-stop moderation tools for safety risks, jailbreaks, and refusals of llms,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Wildguard: Open one-stop moderation tools for safety risks, jailbreaks, and refusals of llms,

Reference 93

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:ed1440dfcbbaa94bfca6c5dcbce4895ba73c24e26759e9a5d83378521789dc6e

Observation 841cab1c-4f80-4c4e-8ac5-8a21afe1e32c · outbound

This paper cites Wildteaming at scale: From in-the-wild jailbreaks to (adversarially) safer language models,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Wildteaming at scale: From in-the-wild jailbreaks to (adversarially) safer language models,

Reference 94

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:f01ba613c8ea46abaae945720d65532c64536e35cf508f5700e80d12669b6a42

Observation 11a2b820-3a99-4a15-85ca-7566c57a6205 · outbound

This paper cites Sciriff: A resource to enhance language model instruction-following over scientific literature,.

Enhancing LLM Metacognition via Cognitive Pairwise Training Sciriff: A resource to enhance language model instruction-following over scientific literature,

Reference 95

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:2b8d3cee41c3e662ae1092bc48aff41cdacbd7eb888059d67a8911b27ee11d5b

Observation 5a889d61-35bc-4534-94a6-cd68941ec8b0 · outbound

This paper cites an unresolved cited work.

Enhancing LLM Metacognition via Cognitive Pairwise Training Unresolved cited work

Reference 96

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:c6795926737a5e9ffabfe6bbdba28d16a000c28b5a5899a907757366bb5a8dac

Observation 550ad1d4-ac42-4ac7-8583-59784a9471c9 · outbound

This paper cites an unresolved cited work.

Enhancing LLM Metacognition via Cognitive Pairwise Training Unresolved cited work

Reference 97

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:2c43702c979f5ae0d67afd736742031265c40d68f4ba3f4df79b781a63092051

Observation 9c923e4e-db04-4b16-9032-a24d3db713bf · outbound

This paper cites I don’t know.

Enhancing LLM Metacognition via Cognitive Pairwise Training I don’t know

Reference 98

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:e12cf23ccdced184d9ae52fd15e0a84812a3bbb298152262feed7437cdb35ed3

Observation 7df0871e-ed7c-4dce-bdd4-6b37b4390253 · outbound

This paper cites For base checkpoints, the dataset loader extracts the raw prompt text from the stored chat-style field and tokenizes it directly, instead of applying a chat template.

Enhancing LLM Metacognition via Cognitive Pairwise Training For base checkpoints, the dataset loader extracts the raw prompt text from the stored chat-style field and tokenizes it directly, instead of applying a chat template

Reference 99

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:5207849037092ffdd23c35db734b9842565d795cd407f43b6b2e167d23f61035

Observation 8f1a6f02-4963-4586-8152-bfcfc5519889 · outbound

This paper cites For each promptx, the policy samples a group ofGresponses{y i}G i=1 from vLLM [52]; in the main Math-RL scriptG= 16, temperature is 0.9, andtop-p=0.95.

Enhancing LLM Metacognition via Cognitive Pairwise Training For each promptx, the policy samples a group ofGresponses{y i}G i=1 from vLLM [52]; in the main Math-RL scriptG= 16, temperature is 0.9, andtop-p=0.95

Reference 100

Resolution
unresolved
no resolver link, observed 2026-06-28T19:01:18.153145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T19:01:18.153145Z digest=sha256:38e67e67e38cc6fad2c1c5ce3d68b641aead81e9235416aa162fe5cda3b4eff2

Pith citing papers

No inbound Pith citation observations are available.