Pith. sign in

Paper Citation Record · LEDGER

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models

As of 7 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 1 inbound Pith citation observation for arXiv:2506.12217.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.12217 v1

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T01:03:43.999176Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T00:55:32.686738Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

60 of 60 outbound references displayed

  • verified exact0
  • verified fuzzy29
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1629fedd-c51a-443c-8c86-a9e44716013d · outbound

This paper cites Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:38.561365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:38.561365Z digest=sha256:7c733425515b8f8f62fa92def8adc10f4d6aa19cf898544c867063f18e05ba4c

Observation c3f47b45-1cf7-4040-b36c-88f44fb55353 · outbound

This paper cites Self-reasoning language models: Unfold hidden reasoning chains with few reasoning catalyst,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Self-reasoning language models: Unfold hidden reasoning chains with few reasoning catalyst,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:50.091457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:38.644078Z digest=sha256:b8bdec333e1a0fe76587de4921a17863114d4fa8c3761c04bbc6c3d7aea0987e

Observation d39c0232-4edf-4f1f-93f0-77c3ec0e014d · outbound

This paper cites Reinforcement learning with verifiable rewards: Grpo’s effective loss, dynamics, and success amplification,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Reinforcement learning with verifiable rewards: Grpo’s effective loss, dynamics, and success amplification,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:38.746258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:38.746258Z digest=sha256:264d01d3bd6c190e9147ad906880b35eadca5b450cd25b898ca15c29ae38819e

Observation 72ccaf44-fc65-488a-a1cd-10b9b7ca1794 · outbound

This paper cites R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:38.824120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:38.824120Z digest=sha256:41114ab10373f5a7ab4acc60ca8008d88fac64973bbc9087fcb5da52a46bdabb

Observation 95f02528-e180-46f1-b36e-a332307da8d4 · outbound

This paper cites Reasoning beyond limits: Advances and open problems for llms,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Reasoning beyond limits: Advances and open problems for llms,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:38.946446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:38.946446Z digest=sha256:9364288a604396fa3e44b265bd9da010c101b8a01f37e94e6e967c67e58cb1c8

Observation 45d137f3-ca87-4e2f-b01a-e084297c6e34 · outbound

This paper cites Crossing the Reward Bridge: Expanding RL with Verifiable Rewards Across Diverse Domains.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Crossing the Reward Bridge: Expanding RL with Verifiable Rewards Across Diverse Domains

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.063087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.063087Z digest=sha256:6588740007af0e87d68f5a2cdf19fc1c20a9c794eea12fa83e69cbc8f045838c

Observation 6ce3f4ec-e201-4c21-8cfa-de0702166cd9 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.150691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.150691Z digest=sha256:3055f84761c6be08a8795ea1f9b75f5b5485b0d85aaa906a67ae95b468ce981e

Observation 00266b23-2c47-4af4-a0f7-ac7c0164fc40 · outbound

This paper cites Understanding R1-Zero-Like Training: A Critical Perspective.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Understanding R1-Zero-Like Training: A Critical Perspective

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.248657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.248657Z digest=sha256:2701677162715919e4593349812f7be6a322a757e9f97f665fee65cc0111c1f2

Observation fec69ab1-664a-4a23-b22f-5e9923cbb2d1 · outbound

This paper cites SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.348782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.348782Z digest=sha256:2aca6ec50b65e2a0a4a26100505ab6dfcc960b11dc875d4d4efdcf118e8d06f7

Observation 39da37bb-30e3-44ac-a74b-d2899675fa2e · outbound

This paper cites TTRL: Test-Time Reinforcement Learning.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models TTRL: Test-Time Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.439450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.439450Z digest=sha256:076b7332f1058c580e9677f71fcf96b023e11de69374624ec5ccccec3adca523

Observation 6e48c6ec-e596-4e7a-a061-ee7584b866ec · outbound

This paper cites Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.555965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.555965Z digest=sha256:50492fca5874d3d758c1e95aeae5c6178b158c48bce12f315f3f6df5510d96f7

Observation 5012415e-37d7-479a-b840-2af01d93001e · outbound

This paper cites Self-Reflection Makes Large Language Models Safer, Less Biased, and Ideologically Neutral.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Self-Reflection Makes Large Language Models Safer, Less Biased, and Ideologically Neutral

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.646395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.646395Z digest=sha256:43bfc31dbefb7061623481920975fd7298d3fd9d85277e65fc5ce17851772ebf

Observation 3a583017-4b0b-4ce9-a51f-739836db59d9 · outbound

This paper cites Dynamic early exit in reasoning models,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Dynamic early exit in reasoning models,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.763786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.763786Z digest=sha256:0931424920df86b645fb434b69486da0bac81f136b12022e5b9f65c94f7c258f

Observation 0ff7c1eb-0746-4af3-a970-991237b9ad2b · outbound

This paper cites Self-Reflection in LLM Agents: Effects on Problem-Solving Performance.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Self-Reflection in LLM Agents: Effects on Problem-Solving Performance

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.875381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.875381Z digest=sha256:7beeb32a87650d8c829002ee6674e3288ea325fb00cf60ea948db17e10bb75ef

Observation e75301d6-8fc0-4c2b-8fd8-258fd41755a5 · outbound

This paper cites Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.949009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.949009Z digest=sha256:48f0f976e0eaf0d551ebf98e26059d9f50fc54726670e9838a4a902a98c270f5

Observation 1c3c4d59-7537-49e7-87a4-48675c718d1f · outbound

This paper cites Steering llama 2 via contrastive activation addition,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Steering llama 2 via contrastive activation addition,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:49.915796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:40.075252Z digest=sha256:75451592405c9162e85d49f157ff70f3d663153783b73c13a1de55dce7a23279

Observation dfcd3ad3-4410-464d-bd53-7cafe43242bd · outbound

This paper cites Generating Wikipedia by Summarizing Long Sequences.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Generating Wikipedia by Summarizing Long Sequences

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:40.150110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:40.150110Z digest=sha256:dd1005fa56562fde80c71655cef173208ae0cbd5d0d3a59156ac1be9d62a6cc2

Observation d7c5a580-fc9f-4a8e-afd2-9006c9ee36f6 · outbound

This paper cites Beyond accuracy: Evaluating the reasoning behavior of large language models - a survey,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Beyond accuracy: Evaluating the reasoning behavior of large language models - a survey,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:49.713626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:40.258094Z digest=sha256:2581c37467893a7555d3bf6c5696b467185a6f5d9a610c0305c20371989a43f0

Observation 6edd9d22-89a5-4dc0-9c6f-af10c5a203a7 · outbound

This paper cites Oat: A research-friendly framework for llm online alignment,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Oat: A research-friendly framework for llm online alignment,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:49.490397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:40.356585Z digest=sha256:b04a321ad0d3223c400e1b00c85b49677440ee4f7c84e52fc5c0651825f909b3

Observation aa276d97-fc20-4b0b-9055-911fb8bdbae9 · outbound

This paper cites Self- consistency improves chain of thought reasoning in language models,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Self- consistency improves chain of thought reasoning in language models,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:49.324127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:40.427956Z digest=sha256:d29eb1bd602548ace585c820ad685d3bc16b046f98187222fb0a1a6046e76f20

Observation f3ca3234-ffce-4943-a683-0ebd222aaa98 · outbound

This paper cites X-reasoner: Towards generalizable reasoning across modalities and domains,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models X-reasoner: Towards generalizable reasoning across modalities and domains,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:49.145880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:40.536878Z digest=sha256:c8b85361ce7f779b8d93d08ff328d84c800350d85b347f780d242dba55fbce0d

Observation 271c3ef1-143a-4feb-a781-27cf896f307c · outbound

This paper cites Reinforcement Learning Enhanced LLMs: A Survey.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Reinforcement Learning Enhanced LLMs: A Survey

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:40.631750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:40.631750Z digest=sha256:43f503812ebdf377e7c2158ed96ed3230ff6fb2b76e66da91c642860a9d0ff18

Observation a66ec951-743b-4545-99c4-8e4e8a841cbf · outbound

This paper cites Absolute Zero: Reinforced Self-play Reasoning with Zero Data.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Absolute Zero: Reinforced Self-play Reasoning with Zero Data

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:40.693792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:40.693792Z digest=sha256:a5011cf83f439e8eb8b38cff745c7c3560c4688b89c9f4e8a0b518431377b065

Observation f9f122bc-638e-4812-94d2-a741f75f391c · outbound

This paper cites When Hindsight is Not 20/20: Testing Limits on Reflective Thinking in Large Language Models.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models When Hindsight is Not 20/20: Testing Limits on Reflective Thinking in Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:40.835799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:40.835799Z digest=sha256:6cef46b0dbfda58342937c9b651041be2034bb646fb540d3cda85c43c257e3f7

Observation d7ed41c2-594d-4bc7-8560-4063ddfa933d · outbound

This paper cites Demystifyinglongchain-of-thoughtreasoninginLLMs,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Demystifyinglongchain-of-thoughtreasoninginLLMs,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:48.981885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:40.918114Z digest=sha256:544b5d8ee10fac769a64fd82fc19a01a5ba44c4d72efce847ec6b47cacc1dff4

Observation 9646293b-ee21-4f86-9fc0-d89a56d2de0f · outbound

This paper cites OpenAI o1 System Card.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models OpenAI o1 System Card

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:41.000635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:41.000635Z digest=sha256:f75d6d281d0bba02203a3afd4d094101fe238283b98065d4d964bde83a6428a1

Observation 469303ad-d1e6-4663-a969-09b9a3aad043 · outbound

This paper cites 2 OLMo 2 Furious.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models 2 OLMo 2 Furious

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:41.101622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:41.101622Z digest=sha256:d30a40935d98e2dae03bc516bf3721d64080daadf08a09ef0ee118c36a85d178

Observation a6bebd1d-2878-441b-b47a-317a95e26e40 · outbound

This paper cites Qwen3technical report,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Qwen3technical report,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:48.803809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:41.187541Z digest=sha256:b1fd107a5529d7cc323ee321871ddb852a5f6c1bef3bf44ecaa9885e9727c420

Observation 7c38a833-c42f-4bdf-b9e3-6d709d13d656 · outbound

This paper cites Measuring mathematical problem solving with the math dataset,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Measuring mathematical problem solving with the math dataset,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:48.604954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:41.281529Z digest=sha256:6c4bdce81b81e22d6cef62a9651e7d6a969238f6ec4754e7684184c56615ed54

Observation c3b0ce6b-1dcd-4094-a847-9ac37c243daf · outbound

This paper cites Qwen2.5 Technical Report.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Qwen2.5 Technical Report

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:41.367515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:41.367515Z digest=sha256:98d297500dd2d8165d2f744460e6e876c5c6e360e82e94f3f73afc7bd237041b

Observation 5f2a934e-8188-4b61-b082-cfc0a740d4f2 · outbound

This paper cites Umap: Uniform manifold approximation and projection,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Umap: Uniform manifold approximation and projection,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:48.402203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:41.455519Z digest=sha256:e2a6839323da335539d872e90996085cb6ce339397607312ac0b51dbb9d24a58

Observation c48b7e33-6880-4a41-9bcc-4e2415ae0261 · outbound

This paper cites Refusal in language models is mediated by a single direction,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Refusal in language models is mediated by a single direction,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:48.223985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:41.581041Z digest=sha256:0cfefd34e577a8281f81ec282f3555fffb90df3daa565e57b7b964fd4270224d

Observation aa3cfd68-4461-42d7-9c7b-067819b1df5e · outbound

This paper cites AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:41.633267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:41.633267Z digest=sha256:f326046794f3f6b7ee1b86eaab78ae0ef5ef630faae68a9dc1c65f93d8df5ab4

Observation 574728a9-8514-40e5-808d-9ebd8bffb441 · outbound

This paper cites GPQA: A graduate-level google-proof q&a benchmark,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models GPQA: A graduate-level google-proof q&a benchmark,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:48.007208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:41.688003Z digest=sha256:c8ddbf7b4a43ceda47c73b0f8b964a09ac2135894a476c34838206065f188539

Observation 663e3187-2685-436c-8ec6-a625da5ab9ca · outbound

This paper cites The Llama 3 Herd of Models.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models The Llama 3 Herd of Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:41.772289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:41.772289Z digest=sha256:45b12cc1aa5c5663d311f4f7ae59584525262a63914bc132c9c4057b10a92735

Observation 176a923f-3bc6-4556-bb92-d9d3e3b7a73e · outbound

This paper cites s1: Simple test-time scaling,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models s1: Simple test-time scaling,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:47.843540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:41.883084Z digest=sha256:5936c63da2d651be05c599447352406ff5cc75aee6e519376957bab2fb9055cb

Observation 455bfbd3-f1f3-4569-9fb0-13256cdc8c65 · outbound

This paper cites Discovering latent knowledge in language models without supervision,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Discovering latent knowledge in language models without supervision,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:47.687530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:41.999426Z digest=sha256:8f4e0b245071232b7cab317e7add80f408a213420db30e6b5892cbca0351dfcf

Observation e080e39d-6fdc-46bd-b468-be482af58d7d · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:42.097166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:42.097166Z digest=sha256:16c13b67726339afb2188babd409723d1bb87647f295a7af956411a408b71078

Observation 066fae06-1e46-4c3e-943f-2ef5c7231f6c · outbound

This paper cites A Language Model's Guide Through Latent Space.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models A Language Model's Guide Through Latent Space

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:42.165517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:42.165517Z digest=sha256:bb4ab8b781cee69d0a11606b657a9e9b458577699ac1efc191978d0c652c6f81

Observation ccda8ba5-f1c8-4351-b338-06c32a8b5839 · outbound

This paper cites Improving Activation Steering in Language Models with Mean-Centring.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Improving Activation Steering in Language Models with Mean-Centring

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:42.231617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:42.231617Z digest=sha256:d426b2450b0da29dc3f68447632107c63a460ddfb6623d871bdbd92519d5a32b

Observation 4b8fa7cc-b175-46cf-a4bc-2e07327b259b · outbound

This paper cites Finding alignments between interpretable causal variables and distributed neural representations,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Finding alignments between interpretable causal variables and distributed neural representations,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:47.540585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:42.293333Z digest=sha256:2a610bd651f3cd5be3a68b8326c71cc3d03cde5522f7561a0cc1b2682c67aa48

Observation 75b62a26-f9ce-45a9-bf6d-25a79750aa6c · outbound

This paper cites Generative agents: Interactivesimulacraofhumanbehavior,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Generative agents: Interactivesimulacraofhumanbehavior,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:47.341340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:42.349684Z digest=sha256:9958e81e39b3a9a54f6c1f264c3fa4be37da00e41ed08fb387e64e053d3e005f

Observation 2e2ee699-326c-43f4-9ce6-e929d40d1b2a · outbound

This paper cites Man is to computer programmer as woman is to homemaker? debiasing word embeddings,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Man is to computer programmer as woman is to homemaker? debiasing word embeddings,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:47.140600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:42.442869Z digest=sha256:417442d792df33890e95dbcbf216d23bd531dbebb9ace23a853207147145a662

Observation 5646bbed-0a47-48ee-92c0-a8de0ca8edb5 · outbound

This paper cites Sparseautoencodersfindhighlyinter- pretable features in language models,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Sparseautoencodersfindhighlyinter- pretable features in language models,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:46.980720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:42.512310Z digest=sha256:0d896e94701500788fb7adef6985fab8ef8e8a233ea0453b6954ba2af4ae6a7b

Observation f7461889-ce63-41fe-89a9-fbd38f2db7d1 · outbound

This paper cites Gold doesn‘t always glitter: Spectral removal of linear and nonlinear guardedattributeinformation,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Gold doesn‘t always glitter: Spectral removal of linear and nonlinear guardedattributeinformation,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:46.832967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:42.632456Z digest=sha256:d632e057c4cf4b9a3d346e28558a72304741fc7005a922b0fc9d6e6b59f6be93

Observation ee86d914-0601-459b-b18b-37f5071d3aff · outbound

This paper cites LEACE: Perfect linear concept erasure in closed form,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models LEACE: Perfect linear concept erasure in closed form,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:46.659998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:42.748817Z digest=sha256:d46c250070d421e38eaa68f24cb0bc8713ab7f0ae21f4588772a8d1f769a5a9f

Observation 4876e0ea-6e93-4259-9ab9-3a14a1faeb66 · outbound

This paper cites Monitoring latent world states in language models with proposi- tional probes,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Monitoring latent world states in language models with proposi- tional probes,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:46.506000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:42.820828Z digest=sha256:e6cb139541cc9e6939f1bc7dcb63fe902e929cc8d17268b5cd7e6f46fbeacede

Observation 14c8a776-3a4d-498f-8938-059650e3aa86 · outbound

This paper cites Let’s verify step by step,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Let’s verify step by step,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:46.336643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:42.869100Z digest=sha256:c686f5a6c4de876eda182314c558c05bbd6e849b70466b4ae8b5fccfe44cd9d7

Observation 45964bdf-c1b5-4127-be86-22247b70986f · outbound

This paper cites Self-refine: Iterative refinement with self-feedback,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Self-refine: Iterative refinement with self-feedback,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:46.194754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:42.995770Z digest=sha256:5dd2a6650b45815ac7159d11e67f7c6039eb1632e52df2ec0f374b2bd2995750

Observation 4f77234f-d40c-4b67-ac9a-698fee7368e6 · outbound

This paper cites Fine-tuning with divergent chains of thought boosts reasoning through self-correction in language models,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Fine-tuning with divergent chains of thought boosts reasoning through self-correction in language models,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:46.041393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:43.095901Z digest=sha256:0998a0625d1967754b76984d3bd9f9786538a43017d8374973035c623abc45ae

Observation b0fac060-c3ff-45bf-be9d-73d9cfe73d47 · outbound

This paper cites STar: Bootstrapping reasoning with reasoning,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models STar: Bootstrapping reasoning with reasoning,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:45.858720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:43.195392Z digest=sha256:7ceeb9b0fba6fee841d0839c6780acfcb971d71ba2d31bd3a916e9b8ea90824d

Observation 2912de11-1a0d-4b74-84d1-31f3dffcfaae · outbound

This paper cites Let’s verify step by step,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Let’s verify step by step,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:45.699379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:43.254472Z digest=sha256:0470e467f4161e23cc8d94518fc95a3e799c9ea82db20c5041ac362aaee0daed

Observation b5856e66-e3bf-4112-b2c0-a838138b03f0 · outbound

This paper cites LLMs Can Easily Learn to Reason from Demonstrations Structure, not content, is what matters!.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models LLMs Can Easily Learn to Reason from Demonstrations Structure, not content, is what matters!

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:43.321333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:43.321333Z digest=sha256:b6b55ce3c3ee351acfe500bc506de121f15073ad780282ca9760f551c6875826

Observation 1af91da7-bec5-4818-8733-323b5f4b5934 · outbound

This paper cites Shorterbetter: Guidingreasoningmodelstofindoptimalinferencelengthforefficient reasoning,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Shorterbetter: Guidingreasoningmodelstofindoptimalinferencelengthforefficient reasoning,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:43.439679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:43.439679Z digest=sha256:d50613ed7bfec4f646cca380fab92a009f63c70adccc23a21a037502a6c94b46

Observation ee2f8254-4724-4d15-b0f4-e5a60cf4666e · outbound

This paper cites Unlocking the capabilities of thought: A reasoning boundary framework to quantify and optimize chain-of-thought,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Unlocking the capabilities of thought: A reasoning boundary framework to quantify and optimize chain-of-thought,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:45.472186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:43.510697Z digest=sha256:cbaa05962028324e16287909c83f142d331d6ffce18ae6b4dd29e15de59399fa

Observation 1e93143a-2574-4054-be9a-79baf7e88ac1 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:43.598058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:43.598058Z digest=sha256:84a77bf700ef5b8746ec32c639dce34c8664834a2f5afa5d6393dbc4a7876e3f

Observation c86d77b5-78b6-4952-baaf-a77e3547715d · outbound

This paper cites Rethinking Reflection in Pre-Training.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Rethinking Reflection in Pre-Training

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:43.678277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:43.678277Z digest=sha256:03dd76c1125c17e6a44f14664f202f5ede211fef3edfe38679adadcd06c35ef7

Observation ab5573a1-46ee-49f2-b945-5d27a0a59f08 · outbound

This paper cites Reflexion: Languageagentswithverbal reinforcement learning,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Reflexion: Languageagentswithverbal reinforcement learning,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:45.213263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:43.774554Z digest=sha256:adfd99a7ce3a543c777aedd04221ddf0bd9cb5b7dda3857ca64b1b4b021340e7

Observation 3fa9b162-1ea8-45d9-b1a8-faa954e327d7 · outbound

This paper cites Reinforcement Learning for Reasoning in Large Language Models with One Training Example.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Reinforcement Learning for Reasoning in Large Language Models with One Training Example

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:43.838361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:43.838361Z digest=sha256:4a33a70ec9573ebf09ee56dd011b275e18c0c044e29415da7fd7f7d61c8cff07

Observation 33193c25-3851-45af-84ae-f7a83be9ac32 · outbound

This paper cites There may not be aha moment in r1-zero-like training — a pilot study.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models There may not be aha moment in r1-zero-like training — a pilot study

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:45.036245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T01:03:43.999176Z digest=sha256:662d071debbbfdfeaf0a279b09e596fb5fb2fde3d7def5d9ac5ac43b9669e9ef

Pith citing papers

Observation 09e6d2ee-3503-4e81-8bf6-3281c83c64a6 · inbound

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges cites this paper.

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges From Emergence to Control: Probing and Modulating Self-Reflection in Language Models

Reference 201

Resolution
unresolved
no resolver link, observed 2026-08-03T00:55:32.686738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:55:32.686738Z digest=sha256:1b6fcd143e8c2be24b427773f3912851e39c9d695e1e7ac5b4d17943349b48eb