Pith. sign in

Paper Citation Record · LEDGER

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models

As of 9 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 1 inbound Pith citation observation for arXiv:2506.12217.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.12217 v1

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T01:03:43.999176Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T00:55:32.686738Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

60 of 60 outbound references displayed

  • verified exact0
  • verified fuzzy29
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1629fedd-c51a-443c-8c86-a9e44716013d · outbound

This paper cites Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:38.561365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:38.561365Z digest=sha256:c61d5b30cb972df2af3a4319a16b1cfe7bc33593540ddf8da22da4cd5768a208

Observation c3f47b45-1cf7-4040-b36c-88f44fb55353 · outbound

This paper cites Self-reasoning language models: Unfold hidden reasoning chains with few reasoning catalyst,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Self-reasoning language models: Unfold hidden reasoning chains with few reasoning catalyst,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:50.091457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:38.644078Z digest=sha256:67b51655698fc59e5612f87de0d483c87754b4f4a4f4bb65732069546ddc9766

Observation d39c0232-4edf-4f1f-93f0-77c3ec0e014d · outbound

This paper cites Reinforcement learning with verifiable rewards: Grpo’s effective loss, dynamics, and success amplification,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Reinforcement learning with verifiable rewards: Grpo’s effective loss, dynamics, and success amplification,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:38.746258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:38.746258Z digest=sha256:a743f33981ed75472f057a326f47715564aab1d82a78f34bc71de94bb036aae8

Observation 72ccaf44-fc65-488a-a1cd-10b9b7ca1794 · outbound

This paper cites R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:38.824120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:38.824120Z digest=sha256:2b4827c3d39d8d3fa0fb448cb613d5477ddfde5be251f1e36674238468faff3f

Observation 95f02528-e180-46f1-b36e-a332307da8d4 · outbound

This paper cites Reasoning beyond limits: Advances and open problems for llms,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Reasoning beyond limits: Advances and open problems for llms,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:38.946446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:38.946446Z digest=sha256:89b5867c0b2238075fc24c64b300e9618d0994c54e6dc75a6ff0151ab35e2202

Observation 45d137f3-ca87-4e2f-b01a-e084297c6e34 · outbound

This paper cites Crossing the Reward Bridge: Expanding RL with Verifiable Rewards Across Diverse Domains.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Crossing the Reward Bridge: Expanding RL with Verifiable Rewards Across Diverse Domains

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.063087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.063087Z digest=sha256:7b606f40612296e38f9522e4390355c3c29ff2cf760e1001120f35274582c871

Observation 6ce3f4ec-e201-4c21-8cfa-de0702166cd9 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.150691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.150691Z digest=sha256:0465ceba4ffa6461afd317287784e4877db98b4e34f60924b7a248c7e5a9dbfe

Observation 00266b23-2c47-4af4-a0f7-ac7c0164fc40 · outbound

This paper cites Understanding R1-Zero-Like Training: A Critical Perspective.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Understanding R1-Zero-Like Training: A Critical Perspective

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.248657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.248657Z digest=sha256:0d202093aa0bf975611a9a37b13c607ea9750adb7067d5c2bd6c37258f2d204c

Observation fec69ab1-664a-4a23-b22f-5e9923cbb2d1 · outbound

This paper cites SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.348782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.348782Z digest=sha256:52444f156089f9579be9de043f3c5876abae145a8c65b0dc9c38a9951a75b6a1

Observation 39da37bb-30e3-44ac-a74b-d2899675fa2e · outbound

This paper cites TTRL: Test-Time Reinforcement Learning.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models TTRL: Test-Time Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.439450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.439450Z digest=sha256:4b81567f098b785034964c8435e5877108efb95c5e269cc99c15a0f0155d9de5

Observation 6e48c6ec-e596-4e7a-a061-ee7584b866ec · outbound

This paper cites Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.555965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.555965Z digest=sha256:a13d051b282a9b3c6f58a7f8a954bdf6aa7664b2168896b0f38ab8262bb501ac

Observation 5012415e-37d7-479a-b840-2af01d93001e · outbound

This paper cites Self-Reflection Makes Large Language Models Safer, Less Biased, and Ideologically Neutral.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Self-Reflection Makes Large Language Models Safer, Less Biased, and Ideologically Neutral

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.646395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.646395Z digest=sha256:f8456a72b1ec6e0121f7b9975d063b0956fae03b3310da7654df7674dfe2969c

Observation 3a583017-4b0b-4ce9-a51f-739836db59d9 · outbound

This paper cites Dynamic early exit in reasoning models,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Dynamic early exit in reasoning models,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.763786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.763786Z digest=sha256:a628a2e133343269970e410c6a38ab790f2026faf5618542eb1c4ee2af421970

Observation 0ff7c1eb-0746-4af3-a970-991237b9ad2b · outbound

This paper cites Self-Reflection in LLM Agents: Effects on Problem-Solving Performance.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Self-Reflection in LLM Agents: Effects on Problem-Solving Performance

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.875381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.875381Z digest=sha256:b6d14f714da150d27717d40bf33d4b83a5374f4721f4c538aaefe94ed8230d29

Observation e75301d6-8fc0-4c2b-8fd8-258fd41755a5 · outbound

This paper cites Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:39.949009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:39.949009Z digest=sha256:7e18b8bdab9e74a82455aff6808552f6259b4e3c61884e86ca6dac57a2f53e26

Observation 1c3c4d59-7537-49e7-87a4-48675c718d1f · outbound

This paper cites Steering llama 2 via contrastive activation addition,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Steering llama 2 via contrastive activation addition,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:49.915796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:40.075252Z digest=sha256:8b17e48e5740c99600953036ef5d0188d5de6be2ff5eb48351e2a8db7d6f68b1

Observation dfcd3ad3-4410-464d-bd53-7cafe43242bd · outbound

This paper cites Generating Wikipedia by Summarizing Long Sequences.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Generating Wikipedia by Summarizing Long Sequences

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:40.150110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:40.150110Z digest=sha256:587275ec765bbf16e3eb501f59915c3ac973d5ac938507f9a7e6bc0d56091a1c

Observation d7c5a580-fc9f-4a8e-afd2-9006c9ee36f6 · outbound

This paper cites Beyond accuracy: Evaluating the reasoning behavior of large language models - a survey,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Beyond accuracy: Evaluating the reasoning behavior of large language models - a survey,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:49.713626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:40.258094Z digest=sha256:71b2b2e05cc3b3a2d8c75712f189f0da8b0d72155fa5f8fdbb98dc5baf74a071

Observation 6edd9d22-89a5-4dc0-9c6f-af10c5a203a7 · outbound

This paper cites Oat: A research-friendly framework for llm online alignment,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Oat: A research-friendly framework for llm online alignment,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:49.490397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:40.356585Z digest=sha256:896db69162611b2f33170173c272a4f218a2f6c7689f46e55903e27205861ea1

Observation aa276d97-fc20-4b0b-9055-911fb8bdbae9 · outbound

This paper cites Self- consistency improves chain of thought reasoning in language models,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Self- consistency improves chain of thought reasoning in language models,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:49.324127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:40.427956Z digest=sha256:5554a0653e7f87c4e3ac7ab9c229c73601177af1cc4884505af43e06e734ce9a

Observation f3ca3234-ffce-4943-a683-0ebd222aaa98 · outbound

This paper cites X-reasoner: Towards generalizable reasoning across modalities and domains,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models X-reasoner: Towards generalizable reasoning across modalities and domains,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:49.145880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:40.536878Z digest=sha256:d1299868b893b21242d172d71dc74362711128e41820428e4162fff1cce9a44e

Observation 271c3ef1-143a-4feb-a781-27cf896f307c · outbound

This paper cites Reinforcement Learning Enhanced LLMs: A Survey.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Reinforcement Learning Enhanced LLMs: A Survey

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:40.631750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:40.631750Z digest=sha256:d63f5190d8a29a2faaa84fcc93313dbe70bcabc13bd692543ca8b3fcface5942

Observation a66ec951-743b-4545-99c4-8e4e8a841cbf · outbound

This paper cites Absolute Zero: Reinforced Self-play Reasoning with Zero Data.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Absolute Zero: Reinforced Self-play Reasoning with Zero Data

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:40.693792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:40.693792Z digest=sha256:03296c6e5a238e02c4f6628c2ba243631156d23a104eba2936bbddf2955cd5f4

Observation f9f122bc-638e-4812-94d2-a741f75f391c · outbound

This paper cites When Hindsight is Not 20/20: Testing Limits on Reflective Thinking in Large Language Models.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models When Hindsight is Not 20/20: Testing Limits on Reflective Thinking in Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:40.835799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:40.835799Z digest=sha256:560f8506031e3f646eba23c8535ec0ad5b0274e5ea8d95fe8a86adda2c96d883

Observation d7ed41c2-594d-4bc7-8560-4063ddfa933d · outbound

This paper cites Demystifyinglongchain-of-thoughtreasoninginLLMs,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Demystifyinglongchain-of-thoughtreasoninginLLMs,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:48.981885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:40.918114Z digest=sha256:da42a91a7a059ada09c9dd1136c64f788468272349bf19011661e8d9054f4464

Observation 9646293b-ee21-4f86-9fc0-d89a56d2de0f · outbound

This paper cites OpenAI o1 System Card.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models OpenAI o1 System Card

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:41.000635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:41.000635Z digest=sha256:8804080f5380abbe67f6da3054d64dce14d7ce6bba3cadb2d945ee2b662f5dfe

Observation 469303ad-d1e6-4663-a969-09b9a3aad043 · outbound

This paper cites 2 OLMo 2 Furious.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models 2 OLMo 2 Furious

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:41.101622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:41.101622Z digest=sha256:1d4c6c769fbbbc985cdece84565ca77d2b59214102a57d035bdf1e90a6ac4c4c

Observation a6bebd1d-2878-441b-b47a-317a95e26e40 · outbound

This paper cites Qwen3technical report,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Qwen3technical report,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:48.803809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:41.187541Z digest=sha256:838d984ae0ffc874e74c2b75415341fd68858d33f7a8345c02c5eaddc00f958b

Observation 7c38a833-c42f-4bdf-b9e3-6d709d13d656 · outbound

This paper cites Measuring mathematical problem solving with the math dataset,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Measuring mathematical problem solving with the math dataset,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:48.604954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:41.281529Z digest=sha256:cd6581178d37705389414d9e99da48d86057640c4c68ad24c7c3b4c4fdcf9561

Observation c3b0ce6b-1dcd-4094-a847-9ac37c243daf · outbound

This paper cites Qwen2.5 Technical Report.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Qwen2.5 Technical Report

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:41.367515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:41.367515Z digest=sha256:0c31ee23972e3251854c5c8b3f49800aab413491876d516e59936e5567bec06e

Observation 5f2a934e-8188-4b61-b082-cfc0a740d4f2 · outbound

This paper cites Umap: Uniform manifold approximation and projection,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Umap: Uniform manifold approximation and projection,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:48.402203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:41.455519Z digest=sha256:5317d24ce2e11c06079800309e8f23a6b567f56bac4936804d537eb266264702

Observation c48b7e33-6880-4a41-9bcc-4e2415ae0261 · outbound

This paper cites Refusal in language models is mediated by a single direction,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Refusal in language models is mediated by a single direction,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:48.223985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:41.581041Z digest=sha256:8391bccd1902b2daa3cd85a97b21fecbd32109f7ff5497950870c0238b4ebb08

Observation aa3cfd68-4461-42d7-9c7b-067819b1df5e · outbound

This paper cites AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:41.633267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:41.633267Z digest=sha256:55f236c15193af44dd3994695b26c1d147b1449559399a941f13aed591e4d578

Observation 574728a9-8514-40e5-808d-9ebd8bffb441 · outbound

This paper cites GPQA: A graduate-level google-proof q&a benchmark,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models GPQA: A graduate-level google-proof q&a benchmark,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:48.007208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:41.688003Z digest=sha256:66cc11e498aecb221b4679ca32bc89705574097750f1ddb73d78bc16bbdabd70

Observation 663e3187-2685-436c-8ec6-a625da5ab9ca · outbound

This paper cites The Llama 3 Herd of Models.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models The Llama 3 Herd of Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:41.772289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:41.772289Z digest=sha256:45b7b7954c3ca376738641b6127a85e35cac081b917e0ce94f4dbb64e55c9250

Observation 176a923f-3bc6-4556-bb92-d9d3e3b7a73e · outbound

This paper cites s1: Simple test-time scaling,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models s1: Simple test-time scaling,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:47.843540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:41.883084Z digest=sha256:2b00d0163242d60ccf4ed58d584a366e1c4f5525e4f75e22302644fd4f07ac8f

Observation 455bfbd3-f1f3-4569-9fb0-13256cdc8c65 · outbound

This paper cites Discovering latent knowledge in language models without supervision,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Discovering latent knowledge in language models without supervision,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:47.687530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:41.999426Z digest=sha256:352f29db2c74dfdf2bdf49b371ce07b74b6ee4d30548a1da2c0e3ea0537b4cfa

Observation e080e39d-6fdc-46bd-b468-be482af58d7d · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:42.097166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:42.097166Z digest=sha256:06e97413b247c358f1fa37c937ed14883b28b5fc3034e77392624bd286258075

Observation 066fae06-1e46-4c3e-943f-2ef5c7231f6c · outbound

This paper cites A Language Model's Guide Through Latent Space.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models A Language Model's Guide Through Latent Space

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:42.165517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:42.165517Z digest=sha256:69fff828e7ab3458a3c1d9f476050182366cfe88b86d6ec9efb19fd4e3c2462c

Observation ccda8ba5-f1c8-4351-b338-06c32a8b5839 · outbound

This paper cites Improving Activation Steering in Language Models with Mean-Centring.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Improving Activation Steering in Language Models with Mean-Centring

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:42.231617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:42.231617Z digest=sha256:e9a2616a10ecace99ba1c5c9689ae4e8c9c2bd753b59f9cbe81d149fe7727179

Observation 4b8fa7cc-b175-46cf-a4bc-2e07327b259b · outbound

This paper cites Finding alignments between interpretable causal variables and distributed neural representations,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Finding alignments between interpretable causal variables and distributed neural representations,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:47.540585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:42.293333Z digest=sha256:24dfc4600de8191c4fb8ac3d460e132b51679ddf5cdb1970c5c1951d1700eafe

Observation 75b62a26-f9ce-45a9-bf6d-25a79750aa6c · outbound

This paper cites Generative agents: Interactivesimulacraofhumanbehavior,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Generative agents: Interactivesimulacraofhumanbehavior,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:47.341340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:42.349684Z digest=sha256:caf6be1fd796b6470127acf41942f4f8f31c8d14c817e2daa70548d1c485680a

Observation 2e2ee699-326c-43f4-9ce6-e929d40d1b2a · outbound

This paper cites Man is to computer programmer as woman is to homemaker? debiasing word embeddings,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Man is to computer programmer as woman is to homemaker? debiasing word embeddings,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:47.140600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:42.442869Z digest=sha256:e676a2789c9b3167853cbd4e53cf5d649b1bd9caac9531cf7f54cec3c609e64f

Observation 5646bbed-0a47-48ee-92c0-a8de0ca8edb5 · outbound

This paper cites Sparseautoencodersfindhighlyinter- pretable features in language models,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Sparseautoencodersfindhighlyinter- pretable features in language models,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:46.980720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:42.512310Z digest=sha256:f373fa8426379de9d12ddfe3dcd55cc58e800e7e33ca2bb64cdab76aa2bd991e

Observation f7461889-ce63-41fe-89a9-fbd38f2db7d1 · outbound

This paper cites Gold doesn‘t always glitter: Spectral removal of linear and nonlinear guardedattributeinformation,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Gold doesn‘t always glitter: Spectral removal of linear and nonlinear guardedattributeinformation,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:46.832967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:42.632456Z digest=sha256:b728d993833b60202d61a894f147175f50bc69536e76dc0bdd871777dea95dac

Observation ee86d914-0601-459b-b18b-37f5071d3aff · outbound

This paper cites LEACE: Perfect linear concept erasure in closed form,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models LEACE: Perfect linear concept erasure in closed form,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:46.659998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:42.748817Z digest=sha256:8d7b608bbdced94bce58a30175a35b10cb8f3bebe759d52062e5fe4a4aa49de6

Observation 4876e0ea-6e93-4259-9ab9-3a14a1faeb66 · outbound

This paper cites Monitoring latent world states in language models with proposi- tional probes,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Monitoring latent world states in language models with proposi- tional probes,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:46.506000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:42.820828Z digest=sha256:a803651e838226650ef40fbd13ed34425c79bf5ddac768535a850f79e58b257d

Observation 14c8a776-3a4d-498f-8938-059650e3aa86 · outbound

This paper cites Let’s verify step by step,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Let’s verify step by step,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:46.336643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:42.869100Z digest=sha256:93eac66849a4eb7628350864c491eddcf72b8ccfdf0d43df4356a0af577eed2d

Observation 45964bdf-c1b5-4127-be86-22247b70986f · outbound

This paper cites Self-refine: Iterative refinement with self-feedback,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Self-refine: Iterative refinement with self-feedback,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:46.194754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:42.995770Z digest=sha256:24086075118f0a9199cc225a67071b4eb0d90b721431ce0e9537a073a588f0c8

Observation 4f77234f-d40c-4b67-ac9a-698fee7368e6 · outbound

This paper cites Fine-tuning with divergent chains of thought boosts reasoning through self-correction in language models,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Fine-tuning with divergent chains of thought boosts reasoning through self-correction in language models,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:46.041393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:43.095901Z digest=sha256:4246b1aababdfc73d778949d956b1be59cb4a38453246b740e8c2aa1351dbe13

Observation b0fac060-c3ff-45bf-be9d-73d9cfe73d47 · outbound

This paper cites STar: Bootstrapping reasoning with reasoning,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models STar: Bootstrapping reasoning with reasoning,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:45.858720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:43.195392Z digest=sha256:6ad456c55263a6d83a8106785215567e56ab7e5a6612adbd24aabf53d52d4ac8

Observation 2912de11-1a0d-4b74-84d1-31f3dffcfaae · outbound

This paper cites Let’s verify step by step,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Let’s verify step by step,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:45.699379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:43.254472Z digest=sha256:8c9c8f1310f85f228c5a813be16b803253ddc54a48f17738fd723bdaebc4b975

Observation b5856e66-e3bf-4112-b2c0-a838138b03f0 · outbound

This paper cites LLMs Can Easily Learn to Reason from Demonstrations Structure, not content, is what matters!.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models LLMs Can Easily Learn to Reason from Demonstrations Structure, not content, is what matters!

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:43.321333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:43.321333Z digest=sha256:ed41eae9a15c237d89b71b63ea8a33f6e4dc890c5f9138255873962065970851

Observation 1af91da7-bec5-4818-8733-323b5f4b5934 · outbound

This paper cites Shorterbetter: Guidingreasoningmodelstofindoptimalinferencelengthforefficient reasoning,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Shorterbetter: Guidingreasoningmodelstofindoptimalinferencelengthforefficient reasoning,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:43.439679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:43.439679Z digest=sha256:ce8775b9f663b58744ecfc5ce755b8f898dc105fd3429d1928cb1f040ce89d48

Observation ee2f8254-4724-4d15-b0f4-e5a60cf4666e · outbound

This paper cites Unlocking the capabilities of thought: A reasoning boundary framework to quantify and optimize chain-of-thought,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Unlocking the capabilities of thought: A reasoning boundary framework to quantify and optimize chain-of-thought,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:45.472186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:43.510697Z digest=sha256:249e98883dafad360620cab5fc551e65260bf220365bc3a25d8c65708aca1a74

Observation 1e93143a-2574-4054-be9a-79baf7e88ac1 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:43.598058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:43.598058Z digest=sha256:faf9e9d61bcb37c082259556c602d53e2cce407c242f1a9373bb015aa1d7f366

Observation c86d77b5-78b6-4952-baaf-a77e3547715d · outbound

This paper cites Rethinking Reflection in Pre-Training.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Rethinking Reflection in Pre-Training

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:43.678277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:43.678277Z digest=sha256:b1c7013226c6bbeda73991294b19a06f9dd6b2c9ea8ac303f0ab024818ba911b

Observation ab5573a1-46ee-49f2-b945-5d27a0a59f08 · outbound

This paper cites Reflexion: Languageagentswithverbal reinforcement learning,.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Reflexion: Languageagentswithverbal reinforcement learning,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:45.213263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:43.774554Z digest=sha256:2ffc35dba56fd7b6a9a4fbeb0c995ce495a722fe1d24aa30bd5a669e313f0048

Observation 3fa9b162-1ea8-45d9-b1a8-faa954e327d7 · outbound

This paper cites Reinforcement Learning for Reasoning in Large Language Models with One Training Example.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Reinforcement Learning for Reasoning in Large Language Models with One Training Example

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:43.838361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:43.838361Z digest=sha256:0c099f13c8b10a222d64ea0468f0ef713890ea9eb7dbae70d6efd16cff80a791

Observation 33193c25-3851-45af-84ae-f7a83be9ac32 · outbound

This paper cites There may not be aha moment in r1-zero-like training — a pilot study.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models There may not be aha moment in r1-zero-like training — a pilot study

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:03:45.036245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T01:03:43.999176Z digest=sha256:354eb9eed69d64925d68de557bc3cf91296dfafe6f01d48dcf266267472031ee

Pith citing papers

Observation 09e6d2ee-3503-4e81-8bf6-3281c83c64a6 · inbound

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges cites this paper.

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges From Emergence to Control: Probing and Modulating Self-Reflection in Language Models

Reference 201

Resolution
unresolved
no resolver link, observed 2026-08-03T00:55:32.686738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:55:32.686738Z digest=sha256:5922212936c8f61230cc6343ebc212b1a09bb80cf05740ac971cf3ba837c72cb