Pith. sign in

Paper Citation Record · LEDGER

Discovering Latent Knowledge in Language Models Without Supervision

As of 6 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 94 inbound Pith citation observations for arXiv:2212.03827.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2212.03827 v2

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-15T20:34:08.207848Z

measured 138 of 138 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 94 of 94 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:14:16.689102Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

44 of 44 outbound references displayed

  • verified exact25
  • verified fuzzy10
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch9

External citation measurements

45
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 7054bd77-617c-4b80-8bfd-00518389ec30 · outbound

This paper cites A General Language Assistant as a Laboratory for Alignment.

Discovering Latent Knowledge in Language Models Without Supervision A General Language Assistant as a Laboratory for Alignment

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-15T20:34:08.305675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:d792f5bf8e783a7ce2b26d33d1fe0f54579c18d6aa86b58d2f533f954bfca5ca

Observation 33d77505-15b6-4fe6-93ce-2a5dfb2592dd · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Discovering Latent Knowledge in Language Models Without Supervision Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-15T20:34:08.237537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:e65e0688101d8ec15d27f1f07575ee4674282d9056e850cf4e0e17fb08b27f54

Observation e7b72cc2-45c0-47b2-95a7-a421c5a34e29 · outbound

This paper cites Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell.

Discovering Latent Knowledge in Language Models Without Supervision Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T20:34:08.352984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:0e6c27d92e0a8b7561b1b5aaafd14f9b888ec93902bb89c504b053298c8bd2e2

Observation 41f2e11f-2b54-4efb-abe6-81569996d8c1 · outbound

This paper cites On the Opportunities and Risks of Foundation Models.

Discovering Latent Knowledge in Language Models Without Supervision On the Opportunities and Risks of Foundation Models

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T20:34:08.278904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:dc0124f2beb7ca81dd7afaff1a7e29bac4cf18125573ae1e8ba17a5cd8597e54

Observation c402b99c-6335-44fc-b359-d92a8708faba · outbound

This paper cites Language Models are Few-Shot Learners.

Discovering Latent Knowledge in Language Models Without Supervision Language Models are Few-Shot Learners

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-15T20:34:08.284137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:349d0c9e0361c4c822a1f6319237aed534c9ce96c38edc82cd6c160fb54115f4

Observation 316431e8-35a7-4f2f-b5c6-35f228550c4d · outbound

This paper cites Deep reinforcement learning from human preferences.

Discovering Latent Knowledge in Language Models Without Supervision Deep reinforcement learning from human preferences

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T08:39:30.882444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:b76c84ff88603104c0f50a1cec9625871336e7ebf69602110f56d4816353e1d4

Observation bc95caa8-bb73-4a7e-b699-88e094c64a50 · outbound

This paper cites Supervising strong learners by amplifying weak experts.

Discovering Latent Knowledge in Language Models Without Supervision Supervising strong learners by amplifying weak experts

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-15T20:34:08.291275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:21624ae6388493de31481be9206174f02e23c279cef48e46b70e2917a8a9e5b8

Observation 5698c28a-791c-4963-9f57-c48273c6b764 · outbound

This paper cites BoolQ: Exploring the Surprising Difficulty of Natural Yes/No Questions.

Discovering Latent Knowledge in Language Models Without Supervision BoolQ: Exploring the Surprising Difficulty of Natural Yes/No Questions

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-15T20:34:08.294891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:b02eb7783868a823235855360dcedf935bee3c7cdb1680a59f7e679f60222dda

Observation 12269bc7-6dd8-4ba7-89e7-e45b05c82c0f · outbound

This paper cites Truthful AI: Developing and governing AI that does not lie.

Discovering Latent Knowledge in Language Models Without Supervision Truthful AI: Developing and governing AI that does not lie

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:34:08.298850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:f27aba81b7c302cade9e81eaee1cd5c0fed15d0310e665ae3abbdd829b3b18d4

Observation 560aaf69-397e-4a37-9c29-567103800a68 · outbound

This paper cites DeBERTa: Decoding-enhanced BERT with Disentangled Attention.

Discovering Latent Knowledge in Language Models Without Supervision DeBERTa: Decoding-enhanced BERT with Disentangled Attention

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-15T20:34:08.302421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:206a5781dd69b222671f9591fd62ff95341bc6bd2cd34d085fa676b570453473

Observation 5e200fab-3720-44fa-80e5-b0ff20d7eda9 · outbound

This paper cites Unsolved Problems in ML Safety.

Discovering Latent Knowledge in Language Models Without Supervision Unsolved Problems in ML Safety

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:45:28.030299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:91fe7f90f8136ec9a3688839006eec330dc450916fe495313907c01d5e226fa6

Observation 548d9f5b-f722-487f-8316-505edf7f8b12 · outbound

This paper cites AI safety via debate.

Discovering Latent Knowledge in Language Models Without Supervision AI safety via debate

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-15T20:34:08.308667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:1a9d1594c3410e610631521b38782c9bfd6eee7ae4b9108ecb7c4152eea34a07

Observation e76b8ba6-ed1d-4f2c-a0a5-65d7d0c74ac0 · outbound

This paper cites Maieutic Prompting: Logically Consistent Reasoning with Recursive Explanations.

Discovering Latent Knowledge in Language Models Without Supervision Maieutic Prompting: Logically Consistent Reasoning with Recursive Explanations

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T20:34:08.311916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:50147a9a69112ea48831a326fa0f71b9b36146f6346906de467417e6209cfed9

Observation 5d0cd27f-e5b7-485f-ae64-cf47f5e8f508 · outbound

This paper cites Alignment of Language Agents.

Discovering Latent Knowledge in Language Models Without Supervision Alignment of Language Agents

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T20:34:08.315226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:4aa921c03d4d2238a7e4395af1e109826e368a9c70d9ca8135bac05f4d7443c4

Observation c03ffb76-70c4-487c-8b84-28b58ab7dc65 · outbound

This paper cites Ground-Truth Labels Matter: A Deeper Look into Input-Label Demonstrations.

Discovering Latent Knowledge in Language Models Without Supervision Ground-Truth Labels Matter: A Deeper Look into Input-Label Demonstrations

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T20:34:08.318512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:790ee846c33d56a03f3ebfa742e0085bfe79da4247cd48a36d58dc1c05c48514

Observation f03c4e3a-47fe-4779-ad06-5deff716a1fa · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Discovering Latent Knowledge in Language Models Without Supervision Adam: A Method for Stochastic Optimization

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-15T20:34:08.321324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:31f7e91b4b373b0b5459c60e38cbb6462da50482d0ebb7fa1bcdc6c070be6ec0

Observation 838a7ba4-b93f-4f1f-9f8e-aa2f772f759b · outbound

This paper cites Scalable agent alignment via reward modeling: a research direction.

Discovering Latent Knowledge in Language Models Without Supervision Scalable agent alignment via reward modeling: a research direction

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-15T20:34:08.324344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:d31838046c4c2ba174faee458012d1e927a0a6c591e3de3b24e4698ec24e1369

Observation ac205680-71ff-44e0-804c-a37f7454ead5 · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

Discovering Latent Knowledge in Language Models Without Supervision RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-15T20:34:08.327577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:5e0094df5d366362f685a362d649f5c0bd9f649d5538d9ea811769fb42e363f1

Observation 475dc91b-c9e4-42ba-b257-490c63ef2b7c · outbound

This paper cites Decoupled Weight Decay Regularization.

Discovering Latent Knowledge in Language Models Without Supervision Decoupled Weight Decay Regularization

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-05-15T20:34:08.330770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:b6eaea8d19c7a03fe65199866d9e7502579a2f9231b11dd321937bcbd355db7b

Observation e7e7ce69-a935-47ec-8949-7b84c00a5a3a · outbound

This paper cites On Faithfulness and Factuality in Abstractive Summarization.

Discovering Latent Knowledge in Language Models Without Supervision On Faithfulness and Factuality in Abstractive Summarization

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:34:08.334377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:0050b993a3183570fdedf82165912f0f059740cb8617f7c8d73fae298cf49889

Observation c3a61323-0f87-4ad2-881d-0289e4a60ec9 · outbound

This paper cites Teaching language models to support answers with verified quotes.

Discovering Latent Knowledge in Language Models Without Supervision Teaching language models to support answers with verified quotes

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-17T11:45:48.341649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:47fd34dcb809cde9dfdbd7a5ccd2bdf9cbe2328ea0e5b3f60d984db8b1305c0d

Observation 2b18b2f4-47b0-4b94-b739-f752e4157c55 · outbound

This paper cites MetaICL: Learning to Learn In Context.

Discovering Latent Knowledge in Language Models Without Supervision MetaICL: Learning to Learn In Context

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T20:34:08.341197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:d2ca700cafec9f1c5a1ce20617d2301cd8cff10d64c7d1f1fb07a30683d23699

Observation 1583a627-e650-41c0-b44b-1d0cf97e21a8 · outbound

This paper cites WebGPT: Browser-assisted question-answering with human feedback.

Discovering Latent Knowledge in Language Models Without Supervision WebGPT: Browser-assisted question-answering with human feedback

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-15T20:34:08.344378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:7b12cb18921f922aabce4bd8d821aea4a615c572aa25c7a9d0075ecc5c7b4338

Observation c95bebdc-9c95-457b-b97a-ebd6e96d1c0d · outbound

This paper cites Training language models to follow instructions with human feedback.

Discovering Latent Knowledge in Language Models Without Supervision Training language models to follow instructions with human feedback

Reference 24

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T20:34:08.347537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:4e94c3af93ab157c1ec928a7949f726b99593b96406f012b6f3e85098cf9f8ff

Observation 326f6614-903a-48fe-96a1-774af62f3fac · outbound

This paper cites Red Teaming Language Models with Language Models.

Discovering Latent Knowledge in Language Models Without Supervision Red Teaming Language Models with Language Models

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-15T20:34:08.350609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:beeb4fd28aadd11f72d81e1aed1a0719368af5a4c85e209cc7437737c2cb7e1a

Observation 96809c38-c7d3-495c-b8ae-2ee028c92ed7 · outbound

This paper cites Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer.

Discovering Latent Knowledge in Language Models Without Supervision Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer

Reference 26

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T20:34:08.242501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:75ddce3cd36a4b22916b8f9344e107633fe4d7cb9f1de9ee423323bd5113a505

Observation d69ab1c6-9bd7-46b7-9bea-97c93036c77b · outbound

This paper cites SQuAD: 100,000+ Questions for Machine Comprehension of Text.

Discovering Latent Knowledge in Language Models Without Supervision SQuAD: 100,000+ Questions for Machine Comprehension of Text

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-05-15T20:34:08.246766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:7d7bae133c5805693d4aecd443fa724d31b7c8a8873cfd7e757a0e7dceffcb27

Observation e07a74a8-8a1d-4c2e-8c5d-5556d2523c75 · outbound

This paper cites Choice of plausible alternatives: An evaluation of commonsense causal reasoning.

Discovering Latent Knowledge in Language Models Without Supervision Choice of plausible alternatives: An evaluation of commonsense causal reasoning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T20:34:08.367484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:2428560c207e158c728ed02d9ef809ecc3f6a8e274d6d0eb62035c5415834b44

Observation 29f5c200-1097-4bf6-bd93-ed2299d2330e · outbound

This paper cites Multitask Prompted Training Enables Zero-Shot Task Generalization.

Discovering Latent Knowledge in Language Models Without Supervision Multitask Prompted Training Enables Zero-Shot Task Generalization

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-05-15T20:34:08.250101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:7e11540dad2273f267d9dae135ab0b7c08a336989766bda785270eef76e81514

Observation a21247f0-bcbd-4776-84ee-cd02e05a69b1 · outbound

This paper cites Learning to summarize from human feedback.

Discovering Latent Knowledge in Language Models Without Supervision Learning to summarize from human feedback

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:46:18.838660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:a966efda693a27115d682b404f6383448d6c8a84f4f89894b1d9299bba9762af

Observation b05242b8-3f58-4cb6-b355-12ec5e3ff3aa · outbound

This paper cites FEVER: a large-scale dataset for Fact Extraction and VERification.

Discovering Latent Knowledge in Language Models Without Supervision FEVER: a large-scale dataset for Fact Extraction and VERification

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-15T20:34:08.258343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:a5146a8bb16ab967e2acf975135eb3edf3b266339c9d2e39391dffb7ada98444

Observation dac2e12b-a208-4dac-be6e-f423cf1e2fe6 · outbound

This paper cites GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding.

Discovering Latent Knowledge in Language Models Without Supervision GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-05-15T20:34:08.261955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:8b56f856c20049b4cc8b0561fb425b1cf769699314275e81354d29a60fe1d3f1

Observation 23635a30-246e-4ba9-b4e8-79ce61bc5709 · outbound

This paper cites Finetuned Language Models Are Zero-Shot Learners.

Discovering Latent Knowledge in Language Models Without Supervision Finetuned Language Models Are Zero-Shot Learners

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-05-15T20:34:08.265277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:933ec4e28ebe107129939d3077914567bd5260c52f64bb346f7f0a0ad9c63736

Observation 8b5fc1b9-5018-4f6b-9899-270224ad3412 · outbound

This paper cites HuggingFace's Transformers: State-of-the-art Natural Language Processing.

Discovering Latent Knowledge in Language Models Without Supervision HuggingFace's Transformers: State-of-the-art Natural Language Processing

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-05-15T20:34:08.268533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:da50aad669d03d7302d0328269427ea62e89b35a2cca6ec5c9a009e28f2d9209

Observation 44ec5d43-9a94-4655-9930-48e61a3a461d · outbound

This paper cites Calibrate Before Use: Improving Few-Shot Performance of Language Models.

Discovering Latent Knowledge in Language Models Without Supervision Calibrate Before Use: Improving Few-Shot Performance of Language Models

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:34:08.272099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:e6480340852774dfa4100222a4c3c38c09cc53817bbb666aa121a90fdddcff80

Observation bab857f6-9b58-4109-b43c-69fa69990089 · outbound

This paper cites Prompt Consistency for Zero-Shot Task Generalization.

Discovering Latent Knowledge in Language Models Without Supervision Prompt Consistency for Zero-Shot Task Generalization

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T20:34:08.275559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:7d985c9cb15455050a68387cded36a6a1480889b107246aa871790e3b0dab2e6

Observation 93f9bf49-bd74-4324-a474-56a74bb934ea · outbound

This paper cites Is 2+2=4? Yes.

Discovering Latent Knowledge in Language Models Without Supervision Is 2+2=4? Yes

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T20:34:08.369552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:6c16b1f2f2795a856b77a9a1c8cd7e9d3607e213f08fb3290946689c9b2ca194

Observation 44ae3498-4e2b-4f8d-b4be-4a3b20ebf014 · outbound

This paper cites hidden states.

Discovering Latent Knowledge in Language Models Without Supervision hidden states

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T20:34:08.371651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:096396692959943fb9422df2fdca9724484f73985f4236f5f2437255930604a2

Observation 669c4cac-f670-41fb-99f1-d862ba1ceac1 · outbound

This paper cites [text] = I loved this movie.

Discovering Latent Knowledge in Language Models Without Supervision [text] = I loved this movie

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T20:34:08.354996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:16009ed72a6cba3364f7fac27d4b862ac9d5b7377e33309a94ee8bb8ef7f38b1

Observation 7e8f239f-3405-4df5-898c-b2342791e464 · outbound

This paper cites ‘ [content].

Discovering Latent Knowledge in Language Models Without Supervision ‘ [content]

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T20:34:08.357051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:1a4a48aff6025fa4810cf8a054eaf785b6982368e048251f92d4647e815bced8

Observation ee62ca3e-f61e-4b6d-a780-8e1243e6fc66 · outbound

This paper cites Here the label is a short sentence.

Discovering Latent Knowledge in Language Models Without Supervision Here the label is a short sentence

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T20:34:08.359284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:928ddd80170f69de66e8c4a536daaa0dbdbf9ad2ac0dfa18a5dfa343b589b427

Observation d7abd8d6-dcc9-42a3-b49b-3cc077d35637 · outbound

This paper cites ‘ [premise].

Discovering Latent Knowledge in Language Models Without Supervision ‘ [premise]

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T20:34:08.361276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:1071bca62b20987c59b808a7a32d54a9a14ed2f130711b0a4e90af668357b1bd

Observation 5b215f2e-1d43-44a0-bf79-6af2a0aa907d · outbound

This paper cites [label]” is “negative.

Discovering Latent Knowledge in Language Models Without Supervision [label]” is “negative

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T20:34:08.363341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:ad89a22510ac4c31207d85fc0ea63df9c4f79908cbb980f9f08316bdf1787788

Observation 4bf9e264-3cfc-4706-9089-47164b0a879e · outbound

This paper cites yes” or “no.

Discovering Latent Knowledge in Language Models Without Supervision yes” or “no

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-15T20:34:08.365373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:34:08.207848Z digest=sha256:e8b4d9ea8b44062c7113bc8385bb32ec7995d3ff6ffad76b61c56f58f50527f3

Pith citing papers

Observation 972a7f42-7d52-416a-9f5e-4c1a96b903d4 · inbound

The Internal State of an LLM Knows When It's Lying cites this paper.

The Internal State of an LLM Knows When It's Lying Discovering Latent Knowledge in Language Models Without Supervision

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-16T00:08:36.576016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-16T00:08:36.530560Z digest=sha256:c5de1bb1f0d9bca96b2737d0960f4170f6020704d9ced3c593e0ec70d0f352c3

Observation a436c8ed-cd9c-44ce-ac8c-599dc9710c33 · inbound

A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions cites this paper.

A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions Discovering Latent Knowledge in Language Models Without Supervision

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:34:08.372263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T02:46:26.957539Z digest=sha256:af849afb411f4ec4eb9909e8dcf6fb3098cc9cfd83f47d3cc83419a699db6b46

Observation 5a4a42f7-f3a3-4b76-a8e5-d13a63289cfc · inbound

Refusal in Language Models Is Mediated by a Single Direction cites this paper.

Refusal in Language Models Is Mediated by a Single Direction Discovering Latent Knowledge in Language Models Without Supervision

Reference 123

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:34:08.372263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T10:47:55.934081Z digest=sha256:05df6bd846058a9826ee80925962c314ab8f4a3658bfbeb7c8980184e55635b3

Observation a856920c-44b4-4c1e-ac61-6f54e23e35f8 · inbound

Training Language Models to Self-Correct via Reinforcement Learning cites this paper.

Training Language Models to Self-Correct via Reinforcement Learning Discovering Latent Knowledge in Language Models Without Supervision

Reference 137

Resolution
metadata mismatch
local_arxiv, observed 2026-05-17T12:04:10.676680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-17T12:04:10.210508Z digest=sha256:665225e78b2512354bfc328962d3789581388a3088b182e9a315fea09ac4cd2d

Observation 99afcef3-01b2-451f-82c4-4b463a8691ba · inbound

Mechanistic Interpretability Needs Philosophy cites this paper.

Mechanistic Interpretability Needs Philosophy Discovering Latent Knowledge in Language Models Without Supervision

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-21T23:50:47.353935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-21T23:49:19.683025Z digest=sha256:8231a6d1457dfc0529e201c431545d4e7d1f42a6f075dcf8d90d7795ac0001f1

Observation 8a1a899e-d56f-4157-9daa-0fef7608adf1 · inbound

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback cites this paper.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Discovering Latent Knowledge in Language Models Without Supervision

Reference 2000

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:16.689102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:16.689102Z digest=sha256:270b9b01b091cb14049503c5deee50884caa93aca42f0ea1c8b999933a7f802b

Observation 42cad935-733f-447c-b499-5832851180d8 · inbound

LENS: Learning Ensemble Confidence from Neural States for Multi-LLM Answer Integration cites this paper.

LENS: Learning Ensemble Confidence from Neural States for Multi-LLM Answer Integration Discovering Latent Knowledge in Language Models Without Supervision

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T11:03:16.510374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:03:16.510374Z digest=sha256:cdff4c6944c27362a10ed7f7e5d76d3a7e59dd42ac0ff995855f7685c15d4d5c

Observation 788a296d-7b76-4c98-8253-351e8993630b · inbound

A Survey on Data Security in Large Language Models cites this paper.

A Survey on Data Security in Large Language Models Discovering Latent Knowledge in Language Models Without Supervision

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T05:04:58.587323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:04:58.587323Z digest=sha256:2b68bd1aea085b3be806bc9bbd6ec11522e27ad275017313447ab214c6e51120

Observation 208a2103-5324-43ae-af81-1de597b8beb4 · inbound

Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM cites this paper.

Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM Discovering Latent Knowledge in Language Models Without Supervision

Reference 224

Resolution
unresolved
no resolver link, observed 2026-08-05T23:13:05.279845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:13:05.279845Z digest=sha256:1d5a66d05f146d5488aa248ee300d537607b2571eccc47bda0bf33ceb8df1b46

Observation 34721bb2-9705-4205-8fb2-c98b1cc3fd4e · inbound

Quantized but Deceptive? A Multi-Dimensional Truthfulness Evaluation of Quantized LLMs cites this paper.

Quantized but Deceptive? A Multi-Dimensional Truthfulness Evaluation of Quantized LLMs Discovering Latent Knowledge in Language Models Without Supervision

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T15:52:44.867137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:52:44.867137Z digest=sha256:7183b626f0db179d34e5997b06705d8b776199483d48bc68e188ae49803a6ece

Observation e0d35989-fd01-498f-bf30-5a86313d7fc2 · inbound

STARE at the Structure: Steering ICL Exemplar Selection with Structural Alignment cites this paper.

STARE at the Structure: Steering ICL Exemplar Selection with Structural Alignment Discovering Latent Knowledge in Language Models Without Supervision

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T14:47:11.367586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:47:11.367586Z digest=sha256:7eb4a3b91ccb57f9e58e4c5d1b7b4a585a53bb5b03c3c29f5202b7a94ab14797

Observation d6239db5-d241-4e32-b4db-0f8ff4549ecf · inbound

Two Causes, Not One: Rethinking Omission and Fabrication Hallucinations in MLLMs cites this paper.

Two Causes, Not One: Rethinking Omission and Fabrication Hallucinations in MLLMs Discovering Latent Knowledge in Language Models Without Supervision

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-05T13:44:11.202844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:44:11.202844Z digest=sha256:9b345d86aa0f7554eca6501049476bd01c73a0e57c83f07eb1bd80e670af73da

Observation db747075-a0cc-407d-971f-457774cc3c43 · inbound

Can LLMs Lie? Investigation beyond Hallucination cites this paper.

Can LLMs Lie? Investigation beyond Hallucination Discovering Latent Knowledge in Language Models Without Supervision

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T10:55:31.201012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:55:31.201012Z digest=sha256:32f49ddb0c85c95c677627fb15df4afe439a2fa47852ad2d30c9c57126b83785

Observation 0965d9b4-8212-4343-8f0a-c922e0bfc76c · inbound

Cross-Layer Attention Probing for Fine-Grained Hallucination Detection cites this paper.

Cross-Layer Attention Probing for Fine-Grained Hallucination Detection Discovering Latent Knowledge in Language Models Without Supervision

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T10:18:02.808211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:18:02.808211Z digest=sha256:cae4994f23fa0f0e921d32a3c1b75b4420aaab302d82f083294aa812962a078e

Observation 737f2084-1b16-47e0-85e0-155f7b7506aa · inbound

Unsupervised Hallucination Detection by Inspecting Reasoning Processes cites this paper.

Unsupervised Hallucination Detection by Inspecting Reasoning Processes Discovering Latent Knowledge in Language Models Without Supervision

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T18:25:37.321564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T18:25:37.321564Z digest=sha256:166f6009bbff1c4ce24b715d6d231cd786da9bc677de93356b1632bf00a2e426

Observation b92854d8-3d78-4b48-9f3d-f77b14696bae · inbound

HalluField: Detecting LLM Hallucinations via Field-Theoretic Modeling cites this paper.

HalluField: Detecting LLM Hallucinations via Field-Theoretic Modeling Discovering Latent Knowledge in Language Models Without Supervision

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:04.927879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:04.927879Z digest=sha256:dfcbfdb86c00b0f5a8102b06cf40c1c81849c50e6b79f066bc7cd2de6a8aeb4f

Observation b9f8d589-fb1d-421f-b447-01e879dbc9e5 · inbound

Neural Message-Passing on Attention Graphs for Hallucination Detection cites this paper.

Neural Message-Passing on Attention Graphs for Hallucination Detection Discovering Latent Knowledge in Language Models Without Supervision

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T13:52:07.804961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:52:07.804961Z digest=sha256:564c36ce5a6a16415a9e17824590d7c89c26a07d142665e1d221419031e781d9

Observation 1d3ccaa2-bfbe-4d8b-b28e-3105ba86be42 · inbound

Geometry of Reason: Spectral Signatures of Valid Mathematical Reasoning cites this paper.

Geometry of Reason: Spectral Signatures of Valid Mathematical Reasoning Discovering Latent Knowledge in Language Models Without Supervision

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-03T13:03:21.532835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:03:21.532835Z digest=sha256:71ad560f733927869dfc19964f9ddb406b3db899b099470408b1731ac678fc15

Observation 5b2de349-72c6-4f7b-ab80-61ce05f197a7 · inbound

No Reliable Evidence of Self-Reported Sentience in Small Large Language Models cites this paper.

No Reliable Evidence of Self-Reported Sentience in Small Large Language Models Discovering Latent Knowledge in Language Models Without Supervision

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T09:31:49.744774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:31:49.744774Z digest=sha256:b22a53a9201f289566371d0bcac885c2aa49c4d89c5d509e47570dce6f3bc27d

Observation fd69cf10-c629-4061-80ef-e2e15dd1ebc3 · inbound

TOPReward: Token Probabilities as Hidden Zero-Shot Rewards for Robotics cites this paper.

TOPReward: Token Probabilities as Hidden Zero-Shot Rewards for Robotics Discovering Latent Knowledge in Language Models Without Supervision

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T21:43:07.066526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:43:07.066526Z digest=sha256:ccff1b62f8975b6987223ea4a20f924302065e25c0c84538e41b9c364f4bf387

Observation bded4634-bc39-47e4-84ff-415f09055973 · inbound

Emergent Manifold Separability during Reasoning in Large Language Models cites this paper.

Emergent Manifold Separability during Reasoning in Large Language Models Discovering Latent Knowledge in Language Models Without Supervision

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:34:08.372263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T20:12:05.809065Z digest=sha256:a2e37ba89d029b05f03018401ae460be9ab598c5cef7deaaa2b8d52167e64949

Observation 4aace6aa-ea9d-41de-b8b7-1bdd9cb7ca87 · inbound

Prompt Injection as Role Confusion cites this paper.

Prompt Injection as Role Confusion Discovering Latent Knowledge in Language Models Without Supervision

Reference 2024

Resolution
malformed identifier
no resolver link, observed 2026-08-02T21:46:13.107901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:46:13.107901Z digest=sha256:d2b2e3e4993cc16b9abfde3a948017e4334c4ab661050f1e9964653c4a86f98c

Observation d64fab97-26cc-4a83-84bc-4d05345c50ef · inbound

How do LLMs Compute Verbal Confidence cites this paper.

How do LLMs Compute Verbal Confidence Discovering Latent Knowledge in Language Models Without Supervision

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-21T10:40:00.611725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-21T10:38:28.533485Z digest=sha256:b4622406cab252966555b1df09e802f6b81e0b284f3c0030537373d4d5045856

Observation 7be395c5-0fd1-4d93-a8ee-072d92950f40 · inbound

To See or To Please: Uncovering Visual Sycophancy and Split Beliefs in VLMs cites this paper.

To See or To Please: Uncovering Visual Sycophancy and Split Beliefs in VLMs Discovering Latent Knowledge in Language Models Without Supervision

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T20:34:08.372263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T09:18:34.853946Z digest=sha256:6f2c0ede85a12aa4e1cceb0f45db179d7f2362e2c2be58d4525764d9a1fd24ed

Observation 8b14138c-2eaf-4c4d-8303-afbd55836c96 · inbound

Interpretable Electrophysiological Features of Resting-State EEG Capture Cortical Network Dynamics in Parkinsons Disease cites this paper.

Interpretable Electrophysiological Features of Resting-State EEG Capture Cortical Network Dynamics in Parkinsons Disease Discovering Latent Knowledge in Language Models Without Supervision

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-13T14:23:49.346727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T14:23:49.346727Z digest=sha256:70f5a2b4755522e5b13750c18c4f36f0423007cbc047208486162ea9876b6759

Observation b2834ec8-8899-4835-81b0-70fbcf06de46 · inbound

Weakly Supervised Distillation of Hallucination Signals into Transformer Representations cites this paper.

Weakly Supervised Distillation of Hallucination Signals into Transformer Representations Discovering Latent Knowledge in Language Models Without Supervision

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:34:08.372263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T19:49:35.865609Z digest=sha256:8967a5f63011c121169f5409a3c44d9efdc9fec4686820ddf9f796bf4fa6ebdd

Observation 5494bc89-f8f3-40da-974c-95101eb4dedd · inbound

The Long Delay to Arithmetic Generalization: When Learned Representations Outrun Behavior cites this paper.

The Long Delay to Arithmetic Generalization: When Learned Representations Outrun Behavior Discovering Latent Knowledge in Language Models Without Supervision

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:34:08.372263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:16:40.430384Z digest=sha256:6087e162d7e8dfc0736775dad8b22086ca0734681b5bd06f254094eae7bad644

Observation 13e3a36c-5c8b-49ba-9e03-63963fc621b9 · inbound

Learning Uncertainty from Sequential Internal Dispersion in Large Language Models cites this paper.

Learning Uncertainty from Sequential Internal Dispersion in Large Language Models Discovering Latent Knowledge in Language Models Without Supervision

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:34:08.372263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-10T08:36:39.242766Z digest=sha256:204f03780f09485f4e1c02b1c8cd845054f4562351638a71ce5c2de298acb11c

Observation 81c5f600-eaa4-4f86-b2a1-6cd0cd3819ec · inbound

How Tokenization Limits Phonological Knowledge Representation in Language Models and How to Improve Them cites this paper.

How Tokenization Limits Phonological Knowledge Representation in Language Models and How to Improve Them Discovering Latent Knowledge in Language Models Without Supervision

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:34:08.372263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-10T06:48:45.333416Z digest=sha256:beccf6e43a7219a0c2e69cfbb3d025751b67ffcfc77f60b48f6b32fbfef0c73e

Observation 794bba6f-e38f-49ce-99b6-7b1ae2127bba · inbound

Are LLM Uncertainty and Correctness Encoded by the Same Features? A Functional Dissociation via Sparse Autoencoders cites this paper.

Are LLM Uncertainty and Correctness Encoded by the Same Features? A Functional Dissociation via Sparse Autoencoders Discovering Latent Knowledge in Language Models Without Supervision

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:34:08.372263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-10T02:51:12.492123Z digest=sha256:0cad18b912b62c9df0697967a8e3565d4258b710e63ea7e4c960012a1ae585ff

Observation 20783210-ed4f-436e-a807-f47148b73a8c · inbound

How LLMs Detect and Correct Their Own Errors: The Role of Internal Confidence Signals cites this paper.

How LLMs Detect and Correct Their Own Errors: The Role of Internal Confidence Signals Discovering Latent Knowledge in Language Models Without Supervision

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:34:08.372263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:30:51.094216Z digest=sha256:7b3295c78bea6b789ba83d770d34c266352aa57b642b0d4542466604e5d669a1

Observation 51eddca8-d954-4d0a-92ae-a5689df686a0 · inbound

Perturbation Probing: A Two-Pass-per-Prompt Diagnostic for FFN Behavioral Circuits in Aligned LLMs cites this paper.

Perturbation Probing: A Two-Pass-per-Prompt Diagnostic for FFN Behavioral Circuits in Aligned LLMs Discovering Latent Knowledge in Language Models Without Supervision

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:34:08.372263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T08:34:14.310656Z digest=sha256:6e1764ac2e9fbcb05336748fabad741915efc1bc39d40fbd95010783a3ef073b

Observation e77e8cdf-ab76-4031-8f23-5d13648ba686 · inbound

Geometric Deviation as an Unsupervised Pre-Generation Reliability Signal: Probing LLM Representations for Answerability cites this paper.

Geometric Deviation as an Unsupervised Pre-Generation Reliability Signal: Probing LLM Representations for Answerability Discovering Latent Knowledge in Language Models Without Supervision

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T20:34:08.372263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-08T17:56:27.360060Z digest=sha256:3fe61e040c68dd1b5ea73e343dbf98f669846e9a82a750de298855c2a09a2411

Observation bf765001-42ef-4f78-92ff-c15d497f1ed4 · inbound

Decodable but Not Corrected by Fixed Residual-Stream Linear Steering: Evidence from Medical LLM Failure Regimes cites this paper.

Decodable but Not Corrected by Fixed Residual-Stream Linear Steering: Evidence from Medical LLM Failure Regimes Discovering Latent Knowledge in Language Models Without Supervision

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:34:08.372263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-08T11:42:16.090162Z digest=sha256:d213de1ca765c399c1b2a702d297fcc7490370bbdde8b26d6e84989a0aa4cd5a

Observation 240a3d52-49f6-40e1-87b3-66b0bce81b0c · inbound

The Geometry of Forgetting: Temporal Knowledge Drift as an Independent Axis in LLM Representations cites this paper.

The Geometry of Forgetting: Temporal Knowledge Drift as an Independent Axis in LLM Representations Discovering Latent Knowledge in Language Models Without Supervision

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:34:08.372263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T01:55:32.734498Z digest=sha256:107d684786d5bad51f4603105a1e397cf644d845e5ddc4ab280d84605b36aff2

Observation 0be2a14b-9cd2-4eef-9ae3-3db49ac3544f · inbound

Repeated-Token Counting Reveals a Dissociation Between Representations and Outputs cites this paper.

Repeated-Token Counting Reveals a Dissociation Between Representations and Outputs Discovering Latent Knowledge in Language Models Without Supervision

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:34:08.372263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T04:19:28.650522Z digest=sha256:4755f0f0fac055ab61c9d0b6529ea77a018ebba5fe9ef98f73bdfbe33adf583b

Observation 148004bd-cf73-4510-ae50-056f92d70652 · inbound

Repeated-Token Counting Reveals a Dissociation Between Representations and Outputs cites this paper.

Repeated-Token Counting Reveals a Dissociation Between Representations and Outputs Discovering Latent Knowledge in Language Models Without Supervision

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-02T14:31:12.034394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:31:12.034394Z digest=sha256:e63cc1ee1b293ec44ea52068c2ee7f994c02dff4f0800f4c08c30cfc41d8c836

Observation 19837d01-1976-41dd-b6a2-05f2d0360a41 · inbound

LLM Agents Already Know When to Call Tools -- Even Without Reasoning cites this paper.

LLM Agents Already Know When to Call Tools -- Even Without Reasoning Discovering Latent Knowledge in Language Models Without Supervision

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T20:34:08.372263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T02:44:00.329276Z digest=sha256:0b1463d6bac5ad2f69f27734e371442399c1fed5ae71b4c300dce5f1d9dcbd77

Observation b8e04aa4-4e78-48dc-ac11-00e62a8cd347 · inbound

LLM Agents Already Know When to Call Tools -- Even Without Reasoning cites this paper.

LLM Agents Already Know When to Call Tools -- Even Without Reasoning Discovering Latent Knowledge in Language Models Without Supervision

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-05-22T10:56:25.877372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-22T10:55:14.037226Z digest=sha256:477548164cca106926f3d1e9ad0beeb1ee6cc0f44035aa2445412f6279e8c991

Observation 0f003387-f22f-4da1-a30d-ae18832fadc6 · inbound

Positive Alignment: Artificial Intelligence for Human Flourishing cites this paper.

Positive Alignment: Artificial Intelligence for Human Flourishing Discovering Latent Knowledge in Language Models Without Supervision

Reference 169

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T20:34:08.372263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-15T05:56:56.902705Z digest=sha256:dd34549586d9d7a6628ea3a48518c5922dd8169b9ad030ca85132b228d28d13a

Observation 68f47cbd-83af-403b-9446-21e8ef19ac46 · inbound

Stories in Space: In-Context Learning Trajectories in Conceptual Belief Space cites this paper.

Stories in Space: In-Context Learning Trajectories in Conceptual Belief Space Discovering Latent Knowledge in Language Models Without Supervision

Reference 70

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T20:34:08.372263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T05:17:34.283917Z digest=sha256:90eff6b166a568a48514b742dfef59cbb71e60cd1bb358cce58806fefae8172f

Observation d235f8c1-d6db-4105-b730-660c8b006b2c · inbound

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces cites this paper.

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces Discovering Latent Knowledge in Language Models Without Supervision

Reference 137

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:34:08.372263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-14T20:17:01.224864Z digest=sha256:33ae79ec8b8f9deee66c5de3353aa06388fbe839a821c441032d6b70cca4469e

Observation f7b26cd2-f8f1-443c-bc01-60f5f5733347 · inbound

PAIR: Prefix-Aware Internal Reward Model for Multi-Turn Agent Optimization cites this paper.

PAIR: Prefix-Aware Internal Reward Model for Multi-Turn Agent Optimization Discovering Latent Knowledge in Language Models Without Supervision

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-20T10:43:12.459720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T10:41:25.205368Z digest=sha256:c4e8727035bd2b7bf3b6858e814265bd7a72a265b1bff7cb43f5fb5655a19579

Observation b12cf2a5-3e99-4956-ab39-40a3e91f015a · inbound

PAIR: Prefix-Aware Internal Reward Model for Multi-Turn Agent Optimization cites this paper.

PAIR: Prefix-Aware Internal Reward Model for Multi-Turn Agent Optimization Discovering Latent Knowledge in Language Models Without Supervision

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T02:23:29.475276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:23:29.475276Z digest=sha256:e46987b6bbcca47be51028301a3215feb3aca19c8e9f3085ec76f09d85b4e361

Observation 8445a34d-820e-40ce-af87-ef1254df4472 · inbound

Trust or Abstain? A Self-Aware RAG Approach cites this paper.

Trust or Abstain? A Self-Aware RAG Approach Discovering Latent Knowledge in Language Models Without Supervision

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-20T23:03:50.377331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T23:02:58.671115Z digest=sha256:e0dfa90a3aaef382e99a9fabb47f999f827ff58f2df17bbe82552614fbabbb59

Observation abc1a637-f32f-4bf0-965d-eb101dd82cdd · inbound

Manifold-Guided Attention Steering cites this paper.

Manifold-Guided Attention Steering Discovering Latent Knowledge in Language Models Without Supervision

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-22T09:14:45.318091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-22T09:14:12.029680Z digest=sha256:fe8da3cf9f50c89072adc8b099df76cb0fcfc72da276112b15ae2fdfd4585bd5

Observation 9a0d9cb4-ca11-4c2e-8336-fb1ffae02cef · inbound

Reading Calibrated Uncertainty from Language Model Trajectories cites this paper.

Reading Calibrated Uncertainty from Language Model Trajectories Discovering Latent Knowledge in Language Models Without Supervision

Reference 29

Resolution
metadata mismatch
local_arxiv, observed 2026-05-25T05:45:23.945287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-25T05:44:23.197975Z digest=sha256:892889532996be7973c85688e8123f10a8981bf8bf2f701d58a8bf935fb2a130

Observation ebe266b5-2bd6-4f20-bde9-98b6b079f2a8 · inbound

Detecting Is Not Resolving: The Monitoring Control Gap in Retrieval Augmented LLMs cites this paper.

Detecting Is Not Resolving: The Monitoring Control Gap in Retrieval Augmented LLMs Discovering Latent Knowledge in Language Models Without Supervision

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-06-29T17:13:44.850347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-29T17:06:46.379906Z digest=sha256:b222dfaffcb9d329e12dfcbf711e0d52bbb6dd25e02b936fb27f8affcbb3c7ec

Observation a37d5119-93fe-4af7-b9a2-f7cd5203f68c · inbound

Prefix-Safe Bayesian Belief Tracking for LLM Reasoning Reliability:Separating Calibration from Ranking cites this paper.

Prefix-Safe Bayesian Belief Tracking for LLM Reasoning Reliability:Separating Calibration from Ranking Discovering Latent Knowledge in Language Models Without Supervision

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-06-29T17:03:40.977377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-29T16:59:25.127619Z digest=sha256:60d55c806efc8716ada12a4333b4d4e370ac7b01179ee2aa1ad0b96d2f881a00

Observation 0d5845af-5a7e-4218-8c30-489698a744e1 · inbound

The Attentional White Bear Effect in Transformer Language Models cites this paper.

The Attentional White Bear Effect in Transformer Language Models Discovering Latent Knowledge in Language Models Without Supervision

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-06-29T12:53:27.042291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T12:43:43.792334Z digest=sha256:f82866f1ccb9d6bf8d1b2e22db77c02cf08bb60f5863552b048f4d42e4ef6d73

Observation 425ad979-64c6-4b4e-bfec-1a9a76639028 · inbound

Can LLMs Use Linguistic Uncertainty Markers to Reliably Reflect Intrinsic Confidence? cites this paper.

Can LLMs Use Linguistic Uncertainty Markers to Reliably Reflect Intrinsic Confidence? Discovering Latent Knowledge in Language Models Without Supervision

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-06-29T12:23:24.421371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T12:18:36.854164Z digest=sha256:26b8e5b6a2653d2f60ac53537e5103c635c9d7beb276f3d6b68acd5294e07cbe

Observation ea620236-5c89-40f9-9a39-07223841e287 · inbound

Scaling Monosemanticity: Extracting Interpretable Features from Claude 3 Sonnet cites this paper.

Scaling Monosemanticity: Extracting Interpretable Features from Claude 3 Sonnet Discovering Latent Knowledge in Language Models Without Supervision

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-06-29T07:53:13.424409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:50:04.813379Z digest=sha256:14d988dced20d5db256d9ad687f0bc134ea473e56c8386a6ce47626367a89552

Observation 032fde70-6b90-4177-963a-7df421816826 · inbound

UniSteer: Text-Guided Flow Matching in Activation Space for Versatile LLM Steering cites this paper.

UniSteer: Text-Guided Flow Matching in Activation Space for Versatile LLM Steering Discovering Latent Knowledge in Language Models Without Supervision

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T08:13:15.611293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T08:06:57.884950Z digest=sha256:04324addd2728025232a8f9e5f4dc1bdbdaa6dd0dd5ad3a32d4beb7f3e76ba91

Observation 68d539dc-ec8c-445b-a156-a838d7880a14 · inbound

CANARY: Zero-Label Detection of Fine-Tuning Contamination in Language Models cites this paper.

CANARY: Zero-Label Detection of Fine-Tuning Contamination in Language Models Discovering Latent Knowledge in Language Models Without Supervision

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T22:26:18.041801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T15:22:01.257829Z digest=sha256:94be7d57f9b7a3f69df8f61ec6b5a3b3603da17c7fb019c947bd88c9064bb369

Observation edcd55e5-5802-4b5e-8faa-c1aea132a358 · inbound

Hallucination Is Linearly Decodable from Mid-Layer Hidden States in Quantized LLMs cites this paper.

Hallucination Is Linearly Decodable from Mid-Layer Hidden States in Quantized LLMs Discovering Latent Knowledge in Language Models Without Supervision

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-06-28T19:22:34.868638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T19:13:23.794792Z digest=sha256:5d020672e36a9bf6b7edc114ff0fa4b7a27eb10a8d647d242bcac91d5e09b17a

Observation 8c7a4ce7-7193-4630-ab35-d123cb01ad32 · inbound

Consistency Training Can Entrench Misalignment cites this paper.

Consistency Training Can Entrench Misalignment Discovering Latent Knowledge in Language Models Without Supervision

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T03:26:28.597517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-28T10:07:31.153337Z digest=sha256:adfe3a509c3738c1d5e2c6e1eb0888d50e70f388493581a6ffe34e313c2edcc4

Observation b70fcb39-61c6-43c1-8e5a-8a6447d5c9fa · inbound

CASS-RTL: Correctness-Aware Subspace Steering for RTL Generation with LLMs cites this paper.

CASS-RTL: Correctness-Aware Subspace Steering for RTL Generation with LLMs Discovering Latent Knowledge in Language Models Without Supervision

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-07-02T15:57:07.351714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T23:06:44.339563Z digest=sha256:8b1d1059f6a3a00c159ee4f352051f0771b0f418e3362cf1628ee41cdce7dc20

Observation 726e4a78-ae6f-4130-850d-d33cb3276854 · inbound

LLM Self-Recognition: Steering and Retrieving Activation Signatures cites this paper.

LLM Self-Recognition: Steering and Retrieving Activation Signatures Discovering Latent Knowledge in Language Models Without Supervision

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-06-28T01:41:29.292254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-28T01:41:03.518190Z digest=sha256:e3ebcb3a004249c65b8ef1e0a1d690883c420fc57c26612d33d602c85a36429d

Observation f1590fe2-9c58-4a74-80cb-8fb82e85583c · inbound

Adversarial Robustness of Activation Steering in Large Language Models cites this paper.

Adversarial Robustness of Activation Steering in Large Language Models Discovering Latent Knowledge in Language Models Without Supervision

Reference 15

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T16:37:09.415704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-27T22:30:36.839964Z digest=sha256:f5d4ce5e737146cc52861510145d1fb4fc7d0dde9319b0511c99f85fd57147f1

Observation d2c0df88-dd3e-4fd3-a531-d1574f04206c · inbound

Now You (Still) See Me: Detecting Evasive Steganographic Payloads in LLMs cites this paper.

Now You (Still) See Me: Detecting Evasive Steganographic Payloads in LLMs Discovering Latent Knowledge in Language Models Without Supervision

Reference 32

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T02:07:33.628376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-27T16:10:18.569412Z digest=sha256:2c590dbe6315a697cc0a3b0ee58ef4e04ef02b3fd56d788c651596ac46082704

Observation 3fe7d966-cb83-4746-a98c-f67677d0a77f · inbound

PRISM: Recovering Instruction Sets from Language Model Activations cites this paper.

PRISM: Recovering Instruction Sets from Language Model Activations Discovering Latent Knowledge in Language Models Without Supervision

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-06-27T17:01:08.112795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-27T16:52:02.948457Z digest=sha256:9efe4576958a380edcc237b2bf72cb858180ce2af34163966acbcf039745f5d0

Observation d2db5217-5ae7-4d8a-98da-78e0d5e50ee5 · inbound

Toward Calibrated, Fair, and accurate Deepfake Detection cites this paper.

Toward Calibrated, Fair, and accurate Deepfake Detection Discovering Latent Knowledge in Language Models Without Supervision

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-06-28T07:11:45.306077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-28T07:05:18.026601Z digest=sha256:9b27cedb0bee490eb65eddfd77a77893793b910f3b9285a1d8cd10f599d1305f

Observation 29b05033-4b28-41f9-8174-14e58829eee9 · inbound

Forecasting Future Behavior as a Learning Task cites this paper.

Forecasting Future Behavior as a Learning Task Discovering Latent Knowledge in Language Models Without Supervision

Reference 59

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T06:07:41.329043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-27T12:55:39.494339Z digest=sha256:70a619ceb24e02ec20aae1a53031e734359c8781b52867949c4c64d765b6f1d0

Observation f33b6223-44aa-466e-92ca-5aa5c8c7b925 · inbound

Pre-Generation Hallucination Detection in Large Language Models via Soft-Target Attention Probing cites this paper.

Pre-Generation Hallucination Detection in Large Language Models via Soft-Target Attention Probing Discovering Latent Knowledge in Language Models Without Supervision

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T07:59:40.433972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T12:19:03.745305Z digest=sha256:253ea5bca5a007ee72d7852a8b483082bad03d0de552ee76bb2c6d4af4eeaa07

Observation 7f67689c-c851-4bea-ab5b-dab52f299521 · inbound

When Agents Commit Too Soon: Diagnosing Premature Commitment in LLM Agents cites this paper.

When Agents Commit Too Soon: Diagnosing Premature Commitment in LLM Agents Discovering Latent Knowledge in Language Models Without Supervision

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-04T10:49:45.709560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T08:32:24.949332Z digest=sha256:a1c770294b0ae2344e9923e5cecd5ce2575c0ed2074a95953fe844515f06089d

Observation a6bb1225-a0a3-4c8e-a376-655bf79aa0e8 · inbound

Plans Don't Persist: Why Context Management Is Load Bearing for LLM Agents cites this paper.

Plans Don't Persist: Why Context Management Is Load Bearing for LLM Agents Discovering Latent Knowledge in Language Models Without Supervision

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-04T10:49:46.382030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T08:27:28.236073Z digest=sha256:d6463589fadaf40b475d4ad2df4b0f9646fb5f967eeac875884a1980bfa8fb04

Observation 1cdbb5a5-2feb-45cb-8bf8-99aafe702d90 · inbound

Perfect Detection, Failed Control: The Geometry of Knowing vs. Steering in Language Models cites this paper.

Perfect Detection, Failed Control: The Geometry of Knowing vs. Steering in Language Models Discovering Latent Knowledge in Language Models Without Supervision

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-04T16:49:57.477292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T00:14:40.674306Z digest=sha256:282c4cc09849b386a5751980c52e31ee931947c45a8aae5fed5e4f7435a15502

Observation 0cfb732d-67cc-4985-a2fc-187b0d3880c8 · inbound

Radical AI Interpretability cites this paper.

Radical AI Interpretability Discovering Latent Knowledge in Language Models Without Supervision

Reference 29

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T13:09:51.066147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-26T05:26:47.854890Z digest=sha256:616a0da8405247d9a471211b96e0dbe581194309c5202dc8fcbd1c6481b4fc92

Observation 815fb3b0-c0b6-4a99-afab-60022749c5a6 · inbound

Auditing Framing-Sensitive Behavioral Instability in Large Language Models for Mental Health Interactions cites this paper.

Auditing Framing-Sensitive Behavioral Instability in Large Language Models for Mental Health Interactions Discovering Latent Knowledge in Language Models Without Supervision

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-07-04T14:09:52.856519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T04:33:07.815172Z digest=sha256:770a15f1936453aae3ece66b64b2efb042c6247daa69a0589dc51b0968923e70

Observation 9c8c957e-203c-4a69-89b5-636ecf12beba · inbound

From Signals to Transfer: A Factorised Study of Probe-Based Uncertainty Estimation in Large Language Models cites this paper.

From Signals to Transfer: A Factorised Study of Probe-Based Uncertainty Estimation in Large Language Models Discovering Latent Knowledge in Language Models Without Supervision

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T18:23:51.030162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-29T05:05:00.406145Z digest=sha256:9e67bc2fca1f40c344abd932d55e0a6bd0c9e6c4c92bb10767b9be6f70fb69d4

Observation 988d2075-4bd9-45aa-902e-1ed74210f20c · inbound

The strength of clinical evidence is recoverable from language model representations but not from their stated grades cites this paper.

The strength of clinical evidence is recoverable from language model representations but not from their stated grades Discovering Latent Knowledge in Language Models Without Supervision

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-06-30T09:34:34.612813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-30T09:30:55.171017Z digest=sha256:10622d130f2ab4dfdd86899326d83514c64a032cad29c5f1deb3a8bed2db068c

Observation b3528d69-d2c3-49fb-8a45-7bc3bad167a1 · inbound

Internal-State Probes Read the Situation, Not the Action: Three Negative Results for Pre-Action Misalignment Monitoring cites this paper.

Internal-State Probes Read the Situation, Not the Action: Three Negative Results for Pre-Action Misalignment Monitoring Discovering Latent Knowledge in Language Models Without Supervision

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T07:44:22.026858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-30T07:37:47.420986Z digest=sha256:293ef7c5c3c561c0cc229b93b46f26ac1028777387addeec4fbf6c3a1b8184e3

Observation 51c3f320-af26-410a-b377-1e688649833e · inbound

Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs cites this paper.

Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs Discovering Latent Knowledge in Language Models Without Supervision

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-07-01T10:35:42.045762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:22:38.232552Z digest=sha256:bff7c5e2b0b1cab26fe77f7cdfd27c293293a1f4409b30e9a0ec46c4b20c87a2

Observation 65575cb8-6020-4fc1-b1c3-ec9f82a82757 · inbound

PRA-RAG: Provably Robust Aggregation in Retrieval-Augmented Generation against Retrieval Corruption cites this paper.

PRA-RAG: Provably Robust Aggregation in Retrieval-Augmented Generation against Retrieval Corruption Discovering Latent Knowledge in Language Models Without Supervision

Reference 60

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T23:47:27.197660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-07-02T23:42:23.695930Z digest=sha256:1502a69c1058c4e80497b414e04db730677df61fd3350f72f61df7a5642d7d81

Observation ac1e7194-fc01-4a14-bd82-fbc71b4e4c05 · inbound

Readable but Not Controllable: Neuron-Level Evidence for Medical LLM Hallucination cites this paper.

Readable but Not Controllable: Neuron-Level Evidence for Medical LLM Hallucination Discovering Latent Knowledge in Language Models Without Supervision

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-02T19:17:17.626552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-02T19:16:31.055024Z digest=sha256:570cca4273b17392b87091a94bd3bae561831b106c00951b0c5d21dcce5f6b37

Observation e99762da-d12e-4ea4-b327-88f774e91423 · inbound

Subliminal Clocks: Latent Time Modelling in Diffusion Language Models cites this paper.

Subliminal Clocks: Latent Time Modelling in Diffusion Language Models Discovering Latent Knowledge in Language Models Without Supervision

Reference 22

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T13:58:21.254599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-07-03T13:48:58.653341Z digest=sha256:22ab57efbbd81e6c7a61786b0a8855ecd87d0b35b27293a5d5741bd72ed3c288

Observation cc5b659a-cfca-421e-948d-a9b4b4bd793c · inbound

Subliminal Clocks: Latent Time Modelling in Diffusion Language Models cites this paper.

Subliminal Clocks: Latent Time Modelling in Diffusion Language Models Discovering Latent Knowledge in Language Models Without Supervision

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T09:07:02.234749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T09:07:02.234749Z digest=sha256:0d62d68988910d56fc2ca38afa0dd21012a3e425639061990afb7df5a4072ec7

Observation fdb5c000-5b4c-40a3-b989-4f544d912bc1 · inbound

Weak-to-Strong Generalization via Direct On-Policy Distillation cites this paper.

Weak-to-Strong Generalization via Direct On-Policy Distillation Discovering Latent Knowledge in Language Models Without Supervision

Reference 89

Resolution
verified exact
local_arxiv, observed 2026-07-07T12:33:45.027092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-07T12:31:42.224094Z digest=sha256:a2f1298c07079ed984e9609eac5dc85980e1a8e9361220f2e144978415bee6f5

Observation 5ddb4ac6-1bdc-4ba8-918d-16f06d406da1 · inbound

Weak-to-Strong Generalization via Direct On-Policy Distillation cites this paper.

Weak-to-Strong Generalization via Direct On-Policy Distillation Discovering Latent Knowledge in Language Models Without Supervision

Reference 85

Resolution
unresolved
no resolver link, observed 2026-07-11T07:01:56.628017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T07:01:56.628017Z digest=sha256:2904913bf1a72513733cd570333a4100778d3086611e9996a67731ae90c845fd

Observation 74b1ff89-6bd2-4fc3-9ea7-260b4678a84c · inbound

Dissociating the Internal Representations of Sycophancy in LLMs cites this paper.

Dissociating the Internal Representations of Sycophancy in LLMs Discovering Latent Knowledge in Language Models Without Supervision

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-09T22:06:35.316057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-09T21:56:32.778349Z digest=sha256:031d77238b3e0258b24bdb980283e126a4b9d0ec8134ccfd68b38d38c70aafa7

Observation 9bbff01e-d216-4a27-9234-b4bd2691c602 · inbound

Dissociating the Internal Representations of Sycophancy in LLMs cites this paper.

Dissociating the Internal Representations of Sycophancy in LLMs Discovering Latent Knowledge in Language Models Without Supervision

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-02T08:11:39.080789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:11:39.080789Z digest=sha256:5206cb950c7d2a86cc325bf9faf8610b81bd9a180fdf962f41cae0e551f93e84

Observation ef966a6b-d11d-4927-a5ff-07acb8c4ad06 · inbound

Prompt Compression via Activation Aggregation cites this paper.

Prompt Compression via Activation Aggregation Discovering Latent Knowledge in Language Models Without Supervision

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-07-10T08:06:57.502233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-07-10T08:03:46.297577Z digest=sha256:83a61914bd56753c4c1983182ff489e70815babe11c2088c8f378a9f8a74b201

Observation 2d105eec-e4fb-4170-aeb5-ce3b335bcbf2 · inbound

Exposure is not manifestation: measurement target and output resolution jointly determine which behavioural-faithfulness evaluator wins cites this paper.

Exposure is not manifestation: measurement target and output resolution jointly determine which behavioural-faithfulness evaluator wins Discovering Latent Knowledge in Language Models Without Supervision

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T07:44:06.996445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:44:06.996445Z digest=sha256:eeeecf2d131b19c23edf4d55cc68397d3d6b5ba56346c82ddf039af98d321c9a

Observation 153b3803-5472-4faf-b276-4b1bb31ccf6d · inbound

The Count Is There, but Misaligned: Understanding and Correcting Counting Failures in VLMs cites this paper.

The Count Is There, but Misaligned: Understanding and Correcting Counting Failures in VLMs Discovering Latent Knowledge in Language Models Without Supervision

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-13T02:15:42.731438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T02:15:42.731438Z digest=sha256:5bdbb741fdebb56f7aa2027da0dae470d1c09b5b1aecd38cb6a84692e4c35f86

Observation 19d968ca-3e05-41c8-a177-2d7064d74d55 · inbound

Confidently Wrong: Detecting Hallucinations in Financial Question Answering from LLM Internal States cites this paper.

Confidently Wrong: Detecting Hallucinations in Financial Question Answering from LLM Internal States Discovering Latent Knowledge in Language Models Without Supervision

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-14T05:45:08.896651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:45:08.896651Z digest=sha256:38d7e6cb7160173aec2287f8d0cff22a68867408d3026316c767fee45ed208d5

Observation 029b3bbe-157e-4230-9e2a-02fd015bb33e · inbound

Inside the Unfair Judge: A Mechanistic Interpretability Account of LLM-as-Judge Bias cites this paper.

Inside the Unfair Judge: A Mechanistic Interpretability Account of LLM-as-Judge Bias Discovering Latent Knowledge in Language Models Without Supervision

Reference 49

Resolution
unresolved
no resolver link, observed 2026-07-14T02:33:34.084111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T02:33:34.084111Z digest=sha256:8211e4a72f8cfb59a3af20a0f6f38505c1ade4bb359944bb0bf6c2400275b618

Observation e0d54499-1682-4ac0-92b3-a092556fbc40 · inbound

The Computational Basis of Confidence in Large Language Models cites this paper.

The Computational Basis of Confidence in Large Language Models Discovering Latent Knowledge in Language Models Without Supervision

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T06:37:40.498943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:37:40.498943Z digest=sha256:1aa8eb18f4176c7586501dca78f010f5af9b8d9f99190fcc9046da9443fcaf86

Observation fc4cf203-9333-48e0-bd82-110f326d7d03 · inbound

The Refusal Residue: When Probes Catch Alignment Faking and When They Don't cites this paper.

The Refusal Residue: When Probes Catch Alignment Faking and When They Don't Discovering Latent Knowledge in Language Models Without Supervision

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T05:33:17.045238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T05:33:17.045238Z digest=sha256:f3536154d9d44504b04e1f9d9fa314d00c64677b081c96661d702e1ad30dff31

Observation d5d3e08a-91a9-4b84-a73b-1f135effb8c3 · inbound

Graded Entity-Familiarity Readouts in Language Models: Polish Adaptation, Cross-Language Robustness, and Refusal Steering cites this paper.

Graded Entity-Familiarity Readouts in Language Models: Polish Adaptation, Cross-Language Robustness, and Refusal Steering Discovering Latent Knowledge in Language Models Without Supervision

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-02T04:52:45.338836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T04:52:45.338836Z digest=sha256:6410e304c8fbff3523c97652db34888ed23a6669f826fbd29f8b23296e5f322d

Observation 6ddf48a2-c533-4b2d-accc-03a15f001caa · inbound

Multi-Head Latent Control: A Unified Interface for LLM Agent Decision Making cites this paper.

Multi-Head Latent Control: A Unified Interface for LLM Agent Decision Making Discovering Latent Knowledge in Language Models Without Supervision

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T02:42:01.203968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T02:42:01.203968Z digest=sha256:6be796f0f4924fee3f77b8167369b583139b8f54f9f9e8395b293d26a29736ec

Observation 2bcafe71-c745-4368-8b24-f0556c70d460 · inbound

The Anatomy of a Truth Direction: Knowledge-Dependent Dimensionality, a Relational Law, and a Convergent Category Geometry in Small Language Models cites this paper.

The Anatomy of a Truth Direction: Knowledge-Dependent Dimensionality, a Relational Law, and a Convergent Category Geometry in Small Language Models Discovering Latent Knowledge in Language Models Without Supervision

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T20:09:48.551226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:09:48.551226Z digest=sha256:ed432a24aca2b5753f8fbab983e0128fdd9ff694c6e6cbc322ae9e6f0c4909b9

Observation fca63d1e-16b8-453c-8966-7f98fb4a1a71 · inbound

Logical Judgments Under Pressure: Diagnosing Syllogistic Stability with Learned Soft Prefixes cites this paper.

Logical Judgments Under Pressure: Diagnosing Syllogistic Stability with Learned Soft Prefixes Discovering Latent Knowledge in Language Models Without Supervision

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T15:40:35.398243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:40:35.398243Z digest=sha256:575c8876f4c9b5b1555240c5743cdc380727bd5b6285567f89b5dd6d450e2e78

Observation 31fcd1be-d249-4209-9f92-08c7bd29507f · inbound

Reading and Steering Representations of Materials-Science Mechanisms in an Open-Weight Language Model cites this paper.

Reading and Steering Representations of Materials-Science Mechanisms in an Open-Weight Language Model Discovering Latent Knowledge in Language Models Without Supervision

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-01T10:59:55.337046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:59:55.337046Z digest=sha256:982f624e0bc5a356b552f0a264bb22011838521506b0fef2a0ebc36865975138

Observation bac76bc1-07cc-4137-a403-25d788855de2 · inbound

Securing Multimodal AI through Internal Information Decomposition cites this paper.

Securing Multimodal AI through Internal Information Decomposition Discovering Latent Knowledge in Language Models Without Supervision

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-02T15:02:58.141704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T15:02:58.141704Z digest=sha256:51367643a806ab67d45453514a5a412f9ee8f2b739e4e52649ec053459e26c62