Pith. sign in

Paper Citation Record · LEDGER

Measuring Progress on Scalable Oversight for Large Language Models

As of 22 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 78 inbound Pith citation observations for arXiv:2211.03540.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2211.03540 v2

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-17T15:01:41.161487Z

measured 120 of 120 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 78 of 78 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:44:02.926095Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact22
  • verified fuzzy7
  • unresolved12
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

32
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation e4174cb8-408e-4da0-b551-8d1f93f2ef5b · outbound

This paper cites The case for aligning narrowly superhuman models , url=.

Measuring Progress on Scalable Oversight for Large Language Models The case for aligning narrowly superhuman models , url=

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.420621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:1ddad771f2fb66b8b7edb5ab5db52c2e3bf1314764860fcea9d9d50c53790288

Observation 43f90ad9-77ef-4b3d-8ef8-da54d4337960 · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-05-17T15:01:41.387249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:9930f02a4a8615799bde904ff2badfe3e2bbb1b00f47fdb3d51a9d3bfdf4a9f7

Observation 1fc12978-fdcc-4fdf-8f40-3a7f0935865b · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-05-17T15:01:41.379940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:dc3ad6e65e98f650f222d2521aa38b174ddc7ebfb6c783764c403fb90fa1b4b2

Observation e61c81fe-cc1d-41c0-ba01-0f0747937c12 · outbound

This paper cites Weld , journal=.

Measuring Progress on Scalable Oversight for Large Language Models Weld , journal=

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.357536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:6fba1fd7322d6f29fb1e12441e1a75cc93f70c3fd980bc40ead063c1180896ff

Observation 6c904c71-1b06-4520-a528-11806f77ca97 · outbound

This paper cites Advances in neural information processing systems , volume=.

Measuring Progress on Scalable Oversight for Large Language Models Advances in neural information processing systems , volume=

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.398455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:916c72c2bcdd79fbb5de3d4524aacb63ce6fd506bf4ff682ac70cfe78d2d6e0f

Observation e38530dd-4572-4342-b936-787ee2a1ad4f · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Measuring Progress on Scalable Oversight for Large Language Models Advances in Neural Information Processing Systems , volume=

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.412471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:59badca5b1656c5a7f1215e3381196ce233d0d71db90ee74f16eb9158ca48dfb

Observation 68d621d7-74b2-410e-a62b-39fdb16acce7 · outbound

This paper cites 2014 , isbn =.

Measuring Progress on Scalable Oversight for Large Language Models 2014 , isbn =

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.361299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:df49d170d6d487958528a0217886b5573e9dfbc1ec7606667bbfece20a409380

Observation 7f12dfe8-a82b-4e83-ae87-a8f1fa0e797c · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-05-17T15:01:41.364577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:c192ad5c9cf84c0ef8339670add2b69475e332d29cdc7df460ed82570272d307

Observation 0e030b36-4d66-468e-904f-da803314b8ce · outbound

This paper cites Organizational behavior and human performance , volume=.

Measuring Progress on Scalable Oversight for Large Language Models Organizational behavior and human performance , volume=

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.368635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:fee4d1059be1e6972876d057ee5ea942974cd76391def74117ee9545c846cff3

Observation 56ee0fd4-2e0e-40eb-ab9f-be6117d81079 · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 23

Resolution
parse uncertain
raw_fallback, observed 2026-05-17T15:01:41.372508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:1ceff89ddd4fec1413d2d923df59af2b88de72c37a8db5018c41d45de4aced32

Observation ebf72b88-e288-4613-9b41-1c5355bd2a76 · outbound

This paper cites Submitted to The Eleventh International Conference on Learning Representations , year=.

Measuring Progress on Scalable Oversight for Large Language Models Submitted to The Eleventh International Conference on Learning Representations , year=

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.376177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:8950cc81c813e6800dc6f4301342092e90f6c4dbab125db3a7e39d70a77b7697

Observation 48075c22-a281-43ab-b94b-91eba9ddc028 · outbound

This paper cites Concrete Problems in AI Safety.

Measuring Progress on Scalable Oversight for Large Language Models Concrete Problems in AI Safety

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.251250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:2ddaf6e33c83799f8b749d2c5a891841689909268af1628447206a35a2ea6fa7

Observation af034535-36a8-4349-aee9-185ad063330c · outbound

This paper cites A General Language Assistant as a Laboratory for Alignment.

Measuring Progress on Scalable Oversight for Large Language Models A General Language Assistant as a Laboratory for Alignment

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.258042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:3740569bf218c911ce72a72aa9fd1478493ecf9604a982c41a1a93fe92795976

Observation 068e9cb5-7109-4594-a000-6428372bbef8 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Measuring Progress on Scalable Oversight for Large Language Models Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.264900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:47e7b4781fc070b2e5f2ccb9da5862d1005ec9cd93a080f04782a3143cdac0a7

Observation 63ac058c-2294-42fd-a11c-5fe3c6dd8d34 · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-05-17T15:01:41.391002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:a4985bf4504a960c6ea56790d955a64ad378de4b87fadbfa4a01c235ac9e082d

Observation 0f41ae8f-6573-400f-b128-5808eaa93162 · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-05-17T15:01:41.394544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:c482719272bc56eb31898209dd166285c843b754782e678dd7f4341ab6bc1bfc

Observation 770ab2e0-d5bd-4de9-b653-926b1649a85f · outbound

This paper cites Supervising strong learners by amplifying weak experts.

Measuring Progress on Scalable Oversight for Large Language Models Supervising strong learners by amplifying weak experts

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.271749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:dbcae4811a8ab344089483d7cf520a41c0eeb29492c46e73ab32d78b2c596b7a

Observation d627035e-63ff-4ce6-a56d-6bd6c823a23a · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-05-17T15:01:41.403924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:bb296d23b4c50e73818b8baf7fe6757abcddf43033b9a8c00c8821e3808e8b9e

Observation d83291cd-833c-4417-aacd-50548c71d4c3 · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-05-17T15:01:41.408201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:d622028bfc8c1afd20f5556eb30eb8cf50403cc3ec22fea612c6176aa5293e9b

Observation 2630f559-5066-4a0a-b0ac-fdb86d473c89 · outbound

This paper cites In: 26th Inter- national Conference on Intelligent User Interfaces, pp.

Measuring Progress on Scalable Oversight for Large Language Models In: 26th Inter- national Conference on Intelligent User Interfaces, pp

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.208363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:f1ffb4ff42da0c68d9a13ffcfb2ef6eef3ee2da2336d978d46b0386084f30318

Observation 2082097a-2392-4c89-b1bc-af4e73fabce2 · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-05-17T15:01:41.416665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:3069f2c7a9958493db156edd25d8facb3095fcdf083a29a6b16997cea527a371

Observation ab1ac030-0194-48be-8b88-645a90aa4394 · outbound

This paper cites Measuring Massive Multitask Language Understanding.

Measuring Progress on Scalable Oversight for Large Language Models Measuring Massive Multitask Language Understanding

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.277357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:9096a68fd573181c9ce63618b6411279be66147547532170a92c55ffa033e13b

Observation 77c61ac6-aff5-4770-b065-45719dbc9a48 · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-05-17T15:01:41.345384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:a7545936e0976894b933fc9b80b7ddf2be2ceb40a93cfbba69903767314ac2ee

Observation 219991db-4899-45af-b29f-a371f39ce014 · outbound

This paper cites Risks from Learned Optimization in Advanced Machine Learning Systems.

Measuring Progress on Scalable Oversight for Large Language Models Risks from Learned Optimization in Advanced Machine Learning Systems

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.284183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:571591b244c149eefe0ebb7fce548605c4c9f30c19b91359b24498ebed3e6e9f

Observation a16fa239-61ad-4915-8a27-069ed6c94bb6 · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-05-17T15:01:41.353088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:881675c05369a5c2dc0e8fa8b58424251de0286eaaa33b195a556334f9bb0aae

Observation 616e3a07-4f6d-4b1a-b950-b71a2e5d4011 · outbound

This paper cites AI safety via debate.

Measuring Progress on Scalable Oversight for Large Language Models AI safety via debate

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.289796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:ef6a1fa00bf631608f59bdc2b94c956c5f2dc66917196f30194ef7c4acb82bb7

Observation 0c268239-7b4f-4d82-b546-4bdca85a5e61 · outbound

This paper cites Language Models (Mostly) Know What They Know.

Measuring Progress on Scalable Oversight for Large Language Models Language Models (Mostly) Know What They Know

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.295310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:498e0df562c3a872f82102fada6558ff59fed8121e115ab2f4f481e412a8349b

Observation 5d9aba1e-ff68-465f-8c5b-fd7298f3886a · outbound

This paper cites Large Language Models are Zero-Shot Reasoners.

Measuring Progress on Scalable Oversight for Large Language Models Large Language Models are Zero-Shot Reasoners

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.301227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:09ba1972b60924a29ab001d3d38d303158c90b584d2cdf0ad56ec1064f076b7b

Observation bb04b6c5-6c29-4d52-ad20-af9f0fb2442f · outbound

This paper cites Towards a Science of Human-AI Decision Making: A Survey of Empirical Studies.

Measuring Progress on Scalable Oversight for Large Language Models Towards a Science of Human-AI Decision Making: A Survey of Empirical Studies

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.307326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:094338dbb58e3bc63f950facbe2127b8dab902f687d626be778c8cb335c34310

Observation a056445e-b407-4ff3-a9f0-721f4559b0bd · outbound

This paper cites Bach, and Jure Leskovec.

Measuring Progress on Scalable Oversight for Large Language Models Bach, and Jure Leskovec

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.233339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:c99a4ba777d609f7105b6db903eeb4e3a876633087c08fb244d4e5df54a754a0

Observation bc09f87c-8d8a-4b89-85ed-780447feea9e · outbound

This paper cites Scalable agent alignment via reward modeling: a research direction.

Measuring Progress on Scalable Oversight for Large Language Models Scalable agent alignment via reward modeling: a research direction

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.314000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:d836f5ffeb22a2eb33c9ddd9379d49871fadb1a9b01ae17f5ee4d892e111122a

Observation bad5145c-a75a-4df8-b393-d975c9763547 · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-05-17T15:01:41.349709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:d839a76c668af487936c192b058f8c0c3abb072fe9f194af623323ec1489e318

Observation 253a3c6d-02c9-40ab-b123-de61f5ba7ff9 · outbound

This paper cites URLhttps://doi.org/10.18653/v1/2022.acl-long.229.

Measuring Progress on Scalable Oversight for Large Language Models URLhttps://doi.org/10.18653/v1/2022.acl-long.229

Reference 54

Resolution
verified exact
doi, observed 2026-05-17T15:01:41.226048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:94a195ff8e4e49ca66f674725e7f0e8190b64ba93d681b0e81b613b2e2eb25fd

Observation 51b8fa22-8dd0-4bcd-b1cc-1ac700418a71 · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 55

Resolution
verified exact
doi, observed 2026-05-17T15:01:41.221106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:56a91217730f766ba96ebf6d722ec40e8e450632f82d6882f5876635cc90a584

Observation 2c94a520-7597-45c9-96ba-0b3323890368 · outbound

This paper cites Show Your Work: Scratchpads for Intermediate Computation with Language Models.

Measuring Progress on Scalable Oversight for Large Language Models Show Your Work: Scratchpads for Intermediate Computation with Language Models

Reference 56

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.319669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:10d681193c874c4cc5f36db668ebe83a80bb2a55e813d6a5bd7ac28707bd8a5c

Observation 90bfc13f-24b3-4db8-a064-b31d33d74640 · outbound

This paper cites Q u ALITY : Question Answering with Long Input Texts, Yes!.

Measuring Progress on Scalable Oversight for Large Language Models Q u ALITY : Question Answering with Long Input Texts, Yes!

Reference 57

Resolution
verified exact
doi, observed 2026-05-17T15:01:41.215737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:7638589e56fedc2e85f394091aab8506a4a6f3a4222916eadb170189870bd536

Observation c03a36e3-1208-41f4-9048-799a93ffb05e · outbound

This paper cites Two-Turn Debate Doesn't Help Humans Answer Hard Reading Comprehension Questions.

Measuring Progress on Scalable Oversight for Large Language Models Two-Turn Debate Doesn't Help Humans Answer Hard Reading Comprehension Questions

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.325393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:b1b7d549d8d950cd7a5287dbb8e1a68ca75da47576ff2fb7065487108d5fccc1

Observation 5c48183f-f710-4029-920c-05f742cf5ff8 · outbound

This paper cites Single-Turn Debate Does Not Help Humans Answer Hard Reading-Comprehension Questions.

Measuring Progress on Scalable Oversight for Large Language Models Single-Turn Debate Does Not Help Humans Answer Hard Reading-Comprehension Questions

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.330513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:00480e756c7599fef02fe1d6dc421919144ba20bf72f90979fdc514e70470038

Observation e3ebd67e-6b71-48ec-9f1b-4f3fc098f26e · outbound

This paper cites Self-critiquing models for assisting human evaluators.

Measuring Progress on Scalable Oversight for Large Language Models Self-critiquing models for assisting human evaluators

Reference 60

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.336065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:8d786ec1313b5947a2dd1946007349c7a4dbe82ac0a55826fcd0263a0371371a

Observation 5d3449ba-ecfb-45eb-b5ca-99e8eb70b6d8 · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-05-17T15:01:41.384078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:65ae6080a2c34bcd7ee712081883625ee50aa834fd33c0d95eaac6c678cc461a

Observation fe0086a4-287f-4f52-b76b-75c9e0f80dde · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

Measuring Progress on Scalable Oversight for Large Language Models Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 62

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.341241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:e857ca6eb008d647961c2fc6bd45768dcb5ed7c7eee5877b21606f60c387c183

Observation ccb6a507-96d6-47ff-926c-880bb6cb9849 · outbound

This paper cites Recursively Summarizing Books with Human Feedback.

Measuring Progress on Scalable Oversight for Large Language Models Recursively Summarizing Books with Human Feedback

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:28:19.819648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:1a038112514f3c2eb212d989e8822c5ffc6cf76a607a3b48a00ce0795887ea0e

Pith citing papers

Observation 310a7225-23d5-403c-b16f-07a60dbc2a9a · inbound

Measuring Faithfulness in Chain-of-Thought Reasoning cites this paper.

Measuring Faithfulness in Chain-of-Thought Reasoning Measuring Progress on Scalable Oversight for Large Language Models

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-11T20:51:38.390234Z digest=sha256:748263fbcf0b690114a1e26ea0f31da38ebf7730dfb47c86f05b2c5909c3e00d

Observation 727345d7-2b3b-41a1-9510-ab5cc62c3528 · inbound

Simple synthetic data reduces sycophancy in large language models cites this paper.

Simple synthetic data reduces sycophancy in large language models Measuring Progress on Scalable Oversight for Large Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-16T14:48:08.508109Z digest=sha256:10a65b7a2d299a112104ecc2cec6acceaf6bc5c131c3b0bc83d353095cd7b9d0

Observation 677afb01-6fb3-4c8a-9183-1496bbf70eb3 · inbound

Llemma: An Open Language Model For Mathematics cites this paper.

Llemma: An Open Language Model For Mathematics Measuring Progress on Scalable Oversight for Large Language Models

Reference 129

Resolution
verified exact
local_arxiv, observed 2026-05-19T08:17:46.335169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-19T08:17:46.055279Z digest=sha256:36bed8a46476a6013c194a5eb560ff8c66ead870f8055be05b112b72d171bd5d

Observation b8b401bf-361d-43d4-993c-33504914f439 · inbound

Towards Understanding Sycophancy in Language Models cites this paper.

Towards Understanding Sycophancy in Language Models Measuring Progress on Scalable Oversight for Large Language Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-11T06:26:29.196349Z digest=sha256:6642ae1b8aeeda3146a551947d54068b32e84934c9de1d2869d50824db0f36db

Observation e7055ce4-1ef6-49b0-a8c9-a3b6dba9cd98 · inbound

A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions cites this paper.

A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions Measuring Progress on Scalable Oversight for Large Language Models

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-13T02:46:26.957539Z digest=sha256:91f5c82a01967e2069db3c8fc8378346175bf3cd472324f6e5bc15d7365fb037

Observation 8601443d-1632-4d69-907c-c0eeff835971 · inbound

TrustLLM: Trustworthiness in Large Language Models cites this paper.

TrustLLM: Trustworthiness in Large Language Models Measuring Progress on Scalable Oversight for Large Language Models

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:17:08.347452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:2f9910639f424d249f308b2b0ca8c321bed9c55e143cd7b6e4bf2e0808c50d13

Observation 901d418f-491f-47db-8464-82980c25dbad · inbound

A Roadmap to Pluralistic Alignment cites this paper.

A Roadmap to Pluralistic Alignment Measuring Progress on Scalable Oversight for Large Language Models

Reference 231

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-16T14:37:53.279275Z digest=sha256:0ba41a32bed3b2b37388fe98cfc91cd7955b79655588cc7e2631ac69ba0a3d58

Observation 2223dc72-1d69-4bda-a082-004977665a4a · inbound

Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models cites this paper.

Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models Measuring Progress on Scalable Oversight for Large Language Models

Reference 108

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T14:43:29.496457Z digest=sha256:fc7560b038c275725e2fe8c9ee025c03db40c5bab50b898f4cf364a87f3c766a

Observation 3e608c7d-1aef-462d-9636-e631bea76776 · inbound

Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models cites this paper.

Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models Measuring Progress on Scalable Oversight for Large Language Models

Reference 267

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T06:38:37.090656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-18T06:38:36.517935Z digest=sha256:b2711bdb2bd40ea19d5b8b909e95477fcbd9da3376188480481488832c4b0535

Observation bb63dc13-39a3-4e88-93b4-b7a2e01e40bf · inbound

AI Safety Landscape for Large Language Models: Taxonomy, State-of-the-art, and Future Directions cites this paper.

AI Safety Landscape for Large Language Models: Taxonomy, State-of-the-art, and Future Directions Measuring Progress on Scalable Oversight for Large Language Models

Reference 74

Resolution
verified exact
local_arxiv, observed 2026-05-23T21:55:50.388446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-23T21:54:26.670284Z digest=sha256:8f6786a8a42e9c4cfce7001a85f3e282c0fe88a4db79e7baba0ecaa53c985de7

Observation c4551775-ebc5-4bce-950f-c5d9e3dcf410 · inbound

Preference Optimization for Reasoning with Pseudo Feedback cites this paper.

Preference Optimization for Reasoning with Pseudo Feedback Measuring Progress on Scalable Oversight for Large Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-12T13:21:05.529270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:21:05.529270Z digest=sha256:e526f8d514002e3fcfbf2b594f1a15fb96faae1421f75d83349a0191022a9862

Observation 14a6d641-dcb1-4ee2-95f4-e69ed53ebdf8 · inbound

Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision cites this paper.

Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision Measuring Progress on Scalable Oversight for Large Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T13:02:57.168900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:02:57.168900Z digest=sha256:36c49bc76669f63c3897320a991a77ee2c74f94c49db84fe3d024b02bcbb98c1

Observation bfec8ed9-71c9-4079-821b-9a38feb8ac92 · inbound

Can an AI Agent Safely Run a Government? Existence of Probably Approximately Aligned Policies cites this paper.

Can an AI Agent Safely Run a Government? Existence of Probably Approximately Aligned Policies Measuring Progress on Scalable Oversight for Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T15:45:49.894196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:45:49.894196Z digest=sha256:987a6c58433d1983df30cc42966c015984d030381d14addab9ee18176d69d0f6

Observation 20742aa9-b3f4-421b-8794-8abe8b5f63a3 · inbound

ProcessBench: Identifying Process Errors in Mathematical Reasoning cites this paper.

ProcessBench: Identifying Process Errors in Mathematical Reasoning Measuring Progress on Scalable Oversight for Large Language Models

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-11T19:37:21.503550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:37:21.503550Z digest=sha256:40fbefca92b78ce679bd41cb95a2e9342ba0a4444522621c72f6ced8003c4d42

Observation d32a8254-b368-4624-95fa-9520e269f644 · inbound

The Superalignment of Superhuman Intelligence with Large Language Models cites this paper.

The Superalignment of Superhuman Intelligence with Large Language Models Measuring Progress on Scalable Oversight for Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.783327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.783327Z digest=sha256:94261aef8c17d337b479293ad0fac66ee44dc0f11676b5514a8da9652ea772d0

Observation 236f959f-c4e5-4f57-8810-375533c6a15a · inbound

Towards AI-$45^{\circ}$ Law: A Roadmap to Trustworthy AGI cites this paper.

Towards AI-$45^{\circ}$ Law: A Roadmap to Trustworthy AGI Measuring Progress on Scalable Oversight for Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T20:13:07.041367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:13:07.041367Z digest=sha256:16320f440d991bce219422f56109f2582efeea4218e70e1574b8eb31aa959d38

Observation c5871bfc-dba2-466e-92b8-e35000899a04 · inbound

The Road to Artificial SuperIntelligence: A Comprehensive Survey of Superalignment cites this paper.

The Road to Artificial SuperIntelligence: A Comprehensive Survey of Superalignment Measuring Progress on Scalable Oversight for Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T10:36:17.106477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T10:36:17.106477Z digest=sha256:d3fa3ccd5fba6a90ca7e4175432f68bbc801833c9cbff6fa281dd126237000e0

Observation f72d5970-6bf3-4fb9-bedc-9c2f0c23309a · inbound

Lies, Damned Lies, and Distributional Language Statistics: Persuasion and Deception with Large Language Models cites this paper.

Lies, Damned Lies, and Distributional Language Statistics: Persuasion and Deception with Large Language Models Measuring Progress on Scalable Oversight for Large Language Models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-11T05:49:34.644435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:49:34.644435Z digest=sha256:f979c4871008f25ce41216306043f19b20719646a2a02f5a51a2bfef0d7b3f76

Observation 0909d486-5115-43f5-9e71-8c6f749473e3 · inbound

Governing AI Agents cites this paper.

Governing AI Agents Measuring Progress on Scalable Oversight for Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T20:33:38.593994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:33:38.593994Z digest=sha256:637d28a15deafb9f9bf92cd87136523cef877b6bc0cd26063b6cb5660a5b76d5

Observation 75af3184-376c-4a08-a8e3-6b1f7060ada3 · inbound

Debate Helps Weak-to-Strong Generalization cites this paper.

Debate Helps Weak-to-Strong Generalization Measuring Progress on Scalable Oversight for Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T17:50:56.287233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:50:56.287233Z digest=sha256:fafa54c90c7ccdef03c909ec56eb124afde45cc4bb65837e6a48192a83cc10a9

Observation da23bdd6-e48f-4ab9-af14-37dbdf74fe4a · inbound

Automated Capability Discovery via Foundation Model Self-Exploration cites this paper.

Automated Capability Discovery via Foundation Model Self-Exploration Measuring Progress on Scalable Oversight for Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T12:17:55.144202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T12:17:55.144202Z digest=sha256:491ec4048762f3410e2f2a1860aed4e40c90512a2e3fff47ba215d555111d85a

Observation dde3f0b1-fc9c-4a43-8dc6-79aff4c43612 · inbound

Standardizing Intelligence: Aligning Generative AI for Regulatory and Operational Compliance cites this paper.

Standardizing Intelligence: Aligning Generative AI for Regulatory and Operational Compliance Measuring Progress on Scalable Oversight for Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T15:05:24.048766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:05:24.048766Z digest=sha256:297b1482b0605d0feb18754862e740087d817513e5289837fcfcde429c5c0174

Observation 9e5401f7-c3a6-44e1-b232-ae7e53c35dca · inbound

Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation cites this paper.

Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation Measuring Progress on Scalable Oversight for Large Language Models

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-21T07:24:12.970748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-21T07:24:12.845841Z digest=sha256:231dd72632fd37da94cd4913acf98f97352fdeb19e148239d1789d6a0ba52dbd

Observation a86f4835-6b31-4184-9d06-8fc9cb6b7139 · inbound

Super Co-alignment of Human and AI for Sustainable Symbiotic Society cites this paper.

Super Co-alignment of Human and AI for Sustainable Symbiotic Society Measuring Progress on Scalable Oversight for Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T10:44:02.926095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:44:02.926095Z digest=sha256:837931992ef3dcc817f31a40cce0c3fa4cfe367f62a05a9fcd11056e80461af1

Observation 97709941-4e7e-403b-8b0e-d25b482973ae · inbound

DeepCritic: Deliberate Critique with Large Language Models cites this paper.

DeepCritic: Deliberate Critique with Large Language Models Measuring Progress on Scalable Oversight for Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T04:42:07.324494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:42:07.324494Z digest=sha256:4b86f45e1321def0c311ee8de6d7dcfdda30110f88d2d4db87715d2809dd159b

Observation a405042c-10b5-4bb6-a6f2-d243d2ed560e · inbound

Understanding LLM Scientific Reasoning through Promptings and Model's Explanation on the Answers cites this paper.

Understanding LLM Scientific Reasoning through Promptings and Model's Explanation on the Answers Measuring Progress on Scalable Oversight for Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-16T04:25:04.351605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:25:04.351605Z digest=sha256:4d23ffd7d195e90082aa4ed159a0c67eb02aaec37f3e2531924f5b24e7ace463

Observation 8b6b4ab2-641c-4bc8-b438-0f2efaac871f · inbound

What Is AI Safety? What Do We Want It to Be? cites this paper.

What Is AI Safety? What Do We Want It to Be? Measuring Progress on Scalable Oversight for Large Language Models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-16T01:01:29.062041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T01:01:29.062041Z digest=sha256:81042698ffca519ca1bb895ff2e1789c6ed4b931172afeffd363d61a9ccb6a53

Observation aa68ffa3-cc7d-4d91-b5c8-109bf5f91d26 · inbound

An alignment safety case sketch based on debate cites this paper.

An alignment safety case sketch based on debate Measuring Progress on Scalable Oversight for Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T23:43:46.401209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:43:46.401209Z digest=sha256:76e3c62e43721eb0be5bda4e0d079677346fb95331256eab02addde7543202d6

Observation e35d027e-89f9-4d94-82ad-e3bb9aee02a8 · inbound

When Models Know More Than They Can Explain: Quantifying Knowledge Transfer in Human-AI Collaboration cites this paper.

When Models Know More Than They Can Explain: Quantifying Knowledge Transfer in Human-AI Collaboration Measuring Progress on Scalable Oversight for Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T10:20:06.236092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:20:06.236092Z digest=sha256:060c53bda120cf34864e6f057d4f1deaebdb70a0c946db938ffdcf5e4a08c287

Observation 48e7a18c-07af-44b3-b642-2af018737cc0 · inbound

Benchmarking Misuse Mitigation Against Covert Adversaries cites this paper.

Benchmarking Misuse Mitigation Against Covert Adversaries Measuring Progress on Scalable Oversight for Large Language Models

Reference 71

Resolution
verified exact
local_arxiv, observed 2026-05-19T10:32:14.636622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-19T10:29:05.104520Z digest=sha256:9a100f74bf3ad51e7915daaa47db7757e60ab809cad9faea1e282d4f978c425e

Observation 329a6965-01d5-4c3c-85e1-d48057160d83 · inbound

Out of Control -- Why Alignment Needs Formal Control Theory (and an Alignment Control Stack) cites this paper.

Out of Control -- Why Alignment Needs Formal Control Theory (and an Alignment Control Stack) Measuring Progress on Scalable Oversight for Large Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T23:26:57.492998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:26:57.492998Z digest=sha256:a7581c3b181eae9682293e1b38c997cdbe6c46a7e43b11e20000fa4cd9ea0ef2

Observation 1bde6da0-f4f3-4a8c-9f28-1afffd67c247 · inbound

Teaching Models to Verbalize Reward Hacking in Chain-of-Thought Reasoning cites this paper.

Teaching Models to Verbalize Reward Hacking in Chain-of-Thought Reasoning Measuring Progress on Scalable Oversight for Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T22:03:51.583235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:03:51.583235Z digest=sha256:cbe050808c75bdf46e4e4b6960aaedff5b5f35cffd82fdf77964974302bb4a67

Observation 5ab5e77a-ee45-4595-9196-49652ab17900 · inbound

Architecting Human-AI Cocreation for Technical Services -- Interaction Modes and Contingency Factors cites this paper.

Architecting Human-AI Cocreation for Technical Services -- Interaction Modes and Contingency Factors Measuring Progress on Scalable Oversight for Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T16:14:52.848885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:14:52.848885Z digest=sha256:430eebecce4f07cc3c295bc3b5cab9e5f9824a7ff691c57994cdb75992d824ba

Observation 23d23f02-0625-4652-9ef7-650e7721348c · inbound

ADEPTS: A Capability Framework for Human-Centered Agent Design cites this paper.

ADEPTS: A Capability Framework for Human-Centered Agent Design Measuring Progress on Scalable Oversight for Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T16:11:46.697917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:11:46.697917Z digest=sha256:514e1d96aac7f7adba09faa67a291211fc8258de357a3d8d1f5b2b8f66f4efa1

Observation c10a1c5d-7669-4c92-aa99-9091e123183d · inbound

Reliable Weak-to-Strong Monitoring of LLM Agents cites this paper.

Reliable Weak-to-Strong Monitoring of LLM Agents Measuring Progress on Scalable Oversight for Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T15:53:47.104785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:53:47.104785Z digest=sha256:bb4dd9ca5eb4450a26be7ea34dc4d97d1707440247b585f5dad895191604a4b5

Observation 17980314-f36b-4b8d-b475-9a95744ea767 · inbound

ACE and Diverse Generalization via Selective Disagreement cites this paper.

ACE and Diverse Generalization via Selective Disagreement Measuring Progress on Scalable Oversight for Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T21:33:19.273101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T21:33:19.273101Z digest=sha256:a536daab1ca391a112f05aabbfef52370086660412666a95bb2a907ba7670027

Observation 507cc51e-c822-4262-8068-1acda3b91d92 · inbound

Certifiable Safe RLHF: Semantic Grounding and Fixed Penalty Constraint Optimization for Safer LLM Alignment cites this paper.

Certifiable Safe RLHF: Semantic Grounding and Fixed Penalty Constraint Optimization for Safer LLM Alignment Measuring Progress on Scalable Oversight for Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T12:27:28.649935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:27:28.649935Z digest=sha256:531d58dd1c7332bcaaf64951acf7171580db84e531d8f5826cc2d2c0f91b4887

Observation 3be9b6ac-ece3-479a-b547-8d6ba6fe8cfc · inbound

Red-Bandit: Test-Time Adaptation for LLM Red-Teaming via Bandit-Guided LoRA Experts cites this paper.

Red-Bandit: Test-Time Adaptation for LLM Red-Teaming via Bandit-Guided LoRA Experts Measuring Progress on Scalable Oversight for Large Language Models

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-21T20:54:21.555096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-21T20:53:58.198974Z digest=sha256:2fdcd0c18422c6337cea2b0ee25131028c4983727c54e1ddca70c8cf2061e4bb

Observation ca93db08-baeb-405b-b3ec-e9cc1a04fb04 · inbound

Learning When to Trust in Contextual Social Bandits cites this paper.

Learning When to Trust in Contextual Social Bandits Measuring Progress on Scalable Oversight for Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-15T12:59:04.484292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T12:59:04.484292Z digest=sha256:d35af2193de5db1509600cf3f35e51210812ee7fc6f91186d4e47b2aaef5d728

Observation 800a08c5-2321-45c7-838a-1a34b6af4792 · inbound

Extrapolating Volition with Recursive Information Markets cites this paper.

Extrapolating Volition with Recursive Information Markets Measuring Progress on Scalable Oversight for Large Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T17:18:34.660366Z digest=sha256:c72e2aa7b786ef235efcf8acfdcdc3266b2095b229a58344af90242ce07a2f92

Observation 3665eacb-86e0-463d-b7d0-64c798b4d4bc · inbound

Auditing and Controlling AI Agent Actions in Spreadsheets cites this paper.

Auditing and Controlling AI Agent Actions in Spreadsheets Measuring Progress on Scalable Oversight for Large Language Models

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T00:18:56.460027Z digest=sha256:9975b9c0253a00d23aa85b28b1e5ad225a279f17fb103f04eece7cde5bd4ee5f

Observation 50419c15-4d9e-43c2-949d-5d1335d125af · inbound

Building a Precise Video Language with Human-AI Oversight cites this paper.

Building a Precise Video Language with Human-AI Oversight Measuring Progress on Scalable Oversight for Large Language Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T00:37:31.858728Z digest=sha256:9db40afffb3fa969622e2491f9b0b14e136f8fa17ce0e2aa37c63b408e83161f

Observation a05e8937-e2ad-401a-9a3e-907f6cbceaf4 · inbound

Structural Enforcement of Goal Integrity in AI Agents via Separation-of-Powers Architecture cites this paper.

Structural Enforcement of Goal Integrity in AI Agents via Separation-of-Powers Architecture Measuring Progress on Scalable Oversight for Large Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-08T06:21:46.364082Z digest=sha256:5953fddd3d7d86ab01d75dcd29ba03f040f739ef4202f21cecf520569e568d1e

Observation b25ff9da-9ea7-4955-b92d-3df7b37fdc41 · inbound

Agentic-imodels: Evolving agentic interpretability tools via autoresearch cites this paper.

Agentic-imodels: Evolving agentic interpretability tools via autoresearch Measuring Progress on Scalable Oversight for Large Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T16:37:43.371592Z digest=sha256:a271b9fa7345b5b52a72b75013733119d9bd7a9dc6161c192b518876422a4a5a

Observation 6b7f1291-242a-4bdd-9a84-c26283012f02 · inbound

Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning cites this paper.

Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning Measuring Progress on Scalable Oversight for Large Language Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-11T00:52:59.406190Z digest=sha256:c3d81051c914329770d029a40ffa5a1a5f6ab2a0bebfdad68024b0daf1cb7441

Observation e88a34e5-5099-4b7d-b469-1edb27a7e224 · inbound

Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning cites this paper.

Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning Measuring Progress on Scalable Oversight for Large Language Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-12T02:23:55.323250Z digest=sha256:b7df3fd328b408d17d71bbd233c2432e158ec4484a4d8274c240dbc1bf71e3dc

Observation 52ba6201-1c33-4492-8663-24ecf82b4e34 · inbound

Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning cites this paper.

Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning Measuring Progress on Scalable Oversight for Large Language Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-13T06:21:00.353334Z digest=sha256:9362d5bd7054b2e15901309fa4aa2ce6aa5ada0fe4ef0be2626f9c8f142f1bfd

Observation 75d69348-83be-4026-bc55-6aff42c4b851 · inbound

Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning cites this paper.

Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning Measuring Progress on Scalable Oversight for Large Language Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-14T20:55:31.770238Z digest=sha256:72b3db980e4c27436d644eb5a4229160a4bc92b985fc5c775f1ecec229d9de87

Observation 96615e44-6514-4e9a-aaa8-e9727849299b · inbound

Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning cites this paper.

Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning Measuring Progress on Scalable Oversight for Large Language Models

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-25T06:00:23.382677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-25T05:57:58.487109Z digest=sha256:6b0f4edb75943c60fcf935dc9c683b0157c768142126086df1af54f7d8587c34

Observation 6265a63d-b96a-473b-91b3-fdccb56ba11f · inbound

Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight cites this paper.

Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight Measuring Progress on Scalable Oversight for Large Language Models

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-11T00:54:25.549158Z digest=sha256:33752ec960244ccec3bb48ac7518b5bcf86dbcdb6c2d7a34dd9e3325a336b901

Observation a7e5db1f-8aee-4d85-8204-4e1a350632ed · inbound

Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight cites this paper.

Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight Measuring Progress on Scalable Oversight for Large Language Models

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-05-21T08:29:52.668438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-21T08:29:09.122055Z digest=sha256:20c28cbc3c6318054d85d5610c2fab91599bbebe681af14c1bcf3b069697718d

Observation a000633a-0eea-4ccf-98db-9324d8d58369 · inbound

LLM Wardens: Mitigating Adversarial Persuasion with Third-Party Conversational Oversight cites this paper.

LLM Wardens: Mitigating Adversarial Persuasion with Third-Party Conversational Oversight Measuring Progress on Scalable Oversight for Large Language Models

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-12T00:51:17.434889Z digest=sha256:042aa91beec0a2e5445d2a7f5c42e18551b40132458a9fda1423b91d2c8a36eb

Observation 7e6c204e-c786-4fb9-ac09-3ce41289aba7 · inbound

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces cites this paper.

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces Measuring Progress on Scalable Oversight for Large Language Models

Reference 211

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-14T20:17:01.224864Z digest=sha256:532d9ea55412c57a66bbec69af5e870363b9b65c2f7183fce860e1ebc4610ddf

Observation cbbff684-9dc2-4824-9d4a-d88844acec80 · inbound

Not Just RLHF: Why Alignment Alone Won't Fix Multi-Agent Sycophancy cites this paper.

Not Just RLHF: Why Alignment Alone Won't Fix Multi-Agent Sycophancy Measuring Progress on Scalable Oversight for Large Language Models

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-14T20:04:57.638215Z digest=sha256:8204359e5b0328a983a20ad3a05929b024c63863f51792223e546230f9e5d3e2

Observation f17d7638-63a1-41b3-a6bd-51446d02fd38 · inbound

Not Just RLHF: Why Alignment Alone Won't Fix Multi-Agent Sycophancy cites this paper.

Not Just RLHF: Why Alignment Alone Won't Fix Multi-Agent Sycophancy Measuring Progress on Scalable Oversight for Large Language Models

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T21:33:46.582396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-20T21:30:30.384184Z digest=sha256:17d1341d2c97a17a98b45787f83ee86f7c7aa5510744e597c8a188c6fc9913ba

Observation 8053ac46-4d95-4aba-8a75-6439a1ad7f3b · inbound

How to Interpret Agent Behavior cites this paper.

How to Interpret Agent Behavior Measuring Progress on Scalable Oversight for Large Language Models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-14T18:23:25.269217Z digest=sha256:a81f55b459e6f28ad64dcde87c8e3d21cedf0178341e7c715bf1fd5bedf4207b

Observation 760313a2-9ef8-4878-92e5-3c63ea3fc415 · inbound

The Behavioral Credibility Trilemma: When Calibrated Autonomy Becomes Impossible cites this paper.

The Behavioral Credibility Trilemma: When Calibrated Autonomy Becomes Impossible Measuring Progress on Scalable Oversight for Large Language Models

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T22:34:01.931399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T22:28:02.493124Z digest=sha256:174e38c8c5a0a9466472b39ba9d2819f39d113557071f04ac9b97fcef909ca42

Observation efa228ea-f0fe-4621-a88a-75d94cf44080 · inbound

The Behavioral Credibility Trilemma: When Calibrated Autonomy Becomes Impossible cites this paper.

The Behavioral Credibility Trilemma: When Calibrated Autonomy Becomes Impossible Measuring Progress on Scalable Oversight for Large Language Models

Reference 1997

Resolution
unresolved
no resolver link, observed 2026-08-02T13:17:56.253641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:17:56.253641Z digest=sha256:fc778be09f2fd1892cd6c1b589aca4da2463f9c6ab22f1f8944b1bc3f50108b0

Observation ada9729b-3f19-4131-b615-3ed97031409a · inbound

ReasonOps: Operator Segmentation for LLM Reasoning Traces cites this paper.

ReasonOps: Operator Segmentation for LLM Reasoning Traces Measuring Progress on Scalable Oversight for Large Language Models

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:13:15.919114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T08:04:44.270536Z digest=sha256:77fad90832a6f1a86503d3af1070c135385d3891a3f3184f8e14240c12c6004d

Observation b082e248-ed58-4ef8-9852-9153da2db123 · inbound

Civilizational Metamaterials: Engineering Coordination Under Capability Gradients and Structural Turbulence cites this paper.

Civilizational Metamaterials: Engineering Coordination Under Capability Gradients and Structural Turbulence Measuring Progress on Scalable Oversight for Large Language Models

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-06-28T19:32:34.344128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-28T19:29:13.317260Z digest=sha256:d4c12b05513f949bd46a32858e76e8a6e874a223e786a8521b1f6bf073d17fd9

Observation 888bdece-b8ad-438b-ba15-21aa95d4705e · inbound

When Helping Hurts and How to Fix It: Multi-Agent Debate for Data Cleaning cites this paper.

When Helping Hurts and How to Fix It: Multi-Agent Debate for Data Cleaning Measuring Progress on Scalable Oversight for Large Language Models

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T23:36:23.859141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-28T14:05:49.696342Z digest=sha256:1aa3188d623e06a5a0a43d8816c3e08bd731f3fa57e3c9254771780bc7b672aa

Observation fee3e933-890f-416d-82ff-4a0a763fc890 · inbound

Sample-Efficient Post-Training for LEGO Spatial-Physics Reasoning cites this paper.

Sample-Efficient Post-Training for LEGO Spatial-Physics Reasoning Measuring Progress on Scalable Oversight for Large Language Models

Reference 61

Resolution
metadata mismatch
local_arxiv, observed 2026-06-28T23:42:49.528693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-28T23:38:38.127345Z digest=sha256:c2d158f50ef741c422e03bf04c351e1c21cf288aa7d52dbb71ab0bb7fc66d7ba

Observation b38cc38a-d24e-42aa-8656-65f8a2ee65ed · inbound

When Behavioral Safety Evaluation Fails: A Representation-Level Perspective cites this paper.

When Behavioral Safety Evaluation Fails: A Representation-Level Perspective Measuring Progress on Scalable Oversight for Large Language Models

Reference 32

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T20:57:23.088762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-27T20:04:17.744876Z digest=sha256:b47fdee3c4beeb1dddb64e654a0e457cbd111a9ba9533c4caaf3978a63f95a50

Observation 340197aa-715d-48ad-8231-e0ea82a8d34f · inbound

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization cites this paper.

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization Measuring Progress on Scalable Oversight for Large Language Models

Reference 123

Resolution
verified exact
local_arxiv, observed 2026-06-27T16:31:02.718795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-27T16:26:34.918099Z digest=sha256:4765a507dc4fcc2901ad37a43336f708ae09ee5b664a88dedb89fec22e780aa9

Observation d21ccf50-e626-4102-b7c9-9ccce5a3d31b · inbound

The Arbiter Agent: Continually Monitoring Multi-Agent Conversations to Detect Emergent Misalignment cites this paper.

The Arbiter Agent: Continually Monitoring Multi-Agent Conversations to Detect Emergent Misalignment Measuring Progress on Scalable Oversight for Large Language Models

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-06-27T13:30:56.585818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-27T13:26:14.457195Z digest=sha256:82c395b43edd25ca68ebf017372961dd0aee9ddad3eda5a24c28cc8d36576bd6

Observation 3bd557f4-6cc1-46f9-858a-d0aee14bd328 · inbound

HANSEL: Extracting Breadcrumbs from Web Agent Trajectories for Interactive Verification cites this paper.

HANSEL: Extracting Breadcrumbs from Web Agent Trajectories for Interactive Verification Measuring Progress on Scalable Oversight for Large Language Models

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-04T02:09:22.850175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-26T19:58:23.151980Z digest=sha256:d7ee5dcab18def84a65699148f49d75980fc3fd3b8a5cd8d32174facad5ad8ce

Observation 03ece39d-ccb5-4cef-bbaa-cb848f64945b · inbound

Regulating AI: Where U.S. State Policy and HCI (Mis)align cites this paper.

Regulating AI: Where U.S. State Policy and HCI (Mis)align Measuring Progress on Scalable Oversight for Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-12T03:28:32.737784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T03:28:32.737784Z digest=sha256:870cac7352d7f23177ce62b67354f9530a86228ae51b77317276ebc66185ec74

Observation 58f09301-6d3f-449b-852e-96cf1488e19c · inbound

Attention Limited Reward Learning cites this paper.

Attention Limited Reward Learning Measuring Progress on Scalable Oversight for Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-11T16:47:52.768236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T16:47:52.768236Z digest=sha256:fb96de94399594d8e011c2418f7d32b5d263d6ebb39a5fb523f9a55be1b174ac

Observation 6af3d5c1-69af-47a7-b4d8-1a4627d85359 · inbound

Weak-to-Strong Generalization via Direct On-Policy Distillation cites this paper.

Weak-to-Strong Generalization via Direct On-Policy Distillation Measuring Progress on Scalable Oversight for Large Language Models

Reference 85

Resolution
verified exact
local_arxiv, observed 2026-07-07T12:33:44.994534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-07T12:31:42.224094Z digest=sha256:5bf4b12744b428a3c2a3aba86b9d5fa1dbc8f729ea8b3a5314637953155efdc0

Observation efc4af27-2ee3-4a29-9127-9092046319d2 · inbound

Weak-to-Strong Generalization via Direct On-Policy Distillation cites this paper.

Weak-to-Strong Generalization via Direct On-Policy Distillation Measuring Progress on Scalable Oversight for Large Language Models

Reference 81

Resolution
unresolved
no resolver link, observed 2026-07-11T07:01:56.628017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T07:01:56.628017Z digest=sha256:011e4319fa1f59521f13fc4c1fead1edf38184e310518c985d0e4f5a972b5f5a

Observation d08d1af3-c0a9-48b4-9d10-950748120de9 · inbound

Measuring Intelligence Beyond Human Scale cites this paper.

Measuring Intelligence Beyond Human Scale Measuring Progress on Scalable Oversight for Large Language Models

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-09T21:26:34.379823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-09T21:21:39.305904Z digest=sha256:61ee4527f8ab0c5adc8b7561988bc4df006724932312d2df76fa1fd38e2b8df2

Observation 63bcabe1-a474-4187-a067-0c06b3a8122b · inbound

Institutional Red-Teaming: Deployment Rules, Not Just Models, Causally Shape Multi-Agent AI Safety cites this paper.

Institutional Red-Teaming: Deployment Rules, Not Just Models, Causally Shape Multi-Agent AI Safety Measuring Progress on Scalable Oversight for Large Language Models

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-09T02:25:55.820061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-09T02:17:34.473738Z digest=sha256:ed0459ad13e9e21508384f107a981e55647b0d3f83eef3c9e5cf0b7a4df30bd2

Observation 7d790367-a3a2-4cfc-9264-d510447bdbd1 · inbound

ScopeJudge: Cost-Aware Pre-Execution Gating for Offensive Security Agents cites this paper.

ScopeJudge: Cost-Aware Pre-Execution Gating for Offensive Security Agents Measuring Progress on Scalable Oversight for Large Language Models

Reference 14

Resolution
metadata mismatch
local_arxiv, observed 2026-07-10T18:37:31.168237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-10T18:29:50.731038Z digest=sha256:dbfb162420807bcbcd19e0276759ad9f3661d7ffc9cc5505313c464c9e904327

Observation 699c0921-412f-408c-af41-6900b077ad48 · inbound

ScopeJudge: Cost-Aware Pre-Execution Gating for Offensive Security Agents cites this paper.

ScopeJudge: Cost-Aware Pre-Execution Gating for Offensive Security Agents Measuring Progress on Scalable Oversight for Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T00:49:48.020275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T00:49:48.020275Z digest=sha256:bdb4786bc77b6d2d182f8fd9f6aa610c7233331e4cfa9f12f6a94cd588777980

Observation 9267c136-32a2-45fd-bbcc-21459b50cce3 · inbound

StealthBench: Measuring Operational Stealth in Autonomous Offensive-Security Agents cites this paper.

StealthBench: Measuring Operational Stealth in Autonomous Offensive-Security Agents Measuring Progress on Scalable Oversight for Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T00:13:38.052542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:13:38.052542Z digest=sha256:0197f0879f6452d7a24bfdabd7f9f87b771f395312d9cc9d841ab7073c47fbb0

Observation b6873bfa-1982-4cee-a43a-d17336c645ae · inbound

Knowing When to Quit: Diagnosing and Training LLMs to Abort Futile Reasoning cites this paper.

Knowing When to Quit: Diagnosing and Training LLMs to Abort Futile Reasoning Measuring Progress on Scalable Oversight for Large Language Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-03T11:36:49.183413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T11:36:49.183413Z digest=sha256:fdefae3583a809ec8cca760bc5c1f851d2093ddc198fc95b73b4c03fded20fbb

Observation 02a26895-fad4-47a8-8063-427b7b116ce1 · inbound

Sharding Prevents LLM Oversight Failures and Adversarial Exploitation cites this paper.

Sharding Prevents LLM Oversight Failures and Adversarial Exploitation Measuring Progress on Scalable Oversight for Large Language Models

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-10T04:33:13.310081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T04:33:13.310081Z digest=sha256:cfae915aae675e7c5514987ccb1f0a15acb6a6f5cb76d9113f76adf6d2a9b3d6

Observation 7aacf96d-82b4-4b27-aabd-7c71d9a5ae21 · inbound

Training AI Scientists to Replicate Research cites this paper.

Training AI Scientists to Replicate Research Measuring Progress on Scalable Oversight for Large Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-14T13:35:09.603186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:35:09.603186Z digest=sha256:3dbd4dcf1656e2c7ec39ff79cc13b765fd715476a438f4c2013e925777ca7232