Pith. sign in

Paper Citation Record · LEDGER

Measuring Progress on Scalable Oversight for Large Language Models

As of 25 July 2026, this Paper Citation Record lists 42 of 42 outbound references and 47 inbound Pith citation observations for arXiv:2211.03540.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2211.03540 v2

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-17T15:01:41.161487Z

measured 89 of 89 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-07-25T06:30:59.84592+00:00

measured 47 of 47 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-15T12:59:04.484292Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T18:37:31.166799Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact22
  • verified fuzzy19
  • unresolved0
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e4174cb8-408e-4da0-b551-8d1f93f2ef5b · outbound

This paper cites The case for aligning narrowly superhuman models , url=.

Measuring Progress on Scalable Oversight for Large Language Models The case for aligning narrowly superhuman models , url=

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.420621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:24c2a44596f5b156ff74e10eb52d9fc611c6c8588a011dc230cf9e9c16638b8e

Observation 43f90ad9-77ef-4b3d-8ef8-da54d4337960 · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.387249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:596fff1985eb4259c5267cad4f879d343a500984b7f0478ed4e249455b78812e

Observation 1fc12978-fdcc-4fdf-8f40-3a7f0935865b · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.379940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:11e2250572e4ac3ea44c755ff7a1a5dff256572723fdec562354915dad63d986

Observation e61c81fe-cc1d-41c0-ba01-0f0747937c12 · outbound

This paper cites Weld , journal=.

Measuring Progress on Scalable Oversight for Large Language Models Weld , journal=

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.357536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:1d8bdb6a00895e70df21519b46f8c39165f860c9bec0aa795fb6342818561242

Observation 6c904c71-1b06-4520-a528-11806f77ca97 · outbound

This paper cites Advances in neural information processing systems , volume=.

Measuring Progress on Scalable Oversight for Large Language Models Advances in neural information processing systems , volume=

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.398455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:df38c39c7f4b4d148cbd9ca5ec20db9efbff11cbdac781ac2368a6fb480a8a2f

Observation e38530dd-4572-4342-b936-787ee2a1ad4f · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Measuring Progress on Scalable Oversight for Large Language Models Advances in Neural Information Processing Systems , volume=

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.412471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:ef7265a61b8d0f85b9d5c50ce6840800b4d30e37c9a466e625c9dc583bed407d

Observation 68d621d7-74b2-410e-a62b-39fdb16acce7 · outbound

This paper cites 2014 , isbn =.

Measuring Progress on Scalable Oversight for Large Language Models 2014 , isbn =

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.361299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:86cedf2b9c20071ca4d90a2a908208f505fbec062b8d4cdb27fe43cf3f218842

Observation 7f12dfe8-a82b-4e83-ae87-a8f1fa0e797c · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.364577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:f16f0edcc7f0ae08ae690021cb1b7f6ae2602c6dbae8462e513281c94f9abd78

Observation 0e030b36-4d66-468e-904f-da803314b8ce · outbound

This paper cites Organizational behavior and human performance , volume=.

Measuring Progress on Scalable Oversight for Large Language Models Organizational behavior and human performance , volume=

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.368635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:e1b4ec3e1ca9257e18eccd9f8eeaed903180fae5e31d2ddb906c7d6ac1606581

Observation 56ee0fd4-2e0e-40eb-ab9f-be6117d81079 · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 23

Resolution
parse uncertain
raw_fallback, observed 2026-05-17T15:01:41.372508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:5bb498014ab0bbf30796447bffcc6d9eaac09d5c650aab98b29b5551919a4d0b

Observation ebf72b88-e288-4613-9b41-1c5355bd2a76 · outbound

This paper cites Submitted to The Eleventh International Conference on Learning Representations , year=.

Measuring Progress on Scalable Oversight for Large Language Models Submitted to The Eleventh International Conference on Learning Representations , year=

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.376177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:2934162190cc26633f0d9b5576bad8586f1e70d6b21d11443e35ade763729ce0

Observation 48075c22-a281-43ab-b94b-91eba9ddc028 · outbound

This paper cites Concrete Problems in AI Safety.

Measuring Progress on Scalable Oversight for Large Language Models Concrete Problems in AI Safety

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.251250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:ded62666529b114584f2e40f48ed3f2a85f46522aec7b44687299e028da43c6e

Observation af034535-36a8-4349-aee9-185ad063330c · outbound

This paper cites A General Language Assistant as a Laboratory for Alignment.

Measuring Progress on Scalable Oversight for Large Language Models A General Language Assistant as a Laboratory for Alignment

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.258042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:0906563d542a55feb0770204496f516aa19d1ebfd6d4d61d5ab12cac4136c3d0

Observation 068e9cb5-7109-4594-a000-6428372bbef8 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Measuring Progress on Scalable Oversight for Large Language Models Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.264900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:8092d450d401595ab59a30d4fc93b63f35cc370e4d8891cf8d4cff12c2e8ebc4

Observation 63ac058c-2294-42fd-a11c-5fe3c6dd8d34 · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.391002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:f24c964195ce3799d60b6364adc429f7e29683712da3d6274a9b0e51aed339a6

Observation 0f41ae8f-6573-400f-b128-5808eaa93162 · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.394544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:1af8a76c1609953dd33d9661fe6eb1871863d35550b11cae46eda809c9f0b96d

Observation 770ab2e0-d5bd-4de9-b653-926b1649a85f · outbound

This paper cites Supervising strong learners by amplifying weak experts.

Measuring Progress on Scalable Oversight for Large Language Models Supervising strong learners by amplifying weak experts

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.271749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:987957ec3cdc1e3fb178cc64f9aeeca8ebff628195447577d8eeb59a328ef1f1

Observation d627035e-63ff-4ce6-a56d-6bd6c823a23a · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.403924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:498051f4ce5a875b89be452fc5280f5ddeca14ad80789b66870851fe6a985106

Observation d83291cd-833c-4417-aacd-50548c71d4c3 · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.408201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:d564e8cacc964d715b4130f77cea815da00b52dc023d2d29fedbf873714668c4

Observation 2630f559-5066-4a0a-b0ac-fdb86d473c89 · outbound

This paper cites In: 26th Inter- national Conference on Intelligent User Interfaces, pp.

Measuring Progress on Scalable Oversight for Large Language Models In: 26th Inter- national Conference on Intelligent User Interfaces, pp

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.208363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:1f382801b93ffa34df89ff75d50768b19f7aa283c824c71fea12b0fe525665fa

Observation 2082097a-2392-4c89-b1bc-af4e73fabce2 · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.416665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:9a961ccfbda47cb6db5c2ed09d7a837ddb6c5a83ef505dbd795155abb5f0d9a8

Observation ab1ac030-0194-48be-8b88-645a90aa4394 · outbound

This paper cites Measuring Massive Multitask Language Understanding.

Measuring Progress on Scalable Oversight for Large Language Models Measuring Massive Multitask Language Understanding

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.277357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:5f1571726aa6e7b5611fd203f7b8631eaefd6d890784e59d60eaacda336770f0

Observation 77c61ac6-aff5-4770-b065-45719dbc9a48 · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.345384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:de2e1a67edeb52ab32de1786f8f3e582fcb095f4d88ceeecf928e0667666c8ff

Observation 219991db-4899-45af-b29f-a371f39ce014 · outbound

This paper cites Risks from Learned Optimization in Advanced Machine Learning Systems.

Measuring Progress on Scalable Oversight for Large Language Models Risks from Learned Optimization in Advanced Machine Learning Systems

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.284183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:92fc9511300384b07e5ddf77899d04f37f046194315054f129867739ae9c6a0a

Observation a16fa239-61ad-4915-8a27-069ed6c94bb6 · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.353088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:68857a0262c13447d2c682d5628aa607fb54571540d792ba4cc7b35bc0f8a9a0

Observation 616e3a07-4f6d-4b1a-b950-b71a2e5d4011 · outbound

This paper cites AI safety via debate.

Measuring Progress on Scalable Oversight for Large Language Models AI safety via debate

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.289796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:46fb504069b563459149cb2384b9225727bc58237db79355f2da0c1ad6e14f24

Observation 0c268239-7b4f-4d82-b546-4bdca85a5e61 · outbound

This paper cites Language Models (Mostly) Know What They Know.

Measuring Progress on Scalable Oversight for Large Language Models Language Models (Mostly) Know What They Know

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.295310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:80e8bdd64c337f39fa36cf077dcee59a26d0f3519feddc36f3f4ac1dd5b57c9c

Observation 5d9aba1e-ff68-465f-8c5b-fd7298f3886a · outbound

This paper cites Large Language Models are Zero-Shot Reasoners.

Measuring Progress on Scalable Oversight for Large Language Models Large Language Models are Zero-Shot Reasoners

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.301227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:5ce6f5b112262eb8865d345b5f7b72b841abb11395e1fbc276680979b9329022

Observation bb04b6c5-6c29-4d52-ad20-af9f0fb2442f · outbound

This paper cites Towards a Science of Human-AI Decision Making: A Survey of Empirical Studies.

Measuring Progress on Scalable Oversight for Large Language Models Towards a Science of Human-AI Decision Making: A Survey of Empirical Studies

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.307326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:d0867f0f1dccd2f2f4dae27c7ce218c41a31b25cc61b2e2f78a4a380c55fd70a

Observation a056445e-b407-4ff3-a9f0-721f4559b0bd · outbound

This paper cites Bach, and Jure Leskovec.

Measuring Progress on Scalable Oversight for Large Language Models Bach, and Jure Leskovec

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.233339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:6c5d70ef8a5e0ad936a57429957a9430d2f0a72b2b704040cb0fcca3cc23e367

Observation bc09f87c-8d8a-4b89-85ed-780447feea9e · outbound

This paper cites Scalable agent alignment via reward modeling: a research direction.

Measuring Progress on Scalable Oversight for Large Language Models Scalable agent alignment via reward modeling: a research direction

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.314000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:1a44541fbcac0e7bf369aa6d1c7950c2eb5a811dfc287e454cf2d7b099f45580

Observation bad5145c-a75a-4df8-b393-d975c9763547 · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.349709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:3927db1229c22bbe6dd7baac307b30f26439543911155c01a475cae27410c665

Observation 253a3c6d-02c9-40ab-b123-de61f5ba7ff9 · outbound

This paper cites URLhttps://doi.org/10.18653/v1/2022.acl-long.229.

Measuring Progress on Scalable Oversight for Large Language Models URLhttps://doi.org/10.18653/v1/2022.acl-long.229

Reference 54

Resolution
verified exact
doi, observed 2026-05-17T15:01:41.226048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:fa9d6845619e016168eed5c5fe353836ef1bda9fc575831c32ba605a9316a49d

Observation 51b8fa22-8dd0-4bcd-b1cc-1ac700418a71 · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 55

Resolution
verified exact
doi, observed 2026-05-17T15:01:41.221106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:5204b035c3ab6204c7c4173b1784f6a2d9b6535fb0a96794ba0a993ef36bfcb1

Observation 2c94a520-7597-45c9-96ba-0b3323890368 · outbound

This paper cites Show Your Work: Scratchpads for Intermediate Computation with Language Models.

Measuring Progress on Scalable Oversight for Large Language Models Show Your Work: Scratchpads for Intermediate Computation with Language Models

Reference 56

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.319669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:c8e4f4657c840be5197234b8e3f19a4b409bc38a51d9e9e57dd734e22050b486

Observation 90bfc13f-24b3-4db8-a064-b31d33d74640 · outbound

This paper cites Q u ALITY : Question Answering with Long Input Texts, Yes!.

Measuring Progress on Scalable Oversight for Large Language Models Q u ALITY : Question Answering with Long Input Texts, Yes!

Reference 57

Resolution
verified exact
doi, observed 2026-05-17T15:01:41.215737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:6a3508d35cb22cfa749cd68f5512cd4f08604053477c287dc458bcafbfd77faf

Observation c03a36e3-1208-41f4-9048-799a93ffb05e · outbound

This paper cites Two-Turn Debate Doesn't Help Humans Answer Hard Reading Comprehension Questions.

Measuring Progress on Scalable Oversight for Large Language Models Two-Turn Debate Doesn't Help Humans Answer Hard Reading Comprehension Questions

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.325393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:ea2f0a4c3c8e1d9a73c107f21340784c80515303160f248b3f98cf241318b632

Observation 5c48183f-f710-4029-920c-05f742cf5ff8 · outbound

This paper cites Single-Turn Debate Does Not Help Humans Answer Hard Reading-Comprehension Questions.

Measuring Progress on Scalable Oversight for Large Language Models Single-Turn Debate Does Not Help Humans Answer Hard Reading-Comprehension Questions

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.330513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:233505f937455e9da3fffd510d396408f458fdae32c55a84362ee9d48e922cdc

Observation e3ebd67e-6b71-48ec-9f1b-4f3fc098f26e · outbound

This paper cites Self-critiquing models for assisting human evaluators.

Measuring Progress on Scalable Oversight for Large Language Models Self-critiquing models for assisting human evaluators

Reference 60

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.336065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:405e5f74fa254cb00f92284645306b00c229f0c065edce73fe04c29148916043

Observation 5d3449ba-ecfb-45eb-b5ca-99e8eb70b6d8 · outbound

This paper cites an unresolved cited work.

Measuring Progress on Scalable Oversight for Large Language Models Unresolved cited work

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T15:01:41.384078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:cbe2114f05bd71b26598ea53049c2d56bc57394eeb6fc01ef88c2998fbace187

Observation fe0086a4-287f-4f52-b76b-75c9e0f80dde · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

Measuring Progress on Scalable Oversight for Large Language Models Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 62

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:01:41.341241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:1bcca8f07e7c5dbbb42ffcfc724f2bafce56a3fa23c04c44afe110ebf01e4ab8

Observation ccb6a507-96d6-47ff-926c-880bb6cb9849 · outbound

This paper cites Recursively Summarizing Books with Human Feedback.

Measuring Progress on Scalable Oversight for Large Language Models Recursively Summarizing Books with Human Feedback

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:28:19.819648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T15:01:41.161487Z digest=sha256:45a0e08e2d52bc7f1507e979fce659783f56b2f55c464567911de846fe394ec7

Pith citing papers

Observation 310a7225-23d5-403c-b16f-07a60dbc2a9a · inbound

Measuring Faithfulness in Chain-of-Thought Reasoning cites this paper.

Measuring Faithfulness in Chain-of-Thought Reasoning Measuring Progress on Scalable Oversight for Large Language Models

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-11T20:51:38.390234Z digest=sha256:1be3bacf76c8a84d990afa392ce96a678119dddcdc0416834fda0e2341023a6f

Observation 727345d7-2b3b-41a1-9510-ab5cc62c3528 · inbound

Simple synthetic data reduces sycophancy in large language models cites this paper.

Simple synthetic data reduces sycophancy in large language models Measuring Progress on Scalable Oversight for Large Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-16T14:48:08.508109Z digest=sha256:fa10e979fe0c50bf8f1ba62b823e7362d94e0e605993506ba822e728ca749de1

Observation 677afb01-6fb3-4c8a-9183-1496bbf70eb3 · inbound

Llemma: An Open Language Model For Mathematics cites this paper.

Llemma: An Open Language Model For Mathematics Measuring Progress on Scalable Oversight for Large Language Models

Reference 129

Resolution
verified exact
local_arxiv, observed 2026-05-19T08:17:46.335169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-19T08:17:46.055279Z digest=sha256:3dbd32a6b53ef2e83b9f5ac5c8c8ffaf95dd64b753bb0e845294feaad3c953bc

Observation b8b401bf-361d-43d4-993c-33504914f439 · inbound

Towards Understanding Sycophancy in Language Models cites this paper.

Towards Understanding Sycophancy in Language Models Measuring Progress on Scalable Oversight for Large Language Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-11T06:26:29.196349Z digest=sha256:1ec36d443f5c8d8b5c7a844e8b18c2d94df1c5b426536f10517409935c39f5a3

Observation e7055ce4-1ef6-49b0-a8c9-a3b6dba9cd98 · inbound

A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions cites this paper.

A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions Measuring Progress on Scalable Oversight for Large Language Models

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-13T02:46:26.957539Z digest=sha256:d6862bb77c01d9fe072d28eb874bd1a6af1db77de76428b95985425e58124954

Observation 8601443d-1632-4d69-907c-c0eeff835971 · inbound

TrustLLM: Trustworthiness in Large Language Models cites this paper.

TrustLLM: Trustworthiness in Large Language Models Measuring Progress on Scalable Oversight for Large Language Models

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:17:08.347452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:0490888c0b11146884576e1896b9ad888ae7033b361a2f82eaa1a61e6ed09024

Observation 901d418f-491f-47db-8464-82980c25dbad · inbound

A Roadmap to Pluralistic Alignment cites this paper.

A Roadmap to Pluralistic Alignment Measuring Progress on Scalable Oversight for Large Language Models

Reference 231

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-16T14:37:53.279275Z digest=sha256:d020c347d5793f19a89f536f1c21599a4251010f77082c0ea2970a21973808a2

Observation 2223dc72-1d69-4bda-a082-004977665a4a · inbound

Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models cites this paper.

Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models Measuring Progress on Scalable Oversight for Large Language Models

Reference 108

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-17T14:43:29.496457Z digest=sha256:45f09fbedb7eb457d165e5ae70cf89b71809abcddf61c335ae842c70ac412094

Observation 3e608c7d-1aef-462d-9636-e631bea76776 · inbound

Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models cites this paper.

Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models Measuring Progress on Scalable Oversight for Large Language Models

Reference 267

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T06:38:37.090656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-18T06:38:36.517935Z digest=sha256:081935680dc88c598bc8497704460abd62f42a83f52baf29610b79345d87fa72

Observation bb63dc13-39a3-4e88-93b4-b7a2e01e40bf · inbound

AI Safety Landscape for Large Language Models: Taxonomy, State-of-the-art, and Future Directions cites this paper.

AI Safety Landscape for Large Language Models: Taxonomy, State-of-the-art, and Future Directions Measuring Progress on Scalable Oversight for Large Language Models

Reference 74

Resolution
verified exact
local_arxiv, observed 2026-05-23T21:55:50.388446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-23T21:54:26.670284Z digest=sha256:70c7117d57a0b74f229ce06c0e978cacbeef0b1d754d54406cf42cf4ea10fcfc

Observation 9e5401f7-c3a6-44e1-b232-ae7e53c35dca · inbound

Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation cites this paper.

Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation Measuring Progress on Scalable Oversight for Large Language Models

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-21T07:24:12.970748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-21T07:24:12.845841Z digest=sha256:129bf7f724523b30f9f8177d8a33275888740a0a60f8ce4cb3e132025e5fc8a1

Observation 48e7a18c-07af-44b3-b642-2af018737cc0 · inbound

Benchmarking Misuse Mitigation Against Covert Adversaries cites this paper.

Benchmarking Misuse Mitigation Against Covert Adversaries Measuring Progress on Scalable Oversight for Large Language Models

Reference 71

Resolution
verified exact
local_arxiv, observed 2026-05-19T10:32:14.636622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-19T10:29:05.104520Z digest=sha256:99bad17b9799c8ce3db15759bc2baa2673a9e0567cca1addd8dbf600e802c86c

Observation 3be9b6ac-ece3-479a-b547-8d6ba6fe8cfc · inbound

Red-Bandit: Test-Time Adaptation for LLM Red-Teaming via Bandit-Guided LoRA Experts cites this paper.

Red-Bandit: Test-Time Adaptation for LLM Red-Teaming via Bandit-Guided LoRA Experts Measuring Progress on Scalable Oversight for Large Language Models

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-21T20:54:21.555096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-21T20:53:58.198974Z digest=sha256:6e72272464fc4952263705207ab42e3b0c8f210812a8b0026d31cc43cdb07afc

Observation ca93db08-baeb-405b-b3ec-e9cc1a04fb04 · inbound

Learning When to Trust in Contextual Social Bandits cites this paper.

Learning When to Trust in Contextual Social Bandits Measuring Progress on Scalable Oversight for Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-15T12:59:04.484292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T12:59:04.484292Z digest=sha256:d463abfb8c4db2704aa727fea770f3827bd0977ea3c3bd6efe933f1a760b6e7d

Observation 800a08c5-2321-45c7-838a-1a34b6af4792 · inbound

Extrapolating Volition with Recursive Information Markets cites this paper.

Extrapolating Volition with Recursive Information Markets Measuring Progress on Scalable Oversight for Large Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-10T17:18:34.660366Z digest=sha256:8801d27153c6d69db1b2972c9925fc6806344cf6285e014a713e0889db9714b3

Observation 3665eacb-86e0-463d-b7d0-64c798b4d4bc · inbound

Auditing and Controlling AI Agent Actions in Spreadsheets cites this paper.

Auditing and Controlling AI Agent Actions in Spreadsheets Measuring Progress on Scalable Oversight for Large Language Models

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-10T00:18:56.460027Z digest=sha256:5d7a22f052866770b8e9f8626c3c8447713ad08c98da11817553955154ed2212

Observation 50419c15-4d9e-43c2-949d-5d1335d125af · inbound

Building a Precise Video Language with Human-AI Oversight cites this paper.

Building a Precise Video Language with Human-AI Oversight Measuring Progress on Scalable Oversight for Large Language Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-10T00:37:31.858728Z digest=sha256:c85510eff0d103c7597e25b927128465fb9bcc9f7e7be6fabfe4c2d7e0c732d1

Observation a05e8937-e2ad-401a-9a3e-907f6cbceaf4 · inbound

Structural Enforcement of Goal Integrity in AI Agents via Separation-of-Powers Architecture cites this paper.

Structural Enforcement of Goal Integrity in AI Agents via Separation-of-Powers Architecture Measuring Progress on Scalable Oversight for Large Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-08T06:21:46.364082Z digest=sha256:ef8568ea93ba258c4a18e4a90c7fb0523089e1b48059d81fd262981f9d70c828

Observation b25ff9da-9ea7-4955-b92d-3df7b37fdc41 · inbound

Agentic-imodels: Evolving agentic interpretability tools via autoresearch cites this paper.

Agentic-imodels: Evolving agentic interpretability tools via autoresearch Measuring Progress on Scalable Oversight for Large Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-07T16:37:43.371592Z digest=sha256:491f2d0017ceb4f0ffd7a7838ec849b25e2f1b63c8511c4b5424c1030f499dea

Observation 6b7f1291-242a-4bdd-9a84-c26283012f02 · inbound

Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning cites this paper.

Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning Measuring Progress on Scalable Oversight for Large Language Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-11T00:52:59.406190Z digest=sha256:adf878bfe6594294b48dd78a3172a26b13603121bc5d5063cdbe900d1b515aca

Observation e88a34e5-5099-4b7d-b469-1edb27a7e224 · inbound

Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning cites this paper.

Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning Measuring Progress on Scalable Oversight for Large Language Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-12T02:23:55.323250Z digest=sha256:ba6adbdbd5d995c141041741e9352d14f2740663a120fd99bf607420e08b5a60

Observation 52ba6201-1c33-4492-8663-24ecf82b4e34 · inbound

Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning cites this paper.

Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning Measuring Progress on Scalable Oversight for Large Language Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-13T06:21:00.353334Z digest=sha256:8d8f67741a7ac2428db921405e6ede064bb8e3e709406b6e792690f2f4b849f1

Observation 75d69348-83be-4026-bc55-6aff42c4b851 · inbound

Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning cites this paper.

Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning Measuring Progress on Scalable Oversight for Large Language Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-14T20:55:31.770238Z digest=sha256:8acebc69e76d48204307c40ba61ff4530df8dfd46550a0f1a3c38922b3357c23

Observation 96615e44-6514-4e9a-aaa8-e9727849299b · inbound

Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning cites this paper.

Extracting Search Trees from LLM Reasoning Traces Reveals Myopic Planning Measuring Progress on Scalable Oversight for Large Language Models

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-25T06:00:23.382677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-25T05:57:58.487109Z digest=sha256:8247f66f953c8042c8150a1112638b8318ad863ed95a812d3d2c2b5816e0e944

Observation 6265a63d-b96a-473b-91b3-fdccb56ba11f · inbound

Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight cites this paper.

Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight Measuring Progress on Scalable Oversight for Large Language Models

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-11T00:54:25.549158Z digest=sha256:754c0ceb527e5284f2c7e551a3278f187fa434bcb3c370cb804397aa79b2eb3a

Observation a7e5db1f-8aee-4d85-8204-4e1a350632ed · inbound

Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight cites this paper.

Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight Measuring Progress on Scalable Oversight for Large Language Models

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-05-21T08:29:52.668438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-21T08:29:09.122055Z digest=sha256:2d598201471e68f12b5c31ecdd499372d3ce3ccd535c0d22edf48dfa24c1a78f

Observation a000633a-0eea-4ccf-98db-9324d8d58369 · inbound

LLM Wardens: Mitigating Adversarial Persuasion with Third-Party Conversational Oversight cites this paper.

LLM Wardens: Mitigating Adversarial Persuasion with Third-Party Conversational Oversight Measuring Progress on Scalable Oversight for Large Language Models

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-12T00:51:17.434889Z digest=sha256:2ba83494f800c907114954033525b0bc8d5d6cfd79fcdc378261f3c67721c001

Observation 7e6c204e-c786-4fb9-ac09-3ce41289aba7 · inbound

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces cites this paper.

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces Measuring Progress on Scalable Oversight for Large Language Models

Reference 211

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-05-14T20:17:01.224864Z digest=sha256:82807adae8a664de32d642806396a74011cff63edd4657363bd6f6ca56e42ab5

Observation cbbff684-9dc2-4824-9d4a-d88844acec80 · inbound

Not Just RLHF: Why Alignment Alone Won't Fix Multi-Agent Sycophancy cites this paper.

Not Just RLHF: Why Alignment Alone Won't Fix Multi-Agent Sycophancy Measuring Progress on Scalable Oversight for Large Language Models

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-14T20:04:57.638215Z digest=sha256:5f8d19fac2a05e98927709091c0f6c9069b6ba35a935b5afaa24110d6daeae3d

Observation f17d7638-63a1-41b3-a6bd-51446d02fd38 · inbound

Not Just RLHF: Why Alignment Alone Won't Fix Multi-Agent Sycophancy cites this paper.

Not Just RLHF: Why Alignment Alone Won't Fix Multi-Agent Sycophancy Measuring Progress on Scalable Oversight for Large Language Models

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T21:33:46.582396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-20T21:30:30.384184Z digest=sha256:0ebba66847e2ce047a3829962171b0ba4fb94179e83c06074a42667a6dfe74db

Observation 8053ac46-4d95-4aba-8a75-6439a1ad7f3b · inbound

How to Interpret Agent Behavior cites this paper.

How to Interpret Agent Behavior Measuring Progress on Scalable Oversight for Large Language Models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:01:41.422297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-05-14T18:23:25.269217Z digest=sha256:02c45ff5e01cc19d093a86480e06174d171b39f312799bc1dabece325fd68056

Observation 760313a2-9ef8-4878-92e5-3c63ea3fc415 · inbound

The Behavioral Credibility Trilemma: When Calibrated Autonomy Becomes Impossible cites this paper.

The Behavioral Credibility Trilemma: When Calibrated Autonomy Becomes Impossible Measuring Progress on Scalable Oversight for Large Language Models

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T22:34:01.931399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-29T22:28:02.493124Z digest=sha256:4b193b24d45cb3eeacfa9644228b5e9660e34b22e3afdb81e1ce83b835797a5e

Observation ada9729b-3f19-4131-b615-3ed97031409a · inbound

ReasonOps: Operator Segmentation for LLM Reasoning Traces cites this paper.

ReasonOps: Operator Segmentation for LLM Reasoning Traces Measuring Progress on Scalable Oversight for Large Language Models

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:13:15.919114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-29T08:04:44.270536Z digest=sha256:a27d4ef93068cdcff660a2258cd11e71b3acb0bd86046a59eaee2072cf01eb91

Observation b082e248-ed58-4ef8-9852-9153da2db123 · inbound

Civilizational Metamaterials: Engineering Coordination Under Capability Gradients and Structural Turbulence cites this paper.

Civilizational Metamaterials: Engineering Coordination Under Capability Gradients and Structural Turbulence Measuring Progress on Scalable Oversight for Large Language Models

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-06-28T19:32:34.344128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-28T19:29:13.317260Z digest=sha256:678ef418e983598f84f21f6abb137ca088664bcc641e04fb21b0a0ec184e2596

Observation 888bdece-b8ad-438b-ba15-21aa95d4705e · inbound

When Helping Hurts and How to Fix It: Multi-Agent Debate for Data Cleaning cites this paper.

When Helping Hurts and How to Fix It: Multi-Agent Debate for Data Cleaning Measuring Progress on Scalable Oversight for Large Language Models

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T23:36:23.859141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-28T14:05:49.696342Z digest=sha256:45c52d66617693acb12faaec295f96229a5c8335c707134adf9d56800b0221b3

Observation fee3e933-890f-416d-82ff-4a0a763fc890 · inbound

Sample-Efficient Post-Training for LEGO Spatial-Physics Reasoning cites this paper.

Sample-Efficient Post-Training for LEGO Spatial-Physics Reasoning Measuring Progress on Scalable Oversight for Large Language Models

Reference 61

Resolution
metadata mismatch
local_arxiv, observed 2026-06-28T23:42:49.528693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-06-28T23:38:38.127345Z digest=sha256:69284fc985eaa272a26af10afbebd9fa2ad4c17dc17095de7473033d46c81601

Observation b38cc38a-d24e-42aa-8656-65f8a2ee65ed · inbound

When Behavioral Safety Evaluation Fails: A Representation-Level Perspective cites this paper.

When Behavioral Safety Evaluation Fails: A Representation-Level Perspective Measuring Progress on Scalable Oversight for Large Language Models

Reference 32

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T20:57:23.088762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-06-27T20:04:17.744876Z digest=sha256:5f48ab19a2dcb3173dfdf105ba7f297ccc13e183d1fb134c0a267ece1f514812

Observation 340197aa-715d-48ad-8231-e0ea82a8d34f · inbound

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization cites this paper.

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization Measuring Progress on Scalable Oversight for Large Language Models

Reference 123

Resolution
verified exact
local_arxiv, observed 2026-06-27T16:31:02.718795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-06-27T16:26:34.918099Z digest=sha256:573058c644652747802f88ad76eb9848c61239715980ee514112f627d7af9796

Observation d21ccf50-e626-4102-b7c9-9ccce5a3d31b · inbound

The Arbiter Agent: Continually Monitoring Multi-Agent Conversations to Detect Emergent Misalignment cites this paper.

The Arbiter Agent: Continually Monitoring Multi-Agent Conversations to Detect Emergent Misalignment Measuring Progress on Scalable Oversight for Large Language Models

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-06-27T13:30:56.585818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T13:26:14.457195Z digest=sha256:e3b6713ebf665345849460af2f17f8cc535487efd1305a90b14df275a1e0424a

Observation 3bd557f4-6cc1-46f9-858a-d0aee14bd328 · inbound

HANSEL: Extracting Breadcrumbs from Web Agent Trajectories for Interactive Verification cites this paper.

HANSEL: Extracting Breadcrumbs from Web Agent Trajectories for Interactive Verification Measuring Progress on Scalable Oversight for Large Language Models

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-04T02:09:22.850175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-26T19:58:23.151980Z digest=sha256:50c942d441599bdea43e33bbf0b94ba3fbc0ee2937bd55a613b8041860ba1cef

Observation 03ece39d-ccb5-4cef-bbaa-cb848f64945b · inbound

Regulating AI: Where U.S. State Policy and HCI (Mis)align cites this paper.

Regulating AI: Where U.S. State Policy and HCI (Mis)align Measuring Progress on Scalable Oversight for Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-12T03:28:32.737784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T03:28:32.737784Z digest=sha256:3436a4b65f73bd9ed2b0ffa31f9ab8c25f042f220bf6b6e2c78b54dc5099a7dc

Observation 58f09301-6d3f-449b-852e-96cf1488e19c · inbound

Attention Limited Reward Learning cites this paper.

Attention Limited Reward Learning Measuring Progress on Scalable Oversight for Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-11T16:47:52.768236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T16:47:52.768236Z digest=sha256:f5413da3c1d34966a5171b9468f9188c0abc818ea18234fcb90878b8f641a1e0

Observation 6af3d5c1-69af-47a7-b4d8-1a4627d85359 · inbound

Weak-to-Strong Generalization via Direct On-Policy Distillation cites this paper.

Weak-to-Strong Generalization via Direct On-Policy Distillation Measuring Progress on Scalable Oversight for Large Language Models

Reference 85

Resolution
verified exact
local_arxiv, observed 2026-07-07T12:33:44.994534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-07-07T12:31:42.224094Z digest=sha256:784daec9bc36f693be05d246633ff8d6b647138b3663ac04efbab76acadb18c1

Observation efc4af27-2ee3-4a29-9127-9092046319d2 · inbound

Weak-to-Strong Generalization via Direct On-Policy Distillation cites this paper.

Weak-to-Strong Generalization via Direct On-Policy Distillation Measuring Progress on Scalable Oversight for Large Language Models

Reference 81

Resolution
unresolved
no resolver link, observed 2026-07-11T07:01:56.628017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T07:01:56.628017Z digest=sha256:51323521d198f3141f2200bc49efd988746884b1c59a64b16598279ff71ddbd9

Observation d08d1af3-c0a9-48b4-9d10-950748120de9 · inbound

Measuring Intelligence Beyond Human Scale cites this paper.

Measuring Intelligence Beyond Human Scale Measuring Progress on Scalable Oversight for Large Language Models

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-09T21:26:34.379823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-07-09T21:21:39.305904Z digest=sha256:3e971e20e57a9083fb3bbdf1fbb4df3d7c7250a90e94e56f2c925986060971d8

Observation 63bcabe1-a474-4187-a067-0c06b3a8122b · inbound

Institutional Red-Teaming: Deployment Rules, Not Just Models, Causally Shape Multi-Agent AI Safety cites this paper.

Institutional Red-Teaming: Deployment Rules, Not Just Models, Causally Shape Multi-Agent AI Safety Measuring Progress on Scalable Oversight for Large Language Models

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-09T02:25:55.820061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=arxiv_source observed=2026-07-09T02:17:34.473738Z digest=sha256:ea91080bd86efbb520bd7e8c5494ebc250b3b93fe6f99a8d7bcbe7a415d08577

Observation 7d790367-a3a2-4cfc-9264-d510447bdbd1 · inbound

ScopeJudge: Cost-Aware Pre-Execution Gating for Offensive Security Agents cites this paper.

ScopeJudge: Cost-Aware Pre-Execution Gating for Offensive Security Agents Measuring Progress on Scalable Oversight for Large Language Models

Reference 14

Resolution
metadata mismatch
local_arxiv, observed 2026-07-10T18:37:31.168237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-07-10T18:29:50.731038Z digest=sha256:f75dccf9d7f888e51e1d6bda8205c4a432ea4bf59f6edd3946440af1e2266bf3