Pith. sign in

Paper Citation Record · LEDGER

The Superalignment of Superhuman Intelligence with Large Language Models

As of 13 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 0 inbound Pith citation observations for arXiv:2412.11145.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.11145 v2

Coverage vector

measured 29 of 29 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T15:18:17.886907Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

29 of 29 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5ee5ceee-70c8-4a54-9594-ff6f9165c7e9 · outbound

This paper cites Large Language Model Alignment: A Survey.

The Superalignment of Superhuman Intelligence with Large Language Models Large Language Model Alignment: A Survey

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.764452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.764452Z digest=sha256:24abfbaed3fffc79456ceb5577e881a90a44c5a97545996ed1d90a73f311e885

Observation 6a4a4bff-479e-4ce9-a2de-474314a45475 · outbound

This paper cites Statistical Rejection Sampling Improves Preference Optimization.

The Superalignment of Superhuman Intelligence with Large Language Models Statistical Rejection Sampling Improves Preference Optimization

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.773756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.773756Z digest=sha256:d947a4ee45381dfa01bc66e3e264ff4679a3d49a8a3e8e44dcc78ff542de7d4f

Observation ab1c205b-8504-432f-98d1-6245523cdc74 · outbound

This paper cites KTO: Model Alignment as Prospect Theoretic Optimization.

The Superalignment of Superhuman Intelligence with Large Language Models KTO: Model Alignment as Prospect Theoretic Optimization

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.778579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.778579Z digest=sha256:66ac2c823bc51f791653c008aa26301f65b176c4ca14943aba1d6d3e67938d39

Observation d32a8254-b368-4624-95fa-9520e269f644 · outbound

This paper cites Measuring Progress on Scalable Oversight for Large Language Models.

The Superalignment of Superhuman Intelligence with Large Language Models Measuring Progress on Scalable Oversight for Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.783327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.783327Z digest=sha256:59e7d19308bf4bdb93018efeed50ff8693f86b2c688cbc9cf22b2cd3bbb9bd3b

Observation 16b50d46-0267-4805-a866-16ef8e1779be · outbound

This paper cites Supervising strong learners by amplifying weak experts.

The Superalignment of Superhuman Intelligence with Large Language Models Supervising strong learners by amplifying weak experts

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.788272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.788272Z digest=sha256:693e9ecb422ee46a90bad77ea027b8d1d76b61ce5a1fbbd2ce723d16bd705f0b

Observation a28e9c60-0409-4a0e-ade6-96397527604a · outbound

This paper cites Scalable agent alignment via reward modeling: a research direction.

The Superalignment of Superhuman Intelligence with Large Language Models Scalable agent alignment via reward modeling: a research direction

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.792623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.792623Z digest=sha256:96fe5212dca2c0cc4274e46c507d17b5ab738fb88a1cad7694b4a63f4c2cd519

Observation 4f0ea86e-646b-42c1-8dd0-c87bb1a3d15c · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

The Superalignment of Superhuman Intelligence with Large Language Models Constitutional AI: Harmlessness from AI Feedback

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.796886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.796886Z digest=sha256:81c1e507d8e51fb9df8d12b151e538c91013a46465d6f854392682866b58e330

Observation 3dd39d37-f285-4817-9a59-e0666b752e76 · outbound

This paper cites AI safety via debate.

The Superalignment of Superhuman Intelligence with Large Language Models AI safety via debate

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.801169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.801169Z digest=sha256:3e8f15d210204613e552c07356dd13d506620e1b7262f21f6bfc293f7be5f7dc

Observation 59ef8334-a65e-445d-b030-20965e2b20c8 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

The Superalignment of Superhuman Intelligence with Large Language Models Training Verifiers to Solve Math Word Problems

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.810199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.810199Z digest=sha256:d075761f92e5ede344041184eb127c23a1fde0884a683d9851aeb1f90eadff24

Observation b66629ff-e189-4614-bdf9-3f3a95ab5386 · outbound

This paper cites LLM Critics Help Catch LLM Bugs.

The Superalignment of Superhuman Intelligence with Large Language Models LLM Critics Help Catch LLM Bugs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.820149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.820149Z digest=sha256:5e02aa7b05e0b6398eaa8ee9c7b8e9614f36c49d142ccd0fa01a1ac4d3c5b887

Observation 2c67b252-b1f5-4a07-bc7a-bbc985a2c63a · outbound

This paper cites Generative Verifiers: Reward Modeling as Next-Token Prediction.

The Superalignment of Superhuman Intelligence with Large Language Models Generative Verifiers: Reward Modeling as Next-Token Prediction

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.824191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.824191Z digest=sha256:b4d62be3b0f4335857564845b0c5d9464be7d2e3827f2cfad4394b2a3b2f3416

Observation 2708c8af-192b-4a68-93db-24b239e5a1e5 · outbound

This paper cites AutoDetect: Towards a unified framework for automated weakness detection in large language models.

The Superalignment of Superhuman Intelligence with Large Language Models AutoDetect: Towards a unified framework for automated weakness detection in large language models

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:18:18.293578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T15:18:17.829537Z digest=sha256:12e73ca703b8bd6a1aa1e2236f3ee5e4b0f29f094cd1836e92ffbc6489233862

Observation ecd67083-f9ee-4f72-baab-447e193754a4 · outbound

This paper cites Language Models Learn to Mislead Humans via RLHF.

The Superalignment of Superhuman Intelligence with Large Language Models Language Models Learn to Mislead Humans via RLHF

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.834330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.834330Z digest=sha256:cf89c98cd68aeed41de70e10194c489cc72c884bc3564c6913b1198136d0ee73

Observation 1047eca7-37bc-493c-b977-30c9fb9dcd0b · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

The Superalignment of Superhuman Intelligence with Large Language Models Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.838574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.838574Z digest=sha256:32d15d3cb048a19869e7fc3711834179421044b29303628db45d939a943b7b18

Observation 2db461c9-03b7-44bb-9755-b95547d1b5d9 · outbound

This paper cites Unveiling the implicit toxicity in large language models.

The Superalignment of Superhuman Intelligence with Large Language Models Unveiling the implicit toxicity in large language models

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:18:18.279193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T15:18:17.842695Z digest=sha256:dcd075a2382e5e57f3767f3a95f13713f9a97dffd76306172788b9a77a24b559

Observation 2d76fa22-507b-429e-a715-bcfd64e27215 · outbound

This paper cites Improving Reward Models with Synthetic Critiques.

The Superalignment of Superhuman Intelligence with Large Language Models Improving Reward Models with Synthetic Critiques

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.846870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.846870Z digest=sha256:fb8d74fc1bda43fd84db298397332459b10126f53c82cb9fc7088e862e64efcb

Observation d9196234-2fa9-4c5c-9a18-9db170704218 · outbound

This paper cites Reinforcement Learning for Generative AI: A Survey.

The Superalignment of Superhuman Intelligence with Large Language Models Reinforcement Learning for Generative AI: A Survey

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.851686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.851686Z digest=sha256:183b8c54740eb40cffc86b8c124bac27d1640fefcab5313c93b4491f6757ffec

Observation c6b3b56d-02ac-4c2f-a0a9-7bef81be63a4 · outbound

This paper cites Learning to refine with fine-grained natural language feedback.

The Superalignment of Superhuman Intelligence with Large Language Models Learning to refine with fine-grained natural language feedback

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:18:18.265397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T15:18:17.856738Z digest=sha256:2b011274b11af1d05b0c90e50a8c8835a36da09d608d4b28948480bca7987e6d

Observation 81eac04b-19eb-454d-b5f6-73eb1e5e90f8 · outbound

This paper cites Improving Model Factuality with Fine-grained Critique-based Evaluator.

The Superalignment of Superhuman Intelligence with Large Language Models Improving Model Factuality with Fine-grained Critique-based Evaluator

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.860973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.860973Z digest=sha256:3688880479207704418908ff8d38a35a30a3838c026f9e568650c12471429535

Observation 5db97a81-6283-4170-99ba-26eda00bbc34 · outbound

This paper cites Bayesian calibration of win rate estimation with LLM evaluators.

The Superalignment of Superhuman Intelligence with Large Language Models Bayesian calibration of win rate estimation with LLM evaluators

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:18:18.251679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T15:18:17.864978Z digest=sha256:4ac7c3ff7c807ae302e6e1c3f6ae3423785983fdc9bbfb3a31c2f43854bafe3c

Observation 717fbb25-ced1-4ed4-b551-15a149fe3a90 · outbound

This paper cites Trust or Escalate: LLM Judges with Provable Guarantees for Human Agreement.

The Superalignment of Superhuman Intelligence with Large Language Models Trust or Escalate: LLM Judges with Provable Guarantees for Human Agreement

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.868903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.868903Z digest=sha256:d7ef9feb7a943f3d656ad38766a03d9c0c77ea5eb284e721a3c56a89ac1a1be0

Observation dfbecd98-f46b-4f7b-92fb-2dc7d38a88b7 · outbound

This paper cites JudgeLM: Fine-tuned Large Language Models are Scalable Judges.

The Superalignment of Superhuman Intelligence with Large Language Models JudgeLM: Fine-tuned Large Language Models are Scalable Judges

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.873601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.873601Z digest=sha256:f609fc4210682bea3ce23999cdd44b902e9b8c1d7bd0a380f31c8862be837180

Observation f631249e-2027-428f-a169-0f54032ef4b2 · outbound

This paper cites Self-critiquing models for assisting human evaluators.

The Superalignment of Superhuman Intelligence with Large Language Models Self-critiquing models for assisting human evaluators

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.877823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.877823Z digest=sha256:9628ba4e479c040d8a4b53c818de689b3d10b90915ac55f3e85edc4a1297ace3

Observation d28700c0-7ecd-4477-8c5e-6f7da6145206 · outbound

This paper cites Lm vs lm: Detecting factual errors via cross examination.

The Superalignment of Superhuman Intelligence with Large Language Models Lm vs lm: Detecting factual errors via cross examination

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:18:18.237875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T15:18:17.882052Z digest=sha256:6a8b35b39a33fe6bd21b348c63e566773ad60f2029afa904a7e0dbd6f0f711a2

Observation 9b7f6cc6-638b-4e6f-8427-d88ade9038f5 · outbound

This paper cites Panacea: Pareto Alignment via Preference Adaptation for LLMs.

The Superalignment of Superhuman Intelligence with Large Language Models Panacea: Pareto Alignment via Preference Adaptation for LLMs

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.886907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.886907Z digest=sha256:bd06b313ffcfb41fdfef6cc6d1782998e8f083fc7b82d0e0d2596b56c33dc6ff

Observation 9c2a3476-5879-499e-9109-0c8b9ff6b1a9 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

The Superalignment of Superhuman Intelligence with Large Language Models Evaluating Large Language Models Trained on Code

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.815334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.815334Z digest=sha256:e725b4c49bf5da3bafb0c40027c5ecbaa06ad0fb2fc6101ac822b70844f57323

Observation 5c7ea3fc-7624-491e-8f72-99b0ad5f281e · outbound

This paper cites A Survey of Large Language Models.

The Superalignment of Superhuman Intelligence with Large Language Models A Survey of Large Language Models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.753976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.753976Z digest=sha256:8964792deeb4417694ef4b9ebc5d521360b4c24e149c63a77263239daaa8425b

Observation b6ced059-cdeb-46d0-bcca-3f7ffa1b1e4b · outbound

This paper cites Proximal Policy Optimization Algorithms.

The Superalignment of Superhuman Intelligence with Large Language Models Proximal Policy Optimization Algorithms

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.768904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.768904Z digest=sha256:327481431a361e67a707ed99bb364460b44808b7212a9ae7cb0be1695171efae

Observation 9719a814-22c5-47cf-b1ad-ab859268a7d5 · outbound

This paper cites Model evaluation for extreme risks.

The Superalignment of Superhuman Intelligence with Large Language Models Model evaluation for extreme risks

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-11T15:18:17.759688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:18:17.759688Z digest=sha256:6628e823678eac184d5c945e9e97a134ff3c9f3c16953959d31ad7d3f2ff80fb

Pith citing papers

No inbound Pith citation observations are available.