Pith. sign in

Paper Citation Record · LEDGER

Aligning Large Language Models with Human: A Survey

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 93 inbound Pith citation observations for arXiv:2307.12966.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.12966 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 93 of 93 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 93 of 93 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:35:47.382195Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

54
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b6d139e9-2d51-4f0d-a1be-8ede276d2ae3 · inbound

A Comprehensive Overview of Large Language Models cites this paper.

A Comprehensive Overview of Large Language Models Aligning Large Language Models with Human: A Survey

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:28:39.510881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-19T20:28:38.900026Z digest=sha256:2336b45bbded4c42805c68f97330788c97a16b142f4bbb75b97f43674844da87

Observation dfce7db8-59cb-4178-b50c-57a452b04b22 · inbound

A Survey on Large Language Model based Autonomous Agents cites this paper.

A Survey on Large Language Model based Autonomous Agents Aligning Large Language Models with Human: A Survey

Reference 177

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:03:01.108498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-15T04:03:00.340349Z digest=sha256:3a2a3dc2192fbebc30c0708ac1c3b56151b2364c6690efe84b2f7dba8449c3b3

Observation 943025cd-c6eb-4bf7-be99-7a2e8cdc359f · inbound

Baichuan 2: Open Large-scale Language Models cites this paper.

Baichuan 2: Open Large-scale Language Models Aligning Large Language Models with Human: A Survey

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-24T06:54:03.668248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-24T06:51:02.531751Z digest=sha256:582c4edf97939d4ee977748bb066ce88d15fbbcc86abbce8c79f846a293defb3

Observation 1ecf09ce-be35-4e4d-98f4-c3dc51f76fb1 · inbound

A Survey on Knowledge Distillation of Large Language Models cites this paper.

A Survey on Knowledge Distillation of Large Language Models Aligning Large Language Models with Human: A Survey

Reference 50

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T23:31:11.557716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-17T23:31:11.213552Z digest=sha256:313e2d049cec8ba2442afab8faa6b6d44f1e84a207c13500a0f1b0c5a3837952

Observation 94ded0ab-ef04-47b5-974d-4ffe0b261b88 · inbound

A Survey on the Memory Mechanism of Large Language Model based Agents cites this paper.

A Survey on the Memory Mechanism of Large Language Model based Agents Aligning Large Language Models with Human: A Survey

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:21:39.695484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-15T07:21:39.440092Z digest=sha256:ba8e8d4f5e28bc4f2be650fee36f34c0606a4ac2a741e119f8fdafcff0c86fca

Observation 7b9e8966-4b40-4226-a713-1a016562f5af · inbound

A Survey on Large Language Models for Code Generation cites this paper.

A Survey on Large Language Models for Code Generation Aligning Large Language Models with Human: A Survey

Reference 279

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:18:06.781203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-13T20:18:06.304134Z digest=sha256:24e738ec74bc7d4afd6faa17f23304adca7067c2131a56955055423cf5a4458d

Observation d195a577-dc7f-4dc9-bf04-853b4077e306 · inbound

PSPO*: An Effective Process-supervised Policy Optimization for Reasoning Alignment cites this paper.

PSPO*: An Effective Process-supervised Policy Optimization for Reasoning Alignment Aligning Large Language Models with Human: A Survey

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T18:20:14.025705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T18:20:14.025705Z digest=sha256:204860d03d65134c5874d81db323f2cb629924739094c25aa303e2d6c218fcd6

Observation 43d4b163-c4ba-45ed-beb8-06a5d4369a6a · inbound

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation cites this paper.

Preference Alignment for Diffusion Model via Explicit Denoised Distribution Estimation Aligning Large Language Models with Human: A Survey

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T14:54:02.519982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:54:02.519982Z digest=sha256:8f6dcd6aea76bac3273ea858798abe06f8eb747aeabc4432b864fd8d349f4c8c

Observation ddda3e63-bb9e-410a-bfaf-eeed711a9384 · inbound

Large Language Model-Brained GUI Agents: A Survey cites this paper.

Large Language Model-Brained GUI Agents: A Survey Aligning Large Language Models with Human: A Survey

Reference 268

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:08:27.878552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-19T11:08:27.472508Z digest=sha256:0bd68e8d8575158cfd8190f14fa888ab2b5e29fe9da06b16876289230de993e7

Observation 272b5568-a605-4c1d-beba-12883454d05f · inbound

Alignment at Pre-training! Towards Native Alignment for Arabic LLMs cites this paper.

Alignment at Pre-training! Towards Native Alignment for Arabic LLMs Aligning Large Language Models with Human: A Survey

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T22:41:16.599664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:41:16.599664Z digest=sha256:3978ce2cd2c8011c15ed873970989cd4c686e0b8c2217fd05c7d3705c8aa6bc2

Observation ad2522f3-e8fe-427f-b5d2-ffa0294e636e · inbound

Practical Considerations for Agentic LLM Systems cites this paper.

Practical Considerations for Agentic LLM Systems Aligning Large Language Models with Human: A Survey

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-11T21:49:09.612383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:49:09.612383Z digest=sha256:71d32f5e195bff67a692160f5c6b4b05628aa80673a757e120509f22bf0d1a6d

Observation 3f1d1fbc-d14c-4bbe-bcbb-c089c1fd6b98 · inbound

MAG-V: A Multi-Agent Framework for Synthetic Data Generation and Verification cites this paper.

MAG-V: A Multi-Agent Framework for Synthetic Data Generation and Verification Aligning Large Language Models with Human: A Survey

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T10:20:15.925641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T10:20:15.925641Z digest=sha256:8eab7ff7c9c504e64ac47c898c626c26b9cb5f1452ec6f9bac74577738a01b4e

Observation 4cf7b096-3905-4aa3-93c8-d2bef147eb84 · inbound

Creating an LLM-based AI-agent: A high-level methodology towards enhancing LLMs with APIs cites this paper.

Creating an LLM-based AI-agent: A high-level methodology towards enhancing LLMs with APIs Aligning Large Language Models with Human: A Survey

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T13:38:04.366729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:38:04.366729Z digest=sha256:db3b9c69997ecb30dbcd87609d1fdfec03fca028f630b51773b30ce145734297

Observation f4e4fe56-7535-424b-b328-10d7df17daf0 · inbound

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs cites this paper.

JailPO: A Novel Black-box Jailbreak Framework via Preference Optimization against Aligned LLMs Aligning Large Language Models with Human: A Survey

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T11:17:49.074644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:17:49.074644Z digest=sha256:c36817e055642e859d393ff6252b4a639362598e10bf670c492fe750c6420fa0

Observation d87c7098-f32e-49c3-ae90-033dfcc11395 · inbound

Sim911: Towards Effective and Equitable 9-1-1 Dispatcher Training with an LLM-Enabled Simulation cites this paper.

Sim911: Towards Effective and Equitable 9-1-1 Dispatcher Training with an LLM-Enabled Simulation Aligning Large Language Models with Human: A Survey

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T10:19:36.603444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:19:36.603444Z digest=sha256:fd3d5932c0ec1ecd8cc0ad9391ece138be9077b61abe80dc1b7426e8102607c7

Observation fec29158-51f2-487a-acd8-66b38529a439 · inbound

ACECode: A Reinforcement Learning Framework for Aligning Code Efficiency and Correctness in Code Language Models cites this paper.

ACECode: A Reinforcement Learning Framework for Aligning Code Efficiency and Correctness in Code Language Models Aligning Large Language Models with Human: A Survey

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T05:44:25.372969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:44:25.372969Z digest=sha256:944273654b00f3111ca522dfae036c32308921a8ba5ca7c8f63bbb2a3da51f66

Observation fba8892d-4e71-4bd4-8f40-9fa3d6560746 · inbound

DiffusionAttacker: Diffusion-Driven Prompt Manipulation for LLM Jailbreak cites this paper.

DiffusionAttacker: Diffusion-Driven Prompt Manipulation for LLM Jailbreak Aligning Large Language Models with Human: A Survey

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T05:31:31.682237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:31:31.682237Z digest=sha256:6f62168deed5d7b32f53ea68a196dff162254bdd2e1cf3b0f76ede780fa93e36

Observation db527423-608b-4f3c-9344-f5c8b43910df · inbound

Emerging Security Challenges of Large Language Models cites this paper.

Emerging Security Challenges of Large Language Models Aligning Large Language Models with Human: A Survey

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T05:23:44.469457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:23:44.469457Z digest=sha256:6790ebc15e4f52ee133a9693d52089e934a69f468c593f50658c541a4caf8c5a

Observation 284ccd89-8ef2-440e-bfc6-ee470834b2dd · inbound

Understanding the Logic of Direct Preference Alignment through Logic cites this paper.

Understanding the Logic of Direct Preference Alignment through Logic Aligning Large Language Models with Human: A Survey

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T05:24:21.038122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:24:21.038122Z digest=sha256:46f8462f7f0c5fbdda1f6e0693102952f9a5c6d3bedcc99bedec1fe84db97140

Observation 24539c06-227e-423c-ba79-bcff85c264fa · inbound

RapGuard: Safeguarding Multimodal Large Language Models via Rationale-aware Defensive Prompting cites this paper.

RapGuard: Safeguarding Multimodal Large Language Models via Rationale-aware Defensive Prompting Aligning Large Language Models with Human: A Survey

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T04:29:04.334021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:29:04.334021Z digest=sha256:dcea4e910c6ab17604ac991a34f0255abf9c5759a7379f74f8363ea1a290fc02

Observation be4b0a9b-b903-43ee-95a2-6f6ae249bd13 · inbound

Enhancing LLM Reasoning with Multi-Path Collaborative Reactive and Reflection agents cites this paper.

Enhancing LLM Reasoning with Multi-Path Collaborative Reactive and Reflection agents Aligning Large Language Models with Human: A Survey

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T22:54:57.943854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:54:57.943854Z digest=sha256:a16c17b42b075a47034c2a4e148ac519ae79994b784b986ca3ca9958bf366482

Observation 8fb260a5-b4ad-42fe-acb9-685cd9b14112 · inbound

Layer-Level Self-Exposure and Patch: Affirmative Token Mitigation for Jailbreak Attack Defense cites this paper.

Layer-Level Self-Exposure and Patch: Affirmative Token Mitigation for Jailbreak Attack Defense Aligning Large Language Models with Human: A Survey

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T22:11:59.904662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:11:59.904662Z digest=sha256:a0679a53a6a4f32b6c48399b10d4b295840ac1254e8b0dade65e1afeb39d256b

Observation 187df938-3367-42db-a558-f61b6475eabf · inbound

Registering Source Tokens to Target Language Spaces in Multilingual Neural Machine Translation cites this paper.

Registering Source Tokens to Target Language Spaces in Multilingual Neural Machine Translation Aligning Large Language Models with Human: A Survey

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T22:05:49.463608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:05:49.463608Z digest=sha256:db2da3c0a3ddb22ae4c41fb1a967f64f96e5be89099956fb710e6afd607d3ff7

Observation b3699032-f98e-44ea-a0fa-27fab98dcc27 · inbound

Open Problems in Machine Unlearning for AI Safety cites this paper.

Open Problems in Machine Unlearning for AI Safety Aligning Large Language Models with Human: A Survey

Reference 137

Resolution
unresolved
no resolver link, observed 2026-08-10T21:24:10.713985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:24:10.713985Z digest=sha256:6391237e880f08e50a34a94c0b4ec627222c0e37770f7d5da97d2b8e0eb926bf

Observation 60075721-f8fd-4e49-83df-a94bba45815e · inbound

A Survey on Responsible LLMs: Inherent Risk, Malicious Use, and Mitigation Strategy cites this paper.

A Survey on Responsible LLMs: Inherent Risk, Malicious Use, and Mitigation Strategy Aligning Large Language Models with Human: A Survey

Reference 204

Resolution
unresolved
no resolver link, observed 2026-08-10T20:05:12.664352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:05:12.664352Z digest=sha256:60c004961b1dcc5984c7e0f8698796795333b90c3b176073b785f5df2d2ce070

Observation c90b588e-6a9e-456a-b9a9-8d39c2371945 · inbound

Gradient-Based Multi-Objective Deep Learning: Algorithms, Theories, Applications, and Beyond cites this paper.

Gradient-Based Multi-Objective Deep Learning: Algorithms, Theories, Applications, and Beyond Aligning Large Language Models with Human: A Survey

Reference 187

Resolution
unresolved
no resolver link, observed 2026-08-10T18:53:50.627408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:53:50.627408Z digest=sha256:50dd6bfb0567ba214e478167a22b21e760d2c3921726cca6ae34d3ebb46be586

Observation 851b166a-73d8-40d1-8f1e-d835b52704ec · inbound

Leveraging Sparsity for Sample-Efficient Preference Learning: A Theoretical Perspective cites this paper.

Leveraging Sparsity for Sample-Efficient Preference Learning: A Theoretical Perspective Aligning Large Language Models with Human: A Survey

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-10T00:11:59.452649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T00:11:59.452649Z digest=sha256:ec41b9fdfeffdaac0dd52819a742e1f5ac42e9e6948fa5f3cec655e9dba7af42

Observation 818504ca-51be-4dbd-bcb2-03e2a091b13b · inbound

LLM Alignment as Retriever Optimization: An Information Retrieval Perspective cites this paper.

LLM Alignment as Retriever Optimization: An Information Retrieval Perspective Aligning Large Language Models with Human: A Survey

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-09T04:08:52.078672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:08:52.078672Z digest=sha256:f4243079f61d1eadc365529b70d417c84f794d26649f78a752f5657dab6392c9

Observation 7b62c1cc-dda2-4c2d-b822-d4c87dd3408f · inbound

LLMs can be easily Confused by Instructional Distractions cites this paper.

LLMs can be easily Confused by Instructional Distractions Aligning Large Language Models with Human: A Survey

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-09T10:50:12.510955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T10:50:12.510955Z digest=sha256:91e56548a7dd3101e2fe3be7748c17ae05583284de960c2d89468e3ba4b12241

Observation 2fc4da3c-2665-48b1-aed9-c336e8ca34f6 · inbound

JBShield: Defending Large Language Models from Jailbreak Attacks through Activated Concept Analysis and Manipulation cites this paper.

JBShield: Defending Large Language Models from Jailbreak Attacks through Activated Concept Analysis and Manipulation Aligning Large Language Models with Human: A Survey

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T12:25:30.662609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T12:25:30.662609Z digest=sha256:fbfb19f6b53731e5ee4662a84a1e1afe0fef4b0e0ab62aba0111eec59b000b03

Observation 1a1e1960-52ff-47ba-87e4-82146ecdbae0 · inbound

Salamandra Technical Report cites this paper.

Salamandra Technical Report Aligning Large Language Models with Human: A Survey

Reference 202

Resolution
unresolved
no resolver link, observed 2026-08-08T04:58:34.907920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:58:34.907920Z digest=sha256:620d0446592cd817fd01aae52d5e41d212bbdaf8697a037d6a77764ff88940d4

Observation 88a9ef1a-fdf6-43df-9a9a-48aa06193a4e · inbound

Self-Consistency of the Internal Reward Models Improves Self-Rewarding Language Models cites this paper.

Self-Consistency of the Internal Reward Models Improves Self-Rewarding Language Models Aligning Large Language Models with Human: A Survey

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T23:18:40.221309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T23:18:40.221309Z digest=sha256:afc37c1cc5160aee05f9fe85ea827bc666a257425f37fe5bcefb87b75a269a11

Observation 7c242e88-f4c8-406b-9c4a-af55d33ae8e0 · inbound

Data2Concept2Text: An Explainable Multilingual Framework for Data Analysis Narration cites this paper.

Data2Concept2Text: An Explainable Multilingual Framework for Data Analysis Narration Aligning Large Language Models with Human: A Survey

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T22:19:01.091274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T22:19:01.091274Z digest=sha256:f886830c7dbc86495a89fbafd46f772020a64cfc979afb85c86c7787a13bdd08

Observation 9ea7a298-1b5d-4e66-81fa-4c2d9a97de88 · inbound

The Science of Evaluating Foundation Models cites this paper.

The Science of Evaluating Foundation Models Aligning Large Language Models with Human: A Survey

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-07T23:35:42.894068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:35:42.894068Z digest=sha256:fc78ac1664a0df6bed7a2c849e4131d34ee74943f71107d927d49cdd2af706d4

Observation eee08696-a2f3-49b1-8d4d-54c818e825cb · inbound

Benchmarking LLM-based Relevance Judgment Methods cites this paper.

Benchmarking LLM-based Relevance Judgment Methods Aligning Large Language Models with Human: A Survey

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-16T12:35:47.382195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:35:47.382195Z digest=sha256:e0526319dd3254e00e9ad5e5e8b730b9cc9e36821c65dcbb98f6cd55ecf99bbf

Observation 2bcbd87d-2d2d-4955-a8fa-8ec4bb1cd94d · inbound

A Survey of Foundation Model-Powered Recommender Systems: From Feature-Based, Generative to Agentic Paradigms cites this paper.

A Survey of Foundation Model-Powered Recommender Systems: From Feature-Based, Generative to Agentic Paradigms Aligning Large Language Models with Human: A Survey

Reference 286

Resolution
unresolved
no resolver link, observed 2026-08-16T11:07:59.593056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:07:59.593056Z digest=sha256:1f6d19862ff630cd6a861cf1f4e9d26e31c72d2556111cc35a6b2936c613667f

Observation 47dca15f-0f4b-40ed-8137-76226752cf07 · inbound

Leveraging LLMs as Meta-Judges: A Multi-Agent Framework for Evaluating LLM Judgments cites this paper.

Leveraging LLMs as Meta-Judges: A Multi-Agent Framework for Evaluating LLM Judgments Aligning Large Language Models with Human: A Survey

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T10:53:24.919360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:53:24.919360Z digest=sha256:7687a9b9fd3418d4a40a23dbc65f54e11459c084f06ea0540f0b20e4b4fdcfd4

Observation 7f60104e-3e48-426c-8b1a-31cc8ce16fdd · inbound

Inducing Robustness in a 2 Dimensional Direct Preference Optimization Paradigm cites this paper.

Inducing Robustness in a 2 Dimensional Direct Preference Optimization Paradigm Aligning Large Language Models with Human: A Survey

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T04:18:40.312439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:18:40.312439Z digest=sha256:5f504f89f1550d5b9786769da30792e8f124b91e88e90941e152e3f6ce87ae68

Observation 91278c53-1752-41d3-905b-8f3c3b775955 · inbound

A Survey on Progress in LLM Alignment from the Perspective of Reward Design cites this paper.

A Survey on Progress in LLM Alignment from the Perspective of Reward Design Aligning Large Language Models with Human: A Survey

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-16T00:52:06.547380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:52:06.547380Z digest=sha256:c2aedcab6774390e349ffa8ae3bf1c5772db37e3fa0989d01bd761e8b965bac8

Observation 320f3550-d278-47c8-92a1-20c5ab014541 · inbound

Aligning Large Language Models with Healthcare Stakeholders: A Pathway to Trustworthy AI Integration cites this paper.

Aligning Large Language Models with Healthcare Stakeholders: A Pathway to Trustworthy AI Integration Aligning Large Language Models with Human: A Survey

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-16T04:33:20.146108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:33:20.146108Z digest=sha256:a90257586d345d83e4c08087f37067de675c9f82a35e301877212aec19cf4fe2

Observation 6cb343b5-b991-4174-b376-c13034426ef1 · inbound

PARM: Multi-Objective Test-Time Alignment via Preference-Aware Autoregressive Reward Model cites this paper.

PARM: Multi-Objective Test-Time Alignment via Preference-Aware Autoregressive Reward Model Aligning Large Language Models with Human: A Survey

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-15T23:52:01.544478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:52:01.544478Z digest=sha256:d2af74486f355600cfbcd71215bdda141d54f553dc757477bb200728d5d39380

Observation fa3bf45e-6bb1-434e-b780-e557cda9ce2d · inbound

DesignFromX: Empowering Consumer-Driven Design Space Exploration through Feature Composition of Referenced Products cites this paper.

DesignFromX: Empowering Consumer-Driven Design Space Exploration through Feature Composition of Referenced Products Aligning Large Language Models with Human: A Survey

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-15T20:54:33.150015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:54:33.150015Z digest=sha256:b60b846cc3e286f73c31fc57a9f56ea760a8c88aee27071fdb99eb89471a24d5

Observation 1558201b-26df-488b-9c14-85b487cd1831 · inbound

IDEAL: Data Equilibrium Adaptation for Multi-Capability Language Model Alignment cites this paper.

IDEAL: Data Equilibrium Adaptation for Multi-Capability Language Model Alignment Aligning Large Language Models with Human: A Survey

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T20:32:06.108893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:32:06.108893Z digest=sha256:15ef3e1b9fea4d85f7a677a286aa1a011f945bc91a8f1a68aff2d86b4337b762

Observation dba8227b-b72d-479f-bde8-1c13692362b6 · inbound

Shallow Preference Signals: Large Language Model Aligns Even Better with Truncated Data? cites this paper.

Shallow Preference Signals: Large Language Model Aligns Even Better with Truncated Data? Aligning Large Language Models with Human: A Survey

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:46.155225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:46.155225Z digest=sha256:897cdd79a2562538faaa35e93fb107ab8baac377505e2988602e9a4443372a17

Observation 609374de-bfb4-44d8-8669-6b56623f2efb · inbound

One Model Transfer to All: On Robust Jailbreak Prompts Generation against LLMs cites this paper.

One Model Transfer to All: On Robust Jailbreak Prompts Generation against LLMs Aligning Large Language Models with Human: A Survey

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:48:18.341107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:48:18.341107Z digest=sha256:477c7377cf7a29beba4cf55f91c39b7e2983fae9428d8fd8f41165f733ecf2bc

Observation 821655a1-b7eb-4960-af6c-5269549aa457 · inbound

Token-level Accept or Reject: A Micro Alignment Approach for Large Language Models cites this paper.

Token-level Accept or Reject: A Micro Alignment Approach for Large Language Models Aligning Large Language Models with Human: A Survey

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:11:31.809976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:11:31.809976Z digest=sha256:eafb774b5f81a2b5f588904e744744503b8f7d6a57959adee67d6cfec3c40adf

Observation ce81cfe1-9c7a-4f3c-8a89-036620d07838 · inbound

Tag-Evol: Achieving Efficient Instruction Evolving via Tag Injection cites this paper.

Tag-Evol: Achieving Efficient Instruction Evolving via Tag Injection Aligning Large Language Models with Human: A Survey

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:06.783191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:39:06.783191Z digest=sha256:4743dbf52b5f056f559adff8af6ffa04658d267df06944e3b246be154ff842cc

Observation 1108cd90-5802-4740-ac3c-6e1f77a773c9 · inbound

Crowd-SFT: Crowdsourcing for LLM Alignment cites this paper.

Crowd-SFT: Crowdsourcing for LLM Alignment Aligning Large Language Models with Human: A Survey

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T10:54:09.009992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:54:09.009992Z digest=sha256:639637e72044253e36f209bce3a8683c51de81a789216f3b2583a03ff83a266c

Observation 40e765d3-f022-43db-b21d-72a3bceeaf90 · inbound

Large Language Models for EEG: A Comprehensive Survey and Taxonomy cites this paper.

Large Language Models for EEG: A Comprehensive Survey and Taxonomy Aligning Large Language Models with Human: A Survey

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:52.449611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:30:52.449611Z digest=sha256:f56ca5ce6f8a4839e31fd5f3976d037c8cd4b257143db43dd4a1edf23676559f

Observation b835bba0-5dc3-49a4-a69b-fdd4e7ec0ca3 · inbound

From Outcomes to Processes: Guiding PRM Learning from ORM for Inference-Time Alignment cites this paper.

From Outcomes to Processes: Guiding PRM Learning from ORM for Inference-Time Alignment Aligning Large Language Models with Human: A Survey

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:57:56.761990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:57:56.761990Z digest=sha256:84ffc4c53049ea6c54967cc254e58885faf9695a1aac5ce020c358acde2b2989

Observation 9bdf70b5-d257-488b-9ee0-c2b0eb1357ce · inbound

Structured Moral Reasoning in Language Models: A Value-Grounded Evaluation Framework cites this paper.

Structured Moral Reasoning in Language Models: A Value-Grounded Evaluation Framework Aligning Large Language Models with Human: A Survey

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:42.983393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:14:42.983393Z digest=sha256:3d5e1bbfcd5a0ce1cc1272d26eb05d11e0a27fe2c34fee25639bf8a761c7ce5e

Observation c851dedf-881e-468f-a1d2-cb025b1299ae · inbound

Beyond Reactive Safety: Risk-Aware LLM Alignment via Long-Horizon Simulation cites this paper.

Beyond Reactive Safety: Risk-Aware LLM Alignment via Long-Horizon Simulation Aligning Large Language Models with Human: A Survey

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T22:42:45.578933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:42:45.578933Z digest=sha256:cc0636d47d0e6d93685e37b7fd543720d8885e3acfa27468c63ca6f308424382

Observation 6cc15457-94e2-4888-8919-1070880e87cf · inbound

CrossPipe: Towards Optimal Pipeline Schedules for Cross-Datacenter Training cites this paper.

CrossPipe: Towards Optimal Pipeline Schedules for Cross-Datacenter Training Aligning Large Language Models with Human: A Survey

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T21:26:21.065188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:26:21.065188Z digest=sha256:ca6aa786005f4a05995b7874f900c75d0a10d681ea8008d664f51108cb2c7650

Observation 3eda6db4-60f4-4369-a821-f0ac6e6a1f7e · inbound

Reliable Annotations with Less Effort: Evaluating LLM-Human Collaboration in Search Clarifications cites this paper.

Reliable Annotations with Less Effort: Evaluating LLM-Human Collaboration in Search Clarifications Aligning Large Language Models with Human: A Survey

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.876176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.876176Z digest=sha256:76a8e1f748b39710c1a0d3ec215f2d4d2dae9f31f2949aa08884efba7ffe58c6

Observation 683d0c86-dafd-4339-a684-3a392aceb7fc · inbound

An Uncertainty-Driven Adaptive Self-Alignment Framework for Large Language Models cites this paper.

An Uncertainty-Driven Adaptive Self-Alignment Framework for Large Language Models Aligning Large Language Models with Human: A Survey

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T14:51:53.937236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:51:53.937236Z digest=sha256:c81c3f9df24305e0a1853fc3d4ba78b06070bf40a3a2eefabb79bfde9f866df9

Observation bffd1bc6-d437-4afa-815a-c1a301d40fde · inbound

Exploring the Impact of Instruction-Tuning on LLM's Susceptibility to Misinformation cites this paper.

Exploring the Impact of Instruction-Tuning on LLM's Susceptibility to Misinformation Aligning Large Language Models with Human: A Survey

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T14:42:56.745397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:42:56.745397Z digest=sha256:c323c2923101ec1718b164d57baccb9d3633c01fc0cb851c2554204097b89243

Observation c03c4251-6a7e-4263-ac2b-26752707d9e7 · inbound

The Fair Game: Auditing & Debiasing AI Algorithms Over Time cites this paper.

The Fair Game: Auditing & Debiasing AI Algorithms Over Time Aligning Large Language Models with Human: A Survey

Reference 128

Resolution
unresolved
no resolver link, observed 2026-08-05T22:45:20.820107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:45:20.820107Z digest=sha256:338316fcc024c7fe1483e8668d28af95ee260b573f0f0531b056b5c1f6cc3641

Observation e421729a-94d4-4be2-a375-cc8ba40258bb · inbound

A Comprehensive Evaluation framework of Alignment Techniques for LLMs cites this paper.

A Comprehensive Evaluation framework of Alignment Techniques for LLMs Aligning Large Language Models with Human: A Survey

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T20:43:17.984113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:43:17.984113Z digest=sha256:8961110391b9bb8260f0b1da8b627d4786b2797067e2277b141b2bb43b1ea21f

Observation 4b92d26f-39a4-4eba-a3f7-10ad453697b9 · inbound

FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain cites this paper.

FActBench: A Benchmark for Fine-grained Automatic Evaluation of LLM-Generated Text in the Medical Domain Aligning Large Language Models with Human: A Survey

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T11:49:37.302932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:49:37.302932Z digest=sha256:05802230ed4edfc3314eaf7eda94c74ae1aa92121e63b90a47ba271890ef4518

Observation 57da1d23-91bb-4077-aafa-9c65c0328894 · inbound

SharedRep-RLHF: A Shared Representation Approach to RLHF with Diverse Preferences cites this paper.

SharedRep-RLHF: A Shared Representation Approach to RLHF with Diverse Preferences Aligning Large Language Models with Human: A Survey

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T10:54:13.130765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T10:54:13.130765Z digest=sha256:3c619dcb8ffefdf5a2bd2154da6696174ab05af0d397aa0bed2387bd3b2c22b2

Observation 7cfc5e63-da03-430b-81d3-7cac7804c278 · inbound

Path to Intelligence: Measuring Similarity between Human Brain and Large Language Model Beyond Language Task cites this paper.

Path to Intelligence: Measuring Similarity between Human Brain and Large Language Model Beyond Language Task Aligning Large Language Models with Human: A Survey

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T15:53:35.838395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:53:35.838395Z digest=sha256:850e5934438d7ef72e764c5da72df326a046fbeee1ba51c29838b551a0384c47

Observation 143d675a-fa5b-49b8-8f3f-97fc9d644632 · inbound

AI Behavioral Science cites this paper.

AI Behavioral Science Aligning Large Language Models with Human: A Survey

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-15T17:27:30.522879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:27:30.522879Z digest=sha256:3c361f5b6032fb0ef4453a399cf3fc7dcc9966078681d6419d384d167a1ea30e

Observation d9fbaf8c-c010-4677-9bb4-9efb238f1cce · inbound

Knowledge-Driven Hallucination in Large Language Models: An Empirical Study on Process Modeling cites this paper.

Knowledge-Driven Hallucination in Large Language Models: An Empirical Study on Process Modeling Aligning Large Language Models with Human: A Survey

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T15:36:34.374898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-18T15:33:29.016355Z digest=sha256:2ca09b7e6436a6f2ae1e91e6a2b7ba085bec59cdbec82bbead95875f0f2f01f0

Observation deaaa5d5-8995-4668-994d-06a8fac37e73 · inbound

Toward Preference-aligned Large Language Models via Residual-based Model Steering cites this paper.

Toward Preference-aligned Large Language Models via Residual-based Model Steering Aligning Large Language Models with Human: A Survey

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T14:43:19.875156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:43:19.875156Z digest=sha256:6ae0d6ea3f5763b337fc21c1482fb393bbab436159f5add4f4fa8b56ba5f6825

Observation a7e5cd87-1075-4fde-ac8c-6632db444924 · inbound

Autonomy Reshapes How Personalization Affects Privacy Concerns and Trust in LLM Agents cites this paper.

Autonomy Reshapes How Personalization Affects Privacy Concerns and Trust in LLM Agents Aligning Large Language Models with Human: A Survey

Reference 98

Resolution
verified exact
arxiv_id, observed 2026-05-18T10:11:13.867179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-18T10:09:47.433345Z digest=sha256:95bf9d4b2cb51f9ba1b3d87b65d437e71f67b79c0e4fcebf249dfe409716b377

Observation be2789a8-ac13-4ff1-b9af-fb424e02b70f · inbound

POPI: Personalizing LLMs via Optimized Natural Language Preference Inference cites this paper.

POPI: Personalizing LLMs via Optimized Natural Language Preference Inference Aligning Large Language Models with Human: A Survey

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-18T05:42:24.403647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-18T05:41:58.231139Z digest=sha256:a0d065f22f910791b48fa4bf28201c8e08e1b5441a2e9a9339df84c162bb9383

Observation 0d685579-39f4-4f85-822c-09b267f8e285 · inbound

Linguistics and Human Brain: A Perspective of Computational Neuroscience cites this paper.

Linguistics and Human Brain: A Perspective of Computational Neuroscience Aligning Large Language Models with Human: A Survey

Reference 202

Resolution
unresolved
no resolver link, observed 2026-08-03T03:22:46.324071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:22:46.324071Z digest=sha256:3680db1290b9d7a5eebf84f378c1c517aa68911e2c004d1b4d248e9bdcf58c4d

Observation f7106de7-e6a0-43d1-bc18-be09eb504468 · inbound

GUI-Libra: Training Native GUI Agents to Reason and Act with Action-aware Supervision and Partially Verifiable RL cites this paper.

GUI-Libra: Training Native GUI Agents to Reason and Act with Action-aware Supervision and Partially Verifiable RL Aligning Large Language Models with Human: A Survey

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-02T20:51:44.365668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T20:51:44.365668Z digest=sha256:7f231d007cadaaa9ebde5d4d4d222a021a10b46078e5214b0463da2601efe98e

Observation 4023afcd-bfba-4924-97b6-1f6598ffc3d5 · inbound

Large Language Model Post-Training: A Unified View of Off-Policy and On-Policy Learning cites this paper.

Large Language Model Post-Training: A Unified View of Off-Policy and On-Policy Learning Aligning Large Language Models with Human: A Survey

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:30:53.819784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T18:28:58.515666Z digest=sha256:ceddd095a475a6b47c1d25c7b98b022c93201851420f50be685a146ada01ca60

Observation 35e1863e-4431-40ce-853c-23a511ae1006 · inbound

Facial-Expression-Aware Prompting for Empathetic LLM Tutoring cites this paper.

Facial-Expression-Aware Prompting for Empathetic LLM Tutoring Aligning Large Language Models with Human: A Survey

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-15T13:40:01.669757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-15T13:39:06.227042Z digest=sha256:c32d519191d17d1cb69774b9b4c808cf766071426eef307be8592c0fb3bba54a

Observation 03466297-b364-41b0-b1b1-1dafa78c695d · inbound

Facial-Expression-Aware Prompting for Empathetic LLM Tutoring cites this paper.

Facial-Expression-Aware Prompting for Empathetic LLM Tutoring Aligning Large Language Models with Human: A Survey

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T18:33:59.804738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:33:59.804738Z digest=sha256:9bb9cd4b3b684c953e4c1547bcf2b12da845a1aaa02f9e79e6ad7623dc3eca25

Observation 885954ad-f753-448e-adb2-695a9b53dfbd · inbound

Consistency Analysis of Sentiment Predictions using Syntactic & Semantic Context Assessment Summarization (SSAS) cites this paper.

Consistency Analysis of Sentiment Predictions using Syntactic & Semantic Context Assessment Summarization (SSAS) Aligning Large Language Models with Human: A Survey

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:00:03.761510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T10:55:20.435471Z digest=sha256:c25c8b5e8e5698f3aeb8fab309e08e1819e471dd5e99faaa96029d9687599c3b

Observation a6cc2b04-a12d-4532-af28-1b95a530c0c1 · inbound

Indirect reciprocity beyond pairwise interactions cites this paper.

Indirect reciprocity beyond pairwise interactions Aligning Large Language Models with Human: A Survey

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-12T00:41:23.645722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-07T14:45:58.815682Z digest=sha256:9b112de9e31367934264fa141def919ad7c8e8e2ae9f4c58a80fd4b1aa85c5d1

Observation 0bda4866-2b6b-404e-8d48-2e25ea1798f8 · inbound

PERSA: Reinforcement Learning for Professor-Style Personalized Feedback with LLMs cites this paper.

PERSA: Reinforcement Learning for Professor-Style Personalized Feedback with LLMs Aligning Large Language Models with Human: A Survey

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:51:42.735240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-09T19:09:07.557773Z digest=sha256:57ee447da559e83a46a7a1b2bd2b8dc00a2411fdae1f57cf23e1c27e6a18975f

Observation 218fb11b-e01e-4cea-83e1-4d84c8eab9c8 · inbound

Response Time Enhances Alignment with Heterogeneous Preferences cites this paper.

Response Time Enhances Alignment with Heterogeneous Preferences Aligning Large Language Models with Human: A Survey

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:45:59.838828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-11T01:04:26.288913Z digest=sha256:9e7325e51441554ef676759c58ae97d5d3f8eb3d21f9974de00861b570d5bf62

Observation 503ca179-0004-414b-bc56-db07cfd005dd · inbound

See What I Mean: Aligning Vision and Language Representations for Video Fine-grained Object Understanding cites this paper.

See What I Mean: Aligning Vision and Language Representations for Video Fine-grained Object Understanding Aligning Large Language Models with Human: A Survey

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-05-20T12:13:16.254561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-20T12:10:54.874012Z digest=sha256:d3be877455e412573e1f0e3948227f41d3b10ff6e2bfea66f6a592e0673cd818

Observation f550640a-e00c-4a1c-95df-9141b09a98ca · inbound

Diagnosing Multi-step Reasoning Failures in Black-box LLMs via Stepwise Confidence Attribution cites this paper.

Diagnosing Multi-step Reasoning Failures in Black-box LLMs via Stepwise Confidence Attribution Aligning Large Language Models with Human: A Survey

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T06:48:05.993884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-20T06:43:23.870127Z digest=sha256:97d8bf47ac0cfeefe59cd286e8b9b7790bbf6e89ac1266361d28ddfc6f4e2f09

Observation 2879b10a-7e1c-4324-a86a-46f1e5885cf1 · inbound

In-Context Reward Adaptation for Robust Preference Modeling cites this paper.

In-Context Reward Adaptation for Robust Preference Modeling Aligning Large Language Models with Human: A Survey

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:23:15.167842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-29T08:19:55.177440Z digest=sha256:a8772c58c49b9083e5562ef305d4e646a99c6c833cca76532df3400684986b77

Observation 55b25e08-35f0-4198-a0fb-18d302586808 · inbound

P$^2$-DPO: Grounding Hallucination in Perceptual Processing via Calibration Direct Preference Optimization cites this paper.

P$^2$-DPO: Grounding Hallucination in Perceptual Processing via Calibration Direct Preference Optimization Aligning Large Language Models with Human: A Survey

Reference 106

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:06:27.161690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-28T11:13:37.291834Z digest=sha256:661585040288d2e47b720cc7b67f41c2826e90af21affc6d167ba66aac3f58a0

Observation f35c814f-da58-476d-ba24-34d59129e525 · inbound

Beyond Averages: Evaluating LLMs on Human Survey Replication at the Distributional Level cites this paper.

Beyond Averages: Evaluating LLMs on Human Survey Replication at the Distributional Level Aligning Large Language Models with Human: A Survey

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:47:29.825064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-27T17:03:42.094786Z digest=sha256:9027f133075b8a7401a2f79c010bd913b0a6dc76ae616d4dfe651ed4cd8767b3

Observation 74afcf28-f983-4edd-926a-f405b7cd92fb · inbound

Emergent Misalignment Can Be Induced by Sycophancy and Reversed via Alignment Gating cites this paper.

Emergent Misalignment Can Be Induced by Sycophancy and Reversed via Alignment Gating Aligning Large Language Models with Human: A Survey

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T00:47:29.982154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-27T17:03:33.199645Z digest=sha256:97c3c2f8e0cd0b548d8911c573074365fd3317c15dc290918f26ca66bfc5b245

Observation ea367c24-37e4-4d5f-b719-ecf7c4e03636 · inbound

Mult-DPO: Multinomial Direct Preference Optimization for Recommender Systems cites this paper.

Mult-DPO: Multinomial Direct Preference Optimization for Recommender Systems Aligning Large Language Models with Human: A Survey

Reference 57

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T03:47:36.181932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-27T14:30:17.544465Z digest=sha256:67830addb16309898565138ed2ab437b7c260dfb3acacb1f0937af82539c7929

Observation 46b00b42-9d1f-43ce-bdcf-16278c9d3683 · inbound

Editorial Alignment: A Participatory Approach to Engaging Editorial Expertise in LLM-mediated Knowledge Dissemination cites this paper.

Editorial Alignment: A Participatory Approach to Engaging Editorial Expertise in LLM-mediated Knowledge Dissemination Aligning Large Language Models with Human: A Survey

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-07-04T05:39:39.611428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-26T15:49:56.843397Z digest=sha256:43922e4f12d93b376430f854e6bf3efcead74ffc3b808ee103cb38ab3b17c4aa

Observation 97c68298-beb0-4af7-95fe-d2cd5e3b9cb5 · inbound

BV-Blend: Uncertainty-Weighted Historical Baselines for Stable Critic-Free RL with Verifiable Rewards cites this paper.

BV-Blend: Uncertainty-Weighted Historical Baselines for Stable Critic-Free RL with Verifiable Rewards Aligning Large Language Models with Human: A Survey

Reference 49

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T12:44:40.090029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-30T10:07:39.554999Z digest=sha256:f91a162cf9daee1cda06ef93e29dc5ec500eef3adf82746c6b885a34e8ebbe39

Observation af9fa2c3-a1f3-4d80-86ef-559e51342893 · inbound

Collective cooperation without individual fidelity in LLM agents cites this paper.

Collective cooperation without individual fidelity in LLM agents Aligning Large Language Models with Human: A Survey

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T03:24:12.463293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T03:20:27.450534Z digest=sha256:bc29706e8b6910820ea330134a682d70830b3abf64c890fff84c0cb8d1660179

Observation a38dba1c-6fdb-4122-bfc3-cb0544ccd2d0 · inbound

Multi-Objective Exploration and Preference Optimization via Mutual Information cites this paper.

Multi-Objective Exploration and Preference Optimization via Mutual Information Aligning Large Language Models with Human: A Survey

Reference 55

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T21:18:58.019477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-07-03T21:17:46.551850Z digest=sha256:194dd940c125fe3fc6fca3412f95dd95ace5729c0e3cd97da1f713f52aa206d6

Observation c34473db-e7db-4b6e-8465-8e8a3320f759 · inbound

Probably Correct Optimal Stable Matching under Two-Sided Uncertainty cites this paper.

Probably Correct Optimal Stable Matching under Two-Sided Uncertainty Aligning Large Language Models with Human: A Survey

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-11T12:55:57.279598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T12:55:57.279598Z digest=sha256:72fe671c5ad25c35297ca80de233c37732d0922926a348e8ff381038664c9804

Observation fa9616f9-19c4-4a9d-9bdf-48b8b0ad9f62 · inbound

SOS-LoRA: Static Orthogonal-Subspace Low-Rank Adaptation with Fixed Multi-Scale Scaling cites this paper.

SOS-LoRA: Static Orthogonal-Subspace Low-Rank Adaptation with Fixed Multi-Scale Scaling Aligning Large Language Models with Human: A Survey

Reference 118

Resolution
unresolved
no resolver link, observed 2026-08-02T09:51:03.551985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T09:51:03.551985Z digest=sha256:b0421cdf8d602ff705e004535f7fa04d7d8f2df60782c42313e7a193c975f687

Observation 2059eb48-5b8f-409d-95e3-8eda809e0613 · inbound

Don't Mix Rewards, Mix Policies: Policy Decomposition and Optimization for Multi-Reward RL cites this paper.

Don't Mix Rewards, Mix Policies: Policy Decomposition and Optimization for Multi-Reward RL Aligning Large Language Models with Human: A Survey

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-03T10:57:59.358866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:57:59.358866Z digest=sha256:530861872d4c8d3cf380f2e08884b82a8206c393a705ef51b547944ea4356711

Observation ef3ba9ef-31fd-468d-a1cf-96a5722b7382 · inbound

Beyond Post-Hoc Temperature Scaling: Bilevel Optimization for LLM Calibration cites this paper.

Beyond Post-Hoc Temperature Scaling: Bilevel Optimization for LLM Calibration Aligning Large Language Models with Human: A Survey

Reference 201

Resolution
unresolved
no resolver link, observed 2026-08-15T14:33:57.705887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:33:57.705887Z digest=sha256:860b116a40aa88dbd0244826f008a2f929217e682435e4216f3568e21f2d6bf7

Observation ac750b83-facc-4f33-8eef-2c51cb07f761 · inbound

Contextual Value Alignment via Multilayer Combinatorial Fusion cites this paper.

Contextual Value Alignment via Multilayer Combinatorial Fusion Aligning Large Language Models with Human: A Survey

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T00:32:35.701976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:32:35.701976Z digest=sha256:5dac3e615c625f5b2976e3f31981f2effa407996fd60e1f9ff44f52bbba7b6fd

Observation 773b2450-6a12-41ba-b2fb-73855c9b8e8c · inbound

LLM Reasoning for Subjective Tasks: Failure Modes, Mitigation, and Dynamic Reasoning Routing cites this paper.

LLM Reasoning for Subjective Tasks: Failure Modes, Mitigation, and Dynamic Reasoning Routing Aligning Large Language Models with Human: A Survey

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-14T04:26:47.828577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:26:47.828577Z digest=sha256:05b4ded2bb8ad633a0bdc5f6dec751290899a588bd7547e6ac893ea44b107483

Observation 1b9abcb5-2103-4c97-b0ac-15d5c1fafb14 · inbound

ThinkRetrieve: Retrieval-Augmented Reasoning Traces for Test-Time Scaling cites this paper.

ThinkRetrieve: Retrieval-Augmented Reasoning Traces for Test-Time Scaling Aligning Large Language Models with Human: A Survey

Reference 230

Resolution
unresolved
no resolver link, observed 2026-08-12T14:10:46.050310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:10:46.050310Z digest=sha256:faecda87fd952688dc3c59a0a964ad8af9bc579bad31c22dec2c4184857dbc30