Pith. sign in

Paper Citation Record · LEDGER

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models

As of 8 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 1 inbound Pith citation observation for arXiv:2506.07165.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.07165 v1

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:48:04.530071Z

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-18T14:02:11.084514Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T14:02:39.792894Z

Reference resolution

45 of 45 outbound references displayed

  • verified exact1
  • verified fuzzy14
  • unresolved29
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 509ef113-699f-40ce-8791-51ccddbc9a56 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:48:04.380172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:48:04.380172Z digest=sha256:6c79de52042322ba3e9f661d9c6d24e280ade12591def99d52e492d89f888faf

Observation 6a775cc4-d1d3-4e49-8a9b-cb25c3ccb4a9 · outbound

This paper cites - (2) Acknowledges both but slight deviations.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models - (2) Acknowledges both but slight deviations

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:48:04.384800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:48:04.384800Z digest=sha256:3889ddc8762f31ad824f37b26dc443e31948e6e58151c1cdec14c3d66147151a

Observation 1b1a3283-9453-4c75-b10b-7acce1c8d99c · outbound

This paper cites Unified Preference Optimization: Language Model Alignment Beyond the Preference Frontier.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unified Preference Optimization: Language Model Alignment Beyond the Preference Frontier

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:48:04.758497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.333706Z digest=sha256:fd7754a160ce6a17b7d3fd61da609995698617ddb13b33fb82f0e9cfb98e4d12

Observation 71380829-229d-44c8-9d61-8c582f171bdb · outbound

This paper cites Based the instruction following rule and given my answer to an instruction, your role is to provide specific and constructive score for me.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Based the instruction following rule and given my answer to an instruction, your role is to provide specific and constructive score for me

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.206741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.394469Z digest=sha256:c4a4908b05dc2060c14a4d6479f2b3683e4f8e6f87c9a3a787dc4bbb7e669d9d

Observation 86162f79-067f-4f88-b0c0-d2f4e8348557 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.909769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.489631Z digest=sha256:ef87ec8e5e3022dff89e673c3e64b5e714020145de13abbcc60ed256e324f219

Observation 4aab7911-1989-4d87-9051-75447098113c · outbound

This paper cites Multi-Objective Alignment of Large Language Models Through Hypervolume Maximization.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Multi-Objective Alignment of Large Language Models Through Hypervolume Maximization

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:48:04.350188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:48:04.350188Z digest=sha256:581a82048d56ed159abc04a8b66e22cc9fab270e7cfd5fc808dfa84228e6524c

Observation fbff26fe-c0c9-46e4-bedc-349cecf533f1 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.804950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.520670Z digest=sha256:79597f53802f54e248750b2267cd71da0382d94e8b23307b2c5c77fda29428db

Observation 603471d7-87ff-4375-9720-010016c55d82 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:05.258845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.360084Z digest=sha256:614d8111667b974b644f4bb8e1c539e7e19666cdabaf47f1a32c94d9fc7ca90e

Observation 5fbb7736-bb7b-42f0-86de-4dcedd084167 · outbound

This paper cites Al ter na tively, you can use a nav iga tion app like Google Maps or Waze to get the most ac cu rate and up -to -date di rec tions.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Al ter na tively, you can use a nav iga tion app like Google Maps or Waze to get the most ac cu rate and up -to -date di rec tions

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:04.774067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.530071Z digest=sha256:96557a74f04a2d66a941bc2fe86a1ee61b2df8054c947d64d2042d9a32a66a95

Observation a6c9871d-7ba5-4d27-ba1d-24fa4ab96c53 · outbound

This paper cites RRHF: Rank Responses to Align Language Models with Human Feedback without tears.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models RRHF: Rank Responses to Align Language Models with Human Feedback without tears

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T05:48:04.370319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:48:04.370319Z digest=sha256:a40eff3cff4dfad5503226ae681e28a4ec8c3f1b8491f9f46372c98d97bd11c5

Observation 1f6ca1b9-2929-4fd4-a204-d7da3e8cf2d7 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:48:04.375331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:48:04.375331Z digest=sha256:8d289ff61f9e80e0e7dbe9fba206cad448ebee513cc365b93e81890806925445

Observation d9e92826-9fb6-46ae-8ebc-a6341bdc5fa9 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:48:04.389549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:48:04.389549Z digest=sha256:e8e3ee7131113fd0b10d76a9484dd5e69b807c3a44b31df52cfd952122997ad7

Observation ed371d59-8070-42de-9b40-2ce71ec1f177 · outbound

This paper cites The response completely missed the essence of what the user wanted.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models The response completely missed the essence of what the user wanted

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.192139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.399006Z digest=sha256:785273e51aaa9f84da3fbc5c239c9bd86f712fdce5c55e15ff5e11ee995bbed4

Observation 0cd52a9f-1b84-4ad1-a6bd-9cbb8fd30d50 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:05.177921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.403467Z digest=sha256:efed6a1ad002eb42c51331de317c0c4ff5a78cb4524030a5c31e71ceed258022

Observation cd53e2c5-9267-4541-9df4-ee050e8602fb · outbound

This paper cites The response did not fully satisfy what the user was looking for.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models The response did not fully satisfy what the user was looking for

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.163942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.408360Z digest=sha256:298f08fc6b9771260ead915fb59c30ec85513114a32cac96d897349720c6e10e

Observation aa9b1903-cb42-4c0a-814c-254e76fae5c4 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:05.150165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.413250Z digest=sha256:37c008d75b8054b24e3b9fccca59792d85df1912301835d038fb92b7e376d40b

Observation cbc3e889-c355-4e81-9ff2-9165a1ad4ad2 · outbound

This paper cites Based the helpfulness rule and given my answer to an instruction, your role is to provide specific and constructive score for me.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Based the helpfulness rule and given my answer to an instruction, your role is to provide specific and constructive score for me

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.136403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.417719Z digest=sha256:699bd22112389a1b9e7d0fe481d2b3bba07731c849b404ae7ea71314d7feec1c

Observation 32325681-7ec1-4dd6-9bdf-f0ea1f3c9990 · outbound

This paper cites All information provided is wrong, false or hallucinated.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models All information provided is wrong, false or hallucinated

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.122505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.422395Z digest=sha256:fa74c23b75c2a62ffc09d6bab3faee759bb573e2584da3405d7d7c83d705016e

Observation 52821ed3-340a-4e97-a005-215c70773b3c · outbound

This paper cites The response may contain multiple instances of hallucinations, false information, misleading information, or irrelevant information.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models The response may contain multiple instances of hallucinations, false information, misleading information, or irrelevant information

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.108613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.426726Z digest=sha256:319c988bedbc0892c518c5c1923c3d21beb31a57ef757a7faa6f0e7b70871102

Observation 58aa1c99-ebb1-497d-9d3e-ebc5d1b74f66 · outbound

This paper cites The response may miss some details, contain misleading information, or minor hallucinations, but is more or less aligned with what the prompt asks for.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models The response may miss some details, contain misleading information, or minor hallucinations, but is more or less aligned with what the prompt asks for

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.094666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.431058Z digest=sha256:b88d6e63fb9d4226487b5a37e55f049712a120d3f7d73001e9e975bf1a51b83d

Observation 7fd95f00-d9a4-45b7-9049-a5a45085108f · outbound

This paper cites It contains no misleading information or hallucinations.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models It contains no misleading information or hallucinations

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.079708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.435493Z digest=sha256:21b6641da156070afc577c0fe35878626695e0b6a54980449d963b85a645985c

Observation 790f3e81-e2ec-499a-aab5-3cfcc4769e69 · outbound

This paper cites preference.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models preference

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.065370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.439958Z digest=sha256:a0a127290afb06283f657a7460c85a4f66ee5355bb07c9d375afdde11bf17545

Observation 0011564e-56b0-4aab-abea-a22f639a26d9 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:05.049744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.445478Z digest=sha256:331333d114e86d59670ed334113749f3942c953e42628287e87cc75e87e4492a

Observation 3f29bfbe-07b8-4c9f-a11d-1b34ce1daa2a · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:05.035054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.449913Z digest=sha256:90ad479372a28ead448c1442029530da52035e20bbef2d22dbfe1ac357938060

Observation dce37129-54b0-45fb-8324-9574dd80a10a · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:05.020579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.454405Z digest=sha256:4d7d0126131f5d84101544eb0dba4a2414c7238b772f1540398300d9ba84692a

Observation 4af259a0-678c-44b6-8a77-ec21b6b0bcc1 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:05.006108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.458745Z digest=sha256:e45f91093d409b6570f13f73f5c50b302c7266b4556bfb08e95120c7e9502fb1

Observation 48b3e66d-2b68-455e-8ba5-0529e032da40 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.992129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.463266Z digest=sha256:66ba4b6656e0fd95bab1dc8d69be02b3828bccad7efa55722a54076c4862492e

Observation 40e32249-f472-4723-8fff-fd3817ceeb14 · outbound

This paper cites “markdown“‘markdown. This is an example of a code block in Markdown. You can see that it is formatted to look like it’s not part of the regular text flow.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models “markdown“‘markdown. This is an example of a code block in Markdown. You can see that it is formatted to look like it’s not part of the regular text flow

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:04.978143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.467611Z digest=sha256:c5dba2179f783c9c14fb05a009199afc14884657b30a9c581ae4ff67355ee02e

Observation 0d84453a-c1c0-4a46-8b93-550b18215763 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.963674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.472349Z digest=sha256:20cf67da27ef970c53c005500e841c3da714708e03a404aa1bf9ffeb4dde99bd

Observation 76909b5b-1482-4ff9-8e44-9f0fcfd9bfd7 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.950425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.476696Z digest=sha256:6f97af1779a39015d84a16b3180aa5128624dfb08b71d8999fc31f5e2c1bb2e6

Observation ebadc2cd-9ee1-40cb-b856-13ae76e0a176 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.937069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.480870Z digest=sha256:899e59e70aea9b42ef6ef6d3270bb52eb79f20be56348a841dd12968460bb4f4

Observation b862183f-ab10-4734-86ba-1069bdc073a7 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.923359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.485127Z digest=sha256:4932f35c927532f20ea68605c84898bf01b2f14b5db1f262cd5d1261d7aebdf9

Observation 41fc8d2a-bdb7-4ad1-a839-ac4003e13cda · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.895292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.494306Z digest=sha256:e70fff88bf7362659aeede47cd4b373fb502777f11c20ab90ef0f7abf0fd3719

Observation 46f191fa-7d8e-42ec-a79a-c6b79afd8bfe · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.881178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.498733Z digest=sha256:f9013ed3cf3102ca2707518238c80d6d9e573016d9e02a28c38c5215b3ee81e2

Observation 8905fa09-1251-4ee1-ad62-a8170e0d10ff · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 39

Resolution
parse uncertain
raw_fallback, observed 2026-08-07T05:48:04.866254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.502975Z digest=sha256:378a57b98df460810fd2c95c7a42ec9e589461f4885188d2c4c3df9b4c994f24

Observation 0666dcdc-6eb4-4a53-9d2a-e916603e2a7f · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.850567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.507704Z digest=sha256:3d4886467d1aea21cfa7c5419e48ffe555e01c25d980a2739abb461330d2aee5

Observation 4b3002d7-2969-41ed-bdf5-5b0cce6667fe · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.834072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.511997Z digest=sha256:c1ea864dee121faea45a9104f24b7d57028ed193738e9b52c286a05b2dcbdccd

Observation 76cc2232-739b-4d7b-8197-e6c39dab316a · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.819470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.516378Z digest=sha256:43a5d70b83a09f522cba7b6cb6b46dfc91a26ef78ce646a9d884cac54edc7d18

Observation d97ec2f6-8e34-480a-9c64-b69b5dd026f1 · outbound

This paper cites an unresolved cited work.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:48:04.788887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.525653Z digest=sha256:0114cfbb4f29925cee378a578032ff3ac761a81ec4bb307fb34a8f125b692065

Observation a13a5ca8-7485-485b-9904-bfad94f57cc4 · outbound

This paper cites Rethinking Bradley-Terry Models in Preference-Based Reward Modeling: Foundations, Theory, and Alternatives.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Rethinking Bradley-Terry Models in Preference-Based Reward Modeling: Foundations, Theory, and Alternatives

Reference 1027

Resolution
unresolved
no resolver link, observed 2026-08-07T05:48:04.364900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:48:04.364900Z digest=sha256:eb4f28d722848c998d489a35a53657316c5c2e659dc208a408b8e8529e36e17b

Observation 10c276b0-1f1d-4f37-ba7b-f54921c2139b · outbound

This paper cites KTO: Model Alignment as Prospect Theoretic Optimization.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models KTO: Model Alignment as Prospect Theoretic Optimization

Reference 1983

Resolution
unresolved
no resolver link, observed 2026-08-07T05:48:04.339130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:48:04.339130Z digest=sha256:20dfb3eae6a03564acf63c737627ff5a0c978a0bbb0449c29f2c09db91ea0334

Observation 46dcb65d-320c-4102-bf03-cf7d6de321d5 · outbound

This paper cites Ryan Park, Rafael Rafailov, Stefano Ermon, and Chelsea Finn.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Ryan Park, Rafael Rafailov, Stefano Ermon, and Chelsea Finn

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.273911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.355394Z digest=sha256:f18975cf18ddb58c4777907064cfe521a3c9358935aa1cf9f27eca22bbf6bbe4

Observation b655c1f6-dca9-4a28-bdb0-b8ca25589367 · outbound

This paper cites Personalized Soups: Personalized Large Language Model Alignment via Post-hoc Parameter Merging.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Personalized Soups: Personalized Large Language Model Alignment via Post-hoc Parameter Merging

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T05:48:04.345187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:48:04.345187Z digest=sha256:bb7d8dcc14e3c3528837925a3f66a489d2f78dd35eaa0a04bf8247e992ccf986

Observation 46af202c-d500-4828-b593-0f2cd9e422e8 · outbound

This paper cites Mohammad Gheshlaghi Azar, Zhaohan Daniel Guo, Bi- lal Piot, Rémi Munos, Mark Rowland, Michal Valko, and Daniele Calandriello.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Mohammad Gheshlaghi Azar, Zhaohan Daniel Guo, Bi- lal Piot, Rémi Munos, Mark Rowland, Michal Valko, and Daniele Calandriello

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.303657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.324206Z digest=sha256:389f3bb2c6a28d4d0cbd09317ce1cf00497b103ed778f774cc755a707552c588

Observation 91b171db-f386-42d6-88ab-6aa8242fb34e · outbound

This paper cites Anirudhan Badrinath, Prabhat Agarwal, and Jiajing Xu.

AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models Anirudhan Badrinath, Prabhat Agarwal, and Jiajing Xu

Reference 4455

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:48:05.288823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:48:04.329332Z digest=sha256:d8e56a127c1d804bd169ddcbe909de802c679568c4a8a3f12b3fb3a4074bebf3

Pith citing papers

Observation 7f8dbf7d-2cd1-4c5d-8052-3748243927f9 · inbound

Failure Modes of Maximum Entropy RLHF cites this paper.

Failure Modes of Maximum Entropy RLHF AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-18T14:02:39.796329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-18T14:02:11.084514Z digest=sha256:9db0e678746fd755b03929785cd2245e627298a313da57804a1aaceff278b439