Pith. sign in

Paper Citation Record · LEDGER

LiPO: Listwise Preference Optimization through Learning-to-Rank

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2402.01878.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.01878 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T14:46:51.815631Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T21:23:27.413513Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2ebd1d91-ecbf-45fb-bfd3-e5688bb9c815 · inbound

UNA: A Unified Supervised Framework for Efficient LLM Alignment Across Feedback Types cites this paper.

UNA: A Unified Supervised Framework for Efficient LLM Alignment Across Feedback Types LiPO: Listwise Preference Optimization through Learning-to-Rank

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-23T21:23:27.417555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-23T21:22:36.970101Z digest=sha256:c531f1dc78fa89db1516c0c921170ab9c37f0c51d44c6e1b4df900e3b74d1704

Observation 52f1e475-40f3-437a-a27d-c611f82978de · inbound

Controllable Protein Sequence Generation with LLM Preference Optimization cites this paper.

Controllable Protein Sequence Generation with LLM Preference Optimization LiPO: Listwise Preference Optimization through Learning-to-Rank

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T14:46:51.815631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:46:51.815631Z digest=sha256:eac12dedd868087602cce5a0616d6e439ce1b9efa8be811ff26e544710d5faa8

Observation 17aeefa1-009b-4909-aa4d-091165f021e5 · inbound

The Differences Between Direct Alignment Algorithms are a Blur cites this paper.

The Differences Between Direct Alignment Algorithms are a Blur LiPO: Listwise Preference Optimization through Learning-to-Rank

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:52:29.529197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-23T03:50:03.720389Z digest=sha256:a5d7498acead5535e97fc6af49ee4efb96b93d3797a50995d4b9bda2f9e1d932

Observation 8b1e524f-e854-4e74-b820-d5a26e47e936 · inbound

LLM Alignment as Retriever Optimization: An Information Retrieval Perspective cites this paper.

LLM Alignment as Retriever Optimization: An Information Retrieval Perspective LiPO: Listwise Preference Optimization through Learning-to-Rank

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-09T04:08:51.724809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:08:51.724809Z digest=sha256:ef4504e541067afababb88871ce8a9305ec48aec2c99eed7d31dd5a3387893f9

Observation f00dd238-8ddf-480e-a080-9759f83a6233 · inbound

PerPO: Perceptual Preference Optimization via Discriminative Rewarding cites this paper.

PerPO: Perceptual Preference Optimization via Discriminative Rewarding LiPO: Listwise Preference Optimization through Learning-to-Rank

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-09T06:01:13.306315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T06:01:13.306315Z digest=sha256:9f820237675ce416b79cb32db4f1b0a8b6daea7d56cb483b84e4a13f28fbc803

Observation ac0a7142-37ff-468f-868a-eb0e608aa41a · inbound

Advancing LLM Safe Alignment with Safety Representation Ranking cites this paper.

Advancing LLM Safe Alignment with Safety Representation Ranking LiPO: Listwise Preference Optimization through Learning-to-Rank

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:18.276005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:18.276005Z digest=sha256:b49306604ad197021d38f965db4f36c3e2cdd062b0b7bc8e109a925d92e6daaf

Observation 6852c46f-e6fb-4d63-972d-aac640c962ee · inbound

MPO: Multilingual Safety Alignment via Reward Gap Optimization cites this paper.

MPO: Multilingual Safety Alignment via Reward Gap Optimization LiPO: Listwise Preference Optimization through Learning-to-Rank

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:28.962243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:28.962243Z digest=sha256:0a9dbdac9aee9467acf62a5b17910b11e0d9c628af1677182dbf466ffd2ef6de

Observation 8d7eba4a-3ff9-4c93-ad87-bddebd038360 · inbound

LPOI: Listwise Preference Optimization for Vision Language Models cites this paper.

LPOI: Listwise Preference Optimization for Vision Language Models LiPO: Listwise Preference Optimization through Learning-to-Rank

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T13:45:55.816724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:45:55.816724Z digest=sha256:a765eb103e5cfa42428799169322d68756ec971c1dec7672388cbcabdba88aff

Observation c104ef4b-df02-4545-b873-a02691f147ea · inbound

HSCR: Hierarchical Self-Contrastive Rewarding for Aligning Medical Vision Language Models cites this paper.

HSCR: Hierarchical Self-Contrastive Rewarding for Aligning Medical Vision Language Models LiPO: Listwise Preference Optimization through Learning-to-Rank

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:00:48.923951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:00:48.923951Z digest=sha256:c9855322e2435025bd3c17f667a02efdc9f8d9a59f2581614c58fd8e4c09e40c

Observation 98d5ad29-7c6a-41bb-900b-d36ab5f2d551 · inbound

LeanPO: Lean Preference Optimization for Likelihood Alignment in Video-LLMs cites this paper.

LeanPO: Lean Preference Optimization for Likelihood Alignment in Video-LLMs LiPO: Listwise Preference Optimization through Learning-to-Rank

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:48.212275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:48.212275Z digest=sha256:5759e6caeeb14550fb74a9f794c1c30a0031535fca046dd624cae61f19e7693b

Observation 9bfe8b37-794b-4f8e-8281-d7588622099f · inbound

HAEPO: History-Aggregated Exploratory Policy Optimization cites this paper.

HAEPO: History-Aggregated Exploratory Policy Optimization LiPO: Listwise Preference Optimization through Learning-to-Rank

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T16:13:08.310472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:13:08.310472Z digest=sha256:75438779a9fc28fdafdba7dbfa80ad2e19f4b566d6d22da68e7def6ffe5c3bdf

Observation 582509b3-ae83-4d8b-ad9a-5603ce94dccc · inbound

Threshold-Guided Optimization for Visual Generative Models cites this paper.

Threshold-Guided Optimization for Visual Generative Models LiPO: Listwise Preference Optimization through Learning-to-Rank

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T17:46:07.941014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-08T17:14:36.632493Z digest=sha256:e74ea05fdf719976c3c3608536b7813924df34b8a900957c9aee1fb4dbe26565

Observation 1838de3e-0a41-455a-b996-f0c3cb00713a · inbound

Response Time Enhances Alignment with Heterogeneous Preferences cites this paper.

Response Time Enhances Alignment with Heterogeneous Preferences LiPO: Listwise Preference Optimization through Learning-to-Rank

Reference 160

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:46:00.020414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-11T01:04:26.288913Z digest=sha256:212a3229d8598cb18fee9bed56d059874e4943e37efde372057d2785161d07b3

Observation c99c18e0-2630-4461-99d9-37a51bcdf627 · inbound

MentalThink: Shaping Thoughts in Mental SVG World cites this paper.

MentalThink: Shaping Thoughts in Mental SVG World LiPO: Listwise Preference Optimization through Learning-to-Rank

Reference 231

Resolution
unresolved
no resolver link, observed 2026-07-12T01:50:59.184754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T01:50:59.184754Z digest=sha256:46ec5348416f0846b19fbbbd0f73169e3940dc26b4e6630f6aa8a6b75c6ba2a7

Observation b3380034-14bd-4832-8393-397a7c40bb0f · inbound

Multi-Turn On-Policy Distillation with Prefix Replay cites this paper.

Multi-Turn On-Policy Distillation with Prefix Replay LiPO: Listwise Preference Optimization through Learning-to-Rank

Reference 88

Resolution
unresolved
no resolver link, observed 2026-07-11T13:53:36.775836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T13:53:36.775836Z digest=sha256:b6d2867e9252fce8f9c97e0e7bc8ccab98432896f08bd255668e3f336f11f517

Observation 910aa229-e412-462f-b930-696349ee7d06 · inbound

Multi-Turn On-Policy Distillation with Prefix Replay cites this paper.

Multi-Turn On-Policy Distillation with Prefix Replay LiPO: Listwise Preference Optimization through Learning-to-Rank

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-02T08:40:41.165786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T08:40:41.165786Z digest=sha256:c1c1caaaa5d546a5c1aa142ce96c6078638c4909e5131e7bcc6aea15d850fa94