Pith. sign in

Paper Citation Record · LEDGER

Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2303.03751.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2303.03751 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T18:43:51.828298Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T13:09:51.325412Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c11513d6-9d16-407a-b038-04becfbf1ecd · inbound

Ruppert-Polyak averaging for Stochastic Order Oracle cites this paper.

Ruppert-Polyak averaging for Stochastic Order Oracle Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T13:55:43.053879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:55:43.053879Z digest=sha256:ebf0ae08f3be20a6cf2e88412fbeace8aedae90b2c25317b6a0fd044f690b56e

Observation 65461040-35cb-4b06-8965-528cc7ed8e69 · inbound

Hierarchical Prompt Decision Transformer: Improving Few-Shot Policy Generalization with Global and Adaptive Guidance cites this paper.

Hierarchical Prompt Decision Transformer: Improving Few-Shot Policy Generalization with Global and Adaptive Guidance Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T04:52:59.579013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:52:59.579013Z digest=sha256:c2a3750fe4a56429ba716642abb53a41817918884b3b12fd47233405deac40d2

Observation 81d80449-fae5-4242-b978-e223ce04e26a · inbound

Steering Large Text-to-Image Model for Abstract Art Synthesis: Preference-based Prompt Optimization and Visualization cites this paper.

Steering Large Text-to-Image Model for Abstract Art Synthesis: Preference-based Prompt Optimization and Visualization Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T18:43:51.828298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:43:51.828298Z digest=sha256:17ce1d17d47f566466dcad5154992a1592443fb51fb71aeea27e943db7df128f

Observation 6e240516-5ea6-4089-99c3-efbd86510ea0 · inbound

Prompt-Tuning Bandits: Enabling Few-Shot Generalization for Efficient Multi-Task Offline RL cites this paper.

Prompt-Tuning Bandits: Enabling Few-Shot Generalization for Efficient Multi-Task Offline RL Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T15:45:06.712104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:45:06.712104Z digest=sha256:f14ae226b1fdafdb077b85720b4b59d1806997f2f98f67cd2dd010dad3727c6d

Observation d54c479e-6732-4401-8172-542ba2789ed3 · inbound

Instant Preference Alignment for Text-to-Image Diffusion Models cites this paper.

Instant Preference Alignment for Text-to-Image Diffusion Models Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T16:51:04.814963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:51:04.814963Z digest=sha256:19c87e72b54ca02fc84f4f761bc5438025c3e94bbb4c9bab672609ee78aa02c1

Observation 3c35af9b-7a2e-4d5e-abd8-1961b0bcda92 · inbound

Oracle Noise: Faster Semantic Spherical Alignment for Interpretable Latent Optimization cites this paper.

Oracle Noise: Faster Semantic Spherical Alignment for Interpretable Latent Optimization Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:06:12.114765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-08T06:58:09.085573Z digest=sha256:df9329a3690ab2e123d8cefb118b502abdc7b23c845bbefbf516c66bf6bcada5

Observation 61092903-4699-4fd0-96ba-ad0feaca4269 · inbound

AdaMeZO: Adam-style Zeroth-Order Optimizer for LLM Fine-tuning Without Maintaining the Moments cites this paper.

AdaMeZO: Adam-style Zeroth-Order Optimizer for LLM Fine-tuning Without Maintaining the Moments Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

Reference 64

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:31:07.649291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-09T19:50:50.653184Z digest=sha256:f50f8e01a0f36086c5328ccbe5311cf224a78cc8e1c5c618be00176dea86db98

Observation 34afb4dd-5ed6-4a44-a72c-c89865faee4c · inbound

Finding Stationary Points by Comparisons cites this paper.

Finding Stationary Points by Comparisons Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T13:09:51.326942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-26T05:26:07.720218Z digest=sha256:43fdd8cd96073b043c409254d058fd768a5fa2eae1da2d4472a9156764649084

Observation 599084a4-c3c0-4de6-99da-26acea2494d4 · inbound

Sculpting NeRF Geometry: Human-Preference Fine-Tuning of a 3D-Aware Face GAN cites this paper.

Sculpting NeRF Geometry: Human-Preference Fine-Tuning of a 3D-Aware Face GAN Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-07-04T12:59:53.187149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-26T05:33:57.816715Z digest=sha256:c55ec1f0ff0ae00ab11a990d47791d88d4302608ed1126b2087cd9830fa22cd9