Pith. sign in

Paper Citation Record · LEDGER

Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2303.03751.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2303.03751 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T15:45:06.712104Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T13:09:51.325412Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6e240516-5ea6-4089-99c3-efbd86510ea0 · inbound

Prompt-Tuning Bandits: Enabling Few-Shot Generalization for Efficient Multi-Task Offline RL cites this paper.

Prompt-Tuning Bandits: Enabling Few-Shot Generalization for Efficient Multi-Task Offline RL Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T15:45:06.712104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:45:06.712104Z digest=sha256:5b7e42b2bfc0958fe4a7b0943048005eaf6a976be64287fb021c1ab2795bca24

Observation d54c479e-6732-4401-8172-542ba2789ed3 · inbound

Instant Preference Alignment for Text-to-Image Diffusion Models cites this paper.

Instant Preference Alignment for Text-to-Image Diffusion Models Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T16:51:04.814963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:51:04.814963Z digest=sha256:5e13e2fd3f3dfcac78799b5f646d48ada59890639cc7f9575709fdaf608e60f3

Observation 3c35af9b-7a2e-4d5e-abd8-1961b0bcda92 · inbound

Oracle Noise: Faster Semantic Spherical Alignment for Interpretable Latent Optimization cites this paper.

Oracle Noise: Faster Semantic Spherical Alignment for Interpretable Latent Optimization Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:06:12.114765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T06:58:09.085573Z digest=sha256:ecfb0a95fdef1e839c3d3c0a49e2e1ffff9464da0b5caea869f9a4edf5fc1bab

Observation 61092903-4699-4fd0-96ba-ad0feaca4269 · inbound

AdaMeZO: Adam-style Zeroth-Order Optimizer for LLM Fine-tuning Without Maintaining the Moments cites this paper.

AdaMeZO: Adam-style Zeroth-Order Optimizer for LLM Fine-tuning Without Maintaining the Moments Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

Reference 64

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:31:07.649291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-09T19:50:50.653184Z digest=sha256:69fbea26e0cc5c1bdde787a83b1ee08498f705398e154c048091a73c3353278d

Observation 34afb4dd-5ed6-4a44-a72c-c89865faee4c · inbound

Finding Stationary Points by Comparisons cites this paper.

Finding Stationary Points by Comparisons Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T13:09:51.326942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T05:26:07.720218Z digest=sha256:b133e66dc678bff552557d4cf0a7c866cc1313cb83d4448de7fe901177be19b5

Observation 599084a4-c3c0-4de6-99da-26acea2494d4 · inbound

Sculpting NeRF Geometry: Human-Preference Fine-Tuning of a 3D-Aware Face GAN cites this paper.

Sculpting NeRF Geometry: Human-Preference Fine-Tuning of a 3D-Aware Face GAN Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-07-04T12:59:53.187149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T05:33:57.816715Z digest=sha256:bf1c119ce99bfaf9ae3757d0943eb0b110e6ba9b07d7df51fd827dcf7d387696