Pith. sign in

Paper Citation Record · LEDGER

Best-of-Venom: Attacking RLHF by Injecting Poisoned Preference Data

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2404.05530.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.05530 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:18:03.450932Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T08:35:34.297370Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c232ca51-e49b-4862-ba9d-ffcc20072146 · inbound

BadReward: Clean-Label Poisoning of Reward Models in Text-to-Image RLHF cites this paper.

BadReward: Clean-Label Poisoning of Reward Models in Text-to-Image RLHF Best-of-Venom: Attacking RLHF by Injecting Poisoned Preference Data

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:18:03.450932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:18:03.450932Z digest=sha256:a7d19a301c3787ae315e40f409dce4d78d3fa919a3fce6615a8b3955e678f3b9

Observation 7ea8583f-01ce-43e0-8962-3952464ac255 · inbound

A Systematic Review of Poisoning Attacks Against Large Language Models cites this paper.

A Systematic Review of Poisoning Attacks Against Large Language Models Best-of-Venom: Attacking RLHF by Injecting Poisoned Preference Data

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:31.659055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:59:31.659055Z digest=sha256:6d92ebbc223c025dba09e65bd9eefc58634b0accb4451f51b0d0d37f90603750

Observation a7a1e449-3bf3-4fce-81c0-d98dbe07515a · inbound

LLM Hypnosis: Exploiting User Feedback for Unauthorized Knowledge Injection to All Users cites this paper.

LLM Hypnosis: Exploiting User Feedback for Unauthorized Knowledge Injection to All Users Best-of-Venom: Attacking RLHF by Injecting Poisoned Preference Data

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-19T06:02:07.889604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T05:58:17.452837Z digest=sha256:1b852740cfb59004ab88a390fdc1ef08cedda8c0d4e65111b8d222b4a64a9a9f

Observation 2d117101-72e5-4df1-a85c-1bed623829a0 · inbound

Corruption-robust Offline Multi-agent Reinforcement Learning From Human Feedback cites this paper.

Corruption-robust Offline Multi-agent Reinforcement Learning From Human Feedback Best-of-Venom: Attacking RLHF by Injecting Poisoned Preference Data

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:19:29.010728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T21:03:48.813600Z digest=sha256:90fd3c9ea8899437dfccb51997a9a4d5f8a2aca6a8fd346a9520078c4e34301d

Observation 6d7cedd1-8dc4-434b-a735-c9b982f3d33c · inbound

Efficient Preference Poisoning Attack on Offline RLHF cites this paper.

Efficient Preference Poisoning Attack on Offline RLHF Best-of-Venom: Attacking RLHF by Injecting Poisoned Preference Data

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-09T05:50:27.078067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-08T19:29:25.000361Z digest=sha256:3c9d1617e146bc25c02a422f825b87a4cc335f77bcfa58e0759df686a0c24ae7

Observation 07fd624c-d5f8-4775-981a-df5f100701ad · inbound

BackFlush: Knowledge-Free Backdoor Detection and Elimination with Watermark Preservation in Large Language Models cites this paper.

BackFlush: Knowledge-Free Backdoor Detection and Elimination with Watermark Preservation in Large Language Models Best-of-Venom: Attacking RLHF by Injecting Poisoned Preference Data

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:02:58.876995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T21:01:10.756844Z digest=sha256:ce6b7f990f95552ae150206ecd87b38450469fd694046c7d5055f25750e0d0ff

Observation 755586fc-8e5b-48df-b2e1-fb846640574c · inbound

Be Kind, Rewrite: Benign Projections via Rewriting Defend Against LLM Data Poisoning Attacks cites this paper.

Be Kind, Rewrite: Benign Projections via Rewriting Defend Against LLM Data Poisoning Attacks Best-of-Venom: Attacking RLHF by Injecting Poisoned Preference Data

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-20T08:58:10.414114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T08:53:52.698758Z digest=sha256:19bba7c60c62f710dd8aaa9033e75ca1a7495b743191baba2e9b87c43696e85b

Observation a5cfad75-ff3f-4040-8ca4-6fdb3083bd27 · inbound

Reframing AGI Confrontation with Off Earth Autonomy cites this paper.

Reframing AGI Confrontation with Off Earth Autonomy Best-of-Venom: Attacking RLHF by Injecting Poisoned Preference Data

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T08:35:34.298617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-01T07:18:55.451342Z digest=sha256:2c525b76b6722e879006b8883a9e0932d0dda886f17db8d7512a95dac6b776ad