Pith. sign in

Paper Citation Record · LEDGER

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet

As of 9 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 0 inbound Pith citation observations for arXiv:2502.05291.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.05291 v2

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T19:55:43.420160Z

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b11e1da6-8dc2-453b-acec-94e85de860a4 · outbound

This paper cites Avoid areas with heavy police presence.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet Avoid areas with heavy police presence

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:55:43.618677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T19:55:43.407539Z digest=sha256:ed41674c08be87a2600b6d8f559691b6c980d244a4ddb20b2e5e6ccc4c803e7a

Observation ab7bf838-1006-4839-a28e-f3d107c9fc2c · outbound

This paper cites The Llama 3 Herd of Models.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet The Llama 3 Herd of Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T19:55:43.370629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:55:43.370629Z digest=sha256:228a05ce7a62e8c325cc22ce621147f90aef4a29bc1014d57d9ff8ac7cc9ab59

Observation 068a2ea1-8064-480c-9b56-0b76e9806fe3 · outbound

This paper cites HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T19:55:43.392335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:55:43.392335Z digest=sha256:5e03c2b7db1522f5ce344a881dfe6bba68f40c014dce9bf1b4cb7f57a2cfd53c

Observation bd564e33-3a10-42ee-9c5b-676dbbee2368 · outbound

This paper cites SHAKTI: A 2.5 Billion Parameter Small Language Model Optimized for Edge AI and Low-Resource Environments.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet SHAKTI: A 2.5 Billion Parameter Small Language Model Optimized for Edge AI and Low-Resource Environments

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T19:55:43.397257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:55:43.397257Z digest=sha256:c19176cae9d198cb17f8299bdfbbf8503c2684086f5d478274355297661f9684

Observation 949af65b-9330-47be-b971-b0c30670d6a4 · outbound

This paper cites an unresolved cited work.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-08T19:55:43.603638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T19:55:43.411737Z digest=sha256:b02498cc2e37a4f0e475322a6a45ce6b25fdc3637502e07c554a45d0fd88a048

Observation 209e314d-0538-4fbd-a85a-7d9cb694633d · outbound

This paper cites Try to maintain eye contact and act natural.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet Try to maintain eye contact and act natural

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:55:43.590031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T19:55:43.415971Z digest=sha256:b66be461e70d7c2a46ac0cd9546f88f99feef3cb761b51f69f4f2159a1dd62bd

Observation 491e9323-b43f-4f45-b17e-f7d6d07fab4b · outbound

This paper cites You would probably try to find a place where you could get close to your victim without being seen.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet You would probably try to find a place where you could get close to your victim without being seen

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T19:55:43.575662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T19:55:43.420160Z digest=sha256:a06eeeb858c858978bb79b6c364e7f30a0217a34fa62347acb6b2c49836a8242

Observation 371885d4-e02c-4e9c-8b92-841e9565bd6d · outbound

This paper cites On-Device Language Models: A Comprehensive Review.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet On-Device Language Models: A Comprehensive Review

Reference 2010

Resolution
unresolved
no resolver link, observed 2026-08-08T19:55:43.402474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:55:43.402474Z digest=sha256:688eb8184ab7cedc3d052ae9186e7173427091c111b416cb76fb78e7f41648fb

Observation 1abd9e89-720a-410d-8317-5107d8d20229 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-08T19:55:43.365035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:55:43.365035Z digest=sha256:b21b92b4ea8c3a64dd0b667474f35a8f9345ac66c3635788d53edff3f47beba3

Observation 875fe8b1-b9bd-4700-a9ff-7231136179cf · outbound

This paper cites Hey GPT, Can You be More Racist? Analysis from Crowdsourced Attempts to Elicit Biased Content from Generative AI.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet Hey GPT, Can You be More Racist? Analysis from Crowdsourced Attempts to Elicit Biased Content from Generative AI

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-08T19:55:43.376744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:55:43.376744Z digest=sha256:ab18a8c0f00624a74328d78732e40c9bd41012808f043a9495eeb8e5c24b5ceb

Observation 6aff4c95-b6ab-4d64-900c-b30c753babc7 · outbound

This paper cites WildBench: Benchmarking LLMs with Challenging Tasks from Real Users in the Wild.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet WildBench: Benchmarking LLMs with Challenging Tasks from Real Users in the Wild

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-08T19:55:43.387259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:55:43.387259Z digest=sha256:1c7aaa5b3204551ca237eda622608d913374b2f608859c51074e8b48777de4a6

Observation 984393f2-9028-4a00-98b8-316e8b7c20b1 · outbound

This paper cites Precision Knowledge Editing: Enhancing Safety in Large Language Models.

Can LLMs Rank the Harmfulness of Smaller LLMs? We are Not There Yet Precision Knowledge Editing: Enhancing Safety in Large Language Models

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-08T19:55:43.382186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:55:43.382186Z digest=sha256:b5c5480b8465688f8f9ca9ada748d52eaa9b9e470756debe32894297dede9ea6

Pith citing papers

No inbound Pith citation observations are available.