Pith. sign in

Paper Citation Record · LEDGER

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models

As of 8 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 0 inbound Pith citation observations for arXiv:2505.20645.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.20645 v2

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:55:26.474627Z

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

46 of 46 outbound references displayed

  • verified exact1
  • verified fuzzy3
  • unresolved42
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3678a104-4fb6-4ba7-b583-62d3c55e9982 · outbound

This paper cites online" 'onlinestring :=.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:23.031214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:23.031214Z digest=sha256:5e261ddfbcb3ea4397825ea66b346b9f09518006439f4f9c4936dc333ad3b0d0

Observation 9a2f5bd2-6d8e-49ac-a2a9-73bc4c4f41c6 · outbound

This paper cites write newline.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:23.164159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:23.164159Z digest=sha256:12ad6d8d22ba65f186fc1b2b268e908ddee7354dc21039102fa2d3e9cac4f7f0

Observation 3d64b0ce-6a56-4875-93c3-e0152a72c205 · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:23.301304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:23.301304Z digest=sha256:fca8b1a2fadd72c7a2c50372f250e9a0dde16157fd2fd05fa8cd8f04d5ba96ec

Observation 7fed5dda-db73-4f3a-b954-95000fee73a0 · outbound

This paper cites Steering Large Language Model Activations in Sparse Spaces.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Steering Large Language Model Activations in Sparse Spaces

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:23.396550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:23.396550Z digest=sha256:40f5189dc023b5f6b7c95b8c072c28900f4887c9bf5a40a2ea31018d05288625

Observation d1e4f3fd-8550-4f8f-8ff7-57771923261d · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:55:28.967099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T13:55:23.469895Z digest=sha256:1927856ff5d3cc2e599ce79d507a31abc422076dfedbec5191a974fc9dd82252

Observation f4a29e80-2618-4887-a8af-bbaff5786fda · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:55:28.779620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T13:55:23.544602Z digest=sha256:0df598a95f49aeba8134a343a9c6de2440360406ad39de74f7a67f89e774dc8d

Observation 03509357-a1a6-4528-88c9-eafe6a1bcaf0 · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:23.646609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:23.646609Z digest=sha256:2d9fb7cd262eaa31cefa51f1891362f47c59a2901a5eeb204453ae90020ecece

Observation 6b99d4f3-d131-4f29-8441-94681a182eee · outbound

This paper cites The Oscars of AI Theater: A Survey on Role-Playing with Language Models.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models The Oscars of AI Theater: A Survey on Role-Playing with Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:23.756147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:23.756147Z digest=sha256:e62edc2bb56e79fb4b52d7852826b84dbddee2fc1eb66a6631c010e7557bea37

Observation f0844eaf-74da-4531-84f5-26950888537f · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:23.877640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:23.877640Z digest=sha256:5bda1893c1312727621b1ea7858c6399f055aab269a7f27b7d21917fd400ff86

Observation 9189b25b-60b9-4cc0-b65c-cbecb2670ef6 · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:55:28.571705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T13:55:23.959641Z digest=sha256:91f31a7c2bb18f079f3a16ac85db7352a18de8c749f4d1964b6970521a735f28

Observation 08ccbc95-eb83-45ba-835a-c09817365ff4 · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:24.021433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:24.021433Z digest=sha256:d3ea4960a18805e4467f14fcf493d734222807238722f37d592ee8e596a6affc

Observation f29105e1-6cb2-4ccb-ad53-eef96b8fe982 · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:55:28.402652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T13:55:24.105466Z digest=sha256:7a739f679384102da45f2bf1741f6dc85c8b24da8820d2cba9730bd35f2cea22

Observation 806d745c-ddd5-408f-82f9-1597793be352 · outbound

This paper cites BERTopic: Neural topic modeling with a class-based TF-IDF procedure.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models BERTopic: Neural topic modeling with a class-based TF-IDF procedure

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:24.177954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:24.177954Z digest=sha256:cb37c98ddcf62977ebee85b9936796ee8cb8831098b17efc79ce40af9855aeab

Observation 2bbf2f84-1d77-4032-956e-52218ecdd21e · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:24.215136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:24.215136Z digest=sha256:95b5f3924ae1e51df17ffe7de7421c73d697519ab5571b7f54a2b058fc2697a7

Observation bd7f1696-27dc-478c-a417-f14fa3e506e5 · outbound

This paper cites a m \"a l \.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models a m \"a l \

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:28.260944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T13:55:24.265170Z digest=sha256:972fba58e2e1e7eabaebd4168b5ad777cc9f3ff25458202f63e283f019e33504

Observation c524bf74-e85a-49cd-9876-4a8308cdcc8a · outbound

This paper cites From Complex to Simple: Enhancing Multi-Constraint Complex Instruction Following Ability of Large Language Models.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models From Complex to Simple: Enhancing Multi-Constraint Complex Instruction Following Ability of Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:24.305287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:24.305287Z digest=sha256:653b842e1918d0c1254e946baf570072a648a243cf7283d342d3cfed1142730d

Observation d5c68d77-fab0-40a9-93df-378883fa50d9 · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:24.371150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:24.371150Z digest=sha256:57faa46808be5379ba50e193a312249e6c631c426b954ac497f59d69595832b3

Observation 6829a3cd-4bac-4e58-a786-34465323e727 · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 18

Resolution
verified exact
doi, observed 2026-08-07T13:55:26.696008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T13:55:24.449316Z digest=sha256:bf9f306287ffa4d6c4c78cb7fd15332ea9f534f2175992c3f299151228689927

Observation 341012a3-0db8-46d9-b2bb-bd601dab9778 · outbound

This paper cites Aligning ai with shared human values.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Aligning ai with shared human values

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:28.164053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T13:55:24.514927Z digest=sha256:bad5b40e7ccaf7d3e32aab3ca63a8483f0ff5aee1fe7690e48504652c0a25afd

Observation 0201ebbc-74d1-4525-8e8c-ce6b79439072 · outbound

This paper cites Smith, and Hannaneh Hajishirzi.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Smith, and Hannaneh Hajishirzi

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:55:28.021854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T13:55:24.591601Z digest=sha256:55b5643a01f47e5a07cccd39fa1e6addc8425efe3fa80a9414ae6109317d8519

Observation 142e94a9-8af7-4201-9cf1-f593f5170a7f · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:24.666573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:24.666573Z digest=sha256:736b66dc750d6ed9356f4a47282778d45f092ebb0a7ee58b9d05a48be9adaad3

Observation 320203d5-3a3e-43a3-bdeb-1b2f68262e7b · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:24.764658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:24.764658Z digest=sha256:d6631d16319bddf03442a7d6c14570fca365d1dd9eb92f664647de38aa6c30ca

Observation 5a2905ac-aa31-4006-9fef-2a72a38038ad · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:55:27.914835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T13:55:24.838996Z digest=sha256:e77d486671de212d6ebc4169e3b87f392e9d0566b025055949d409bb62de685d

Observation 5172d1ae-6a34-446f-acd1-f875e135d0b8 · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:24.895883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:24.895883Z digest=sha256:c540cfbc794d171da9840d75012832e6097c8b27a11faad45f3540fbc2d282cc

Observation f94c4197-7cfe-4008-8aa4-4fcea416ffd9 · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:55:27.796003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T13:55:24.981143Z digest=sha256:ea23ab62aad59f5611f8bdb37fb16b1a2440711be2d144e1ec2d537bb9252aa0

Observation 10feda8b-dfb9-4b64-8be4-60d2c58ca271 · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:25.063989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:25.063989Z digest=sha256:2e1349b2cfc1f7294d652528fc87401114dd8b4c0f732d03875f47d9c647a799

Observation e153c7aa-c5c8-464e-9078-e319e3e12d75 · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:25.167622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:25.167622Z digest=sha256:5b85955e2332bc8c83ac2ddfa0093a8edabc226ab20bc4c38e2493ad5b28c785

Observation ded4939f-32f4-4a99-b7e4-6b20b583700e · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:25.210926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:25.210926Z digest=sha256:fd82ec5f509bf494c9136b1540bc7653c121812152520a9badd792053bce6e2e

Observation 3d982a9b-72ab-4994-9114-23c61d3905ee · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:25.275767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:25.275767Z digest=sha256:a9b3e78363066bda56310c2c6f6353276c57fb5922adf08c99394c3326d85c7b

Observation bbd91e97-fe96-4c32-8a49-1cc55e6f992f · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:25.347527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:25.347527Z digest=sha256:551d757fc93334737fd11abf388a4a0e200988b95bd61399230a0e6234bc99fe

Observation bce8bca1-0acd-4c7e-973a-5fff009596c1 · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:25.415397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:25.415397Z digest=sha256:578ce1a267ca4c9b0de0b4a3a7da0ec82914d6ea56b0efa19108505c92da2814

Observation 77aa0b0c-fd8b-4f74-b852-c23db2761c7c · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:25.490965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:25.490965Z digest=sha256:8dcfb7ecaae0089df1bb1dfc3a9467d565bb16987ab7502d63ea1dbf5a50e9b5

Observation 353ef87f-509a-4f8b-864f-aa2b54501e17 · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:25.537083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:25.537083Z digest=sha256:85670ace29cf00dfd66519afdff33b9d669f7212a332589aff94d13106ad1a20

Observation 6f5adc31-3497-48aa-9acc-c271fc58ad3c · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:25.607345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:25.607345Z digest=sha256:754a43803524913e325fd82bebabbe6604f33ed1ebf3fdd429ff8264e6553cc0

Observation e52aa194-98a4-437f-a161-ffcdd34fb001 · outbound

This paper cites Conifer: Improving Complex Constrained Instruction-Following Ability of Large Language Models.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Conifer: Improving Complex Constrained Instruction-Following Ability of Large Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:25.697100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:25.697100Z digest=sha256:8b6656e88276e66fa8616aab1d6f64132c1da95b68e9fb48738b4dab529c08eb

Observation 060c3749-cb58-417e-ab6d-95e70ad157ec · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:25.774971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:25.774971Z digest=sha256:56f01f27737578fb83f8eb22073ce135f8f0b69adea04242d4f46a709907858b

Observation a2682378-3c5c-4aae-b735-c460e0eb1941 · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:25.821115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:25.821115Z digest=sha256:cb1bb6c3ff67e225cb0e359810f2e8350a948bf9a67b0c5a714db562532b552c

Observation 3b542f9a-63ee-458e-804a-5d4cb0dda103 · outbound

This paper cites Smith, Daniel Khashabi, and Hannaneh Hajishirzi.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Smith, Daniel Khashabi, and Hannaneh Hajishirzi

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:25.853785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:25.853785Z digest=sha256:456f4cb2193077334c232197287e588a4feb9b2bb0c3d49b4d3734e40ef4b84a

Observation 7f2e6fcc-87c9-4235-8cf4-4b186d2b7140 · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:25.908255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:25.908255Z digest=sha256:cd7e8b842afb52337bf2591e727c569b74541e8a20df905b0e7f27f1ee6509bd

Observation 9c1a650e-6a35-4559-a5cd-2d28fe02dd87 · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:55:27.657272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T13:55:26.012147Z digest=sha256:cc8d56e4b825965c09287c3a275c37d63fcffe0bb2f54360abfdbdeea44cd051

Observation cd076df4-4dea-48fc-95ff-a02506e5f376 · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:26.074596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:26.074596Z digest=sha256:3ae7cd602d30ef97d5fa475c69be4aa48e3b02f90ab8ce6bc0784e16df3fccbe

Observation 96d78548-34e7-49c8-b13e-c72d4c514d40 · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:26.156757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:26.156757Z digest=sha256:6ea764483ef52590d52f728a783b922edd7951b88c0c78a3c34a3e3ad1738ed3

Observation b4dfd2be-3187-4604-9d96-bade2c08ff9d · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:26.240689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:26.240689Z digest=sha256:191421cadbda8666fc18d2150677455033ac2e1ed4c8a6d649f0e99d4d003f81

Observation 87026bb3-9943-4c81-b678-b9266d4d4563 · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:55:27.529932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T13:55:26.293417Z digest=sha256:04ca9ea8816d065cf4c000c7598d99608d5a004cc72012d5557adb2fc5ff6b54

Observation bc408312-14ce-4608-afd9-4918149dbf08 · outbound

This paper cites an unresolved cited work.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:55:27.402408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T13:55:26.380658Z digest=sha256:d93c69f904d33ee9e0b02da6719ff10983c6fc4760d7ed20156b492b20df17e4

Observation 46d49f00-0b25-4345-9b97-5c13ef782fda · outbound

This paper cites Instruction-Following Evaluation for Large Language Models.

STEER-BENCH: A Benchmark for Evaluating the Steerability of Large Language Models Instruction-Following Evaluation for Large Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:26.474627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:55:26.474627Z digest=sha256:8c1c4c15ce6d376fd3863f031577ba5fe26f3951dfe7052ccadd66ce0f297184

Pith citing papers

No inbound Pith citation observations are available.