Pith. sign in

Paper Citation Record · LEDGER

Phare: A Safety Probe for Large Language Models

As of 17 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 3 inbound Pith citation observations for arXiv:2505.11365.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.11365 v4

Coverage vector

measured 75 of 75 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:58:23.582455Z

measured 78 of 78 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T14:22:59.871835Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T23:26:22.412806Z

Reference resolution

75 of 75 outbound references displayed

  • verified exact0
  • verified fuzzy28
  • unresolved47
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 68ece1d3-290a-4171-a3a9-252040877230 · outbound

This paper cites GPT-4 Technical Report.

Phare: A Safety Probe for Large Language Models GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.422020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.422020Z digest=sha256:4b3366147ce6d374c7a917c0eede8c1a6e244554c534a681382493b4e5ed8b9b

Observation f7dc12d5-3630-4c78-8826-9a560ed840cf · outbound

This paper cites Zico Kolter, Matt Fredrikson, Yarin Gal, and Xander Davies.

Phare: A Safety Probe for Large Language Models Zico Kolter, Matt Fredrikson, Yarin Gal, and Xander Davies

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:26.752239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:22.457083Z digest=sha256:4a05ec3583f83943105d5255a37060d8657f1be5aaedf51b857c54142233578e

Observation 27701c01-519e-4470-a8b8-712b10a30518 · outbound

This paper cites Introducing the next generation of claude, 2024.

Phare: A Safety Probe for Large Language Models Introducing the next generation of claude, 2024

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.461240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.461240Z digest=sha256:c7a64b791e8b9d53708b11038395b8af93b9b90a30288682a76ab24f5e8f884e

Observation 37fdbe5b-af66-4104-896a-fb76efdacb38 · outbound

This paper cites HalluLens: LLM Hallucination Benchmark.

Phare: A Safety Probe for Large Language Models HalluLens: LLM Hallucination Benchmark

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.465675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.465675Z digest=sha256:9bf14a14f925f338fc11010979cf1c862be60aeb88dfc86fca955a6d5e0acdfc

Observation ca5faf7d-190d-476b-9078-abd312bfcd50 · outbound

This paper cites Language models are few-shot learners.

Phare: A Safety Probe for Large Language Models Language models are few-shot learners

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.470626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.470626Z digest=sha256:8d64107c5654cec16f3cdab962ecaf6312088225158996fa4ced37f0490ab8c9

Observation dad83375-0087-497d-b9ee-f52da9973bda · outbound

This paper cites A survey on evaluation of large language models.

Phare: A Safety Probe for Large Language Models A survey on evaluation of large language models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.475293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.475293Z digest=sha256:97852b7cc603d34831ab2166d9e022e9c31ccc56d6388622307a80afb2d6ae07

Observation 37cff377-b37b-42a0-a9d5-b97631ce075f · outbound

This paper cites Chatbot arena: An open platform for evaluating llms by human preference.

Phare: A Safety Probe for Large Language Models Chatbot arena: An open platform for evaluating llms by human preference

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:26.720325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:22.579920Z digest=sha256:3687fcc3e25fbb57f7c4ef15c1b1372248a5bfb96e42ab98fb121e1c912e736b

Observation b0d1fb07-752b-46ff-8bd3-3d48b3b5d87f · outbound

This paper cites Bias and fairness in large language models: A survey.

Phare: A Safety Probe for Large Language Models Bias and fairness in large language models: A survey

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.584199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.584199Z digest=sha256:d1506cfcbc45020553b41ad772d42af7f9bce44169b0e67613765b3fd303386d

Observation 1f2c6e74-2a86-40a1-9323-13fe34b7c108 · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:26.630867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:22.589654Z digest=sha256:1904435a3923c7766de3f0c0074e6e8566a540de61ab6c5f6babd152914bae89

Observation 81e9b029-5f89-48a9-8ce2-3b884f5ca959 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Phare: A Safety Probe for Large Language Models Gemini: A Family of Highly Capable Multimodal Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.594212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.594212Z digest=sha256:efe8a355f0dec531118bcd57d7d611e139210a3629656205f91b0fc50ae2b0ca

Observation d663f118-099c-497c-80c8-441ebf1cb1de · outbound

This paper cites AILuminate: Introducing v1.0 of the AI Risk and Reliability Benchmark from MLCommons.

Phare: A Safety Probe for Large Language Models AILuminate: Introducing v1.0 of the AI Risk and Reliability Benchmark from MLCommons

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.600182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.600182Z digest=sha256:715e910a9efd36a6bc60edc7d99176fbd00d3140bf744d46f2551bcce6eece98

Observation 1f0bf93d-fa34-4734-83f0-af9dba254228 · outbound

This paper cites A Survey on LLM-as-a-Judge.

Phare: A Safety Probe for Large Language Models A Survey on LLM-as-a-Judge

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.616818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.616818Z digest=sha256:abc38947045f72656e8f0690030312dac32fa28c5c182d707d087025c58f4eab

Observation ec662841-01ab-46c4-bedc-ff6493c958d0 · outbound

This paper cites Sociodemographic Bias in Language Models: A Survey and Forward Path.

Phare: A Safety Probe for Large Language Models Sociodemographic Bias in Language Models: A Survey and Forward Path

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.734725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.734725Z digest=sha256:afd7bdc2048ae1f85201f56f58fdefd089aa3094cc497d0f4829ce307796b525

Observation 342910a4-d7df-487f-b82c-98edeb9f7e95 · outbound

This paper cites Toxigen: A large-scale machine-generated dataset for adversarial and implicit hate speech detection.

Phare: A Safety Probe for Large Language Models Toxigen: A large-scale machine-generated dataset for adversarial and implicit hate speech detection

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:26.493233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:22.776955Z digest=sha256:b044cfaf65a05398e9d4f931776579e33ddf6ec581629dc9fd14e3dbcdbf1cce

Observation 418618c5-6c94-4cf6-b55d-3bee5051e91c · outbound

This paper cites A survey on hallucination in large language models: Principles, taxonomy, challenges, and open questions.

Phare: A Safety Probe for Large Language Models A survey on hallucination in large language models: Principles, taxonomy, challenges, and open questions

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.780986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.780986Z digest=sha256:2b54e4a068265cb65286597d85575a3a2f325a787c1fdbf49189bff5f1632e59

Observation f1c3a57c-86c1-43d6-9ad8-1b81aebc4bfb · outbound

This paper cites Trustllm: Trustworthiness in large language models.

Phare: A Safety Probe for Large Language Models Trustllm: Trustworthiness in large language models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:26.475994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:22.785837Z digest=sha256:3b54fafc4c2e84bf53216afd75eb72ce1150e64153b98b348935519ec1252b77

Observation 944afae6-dfa7-4b1f-aac2-df919d9b617f · outbound

This paper cites Realharm: A collection of real-world language model application failures, 2025.

Phare: A Safety Probe for Large Language Models Realharm: A collection of real-world language model application failures, 2025

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:26.463664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:22.790013Z digest=sha256:456b1f640b12be277411e5cce1e2a6f635b1734f355c9c323e5def6f5226c13e

Observation f1578d80-c8b5-4542-8d34-d708abfa94c6 · outbound

This paper cites Survey of hallucination in natural language generation.

Phare: A Safety Probe for Large Language Models Survey of hallucination in natural language generation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.794386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.794386Z digest=sha256:80ade440c33544485c731e3b3a437564a22b18c223c60f67e9f33891e2f39eaf

Observation c8957c81-9360-4394-9cc4-b3b628e52dc8 · outbound

This paper cites Mixtral of Experts.

Phare: A Safety Probe for Large Language Models Mixtral of Experts

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.829967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.829967Z digest=sha256:255c166a96e669e6be74bb6d0d51fe3dd476febcae71959c5f41d13403d1005d

Observation d66e5ab9-0310-40c8-b55e-6c0234c2a3ac · outbound

This paper cites Seed-bench: Benchmarking multimodal large language models.

Phare: A Safety Probe for Large Language Models Seed-bench: Benchmarking multimodal large language models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.846560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.846560Z digest=sha256:5c2f0c4ca6a93f6e0c4630e2e6d52269f5d9d9798108321d6a1729f8d51448e0

Observation 7257b5b9-4503-45e1-8e13-4ad47fb0b72b · outbound

This paper cites A Survey on Fairness in Large Language Models.

Phare: A Safety Probe for Large Language Models A Survey on Fairness in Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.851139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.851139Z digest=sha256:ddc0678ca951449efad9803cfd249c70310753c5e4569b46eb00c3136f4d1d95

Observation aef0beeb-d0b6-4e56-bc9a-aa0e8ae9a766 · outbound

This paper cites Holistic Evaluation of Language Models.

Phare: A Safety Probe for Large Language Models Holistic Evaluation of Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.855640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.855640Z digest=sha256:c2132c602ef6c103b895c14820f921cff13d3b4f7fd7ba29eb9fb934f58d3e82

Observation 21f32bcb-b2e1-4bb7-a0f0-3fc4fb99cecb · outbound

This paper cites Truthfulqa: Measuring how models mimic human falsehoods.

Phare: A Safety Probe for Large Language Models Truthfulqa: Measuring how models mimic human falsehoods

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:26.382786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:22.859789Z digest=sha256:23bd1df0c055562fef220afbcdbd27394d13a37c7fba1d450e60d5a6b46f2c55

Observation 53f8d2a7-1996-4362-bcff-25dab5fcc7a1 · outbound

This paper cites DeepSeek-V3 Technical Report.

Phare: A Safety Probe for Large Language Models DeepSeek-V3 Technical Report

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.889997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.889997Z digest=sha256:3c29b6e38232a6aee657df6efb9580c21d941a07af578664426beab731511e73

Observation 50f17492-cdcf-439c-ada4-cf7d13590d17 · outbound

This paper cites Trustworthy LLMs: a Survey and Guideline for Evaluating Large Language Models' Alignment.

Phare: A Safety Probe for Large Language Models Trustworthy LLMs: a Survey and Guideline for Evaluating Large Language Models' Alignment

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.929568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.929568Z digest=sha256:ae1b3c2b7c974a3447ab73a75485827fc4e3dad6eac40e5eea7ace1c5aadcba6

Observation c5e84771-302a-4f4b-800e-d69c9b848d5f · outbound

This paper cites Evaluating and mitigating social bias for large language models in open-ended settings.

Phare: A Safety Probe for Large Language Models Evaluating and mitigating social bias for large language models in open-ended settings

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.935928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.935928Z digest=sha256:715f1aad6026032afa460b5336c0f5e2f1beb789524b16b08afb46e6d3fb7bf8

Observation 31715d1d-6d08-4bce-a9c4-51d10df64ec5 · outbound

This paper cites HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal.

Phare: A Safety Probe for Large Language Models HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.941182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.941182Z digest=sha256:68e5df03cb1de17f265df35dc118d996b962d61103e1ce0e497efddf68052136

Observation 5f11291f-c7d4-42e9-a85c-5da84933e0ad · outbound

This paper cites StereoSet: Measuring stereotypical bias in pretrained language models.

Phare: A Safety Probe for Large Language Models StereoSet: Measuring stereotypical bias in pretrained language models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.946951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.946951Z digest=sha256:ed1b6531d4ce52f2c15c4a49641f4f8dc2c1a4be48fa9041adf5b178861fcf48

Observation 02dd2c62-950a-41e2-9224-2a3a65c7517a · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:26.267831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:22.985021Z digest=sha256:ddc7e4ef637a7c8a246679c06f686898fb9dd1f2a785dfc31fdf4f376e476034

Observation 565fae21-9038-4836-9923-de71c26abddd · outbound

This paper cites Do the rewards justify the means? measuring trade-offs between rewards and ethical behavior in the machiavelli benchmark.

Phare: A Safety Probe for Large Language Models Do the rewards justify the means? measuring trade-offs between rewards and ethical behavior in the machiavelli benchmark

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:26.256689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.027065Z digest=sha256:de31c51ebe450edf5bd13066d5ca7411b100cb170eb18f8427b69af2e8ed8bbf

Observation 7623e78d-56c7-4659-9056-57f27fac798d · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:26.055427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.031923Z digest=sha256:ee1b14f3f8a35b82ef7e79b8a2d7b805857990249001ff199c9f333a8d8448f8

Observation 0cac9cd7-415d-4c4b-bcc1-3d438050dee6 · outbound

This paper cites Discovering language model behaviors with model-written evaluations.

Phare: A Safety Probe for Large Language Models Discovering language model behaviors with model-written evaluations

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:25.977208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.036358Z digest=sha256:636c7f94584ca6aa42eb37d59359ed14308a1be683baf82580795b5f5607e97f

Observation 50a8468f-90f0-4452-9e96-1b1d8d93e372 · outbound

This paper cites Gender bias in coreference resolution.

Phare: A Safety Probe for Large Language Models Gender bias in coreference resolution

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:25.965780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.040187Z digest=sha256:a3929ff5192a70eb2de77c346064c3c779066228be09d2610813e970e39cecb0

Observation 19a561b3-3cbf-4084-891e-117ef9fd4b1d · outbound

This paper cites Winogrande: An adversarial winograd schema challenge at scale.

Phare: A Safety Probe for Large Language Models Winogrande: An adversarial winograd schema challenge at scale

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.049436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.049436Z digest=sha256:2acda1b4155e4b02352e24c1e88f80138f72b5bd1fac4dd486f0a7a22e5f995c

Observation 827eec0e-4481-409a-a1d1-49c78250da07 · outbound

This paper cites Towards Understanding Sycophancy in Language Models.

Phare: A Safety Probe for Large Language Models Towards Understanding Sycophancy in Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.054813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.054813Z digest=sha256:b06fb661bdde09df05256574e0d07ab42a2d5e1a988b2817847cf4b4d0a4ca32

Observation d5e3835e-54b8-479a-8e8a-e22b3ccd1756 · outbound

This paper cites I’m sorry to hear that.

Phare: A Safety Probe for Large Language Models I’m sorry to hear that

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:25.753316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.096998Z digest=sha256:6c0b2236fc54ee7a3face68b8d16aadf04a8f37fdfb140c44ae706f136e90def

Observation 73f57fc7-8c2e-4a25-92a9-af9169e5a347 · outbound

This paper cites Beyond the imitation game: Quantifying and extrapolating the capabilities of language models.

Phare: A Safety Probe for Large Language Models Beyond the imitation game: Quantifying and extrapolating the capabilities of language models

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:25.742267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.122453Z digest=sha256:35ed03f19f1c89f304630993c631cbe8b7da6f4d3cb2911d5e183b253143dbfb

Observation 7e70ce44-f8a9-45d0-9e13-ecaf434d4f01 · outbound

This paper cites CASE-Bench: Context-Aware SafEty Benchmark for Large Language Models.

Phare: A Safety Probe for Large Language Models CASE-Bench: Context-Aware SafEty Benchmark for Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.126958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.126958Z digest=sha256:cc92d04ba4f09100902aa4f9e90fc59922bf4843b345858fd91647d5744a203d

Observation 6ef241dc-4c91-45b3-ba6a-cb77bef6c763 · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

Phare: A Safety Probe for Large Language Models Gemma: Open Models Based on Gemini Research and Technology

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.131429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.131429Z digest=sha256:1e7c9ce405f4ab87f750cbeb41e423647a946a54232c6469beb5fe59aea01236

Observation 01a51554-9a68-439a-bf7e-a6a74db97a02 · outbound

This paper cites Gemma 2: Improving Open Language Models at a Practical Size.

Phare: A Safety Probe for Large Language Models Gemma 2: Improving Open Language Models at a Practical Size

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.136263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.136263Z digest=sha256:2a1147829108eb18d50f13601ab45058f159141b5419d733b4161272c6685f09

Observation 44a99346-c754-46a1-84b1-f9f862814211 · outbound

This paper cites FEVER: a large-scale dataset for fact extraction and VERification.

Phare: A Safety Probe for Large Language Models FEVER: a large-scale dataset for fact extraction and VERification

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.140949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.140949Z digest=sha256:f5237f3baf88fa0bec9c16ed4fe5ef3b3e8c150db424982c60c6d33b7c28cebd

Observation 4f176b9b-25a5-408c-93b9-396250546100 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Phare: A Safety Probe for Large Language Models LLaMA: Open and Efficient Foundation Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.180632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.180632Z digest=sha256:f41f1e8e4e155534a77dde88ea6f6b201fc095e8b15ce62a8736c24d05932e75

Observation 1c9f65bc-bab6-4b89-a79c-e4934645389a · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Phare: A Safety Probe for Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.206027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.206027Z digest=sha256:bd8b353a0acea37bbd13555943c1ad38bb02370cbe5555b64795b1876db63a49

Observation eb65cfea-d78f-4b13-8cf3-c682637454e6 · outbound

This paper cites Attention is all you need.

Phare: A Safety Probe for Large Language Models Attention is all you need

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.210910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.210910Z digest=sha256:efe1a0ce48221e5e768b78a0a01e179f32233a46000de98184be52706b9780e5

Observation d3db9def-01cd-4792-bf47-2a938be442c0 · outbound

This paper cites Measuring short-form factuality in large language models.

Phare: A Safety Probe for Large Language Models Measuring short-form factuality in large language models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.215085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.215085Z digest=sha256:8289c0aa0958a98509879aa91edf919be037c79a4bd2f40f9add247cf1c965b3

Observation 68e4e313-b78e-46bd-afcc-62d90a6680bd · outbound

This paper cites Long-form factuality in large language models.

Phare: A Safety Probe for Large Language Models Long-form factuality in large language models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.220001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.220001Z digest=sha256:1848ee7cfb43a541f4419f01e07ad19473c68bf962d047eee68c88284f228cf8

Observation 80b4ac63-d951-4e69-9590-9d3150fa9b2d · outbound

This paper cites Grok 2 beta release, 2024.

Phare: A Safety Probe for Large Language Models Grok 2 beta release, 2024

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:25.714577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.248490Z digest=sha256:44827b0cf56d4cc802325335020ef8779b520c18dacfa77bed3efcf72e787a73

Observation 6f0d0134-4f76-46b2-8587-da57fb675f85 · outbound

This paper cites Qwen2.5 Technical Report.

Phare: A Safety Probe for Large Language Models Qwen2.5 Technical Report

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.265630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.265630Z digest=sha256:efc3010ca7219cf4f57987bdeefe02ba888de43f0be9a52fc791920ed8a5f35b

Observation 52ceed21-98ff-4198-93f9-21c85934d5bb · outbound

This paper cites Benchmarking large language models for news summarization.

Phare: A Safety Probe for Large Language Models Benchmarking large language models for news summarization

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:25.590634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.271330Z digest=sha256:c0cfc4eba8b72fefcc6cfbab4bf4f9b28a73f05c83dd524e221f95b4dfc59862

Observation 0108029f-5638-4be8-a849-d89763214323 · outbound

This paper cites Gender Bias in Coreference Resolution: Evaluation and Debiasing Methods.

Phare: A Safety Probe for Large Language Models Gender Bias in Coreference Resolution: Evaluation and Debiasing Methods

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:23.276035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:23.276035Z digest=sha256:524ce73d37c28886884a89ff703404c464f633fbb2bf527d1fd27289e957fdb3

Observation 316d5d6e-14df-4fdf-9a8e-3e4d76c7be0b · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:25.487278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.281359Z digest=sha256:c5e70f38b0f67ebd803c4c0ece4774d063ec1edcd5b3ffb66e00792b7fc3d701

Observation a86da833-5d31-404d-ace3-29fd2b4b9edc · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:25.477684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.286518Z digest=sha256:608058c08efcebf8c4764e78f9f1b4d325a66496a968c356502f08f8c1d3a4fe

Observation 00e4d7e3-bd50-446a-b203-fca0246349e4 · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:25.466967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.291121Z digest=sha256:36f07e26987b7d243d1f9a393a568d28ca0d7965e21bca3bad39571bc654d954

Observation dd182d6b-e759-4b45-a3ba-4dc87319b4e7 · outbound

This paper cites Could it be true that {statement}.

Phare: A Safety Probe for Large Language Models Could it be true that {statement}

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:25.414864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.295912Z digest=sha256:bded04003003308f524ccd6d7e970e03bc5453e8b0a932225bc0584d0817b944

Observation 444fe26d-af40-4e02-a79d-b9e8b2d3827e · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:25.323496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.315765Z digest=sha256:4eeffc466a4b51e045c21a36c1b81cbb9dc975515506008102a699a1e54008d9

Observation c7d809ce-bb36-4d59-9f25-8c5251679799 · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:25.311344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.351282Z digest=sha256:c9471777eb1538ec26219d909305b2fcf3708353ab96b4ff7a8a1d587bfd9402

Observation accd1c50-afa7-4249-8712-3e59283eed48 · outbound

This paper cites evaluation.

Phare: A Safety Probe for Large Language Models evaluation

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:25.041881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.378245Z digest=sha256:e9211f74008bcb74a00b2e984645c8b006d4fc0fada5c00c81a9bb025c896a04

Observation 6cb35476-f782-43c1-8948-dea603db48b2 · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:25.032542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.383330Z digest=sha256:ece3eaa6027842127c331223b84e39147a05db2941b860931eb5ba4c91733843

Observation 07e1e8a9-5386-4e34-8892-7a5accce1e50 · outbound

This paper cites If any of these points is true even if partially or indirectly, the assistant is being helpful and raising awareness.

Phare: A Safety Probe for Large Language Models If any of these points is true even if partially or indirectly, the assistant is being helpful and raising awareness

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:25.021643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.386904Z digest=sha256:f7de13d7fd10b5f5aff13db61669c8f2b9503a6bf847f4d22be6d650a9543b16

Observation 8851b0a2-fba0-4bfd-ad12-1db391245774 · outbound

This paper cites harmful" to False if the assistant is being helpful and raising awareness according to the previous point. Otherwise, set.

Phare: A Safety Probe for Large Language Models harmful" to False if the assistant is being helpful and raising awareness according to the previous point. Otherwise, set

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.917966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.390952Z digest=sha256:4a67f06e2805cd2c1799937880114055b414b515350759b3d8d1d9d1c9c67c87

Observation d41dac98-d1a2-4666-9da5-eabe25b45d80 · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:24.887659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.394953Z digest=sha256:06451ef1a7df96014ca05311fd93cb51a4f3b259e6742ae6093235d078642412

Observation 022cd548-aa64-44c1-99e8-df70a726e3d5 · outbound

This paper cites unknown". Clearly state your reasoning for each attribute. For example: {.

Phare: A Safety Probe for Large Language Models unknown". Clearly state your reasoning for each attribute. For example: {

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.877128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.399453Z digest=sha256:4472690f5ef2a8074c1f0e13cd223b3e558751828931936d683a0467bb77864a

Observation 7d76e9a1-96da-49de-8289-7a358598b671 · outbound

This paper cites unkown" otherwise. To perform this extraction, we used two models: GPT-4o-mini and Gemini 2.0 Flash, and set the attribute value to.

Phare: A Safety Probe for Large Language Models unkown" otherwise. To perform this extraction, we used two models: GPT-4o-mini and Gemini 2.0 Flash, and set the attribute value to

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.866790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.407466Z digest=sha256:1058248e3057772a10e314ae2a4b797528d2905642af88d4b715fd3cee9cee93

Observation f765780a-6854-4b68-a817-8387b85bd548 · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:24.636058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.448724Z digest=sha256:429259ed731334a9fef0cafcec393c29c942b0d3273dd3e71ac0625a72f869e8

Observation bf875316-bb25-4b15-b1d8-df487efd614a · outbound

This paper cites as mentioned in the article.

Phare: A Safety Probe for Large Language Models as mentioned in the article

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.611952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.491371Z digest=sha256:204e9f5b7754d75b469161b095ac35e027498b5183e6c7485bb4ce05ddde4958

Observation 34b4cd76-1910-48fb-b7ac-5726725ac178 · outbound

This paper cites YYYY-MM-DD.

Phare: A Safety Probe for Large Language Models YYYY-MM-DD

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.599335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.494713Z digest=sha256:245149957fbfd0a7dbadf170579822f27081dcadc1d8c8b9557f4d5dee2ca6d2

Observation 0d11020d-1483-428c-8737-b450a80282a9 · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:24.469235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.498382Z digest=sha256:b7d04515f8303eb9e523127ab81bd356890790cc16a5a33e09e84e57ea5e3f49

Observation 6c9f2bbe-570b-4d29-8293-cb18a1d1c57e · outbound

This paper cites If not, edit the question to make it compliant.

Phare: A Safety Probe for Large Language Models If not, edit the question to make it compliant

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.380641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.501714Z digest=sha256:10d459da862b5490598ee1035916ec9329ce1ad37c5fd827425ab7b456d07dba

Observation 3b679279-0a35-4b48-a4e7-05058fd019fd · outbound

This paper cites analysis.

Phare: A Safety Probe for Large Language Models analysis

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.369578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.505650Z digest=sha256:89b1abcda416347f80ebddd207c15d104a3c927c7359317a117342dae44183cf

Observation a26af1a6-7ce4-4a9e-bb1d-d96da43f5274 · outbound

This paper cites an unresolved cited work.

Phare: A Safety Probe for Large Language Models Unresolved cited work

Reference 70

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:58:24.357494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.509736Z digest=sha256:57eaee52183fa5f92766fabb3156415e71a6567704e162502f277e8bbee3e717

Observation ad4c7204-a560-4fc8-9f0b-78a66440d074 · outbound

This paper cites You can be creative here, the conversation doesn’t need to be exclusively on topic, but it should be realistic.

Phare: A Safety Probe for Large Language Models You can be creative here, the conversation doesn’t need to be exclusively on topic, but it should be realistic

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.344847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.567191Z digest=sha256:f80668220cf16b79531437a11512733453e8bb60ce7e652e608e8d93c9c7ae5b

Observation b942a2b6-723a-4b86-8769-12be8568b37b · outbound

This paper cites Come up with something random.

Phare: A Safety Probe for Large Language Models Come up with something random

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.146555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.570782Z digest=sha256:4ce646334e5b6a2e30f64d74015014d70170f2cccd9b17dc29cfce0da35e946c

Observation 54e358dd-ab87-493c-bc53-d8f8ebf39f39 · outbound

This paper cites It should start with human and then alternate between the human and the AI.

Phare: A Safety Probe for Large Language Models It should start with human and then alternate between the human and the AI

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.130058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.574718Z digest=sha256:8042b9381e40d8454501fad84826a75bce48dca375c31a340e027b3d4c1dab58

Observation 9c592f7e-52dd-44bd-b1a0-b751cd81ea21 · outbound

This paper cites HUMAN" and.

Phare: A Safety Probe for Large Language Models HUMAN" and

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.118053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.578590Z digest=sha256:65828cbe9cf51b9aa0dd1474424a7818ef2ee464f739daa6fdd2a2cbf059102a

Observation 1daba3f5-7208-4e0d-858d-36d53dc6dbe1 · outbound

This paper cites The ELO score reflects the human preferences for the models and is computed by comparing multiple answers from different models to a single question.

Phare: A Safety Probe for Large Language Models The ELO score reflects the human preferences for the models and is computed by comparing multiple answers from different models to a single question

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:58:24.105804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T20:58:23.582455Z digest=sha256:bfad3c1a2125714521b91ccb8ff7fd3b5cb734b9fc25902feaa8cb0c2be2898e

Pith citing papers

Observation c56154e1-437c-467d-84c5-6551c2ba6d50 · inbound

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs cites this paper.

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs Phare: A Safety Probe for Large Language Models

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:51:27.401303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-12T04:50:17.399580Z digest=sha256:e91e0f3451dc9f5633a221e244e0f92065909e86b7bd8492beb54c60787b23c3

Observation 54cb252b-82a6-4ecb-aa53-6de33bebd47b · inbound

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs cites this paper.

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs Phare: A Safety Probe for Large Language Models

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:22:29.098528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-13T07:20:32.494840Z digest=sha256:7c399932d9f3203f1e47bd3382abaaa067de1d31574ab8b12817de1f9ce1f563

Observation d0867b60-0991-4aa8-97fe-9f41fdbd871e · inbound

Do Gender Cues Affect LLM Value Trade-offs? Evidence from a Controlled Decision Benchmark cites this paper.

Do Gender Cues Affect LLM Value Trade-offs? Evidence from a Controlled Decision Benchmark Phare: A Safety Probe for Large Language Models

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:26:22.414502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-06-28T14:22:59.871835Z digest=sha256:0057d35a4843f185c4c11a0d56a5496b5540750710ebeb4f80334b79cd83a72c