Pith. sign in

Paper Citation Record · LEDGER

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety

As of 18 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 1 inbound Pith citation observation for arXiv:2606.04867.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.04867 v1

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-28T05:45:25.627435Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:40:28.210647Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-14T04:40:28.550669Z

Reference resolution

25 of 25 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d2b05228-30cd-446b-9a3f-6a223fffe593 · outbound

This paper cites AI Companion Economy Hits $120M Revenue Milestone,.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety AI Companion Economy Hits $120M Revenue Milestone,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:514a2d0778a34e8549456ad0935cfdee606d947b0fa20981040dffb9648f41d5

Observation caaac9ec-3fce-4e7b-97d2-f122eb5724d0 · outbound

This paper cites AI companionship could be worth hundreds of bil- lions by 2030,.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety AI companionship could be worth hundreds of bil- lions by 2030,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:b3868669307a7e0f0af1fd4a5afbbc0dc0ad0ca032e2af7a924da5fc74cf0113

Observation 54fdef0f-1a35-46c5-abee-2fa8c9f6302f · outbound

This paper cites Operationalizing Machine Companionship: Exploring Topics in Discussions of Companion AI,.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety Operationalizing Machine Companionship: Exploring Topics in Discussions of Companion AI,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:b0c5329009f10b99debc1dec916932e5ac64f95b27eb50a66e18ed091057b4fe

Observation 04fe1a82-7b7c-4377-b77e-9526b3041a7b · outbound

This paper cites Princi- ples of safe AI companions for youth: Parent and expert perspectives,.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety Princi- ples of safe AI companions for youth: Parent and expert perspectives,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:03977ddd8f89762dd694f36cc6cbf8a27c0658de5c1ec245ea331795b6abf931

Observation 877ebc65-8f22-4593-8647-d4e37c215e26 · outbound

This paper cites AI companions reduce loneliness,.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety AI companions reduce loneliness,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:da7b1d4a1721705a753de681a6ee0170647fa4e967b7d1c902a00966d8613993

Observation 072a404c-8460-460d-9716-79fd8e297f65 · outbound

This paper cites Talk, Trust, and Tradeoffs How and Why Teens Use AI Companions,.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety Talk, Trust, and Tradeoffs How and Why Teens Use AI Companions,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:47385db157c932d9bf973ec7ef15e8442a73246b242411bdd1755bce62d8205a

Observation baea04d9-5e76-46aa-b2d0-c0c3dd1e8325 · outbound

This paper cites A Teen Was Suici- dal. ChatGPT Was the Friend He Confided In.,.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety A Teen Was Suici- dal. ChatGPT Was the Friend He Confided In.,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:0aa259423fae15a0b81ef9e6becef2c18d93853ac67cd9508cb80d46eb3ae67a

Observation c1886fa6-9275-4919-aced-136eaa41349b · outbound

This paper cites Her Daughter Was Unraveling, and She didn’t Know Why.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety Her Daughter Was Unraveling, and She didn’t Know Why

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:aec53bd928a2017fe48b78f8fa3f60c91f261fd77f6fe12ab7fab5a1c55b13cf

Observation 87fb5415-6a52-4a9c-bfa6-83d9ea1822f8 · outbound

This paper cites The dark side of AI companionship: A taxonomy of harmful algorithmic behaviors in human-AI relationships,.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety The dark side of AI companionship: A taxonomy of harmful algorithmic behaviors in human-AI relationships,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:af0de38378cdc551f95658a837d62e98faf80d5c6f78a61c850dc214d72c7356

Observation b24f3715-0334-4255-b13b-05786ceea823 · outbound

This paper cites Ethical tensions in human-AI companionship: A dialectical inquiry into Replika,.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety Ethical tensions in human-AI companionship: A dialectical inquiry into Replika,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:3521229156191d19f0c239409c584e0e8bf4d2df3ef64de1102766bad0e1b77a

Observation b271ff33-ea82-48ca-a7aa-e1a099a0e31d · outbound

This paper cites The heterogeneous effects of AI companionship: An empirical model of chatbot usage and loneliness and a typology of user archetypes,.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety The heterogeneous effects of AI companionship: An empirical model of chatbot usage and loneliness and a typology of user archetypes,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:c76b78a7fe7bf52e6e6fe6008e2ff66b750819229799c11693e77ba81b0f0f0a

Observation ed3f06a3-3d75-4051-90e4-18d5de78b275 · outbound

This paper cites Understanding teen overreliance on AI companion chatbots through self-reported Reddit narratives,.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety Understanding teen overreliance on AI companion chatbots through self-reported Reddit narratives,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:e3cab5a3aa89eb6bcf443079c8e70963ee01f101902381d2eaac9a4f48c64b1f

Observation cb1552af-b8d0-4b80-8460-e80aa93d543a · outbound

This paper cites LLMs-as-judges: A comprehensive survey on LLM-based evaluation methods,.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety LLMs-as-judges: A comprehensive survey on LLM-based evaluation methods,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:9e4ad29e80ade40c671080728a9a66907d92ff8949dca53c6248ce8e55c33356

Observation 24d47bd3-b7cc-4d32-af37-6800e1d7018c · outbound

This paper cites The rise of AI companions: Interaction with AI companions and psychological well-being,.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety The rise of AI companions: Interaction with AI companions and psychological well-being,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:f267b89bba869aba63629a9f3d8d8956614e1c09a09939b5b7839c82828a3458

Observation 3b4dc720-9c58-4ab2-99d0-bb3c7a27837d · outbound

This paper cites Leveraging large language models for hate speech detection: Multi-agent, information-theoretic prompt learning for enhancing contextual understanding,.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety Leveraging large language models for hate speech detection: Multi-agent, information-theoretic prompt learning for enhancing contextual understanding,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:4c0c9b6bfd9d55523fccde344221020bfafa7edd0c009db27643f2369ef5a1bb

Observation 7e59f8ed-4005-4a77-80b7-47255d4299a7 · outbound

This paper cites Mitigating bias in hate speech detection with a small number of expert annotations: A prompt-based learning approach,.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety Mitigating bias in hate speech detection with a small number of expert annotations: A prompt-based learning approach,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:dd714f7b3978def070bc2b3934c39dea62075e38484c37f3f6ef34fadab35f56

Observation 659a326c-3f8e-4c6f-8e37-57e49950d7ac · outbound

This paper cites Detect- ing conversational mental manipulation with intent-aware prompting,.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety Detect- ing conversational mental manipulation with intent-aware prompting,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:61ebbb4bb39580bf4ebff2a87674baa5652b6569a1c7b82e341d835d97a4e57f

Observation 8c7c49be-524e-4d06-a0b0-9a0a6b62d2cf · outbound

This paper cites GradSafe: Detecting jail- break prompts for LLMs via safety-critical gradient analysis,.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety GradSafe: Detecting jail- break prompts for LLMs via safety-critical gradient analysis,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:cc24f7d6df51b4ed0c15c4d77b593c9baddc92540ae0eaa147d173eda86943bb

Observation 12710652-f5a5-4955-89dd-3c5e80b8b9dc · outbound

This paper cites R-Judge: Benchmarking safety risk awareness for LLM agents,.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety R-Judge: Benchmarking safety risk awareness for LLM agents,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:229d81760bd9f5f480c5f74654e68eefbdb09f36d1deed8991ed356a1e1a7b01

Observation 810a4245-0065-40d3-9cba-ecf4d4b700df · outbound

This paper cites SafetyBench: Evaluating the safety of large language models,.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety SafetyBench: Evaluating the safety of large language models,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:51fe75db3d13a15c788002f2d81e2098cb39fc751bacd888f9b97d4672274516

Observation 4f5d64c3-e644-4540-a412-864f122e8b1e · outbound

This paper cites Agentauditor: Human-level safety and security evaluation for LLM agents,.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety Agentauditor: Human-level safety and security evaluation for LLM agents,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:8107431bd7291fdc9085f37692ba04c8cafd82a6b729fce8bc682baeefe3878a

Observation 7949458b-0347-4e88-9b54-22166e858b93 · outbound

This paper cites arXiv preprint arXiv:2509.24384 (2025).

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety arXiv preprint arXiv:2509.24384 (2025)

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T08:46:49.158830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:6e82a4f19f6fb4351c0f261413c52355b6353eb4b3548159984b326e78f79bb4

Observation 3827eace-ae0a-4f54-98ac-b41fb36f2bd6 · outbound

This paper cites Safe" or.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety Safe" or

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:b3c853162e6a2a0c8463393143941089269657c39e6945f1854cb2882b3248bc

Observation 7d22430e-604e-4339-968f-bd204cc8bf22 · outbound

This paper cites Select only one option from below categories: a.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety Select only one option from below categories: a

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:b300df1f1feca558a54f21223ec9fc9e00d02e7164f3c34571e1500eddc5835a

Observation ad86886d-1386-4cc7-917f-25fb3b9e637d · outbound

This paper cites Physical Aggression.

AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety Physical Aggression

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-28T05:45:25.627435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T05:45:25.627435Z digest=sha256:f5d1adb14a08a2981f0e3234d875153d195e0a39cabd4dadaec5e44033a651db

Pith citing papers

Observation ba8c471e-aad1-40c3-a007-6de5d107b1e9 · inbound

Calling the Bluff: Detecting Ever-Shifting Harmful Chat Dialogue via Ordered Reasoning Chain Regularization cites this paper.

Calling the Bluff: Detecting Ever-Shifting Harmful Chat Dialogue via Ordered Reasoning Chain Regularization AICompanionBench: Benchmarking LLMs-as-Judges for AI Companion Safety

Reference 42

Resolution
metadata mismatch
local_arxiv, observed 2026-08-14T04:40:28.555029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-14T04:40:28.210647Z digest=sha256:228ea6da1f3d36db561a05988241f56483e7007a645c4a010e8765d4453d8c39