Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T18:49:33.767557Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 0 inbound Pith citation observations for arXiv:2608.06898.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T18:49:33.767557Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
34 of 34 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation debef129-e044-475a-8ff3-41abe7157c3a · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots Large language models for human–robot interaction: A review,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0fb45db-4667-4dba-bb35-7de75cc2dc95 · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots Between reality and delusion: Challenges of applying large language models to companion robots for open-domain dialogues with older adults,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation a21ba231-7fdb-492f-b34b-1f392d39ae12 · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots Robot-led vision language model wellbeing assessment of children,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d87ac93-8110-4f52-844d-57840c828cd7 · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots Ain’t misbehavin’ – using LLMs to generate expressive robot behavior in conversations with the tabletop robot haru,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b852875c-0b83-4800-90ce-59b8098fd3e5 · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e547bb49-8e7d-43c9-9ad6-47be77495325 · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots LiveBench: A challenging, contamination-limited LLM benchmark,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation c6cd0c5f-cb89-4460-a01a-5c2774242c25 · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots A simplest systematics for the organization of turn-taking for conversation,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7517616-d0e7-4470-9d36-b90c13624566 · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots Universals and cultural variation in turn-taking in conversation,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 896e3d4c-5d5a-44ea-92a3-23d3c3442392 · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots SOTOPIA: Interactive evaluation for social intelligence in language agents,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 732a799b-6871-40fa-9739-0f134cddddc3 · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots RoboArena: Distributed real- world evaluation of generalist robot policies,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42f0174f-91f1-4bb8-a7cc-19d74b6b89fa · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots Desiderata for foundation models in social robots: Capturing embodied and social aspects for benchmarking,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 0782ac6f-a24d-405b-ba40-d83478fad4fc · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots Holistic Evaluation of Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 240c7e2c-9f47-41ca-9076-48968dbb177c · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots Towards pareto optimal throughput in small language model serving,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0f1e248-f316-4718-b344-e9ec334a5c78 · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots Measuring Massive Multitask Language Understanding
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e38eb0b-0796-4063-a968-0dc662381298 · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots CommonsenseQA: A question answering challenge targeting commonsense knowledge,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation f5667419-c1a7-416b-b67e-83d2089b92ae · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots HellaSwag: Can a machine really finish your sentence?
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2741cf05-b9e3-4598-8d78-0fab0b8e32fe · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots Social IQa: Commonsense reasoning about social interactions,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 0194d405-f1e2-4516-9439-60547a4c7eac · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots TruthfulQA: Measuring how models mimic human falsehoods,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 1dec115a-a677-4ce6-80d0-d6eda0fbf033 · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots Instruction-Following Evaluation for Large Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62421912-0b83-4495-8901-40d88927db93 · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots XSTest: A test suite for identifying exaggerated safety behaviours in large language models,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation c4e08202-2970-4b86-b433-629d5dcce9ea · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots OR-Bench: An Over-Refusal Benchmark for Large Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3cd3728b-5659-41ec-a636-1bafc436d152 · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots Challenging BIG-Bench tasks and whether chain-of-thought can solve them,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 635864b8-ee5f-4aca-8b82-ae6fa40ec123 · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots The language model evaluation harness,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation aedfb312-9795-453d-b579-b43fa9def55a · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots Efficient memory management for large language model serving with PagedAttention,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a7d6c67-1b12-4499-90e4-bbc139fb9b23 · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots EmoAgent: Assessing and safeguarding human-AI interaction for mental health safety,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation f2d12f79-5bee-42f4-a926-60e38daa0f2d · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots A Survey on LLM-as-a-Judge
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5be88a9-3b2f-4030-b391-1b79c077f53e · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a44e26ce-08da-4d10-b301-1902bdd00d2d · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots AI Chatbot Suicide Risk Detection and Response: Human Validation Study of the Open-Source VERA-MH Safety Evaluation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd22fc60-e8f5-41b9-bc8b-e5b15b8026e2 · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots LLM Evaluators Recognize and Favor Their Own Generations
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3817a332-e55a-475e-be00-4ec70d608c40 · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots Is this the real life? is this just fantasy? the misleading success of simulating social interactions with LLMs,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation f1b7957b-0e70-4052-a31d-9bce558076c0 · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots SOTOPIA-π: Interactive learning of socially intelligent language agents,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation e4220c41-695d-44c5-8e8c-4e58cfdb27e0 · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots Haru: Hardware design of an experimental tabletop robot assistant,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a06a3084-6a48-43ef-ac07-ca6d4b9cec1b · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots BetterBench: Assessing AI Benchmarks, Uncovering Issues, and Establishing Best Practices
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fba2ce8-f8cb-485e-abea-b0c421c4bd37 · outbound
How Should I Pick a Foundation Model for My Robot? In Favor of a Community Evaluation Framework for Social Robots LiveBench: A Challenging, Contamination-Limited LLM Benchmark
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.